The AI week, distilled.
Week 23 · 2026
This week in non‑Microsoft AI

OpenAI, Anthropic, and Google push hybrid deployment, security, and smaller multimodal models

OpenAI surfaced new flagship model and multimodal updates and signaled stronger hybrid deployment ambitions via a Dell partnership. Anthropic expanded a security-focused model preview and published new safety evaluation methods, while Google emphasized smaller multimodal models and continued alignment research.

01

OpenAI lists GPT‑5.5, Images 2.0, and Dell deal

OpenAI’s newsroom added announcements for GPT‑5.5 and ChatGPT Images 2.0 and highlighted a partnership with Dell to bring Codex capabilities to hybrid and on‑prem enterprise environments.

  • Treat the Dell partnership as a new procurement path for regulated Czech workloads that cannot rely on public cloud connectivity or cross-border data flows.
  • Re-check your model evaluation baselines because GPT‑5.5 implies another performance step that can change cost, latency, and quality assumptions in existing copilots and automations.
  • Plan for image-centric workflows (document visual QA, design review, inspection) because Images 2.0 signals continued product investment in multimodal usage inside ChatGPT.
02

Anthropic expands Project Glasswing and Claude Mythos Preview

Anthropic announced an expansion of Project Glasswing, and Cybersecurity Dive reported Anthropic is widening access to Claude Mythos Preview to additional organizations in government and critical infrastructure contexts.

  • Use Glasswing/Mythos positioning as a vendor-comparison input if your SOC or DevSecOps teams want AI tuned for vulnerability discovery and security analysis rather than general chat.
  • Ask vendors for evidence of critical-infrastructure deployment patterns and controls because Anthropic is explicitly targeting high-assurance environments.
  • Map this to your software assurance program by testing whether Claude Security-style code scanning can reduce time-to-triage for codebase and dependency risk in Czech regulated sectors.
03

Cybersecurity Dive details Mythos expansion to more organizations

Cybersecurity Dive reported Anthropic expanded Claude Mythos Preview access to roughly 150 additional organizations as part of its security initiative.

  • Benchmark security model outputs against your internal red-team playbooks because vendor claims around vulnerability hunting often fail on Czech-language context and local stacks.
  • Treat expanded access as a maturity signal when selecting a provider for security use cases that require repeatability, audit logs, and controlled rollout.
  • Use this as leverage in procurement to require clear data handling, retention, and incident response terms specific to security telemetry and code.
04

Google releases Gemma 4 12B for laptop-class multimodal use

Google introduced Gemma 4 12B, an encoder-free multimodal model designed to run efficiently on laptops while supporting audio and vision inputs.

  • Pilot on-device deployments for sensitive documents and field scenarios where Czech data governance limits cloud processing or connectivity is unreliable.
  • Factor simpler architectures into endpoint TCO because removing separate encoders can reduce memory overhead and deployment complexity across large device fleets.
  • Evaluate Gemma alongside Llama/Mistral alternatives if you need an open-model route with potential EU hosting, fine-tuning, and controllable inference.
05

Google Cloud roundup highlights Gemini 3.5 and Gemini Omni

Google Cloud’s AI roundup emphasized the Gemini 3.5 family and Gemini Omni as key recent updates across Google’s product stack.

  • Use this as a trigger to validate tool-use and agent orchestration capabilities in your target workloads because Google is framing Gemini as action-oriented rather than chat-only.
  • Reassess hyperscaler selection criteria if you want a tighter coupling between models, governance controls, and managed services under one contract.
  • Ask explicitly about EU data residency options and auditability in Gemini-enabled services before expanding to regulated Czech workloads.
06

Anthropic Institute publishes recursive self-improvement risk testing

Anthropic described a recurring internal evaluation that asks Claude to optimize training code for a small model to probe potential recursive self-improvement behavior under constraints.

  • Use the published methodology as due-diligence material for your AI risk assessments and board reporting because it provides concrete tests rather than general assurances.
  • Translate the approach into internal controls by separating productivity automation (code optimization) from autonomy escalation (unbounded training changes).
  • Prefer vendors that publish repeatable safety evaluation procedures if you expect upcoming EU AI Act enforcement to demand documented risk management.