Models & Assistants
New AI models, model families, assistants, and user-facing capability changes.
GPT-6 Astra is now the default model for ChatGPT Work and Codex: here's what enterprise teams need to know
OpenAI's September 9 post sets GPT-6 Astra as the default for Work and Codex, with benchmark scores, case studies, and an admin-enable requirement.
OpenAI's unreleased model solves a Millennium Prize maths problem: and the controversy may matter as much as the proof
10,000 AI agents, 88 hours, and a Lean-verified proof of finite-time singularity in Navier: Stokes. Here's what happened and why it matters.
ChatGPT Images 2.5 is here: faster generation, a sketch tool, and two new API models
OpenAI's Images 2.5 replaces 2.0 across all ChatGPT tiers and the API, cutting latency by 50% and adding Sketch, comment editing, and two new model IDs.
GPT-6 Astra is now available to all ChatGPT Plus, Pro, Business, and Enterprise users
GPT-6 Astra expands beyond its limited preview to all paid ChatGPT tiers, the OpenAI API, Azure, and AWS Bedrock: with faster computer use and new Codex note-keeping.
GPT-6 Astra is here: OpenAI's most capable model yet, and a possible AGI milestone
OpenAI launched GPT-6 Astra on 3 September 2026, with Greg Brockman calling it a generational leap and suggesting it may represent AGI.
Gemini 3.8 Flash is now generally available, with a locked-down cyber sibling for vetted defenders
Google's Gemini 3.8 Flash launched on 2 September 2026 at the same $0.75/$3.75 introductory price as 3.7 Flash, with stronger agentic and coding scores.
Claude Fable 5.1 and Mythos 5.1 are now generally available: same price, much cheaper caching
Anthropic launches Claude Fable 5.1 with cache reads cut 75% to $0.25/MTok, saving up to 45% on agentic workloads. Mythos 5.1 stays restricted.
Anthropic opens 10,000 free and discounted Claude seats for scientists
Anthropic's new Claude Team plan gives verified researchers free or $15/month premium seats, expands science credits to $50k, and enrolls first Mythos users.
Gemini 3.7 Flash is now generally available, alongside new video controls in Gemini Omni 1.1 Flash
Google's Gemini 3.7 Flash hits the API with big coding and agent gains at half price until end of 2026, plus Omni 1.1 Flash adds 4K video tools.
Claude now shares one memory across chat and Cowork, with full user control
Anthropic unified Claude's memory across Chat and Cowork on 25 August 2026, with topic-by-topic editing, sensitive topic opt-in, and a legacy export deadline.
OpenAI previews Ultrafast: GPT-5.6 Sol at 750 tokens per second, powered by Cerebras
OpenAI's new Ultrafast API tier runs GPT-5.6 Sol up to 14× faster than standard, hitting 750 tokens/sec via Cerebras hardware.
Claude improved a 160-year-old maths bound by 25 percentage points in 36 hours
An unreleased Claude research model advanced the Riemann zeta function lower bound from 41.6% to 67.2%, the largest single-step gain in history.
GPT-Live voice gets file uploads, Projects integration, and becomes the default for Enterprise, Edu and Healthcare
ChatGPT's GPT-Live voice mode now supports file uploads mid-conversation, works inside Projects, and is the default voice experience for enterprise workspaces.
OpenAI is retiring o3 from ChatGPT on 26 August 2026
OpenAI has set 26 August 2026 as the hard retirement date for o3 in ChatGPT, following a 90-day sunset period announced on 28 May 2026.
ChatGPT's 6 August 2026 update: what changed for paid and free users
GPT-5.6 Sol gets 68% fewer factual errors and a reasoning slider for Plus/Pro; Free users move to Luna with unlimited text chats.
Gemini 3.6 Flash and 3.5 Flash-Lite are now generally available
Google's new Gemini 3.6 Flash uses 17% fewer output tokens, scores 83% on computer use, and costs $1.50/$7.50 per million tokens.
Google DeepMind's Gemini Robotics 2: Whole-Body Control, Multi-Robot Teams, and Fast Adaptation
Google DeepMind's Gemini Robotics 2 is a three-model suite giving humanoid robots whole-body control, multi-step planning, and rapid adaptation to new hardware.
Gemini 3.1 Flash-Lite is now generally available — Google's cheapest production-ready model in the 3.x family
Google's Gemini 3.1 Flash-Lite has graduated from preview to GA, locking in pricing at $0.25/1M input and $1.50/1M output tokens.
Google's Gemini 3.5 Pro Is Months Behind Schedule — and the Coding Gap Is the Core Problem
Bloomberg reports Google's flagship Gemini 3.5 Pro is significantly delayed due to coding performance shortfalls, raising competitive alarms vs. OpenAI and Anthropic.
GPT-Live-1 is here: OpenAI's full-duplex voice model replaces Advanced Voice Mode globally
OpenAI's GPT-Live-1 and GPT-Live-1 mini launch July 8, replacing Advanced Voice Mode with simultaneous listen-and-speak AI powered by GPT-5.5 in the background.
ChatGPT's Model Picker Just Got a Rethink: Here's What the New Effort Tiers Mean for You
OpenAI is replacing named-model selection in ChatGPT with effort-based tiers — Instant, Medium, High, and more — rolling out to Plus and Pro users now.
Claude Fable 5 is here: Anthropic's first public Mythos-class model, with a safety wall built in
Anthropic launches Claude Fable 5 with a 1M-token context window, $10/$50 pricing, and a safety-classifier fallback — plus a restricted Mythos 5 for Project Glasswing partners.
OpenAI is retiring GPT-4.5 and o3 from ChatGPT — here's what changes and when
GPT-4.5 leaves ChatGPT on June 27, 2026 and o3 on August 26. No API changes. Here's what it means for you.
Anthropic found a hidden 'workspace' inside Claude — and built a tool to read it
Anthropic's J-lens research reveals a small internal neural workspace in Claude that mirrors neuroscience's global workspace theory, with real safety implications.
GPT-5.4 Mini and Nano: OpenAI's Fastest Small Models Are Built for the Age of AI Agents
OpenAI releases GPT-5.4 mini and nano — faster, cheaper models designed for agentic workflows, with near-flagship performance at a fraction of the cost.
GPT-Rosalind gets its first major upgrade: agentic coding, new benchmarks, and global research access
OpenAI's life-sciences model gains GPT-5.5 agentic capabilities, three new benchmarks, two Codex plugins, and opens to eligible organisations worldwide.
Claude Design now keeps your brand consistent across sessions and syncs directly with Claude Code
Claude Design moves to beta with design system persistence, WYSIWYG canvas editing, bidirectional Claude Code sync, and a new desktop app home.
GPT-5.5 Instant gets more readable — and Canvas is gone
OpenAI has improved GPT-5.5 Instant's response quality and quietly retired Canvas in favour of native writing and code blocks in chat.
GPT-5.2 Is Gone from ChatGPT — What the New 90-Day Sunset Policy Means for You
As of June 12, 2026, GPT-5.2 Instant, Thinking, and Pro are retired from ChatGPT, with all conversations auto-migrated to GPT-5.5 equivalents.
Microsoft builds its own AI models: MAI-Thinking-1 and MAI-Code-1 explained
Microsoft unveiled seven in-house MAI models at Build 2026, cutting reliance on OpenAI with its own reasoning and coding models.
Anthropic Files Confidential IPO Papers with the SEC, Targeting a Trillion-Dollar Debut
Anthropic has submitted a confidential S-1 to the SEC, setting up a potential October 2026 IPO at a valuation above $1 trillion.
Claude Opus 4.8 is here: faster, smarter, and significantly better at the work that matters
Anthropic's Claude Opus 4.8 brings stronger coding, agentic tasks, and professional workflows — at the same price as its predecessor.
Gemini 3.5 Flash: Google's fastest model is now its smartest too
Google launched Gemini 3.5 Flash at I/O 2026 — a model that outperforms the previous Pro tier at 4x the speed, now the default across Search and the Gemini app.
Gemini Omni: Google's Multimodal Video Model That Lets You Edit With Conversation
Google's Gemini Omni launches at I/O 2026, replacing Veo in the Gemini app with multimodal video generation and conversational editing from any input type.
GPT-5.5 Instant is now the default ChatGPT model — here's what changed
OpenAI replaced GPT-5.3 Instant with GPT-5.5 Instant as ChatGPT's default model on May 5, 2026, bringing fewer hallucinations and tighter responses.
Microsoft launches MAI-Transcribe-1, MAI-Voice-1, and MAI-Image-2 on Microsoft Foundry
Microsoft's new in-house AI models for speech, voice, and image generation are now in public preview on Microsoft Foundry — here's what they do and who they're for.
Anthropic launches Claude Opus 4.6 with 1M token context, agent teams, and top benchmark scores
Anthropic's new flagship model brings a 1M token context window, multi-agent coordination, and state-of-the-art results on coding and reasoning benchmarks.