MODEL

Claude Opus 5.5

modeltopic-noteanthropic

Overview

Claude Opus 5.5 is Anthropic‘s Sept 22, 2026 Opus-tier frontier release, framed by the company as “Claude Fable 5.1-level intelligence” at 40% lower run cost than Claude Opus 5 and >30% faster output. Per-token API pricing dropped from $5/$25 to $4/$20 per Mtok (a 20% cut), while cache reads dropped to $0.20/Mtok — a 60% cut. The “40% less to run” figure is Anthropic’s own composite; the per-token comparison is 20% and the cache-read discount is where the rest lives. Agentic workloads (multi-turn, prompt-cache-heavy) capture the full 60% cache-read cut, while single-shot API calls only see the 20%. Claude Code v2.1.280 swapped the Opus default to claude-opus-5-5 the same day.

Timeline

  • 2026-09-23-AI-Digest — Anthropic releases Claude Opus 5.5 on Sept 22, framed as Claude Fable 5.1-level intelligence at 40% lower run cost than Claude Opus 5 and >30% faster output. Per-token pricing drops from $5/$25 to $4/$20 per Mtok (a 20% cut); cache reads drop to $0.20/Mtok — a 60% cut. Anthropic also lifted five-hour usage caps on Pro/Max/Team/Enterprise seats and flagged Sonnet 5.5 and Haiku 5.5 for the coming weeks. Same-day, Claude Code v2.1.280 swapped the Opus default over to claude-opus-5-5. Lands the same day as OpenAI‘s GPT-6 Sol / GPT-6 Luna release — the sub-24h dual launch with matched per-token cuts is the coordinating datapoint.
  • 2026-09-24-AI-Digest — Claude Opus 5.5 is the model running the 21-hour autonomous Anthropic agent-loop that surfaced a novel RNA-repeat + reverse-transcriptase enzyme system with Feng Zhang’s on-the-record “genuinely intriguing” quote. Load-bearing framing: the agent loop (not a specialised model) produced the candidate — Opus 5.5 acts here as general-purpose reasoning substrate, not a bespoke biology model. Also referenced as reasoning-tier pricing anchor in the “Tokens too cheap to meter” HN discussion at $4/$20, where the tier-specific-not-market-wide framing is load-bearing (Luna/Flash-tier commodity vs Opus 5.5/gpt-5-high reasoning-tier). Aider polyglot top-5 rescore hasn’t rotated to include Opus 5.5 yet — measurement lag from the Sept 22 paired-launch continues.
  • 2026-09-25-AI-Digest — Opus 5.5 surfaces in two HN threads today. (1) “Opus 5.5 is good at explainer videos” (~172 pts · ~100 cmts, launchvideo.io) — practitioner writeup showing Claude Opus 5.5 producing end-to-end explainer videos, another datapoint that frontier LLMs are crossing into multi-modal video-production workflows practitioners are actually shipping into. (2) Opus 5.5 is again referenced as the model behind the 21-hour autonomous ART-enzyme agent loop, now under Bloomberg’s replication-doubt beat — Anthropic’s own methodology write-up disclosing 10 internal reruns that failed to rediscover the ART system. Load-bearing framing to carry: the capability demonstration (long-horizon autonomous DNA search producing a plausible novel candidate) survives; the validated-new-gene-editing-platform second-order framing does not. Opus 5.5 remains general-purpose reasoning substrate, not a bespoke biology model.
  • 2026-09-27-AI-Digest — Simon Willison published on 2026-09-26 a short piece using Claude Opus 5.5 to author an animated HTML5-canvas pixel-art kākāpō clip and Claude Code driving Playwright to record it — a small but concrete “Claude Code as creative-tool driver” data point for the growing corpus of agentic-coding-in-the-wild examples. Load-bearing softener: one hobbyist example does not settle the “which model is best at driving Playwright” question, and Willison’s post is a keynote demo rather than a systematic comparison. Reframe worth carrying: Claude Opus 5.5 + Claude Code + Playwright is a working end-to-end animation pipeline in Willison's hands today, not Claude has decisively won the tool-driving-model tier. Ties to Google‘s early-Gemini-4 post-training preview roadmap flagged by DeepMind CEO Koray Kavukcuoglu at The Information’s AI Agenda Live Summit last week (already-reported: 2026-09-26-AI-Digest); the two threads together sketch the tool-driving-model axis — Willison’s post is the shipped example, Kavukcuoglu’s summit statement is the roadmap. Log against MOC - Agentic Coding and MOC - Developer Tools.

Key Developments

  1. The “40% Less to Run” Framing Decomposes as 20% Per-Token + 60% Cache Reads (September 22, 2026): Anthropic’s headline composite is real but the underlying levers differ by workload. Per-token pricing moved $5/$25 → $4/$20 (a 20% cut) and cache reads dropped to $0.20/Mtok — a 60% cut. Agentic workloads (multi-turn, prompt-cache-heavy) capture the full 60% cache-read cut; single-shot API calls only see the 20%. Anyone quoting the 40% number for isolated one-shot cost comparisons is off by 2×.

  2. Same-Day Substrate Cutover in Claude Code v2.1.280 (September 22, 2026): Claude Code v2.1.280 swapped the Opus default over to claude-opus-5-5 on the same tag that shipped the model — same-day model-and-substrate delivery, tightening the pattern the Claude Opus 5 / v2.1.219 launch already established.

  3. Sub-24h Paired Frontier Launch With OpenAI (September 22, 2026): OpenAI shipped GPT-6 Sol and GPT-6 Luna the same day with matched per-token cuts. Simon Willison flagged the sub-24h dual launch as unprecedented; near-simultaneous frontier releases have clustered before, but a sub-day dual launch with paired price cuts hasn’t. The coordinating framing to carry: coordinating datapoint on the price curve, not price war ignition.

  4. Five-Hour Usage Caps Lifted on Pro/Max/Team/Enterprise (September 22, 2026): Anthropic also lifted five-hour usage caps on Pro/Max/Team/Enterprise seats and flagged Sonnet 5.5 and Haiku 5.5 for the coming weeks — the tier-consistent-pricing pattern from Claude Opus 5 holds while the usage-envelope loosens on the paid seats.

  • 2026-09-29-AI-Digest — Claude Sonnet 5.5 ships Sept 28 with GDPval-AA v2.1 landing within two Elo of Claude Opus 5.5, and Terminal-Bench 4.0 jumping 10.3% → 70.6% on the Sonnet tier at unchanged $2/$10 per Mtok list price. Opus 5.5 surfaces as the ceiling comparator: the Sonnet-tier is now within measurement noise on GDPval-AA v2.1, which reframes the “when to default to Opus” calculus for tool-use-heavy agent stacks. Load-bearing softener: this is GDPval-AA v2.1 (agentic knowledge-work eval) — coding-specific benchmarks like SWE-Bench Pro have not been reported by Anthropic for Sonnet 5.5 yet, so the “collapse” is agentic-workflow-and-tool-use-scoped, not deep-coding-scoped. The Sonnet-5.5 free-tier claude.ai default deployment on the same day is what turns the capability delta into a practitioner-visible change. Reframe: Sonnet-tier is within 2 Elo of Opus on GDPval-AA v2.1 at Sonnet pricing, not Anthropic obsoleted Opus 5.5. The Sept-14 2026-09-15-AI-Digest flag that Sonnet 5.5 and Haiku 5.5 would ship “in the coming weeks” resolves as 2 weeks 6 days. Log against MOC - Major Companies and MOC - Agentic Coding.

  • 2026-10-02-AI-Digest — Claude Opus 5.5 surfaces as (1) comparator anchor in Cloudflare‘s Clef launch — Clef is sized and positioned for the guardrail / classification tier that sits in front of LLMs like Opus 5.5 or Gemini 4 Argon — and (2) as the first explicit mid-tier-step-up data point in today’s Key Takeaway on the convergent $2/$10 floor. Load-bearing framing to carry: Opus 5.5’s step-up curve plus Gemini 4 Argon‘s $4/$20 standard rate are two independent data points suggesting the convergent-mid-tier $2/$10 line will not hold through 2027 as a standard rate — intro periods will. Reframe: Opus 5.5 is one of two labs with published step-up curves; signal, not confirmation, not the mid-tier floor has been abandoned. No fresh model-side action today; log as pricing-floor comparator anchor + control-plane-tier comparator. Log against MOC - Major Companies.

See also: Anthropic, Claude Fable 5.1, Claude Opus 5, Claude Code, GPT-6 Sol, GPT-6 Luna, MOC - Major Companies, MOC - Agentic Coding.