MODEL
Claude Sonnet 5
Overview
Claude Sonnet 5 is Anthropic‘s June 2026 mid-tier flagship — the successor to the Sonnet 4.x line — launched June 30, 2026 with a native 1M-token context window, stronger agentic-reasoning and tool-use benchmarks than the prior Sonnet tier, and promotional pricing of $2 input / $10 output per Mtok through August 31, 2026 (reverting to $3/$15 after) — roughly half the standing Claude Opus 4.8 price. Sonnet 5 is the first Sonnet-tier model to match or edge Opus-tier scores on published knowledge-work and tool-use benchmarks while still trailing on deep coding, and its same-day availability inside Claude Code v2.1.197 as the default model collapses the model-to-tooling lag to zero.
Timeline
- 2026-07-01-AI-Digest — Anthropic launches Claude Sonnet 5 on June 30 with a native 1M-token context window and promotional pricing of $2/$10 per Mtok through Aug 31 (then $3/$15). Independent-outlet benchmark reporting shows Sonnet 5 matches Claude Opus 4.8 on HLE-with-tools (57.4 vs 57.9), edges it on GDPval-AA v2 (1,618 vs 1,615) — the first time a Sonnet-tier model has outscored an Opus-tier model on any published benchmark — and still trails on SWE-bench Pro (63.2 vs 69.2). Available same-day in Claude Code
v2.1.197as the default model with 1M-context access gated on the version bump. The narrow read: Anthropic closes the Sonnet-to-Opus quality gap on knowledge-work and tool-use benchmarks specifically, not on deep coding. The structural read: for tool-use-heavy agent stacks — the ones that dominate the enterprise-agent surface — Sonnet 5 makes the “default your agent to Opus” calculus harder to justify at 2× the price, while the coding-agent case for Opus 4.8 stays intact. Ships alongside Claude Science in beta, joining the Claude Code / Claude Design / Claude for Small Business workflow-surface cluster. - 2026-07-03-AI-Digest — Sonnet 5 surfaces today as the effective-cost comparator anchor in OpenAI‘s GPT-5.6 Sol / Terra / Luna preview framing: the digest explicitly reframes the Sonnet-5-vs-Terra effective-cost story as a three-variable comparison (tokenizer ratio × per-token rate × cache-reuse rate) rather than the two-column table the initial promo-pricing analysis assumed — extending the 2026-07-02-AI-Digest Willison tokenizer-inflation reading by adding cache-reuse as the third variable. Same digest: Sonnet 5 is called out in the Aider polyglot freeze framing as one of the two frontier releases inside the freeze window (alongside GPT-5.6 Sol preview) that haven’t yet posted polyglot numbers — the 1–3-week typical inclusion lag is the load-bearing detail. No fresh Sonnet-5-specific launch or feature today; comparator-anchor role only.
- 2026-07-02-AI-Digest — Simon Willison measures Sonnet 5’s new tokenizer producing ~1.4× more tokens on English prose, ~1.33× on Spanish, and ~1.28× on Python than Sonnet 4.6 for the same input — the $2/$10 promo through August 31 (reverting to $3/$15 after) is effectively a ~30% stealth per-request price increase on top-line English workloads once tokenizer inflation is priced in. Finout independently corroborates the tokenizer change and the practical-cost implication; the “$2/$10 headline is cost-neutral” framing does not survive contact with the tokenizer numbers. Willison also flags that Sonnet 5’s agentic tool and platform-features surface is unchanged from Sonnet 4.6 — this is a performance-only upgrade wrapped in an aggressive-looking price line, not a new-capability release. The structural read worth carrying: promotional per-token pricing is now a two-variable problem for cost-model comparisons, not a one-variable one — the corpus should carry “tokens per input” alongside ”$ per token” whenever a new model’s pricing is compared to its predecessor’s.
- 2026-07-09-AI-Digest — Sonnet 5 surfaces today as the “cheaper worker” tier in The Decoder’s writeup of Anthropic‘s Advisor and Orchestrator cost patterns pushed through Claude Managed Agents. Advisor (Sonnet 5-first, calls Claude Fable 5 for guidance) reaches ~92% of Fable-solo on SWE-Bench Pro at ~63% of the cost, using ~1 Fable call per task. Orchestrator (Fable plans, Sonnet workers execute) hits ~96% of Fable on BrowseComp at ~46% of the cost, spreading Fable’s reasoning cost across a Sonnet worker pool. Narrow read: these are Anthropic-reported numbers on two specific benchmarks — Sonnet 5’s Advisor / Orchestrator roles are the load-bearing evidence that the Sonnet-tier is now positioned as the default execution tier underneath the flagship rather than a discount alternative to it. Structural read worth carrying: paired against today’s OpenAI GPT-Live-1 → GPT-5.5 delegation shape, manager-delegates-to-cheaper-worker is becoming the default agentic architecture cross-lab, with Sonnet 5 as the Anthropic side’s cheaper worker of choice — the $2/$10 promo pricing is what makes the Orchestrator math work. Watch whether the Aug 31 reversion to $3/$15 shifts the Orchestrator cost-savings math meaningfully. Same digest: Grok 4.5 and GPT-5.6 Sol land as external comparators for Sonnet 5’s positioning in the Q3 pricing lineup — Terra at $2.50/$15 positioned as “half of Sol, matches GPT-5.5” now sits right on top of Sonnet 5’s promo price band.
- Tokenizer Inflation Turns $2/$10 Promo into a ~30% Stealth Price Increase (July 2, 2026): Simon Willison‘s post-launch measurement finds Sonnet 5’s tokenizer producing ~1.4× more tokens on English prose, ~1.33× on Spanish, and ~1.28× on Python than Sonnet 4.6 for the same input; Finout independently corroborates. On top-line English workloads the promotional $2/$10 pricing through August 31 reads as a ~30% stealth per-request price increase once tokenizer inflation is priced in. Sonnet 5’s agentic tool and platform-features surface is unchanged from Sonnet 4.6 — a performance-only upgrade, not a new-capability release. The corpus carry-forward: promotional per-token pricing is now a two-variable comparison (“tokens per input” alongside ”$ per token”), not a single-variable one.
Key Developments
-
First Sonnet-Tier Model to Outscore an Opus-Tier Model on a Published Benchmark: Sonnet 5’s GDPval-AA v2 edge over Claude Opus 4.8 (1,618 vs 1,615) is the first time a Sonnet-tier model has crossed above an Opus-tier score on any public benchmark. Matches Opus 4.8 on HLE-with-tools (57.4 vs 57.9) and still trails on SWE-bench Pro (63.2 vs 69.2). Capability parity is knowledge-work-and-tool-use-scoped, not coding-scoped.
-
Promo Pricing Re-Anchors the Default-Agent-Tier Decision on Tool Use: $2/$10 per Mtok through Aug 31 (then $3/$15) is roughly half the standing Claude Opus 4.8 price. For tool-use-heavy enterprise agent stacks the “default to Opus” calculus gets harder; for deep-coding-agent workflows Opus 4.8 stays intact. The 60-day watch item is whether the promotional pricing ends on Aug 31 as scheduled or gets extended — the reversion to $3/$15 is the load-bearing detail for practitioner-cost planning.
-
Native 1M Context and Same-Day Claude Code Availability: Native 1M-token context gated on Claude Code
v2.1.197; earlier2.1.xbuilds fall back to standard windows. Same-day CLI availability collapses the model-to-tooling lag to zero and continues the pattern of Claude Code as the launch surface for the model rather than a downstream integration.
- 2026-08-09-AI-Digest — Sonnet 5 is one of the three Anthropic frontier SKUs audited by Trajectory Labs alongside Claude Fable 5 and Claude Opus 5 as part of the Aug 8 Auto Mode default-on announcement — 72 attack scenarios × 10 runs against Fable 5 / Opus 5 / Sonnet 5 with Auto Mode engaged logs 0/720 successful prompt-injection attacks (vs 5.83% success rate against pre-classifier GPT-5.6 Sol baseline). Narrow read: third-party 0/720 result on a specific test suite — the strongest cross-SKU injection-audit datapoint Sonnet 5 has appeared in, and the load-bearing complement to Sonnet 5’s role as the default execution tier inside Anthropic’s Advisor / Orchestrator agent patterns from 2026-07-09-AI-Digest. Structural read: Sonnet 5’s inclusion in the Trajectory Labs audit sharpens its enterprise-execution positioning — Auto Mode’s 0/720 result on Sonnet 5 alongside Fable 5 / Opus 5 means the classifier-not-approval-gate posture Anthropic is shipping Aug 14 for Pro / Max / Team covers the tier that most enterprise agent stacks are running as their default worker, not only the flagship. No fresh Sonnet 5 product action today; log as part of the Auto Mode audit stack.
- 2026-08-12-AI-Digest — Anthropic on Aug 11 makes Sonnet 5’s $2 / $10 introductory pricing permanent and cancels the previously-scheduled Sept 1 step-up to $3 / $15 — the
@claudeaiX post is unambiguous (“will remain unchanged”). First frontier lab to publicly un-schedule a mid-2026 price increase. Load-bearing framing to carry: “un-scheduling a published increase,” not “a price cut” — the floor was already $2 / $10 since launch; what moves is the previously-announced ceiling, and coverage that reports “Anthropic drops Sonnet 5 price” has the sign wrong. Structural read the corpus carries: Sonnet 5 is now positioned to compete on cost-per-workflow-token for agent traffic while Claude Opus 5 and Claude Mythos 5 continue at their standing rates — mid-tier competing on a different axis than the top-of-stack, and the 2026-07-09-AI-Digest Advisor / Orchestrator cost-savings math (~63% of Fable-solo cost on SWE-Bench Pro) survives the reversion-day pass without the Sept 1 recalibration. Sonnet 5 remains absent from the Aider polyglot leaderboard — the price move is a demand / capacity call, not a fresh-benchmark-tied repricing. Bundle carefully with the 2026-07-02-AI-Digest Willison tokenizer-inflation ~30% stealth-increase datum: the headline price is now durable, but the effective per-request cost still needs the tokenizer variable priced in. 30 / 60 / 90-day watch: whether OpenAI un-schedules its own mid-tier pricing on GPT-5.6 Sol / GPT-5.6 Luna; whether Sonnet 5 finally appears on Aider’s polyglot board with the pricing framed as durable; whether Anthropic couples this with a batch / cache-write discount refresh.
- $2 / $10 Made Permanent, Sept 1 $3 / $15 Step-Up Cancelled — First Frontier Lab to Un-Schedule a Published Increase (August 11, 2026): The Aug 11 move is “cancels the ceiling,” not “cuts the price” — the floor was already $2 / $10 at launch. Sonnet 5 continues competing on cost-per-workflow-token for agent traffic while Claude Opus 5 and Claude Mythos 5 hold their standing rates. Structural read: the Advisor / Orchestrator manager-delegates-to-cheaper-worker economics from 2026-07-09-AI-Digest now sits on durable Sonnet-5 pricing rather than a temporary promo window, and the corpus should stop counting down to Aug 31 in cost projections. The tokenizer-inflation caveat from 2026-07-02-AI-Digest still holds — effective per-request cost is headline × tokens-per-input × cache-reuse — but the headline commitment is now permanent. 30-day watch: whether OpenAI matches on the mid-tier; whether the Aider polyglot board finally admits a Sonnet 5 submission now that the pricing is durable.
- 2026-08-13-AI-Digest — Sonnet 5 is named alongside Claude Haiku 4.5 and Sonnet 4.6 as one of the Claude models covered by Anthropic‘s Aug 11 global watermarking commitment — invisible machine-readable watermarks embedded in text output across the Platform API, claude.ai, Claude Code, and cloud partners; C2PA-signed provenance on generated files (.svg/.png/.jpg); “may persist through some editing.” Older Claude models exempt during the transition — the load-bearing distinction is the model-version cutoff, not the geography, with global rollout not EU-only geofencing. Narrow read to carry: Sonnet 5 sits inside the watermarked cohort by virtue of being a post-Aug-2 Claude release — the version-cutoff enforcement policy pulls it into the same detection surface as Sonnet 4.6 and Haiku 4.5. Structural read: the same-day arXiv:2608.09867 encrypted-CoT extraction preprint (Panfilov et al., 367 PII + 182 credentials recovered from 315,000+ decoded blocks in public logs) frames Sonnet 5’s encrypted-CoT surface as one of the API responses now known to be portable across sessions/users/models within a family — providers have patched, patch status varies. No fresh Sonnet 5 product action today; log as watermarked-cohort inclusion + encrypted-CoT class reference. 30 / 60 / 90-day watch: whether cross-tool editing chains preserve the C2PA signal on Sonnet 5 outputs at practitioner-relevant fidelity; whether Anthropic’s CoT-portability patch specifics get published for Sonnet 5’s response class.
Related
See also: Anthropic, Claude Opus 4.8, Claude Code, Claude Design, Claude for Small Business, MOC - Major Companies, MOC - Agentic Coding, MOC - Developer Tools.