
TL;DR: „Claude Sonnet 5 (June 30, 2026) brings near-Opus-4.8 agentic performance at Sonnet pricing: $2/$10 per M tokens (intro until Aug 31), then $3/$15. It becomes the default in Claude Code and on Free/Pro – agentic AI shifts from Opus-tier premium to standard toolkit."
— Till FreitagWhat Anthropic shipped today
On June 30, 2026, Anthropic released Claude Sonnet 5 – the most agentic Sonnet model yet. Performance close to Opus 4.8, priced at Sonnet levels, available everywhere on day one: Free, Pro, Max, Team, Enterprise, Claude Code, and the API as claude-sonnet-5.
| Model | Input / Output (per 1M tokens) | Availability |
|---|---|---|
| Sonnet 5 (intro until Aug 31, 2026) | $2 / $10 | Free, Pro, Max, Team, Enterprise, Claude Code, API |
| Sonnet 5 (standard from Sep 1, 2026) | $3 / $15 | same |
| Opus 4.8 (reference) | $5 / $25 | Premium |
Why this is a turning point
The Sonnet 3.x generation was the on-ramp to agentic coding and tool use in 2024/25. After that, the biggest agentic jumps mostly happened in the Opus class – with Opus-class pricing to match.
Sonnet 5 closes that gap. Anthropic itself frames it as "close to Opus 4.8 performance, at lower prices." In practice: workloads that previously required Opus (long coding sessions, computer use, multi-step agents) are now economical on Sonnet.
This isn't only a price update – it's a default shift. Sonnet is now the default model in Free and Pro plans. Millions of users get agentic behavior with no opt-in.
The benchmarks (per Anthropic)
Anthropic compares Sonnet 5 against Sonnet 4.6 and Opus 4.8 in the launch post:
- BrowseComp (agentic search): Sonnet 5 strictly improves over 4.6 and approaches Opus 4.8 at higher effort levels.
- OSWorld-Verified (computer use): same pattern – Sonnet 4.6 trailed Opus 4.8, Sonnet 5 and Opus 4.8 now cover one range at different price/effort points.
- Humanity's Last Exam, SWE-bench, tool use: per the system card, jumps across the board vs. 4.6, plus lower hallucination and sycophancy rates.
If you want methodology context, we wrote about how AI benchmarks actually work (and where they break) – including Goodhart's Law and live arenas.
What early-access partners report
The launch post quotes Lovable, ClickHouse, Pace, Eve, and others. Common thread:
- "Same output, fewer steps." Sonnet 5 reaches comparable results with fewer iterations – which matters because token consumption, not raw model quality, drives real cost.
- "Stays on plan." On multi-step workflows (Salesforce updates + outreach, PR pipelines, insurance FNOL) Sonnet 5 finishes end-to-end more often instead of stalling.
- "Checks its own output." Self-verification without explicit prompting – exactly the behavior people used to coax out with "please double-check your answer" reflection tricks (see our notes on Lovable's model routing).
What changes for builders
Three concrete consequences if you ship on Claude today:
- Switch your default model in Claude Code. For most coding sessions, Sonnet 5 is enough. Keep Opus 4.8 for the hardest brownfield tasks and anything that needs the top of the accuracy curve.
- Effort levels, not model switching. Sonnet 5 supports multiple effort levels (up to
xhigh). Instead of ping-ponging between Sonnet and Opus, stay on one model and dial effort. - Take the tokenizer note seriously. Sonnet 5 uses a new tokenizer; the same input can map to roughly 1.0–1.35× more tokens. Intro pricing is set so the transition is roughly cost-neutral – re-evaluate your pipelines once standard pricing kicks in.
How it fits the AI race
Sonnet 5 is Anthropic's second major move in three weeks – after Claude Fable 5 & Mythos 5. Fable took the Mythos class GA; Sonnet 5 now pushes the agentic mid-tier forward. Together, the lineup gets very clear:
- Sonnet 5 – default agent for the broad market.
- Opus 4.8 – high-accuracy brownfield and reasoning.
- Fable 5 / Mythos 5 – frontier and Project Glasswing use cases.
We've added Sonnet 5 to our timeline as the next milestone: The AI race in 43+ milestones.
What we'll be watching
- Real cost curves after the tokenizer change, especially for long agent runs.
- Behavior in Claude Code over very long sessions – does self-verification hold up over hours?
- Downstream effect of the Free/Pro default shift on the ecosystem, including Lovable, which is likely to route many flows through Sonnet 5 by default.








