Stylized number 5 made of orange ribbons and gears – cover for Claude Sonnet 5

    Claude Sonnet 5: Agentic AI Goes Mainstream

    30. Juni 20263 min read
    Till Freitag

    TL;DR:Claude Sonnet 5 (June 30, 2026) brings near-Opus-4.8 agentic performance at Sonnet pricing: $2/$10 per M tokens (intro until Aug 31), then $3/$15. It becomes the default in Claude Code and on Free/Pro – agentic AI shifts from Opus-tier premium to standard toolkit."

    Till Freitag

    What Anthropic shipped today

    On June 30, 2026, Anthropic released Claude Sonnet 5 – the most agentic Sonnet model yet. Performance close to Opus 4.8, priced at Sonnet levels, available everywhere on day one: Free, Pro, Max, Team, Enterprise, Claude Code, and the API as claude-sonnet-5.

    ModelInput / Output (per 1M tokens)Availability
    Sonnet 5 (intro until Aug 31, 2026)$2 / $10Free, Pro, Max, Team, Enterprise, Claude Code, API
    Sonnet 5 (standard from Sep 1, 2026)$3 / $15same
    Opus 4.8 (reference)$5 / $25Premium

    Why this is a turning point

    The Sonnet 3.x generation was the on-ramp to agentic coding and tool use in 2024/25. After that, the biggest agentic jumps mostly happened in the Opus class – with Opus-class pricing to match.

    Sonnet 5 closes that gap. Anthropic itself frames it as "close to Opus 4.8 performance, at lower prices." In practice: workloads that previously required Opus (long coding sessions, computer use, multi-step agents) are now economical on Sonnet.

    This isn't only a price update – it's a default shift. Sonnet is now the default model in Free and Pro plans. Millions of users get agentic behavior with no opt-in.

    The benchmarks (per Anthropic)

    Anthropic compares Sonnet 5 against Sonnet 4.6 and Opus 4.8 in the launch post:

    • BrowseComp (agentic search): Sonnet 5 strictly improves over 4.6 and approaches Opus 4.8 at higher effort levels.
    • OSWorld-Verified (computer use): same pattern – Sonnet 4.6 trailed Opus 4.8, Sonnet 5 and Opus 4.8 now cover one range at different price/effort points.
    • Humanity's Last Exam, SWE-bench, tool use: per the system card, jumps across the board vs. 4.6, plus lower hallucination and sycophancy rates.

    If you want methodology context, we wrote about how AI benchmarks actually work (and where they break) – including Goodhart's Law and live arenas.

    What early-access partners report

    The launch post quotes Lovable, ClickHouse, Pace, Eve, and others. Common thread:

    • "Same output, fewer steps." Sonnet 5 reaches comparable results with fewer iterations – which matters because token consumption, not raw model quality, drives real cost.
    • "Stays on plan." On multi-step workflows (Salesforce updates + outreach, PR pipelines, insurance FNOL) Sonnet 5 finishes end-to-end more often instead of stalling.
    • "Checks its own output." Self-verification without explicit prompting – exactly the behavior people used to coax out with "please double-check your answer" reflection tricks (see our notes on Lovable's model routing).

    What changes for builders

    Three concrete consequences if you ship on Claude today:

    1. Switch your default model in Claude Code. For most coding sessions, Sonnet 5 is enough. Keep Opus 4.8 for the hardest brownfield tasks and anything that needs the top of the accuracy curve.
    2. Effort levels, not model switching. Sonnet 5 supports multiple effort levels (up to xhigh). Instead of ping-ponging between Sonnet and Opus, stay on one model and dial effort.
    3. Take the tokenizer note seriously. Sonnet 5 uses a new tokenizer; the same input can map to roughly 1.0–1.35× more tokens. Intro pricing is set so the transition is roughly cost-neutral – re-evaluate your pipelines once standard pricing kicks in.

    How it fits the AI race

    Sonnet 5 is Anthropic's second major move in three weeks – after Claude Fable 5 & Mythos 5. Fable took the Mythos class GA; Sonnet 5 now pushes the agentic mid-tier forward. Together, the lineup gets very clear:

    • Sonnet 5 – default agent for the broad market.
    • Opus 4.8 – high-accuracy brownfield and reasoning.
    • Fable 5 / Mythos 5 – frontier and Project Glasswing use cases.

    We've added Sonnet 5 to our timeline as the next milestone: The AI race in 43+ milestones.

    What we'll be watching

    • Real cost curves after the tokenizer change, especially for long agent runs.
    • Behavior in Claude Code over very long sessions – does self-verification hold up over hours?
    • Downstream effect of the Free/Pro default shift on the ecosystem, including Lovable, which is likely to route many flows through Sonnet 5 by default.

    Sources

    TeilenLinkedInWhatsAppE-Mail

    Related Articles

    A stylized five made of butterflies – visual for Claude Fable 5
    June 9, 20266 min

    Claude Fable 5 & Mythos 5: When AI Shifts from Tasks to Responsibilities

    Anthropic launches Claude Fable 5 and Mythos 5 – SOTA on almost all benchmarks. More interesting than the numbers: The s

    Read more
    OpenClaw Pricing Shock: How to Avoid the $500 Bill
    April 5, 20262 min

    OpenClaw Pricing Shock: How to Avoid the $500 Bill

    Anthropic just killed third-party tool coverage under Claude subscriptions. If you're running OpenClaw without prep, you

    Read more
    From Chat to Workflow: How Anthropic Is Turning Claude Into a Digital Coworker
    March 30, 20262 min

    From Chat to Workflow: How Anthropic Is Turning Claude Into a Digital Coworker

    Dispatch, Computer Use, persistent tasks – Anthropic is layering capabilities in an order that's no accident. A strategi

    Read more
    Why We Switched from ChatGPT to Claude – and What We Learned About LLMs Along the Way
    February 20, 20265 min

    Why We Switched from ChatGPT to Claude – and What We Learned About LLMs Along the Way

    We worked with ChatGPT for 18 months – then switched to Claude. Here's our honest comparison of all major LLMs and why C

    Read more
    Visualization of a large pale neural network sphere and a smaller bright sphere in cyan/yellow – the shrinking frontier of open models
    June 8, 20265 min

    Nex-N2-Pro: How the Open-Model Frontier Shrunk 75 % in Six Weeks

    Six weeks ago, DeepSeek-V4-Pro with 1.6 trillion parameters was the largest open-weight model ever released. Today, Nex-

    Read more
    Why 🦞 Became the Secret Handshake of the Agentic AI Movement
    May 19, 20263 min

    Why 🦞 Became the Secret Handshake of the Agentic AI Movement

    How a crustacean became the tribal emoji of the agentic AI scene – from Anthropic memes to X bios full of lobster claws.

    Read more
    Visualization of Kimi K2.6 long-horizon agents: a Moonshot crescent symbol alongside distributed sub-agent nodes over a coordination gridDeep Dive
    April 21, 20268 min

    Kimi K2.6: The Most Interesting AI Optimization in 2026 Isn't Intelligence – It's Duration

    Moonshot AI open-sourced Kimi K2.6 yesterday. 1 trillion parameters, 300 sub-agents, 13 hours of autonomous code refacto

    Read more
    Editorial illustration of the Claude Design launch – warm sand-tone background with the rust-orange Claude spark motif, glassmorphic UI panels showing a wireframe, color tokens, and a dashboard mockup, with subtle Adobe-red and Figma-purple accents hinting at the market disruption.
    April 17, 20265 min

    Claude Design Is Here: How Anthropic Labs Wiped $30B Off Figma, Adobe and Wix in a Single Day

    On April 17, 2026, Anthropic launched Claude Design – the first Anthropic Labs product for visual work. Powered by Opus

    Read more
    Claude Opus 4.7 Is Here: What Premium Teams Need to Know About the Tokenizer, xhigh, and Spend Controls
    April 17, 20265 min

    Claude Opus 4.7 Is Here: What Premium Teams Need to Know About the Tokenizer, xhigh, and Spend Controls

    Anthropic just released Claude Opus 4.7. Same price as 4.6, but noticeably better at coding, agents, and visual output.

    Read more