Anthropic launches Claude Sonnet 5, saying it nears Opus 4.8 performance at lower prices and is substantially better than Sonnet 4.6 for agentic work
Claude Sonnet 5 is built to be the most agentic Sonnet model yet. It can make plans, use tools like browsers and terminals …
Anthropic
Context & Ripple Effects
Anthropic’s Sonnet line has repeatedly absorbed capabilities previously associated with its flagship Opus tier: Claude 3.5 Sonnet was positioned ahead of Claude 3 Opus in some tests, and Sonnet 4.6 added coding, computer-use, instruction-following, and a beta 1M-token context window.
Sonnet 5 extends that product arc toward agentic work and tool use. The significance is not merely a new model release, but Anthropic’s claim that a lower-priced tier is approaching the performance of a newer premium tier for work that involves planning and operating software tools.
First-order effects
Anthropic can offer customers a lower-priced Claude option for agentic workflows while reserving Opus as its premium tier, assuming its stated performance and pricing positioning holds in deployment.
Developers using Sonnet 4.6 for coding, browser, terminal, or other tool-mediated tasks gain a new default candidate designed around those workflows.
Second-order effects
The narrower gap between Sonnet and Opus increases pressure on Anthropic to differentiate its flagship models by reliability, capability ceilings, or specialized use cases rather than model-family labels alone.
Competing model providers will face stronger demand to show not just benchmark gains but cost-effective performance in long-running, tool-using workflows, where per-task economics matter to customers.
Third-order effects
If capable agentic behavior continues moving from flagship models into lower-priced tiers, AI adoption may shift from experimentation toward broader operational deployment, because more workloads become economically viable to automate or augment.
The resulting differentiation battle is likely to center on dependable execution with tools, context handling, and deployment controls—not simply the release of ever-larger general-purpose models.
The trend: This is part of the continuing commoditization of frontier-model capability into lower-cost model tiers, with agentic software work becoming the key competitive proving ground.
Yeah Sonnet 5 furthers the case. They stumbled on the next paradigm of training that is way beyond language modeling and the remaining lead is much less obvious.
what is the fucking point of saying this for Opus specifically? all compared models are “reference”. these jerks are finding new ways to trigger me [image]
Sonnet 5.0 just dropped. Early vibes > huge token guzzler, promotion pricing helps > step above 4.6, but not really as good as Opus 4.8 Opus 4.8 still seems better than Sonnet 5 in terms of cost/performance trade-off. Will confirm on LiveBench shortly
While this isn't the release we're waiting on from Anthropic (wen Fable!?) it's absolutely welcome! Sonnet 5 is here folks 🔥 It's more Agentic, it's faster than Opus 4.8 obviously while also being fairly close in performance!? [image]
BREAKING claude releases Sonnet 5. Here is what they didn't say. Opus 4.8 is completely nerfed into the ground since these benchmarks. almost unusable. this means Sonnet 5 will become the new standard moving forward until we get fable 5 back.
@claudeai Sonnet 5 climbed hard on agentic search. Huge implecations for agentic-workflow. The old Sonnet is basically dead weight on this point. [image]
It happened. Claude Sonnet 5 released. Dominates Opus 4.6. Almost as good as Opus 4.8. A fraction of the price. It's also significantly faster than Opus, while being incredibly agentic. Without a doubt the new go to model for Hermes/OpenClaw Do yourself a favor and do this [image…
Sonnet 5 also holds up better in agents that run unattended. It keeps state across many steps, recovers from errors without losing the thread, and checks its own work as it goes. More multi-step runs finish correctly the first time.
Claude Sonnet 5 is live. It does Opus-level work at Sonnet-level pricing. Sonnet jumps from 5.3% to 13.5% on @Zapier's AutomationBench, scoring more than double Sonnet 4.6 on multi-step workflows. It's got all the speed and none of the stalling. It trusts what it finds, knows [im…
Introducing Claude Sonnet 5, our most agentic Sonnet yet. It makes plans, uses tools like browsers and terminals, and runs autonomously at a level that just a few months ago required larger and more expensive models. [video]
Sonnet 5 is a substantial improvement over Sonnet 4.6 on reasoning, tool use, coding, and knowledge work. Its performance is close to Opus 4.8, at lower prices. [image]
Sonnet 5 — SURPRISE: it's their most agentic model yet (in fact, Sonnet 5 also made this chart) — near Opus — $2/$10 until Aug 31 — recommended for coding — www.anthropic.com/news/claude- ... [image]
We've been running Anthropic's Claude Sonnet 5 through the Box AI Complex Work Eval, our agentic benchmark that puts models through real enterprise document work end-to-end. Sonnet 5 holds frontier-class quality on complex multi-step work and pulls ahead of Sonnet 4.6 in several
Here is my first assessment of Sonnet 5: Sonnet 5 is better than Sonnet 4.6. Who would have thought? But jokes aside: Unfortunately, it is weaker than Opus 4.8 across all evals. Why they nevertheless labeled the latest Sonnet 5 iteration with a “5”, even though “4.8” would have […