/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

days · browse · Enter similar · o open

Anthropic launches Claude Sonnet 5, saying it nears Opus 4.8 performance at lower prices and is substantially better than Sonnet 4.6 for agentic work

Claude Sonnet 5 is built to be the most agentic Sonnet model yet.  It can make plans, use tools like browsers and terminals …

Anthropic

Context & Ripple Effects

Anthropic’s Sonnet line has repeatedly absorbed capabilities previously associated with its higher-end models: Claude 3.5 Sonnet was presented as surpassing Claude 3 Opus on some measures, and Sonnet 4.6 later added coding, computer-use, instruction-following, and a beta 1M-token context window.

Sonnet 5 extends that product arc by emphasizing planning and tool use while being positioned close to Opus 4.8 at a lower price. The significance is not simply another model release, but a further compression of the gap between Anthropic’s mid-tier and flagship offerings for agentic workloads.

First-order effects

  • Anthropic can steer cost-sensitive developers and enterprises toward Sonnet 5 for browser, terminal, and other tool-using workflows that might otherwise have required an Opus-class model.
  • Sonnet 4.6 becomes the clearly superseded Sonnet option for agentic work, increasing pressure on existing users to evaluate an upgrade.

Second-order effects

  • A lower-priced model positioned near flagship performance raises the competitive bar for rival model providers: they must defend both agent reliability and the economics of deploying it, rather than benchmark performance alone.
  • Customers building agentic products gain a stronger incentive to move experimentation and production workloads from isolated chat use toward tool-connected automation, where model reliability and cost compound across repeated steps.

Third-order effects

  • If lower-priced tiers continue to inherit frontier capabilities, model vendors may differentiate less through a single premium flagship and more through the reliability, controls, and developer experience surrounding agent deployment.
  • The product hierarchy could increasingly be organized around workload economics and operational trust—how safely and consistently a model plans and acts with tools—rather than a simple distinction between “standard” and “frontier” models.

The trend: This is one data point in the commoditization of frontier-model capability into lower-cost tiers, with agentic tool use becoming the key competitive workload.

Discussion

  • @claudeai Claude on x
    Introducing Claude Sonnet 5, our most agentic Sonnet yet. It makes plans, uses tools like browsers and terminals, and runs autonomously at a level that just a few months ago required larger and more expensive models. [video]
  • @claudedevs @claudedevs on x
    Claude Sonnet 5 is here. Top-tier performance on coding and tool use at Sonnet pricing, with a 1M context window. It's the new default in Claude Code for Pro users, and available everywhere on the Claude Platform, including the API and Managed Agents.
  • @kimmonismus @kimmonismus on x
    Here we go: Sonnet 5 is live: The tl;dr • Anthropic calls it the most agentic Sonnet yet • Near Opus 4.8-level performance, but cheaper • Strong gains in reasoning, tool use, coding, and knowledge work • Default model for Free and Pro users • Available in Claude Code and [image]
  • @simonw Simon Willison on x
    Notes (and a Pelican) on Claude Sonnet 5 - the new tokenizer makes it ~1.4x more expensive for English, ~1.33x more expensive for Spanish but roughly the same price for Simplified Mandarin https://simonwillison.net/...
  • @artificialanlys @artificialanlys on x
    Claude Sonnet 5 achieves 53 on the Artificial Analysis Intelligence Index, but without promotional pricing will cost more per task than Opus 4.8 We supported @AnthropicAI to evaluate Claude Sonnet 5 ahead of release: with max effort it improves 6 points over Sonnet 4.6 to [image]
  • @bindureddy Bindu Reddy on x
    Sonnet 5.0 just dropped. Early vibes > huge token guzzler, promotion pricing helps > step above 4.6, but not really as good as Opus 4.8 Opus 4.8 still seems better than Sonnet 5 in terms of cost/performance trade-off. Will confirm on LiveBench shortly
  • @dorialexander Alexander Doria on x
    Yeah Sonnet 5 furthers the case. They stumbled on the next paradigm of training that is way beyond language modeling and the remaining lead is much less obvious.
  • @banteg @banteg on x
    noticed that this bench was run with safeguards turned off. on a production version it scores zero. same situation as a third-party eval finding that fable rejects 99.5% of prompts while anthropic posts self-reported scores achieved on a non-production config. this is
  • @cognition @cognition on x
    Claude Sonnet 5 is now available in Devin Desktop and Devin CLI. Sonnet 5 pairs frontier-level coding performance with a more affordable price point and outperforms Opus 4.8 on FrontierCode Extended. [image]
  • @levie Aaron Levie on x
    We've been running Anthropic's Claude Sonnet 5 through the Box AI Complex Work Eval, our agentic benchmark that puts models through real enterprise document work end-to-end.  Sonnet 5 holds frontier-class quality on complex multi-step work and pulls ahead of Sonnet 4.6 in several…
  • @daniellefong Danielle Fong on x
    sonnet 5 release vibes are bad of course because of the white elephant. but it's quite a good model, if what you like are nervous models
  • @altryne @altryne on x
    While this isn't the release we're waiting on from Anthropic (wen Fable!?) it's absolutely welcome!  Sonnet 5 is here folks 🔥 It's more Agentic, it's faster than Opus 4.8 obviously while also being fairly close in performance!?
  • @cursor_ai @cursor_ai on x
    Claude Sonnet 5 is now available in Cursor. On CursorBench, it's a meaningful step up from Sonnet 4.6: 57% vs. 49%. [image]
  • @claudedevs @claudedevs on x
    Sonnet 5 also holds up better in agents that run unattended. It keeps state across many steps, recovers from errors without losing the thread, and checks its own work as it goes. More multi-step runs finish correctly the first time.
  • @altryne @altryne on x
    Going to give this a try and definitely cover this on ThursdAI [image]
  • @claudeai Claude on x
    Sonnet 5 is a substantial improvement over Sonnet 4.6 on reasoning, tool use, coding, and knowledge work. Its performance is close to Opus 4.8, at lower prices. [image]
  • @amazon @amazon on x
    🚨 Claude Sonnet 5 just landed on Amazon Bedrock and Claude Platform on @awscloud. @AnthropicAI's most capable Sonnet model yet. We're talking advanced intelligence for coding, agents, and professional work. Available now.
  • @kimmonismus @kimmonismus on x
    OpenAI achieved a much more significant breakthrough today. Sonnet 5 is an average release. But the fact that OpenAI, according to The Information, has managed to more than halve the inference costs of its current models through a new approach to inference optimization is [image]
  • @teortaxestex @teortaxestex on x
    The one bench nobody wants to hill-climb
  • @rohanpaul_ai Rohan Paul on x
    @claudeai Sonnet 5 climbed hard on agentic search. Huge implecations for agentic-workflow. The old Sonnet is basically dead weight on this point. [image]
  • @deredleritt3r Prinz on x
    Notice that Sonnet 5 scores worse than Opus 4.8 on every single benchmark (except GDPval, on which it's 3 points higher - nothing material). This is in line with my suspicion that we have an unofficial moratorium on frontier model releases in the U.S. until the Fable 5/GPT-5.6
  • @_maxblade Max Blade on x
    BREAKING claude releases Sonnet 5. Here is what they didn't say. Opus 4.8 is completely nerfed into the ground since these benchmarks. almost unusable. this means Sonnet 5 will become the new standard moving forward until we get fable 5 back.
  • @banteg @banteg on x
    welcome to benchnerfing era, sonnet 5 weaker than sonnet 4.6 [image]
  • @mweinbach Max Weinbach on x
    Sonnet 5 medium is better than GLM 5.2 high and roughly the same price hilarious tbh
  • @openrouter @openrouter on x
    Claude Sonnet 5 is rolling out on OpenRouter with a promo price: $2/M in and $10/M out! It boosts agentic coding and pro workflows w/ flagship intelligence at Sonnet pricing. In early tests, agents were more reliable, faster, and easier to trust with larger tasks than 4.6. [image…
  • @teortaxestex @teortaxestex on x
    what is the fucking point of saying this for Opus specifically? all compared models are “reference”. these jerks are finding new ways to trigger me [image]
  • @wadefoster Wade Foster on x
    Claude Sonnet 5 is live. It does Opus-level work at Sonnet-level pricing. Sonnet jumps from 5.3% to 13.5% on @Zapier's AutomationBench, scoring more than double Sonnet 4.6 on multi-step workflows. It's got all the speed and none of the stalling. It trusts what it finds, knows [im…
  • @afinetheorem Kevin A. Bryan on x
    Sonnet 5 has...limitations. On my geography benchmark, it's at the level of GPT-4o or Gemini 2.0 Flash. Fable, btw, was right near the top...(41%, 56%, 61%, 800). Surprisingly bad - has a Haiku-level feel to it. This is with API default adaptive: high reasoning, btw. [image]
  • @github @github on x
    📣 @AnthropicAI's Claude Sonnet 5 is now generally available and rolling out in GitHub Copilot. Early testing for Claude Sonnet 5 showed: • strong results across a range of coding scenarios, particularly for performance on CLI-style tasks. • excellent prompt-cache utilization [vid…
  • @iharnoorsingh Harnoor Singh on x
    Sonnet 5 is here, closer power to Opus but 2x cheaper!!
  • @luke_metro @luke_metro on x
    the marketing team knew what they were doing by starting with a GIF of a giant 5 😭
  • @alexstamos Alex Stamos on x
    New shibboleth just dropped. [image]
  • @alexfinn Alex Finn on x
    It happened. Claude Sonnet 5 released. Dominates Opus 4.6. Almost as good as Opus 4.8. A fraction of the price. It's also significantly faster than Opus, while being incredibly agentic. Without a doubt the new go to model for Hermes/OpenClaw Do yourself a favor and do this [image…
  • @deryatr_ Derya Unutmaz on x
    Sonnet 5: less for more $$$. Thanks, but I'll skip this amazing deal, dear Claude! [image]
  • @arena @arena on x
    Claude Sonnet 5 is ready for you in the Agent Arena! In Agent Arena, we measure models on millions of real-world, long-horizon agentic tasks from a global community of users. Models can access web search, filesystem, and terminal tools to complete complex workflows. The [image]
  • @kimmonismus @kimmonismus on x
    Here is my first assessment of Sonnet 5...Unfortunately, it is weaker than Opus 4.8 across all evals.  Why they nevertheless labeled the latest Sonnet 5 iteration with a “5”, even though “4.8” would have been more fitting, is beyond me.  Normally, major version jumps in particula…
  • Rahul Patil Rahul Patil on linkedin
    Claude Sonnet 5 is available today.  Over the past year, the clearest capability gains have come from our largest models. …
  • @timkellogg.me Mr. Tim on bluesky
    Sonnet 5  —  SURPRISE: it's their most agentic model yet (in fact, Sonnet 5 also made this chart)  — near Opus  — $2/$10 until Aug 31  — recommended for coding  —  www.anthropic.com/news/claude- ...  [image]
  • r/claude r on reddit
    Introducing Claude Sonnet 5
  • r/singularity r on reddit
    Introducing Claude Sonnet 5
  • r/ClaudeAI r on reddit
    Introducing Claude Sonnet 5