/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

days · browse · Enter similar · o open

Anthropic announces Claude 3 Opus, Sonnet, and Haiku, aiming to reduce AI model hallucinations; Opus and Sonnet are available now, and Haiku in the coming weeks

The startup says new versions of Claude will be twice as likely to answer a question correctly.

Bloomberg Rachel Metz

Context & Ripple Effects

Claude 3 is Anthropic’s initial three-tier release: Opus and Sonnet target capability, while Haiku was positioned shortly afterward for high-volume, latency-sensitive use cases. The common claim across the launch is that reducing incorrect answers is central to making the models more dependable.

The release also established a rapid upgrade cycle. Within months, Anthropic said Claude 3.5 Sonnet surpassed the earlier Opus flagship, while the later Artifacts launch shifted attention from model responses toward interactive outputs inside Claude.

First-order effects

  • Anthropic immediately expands Claude into distinct performance and deployment tiers, giving users a choice between available Opus and Sonnet models while awaiting Haiku.
  • The company makes answer accuracy a core product claim, raising the practical bar for teams evaluating Claude for tasks where unsupported responses are costly.

Second-order effects

  • Model providers competing for the same users face pressure to differentiate on both reliability and fit-for-purpose tiers, rather than benchmark capability alone.
  • Customers can more readily match model capability to workload needs; Haiku’s subsequent positioning around speed and volume makes latency and operating economics part of the selection decision.

Third-order effects

  • If tiered releases continue to improve reliability, competition is likely to move toward AI portfolios optimized for specific workflows, not one universal flagship model.
  • The durable measure of progress becomes cost per dependable outcome: headline model gains matter only insofar as users can deploy them with fewer errors and acceptable latency.

The trend: This is one data point in the industrialization of generative AI, where vendors package improving model reliability into workload-specific tiers and product experiences.

Discussion

  • @alexrkonrad Alex Konrad on x
    News: Anthropic has released Claude 3, a trio of AI models it says can outperform rivals like OpenAI's GPT-4 and Google's Gemini 1 Ultra. @kenrickcai and I spoke to cofounders Dario and Daniela Amodei about the release for @Forbes. https://www.forbes.com/...
  • @alexrkonrad Alex Konrad on x
    @kenrickcai @Forbes Anthropic's new flagship model, Claude 3 Opus, beat GPT-4 and Gemini on a number of benchmarks. But it's pricy, and CEO Amodei admitted it's unknown how it fully stacks up against unreleased models like OpenAI's GPT 4 Turbo or Google's Gemini 1.5 Ultra. https:…
  • @alexrkonrad Alex Konrad on x
    @kenrickcai ... We spoke to Anthropic about perceptions from some that it's models have degraded over time; on the LMSYS leaderboard, Claude 1 ranks higher than Claude 2. Amodei said Claude 3 has been trained to generate far fewer “incorrect refusals” than its predecessor, withou…
  • @mattshumer_ Matt Shumer on x
    Holy shit. Anthropic's Claude 3 beat GPT-4! Testing the model now.
  • @krishnanrohit Rohit on x
    Claude 3 is out apparently. Is it better than GPT-4 or Gemini? [image]
  • @jackclarksf Jack Clark on x
    Thrilled about these new models - I've been playing around with Claude 3 Opus a lot and it's very capable and useful. Like with most frontier models, it has chewed through a bunch of evals so we need to now build more complicated evals to better understand its capabilities.
  • @emollick Ethan Mollick on x
    And then there were three... I got access to the new Anthropic Claude 3 AI a few days ago, so not enough time for a full review, but it was obvious it was GPT-4 class even before they released the testing stats. At the same time, like Gemini Advanced, it doesn't blow GPT-4 away. …
  • @anthropicai @anthropicai on x
    Haiku is the fastest and most cost-effective model on the market for its intelligence category. For the vast majority of workloads, Sonnet is 2x faster than Claude 2 and Claude 2.1, while Opus is about the same speed as past models.
  • @anthropicai @anthropicai on x
    Opus and Sonnet are accessible in our API which is now generally available, enabling developers to start using these models immediately. Sonnet is powering the free experience on https://claude.ai/, with Opus available for Claude Pro subscribers.
  • @anthropicai @anthropicai on x
    Claude 3 offers sophisticated vision capabilities on par with other leading models. The models can process a wide range of visual formats, including photos, charts, graphs and technical diagrams. [video]
  • @anthropicai @anthropicai on x
    Today, we're announcing Claude 3, our next generation of AI models. The three state-of-the-art models—Claude 3 Opus, Claude 3 Sonnet, and Claude 3 Haiku—set new industry benchmarks across reasoning, math, coding, multilingual understanding, and vision. [image]
  • @emollick Ethan Mollick on x
    The GPT-4 benchmark has now been beaten by the two other leading AI companies (even if not by a huge margin). It is very much OpenAI's move.
  • r/LocalLLaMA r on reddit
    Claude3 release