/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

days · browse · Enter similar · o open

Anthropic's Claude 3.7 Sonnet reportedly cost “a few tens of millions of dollars” to train, similar to Claude 3.5 and cheaper than GPT-4, which cost over $100M

“Assuming Claude 3.7 Sonnet indeed cost just ‘a few tens of millions of dollars’ to train, not factoring in related expenses, it's a sign of how relatively cheap it's becoming to release state-of-the-art models.”  —  https://techcrunch.com/... X: Ethan Mollick / @emollick : After publishing the post, I was contacted by Anthropic who told me that Sonnet 3.7 would not be considered a 10^26 FLOP model and cost a few tens of millions of dollars, though future models will be much bigger. I updated the post to reflect this, though it doesn't change much. Piotr Cieluchowski / @thisguyoftheai : So, it turns out training the latest AI wonder from Anthropic was a budget affair—just a “few tens of millions” and a sprinkle of computing power. Who knew revolutionary tech could be so affordable? Dive into the genius of Claude 3.7 Sonnet here: https://techcrunch.com/... Peter Wildeford / @peterwildeford : Pretty interesting revelations here: - Claude 3.7 is much smaller model size than Grok yet better? - Trained for “few tens of millions of dollars” (single model training run size, this would be apples-to-apples comparison with DeepSeek's $6M) - Future models will be much bigger Tim Crawford / @tcrawford : Lowering the cost will speed up innovation and potentially open the field to new entrants. Anthropic's latest flagship AI might not have been incredibly costly to train. https://techcrunch.com/... #CIO #AI

TechCrunch Kyle Wiggers

Discussion

  • @emollick Ethan Mollick on x
    After publishing the post, I was contacted by Anthropic who told me that Sonnet 3.7 would not be considered a 10^26 FLOP model and cost a few tens of millions of dollars, though future models will be much bigger. I updated the post to reflect this, though it doesn't change much.
  • @thisguyoftheai Piotr Cieluchowski on x
    So, it turns out training the latest AI wonder from Anthropic was a budget affair—just a “few tens of millions” and a sprinkle of computing power. Who knew revolutionary tech could be so affordable? Dive into the genius of Claude 3.7 Sonnet here: https://techcrunch.com/...
  • @peterwildeford Peter Wildeford on x
    Pretty interesting revelations here: - Claude 3.7 is much smaller model size than Grok yet better? - Trained for “few tens of millions of dollars” (single model training run size, this would be apples-to-apples comparison with DeepSeek's $6M) - Future models will be much bigger
  • @tcrawford Tim Crawford on x
    Lowering the cost will speed up innovation and potentially open the field to new entrants. Anthropic's latest flagship AI might not have been incredibly costly to train. https://techcrunch.com/... #CIO #AI