/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

days · browse · Enter similar · o open

DeepSeek details V3.1 and says it surpasses R1 on key benchmarks and is customized to work with next-gen Chinese-made AI chips, after unveiling it on August 19

DeepSeek unveiled an update to an older model that it says surpasses the seminal R1 on key benchmarks, keeping the Chinese startup …

Bloomberg

Context & Ripple Effects

V3.1 arrived first as a sparse release with a longer context window; this update supplies the missing performance and hardware-positioning details around the August 19 V3.1 launch.

The claim extends DeepSeek’s recent effort to improve R1’s math, programming and logic performance, while shifting attention from model quality alone to where the model can run: its earlier R1 update set that performance arc.

First-order effects

  • DeepSeek can present V3.1 as both an upgrade over R1 on selected benchmarks and a model tailored for next-generation Chinese-made AI chips, giving users a more specific deployment option.
  • Chinese AI-chip makers gain a named model-compatibility claim that can help validate their platforms for customers evaluating inference deployments.

Second-order effects

  • Model providers and chip vendors serving the same market face pressure to show comparable software optimization and benchmark performance, not merely raw hardware specifications.
  • Buyers may place greater weight on tested model-chip combinations when choosing AI infrastructure, favoring suppliers able to support an integrated stack.

Third-order effects

  • If such co-optimization becomes routine, AI competition can increasingly turn on full-stack deployment—models, inference software and locally available chips—rather than frontier-model claims in isolation.
  • That shift could broaden the role of second-source compute in procurement, though DeepSeek’s benchmark and compatibility assertions alone do not establish broad customer adoption.

The trend: This is one data point in the move toward regionally integrated AI stacks, where model advances are paired with optimization for alternative compute supply.

Discussion

  • @segyges SE Gyges on bluesky
    new deepseek.  if it's a top tier coding model as rumored i am probably abandoning western labs for the forseeable future huggingface.co/deepseek-ai/...
  • @deepseek_ai @deepseek_ai on x
    Introducing DeepSeek-V3.1: our first step toward the agent era! 🚀 🧠 Hybrid inference: Think & Non-Think — one model, two modes ⚡️ Faster thinking: DeepSeek-V3.1-Think reaches answers in less time vs. DeepSeek-R1-0528 🛠️ Stronger agent skills: Post-training boosts tool use and
  • @artificialanlys @artificialanlys on x
    DeepSeek launches V3.1, unifying V3 and R1 into a hybrid reasoning model with an incremental increase in intelligence Incremental intelligence increase: Initial benchmarking results for DeepSeek V3.1 show Artificial Analysis Intelligence Index of 60 in reasoning mode, up from [im…
  • @_akhaliq @_akhaliq on x
    DeepSeek-V3.1 ball bouncing inside a spinning hexagon with @FireworksAI_HQ in anycoder, one shot [video]
  • @scaling01 @scaling01 on x
    DeepSeek-V3.1 on par with o3, Opus 4 and Gemini 2.5 Pro Preview on coding It achieves a 76.3% score on Aider Polyglot with Thinking [image]
  • @deepseek_ai @deepseek_ai on x
    Tools & Agents Upgrades 🧰 📈 Better results on SWE / Terminal-Bench 🔍 Stronger multi-step reasoning for complex search tasks ⚡️ Big gains in thinking efficiency 3/5 [image]
  • @scaling01 @scaling01 on x
    DeepSeek V3.1 showing only minor improvements over V3 in the Artificial Intelligence Index [image]
  • @nadzi_mouad Mouad on x
    DeepSeek-V3.1 benchmarks just dropped and... holy efficiency batman 🦇 • 66% on SWE-bench (best open model) • 5.5x faster at terminal tasks than R1 • 3.4x better at web browsing tasks • Still just 37B active params per token This is what happens when you solve hybrid AI [image]
  • @byintes_ Bryan on x
    DeepSeek v3.1 just dropped. Key highlights: - Big improvements on coding, agentic and reasoning vs older Deepseek R1. Open source SOTA - Still slightly behind other closed SOTA models in benchmarks. E.g. 66.0% on SWE-Bench Verified vs GPT-5's 74.9% and Opus 4.1's 74.5% - [image]
  • @deepseek_ai @deepseek_ai on x
    API Update ⚙️ 🔹 deepseek-chat → non-thinking mode 🔹 deepseek-reasoner → thinking mode 🧵 128K context for both 🔌 Anthropic API format supported: https://api-docs.deepseek.com/ ... ✅ Strict Function Calling supported in Beta API: https://api-docs.deepseek.com/ ... 🚀 More API resour…
  • @zoomyzoomm @zoomyzoomm on x
    DeepSeek V3.1 scores higher than GPT-4.5 on coding benchmarks. And costs 180x less. [image]
  • @scaling01 @scaling01 on x
    This is huge. DeepSeek-V3.1 is on par with OpenAI models in terms of reasoning efficiency [image]
  • @deepseek_ai @deepseek_ai on x
    Pricing Changes 💳 🔹 New pricing starts & off-peak discounts end at Sep 5th, 2025, 16:00 (UTC Time) 🔹 Until then, APIs follow current pricing 📝 Pricing page: https://api-docs.deepseek.com/ ... 5/5 [image]
  • @teortaxestex @teortaxestex on x
    And so we know what V3.1 is. Yes, it's an agent. - continued long-context pretrain for 840B tokens - *significant* gains in agentic regimes (I've hopefully accurately aggregated some tables) They responded to GLM&Kimi. ...They didn't announce V4. [image]
  • @deepseek_ai @deepseek_ai on x
    Model Update 🤖 🔹 V3.1 Base: 840B tokens continued pretraining for long context extension on top of V3 🔹 Tokenizer & chat template updated — new tokenizer config: https://huggingface.co/... 🔗 V3.1 Base Open-source weights: https://huggingface.co/... 🔗 V3.1 Open-source weights:
  • @kimmonismus @kimmonismus on x
    DeepSeek v3.1 is a massive upgrade! DeepSeek-V3.1 / DeepSeek-V3-(0324) SWE: 66.1 - 43.4 WE-bench Multilingual 54.5 - 29.3 Terminal-Bench 31.3 - 13.3 Really looking forward to DeepSeek r2! [image]