DeepSeek rolls out the official V4 Flash API in public beta, touting enhanced agent capabilities and benchmark scores “far surpassing” V4 Pro Preview
Bloomberg Newley Purnell
Context & Ripple Effects
DeepSeek introduced V4 Pro and V4 Flash as previews in April, while characterizing V4 Pro as still behind the state of the art. The public-beta API is the next commercialization step for that V4 preview lineup.
The release also follows DeepSeek's work on DSpark inference acceleration for V4 models, which the company said was designed to reduce inference latency. Together, the updates make agent-oriented performance a more practical product question rather than solely a model-preview claim.
First-order effects
- Developers can now access V4 Flash through an official public-beta API, with DeepSeek positioning it for stronger agent use cases.
- DeepSeek is using claimed benchmark gains over V4 Pro Preview to differentiate Flash from its earlier preview offering; those claims will now be easier for API users to test in workloads.
Second-order effects
- Customers evaluating V4 Pro Preview may reassess which V4 variant to build on, especially for agent workflows where tool use and response speed matter.
- The beta puts pressure on competing model API providers to show comparable agent performance in deployable products, not just in model announcements.
Third-order effects
- If successive model releases move quickly from preview to public APIs, competition will increasingly center on the reliability, evaluation and operational control of agent-facing services.
- Broader agent deployment raises the importance of access and usage governance as model providers expose more capable systems to third-party developers.
The trend: Frontier-model vendors are turning benchmark-led model iterations into public API products aimed at powering embedded AI agents.
Related: Embedded AI agents · Frontier-model access governance · DeepSeek · DeepSeek V4 Pro and V4 Flash preview · DeepSeek's DSpark framework for V4 models
Related Coverage
- Change Log — Date: 2026-07-31 DeepSeek-V4-Flash Update The official release of the DeepSeek-V4-Flash API is now in public beta. DeepSeek API Docs
- DeepSeek releases beta version of V4 models as AI price war heats up Nikkei Asia · Cissy Zhou
- DeepSeek-V4-Flash Update — The official release of the DeepSeek-V4-Flash API is now in public beta. DeepSeek API Docs
- DeepSeek updates V4-Flash API for coding agents without changing architecture RuntimeWire · Ryan Merket
- DeepSeek-V4-Flash Update Hacker News
- I put ChatGPT Work up against Claude Cowork, and the winner wasn't who I expected XDA Developers · Mahnoor Faisal
- Use this model — DeepSeek-V4-Flash-0731 … DeepSeek-V4-Flash-0731 outperforms DeepSeek-V4-Pro … Hugging Face
- New Deepseek Flash model matches OpenAI's GPT-5.6 Luna at roughly 60 percent lower cost The Decoder · Thomas Joos
- DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis Hacker News
- Hacker uses DeepSeek AI to autonomously attack vulnerable servers BleepingComputer · Lawrence Abrams
- DeepSeek Releases Official V4-Flash Model as China's AI Race Intensifies Caixin Global · Yang Zirui
- 🐋 DeepSeek Answered OpenAI's Price Cut Overnight Superintelligence
- DeepSeek-V4-Flash-0731 and Inkling Small: Smaller, but Better? The Kaitchup · Benjamin Marie
- DeepSeek Upgrades DeepSeek-V4-Flash-0731 with Major Agentic and Coding Gains MarkTechPost · Asif Razzaq
Analysis
Discussion
-
r/singularity
r
on reddit
There's a new Deepseek v4 flash in town!
-
@deepseek_ai
@deepseek_ai
on x
⚠️ Note 🔷 DeepSeek-V4-Flash-0731 keeps the exact same model architecture and size as the preview version. 🔷 Today's upgrade applies ONLY to the DeepSeek-V4-Flash API. The DeepSeek-V4-Pro API and App/Web models remain unchanged for now. The official release of DeepSeek-V4-Pro is c…
-
@ns123abc
Nik
on x
Sir... DeepSeek just price mogged us AGAIN [image]
-
@weiliu99
Wei Liu
on x
Enjoy the latest V4Flash: fast, smart, and agentic.
-
@unslothai
@unslothai
on x
@deepseek_ai If DeepSeek-V4-Flash is this good and this small, imagine DeepSeek-V4-Pro! 🤯 And imagine running Flash locally on your own device! 🔥 We can't wait to make quants so everyone can run it locally!!!
-
@aisearchio
@aisearchio
on x
How is this even possible?! Deepseek drops V4 Flash which is as good as Opus 4.8 but 100x cheaper and way smaller.
-
@jun_song
Jun Song
on x
Frontier labs are dead now.
-
@teortaxestex
@teortaxestex
on x
> these scores > our upcoming DeepSeek Harness (minimal mode) Who was saying that DeepSeek doesn't understand the importance of agents? Huh? Huh? I don't hear you! [image]
-
@andrewzeng17
Weihao Zeng
on x
Try our model on general agent tasks —making intelligence more accessible to everyone 🥳🥳🥳
-
@tech2wild
@tech2wild
on x
It's literally beating 700B-1TB Parameter Models in all benchmarks lol. GLM/CHATGPT 5.2 Quality on 2 x DGX Sparks at 70 tok/s.... This is a Qwen 3.5 27B Moment...
-
@davidondrej1
David Ondrej
on x
what the fuck
-
@apples_jimmy
@apples_jimmy
on x
I think chinas going to ignore that petition
-
@miaai_lab
Mia
on x
Can't wait for the release of the weights! @deepseek_ai
-
@zrdianjiao
@zrdianjiao
on x
Same architecture, same size as the preview — Terminal Bench 61.8 → 82.7, DeepSWE 7.3 → 54.4. All of it from post-training. Very impressive. Congrats! 👏
-
@kimmonismus
@kimmonismus
on x
This is insane! not kidding at all. DeepSeek v4 *Flash* super close to Opus 4.8, Insane upgrade …
-
@zephyr_z9
@zephyr_z9
on x
Big whale drop
-
@deepseek_ai
@deepseek_ai
on x
🚀 DeepSeek-V4-Flash Official API is now LIVE in public beta! …
-
@mrahmadawais
Ahmad Awais
on x
@deepseek_ai New DeepSeek V4 Flash is already live in @CommandCodeAI 10x usage $1 Go plan $10 in free auto credits. Running internal benchmarks now. Hard to believe how much better it is. With tool and cache repairs in our harness it's gonna outperform a lot of models.
-
@jrw14.whnc.me
Joshua White
on bluesky
Uh. This is the biggest news in AI today, and it won't be particularly close. Deepseek's FLASH model, 200ish billion parameters, is obliterating models 3-5x it's size and costs pennies. [embedded post]
-
r/DeepSeek
r
on reddit
DeepSeek-V4-Flash API is now in public beta
-
r/DeepSeek
r
on reddit
DeepSeek-V4-Flash Update — The official release of the DeepSeek-V4-Flash API is now in public beta.
-
r/singularity
r
on reddit
DeepSeek-V4-Flash Official API is now LIVE in public beta! Massive upgrades for flash model.
-
r/accelerate
r
on reddit
DeepSeek V4 Flash Update
-
r/LocalLLaMA
r
on reddit
DeepSeek-V4-Flash-0731 is going to cause another market crash.
-
r/LocalLLaMA
r
on reddit
DeepSeek-V4-Flash has been updated, “The official release of DeepSeek-V4-Pro will follow soon”
-
r/LocalLLaMA
r
on reddit
The official release Deepseek V4 flash is live on the API
-
@jason
@jason
on x
It's happening Default will be local model by next year for most corporate workers
-
@emostaque
Emad
on x
As shocking as the Kimi K3 release. Massive performance gain was just with post-training Model is 3x smaller than GLM 5.2 (10x smaller than K3) & works on a MacBook / Spark This is Q1 flagship (Opus 4.6/GPT 5.4) level for < $0.28/m tokens (100x cheaper)
-
@artificialanlys
@artificialanlys
on x
DeepSeek V4 Flash 0731 scores 50 on the Artificial Analysis Intelligence Index …
-
@jason
@jason
on x
Open Source is winning bigly Tokens will be 100x cheaper in 24 months You're going to run 50% of tokens on your local hardware unmetered @Dell, @nvidia and @Apple are the big winners in this trend
-
@theahmadosman
Ahmad
on x
DeepSeek V4 Flash is ~70% smaller in size than GLM 5.2 It also beats GLM 5.2 which was the SoTA model just about a month ago In this video I explained how we will have GLM 5.2 level intelligence on a single RTX 5090 in < 18 months Looks like it's happening way sooner than that [i…
-
@haider1
Haider
on x
what DeepSeek managed with v4-flash is completely insane roughly 300B parameters, with the same architecture and model size as the v4-flash-preview, yet Terminal-Bench jumped from 61.8 to 82.7 and DeepSwe from 7.3 to 54.4 all of that at just $0.18/m output tokens i still don't un…
-
@bijanbowen
@bijanbowen
on x
Either Deepseek flash is benchmaxxed on my specific prompts or it's an absolute freak of nature
-
@cline
@cline
on x
DeepSeek silently updated their changelog with a new V4-Flash upgrade 1 hour ago. Their new Terminal-Bench score is 82.7, a massive +25.8 point leap from its initial April preview score of 56.9. Currently only available via their API, open weights release will follow shortly. [im…
-
@arena
@arena
on x
DeepSeek-V4-Flash by @deepseek_ai is in the Arena! In Agent Arena, we measure models on millions …
-
@teortaxestex
@teortaxestex
on x
DeepSeek-V4-Flash is updated. «DSBench-FullStack is an internal full-stack development test set, and DSBench-Hard is an internal Coding Agent hard-problem test set» «The official release of DeepSeek-V4-Pro will follow soon» These scores are... quite a lot better than Preview. [im…
-
@tech2wild
@tech2wild
on x
As Much As I LOVE GLM 5.2, you have to weigh the odds here. You go from 35 tok/s to 70 tok/s. 300k Ctx to 1M ctx. 4 Sparks to 2 Sparks. The Comparisons are crazy. I am loading this up NOW REPO updates coming soon !
-
@elshayib_
@elshayib_
on x
If the flash model is beating GLM 5.2 then the pro version is gonna be up there competing with Fable 5 and GPT-5.6 and of course at a fraction of the price and also a much smaller model so running it locally is easier.
-
@kimmonismus
@kimmonismus
on x
This is just insane. For me the flash release is another DeepSeek moment. Intelligence so cheap and so good that it's literally too cheap to meter. Many don't get how insane this release is. [image]
-
r/LocalLLaMA
r
on reddit
deepseek-ai/DeepSeek-V4-Flash- 0731 on Huggingface
-
@davidondrej1
David Ondrej
on x
makes you realize how trash Sonnet 5 is... lol [image]
-
@yuchenj_uw
Yuchen Jin
on x
Insane Terminal-Bench 2.1 result. 1. small models like DeepSeek V4-Flash are powerful. 2. open-source models are forcing intelligence too cheap to meter. [image]
-
@shreymodi13
Shrey Modi
on x
the updated model is 100x cheaper than opus 4.8 with similar capabilities! this is absolutely insane. waiting for the official Deepseek V4-Pro launch, fable level accuracy at 1/100th the cost?
-
@michaeljburry
Cassandra Unchained
on x
The real news is OpenAI slashing and burning its prices to prepare for this. #DeepSeek releases beta version of V4 models as AI price war heats up. https://asia.nikkei.com/...
-
r/singularity
r
on reddit
Weights of Deepseek v4 flash 0731 have been released!!!