Google unveils Gemini 3 Flash, which it says has Pro-grade reasoning with lower latency, outperforming 2.5 Pro “while being 3x faster at a fraction of the cost”
Following last month's launch of Gemini 3 Pro, Google today announced Gemini 3 Flash for consumers and developers.
9to5Google Abner Li
Context & Ripple Effects
Gemini’s Flash tier has long been positioned as the lighter, cheaper counterpart to its larger models, beginning with Gemini 1.5 Flash’s lower-cost multimodal approach. Google later expanded that segmentation with 2.5 Flash, Pro and Flash-Lite models made generally available together.
This release narrows the practical distinction between the fast tier and the reasoning-focused tier for both consumer and developer use. It establishes the performance-and-latency claim against Gemini 2.5 Pro that later Flash and Flash-Lite updates will build on.
First-order effects
- Consumers and developers gain a new Gemini option that Google positions for stronger reasoning without the latency profile of a larger Pro model.
- Google can steer workloads that previously required its higher-end 2.5 Pro model toward Flash when speed and cost are the deciding constraints.
Second-order effects
- The claim raises the competitive bar on price-performance: rival model providers must increasingly compare not only top-end capability, but also reasoning quality at production latency.
- For developers, model selection becomes more workload-specific, with fast models becoming more credible for tasks formerly reserved for premium reasoning endpoints.
Third-order effects
- If this product cadence holds, the premium-model category may become less defined by raw capability and more by the residual workloads that cannot be served economically by optimized fast models.
- The sequence points to a more deliberately tiered AI API market, where providers use model families to segment quality, latency and cost rather than offer a simple flagship-versus-budget choice.
The trend: Frontier-model vendors are compressing stronger reasoning into lower-latency, lower-cost tiers to widen the set of AI workloads that can run economically at scale.
Related: Compute-to-API flywheel · Integrated AI Stack · Gemini · Google makes Gemini 2.5 Flash and Pro generally available · Google launches Gemini 3.1 Flash-Lite
Related Coverage
- You can try Google's new Gemini 3 Flash AI model today for free - it's even in Search's AI Mode ZDNET · Webb Wright
- Google Says Its New Gemini 3 Flash AI Model Is Better and Faster Than 2.5 Pro CNET · Imad Khan
- Google's Gemini 3 Flash model outperforms GPT-5.2 in some benchmarks Engadget · Igor Bonifacic
- Google's Gemini 3 Flash makes a big splash with faster responsiveness and superior reasoning SiliconANGLE · Mike Wheatley
- Gemini 3 Flash comes to the Gemini app The Keyword · Josh Woodward
- Gemini 3 Flash is rolling out globally in Google Search The Keyword · Robby Stein
- Google rolls out Gemini 3 Flash to AI Mode in Search globally Search Engine Land · Danny Goodwin
- Humanity's Last Exam — HLE has been finalized to 2,500 questions. SEAL Leaderboard
- Build with Gemini 3 Flash, frontier intelligence that scales with you The Keyword · Logan Kilpatrick
- Google Releases More Efficient Gemini 3 AI Model Across Products Bloomberg · Carmen Arroyo
- Google's new Gemini 3 Flash is fast, cheap and everywhere Axios · Megan Morrone
- Google releases Gemini 3 Flash, promising improved intelligence and efficiency Ars Technica · Ryan Whitwam
- Google's New Gemini 3 Flash Rivals Frontier Models at a Fraction of the Cost The New Stack · Frederic Lardinois
- Gemini 3 Flash is here, bringing a ‘huge’ upgrade to the Gemini app The Verge · Emma Roth
- Gemini 3 Flash is now available in Gemini CLI Google Developers Blog · Taylor Mullen
- Google reveals Gemini 3 Flash to speed up AI search and beefs up image generation Digital Trends · Nadeem Sarwar
- Gemini 3 Flash: frontier intelligence built for speed Hacker News
Discussion
-
@noamshazeer
Noam Shazeer
on x
Gemini 3 Flash is live. ⚡️ We've packed Gemini 3's Pro-grade reasoning into a leaner model with Flash-level latency, efficiency, and cost. It's my favorite model to use - the latency feels like a real conversation, with the deep intelligence intact. Available in the API, Gemini
-
@scaling01
@scaling01
on x
slightly more expensive than Gemini 2.5 Flash which was $0.5/$2.5, but it's a much stronger and also larger model, so that's fine I guess but I honestly feared they would do: $0.75/$4.5 or even $1/$6
-
@scaling01
@scaling01
on x
Gemini 3 Flash Pricing: $0.5 Input $3.0 Output
-
@patelnamra573
Namra Patel
on x
Gemini 3 flash is launched benchmarks are insane HLE 33% without tool, SWE Bench 78% And its cheaper as well Deepmind on fire 🤯 [image]
-
@joncallahan
Jon Callahan
on x
Gemini 3 flash is live with solid vision support as 2.5 flash was known for [image]
-
@kimmonismus
@kimmonismus
on x
Gemini 3.0 Flash is an absolutely fantastic release. Consider this: It costs a quarter (1/4) of what Gemini 3.0 Pro costs and achieves similar results to the Pro model in almost all benchmarks, such as HLE and ARC-AGI 2. In other benchmarks, it even outperforms the more [image]
-
@scaling01
@scaling01
on x
You don't understand Gemini 3 Flash is a bigger deal than Pro [image]
-
@ai_for_success
AshutoshShrivastava
on x
Official: Google launched Gemini 3 Flash ⚡⚡⚡ TL;DR: - Cheap, fast and intelligent model. - 3x faster than 2.5 Pro - Outperforms Gemini 2.5 Pro across benchmarks - 90.4% GPQA Diamond, 78% SWE-bench Verified I've been testing this for a while, thanks to the Google DeepMind [video]
-
@swalliec69635
Alex
on x
Gemini 3.0 Flash >Gemini 3.0 Pro [image]
-
@jackwoth98
Jack Wotherspoon
on x
Gemini 3 Flash now available in Gemini CLI ⚡️⚡️⚡️ ❯ npm install -g @google/gemini-cli@latest Gemini 3 Flash is fast, cheap... and really good at coding! 🤯 Available now for paid customers + allow-listed free tier users. Rolling out to rest of free tier soon! Enable “Preview [imag…
-
@ntaylormullen
N. Taylor Mullen
on x
⚡️⚡️⚡️ Gemini 3 Flash + Gemini CLI is now available! Super fast and super smart!! ❯ npm install -g @google/gemini-cli@latest - ✅ Available Now: Paid customers + Allow listed free tier - 📷 Coming Soon: More availability to free tier users Enable “Preview features” in [image]
-
@aibattle_
@aibattle_
on x
Gemini 3 Flash and Gemini 2.5 Flash pricing comparison [image]
-
@arkitus
Ali Eslami
on x
Only 29 days after Pro: Gemini 3 Flash is just as smart (1477 LMSYS Elo), 4x cheaper, and sooo much faster! ⚡ [video]
-
@dynamicwebpaige
@dynamicwebpaige
on x
🙌 Gemini 3.0 Flash is finally here! Flash is superior to Gemini 2.5 Pro in speed and performance (I've found it to be even better than Gemini 3.0 Pro, in some use cases! 🤯) - and it's super, super speedy and cost-effective. 👇Try it in @GoogleAIStudio Build today! [video]
-
@addyosmani
Addy Osmani
on x
Introducing Gemini 3 Flash! ⚡️⚡️⚡️ Frontier intelligence built for speed at a fraction of the cost. Here's ~4 minutes of demos. [video]
-
@sundarpichai
Sundar Pichai
on x
We're back in a Flash ⚡ Gemini 3 Flash is our latest model with frontier intelligence built for lightning speed, and pushing the Pareto Frontier of performance and efficiency. It outperforms 2.5 Pro while being 3x faster at a fraction of the cost. With this release, Gemini 3's [i…
-
@demishassabis
Demis Hassabis
on x
For a fast model, Gemini 3 Flash offers incredible performance, allowing us to provide frontier intelligence to everyone globally. Try the ‘fast’ mode from the model picker in the @GeminiApp - it's shockingly speedy AND smart. Best pound-for-pound model out there ⚡️⚡️⚡️ [image]
-
@mweinbach
Max Weinbach
on x
I've had early access to Gemini 3 Flash and holy crap is this model great It's nearly as intelligent as Gemini 3 Pro, but it's 1/4th the price and like 2x faster. I've been using it in OpenCode and sure, I need to follow up more, but it's so much faster I don't even care. [image]
-
@osanseviero
Omar Sanseviero
on x
Introducing Gemini 3 Flash ⚡️Performance close to Gemini 3 Pro, with great multimodal and tool use quality ⚡️3x faster than Gemini 2.5 Pro, while cheaper and better at most benchmarks ⚡️LMArena score of 1477 (top 3 model) The time to build is now (and yes, there's a free tier)
-
@techikansh
@techikansh
on x
Gemini 3.0 Flash‼️ Input : $0.30 Output : $2.50 [image]
-
@omarsar0
Elvis
on x
Damn! Gemini 3 Flash is no joke. Faster and cheaper, while demonstrating remarkable reasoning capabilities. Amazing that we have models of this caliber with multimodal and agentic capabilities. Time to build! Stay tuned for more of my thoughts on this model. [image]
-
@ammaar
Ammaar Reshi
on x
Gemini 3 Flash is here! ⚡️ It combines advance coding skills, low latency performance, all at a fraction of the cost. Watch its incredible frontend design and coding capabilities in action below where it generates 3 different interactive UIs (video NOT sped up!) [video]
-
@sam_witteveen
Sam Witteveen
on x
Gemini 3 Flash is here! Google just dropped what could become your new daily driver model for AI applications. ⚡ [image]
-
@rseroter
Richard Seroter
on x
🚨Let's GO. Gemini 3 Flash is now available🚨 It's fast, cheap, high-performing, and intelligent. Check out these benchmarks, and start using it TODAY in all your favorite @Google surfaces like @antigravity, Gemini CLI, @GoogleAIStudio, the @GeminiApp and more. [image]
-
@kat_kampf
Kat Kampf
on x
Gemini 3 Flash is here! Frontier intelligence with the speed and scale of a Flash model ⚡️ [image]
-
@engineerrprompt
@engineerrprompt
on x
Gemini 3 Flash is live! Big thanks to @googleaidevs & @GoogleDeepMind for the early access. I've been putting it through its paces (full video in the next post). The Gemini family continues to dominate the Pareto Frontier. The cost-to-performance ratio here is incredible, [image]
-
@melvinjohnsonp
Melvin Johnson
on x
Excited to announce our latest Gemini 3 family addition — 3.0 Flash which offers frontier intelligence at scale at a fraction of the cost. ⚡️⚡️⚡️ It pushes the pareto frontier of quality vs cost and speed. https://x.com/... [image]
-
@saboo_shubham_
Shubham Saboo
on x
Gemini 3 Flash is here ⚡️ And the numbers look insane. Have been using it for a while now, it's a really good model. [image]
-
@jeffdean
Jeff Dean
on x
Not only is Gemini 3 Flash higher quality than 2.5 Pro it is also 3x faster (based on Artificial Analysis benchmarking) at a fraction of the cost. Here's a side-by-side demo. [video]
-
@fhinkel
Franziska Hinkelmann, PhD
on x
Gemini 3 Flash is now available in Gemini CLI ⚡️ ❯ npm install -g @google/gemini-cli@latest Enable “Preview features” in /settings, and go FAST! https://developers.googleblog.com/ ... [image]
-
@jeffdean
Jeff Dean
on x
One of the things we strive to do with each new Gemini release is to make the new Flash model as good or better than the previous model's Pro model. Gemini 3 Flash exceeds Gemini 2.5 Pro on nearly every metric, often by very large margins, and almost matches Gemini 3 Pro on most …
-
@rmstein
Robby Stein
on x
1/2 Gemini 3 Flash is here ⚡⚡⚡ Starting today, we're rolling it out globally in Search as the default model in AI Mode ( https://www.google.com/ai). 3 Flash brings the sophisticated reasoning capabilities of Gemini 3 to Search, without compromising speed. This means AI Mode [vide…
-
@scaling01
@scaling01
on x
Gemini 3 Flash Benchmarks [image]
-
@thefox
Nick Fox
on x
Gemini 3 Flash is here, rolling out now in AI Mode in Search! ⚡️⚡️⚡️ 3 Flash brings frontier model intelligence to everyone...with the speed and accuracy of Google Search. Excited to ship it globally! Try it now at https://google.ai/. [video]
-
@joshwoodward
Josh Woodward
on x
Introducing Gemini 3 Flash. You asked for Pro smarts at Flash speeds. This model is smarter than 2.5 Pro and 3x faster than 3 Pro. Enjoy! ⚡⚡⚡ It's free for everyone in @GeminiApp! [image]
-
@koraykv
Koray Kavukcuoglu
on x
Gemini 3 Flash is here. ⚡⚡⚡ Pro-grade reasoning with Flash-level speed and efficiency. It's rolling out today globally as the default model on Gemini App and Search AI Mode. Learn more: https://blog.google/... [video]
-
@officiallogank
Logan Kilpatrick
on x
Introducing Gemini 3 Flash, our frontier intelligence model, available at scale for everyone. It excels at coding, tool calling, and is stronger than 2.5 Pro across most metrics!! ⚡️ Available in the API at $0.50 in / 1M tokens and $3.00 out / 1M tokens across. [image]
-
@googledeepmind
@googledeepmind
on x
Gemini 3 Flash gives you frontier intelligence at a fraction of the cost. ⚡ Here's how it's built for speed and scale 🧵 [image]
-
@officiallogank
Logan Kilpatrick
on x
Gemini 3 Flash punches way above its weight class, surpassing 2.5 Pro on many benchmarks, while being much cheaper, faster, and more token efficient. [image]
-
@sundarpichai
Sundar Pichai
on x
We're back in a Flash ⚡ Gemini 3 Flash is our latest model with frontier intelligence built for lightning speed, and pushing the Pareto Frontier of performance and efficiency. It outperforms 2.5 Pro while being 3x faster at a fraction of the cost. With this release, Gemini 3's [i…
-
@jeffdean
Jeff Dean
on x
We've pushed out the Pareto frontier of efficiency vs. intelligence again. With Gemini 3 Flash ⚡️, we are seeing reasoning capabilities previously reserved for our largest models, now running at Flash-level latency. This opens up entirely new categories of near real-time [image]
-
@googleai
@googleai
on x
We're expanding the Gemini 3 family with the launch of Gemini 3 Flash. This model: — Combines Gemini 3's Pro-grade reasoning with Flash-level latency, efficiency, and cost — Delivers frontier-level performance on PHD-level reasoning and knowledge benchmarks — Is our most [video]
-
r/singularity
r
on reddit
Google releases Gemini 3 Flash: Ranks #3 on LMArena (above Opus 4.5), scores 99.7% on AIME and costs $0.50/1M plus Benchmarks.
-
r/Bard
r
on reddit
Gemini 3 Flash: frontier intelligence built for speed