/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

days · browse · Enter similar · o open

Google unveils Gemini 3 Flash, which it says has Pro-grade reasoning with lower latency, outperforming 2.5 Pro “while being 3x faster at a fraction of the cost”

Following last month's launch of Gemini 3 Pro, Google today announced Gemini 3 Flash for consumers and developers.

9to5Google Abner Li

Context & Ripple Effects

Gemini’s Flash tier has long been positioned as the lighter, cheaper counterpart to its larger models, beginning with Gemini 1.5 Flash’s lower-cost multimodal approach. Google later expanded that segmentation with 2.5 Flash, Pro and Flash-Lite models made generally available together.

This release narrows the practical distinction between the fast tier and the reasoning-focused tier for both consumer and developer use. It establishes the performance-and-latency claim against Gemini 2.5 Pro that later Flash and Flash-Lite updates will build on.

First-order effects

  • Consumers and developers gain a new Gemini option that Google positions for stronger reasoning without the latency profile of a larger Pro model.
  • Google can steer workloads that previously required its higher-end 2.5 Pro model toward Flash when speed and cost are the deciding constraints.

Second-order effects

  • The claim raises the competitive bar on price-performance: rival model providers must increasingly compare not only top-end capability, but also reasoning quality at production latency.
  • For developers, model selection becomes more workload-specific, with fast models becoming more credible for tasks formerly reserved for premium reasoning endpoints.

Third-order effects

  • If this product cadence holds, the premium-model category may become less defined by raw capability and more by the residual workloads that cannot be served economically by optimized fast models.
  • The sequence points to a more deliberately tiered AI API market, where providers use model families to segment quality, latency and cost rather than offer a simple flagship-versus-budget choice.

The trend: Frontier-model vendors are compressing stronger reasoning into lower-latency, lower-cost tiers to widen the set of AI workloads that can run economically at scale.

Discussion

  • @noamshazeer Noam Shazeer on x
    Gemini 3 Flash is live. ⚡️ We've packed Gemini 3's Pro-grade reasoning into a leaner model with Flash-level latency, efficiency, and cost. It's my favorite model to use - the latency feels like a real conversation, with the deep intelligence intact. Available in the API, Gemini
  • @scaling01 @scaling01 on x
    slightly more expensive than Gemini 2.5 Flash which was $0.5/$2.5, but it's a much stronger and also larger model, so that's fine I guess but I honestly feared they would do: $0.75/$4.5 or even $1/$6
  • @scaling01 @scaling01 on x
    Gemini 3 Flash Pricing: $0.5 Input $3.0 Output
  • @patelnamra573 Namra Patel on x
    Gemini 3 flash is launched benchmarks are insane HLE 33% without tool, SWE Bench 78% And its cheaper as well Deepmind on fire 🤯 [image]
  • @joncallahan Jon Callahan on x
    Gemini 3 flash is live with solid vision support as 2.5 flash was known for [image]
  • @kimmonismus @kimmonismus on x
    Gemini 3.0 Flash is an absolutely fantastic release. Consider this: It costs a quarter (1/4) of what Gemini 3.0 Pro costs and achieves similar results to the Pro model in almost all benchmarks, such as HLE and ARC-AGI 2. In other benchmarks, it even outperforms the more [image]
  • @scaling01 @scaling01 on x
    You don't understand Gemini 3 Flash is a bigger deal than Pro [image]
  • @ai_for_success AshutoshShrivastava on x
    Official: Google launched Gemini 3 Flash ⚡⚡⚡ TL;DR: - Cheap, fast and intelligent model. - 3x faster than 2.5 Pro - Outperforms Gemini 2.5 Pro across benchmarks - 90.4% GPQA Diamond, 78% SWE-bench Verified I've been testing this for a while, thanks to the Google DeepMind [video]
  • @swalliec69635 Alex on x
    Gemini 3.0 Flash >Gemini 3.0 Pro [image]
  • @jackwoth98 Jack Wotherspoon on x
    Gemini 3 Flash now available in Gemini CLI ⚡️⚡️⚡️ ❯ npm install -g @google/gemini-cli@latest Gemini 3 Flash is fast, cheap... and really good at coding! 🤯 Available now for paid customers + allow-listed free tier users. Rolling out to rest of free tier soon! Enable “Preview [imag…
  • @ntaylormullen N. Taylor Mullen on x
    ⚡️⚡️⚡️ Gemini 3 Flash + Gemini CLI is now available! Super fast and super smart!! ❯ npm install -g @google/gemini-cli@latest - ✅ Available Now: Paid customers + Allow listed free tier - 📷 Coming Soon: More availability to free tier users Enable “Preview features” in [image]
  • @aibattle_ @aibattle_ on x
    Gemini 3 Flash and Gemini 2.5 Flash pricing comparison [image]
  • @arkitus Ali Eslami on x
    Only 29 days after Pro: Gemini 3 Flash is just as smart (1477 LMSYS Elo), 4x cheaper, and sooo much faster! ⚡ [video]
  • @dynamicwebpaige @dynamicwebpaige on x
    🙌 Gemini 3.0 Flash is finally here! Flash is superior to Gemini 2.5 Pro in speed and performance (I've found it to be even better than Gemini 3.0 Pro, in some use cases! 🤯) - and it's super, super speedy and cost-effective. 👇Try it in @GoogleAIStudio Build today! [video]
  • @addyosmani Addy Osmani on x
    Introducing Gemini 3 Flash! ⚡️⚡️⚡️ Frontier intelligence built for speed at a fraction of the cost. Here's ~4 minutes of demos. [video]
  • @sundarpichai Sundar Pichai on x
    We're back in a Flash ⚡ Gemini 3 Flash is our latest model with frontier intelligence built for lightning speed, and pushing the Pareto Frontier of performance and efficiency. It outperforms 2.5 Pro while being 3x faster at a fraction of the cost. With this release, Gemini 3's [i…
  • @demishassabis Demis Hassabis on x
    For a fast model, Gemini 3 Flash offers incredible performance, allowing us to provide frontier intelligence to everyone globally. Try the ‘fast’ mode from the model picker in the @GeminiApp - it's shockingly speedy AND smart. Best pound-for-pound model out there ⚡️⚡️⚡️ [image]
  • @mweinbach Max Weinbach on x
    I've had early access to Gemini 3 Flash and holy crap is this model great It's nearly as intelligent as Gemini 3 Pro, but it's 1/4th the price and like 2x faster. I've been using it in OpenCode and sure, I need to follow up more, but it's so much faster I don't even care. [image]
  • @osanseviero Omar Sanseviero on x
    Introducing Gemini 3 Flash ⚡️Performance close to Gemini 3 Pro, with great multimodal and tool use quality ⚡️3x faster than Gemini 2.5 Pro, while cheaper and better at most benchmarks ⚡️LMArena score of 1477 (top 3 model) The time to build is now (and yes, there's a free tier)
  • @techikansh @techikansh on x
    Gemini 3.0 Flash‼️ Input : $0.30 Output : $2.50 [image]
  • @omarsar0 Elvis on x
    Damn! Gemini 3 Flash is no joke. Faster and cheaper, while demonstrating remarkable reasoning capabilities. Amazing that we have models of this caliber with multimodal and agentic capabilities. Time to build! Stay tuned for more of my thoughts on this model. [image]
  • @ammaar Ammaar Reshi on x
    Gemini 3 Flash is here! ⚡️ It combines advance coding skills, low latency performance, all at a fraction of the cost. Watch its incredible frontend design and coding capabilities in action below where it generates 3 different interactive UIs (video NOT sped up!) [video]
  • @sam_witteveen Sam Witteveen on x
    Gemini 3 Flash is here! Google just dropped what could become your new daily driver model for AI applications. ⚡ [image]
  • @rseroter Richard Seroter on x
    🚨Let's GO. Gemini 3 Flash is now available🚨 It's fast, cheap, high-performing, and intelligent. Check out these benchmarks, and start using it TODAY in all your favorite @Google surfaces like @antigravity, Gemini CLI, @GoogleAIStudio, the @GeminiApp and more. [image]
  • @kat_kampf Kat Kampf on x
    Gemini 3 Flash is here! Frontier intelligence with the speed and scale of a Flash model ⚡️ [image]
  • @engineerrprompt @engineerrprompt on x
    Gemini 3 Flash is live! Big thanks to @googleaidevs & @GoogleDeepMind for the early access. I've been putting it through its paces (full video in the next post). The Gemini family continues to dominate the Pareto Frontier. The cost-to-performance ratio here is incredible, [image]
  • @melvinjohnsonp Melvin Johnson on x
    Excited to announce our latest Gemini 3 family addition — 3.0 Flash which offers frontier intelligence at scale at a fraction of the cost. ⚡️⚡️⚡️ It pushes the pareto frontier of quality vs cost and speed. https://x.com/... [image]
  • @saboo_shubham_ Shubham Saboo on x
    Gemini 3 Flash is here ⚡️ And the numbers look insane. Have been using it for a while now, it's a really good model. [image]
  • @jeffdean Jeff Dean on x
    Not only is Gemini 3 Flash higher quality than 2.5 Pro it is also 3x faster (based on Artificial Analysis benchmarking) at a fraction of the cost. Here's a side-by-side demo. [video]
  • @fhinkel Franziska Hinkelmann, PhD on x
    Gemini 3 Flash is now available in Gemini CLI ⚡️ ❯ npm install -g @google/gemini-cli@latest Enable “Preview features” in /settings, and go FAST! https://developers.googleblog.com/ ... [image]
  • @jeffdean Jeff Dean on x
    One of the things we strive to do with each new Gemini release is to make the new Flash model as good or better than the previous model's Pro model. Gemini 3 Flash exceeds Gemini 2.5 Pro on nearly every metric, often by very large margins, and almost matches Gemini 3 Pro on most …
  • @rmstein Robby Stein on x
    1/2 Gemini 3 Flash is here ⚡⚡⚡ Starting today, we're rolling it out globally in Search as the default model in AI Mode ( https://www.google.com/ai). 3 Flash brings the sophisticated reasoning capabilities of Gemini 3 to Search, without compromising speed. This means AI Mode [vide…
  • @scaling01 @scaling01 on x
    Gemini 3 Flash Benchmarks [image]
  • @thefox Nick Fox on x
    Gemini 3 Flash is here, rolling out now in AI Mode in Search! ⚡️⚡️⚡️ 3 Flash brings frontier model intelligence to everyone...with the speed and accuracy of Google Search. Excited to ship it globally! Try it now at https://google.ai/. [video]
  • @joshwoodward Josh Woodward on x
    Introducing Gemini 3 Flash. You asked for Pro smarts at Flash speeds. This model is smarter than 2.5 Pro and 3x faster than 3 Pro. Enjoy! ⚡⚡⚡ It's free for everyone in @GeminiApp! [image]
  • @koraykv Koray Kavukcuoglu on x
    Gemini 3 Flash is here. ⚡⚡⚡ Pro-grade reasoning with Flash-level speed and efficiency. It's rolling out today globally as the default model on Gemini App and Search AI Mode. Learn more: https://blog.google/... [video]
  • @officiallogank Logan Kilpatrick on x
    Introducing Gemini 3 Flash, our frontier intelligence model, available at scale for everyone. It excels at coding, tool calling, and is stronger than 2.5 Pro across most metrics!! ⚡️ Available in the API at $0.50 in / 1M tokens and $3.00 out / 1M tokens across. [image]
  • @googledeepmind @googledeepmind on x
    Gemini 3 Flash gives you frontier intelligence at a fraction of the cost. ⚡ Here's how it's built for speed and scale 🧵 [image]
  • @officiallogank Logan Kilpatrick on x
    Gemini 3 Flash punches way above its weight class, surpassing 2.5 Pro on many benchmarks, while being much cheaper, faster, and more token efficient. [image]
  • @sundarpichai Sundar Pichai on x
    We're back in a Flash ⚡ Gemini 3 Flash is our latest model with frontier intelligence built for lightning speed, and pushing the Pareto Frontier of performance and efficiency. It outperforms 2.5 Pro while being 3x faster at a fraction of the cost. With this release, Gemini 3's [i…
  • @jeffdean Jeff Dean on x
    We've pushed out the Pareto frontier of efficiency vs. intelligence again. With Gemini 3 Flash ⚡️, we are seeing reasoning capabilities previously reserved for our largest models, now running at Flash-level latency. This opens up entirely new categories of near real-time [image]
  • @googleai @googleai on x
    We're expanding the Gemini 3 family with the launch of Gemini 3 Flash. This model: — Combines Gemini 3's Pro-grade reasoning with Flash-level latency, efficiency, and cost — Delivers frontier-level performance on PHD-level reasoning and knowledge benchmarks — Is our most [video]
  • r/singularity r on reddit
    Google releases Gemini 3 Flash: Ranks #3 on LMArena (above Opus 4.5), scores 99.7% on AIME and costs $0.50/1M plus Benchmarks.
  • r/Bard r on reddit
    Gemini 3 Flash: frontier intelligence built for speed