/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

days · browse · Enter similar · o open

Google prices Gemini 3.6 Flash lower than 3.5 Flash, at $1.50/1M input tokens and $7.50/1M output tokens, and Gemini 3.5 Flash-Lite at $0.30/1M and $2.50/1M

As we wait for 3.5 Pro, Google today announced Gemini 3.6 Flash and 3.5 Flash-Lite, while providing updates on what comes next.

9to5Google Abner Li

Context & Ripple Effects

Google’s Flash line has repeatedly paired smaller or faster models with lower prices, including an earlier Flash-8B variant positioned with a 50% price cut and higher rate limits.

The immediate backdrop is Gemini 3.5 Flash’s $1.50 input and $9 output pricing in May; the higher 3.5 Flash output rate made the new output-price reduction the material change in this release.

First-order effects

  • Google lowers Gemini 3.6 Flash’s output-token price to $7.50 per million from Gemini 3.5 Flash’s $9, while holding the input rate at $1.50 per million.
  • Gemini 3.5 Flash-Lite establishes a separate low-cost tier at $0.30 per million input tokens and $2.50 per million output tokens, giving API buyers a cheaper option for cost-sensitive workloads.

Second-order effects

  • Developers using output-heavy applications can reduce Gemini spend by moving from 3.5 Flash to 3.6 Flash, subject to their own quality and migration testing.
  • The widened Flash/Flash-Lite menu makes model routing more attractive: customers can reserve the higher-priced tier for tasks that need it and direct routine requests to Flash-Lite.

Third-order effects

  • If successive Flash releases keep improving price-performance, API-model competition will increasingly turn on effective inference cost rather than a single published token rate.
  • A more segmented portfolio can make Google’s model platform stickier, but it also raises the importance of clear workload benchmarks and routing tools for customers comparing tiers.

The trend: This is another step in the shift toward tiered, continually repriced AI inference offerings built around matching workloads to cost and latency targets.

Discussion

  • @bindureddy Bindu Reddy on x
    Gemini 5.6 flash has a new low price :) Literally the BEST CHAT model in the world 🚀🚀 [image]
  • @mweinbach Max Weinbach on x
    While I didn't have early access to this one, I'm excited to try it! Big Gemini 3.5 Flash fan, but GPT 5.6 Luna stole me away because it was cheaper and faster Gemini 3.6 Flash seems to maybe bring it back to Gemini!
  • @antigravity @antigravity on x
    Gemini 3.6 Flash is live in Antigravity! ⚡️ Building on 3.5 Flash feedback, it consumes up to 17% fewer output tokens while completing complex workflows in fewer reasoning steps and tool calls. [image]
  • @caseynewton Casey Newton on bluesky
    Frontier labs don't generally announce when they are starting pre-training.  Among other things it serves as a kind of anti-hype for your current-generation models — and Gemini 3.5 Pro has enough anti-hype around it already [embedded post]
  • @synthwavedd Leo on x
    Gemini 3.6 Flash benchmarks are out, and it's... beaten by other models on code tasks, and is only really consistently SoTA on vision and context benchmarks. But hey, 3.1 Pro is now so old 3.6 Flash outperforms it across the board 😭 [image]
  • @arrakis_ai Choi on x
    A lot of people are saying Google is falling behind after Gemini 3.6 Flash. I think they're reading it the wrong way. To me, Google has changed its strategy. Yes, Gemini is behind GPT-5.6 Luna, Grok 4.5, and Claude Sonnet 5 in coding. But it leads in computer use, visual