Google releases Gemini 2.0 Flash via its API, an experimental Gemini 2.0 Pro version via its apps, Gemini 2.0 Flash Thinking, and 2.0 Flash-Lite in AI Studio
Gemini 2.0 AI updates include cheaper access for developers, and AI that can use other Google apps like YouTube.
The VergeJess Weatherbed
Context & Ripple Effects
This expands the Gemini 2.0 rollout that had already put Flash on Google’s developer surfaces with multimodal generation and third-party-service use. The new lineup separates a broadly available API model from experimental app access and a lower-cost option, while extending Gemini’s reach into Google-owned services such as YouTube.
Developers can choose among Gemini 2.0 variants with different cost and reasoning trade-offs, including lower-cost Flash-Lite access in AI Studio, rather than treating Flash as a single offering.
Google app users gain experimental access to Gemini 2.0 Pro and a model able to draw on other Google apps, making Gemini more directly embedded in Google’s own product surface.
Second-order effects
Model selection becomes a practical application-design decision: developers can reserve higher-capability or thinking-oriented variants for harder tasks and use Lite for cost-sensitive workloads.
Rival AI platforms face pressure to match both lower-priced API tiers and tighter links between assistants and their existing application ecosystems.
Third-order effects
Google is moving toward a portfolio model in which speed, cost, reasoning behavior and distribution channel are separate product levers—not a single flagship-model contest.
If this tiering persists, AI application economics will increasingly depend on routing work across model classes and on access to the surrounding workflow surfaces, favoring providers that control both APIs and widely used apps.
The trend: Frontier AI providers are turning model releases into tiered, workflow-connected platforms that compete on deployment cost and distribution as much as raw capability.
3/ One more thing - our latest experimental 2.0 Flash Thinking model with advanced reasoning is available to all @Geminiapp users. Look for it in the drop down, including a version that connects to apps like YouTube, Search + Maps. More details on everything we're announcing
2/ We're introducing a new model Flash-Lite, which is extremely efficient and capable, in public preview. Plus an experimental version of Gemini 2.0 Pro, our best model for coding performance and complex prompts, is now available. Gemini Advanced users can try it out in the [imag…
⚡️Gemini 2.0 Flash is now GA - Significant performance improvements over 1.5 Pro and Flash - Simplified pricing - now the price for < 128K and > 128K tokens is the same 🔦 We are also showcasing Flash-Lite Preview and Pro Experimental
Google's full release of Gemini 2.0 Flash is a great thing for the voice AI ecosystem. Up to this point, almost every production voice AI agent has used GPT-4o. Voice AI apps need an LLM with fast TTFT, good instruction following, reliable function calling, and natural
The effect of Open Source AI is that it forces massive corporations to keep up. NOW it's ok to show a model's reasoning after DeepSeek R1 championed it. “OpenAI” and now Google will show reasoning. In good models the Reasoning is better than the answer https://blog.google/...
Announcing updates to the Gemini model family: -Gemini 2.0 Flash is now generally available -Gemini 2.0 Pro Experimental is available in AI Studio and Vertex AI -Gemini 2.0 Flash-Lite, an efficient, cost-effective model now in public preview Dive in ↓ https://developers.googleblo…
Great news for developers - Gemini 2.0 Flash is now generally available and we debuted Gemini 2.0 Flash-Lite, our most cost-efficient model yet. https://blog.google/...
The Gemini 2.0 Flash pricing is crazy It's cheaper than GPT-4o mini in every tier while being as intelligent as ChatGPT-4o (which is actually more expensive than the numbers in the image) And there's an even cheaper tier that's 1/2 GPT-4o-mini and about equal [image]
Gemini's just launched their new Flash models. They are cheaper, better and have 8x the context of GPT 4o-mini! Per million input (cached), input and output tokens: Gemini 2 Flash Lite: $0.01875, $0.075, $0.30 Gemini 2 Flash: $0.025, $0.1, $0.40 GPT 4o-mini: $0.075, $0.15, $0.60 …
We're making Gemini 2.0 Flash available to all Perplexity Pro users. This is the first time we're bringing a Gemini model to Perplexity. Flash 2.0 is an incredible multimodal cost-efficient model. Update the app and find it in settings or “rewrite” option at the end of an answer.…
Next week is going to be massive, we might get: - o3-mini - Gemini 2.0 Pro: Probably new experimental model, this will be huge. - Gemini 2.0 Flash - - - - : This will be a new model with a new name Try to fill in the blanks 😉 (I know the name, but will wait for the leak to
NEW from @GoogleAI Gemini 2 pro is official, Gemini 2.0 flash in in GA and there's a new model! Flash-lite (in preview) Pricing / Value is crazy for FlashLite - just 7 cents per 1M tokens, while performing better than Flash 1.5 (at the same speeds, 1M context multimodal) 🥵 [image…
Today is the day I've been looking forward to for almost a year now... Say hello to Gemini 2.0 Flash, Gemini 2.0 Flash-Lite, and Gemini 2.0 Pro, our strongest lineup of models ever, available to all developers. 🧵 https://developers.googleblog.com/ ...
🚨 Breaking : Google DeepMind has launched three new model today. - Gemini 2.0 Flash - Gemini 2.0 Flash-Lite Preview - Gemini 2.0 Pro Experimental Few day back I told you guys about new model coming and it was Lite preview model . I don't lie 😁 [image]
We're also excited to reveal 2.0 Flash-Lite. ⚡ It has better quality than 1.5 Flash, at similar cost and speed, and comes with: 🚀 A 1 million token context window 🖼️ Multimodal input 📝 Text output Now available in @Google AI Studio and @GoogleCloud's #VertexAI platform in [image]
2.0 Pro Experimental is our best model yet for coding and complex prompts, refined with your feedback. 🤝 It has a better understanding of world-knowledge and comes with our largest context window yet of 2 million tokens - meaning it can analyze large amounts of information. [vide…
Gemini 2.0 is now available to everyone. ✨ ⚡ Start using an updated 2.0 Flash in @Google AI Studio, @GoogleCloud's #VertexAI and in @GeminiApp. We're also introducing: 🔵 2.0 Pro Experimental, which excels at coding. 🔵 2.0 Flash-Lite, our most cost-efficient model yet. 🔵 [image]
“So today, we're introducing 2.0 Flash-Lite, a new model that has better quality than 1.5 Flash, at the same speed and cost” I think that's 7.5c/million input tokens, 30c/million output tokens - half the price of OpenAI's GPT-4o mini
2.0 Flash is optimized for high-volume, high-frequency tasks at scale - making it our highly efficient workhorse model for developers. Our latest version shows improved performance in key benchmarks. It's now available for more people across @Google AI products → [image]
2.0 Flash-Lite, a new model, outperforms Flash 1.5 at a similar cost and speed, with a 1 million token context window and multimodal input. 🔦 It's now available in Google AI Studio & Vertex AI in public preview. Learn more → https://developers.googleblog.com/ ...
And 2.0 Pro Experimental is our best model yet for coding and complex prompts, with better understanding and reasoning of world-knowledge. It has our largest context window at 2 million tokens, enabling it to analyze and understand large amounts of info. 🤯 [image]
First up, an updated 2.0 Flash will be generally available to more people across our AI products, including: ✨ Gemini API in Google AI Studio & Vertex AI ✨ Both 2.0 Flash Thinking Experimental models and 2.0 Pro Experimental are rolling out to the Gemini app [image]
Gemini 2.0 Pro Experimental is available now to devs in Google AI Studio & Vertex AI, and to Gemini Advanced users on desktop and mobile. Learn more about all our latest 2.0 models ↓ https://blog.google/...
Today we're expanding the Gemini 2.0 family with new options and broader availability. This builds on the first model we launched in December: 2.0 Flash, our model with low latency and better performance ⚡ Read more on today's launches ⬇️ [image]