Google rolls out Gemini 2.5 Flash in preview, with support for Canvas for working on text or code; developers can disable thinking or set a token limit for it
Google's Gemini AI may have had a slow start, but it has been anything but in 2025. Barely a week goes by that another model …
The notable addition is not only a preview model but controls over its reasoning behavior alongside Canvas support for text and code. That makes model behavior, latency and output cost more directly configurable within the same Google workflow.
First-order effects
Developers using the Gemini preview can choose whether to invoke thinking and cap its token use, giving them direct control over the trade-off between reasoning depth and resource consumption.
Canvas users gain a Gemini 2.5 Flash option for drafting and coding work, bringing the preview model into an interactive production surface rather than limiting it to an API workflow.
Second-order effects
Applications built on Gemini can tune reasoning per task instead of treating it as a fixed model property, which increases pressure on rival developer platforms to expose similarly granular controls.
Putting model selection and work-surface features together can make Google’s Gemini environment stickier for teams that want to move from generation to text or code editing without changing tools.
Third-order effects
If reasoning controls become standard, competition will increasingly center on configurable efficiency and workflow integration, not just headline model capability.
The pattern points toward AI workspaces that package models, agent-like task execution and editing surfaces together; the eventual differentiator may be how well providers let customers govern quality, speed and usage.
The trend: This is one data point in the consolidation of AI models into configurable, task-oriented work surfaces for developers and knowledge workers.
Gemini 2.5 Flash is here, our first unified reasoning model with thinking budgets. 🔥 It's on the perato frontier and punches above its price and size!! https://developers.googleblog.com/ ...
Introducing Gemini 2.5 Flash 🔥 ⚡️High quality for best cost 🤔Dynamic thinking, with fine-grained thinking budget 🚅Available on AI Studio Blog: https://developers.googleblog.com/ ... AI Studio: https://aistudio.google.com/ ... [image]
Gemini 2.5 Flash just dropped. ⚡ As a hybrid reasoning model, you can control how much it ‘thinks’ depending on your 💰 - making it ideal for tasks like building chat apps, extracting data and more. Try an early version in @Google AI Studio → https://ai.dev/ [image]
Props to Google for including O4-mini in their Flash 2.5 release. A model released *yesterday*, while some companies only compare to their own models. Looking good gemini. [image]
Gemini Flash 2.5 is price-performance ratio next level. Absolutely nuts. Google cooked - again. Makes you wonder what Gemini 3.0 will be capable of. h/t @AnalogPvt [image]
Gemini 2.5 Flash is our fast, cost-efficient thinking model. 🧠 It's our first hybrid reasoning model, so the model adapts the amount that it thinks based on the prompt.
With 2.5 Flash in preview, developers can also control how much the model reasons so people can optimize across quality, cost, and latency. We can't wait to see how you put Gemini 2.5 Flash to work in your apps. https://developers.googleblog.com/ ...
Today, developers can start building with Gemini 2.5 Flash, now in preview in the Gemini API in Google AI Studio and Vertex AI. It's also available to everyone in the @GeminiApp. https://blog.google/...