Google releases Gemini 1.5 Flash-8B, a smaller and faster 1.5 Flash variant with a 50% lower price, 2x higher rate limits, and lower latency on small prompts
50% lower price (vs 1.5 Flash) — 2x higher rate limits (vs 1.5 Flash) … Forums: r/Bard : Gemini 8b out of preview available for production use through api
Google Developers Blog
Related Coverage
- Google launches Gemini 1.5 Flash-8B with enhanced performance and cost efficiency Financial Express
- Google democratizes AI with Gemini 1.5 Flash-8B, the cheapest Gemini model to date Neowin
- Google's Gemini 1.5 Flash-8B Hits Production: High Volume, Low Cost AI WinBuzzer
- Google's lightweight Gemini 1.5 Flash-8B hits general availability SiliconANGLE
- Gemini 1.5 Flash-8B is now production ready (via) Gemini 1.5 Flash-8B is “a smaller … Simon Willison's Weblog
- If you thought that Gemini 1.5 Flash was fast and cost-efficient, you haven't seen *anything* yet: say hello to Gemini 1.5 Flash-8B! 🙌 … Paige Bailey
- Say hello to Gemini 1.5 Flash-8B ⚡ , now available for production usage with: — 50% lower price (vs 1.5 Flash) — 2x higher rate limits (vs 1.5 Flash) … Logan Kilpatrick
Discussion
-
@officiallogank
Logan Kilpatrick
on x
Say hello to Gemini 1.5 Flash-8B ⚡️, now available for production usage with: - 50% lower price (vs 1.5 Flash) - 2x higher rate limits (vs 1.5 Flash) - lower latency on small prompts (vs 1.5 Flash) https://developers.googleblog.com/ ...
-
@natolambert
Nathan Lambert
on x
Probably good for data filtering/classification pipelines that you need to scale up, or at least try.
-
@altryne
Alex Volkov
on x
Basically “free” intelligence is already here. “When using Caching, Gemini 1.5 Flash-8B costs $0.01 / million tokens 🤯” New Gemini 1.5 flash 8B is... you can literally use it all day long and not even notice in your bank account [image]
-
@maxwinebach
Max Weinbach
on x
Genini 1.5 Flash 8B is insane for a few reasons. 1 million token context window full multimodal 8B parameter Fast af CHEAP And quality isn't amazing vs other models but it's good enough with proper prompting and so so cheap
-
@_arohan_
Rohan Anil
on x
Flash8B General Availability: We originally trained Flash 8B giving it all our algorithmic efficiency improvements to pack as much as possible in a small form factor which then was scaled up to Flash On benchmarks, it is closely matching Flash announced during May at I/O [image]
-
@officiallogank
Logan Kilpatrick
on x
When using Caching, Gemini 1.5 Flash-8B costs $0.01 / million tokens 🤯 [image]
-
r/Bard
r
on reddit
Gemini 8b out of preview available for production use through api