/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

days · browse · Enter similar · o open

Google releases Gemini 1.5 Flash-8B, a smaller and faster 1.5 Flash variant with a 50% lower price, 2x higher rate limits, and lower latency on small prompts

50% lower price (vs 1.5 Flash)  — 2x higher rate limits (vs 1.5 Flash) … Forums: r/Bard : Gemini 8b out of preview available for production use through api

Google Developers Blog

Discussion

  • @officiallogank Logan Kilpatrick on x
    Say hello to Gemini 1.5 Flash-8B ⚡️, now available for production usage with: - 50% lower price (vs 1.5 Flash) - 2x higher rate limits (vs 1.5 Flash) - lower latency on small prompts (vs 1.5 Flash) https://developers.googleblog.com/ ...
  • @natolambert Nathan Lambert on x
    Probably good for data filtering/classification pipelines that you need to scale up, or at least try.
  • @altryne Alex Volkov on x
    Basically “free” intelligence is already here. “When using Caching, Gemini 1.5 Flash-8B costs $0.01 / million tokens 🤯” New Gemini 1.5 flash 8B is... you can literally use it all day long and not even notice in your bank account [image]
  • @maxwinebach Max Weinbach on x
    Genini 1.5 Flash 8B is insane for a few reasons. 1 million token context window full multimodal 8B parameter Fast af CHEAP And quality isn't amazing vs other models but it's good enough with proper prompting and so so cheap
  • @_arohan_ Rohan Anil on x
    Flash8B General Availability: We originally trained Flash 8B giving it all our algorithmic efficiency improvements to pack as much as possible in a small form factor which then was scaled up to Flash On benchmarks, it is closely matching Flash announced during May at I/O [image]
  • @officiallogank Logan Kilpatrick on x
    When using Caching, Gemini 1.5 Flash-8B costs $0.01 / million tokens 🤯 [image]
  • r/Bard r on reddit
    Gemini 8b out of preview available for production use through api