/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

days · browse · Enter similar · o open

Google announces Gemini 1.5 Flash, which is more lightweight and cheaper than Gemini Pro, but has the same multimodal capabilities and 1M-token context window

Engadget Pranav Dixit

Discussion

  • @michaelsayman Michael Sayman on threads
    Huge.  Gemini Pro 1.5 Flash will start at $0.35 / 1M tokens.  For comparison, OpenAI GPT-4o starts at $5.00 / 1M tokens.
  • @hoitab @hoitab on threads
    It's not just fast and lightweight.  It will also have loooooonnnnggg context.  GoogleIO
  • @mehtology Ivan Mehta on threads
    Congratulations to Google on launching more names.
  • @google @google on x
    The Gemini era is here, bringing the magic of AI to the tools you use every day. Learn more about all the announcements from #GoogleIO → https://blog.google/... [video]
  • @googledeepmind @googledeepmind on x
    Today, we're excited to introduce a new Gemini model: 1.5 Flash. ⚡ It's a lighter weight model compared to 1.5 Pro and optimized for tasks where low latency and cost matter - like chat applications, extracting data from long documents and more. #GoogleIO [image]
  • @oriolvinyalsml Oriol Vinyals on x
    Live from #GoogleIO, we're announcing significant updates to our Gemini family of models. More multimodal, real time, faster, better, longer context... We can't wait to see the amazing things people will do with this tech, as we continue making it more helpful and useful for [ima…
  • @googledeepmind @googledeepmind on x
    We've built a range of AI systems that can: 🔵 Turn vision and language into action for robots 🔵 Navigate complex virtual 3D environments 🔵 Solve Olympiad-level math problems And more. #GoogleIO https://twitter.com/... [image]
  • @demishassabis Demis Hassabis on x
    Making great progress on the Gemini Era. At #GoogleIO we shared 2M long context breakthrough with 1.5 Pro and announced Gemini 1.5 Flash, a lighter-weight multimodal model with long context designed to be fast and cost-efficient to serve at scale. More: https://blog.google/... [i…
  • @simonw Simon Willison on x
    Interesting that Google used the same pricing multiple for Gemini Pro 1.5 vs the new Gemini Flash - Flash is exactly 1/10th the price of Pro: https://ai.google.dev/pricing
  • @google @google on x
    Both Gemini 1.5 Pro and 1.5 Flash are natively multimodal with a massive 1 million token context window. Sign up to try the new 2 million context window in Gemini 1.5 Pro in Google AI Studio → https://ai.google.dev/ #GoogleIO [image]
  • @simonw Simon Willison on x
    The llm-gemini model now supports the new inexpensive Gemini 1.5 Flash model: pipx install llm llm install llm-gemini —upgrade llm keys set gemini # paste API key here llm -m gemini-1.5-flash-latest ‘a short poem about otters’
  • @google @google on x
    Built with the same technology as Gemini, our Gemma family of models offers industry-leading performance in lightweight 7B and 2B sizes. Today, we're introducing PaliGemma, our first vision-language open model. #GoogleIO [image]
  • @google @google on x
    At #GoogleIO @GoogleDeepMind introduced a series of updates across the Gemini family of models, including our new lighter-weight model 1.5 Flash. We also shared new research that is helping guide the future of AI assistants. https://blog.google/...
  • @demishassabis Demis Hassabis on x
    We think of @GoogleDeepMind as the engine room of @Google in the AI era. Thrilled to share our vision at #GoogleIO incl the latest Gemini model 1.5 Flash, Project Astra our universal AI agent effort, our new gen video model Veo, Imagen 3 & lots more! https://deepmind.google/ [ima…
  • @zacharynado Zachary Nado on x
    squeezing model sizes down is just as important as scaling up in my opinion, and 1.5 Flash ⚡️ is so incredibly capable while so small and cheap it's been blowing our minds 🤯 it has been an incredible privilege and so much fun building this model (sometimes too much fun)! ⚡️
  • @jerryjliu0 Jerry Liu on x
    Gemini 1.5 Pro is $3.50 up to 128k tokens, $7 after Gemini 1.5 Flash is $0.35 up to 128k tokens (and presumably double after?) Assuming the last is true, this is the cheapest 1M token model out there. Game changer [image]
  • @googledevs @googledevs on x
    Today we launched Gemini 1.5 Flash in public preview at #GoogleIO, and Gemini 1.5 Pro will be coming with a two million tokens context window (behind waitlist). Both models are available in 200+ regions, including many countries in Europe. Try it out → http://aistudio.google.com/…
  • @google @google on x
    Today we're announcing Gemini 1.5 Flash, optimized for narrower or high-frequency tasks. Both Gemini 1.5 Pro and 1.5 Flash are now available in over 200 countries and territories. #GoogleIO https://blog.google/... [image]
  • @anshelsag Anshel Sag on x
    Why isn't @Google talking about the cost of these models? I think that @OpenAI did a much better job of communicating those improvements.
  • @deliprao @deliprao on x
    Google : OpenAI :: Flash : Turbo
  • @google @google on x
    Introducing Gemini 1.5 Flash ⚡ It's a lighter-weight model, optimized for tasks where low latency and cost matter most. Starting today, developers can use it with up to 1 million tokens in Google AI Studio and Vertex AI. #GoogleIO [image]
  • @daveyalba Davey Alba on x
    DeepMind CEO Demis Hassabis has taken the stage at #GoogleIO for the first time. He introduces Gemini 1.5 Flash, which Google says is the fastest AI model available through its API, lighter weight than 1.5 Pro but still adept enough for high-frequency tasks [image]
  • @mishaalrahman Mishaal Rahman on x
    Gemini 1.5 Flash is a lightweight version of 1.5 while still supporting multimodal reasoning and long context windows (1M tokens). Developers can sign up to try 2M tokens. This is a model designed for reduced latency applications. [image]
  • @lanceulanoff Lance Ulanoff on x
    Google introduces Gemini 1.5 Flash, a lighter-weight model compared to Pro. Still multimodal. #GoogleIO [image]
  • @officiallogank Logan Kilpatrick on x
    Hello world to Gemini 1.5 Flash 📸 Flash comes natively with multi-modal capabilities, up to a 2 million context window, higher rate limits, lower cost, and more.
  • @jordannovet Jordan Novet on x
    did Demis just wait for the crowd to clap for him after he announced simply the launch of ‘Gemini 1.5 Flash’ with zero context