/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

days · browse · Enter similar · o open

Google unveils updates to its Gemma family of open models, including Gemma 2 2B, which it claims surpasses GPT-3.5 and Mixtral 8x7B on the LMSYS Chatbot Arena

Google has unveiled updates to its Gemma 2 family of open-source language models, focusing on improved performance, safety, and transparency.

The Decoder Jonathan Kemper

Discussion

  • @lmsysorg @lmsysorg on x
    Congrats @GoogleDeepMind on the Gemma-2-2B release! Gemma-2-2B has been tested in the Arena under “guava-chatbot”. With just 2B parameters, it achieves an impressive score 1130 on par with models 10x its size! (For reference: GPT-3.5-Turbo-0613: 1117, Mixtral-8x7b: 1114). This [i…
  • @googledeepmind @googledeepmind on x
    We're also introducing ShieldGemma: a series of state-of-the-art safety classifiers designed to filter harmful content. 🛡️ These target hate speech, harassment, sexually explicit material and more, both in the input and output stages. [image]
  • @reach_vb @reach_vb on x
    Google just dropped Gemma 2 2B! 🔥 > Scores higher than GPT 3.5, Mixtral 8x7B on the LYMSYS arena > MMLU: 56.1 & MBPP: 36.6 > Beats previous (Gemma 1 2B) by more than 10% in benchmarks > 2.6B parameters, Multilingual > 2 Trillion tokens (training set) > Distilled from Gemma 2 [ima…
  • @googledeepmind @googledeepmind on x
    Finally, we're announcing Gemma Scope, a set of tools to help researchers examine how Gemma 2 makes decisions. 🔍 It's a comprehensive, open suite of sparse autoencoders - specialized neural networks that zoom into the model's inner workings and make them more interpretable. [imag…
  • @gblazex Blaze on x
    The fact that we have a TINY model runnable on basically any smartphone + THIS closely aligned with human preferences & everyday usefulness is amazing. And all of it Open Weights permissive license incl. commercial use & fine-tuning. I mean come on... [image]
  • @chris_j_paxton Chris Paxton on x
    It's old news now but GPT 3.5 - chatgpt - was the one that kicked off the whole LLM craze. to see it being beaten by a 2B paramater model is mind-blowing. This really means we can make every single thing we own (somewhat) intelligent [image]
  • @googledeepmind @googledeepmind on x
    We're welcoming a new 2 billion parameter model to the Gemma 2 family. 🛠️ It offers best-in-class performance for its size and can run efficiently on a wide range of hardware. Developers can get started with 2B today → https://developers.googleblog.com/ ... [image]