/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

days · browse · Enter similar · o open

Nvidia and Mistral release Mistral NeMo, a 12B-parameter language model with a 128K-token context window, available under the Apache 2.0 open-source license

Mistral NeMo: our new best small model.  A state-of-the-art 12B model … Jonathan Kemper / The Decoder : Mistral releases three new LLMs for math, code and general tasks X: Prince Canuma / @prince_canuma : Mistral NeMo Instruct (Q4) is blazing fast 🔥 Running locally on M3 Max at 37 tokens/s using MLX 🚀 > pip install fastmlx And install mlx-lm from source (PR #895) [video] Anshel Sag / @anshelsag : This is a very interesting partnership and one that will likely only increase NVIDIA's relevance. Simon Willison / @simonw : “Mistral NeMo: our new best small model. A state-of-the-art 12B model with 128k context length, built in collaboration with NVIDIA, and released under the Apache 2.0 license.” Sounds like there's a new OpenAI model coming later too, a 4o-mini replacement for GPT-3.5 Turbo Sankalp / @dejavucoder : >128k context window >beats competitors in benchmarks >multi-lingual; >new tiktoken-based tokenizer called tekken trained over multiple languages, >weights on hf [image] @reach_vb : Let's goooo! Nvidia & Mistral release Mistral NeMo 12B 🔥 > Apache 2.0 licensed w/ 128K contact > Beats Llama 3 8B, Gemma 2 9B > Multilingual - EN, FR, DE, ES, IT, PT, CN, JP, KR, AR ⚡️ > Base + Instruct model (uncensored) > Supports function calling > Trained on both [image] Alex Volkov / @altryne : Mistral team is on fire this week damn Colab with @nvidia releasing a 128k!! 12B new open source model that's likely to be the next small-ish OSS darling called NeMo 👏 [image] Arthur Mensch / @arthurmensch : Today, we're announcing Mistral NeMo, a tiny multilingual model, 128k context length, trained with quantization awareness in collaboration with the NVIDIA research team. Kyle Russell / @kylebrussell : “Mistral NeMo was trained with quantisation awareness, enabling FP8 inference without any performance loss.”

VentureBeat Michael Nuñez

Discussion

  • @prince_canuma Prince Canuma on x
    Mistral NeMo Instruct (Q4) is blazing fast 🔥 Running locally on M3 Max at 37 tokens/s using MLX 🚀 > pip install fastmlx And install mlx-lm from source (PR #895) [video]
  • @anshelsag Anshel Sag on x
    This is a very interesting partnership and one that will likely only increase NVIDIA's relevance.
  • @simonw Simon Willison on x
    “Mistral NeMo: our new best small model. A state-of-the-art 12B model with 128k context length, built in collaboration with NVIDIA, and released under the Apache 2.0 license.” Sounds like there's a new OpenAI model coming later too, a 4o-mini replacement for GPT-3.5 Turbo
  • @dejavucoder Sankalp on x
    >128k context window >beats competitors in benchmarks >multi-lingual; >new tiktoken-based tokenizer called tekken trained over multiple languages, >weights on hf [image]
  • @reach_vb @reach_vb on x
    Let's goooo! Nvidia & Mistral release Mistral NeMo 12B 🔥 > Apache 2.0 licensed w/ 128K contact > Beats Llama 3 8B, Gemma 2 9B > Multilingual - EN, FR, DE, ES, IT, PT, CN, JP, KR, AR ⚡️ > Base + Instruct model (uncensored) > Supports function calling > Trained on both [image]
  • @altryne Alex Volkov on x
    Mistral team is on fire this week damn Colab with @nvidia releasing a 128k!! 12B new open source model that's likely to be the next small-ish OSS darling called NeMo 👏 [image]
  • @arthurmensch Arthur Mensch on x
    Today, we're announcing Mistral NeMo, a tiny multilingual model, 128k context length, trained with quantization awareness in collaboration with the NVIDIA research team.
  • @kylebrussell Kyle Russell on x
    “Mistral NeMo was trained with quantisation awareness, enabling FP8 inference without any performance loss.”
  • r/SillyTavernAI r on reddit
    Mistral partners with Nvidia to release Nemo, a 12B model outperformming Gemma and Llama-3 8B