Nvidia launches Nemotron 3, a family of AI models using a hybrid mixture-of-experts architecture and the Mamba-Transformer design, in 30B, 100B, and ~500B sizes
Nvidia launched the new version of its frontier models, Nemotron 3, by leaning in on a model architecture that the world's …
VentureBeat Emilia David
Related Coverage
- As Meta fades in open-source AI, Nvidia senses its chance to lead ZDNET · Tiernan Ray
- Nvidia's Open Source Play Isn't About Openness Implicator.ai · Maria Garcia
- Nvidia Becomes a Major Model Maker With Nemotron 3 Wired · Will Knight
- NVIDIA Debuts Nemotron 3 Family of Open Models TechPowerUp
- Sorry, you have been blocked — You are unable to access theneuron.ai — Why have I been blocked? theneuron.ai
- NVIDIA Nemotron 3 Models Announced: Open AI Models In Nano, Super, Ultra Sizes, 4x Faster Vs Nemotron 2 Wccftech · Hassan Mujtaba
- Inside NVIDIA Nemotron 3: Techniques, Tools, and Data That Make It Efficient and Accurate NVIDIA Technical Blog
Discussion
-
@artificialanlys
@artificialanlys
on x
NVIDIA has just released Nemotron 3 Nano, a ~30B MoE model that scores 52 on the Artificial Analysis Intelligence Index with just ~3B active parameters Hybrid Mamba-Transformer architecture: Nemotron 3 Nano combines the hybrid Mamba-Transformer approach @NVIDIAAI has used on [ima…
-
@artificialanlys
@artificialanlys
on x
NVIDIA focused on efficiency as well as intelligence with Nemotron 3 Nano, and it presents an attractive trade-off between speed and capability. In pre-release testing of the @DeepInfra serverless endpoint, we saw output speeds of ~380 tokens per second [image]
-
@nvidiaaidev
@nvidiaaidev
on x
✨ Meet our new open family of models: @NVIDIA Nemotron 3 Open in weights, data, tools, and training, Nemotron 3 is built for multi-agent apps and features: • An efficient hybrid Mamba‑Transformer MoE architecture • 1M token context for long-term memory and improved reasoning [vid…
-
@natolambert
Nathan Lambert
on x
It's an honor to be competing with Nvidia for the best models with open data, checkpoints, and code. Super excited about Nemotron 3 and Nvidia's new focus on fully open models in 2025.
-
@leonderczynski
Leon Derczynski
on x
New: Nemotron v3 is open, fastest, highest benchmark scoring. Nemotron v3 Nano delivers 4x higher throughput than Nemotron 2 Nano & delivers most tokens per second at scale using hybrid mamba/transformer MoE architecture - state space models are the way! https://research.nvidia.c…
-
@ctnzr
Bryan Catanzaro
on x
Nemotron 3 Nano is competitive with other leading open source models, but 1.5-3.3X faster. [image]
-
@ctnzr
Bryan Catanzaro
on x
Today, @NVIDIA is launching the open Nemotron 3 model family, starting with Nano (30B-3A), which pushes the frontier of accuracy and inference efficiency with a novel hybrid SSM Mixture of Experts architecture. Super and Ultra are coming in the next few months. [image]
-
r/LocalLLaMA
r
on reddit
NVIDIA releases Nemotron 3 Nano, a new 30B hybrid reasoning model!