/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

days · browse · Enter similar · o open

Nvidia launches Nemotron 3, a family of AI models using a hybrid mixture-of-experts architecture and the Mamba-Transformer design, in 30B, 100B, and ~500B sizes

Nvidia launched the new version of its frontier models, Nemotron 3, by leaning in on a model architecture that the world's …

VentureBeat Emilia David

Discussion

  • @artificialanlys @artificialanlys on x
    NVIDIA has just released Nemotron 3 Nano, a ~30B MoE model that scores 52 on the Artificial Analysis Intelligence Index with just ~3B active parameters Hybrid Mamba-Transformer architecture: Nemotron 3 Nano combines the hybrid Mamba-Transformer approach @NVIDIAAI has used on [ima…
  • @artificialanlys @artificialanlys on x
    NVIDIA focused on efficiency as well as intelligence with Nemotron 3 Nano, and it presents an attractive trade-off between speed and capability. In pre-release testing of the @DeepInfra serverless endpoint, we saw output speeds of ~380 tokens per second [image]
  • @nvidiaaidev @nvidiaaidev on x
    ✨ Meet our new open family of models: @NVIDIA Nemotron 3 Open in weights, data, tools, and training, Nemotron 3 is built for multi-agent apps and features: • An efficient hybrid Mamba‑Transformer MoE architecture • 1M token context for long-term memory and improved reasoning [vid…
  • @natolambert Nathan Lambert on x
    It's an honor to be competing with Nvidia for the best models with open data, checkpoints, and code. Super excited about Nemotron 3 and Nvidia's new focus on fully open models in 2025.
  • @leonderczynski Leon Derczynski on x
    New: Nemotron v3 is open, fastest, highest benchmark scoring. Nemotron v3 Nano delivers 4x higher throughput than Nemotron 2 Nano & delivers most tokens per second at scale using hybrid mamba/transformer MoE architecture - state space models are the way! https://research.nvidia.c…
  • @ctnzr Bryan Catanzaro on x
    Nemotron 3 Nano is competitive with other leading open source models, but 1.5-3.3X faster. [image]
  • @ctnzr Bryan Catanzaro on x
    Today, @NVIDIA is launching the open Nemotron 3 model family, starting with Nano (30B-3A), which pushes the frontier of accuracy and inference efficiency with a novel hybrid SSM Mixture of Experts architecture. Super and Ultra are coming in the next few months. [image]
  • r/LocalLLaMA r on reddit
    NVIDIA releases Nemotron 3 Nano, a new 30B hybrid reasoning model!