/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

days · browse · Enter similar · o open

Nvidia says its inference accelerator Groq 3 LPX has entered full production, with Nebius signing on as the first customer, and SpaceX will deploy Vera CPUs

Chipmaker Nvidia Corp. says its dedicated artificial intelligence inference accelerator Groq 3 LPX has now entered full production …

SiliconANGLE Mike Wheatley

Context & Ripple Effects

Nvidia’s production milestone converts its December 2025 licensing deal with Groq into a shipping inference system. The company had outlined the 256-LPU rack in March, framing availability for the second half of 2026.

The move arrives alongside a reported 3,400-token-per-second benchmark result for the rack, while Nebius provides the first named cloud deployment. SpaceX’s Vera CPU commitment broadens the announcement from an accelerator sale to a compute-platform adoption.

First-order effects

  • Nebius becomes the first named AI cloud customer for Groq 3 LPX, giving Nvidia an initial production deployment for the Groq-derived inference rack.
  • SpaceX’s planned Vera CPU deployment gives Nvidia a named buyer for its CPU platform alongside the LPX inference system.

Second-order effects

  • Nebius can differentiate its AI-cloud offering around LPX throughput, placing competing inference providers under pressure to show comparable performance on long-context workloads.
  • Nvidia can sell inference capacity through both specialized LPX racks and Vera-based systems, increasing the commercial importance of system-level deployments rather than standalone chips.

Third-order effects

  • If early cloud and large-scale customer deployments expand, Nvidia’s Groq licensing bet points toward a more vertically integrated inference stack spanning specialized accelerators, CPUs and rack systems.
  • The market’s value capture shifts toward operators that can turn inference hardware performance into available, deployable cloud capacity rather than simply procure accelerators.

The trend: AI infrastructure suppliers are packaging specialized inference silicon, CPUs and full systems to compete for production workloads at cloud and large-scale operators.

Discussion

  • @elonmusk Elon Musk on x
    SpaceX, in partnership with Nvidia, has designed a space-optimized Vera Rubin NVL72 system for launch to orbit in Q4 next year, with significant scale in 2028
  • @zephyr_z9 @zephyr_z9 on x
    GROQ 3 LPX in Full Production LFG!!!
  • @benitoz Ben Pouladian on x
    NVIDIA's embargo lifted. Three calls we dated shipped today: Groq LPX, Spectrum X at 128k GPUs a rail, Vera CPUs at SpaceX While the circular-financing crowd argued, $NVDA raised next year's prices 15% The Audit grades our record before Wednesday https://bepresearch.substack.com/…
  • @demian_ai Dylan on x
    you read that right, Nebius is the first AI cloud to put NVIDIA Groq 3 LPX into production on @nebiustf Today NVIDIA also confirmed LPX is in full production, the first fruit of the Groq licensing deal from last year. Over the last year NVIDIA stopped just selling accelerators
  • @michaeldell Michael Dell on x
    The next wave of AI is about deploying incredible innovation at massive scale. Excited to work with @GroqInc and @nvidia to make that happen with @Dell. 🚀 https://x.com/...
  • @benitoz Ben Pouladian on x
    HotChips warm up: NVIDIA just posted the slide. Groq 3 LPX. 3,431 tokens a second. 4x the next public endpoint. Gemma 4 31B at 100,000-token context. In full production. Faster than Cerebras 😮🔥 $NVDA
  • @benbajarin Ben Bajarin on x
    $NBIS adopting $NVDA Groq 3 LPX as the market for premium inference tokens is starting to shape up. https://nebius.com/...
  • @patrickmoorhead Patrick Moorhead on x
    Groq meet Groq.
  • @vikramskr Vikram Sekar on x
    NVIDIA's talk about Vera CPUs at Hot Chips is a series of chart murders like the one below, which is disappointing. The subtext for comparisons is important. “Comparison based on relative performance for individual SPECrate®2026_int_base benchmark ratios between an NVIDIA Vera
  • @nvidianewsroom @nvidianewsroom on x
    NEWS: NVIDIA Groq 3 LPX is now in full production. NVIDIA Vera Rubin NVL72 is the foundation of every AI factory. Paired with Groq 3 LPX, it unlocks faster, smarter agents and breakthrough user experiences. Through extreme co-design across seven chips and five purpose-built
  • @nvidianewsroom @nvidianewsroom on x
    The first CPU built for agents is now at work at massive scale. @SpaceXAI is deploying NVIDIA Vera to power agentic AI — faster agents, GPUs fully utilized, and a single architecture from Earth to orbit. Read more ⬇️ https://nvidianews.nvidia.com/ ...
  • @groqinc @groqinc on x
    We are thrilled to announce that Groq will be among the first adopters of NVIDIA Groq 3 LPX, deploying it alongside NVIDIA Vera Rubin NVL72 in our purpose-built AI inference Cloud. Groq is working with Dell Technologies to deploy NVIDIA Groq 3 LPX. When Groq brings NVIDIA Groq 3
  • Stuart Pitts Stuart Pitts on linkedin
    Big day.  Nebius will be the first AI cloud to deploy NVIDIA Groq 3 LPX alongside Vera Rubin NVL72 for Nebius Token Factory. …
  • @tobiasmann Tobias Mann on bluesky
    How potent are Nvidia's Groq-3 LPX racks?  Very, but Gemma 4 31B tells an idealized story.  The real trick will be how efficiency it scales to MoE.  —  My full 1,200 word analysis only @theregister.com  —  www.theregister.com/systems/ 2026...