Nvidia says its inference accelerator Groq 3 LPX has entered full production, with Nebius signing on as the first customer, and SpaceX will deploy Vera CPUs
Chipmaker Nvidia Corp. says its dedicated artificial intelligence inference accelerator Groq 3LPX has now entered full production …
SiliconANGLEMike Wheatley
Context & Ripple Effects
Nvidia’s production milestone converts its December 2025 licensing deal with Groq into a shipping inference system. The company had outlined the 256-LPU rack in March, framing availability for the second half of 2026.
The move arrives alongside a reported 3,400-token-per-second benchmark result for the rack, while Nebius provides the first named cloud deployment. SpaceX’s Vera CPU commitment broadens the announcement from an accelerator sale to a compute-platform adoption.
First-order effects
Nebius becomes the first named AI cloud customer for Groq 3 LPX, giving Nvidia an initial production deployment for the Groq-derived inference rack.
SpaceX’s planned Vera CPU deployment gives Nvidia a named buyer for its CPU platform alongside the LPX inference system.
Second-order effects
Nebius can differentiate its AI-cloud offering around LPX throughput, placing competing inference providers under pressure to show comparable performance on long-context workloads.
Nvidia can sell inference capacity through both specialized LPX racks and Vera-based systems, increasing the commercial importance of system-level deployments rather than standalone chips.
Third-order effects
If early cloud and large-scale customer deployments expand, Nvidia’s Groq licensing bet points toward a more vertically integrated inference stack spanning specialized accelerators, CPUs and rack systems.
The market’s value capture shifts toward operators that can turn inference hardware performance into available, deployable cloud capacity rather than simply procure accelerators.
The trend:AI infrastructure suppliers are packaging specialized inference silicon, CPUs and full systems to compete for production workloads at cloud and large-scale operators.
SpaceX, in partnership with Nvidia, has designed a space-optimized Vera Rubin NVL72 system for launch to orbit in Q4 next year, with significant scale in 2028
NVIDIA's embargo lifted. Three calls we dated shipped today: Groq LPX, Spectrum X at 128k GPUs a rail, Vera CPUs at SpaceX While the circular-financing crowd argued, $NVDA raised next year's prices 15% The Audit grades our record before Wednesday https://bepresearch.substack.com/…
you read that right, Nebius is the first AI cloud to put NVIDIA Groq 3 LPX into production on @nebiustf Today NVIDIA also confirmed LPX is in full production, the first fruit of the Groq licensing deal from last year. Over the last year NVIDIA stopped just selling accelerators
The next wave of AI is about deploying incredible innovation at massive scale. Excited to work with @GroqInc and @nvidia to make that happen with @Dell. 🚀 https://x.com/...
HotChips warm up: NVIDIA just posted the slide. Groq 3 LPX. 3,431 tokens a second. 4x the next public endpoint. Gemma 4 31B at 100,000-token context. In full production. Faster than Cerebras 😮🔥 $NVDA
NVIDIA's talk about Vera CPUs at Hot Chips is a series of chart murders like the one below, which is disappointing. The subtext for comparisons is important. “Comparison based on relative performance for individual SPECrate®2026_int_base benchmark ratios between an NVIDIA Vera
NEWS: NVIDIA Groq 3 LPX is now in full production. NVIDIA Vera Rubin NVL72 is the foundation of every AI factory. Paired with Groq 3 LPX, it unlocks faster, smarter agents and breakthrough user experiences. Through extreme co-design across seven chips and five purpose-built
The first CPU built for agents is now at work at massive scale. @SpaceXAI is deploying NVIDIA Vera to power agentic AI — faster agents, GPUs fully utilized, and a single architecture from Earth to orbit. Read more ⬇️ https://nvidianews.nvidia.com/ ...
We are thrilled to announce that Groq will be among the first adopters of NVIDIA Groq 3 LPX, deploying it alongside NVIDIA Vera Rubin NVL72 in our purpose-built AI inference Cloud. Groq is working with Dell Technologies to deploy NVIDIA Groq 3 LPX. When Groq brings NVIDIA Groq 3
How potent are Nvidia's Groq-3 LPX racks? Very, but Gemma 4 31B tells an idealized story. The real trick will be how efficiency it scales to MoE. — My full 1,200 word analysis only @theregister.com — www.theregister.com/systems/ 2026...