Nvidia unveils Grace, a high-performance Arm-based server CPU for large-scale neural network workloads, expected to become available in Nvidia products in 2023
Kicking off another busy Spring GPU Technology Conference for NVIDIA, this morning the graphics and accelerator designer …
Context & Ripple Effects
Grace marks Nvidia's move from supplying accelerators alongside other vendors' server CPUs toward offering its own Arm-based CPU for neural-network systems. The initial product was later expanded into a 144-core Grace CPU Superchip, showing that the CPU was a product line rather than a one-off component.
The arc then shifted from a standalone CPU to tightly coupled compute packages: Grace Hopper combined Grace with a GPU, and Blackwell's GB200 again paired GPUs with Grace. That makes this announcement the starting point for Nvidia's CPU-GPU system strategy.
First-order effects
- Nvidia adds an Arm-based server-CPU option to its data-center portfolio, giving customers building large neural-network workloads a CPU designed around Nvidia's own products.
- Server buyers planning Nvidia-based AI systems gain a new CPU road map, while Nvidia takes on responsibility for the host processor as well as the accelerator.
Second-order effects
- Nvidia's subsequent Grace Superchip and Grace Hopper products turn CPU selection into part of a broader Nvidia platform decision, increasing the value of validated CPU-GPU combinations for buyers.
- Arm's server-core roadmap gains a prominent deployment path: Neoverse V2 was slated to power the upcoming Grace CPU, linking Arm's cloud and HPC ambitions to Nvidia's system products.
Third-order effects
- If the Grace-to-GH200-to-GB200 progression holds, AI infrastructure will be organized increasingly around heterogeneous superchips rather than separately sourced CPUs and GPUs.
- That integration gives Nvidia more control over the compute stack, while narrowing the portion of an AI server that customers can choose independently.
The trend: Nvidia is moving from GPU supplier to provider of integrated Arm CPU-GPU systems for AI and high-performance computing.