AI news story

GTC 2026: With Groq 3 LPX, Nvidia adds dedicated inference hardware to its platform for the first time

At GTC 2026, Nvidia expanded the Vera Rubin platform it introduced at CES with custom CPU racks, dedicated inference chi…

  • Hardware
  • Source: The Decoder
  • Published: 2026-03-17

Editor's take

Nvidia's GTC 2026 announcement introduces specialized inference hardware, the Groq 3 LPX, to its Vera Rubin platform, alongside custom CPU racks and an inference OS.

This move signifies Nvidia's strategic pivot to address the rapidly growing demand for efficient AI inference, a crucial bottleneck for deploying large language models like Meta's Llama 3 or Mistral AI's Mixtral in production environments. By offering dedicated inference silicon, Nvidia aims to capture a larger share of the inference market, directly challenging specialized inference chip makers and potentially impacting cloud providers' internal silicon efforts.

The success of Groq 3 LPX will hinge on its performance benchmarks against existing GPUs and ASICs in real-world inference workloads, as well as the adoption rate of its new inference OS and open model alliances. Future developments to monitor include Nvidia's pricing strategy for these new components and the competitive responses from Intel, AMD, and other AI hardware developers.