Open account

Nvidia’s $20 Billion Groq Bet Is Going Live Before the End of 2026. Here’s What It Means for Investors.

Nvidia Corporation announced that its new low-latency AI inference system, Groq 3 LPX, has entered full production and will launch commercially before the end of 2026. Nebius has been named the first AI cloud provider to deploy the technology via its Nebius Token Factory platform. The development follows Nvidia's $20 billion cash licensing deal inked in December 2025, which included hiring key Groq leadership and engineering personnel while maintaining Groq as an independent entity. The Groq 3 LPX system features liquid-cooled racks containing 256 Language Processing Units (LPUs) integrated into Nvidia's Vera Rubin architecture. Designed to operate alongside existing GPU clusters without requiring alterations to standard CUDA software workflows, the architecture reportedly boosts inference throughput by up to 35 times per megawatt. This high-efficiency architecture addresses key power-supply bottlenecks in modern data centers while solidifying Nvidia's competitive moat against rising non-GPU inference architectures as industry focus shifts from AI training to low-latency deployment.

Category

NVIDIA

Sentiment

Bullish

Event

Product launch

Reading time

1 min