The Hallway Track
Product Launches

How NVIDIA Groq 3 LPX Unlocks Ultrafast Interactivity at Long Context on NVIDIA Vera Rubin

NVIDIA Developer Blog · Aug 24, 2026 · Product Launches

NVIDIA Groq 3 LPX accelerator enables ultrafast interactive inference on Vera Rubin NVL72 at long context

“NVIDIA Vera Rubin NVL72, the most versatile machine ever built, delivering high throughput and interactivity across the widest range of AI workloads—from small to large models, both open and closed.”

NVIDIA has detailed Groq 3 LPX, a dedicated interactive AI inference accelerator designed for its Vera Rubin NVL72 platform. The chip extends the NVL72's capabilities by enabling ultrafast interactivity specifically at long context lengths, addressing a key bottleneck in deploying large models for real-time applications. This signals NVIDIA's continued push to dominate the AI inference hardware stack with purpose-built silicon alongside its GPU line.

NVIDIA inference hardware Vera Rubin Groq accelerator

Watch / read the original source →