NVIDIALaunchDevelopersAugust 24, 2026

NVIDIA Groq 3 LPX enters full production for Vera Rubin

NVIDIA announced Groq 3 LPX, its interactive AI inference accelerator for the Vera Rubin platform, is in full production. In an Artificial Analysis benchmark running Gemma 4 31B, it delivered 3,431 output tokens per second on 100K context, 4x faster than the nearest alternative. Nebius is the first AI cloud to adopt it.

3 sources

NVIDIA by email

Get an email when NVIDIA has news

No news that day, no email.

More stories today

Open the live feed
NVIDIA Groq 3 LPX enters full production for Vera Rubin