NVIDIALaunchDevelopersAugust 24, 2026

NVIDIA Groq 3 LPX in full production, hits 3,431 tokens/s on Gemma 4 31B

NVIDIA announced Groq 3 LPX, a token-generation accelerator for Vera Rubin, is in full production. Artificial Analysis measured 3,431 output tokens/s on Gemma 4 31B with 100K context, 4x faster than the nearest alternative. Nebius is the first AI cloud to adopt it.

How this story unfolded

same day · 3 reports · 2 community posts · 5 of 6 shown

  1. Aug 24
  2. Aug 25

NVIDIA by email

Get an email when NVIDIA has news

No news that day, no email.

More stories today

Open the live feed