NVIDIALaunchDevelopersAugust 24, 2026

NVIDIA Groq 3 LPX enters full production for agentic AI

NVIDIA announced Groq 3 LPX, an interactive inference accelerator for Vera Rubin, is in full production. In an Artificial Analysis benchmark running Gemma 4 31B, it delivered 3,400 output tokens per second for 100,000-token contexts, 4x faster than the nearest alternative. Nebius is the first AI cloud to adopt it.

How this story unfolded

3 days · 7 reports · 2 community posts · 9 of 10 shown

  1. Aug 24
  2. Aug 27

NVIDIA by email

Get an email when NVIDIA has news

No news that day, no email.

More stories today

Open the live feed