AI3 min read
NVIDIA says Groq 3 LPX is in full production for agentic inference
At Hot Chips, NVIDIA called Groq 3 LPX an interactive inference accelerator that extends Vera Rubin. On Gemma 4 31B with 100k context, Artificial Analysis measured 3,400 output tokens per second.
