NVIDIA Groq 3 LPX hits 3,431 tokens/sec on 100K context benchmark with Gemma 4 31B
NVIDIA's Groq 3 LPX accelerator hit 3,431 output tokens per second on the Artificial Analysis 100K context benchmark with Gemma 4 31B. The result enables responsive multi-agent systems that can process full session histories at long context lengths.