§ models · storyline
Nvidia Groq 3 LPX racks deliver 3,400 tokens per second
Nvidia's Groq 3 LPX racks achieve 3,400 tokens per second on Artificial Analysis benchmarks running Gemma 4 31B with a 100,000-token input sequence.
The Register: Nvidia says its Groq 3 LPX racks delivered 3,400 tokens per second in an Artificial Analysis benchmark running Gemma 4 31B with a 100,000-token input sequence — Nvidia's $20 billion bet on Groq's LPU tech sure looks like it was a good one.
On Monday, the GPU giant offered the first glimpse …
§ sources1 publication · timeline below