Nvidia says its Groq 3 LPX racks delivered 3,400 tokens per second in an Artificial Analysis benchmark running Gemma 4 31B with a 100,000-token input sequence (The Register)

Techmeme ·

Nvidia disclosed early Groq 3 LPU benchmark results showing unusually high long-context inference throughput for Gemma 4 31B.

Categories: Money & Moves

Excerpt

<a href="https://www.theregister.com/systems/2026/08/24/what-nvidias-first-groq-3-lpu-benchmarks-do-and-dont-tell-us-about-its-20b-gamble/5291880"><img align="RIGHT" border="0" hspace="4" src="http://www.techmeme.com/260824/i25.jpg" vspace="4" /></a> <p><a href="https://www.techmeme.com/260824/p25#a260824p25" title="Techmeme permalink"><img height="12" src="http://www.techmeme.com/img/pml.png" style="border: none; padding: 0; margin: 0;" width="11" /></a> <a href="https://www.theregister.com/">The Register</a>:<br /> <span style="font-size: 1.3em;"><b><a href="https://www.theregister.com/systems/2026/08/24/what-nvidias-first-groq-3-lpu-benchmarks-do-and-dont-tell-us-about-its-20b-gamble/5291880">Nvidia says its Groq 3 LPX racks delivered 3,400 tokens per second in an Artificial Analysis benchmark running Gemma 4 31B with a 100,000-token input sequence</a></b></span>&nbsp; &mdash;&nbsp; Nvidia's $20 billion bet on Groq's LPU tech sure looks like it was a good one.&nbsp; On Monday, the GPU giant offered the first glimpse &hellip; </p>