Nvidia says its Groq 3 LPX racks delivered 3,400 tokens per second in an Artificial Analysis benchmark running Gemma 4 31B with a 100,000-token input sequence (The Register)

1 hour ago 1
Add to circle

The Register:
Nvidia says its Groq 3 LPX racks delivered 3,400 tokens per second in an Artificial Analysis benchmark running Gemma 4 31B with a 100,000-token input sequence  —  Nvidia's $20 billion bet on Groq's LPU tech sure looks like it was a good one.  On Monday, the GPU giant offered the first glimpse …

Read Entire Article