Nvidia says its Groq 3 LPX is four times faster than Cerebras, but the math is more complicated
The Decoderen

Nvidia is moving its Groq 3 LPX inference chip into full production and reports 3,400 tokens per second on Gemma 4 31B, four times faster than Cerebras. But the numbers don't tell the whole story. Nvidia needs at least 64 accelerators to get there, while Cerebras needs only one or two, according to The Register. How well the architecture scales with large MoE models remains an open question. The article Nvidia says its Groq 3 LPX is four times faster than Cerebras, but the math is more complicated appeared first on The Decoder .
This is a short summary published by AI Global Wire. The full article is owned and hosted by The Decoder — open it there to read it in full.
Read the full story at The Decoder- Verktyg
Related AI news
- Apple unveils M6, a 2nm chip with a 12-core CPU and GPU, and up to 32GB of unified memory, saying it provides "the world's fastest single-threaded performance" (Jay Peters/The Verge)Techmeme · August 25, 2026
- Apple unveils a Mac mini with M6 and M5 Pro, with up to 4x faster AI performance and up to 2x faster graphics, for $899+ with M6 and $1,699+ with M5 Pro (Apple)Techmeme · August 25, 2026
- Apple updates the Mac Studio with M5 Max and M5 Ultra, with up to 4.3x faster AI performance, faster graphics, and up to 512GB of unified memory for $2,499+ (Apple)Techmeme · August 25, 2026
- Nucleus Security launches Helix, an AI engine for exposure managementSiliconANGLE · August 25, 2026
- Ropedia launches next-gen wearable capture device for robotic AI training dataSiliconANGLE · August 25, 2026
- Nvidia NemoClaw flaw let attackers poison the model behind a developer’s AI agentSiliconANGLE · August 25, 2026