
NVIDIA's Groq 3 LPX Enters Mass Production, Hits 3,431 Tokens Per Second in Benchmark
NVIDIA's Groq 3 LPX, now in mass production, splits AI inference work between GPUs and LPUs to reach 3,431 output tokens per second. Here's what that number means, where the benchmark falls short, and what's behind the $17 billion deal.








