Nvidia's Groq 3 LPX accelerator enters full production; Nebius first customer, SpaceXAI to adopt Vera CPUs
Nvidia announced that its Groq 3 LPX inference accelerator has entered full production, with Nebius as the first customer. SpaceXAI will adopt Vera CPUs. In a separate benchmark, Nvidia reported that Groq 3 LPX racks achieved 3,400 tokens per second running Gemma 4 31B with a 100,000-token input sequence.
Coverage timeline
Techmeme
Mike Wheatley / SiliconANGLE : Nvidia says its inference accelerator Groq 3 LPX has entered full production and Nebius has signed on as the first customer; SpaceXAI will adopt Vera CPUs — Chipmaker Nvidia Corp. says its dedicated artificial intelligence inference accelerator Groq 3 LPX has now entered full production …

Techmeme
The Register : Nvidia says its Groq 3 LPX racks delivered 3,400 tokens per second in an Artificial Analysis benchmark running Gemma 4 31B with a 100,000-token input sequence — Nvidia's $20 billion bet on Groq's LPU tech sure looks like it was a good one. On Monday, the GPU giant offered the first glimpse …
