Back to News

Nvidia's Groq 3 LPX accelerator enters full production; Nebius first customer, SpaceXAI to adopt Vera CPUs

#nvidia#groq#inference#accelerator

Nvidia announced that its Groq 3 LPX inference accelerator has entered full production, with Nebius as the first customer. SpaceXAI will adopt Vera CPUs. In a separate benchmark, Nvidia reported that Groq 3 LPX racks achieved 3,400 tokens per second running Gemma 4 31B with a 100,000-token input sequence.

Coverage timeline

  1. Techmeme

    Mike Wheatley / SiliconANGLE : Nvidia says its inference accelerator Groq 3 LPX has entered full production and Nebius has signed on as the first customer; SpaceXAI will adopt Vera CPUs — Chipmaker Nvidia Corp. says its dedicated artificial intelligence inference accelerator Groq 3 LPX has now entered full production …

  2. Techmeme

    The Register : Nvidia says its Groq 3 LPX racks delivered 3,400 tokens per second in an Artificial Analysis benchmark running Gemma 4 31B with a 100,000-token input sequence — Nvidia's $20 billion bet on Groq's LPU tech sure looks like it was a good one. On Monday, the GPU giant offered the first glimpse …