| Mobile Web

Nvidia begins mass production of Groq 3 LPX for AI agents

Nvidia has shifted its Groq 3 LPX inference accelerator for AI agents into mass production. The chip extends Nvidia\'s Vera Rubin data center platform and targets high-speed token generation to reduce decode latency in agent workflows. In an Artificial Analysis benchmark, it produced 3,400 tokens per second running the open-source Gemma 4 31B model with a 100,000-token context window. Nebius Group is among early adopters, and SpaceX was named a new flagship customer.