Search results for Nebius Token Factory
AI & Enterprise
Nvidia begins mass production of Groq 3 LPX for AI agents
Nvidia has shifted its Groq 3 LPX inference accelerator for AI agents into mass production. The chip extends Nvidia\'s Vera Rubin data center platform and targets high-speed token generation to reduce decode latency in agent workflows. In an Artificial Analysis benchmark, it produced 3,400 tokens per second running the open-source Gemma 4 31B model with a 100,000-token context window. Nebius Group is among early adopters, and SpaceX was named a new flagship customer.
Industry
Nvidia unveils four new products ahead of earnings, focusing on token generation speed
Nvidia unveiled new AI inference accelerators and network infrastructure products at Hot Chips 2026 in the United States on Aug. 25, a day before its second-quarter earnings. The company said interactive inference accelerator Nvidia Grok 3 LPX has entered mass production, alongside Spectrum-X Multiplane Ethernet scaling and ScaleIn infrastructure acceleration technology. The announcements focused on inference, particularly token generation speed, and also highlighted inference speed, network scalability and agent-focused CPUs within the Vera Rubin platform.
AI & Enterprise
Nebius to buy AI inference optimisation firm Eigen AI for $643 million
AI cloud platform company Nebius Group said on May 1 it will acquire AI inference and model optimisation company Eigen AI for about $643 million. The purchase price will be paid in cash and Class A shares. Eigen AI’s optimisation technology will be integrated into Nebius’ managed inference platform, Nebius Token Factory. The Eigen AI founding team is made up of alumni of MIT’s HAN Lab.