NVIDIA Announces Full-Scale Production of Groq 3 LPX Racks for Nebius Deployment

iconChaincatcher
Share
AI summary iconSummary
NVIDIA Senior Director Dion Harris confirmed on Monday that Groq 3 LPX racks are now in full production for Nebius, a new cloud provider. This deployment follows NVIDIA’s $20 billion licensing agreement with Groq in December 2025. New token listings from cloud providers may signal improved infrastructure. The Groq 3 LPX, launched in March, features 500 MB of SRAM per chip and processes 3,400 tokens per second. Low-latency chips enable cloud providers to better serve latency-sensitive users. Inflation data could influence data center investment trends. Groq founder Jonathan Ross now leads software architecture at NVIDIA. The company plans to allocate a quarter of its data center space for Groq chips.

ChainCatcher reports that on Monday local time, Dion Harris, Senior Director at NVIDIA, announced that the Groq 3 LPX rack has entered full-scale mass production and will be deployed in the data centers of the new cloud service provider Nebius, with an expected launch later this year. This marks the commercialization of technology acquired by NVIDIA following its approximately $20 billion technology licensing deal with Groq in December last year. Several key Groq employees have joined NVIDIA, with founder Jonathan Ross serving as NVIDIA’s Chief Software Architect. NVIDIA is accelerating the production of Groq chips and delivering products to customers, underscoring the importance of low-latency inference. The Groq 3 LPX, released in March this year, is an inference accelerator for the Vera Rubin platform, integrating 500 megabytes of high-speed SRAM directly on the chip die to reduce memory bottlenecks. Each LPX rack can integrate 256 Groq 3 chips; according to NVIDIA’s benchmark tests, it can process 3,400 tokens per second. The chip is manufactured by Samsung. Harris stated that low-latency chips are not intended to replace GPUs but rather to use the appropriate processor for different stages of workloads, enabling cloud providers to offer premium pricing tiers to latency-sensitive users. NVIDIA CEO Jensen Huang plans to allocate approximately one-quarter of his data center space for AI programming applications to Groq chips, with the remainder dedicated to Vera Rubin systems, and expects combined sales of Blackwell and Vera Rubin systems to reach $1 trillion by 2027.

Disclaimer: The information on this page may have been obtained from third parties and does not necessarily reflect the views or opinions of KuCoin. This content is provided for general informational purposes only, without any representation or warranty of any kind, nor shall it be construed as financial or investment advice. KuCoin shall not be liable for any errors or omissions, or for any outcomes resulting from the use of this information. Investments in digital assets can be risky. Please carefully evaluate the risks of a product and your risk tolerance based on your own financial circumstances. For more information, please refer to our Terms of Use and Risk Disclosure.