Nvidia’s Groq 3 LPX racks go into full production after $20B Groq deal

TL;DR Summary
Nvidia says its Groq 3 LPX rack is in full production following its $20 billion acquisition of Groq, to be deployed with Vera CPUs and Rubin GPUs at neocloud Nebius and go online later this year, underscoring the push for ultra-low-latency AI inference; the LPX rack uses 256 Groq 3 chips with 500 MB SRAM per chip and can deliver about 3,400 tokens per second, with Samsung manufacturing Groq chips and TSMC handling Nvidia GPUs as Nvidia ramps Vera Rubin shipments toward multi-year sales.
- Nvidia says Groq racks will be online this year following $20 billion purchase CNBC
- NVIDIA Groq 3 LPX Now in Full Production With World-Class Speed for Agentic AI NVIDIA Newsroom
- AgentX - InferenceXv3: Does CUDA Moat Hold up in Agentic Inferencing? SemiAnalysis
- SpaceXAI to use Nvidia's Vera CPUs for agentic AI applications Seeking Alpha
- Nvidia Announces New Customers For Vera CPU, Groq LPX Racks The Information
Reading Insights
Total Reads
0
Unique Readers
6
Time Saved
2 min
vs 3 min read
Condensed
85%
548 → 84 words
Want the full story? Read the original article
Read on CNBC