Mark RadarMARK RADAR
About
EN
Sign in

Nvidia Puts Groq 3 LPX Into Production as SpaceXAI Adopts Vera CPU

1 reports · First detected 2026-08-25 · Last active 2026-08-25

Nvidia’s Vera Rubin is a rack-scale computing platform designed for agentic AI, combining the Vera CPU, Rubin GPUs and Groq 3 LPX inference accelerators. Unlike conventional chatbot workloads, AI agents must coordinate tools, execute code, process data and run simulations between model calls. That shifts more of the bottleneck to CPU orchestration and low-latency token generation, making tightly integrated infrastructure critical as AI systems move from training into interactive, commercial deployments.

On Aug. 24, 2026, Nvidia said SpaceXAI would deploy Vera CPUs as it expands the infrastructure behind Grok toward gigawatts of computing capacity, with an optimized Vera Rubin NVL72 also planned for its first-generation Starmind AI satellite. Nvidia separately put Groq 3 LPX into full production. In a Gemma 4 31B test using a 100,000-token context, the system generated 3,400 output tokens per second, four times the nearest alternative. Nebius will be the first AI cloud provider to add it to its Token Factory inference platform.

All Coverage

1 original reports

The Backstory

The history behind this event

No historical echoes for this signal

Mark Radar|MARK RADAR

If you search news on Google, you can set Mark Radar as a preferred source—our coverage will show up more often in your results. Set as preferred source on Google →

All times are in Taipei time (GMT+8)