
An FPGA bridges the synchronous LPU domain and the asynchronous world of host I/O and GPU hand-offs, and Nvidia's Dynamo runtime, together with an LPU extension to CUDA, orchestrates the split. The company put the gains from these modes at roughly three-to-five-times over Rubin alone on a two-trillion-parameter workload with a 400K-token cached context, all Nvidia-measured.
Cerebras used the same Hot Chips session to present its CS4 wafer-scale system, which chief system architect Jean-Philippe Fricker said runs up to 30 times faster than GPUs and doubles the token rate of the CS3 while carrying 10 times the token capacity. Each CS4 rack packs three wafer-scale engines into a new modular platform Cerebras calls Nexus, built around pluggable compute "backpacks" that separate power, compute, and I/O, and Fricker put its memory bandwidth at 43 PB/s, which he told the audience was "2,000 times higher memory bandwidth than Nvidia's next-generation Rubin chip." Cerebras also has a partner for the prefill side of the same problem: it agreed in July to pair AMD Helios GPUs for prefill with its wafer-scale engines for decode, the same division of labor Nvidia now builds in-house with Groq.
Nvidia pulled the Rubin CPX, its own GDDR7-based long-context accelerator, to focus on shipping the LPU this year, a decision VP Ian Buck laid out at GTC 2026 . The $20 billion deal that produced the LP30 was structured as a non-exclusive IP license plus the hiring of Ross, president Sunny Madra, and most of Groq's engineers, a form that avoided a formal merger review. Arsovski opened the Hot Chips talk by calling it "a pinch me moment for the Groq team that's now integrated into the Nvidia group."
Senators Elizabeth Warren and Richard Blumenthal wrote to the FTC and to Nvidia in early 2026, arguing the arrangement acquired Groq "in all but name," and no formal, deal-specific investigation has been confirmed as of late August.
(Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) (Image credit: Nvidia) Image 1 of 44 View Original
TOPICS Nvidia See all comments (0) Luke James Social Links Navigation Contributor Luke James is a freelance writer and journalist. Although his background is in legal, he has a personal interest in all things tech, especially hardware and microelectronics, and anything regulatory.
Key considerations
- Investor positioning can change fast
- Volatility remains possible near catalysts
- Macro rates and liquidity can dominate flows
Reference reading
- https://www.tomshardware.com/tech-industry/semiconductors/SPONSORED_LINK_URL
- https://www.tomshardware.com/tech-industry/semiconductors/nvidia-presents-groq-3-lpx-architecture-and-unveils-its-first-third-party-inference-benchmark#main
- https://www.tomshardware.com/my-account
- Enthusiast turns a Lenovo Yoga laptop, an M.2 slot, and AMD Radeon RX 7900 XT into 'the world's stupidest' desktop for local AI chatbots — M.2 franken-rig cripp
- Into the Omniverse: How Open World Models Push the Frontier of Physical AI
- Rockstar releases statement after a week of GTA VI leaks, avoids mentioning leaker's demands — says that gameplay leaks have been ‘heartbreaking for our team’
- Mad scientist makes LEDs in his backyard — semiconductor wizard who made his own RAM turns his attention to using bathroom eBay laser to etch sapphire wafers
- Windows veteran's vibe-coded Task Manager now also runs on Mac and Linux — downloadable app is the result of a 107-page spec fed to Claude Code
Informational only. No financial advice. Do your own research.