Groq raises $650M following massive Nvidia talent and IP deal
AI chipmaker Groq has secured $650 million in new funding to continue scaling its inference cloud business following the transition of its original hardware IP to Nvidia. The company has undergone an executive leadership refresh and expanded its global data center footprint to 13 locations to serve a growing base of inference developers.
Key Takeaways
- New funding led by Disruptive and Infinitum to support pivot toward a "neocloud" inference-as-a-service model.
- Nvidia now holds non-exclusive licenses for Groq’s LPU hardware IP, launching the Nvidia Groq 3 LPX hardware system in March 2026.
- Post-pivot leadership includes CEO Doug Wightman, COO Alan Rice (ex-Meta), and CTO Sinclair Schuller.
- Groq currently operates 13 data centers globally and serves over five million developers processing trillions of tokens weekly.
Why It Matters
Groq’s recapitalization signals that specialized inference clouds remain high-priority assets even as hardware IP commoditizes. While Nvidia has integrated Groq’s high-bandwidth SRAM architecture into its Rubin platform to solve the "memory wall" for trillion-parameter models, Groq is attempting to maintain its moat through geographic scale and a developer-first cloud abstraction. This survival strategy demonstrates that the "not-acqui-hire" model—where giants strip talent and IP while leaving the entity intact—does not necessarily end a startup’s independence. Watch Groq's ability to maintain performance leads over Nvidia's native LPX systems to gauge the viability of independent inference providers.
Additional Context
The trend of "not-acqui-hire" deals has reshaped the AI competitive landscape over the last year as big tech firms seek to bypass antitrust scrutiny. Per Forbes (June 2026), Meta executed a similar $14.3 billion deal with Scale AI in June 2025, taking a 49% stake and hiring founder Alexandr Wang to lead a internal unit. Despite losing key leadership, Scale AI reportedly reached a $1 billion revenue run rate by mid-2026, suggesting that infrastructure-heavy startups can survive the departure of their founding teams if their data and cloud divisions remain robust. Simultaneously, the center of gravity in the AI market is shifting from Model Training to Model Inference. Per industry analysis (May 2026), inference is projected to represent two-thirds of all AI compute demand by the end of 2026, up from just one-third in 2023. This volume-driven shift has sparked massive funding for specialized infrastructure players. For instance, Baseten entered talks for a $1 billion raise at an $11 billion valuation in May 2026, according to Cyber News Centre, highlighting the premium margins associated with optimizing real-time token generation. Nvidia’s integration of Groq’s technology has already reached the hardware level. At GTC 2026, Nvidia unveiled the Groq 3 LPX rack—the first product resulting from the licensing deal—which promises 35x higher throughput per megawatt for trillion-parameter models compared to the prior Blackwell architecture, per The Decoder (March 2026). This move effectively forces independent providers like Groq to pivot away from selling silicon toward high-value, managed-service layers like GroqCloud to remain competitive against the vertically integrated hyperscalers.
Read full article at techcrunch.com
Get this in your inbox → Subscribe
Enjoy our coverage?
Add StreamingMeme as a preferred source on Google to see more of our streaming news at the top of your Search results.
Add as preferred source