$350M Series A Completes Groq's Pivot to Nvidia Neocloud
Groq closed a $350 million Series A round at a $3.5 billion valuation, finalizing its transition from an AI chipmaker to an Nvidia-powered inference cloud.
AI infrastructure provider Groq has closed a $350 million Series A funding round led by Disruptive, with planned participation from Nvidia. The investment values the company at $3.5 billion, a 49% decrease from its $6.9 billion peak in September 2025. The repricing finalizes a fundamental shift in Groq’s business model from a proprietary silicon designer to an Nvidia-powered cloud inference provider.
The Hardware Licensing Transition
The pivot from manufacturing hardware to operating a neocloud stems from a December 2025 non-exclusive technology licensing agreement with Nvidia. The deal, estimated at $20 billion in cash, transferred Groq’s core inference architecture to Nvidia. As part of the arrangement, Groq’s founding technical leadership, including founder-CEO Jonathan Ross and president Sunny Madra, joined Nvidia.
Nvidia quickly integrated the acquired architecture into its own product stack, resulting in the March 2026 debut of the Groq 3 LPU (LPX platform).
Following the leadership departures, Groq restructured under Executive Chairman Alex Davis and CEO Adam Winter. The company now operates as an independent NVIDIA Cloud Partner (NCP), shifting its engineering focus entirely to “inference-as-a-service.” Instead of fabricating Language Processing Units (LPUs), Groq now hosts medium and large Nvidia GPU clusters to serve external developer workloads.
Infrastructure and Capacity Scaling
The Series A brings Groq’s neocloud infrastructure funding to $1 billion since June 2026. The capital is strictly allocated to scaling the company’s physical data center footprint.
Groq currently operates 13 data centers distributed across North America, Europe, the Middle East, and Asia Pacific. The company plans to scale its total operational capacity from 54 megawatts to more than 200 megawatts by 2027. This infrastructure expansion is designed to support a customer base that the company reports at over 6 million developers and thousands of AI-native enterprise clients.
For engineering teams, Groq’s transition consolidates the hardware market while altering the execution environment for existing deployments. If your application relies on Groq’s APIs for high-speed AI inference, your production workloads are now executing on Nvidia systems rather than proprietary LPUs. You should audit your performance profiles to account for any latency or throughput variance introduced by the underlying hardware shift.
Get Insanely Good at AI
The book for developers who want to understand how AI actually works. LLMs, prompt engineering, RAG, AI agents, and production systems.
Keep Reading
How to launch Hugging Face models in SageMaker Studio
You will learn how to use the new Hugging Face integration to automatically provision and deploy open-source models directly into Amazon SageMaker Studio.
Etched Sohu Chip Hits 500K Llama Tokens/Sec in $10.3B Round
The hardware startup Etched secured a $300 million Series C to manufacture fixed-function inference ASICs optimized exclusively for transformer architectures.
$1B Series F Lands SambaNova SN50 Inference at JPMorgan
SambaNova Systems has raised a $1 billion Series F round at an $11 billion valuation to scale production of its SN50 enterprise AI inference infrastructure.
$1B in Transformer ASIC Orders Drives Etched to $5B Valuation
Silicon startup Etched has emerged from stealth with a $5 billion valuation and $1 billion in contracted revenue for its transformer-specific Sohu chips.
Groq Lands $650M to Scale Neocloud Inference Infrastructure
Following a $20 billion IP deal with Nvidia that drained its founding team, Groq has raised $650 million to rebuild as a dedicated inference cloud provider.