Groq Raises $650M to Scale Its AI Inference Cloud Business
- Karan Bhatia

- Jun 23
- 2 min read

Groq, building fast, low-cost inference, has announced $650 million in new growth capital to accelerate the expansion of its AI inference cloud. The round was led by Disruptive and Infinitum, with participation from investors who elected to reinvest in the company.
Groq’s current growth phase gained momentum following a licensing agreement with NVIDIA in late 2025 and the subsequent introduction of NVIDIA’s LPX platform, which incorporates Groq’s inference technology. These developments reinforced the company’s focus on a single objective: building a global AI inference cloud capable of serving the growing demand for large-scale AI workloads.
Today, Groq operates 13 data centers across North America, Europe, the Middle East, and Asia-Pacific, supporting millions of developers and thousands of AI-native companies that process trillions of tokens each week. The latest funding will be used to expand this infrastructure footprint, deploy next-generation LPX systems, and accelerate capacity growth. With plans to scale toward 200 MW by 2027, Groq is positioning itself to meet the rapidly increasing demand for AI inference infrastructure.
New Leadership Team.
Groq’s growth strategy is supported by deep expertise in AI inference, infrastructure operations, and enterprise software. The company believes its experience operating LPUs at scale provides a meaningful advantage as demand for AI inference infrastructure continues to accelerate.
To support its next phase of expansion, Groq has assembled a leadership team with backgrounds spanning hyperscale infrastructure, cloud platforms, enterprise software, and large-scale operations. This combination of technical, operational, and commercial experience is intended to strengthen execution as the company expands its global infrastructure footprint and accelerates adoption of its AI inference platform.
With expertise across inference operations, data center infrastructure, product development, and company building, the leadership team is positioned to guide Groq through its next stage of growth and establish a stronger presence in the rapidly evolving AI infrastructure market.
The Inference Opportunity.
As AI adoption expands, inference, the process of running AI models in production, is expected to become one of the largest segments of the AI infrastructure market. While much of the industry’s focus has historically been on model training, long-term demand is increasingly shifting toward the compute required to serve AI applications at scale.
Success in inference depends on delivering low-latency, reliable, and cost-efficient performance for real-world workloads. Groq has built its platform around these requirements, focusing on infrastructure designed specifically for high-volume inference. As AI moves from experimentation to widespread deployment, the company aims to capitalize on the growing demand for scalable and efficient inference infrastructure.


