GroqCloud
A managed inference cloud offering API access to leading open models with fast token generation, serving developers and enterprises that need low-latency, cost-efficient inference at scale.
Groq is an AI inference company founded in 2016 by former Google engineers. It pioneered the Language Processing Unit (LPU), a deterministic accelerator designed for running large language models with very low latency, and operates GroqCloud, an inference cloud used by more than five million developers. In December 2025 NVIDIA licensed Groq's inference technology on a non-exclusive basis, and Groq continues as an independent inference cloud business.
Groq is an AI inference company founded in 2016 by former Google engineers. It pioneered the Language Processing Unit (LPU), a deterministic accelerator designed for running large language models with very low latency, and operates GroqCloud, an inference cloud used by more than five million developers. In December 2025 NVIDIA licensed Groq's inference technology on a non-exclusive basis, and Groq continues as an independent inference cloud business.
Groq runs 13 data centres across North America, Europe, the Middle East and Asia-Pacific, including a facility in Dammam, Saudi Arabia, that has served production traffic since early 2025 under a USD 1.5 billion agreement with the Kingdom. Groq is an official inference provider to HUMAIN, giving regional enterprises and government bodies access to high-speed inference with in-country data processing.
A managed inference cloud offering API access to leading open models with fast token generation, serving developers and enterprises that need low-latency, cost-efficient inference at scale.
Purpose-built inference accelerators with on-chip SRAM and deterministic execution; Groq is among the first providers bringing the NVIDIA Groq 3 LPX platform to market alongside Vera Rubin systems.
Dedicated bare-metal inference capacity tuned for speed and reliability, giving organisations full control over their environment for sovereign, regulated or high-volume workloads such as national AI services.
A tested inference stack that turns dedicated capacity into production-ready performance without in-house infrastructure expertise, plus enterprise-grade governance, auditability and control for organisations with strict compliance requirements.
Regional inference capacity delivered with partners such as HUMAIN in Saudi Arabia, supporting Arabic language models and day-zero availability of new open models for local enterprises and government.
Groq's entire business is AI inference. Its LPU architecture avoids the caches and branch prediction of general-purpose processors, delivering predictable, high-throughput token generation for large language models. The company positions inference, not training, as the bottleneck for AI adoption, and is scaling capacity from 57 MW toward more than 200 MW to serve real-time and agentic workloads.
GroqCloud hosts leading open-weight models and combines LPUs with NVIDIA GPUs in disaggregated configurations following Groq's admission to the NVIDIA Cloud Partner programme in 2026. In Saudi Arabia, Groq's Dammam facility powers HUMAIN's inference services, including HUMAIN One, and delivered OpenAI's gpt-oss open models locally on day one.
Where Groq fits in the four-layer AI stack we bring to partners.
Explore the layerLed by Disruptive, the round funds expansion of Groq's inference cloud, which now combines Groq 3 LPUs with NVIDIA GPUs across 13 data centres, following Groq's admission to the NVIDIA Cloud Partner programme.
Growth capital led by Disruptive and Infinitum accelerates expansion of GroqCloud toward 200 MW of capacity by 2027, serving Fortune 500 enterprises and more than five million developers.
NVIDIA licensed Groq's inference technology on a non-exclusive basis; founder Jonathan Ross and other leaders joined NVIDIA, while Groq continues as an independent company operating GroqCloud without interruption for customers.
Groq's inference infrastructure underpins HUMAIN One, HUMAIN's real-time AI operating system for enterprises, extending the partnership under which Groq serves as an official inference provider in Saudi Arabia.
Architecture, sizing and proof-of-concept support so partners propose the right configuration first time.
02Certification paths, workshops and demo access that build partner capability around the portfolio.
03Local stock, professional services, escalation and RMA handling across Saudi Arabia and the region.
Speak with our team about availability, solution design and partner enablement.