← All organizations

AI inference infrastructure

Groq

Groq operates an AI inference cloud for developers and enterprises running AI models. Its platform combines three layers: GroqMetal supplies dedicated bare-metal infrastructure, GroqCore adds a production inference stack, and GroqAssured adds enterprise governance and auditability. Customers can control dedicated computing capacity or use the inference stack without managing the underlying infrastructure themselves.

Founded by Jonathan Ross and Douglas Wightman, Groq grew from custom chip development into cloud infrastructure. Adam Winter became CEO in 2026. Its Language Processing Unit, or LPU, uses a compiler to schedule computation and data movement across chips, with deterministic execution and on-chip SRAM. This software-controlled architecture makes execution timing predictable and places memory alongside compute, rather than relying on runtime scheduling to coordinate each operation.

In December 2025, NVIDIA entered a non-exclusive licensing agreement for Groq’s LPU technology and hired certain employees. The transaction involved $17 billion in consideration, including imputed interest; NVIDIA purchased no equity interests, existing products or customer contracts. Groq continues operating its cloud business. In August 2026, the company reported more than six million developers, thousands of AI-native companies and Fortune 500 enterprise customers, with 13 data centers across North America, Europe, the Middle East and Asia-Pacific.

groq.com

1 talk

Newest first

1 speaker at AIE

Affiliations reflect their AIE appearances, not necessarily current employment.

Messages from the stage

Inference speed through a processor-race analogy

Madra compared inference improvements with the historical race toward gigahertz microprocessors, citing a reported greater-than-50% speed increase for Llama 3 8B between April and June.

Affiliations reflect each recorded session, not necessarily current employment.

Company sources · checked 2026-08-28