Companies

Groq

groq.com

Groq builds LPU-powered cloud infrastructure that delivers fast, affordable AI inference for developers and enterprises.

HQSan Jose, California, United States
Employees51-200
Funding$1.75B
Valuation$6.9B
Revenue$500M (2025)
1 active role
Profile 6mo agoJobs checked 56m ago
AI / MLAI ApplicationAPI / PlatformSeries D+$1B+

About

Groq builds specialized LPU hardware and GroqCloud infrastructure for real-time AI inference. It sells to developers, AI-native companies, and Fortune 500 enterprises, differentiating through high-speed, cost-efficient inference delivered through a global cloud platform.

Market

Groq competes in the AI-inference infrastructure and cloud market, spanning specialized accelerator hardware, inference-as-a-service, and developer access to hosted models. It differentiates through its proprietary LPU architecture and GroqCloud platform, emphasizing high throughput, low latency, affordability, and energy efficiency for real-time AI workloads. Its competitive set therefore includes both accelerator companies such as NVIDIA, Cerebras, and SambaNova and hyperscale cloud providers such as Google, AWS, and Microsoft.

Target Customers

Groq primarily targets AI developers and organizations—from startups and independent builders to large enterprises, including Fortune 500 companies—that need fast, affordable, low-latency inference for generative-AI applications. Its platform is aimed at technical teams and enterprise buyers deploying AI applications at scale.

At a Glance

Problem

Modern AI applications need to serve model outputs quickly, consistently, and at a controllable cost. Conventional infrastructure can make real-time inference expensive and operationally difficult, especially when every user interaction consumes tokens and latency directly affects the experience. Groq addresses the combined pain of response speed, energy efficiency, and predictable serving economics rather than model training itself. Its strongest use cases are latency-sensitive applications such as call-center and sales copilots, real-time customer experiences, voice interfaces, and AI agents.

The economic value is clearest when inference is continuous and interactive: faster responses can improve engagement, while predictable pricing avoids paying for idle capacity or absorbing large swings in infrastructure cost. Groq positions its offering around linear, predictable pricing with no hidden costs or idle infrastructure, making real-time AI more practical to deploy at scale.

Product / Service

Groq builds a purpose-designed Language Processing Unit, or LPU, together with the software stack needed to run AI inference. The company operates its own inference hardware and offers GroqCloud as a token-as-a-service platform, allowing developers to run open-source large language models through an API rather than purchasing and operating all the underlying compute. Groq provides the service in public, private, and co-cloud configurations, while also offering cloud and on-premise solutions for scaled deployments.

The product is optimized for fast, scalable inference: the LPU-based stack runs in data centers and returns low-latency responses, while GroqCloud gives developers access to performance and costs they can plan around. The benefit is a serving layer designed specifically for production AI workloads, particularly applications where response time, throughput, and predictable cost matter more than general-purpose compute flexibility.

Market

Groq competes in AI inference infrastructure, spanning specialized AI accelerators, inference clouds, and on-premise systems. Its direct hardware and infrastructure comparison set includes Cerebras and SambaNova, while Nvidia remains the dominant alternative through GPU-based and rack-scale inference systems. More broadly, Groq competes with traditional GPU infrastructure wherever developers choose how to run production models. Its differentiation is an inference-specific LPU architecture and an integrated cloud delivery model focused on speed, efficiency, and predictable economics.

Groq is operating at substantial commercial scale rather than appearing pre-launch: company materials report 13 data centers across North America, Europe, the Middle East, and Asia-Pacific, more than five million developers, thousands of AI-native companies, and trillions of tokens consumed each week. The available evidence does not disclose revenue, but the developer, enterprise, and infrastructure footprint demonstrates meaningful adoption of GroqCloud and its inference platform.

Founders & Leadership

Jonathan RossFounder
Founder; Chief Software Architect, NVIDIA
Doug WightmanFounder
Founder
Adam WinterChief Executive Officer
Matt EngChief Financial Officer

Funding History

2016-12
Series A$10.3M

Social Capital

2018-09
Series B$52.3M

Social Capital

2020-08
Series BUndisclosed

Not disclosed

2021-04
Series C$300M

Tiger Global Management, D1 Capital Partners, BlackRock

2024-08
Series D$640M

Cisco Investments, BlackRock

2025-09
Series E$750M

Disruptive

2026-06
Series E$650M

Disruptive, Infinitum Partners

Recent News

2026-06-22funding
Groq Raises $650M to Scale Its AI Inference Cloud Business

Groq announced $650 million in growth capital led by Disruptive and Infinitum to expand its global AI inference cloud. The company said it operates 13 data centers, serves more than five million developers, and is targeting 200 MW of capacity by the end of 2027.

2026-06-22
AI chipmaker Groq confirms $650M raise, re-staffs after Nvidia's $20B not-acqui-hire deal

TechCrunch reported that Groq raised $650 million and pivoted more heavily toward its neocloud inference business after its agreement with Nvidia. Groq also added executives including COO Alan Rice, CTO Sinclair Schuller, and CPO Rakesh Malhotra.

2025-12-24
Nvidia buying AI chip startup Groq for about $20 billion, its largest deal

CNBC reported that Nvidia was buying Groq's assets for approximately $20 billion, describing it as Nvidia's largest deal on record.

2025-12-24partnership
Groq and Nvidia Enter Non-Exclusive Inference Technology Licensing Agreement to Accelerate AI Inference at Global Scale

Groq announced a non-exclusive licensing agreement allowing Nvidia to use Groq's inference technology. Nvidia later incorporated the technology into its next-generation LPX platform, according to Groq.

2025-12-18partnership
Groq Partners with U.S. Department of Energy to Advance AI Inference and Next-Generation Computing Infrastructure

Groq announced a partnership with the U.S. Department of Energy focused on advancing AI inference and next-generation computing infrastructure.

2025-10-20partnership
IBM and Groq Partner to Accelerate Enterprise AI Deployment with Speed and Scale

IBM and Groq announced a strategic go-to-market and technology partnership intended to provide clients with access to Groq inference technology. The companies planned to integrate and enhance open-source vLLM with Groq's LPU architecture, alongside IBM Granite models.

2025-10-03
AI chip startup Groq plans to establish more than a dozen data centers next year

Data Center Dynamics reported that Groq had already established 12 data centers in 2025 and planned to add more than a dozen the following year, highlighting the company's infrastructure expansion.

2025-09-26partnership
McLaren Racing announces Groq as an Official Partner of the McLaren Formula 1 Team

McLaren Racing announced Groq as an official partner of its Formula 1 team, using Groq's inference capabilities for decision-making, analysis, development, and real-time insights.

2025-08-05
Groq Recognized in 2025 Gartner® Cool Vendor in AI

Groq announced that it had been recognized in Gartner's 2025 Cool Vendor report in AI.

Active Roles

1
United States/Engineering/4d ago

Business Model

Groq primarily monetizes cloud inference services, charging customers for AI model usage through token-based pricing for input and output. It also supports enterprise and on-premises AI deployments built on its hardware and software platform.

Products

Groq LPU Inference EngineGroqCloudLPU-based AI inference infrastructureGroqCloud developer playground and self-serve access

Customers

Solomei AIStackAIReBlinkRecallStats PerformMem0StashAutonomaCallimacusGPTZeroPerigonScreenAppUnifonicTenaliWillowPGA of AmericaaiXplainArgonne National LaboratoryBell

Tech Stack

Groq LPU (Language Processing Unit)GroqCloud inference cloudLarge language models (LLMs) and other leading modelsData-center AI inference infrastructureSelf-serve developer playground and inference-engine access

Competitors

NVIDIA
Cerebras
SambaNova
Google Cloud
Amazon Web Services (AWS)
Microsoft Azure
OpenRouter

Key Investors

Disruptive, BlackRock, Neuberger Berman, DTCP, Samsung Catalyst Fund, Cisco Investments, D1 Capital Partners, Altimeter, 1789 Capital, Infinitum