Developer docs
API playgroundTry for free, no card

Search company profiles

Groq

Full company profile

uuid00000ri

Namestring
Groq
Legal namestring
Groq, Inc.
Websiteurl
groq.com
Company typeenum
Private
Founded yearint
2016
Descriptiontext

Groq, Inc. is a privately held AI infrastructure company founded in 2016 that designs the LPU (Language Processing Unit), a deterministic, SRAM-based, statically scheduled inference accelerator, and operates GroqCloud, a hyperscale AI inference cloud spanning 13 data centers across North America, Europe, the Middle East, and Asia-Pacific. The company's technical core is custom silicon purpose-built for low-latency, high-throughput LLM serving — featuring an on-chip SRAM memory architecture with up to ~7x the memory bandwidth of leading GPUs and a synchronous low-diameter interconnect that scales to thousands of chips. On top of the LPU, GroqCloud offers an OpenAI-compatible REST API, a developer/enterprise tiered product with a Free, Developer (pay-per-token), and Enterprise (custom commitment) plan, plus add-on modules including Compound agentic AI, Tool Use with built-in and MCP-compatible remote tools, Speech-to-Text (Whisper), Text-to-Speech (Orpheus), OCR/Vision, Reasoning, Content Moderation, Structured Outputs, Prompt Caching, and LoRA Inference, served across Performance, Flex, and Batch processing tiers.

In December 2025, Groq entered a ~$20 billion non-exclusive LPU technology licensing and acqui-hire arrangement with Nvidia: founder/CEO Jonathan Ross and president Sunny Madra departed to Nvidia while Groq retained full ownership of its LPU IP and continued independently under interim CEO Adam Winter and CFO Matt Eng. The company subsequently rebranded its strategy as 'Groq 2.0', pivoting from standalone chip sales to an AI inference neocloud model, closing a $650 million growth-capital round in June 2026 (led by Disruptive and Infinitum) to scale to 200MW of compute capacity by end-2027. Groq monetizes through usage-based developer pricing, prepaid Service Credits, enterprise subscription contracts, and — historically — chip and IP licensing. Its customer base spans 3M+ developers and thousands of AI-native companies, alongside named enterprise customers such as Dropbox, Vercel, Chevron, Volkswagen, Canva, Robinhood, Riot Games, Workday, Ramp, McLaren F1, and the PGA of America; Forbes reported revenue near $100 million at the time of the December 2025 Nvidia transaction.

Short descriptiontext

Groq designs the LPU (Language Processing Unit), a purpose-built inference accelerator, and operates GroqCloud, a hyperscale AI inference cloud across 13 global data centers serving 3M+ developers, thousands of AI-native companies, and enterprise customers including Dropbox, Chevron, Volkswagen, and Workday.

Operating statusenum
Operating
Ownership categoryenum
Headcount rangeband
101–250
akta.pro rankint
HeadquartersSan Jose, United States
HQ citystring
San Jose
HQ countrystring
United States
HQ regionstring
North America
Markets served

Serves global market

Offices5 records

Each record includes

City, Country, Type, Description, Source

Keyword5 values
AI inference cloud, custom inference silicon, LLM serving platform, language processing unit, neocloud infrastructure
Industry4 codes
1Model Governance, Risk & Compliance (GRC) Platforms
CodeHDAAAKAAPrimaryYes
2AI Governance, Risk & Compliance (GRC) Platforms
CodeHDAAAMAAPrimaryNo
3Enterprise AI Governance, Risk & Compliance Platforms (Model Risk, Audit, Policies)
CodeHDAEANAEPrimaryNo
4Cloud Compliance, Audit & Continuous Controls Monitoring (CCM/GRC)
CodeHDABAHAIPrimaryNo
NAICS code3 codes
  • Software Publishers513210
  • Computer Systems Design and Related Services54151
  • Computer Systems Design Services541512
SIC code2 codes
  • Services-Prepackaged Software7372
  • Services-Computer Integrated Systems Design7373
Product category
AI Inference Cloud Infrastructure
GTM motion5 records

Each record includes

Type, Description, Source

Revenue model6 records
1Developer tier — pay per token
TypeUsage Based
Description

Self-serve Developer plan charges customers per token consumed for LPU-powered inference, with higher token limits, chat support, and access to Flex and Batch processing tiers.

console.groq.com
2Enterprise tier — committed contracts
TypeSubscription Recurring
Description

Enterprise plan sold via Order Form with custom commitments, scalable capacity, dedicated support, LoRA Inference, SSO/SCIM, 90-day audit logs, and prepaid Service Credits; supports multi-year, committed-spend contracts.

console.groq.com
3Prepaid Service Credits
TypeUsage Based
Description

Customers can prepay for Cloud Services by purchasing Service Credits (with optional auto-reload Balance Maintenance), which can be redeemed against usage; Service Credits expire 1 year after purchase.

console.groq.com
4Free tier
TypeFreemium
Description

Free tier for developers to build and test on Groq APIs with community support; functions as a top-of-funnel acquisition channel.

groq.com
5Historical hardware / IP licensing (transitioning)
TypeLicensing Royalties
Description

Prior revenue included chip sales and a $20B non-exclusive licensing deal with Nvidia for Groq's LPU technology (December 2025); the company is now pivoting to AI inference cloud and infrastructure services rather than standalone chip sales.

larepublica.co
6Neocloud / AI inference cloud services
TypeManaged Services
Description

Strategic shift to operating 13 LPU-powered data centers across North America, Europe, the Middle East, and APAC, processing trillions of AI tokens weekly for developers and AI-native companies; aiming to scale to 200MW of capacity by end of 2027.

techmeme.com
Cost components6 values
Technology or R&D, Infrastructure, Personnel, Supply Chain, Operations, Marketing or Sales
Pricing details5 tiers
1Free — build and test on Groq APIs with community support
ModelFreemiumBilling cadencePay-as-you-go
Notes

$0; community support; intended for getting started.

console.groq.com
2Developer — pay per token, higher limits, chat support
ModelUsage-basedBilling cadencePay-as-you-go
Notes

Pay per token; higher token limits; chat support; access to Flex and Batch processing; spend limits; 7-day audit logs.

console.groq.com
3Enterprise — custom contract with dedicated support and SSO/SCIM
ModelSubscriptionBilling cadenceMulti-year contract
Notes

Everything in Developer plus scalable capacity, dedicated support, LoRA Inference, SSO & SCIM, 90-day audit logs; pricing negotiated via Order Form.

console.groq.com
4Batch processing — 50% discount, 24h-7d turnaround
ModelUsage-basedBilling cadencePay-as-you-go
Notes

Asynchronous Batch API returns completions within 24 hours to 7 days at a 50% discount versus on-demand pricing.

console.groq.com
5Prepaid Service Credits (with optional auto-reload)
ModelOtherBilling cadencePay-as-you-go
Notes

Prepaid credits redeemable against Cloud Services; expire 1 year after purchase/issuance; non-transferable; optional Balance Maintenance Service auto-reloads credits on a recurring basis.

console.groq.com
GTM typeB2B
B2B
Offering typeSoftware
Software
Brand1 of 4 records shown
1GroqCloud
Description

Cloud-based AI inference platform / console and APIs (including Groq Playground, GroqChat and the OpenAI-compatible Groq API) that runs on Groq's LPU-based stack.

groq.com
+3 more records
Core offering1 text field

Groq develops and operates the LPU (Language Processing Unit), a custom inference accelerator silicon pioneered in 2016, and delivers it as a vertically integrated AI inference cloud called GroqCloud. GroqCloud exposes an OpenAI-compatible API plus developer and enterprise consoles that serve open and proprietary LLMs, multimodal, speech, and agentic AI workloads across 13 global data centers with deterministic low latency.

Differentiator
Functional benefit
Problem solved
Quantifiable outcome1 of 5 values shown
  • 7.41x faster chat speed and 89% cost reduction, enabling a 3x increase in token consumption (Fintool)
+4 more records
Product overview1 text field

Groq is a single-vendor, vertically integrated AI inference platform: its purpose-built LPU (Language Processing Unit) silicon powers the GroqCloud developer and enterprise platform, which together form the core of its offering. The portfolio combines a hardware layer (LPU architecture, deployed across 13 data centers globally targeting 200MW by 2027) with a software/platform layer (GroqConsole, GroqCloud, GroqChat, Groq Playground) and a set of add-on modules — Tool Use (with built-in tools like Web Search, Code Execution, Wolfram Alpha, and Browser Search), Remote Tools and MCP, Compound (agentic AI), Speech to Text (Whisper), Text to Speech (Orpheus), OCR/Image Recognition, Reasoning, Content Moderation, Structured Outputs, Prompt Caching, and LoRA Inference. Service is offered in three billing tiers (Free, Developer, Enterprise) with Performance, Flex, and Batch processing tiers, and an OpenAI-compatible REST API plus Python/JS SDKs and a partner catalog of coding-agent integrations (Factory Droid, OpenCode, Kilo Code, Roo Code, Cline).

Product and service3 records
1GroqCloud
CategoryAI Inference Platform
Description

Cloud-based AI inference platform that runs LPU-powered inference across 13 data centers, serving developers and enterprises with low-latency, low-cost access to a catalog of open and proprietary LLMs, vision, speech, and agentic AI models via an OpenAI-compatible API.

2Groq LPU (Language Processing Unit)
CategoryAI Inference Silicon
Description

Custom-built inference accelerator chip with a deterministic, statically scheduled, SRAM-based architecture optimized for low-latency, high-throughput LLM and multimodal inference at scale.

3GroqCloud Console
CategoryDeveloper Platform Console
Description

Online developer and administrator console for managing GroqCloud accounts, organizations, API keys, billing, rate limits, usage, logs, and audit logs; includes GroqChat chat interface and Groq Playground prompt experimentation interface.

Scale indicator12 records

Each record includes

Type, Value, Description, Source

Partnership6 partners
Strategic tierStrategic Supplier — Exclusive Rack-Scale System Integrator For Groq Lpu-Based Products.TypeTechnology or IntegrationAnnounced on2026-04-01
Description

Selected as the exclusive rack-scale supplier for the Groq 3 LPX rack, which packs 256 language processing units per rack with 640 Tb/s scale-up bandwidth. Initial shipments of approximately 6,000 racks are scheduled for Q3 2026, followed by 10,000 racks in 2027. Foxconn plans to double its data center cabinet production capacity to 2,000 per week to support this and other demand.

Strategic tierFlagship Strategic Partner — Entered A ~$20B Non-Exclusive Licensing And Acqui-Hire Deal For Groq'S Lpu Technology.TypeStrategic or Co-development PartnerAnnounced on2025-12-01
Description

Signed a non-exclusive licensing agreement for Groq's LPU chip technology in December 2025 (reportedly $17B–$20B), and hired away Groq's founder/CEO Jonathan Ross, president Sunny Madra, and other key engineering talent. Groq retained full ownership of its LPU IP. Nvidia productized the licensed technology as the Groq 3 LPX / LP30, manufactured by Samsung on 4nm, and integrated it into its Vera Rubin rack-scale platform. The deal is under U.S. Senate antitrust scrutiny (Senators Warren and Blumenthal).

Strategic tierCore Manufacturing Partner — Fabricates Groq'S Lpu Silicon On Samsung'S 4Nm Process Node.TypeOEM/ Whitelabel/ Licensing PartnerAnnounced on2025-09-01
Description

Manufactures Groq's third-generation LPU chips at its Pyeongtaek campus using the 4nm (SF4) process. Samsung secured a deal to produce Groq's AI6-equivalent LPU for AI inference; Groq is one of the foundry customers (alongside Tesla) that has shifted to Samsung from TSMC for advanced-node AI chip production.

Strategic tierFlagship Partnership — Official Partner Of The Mclaren Formula 1 Team.TypeGTM or Marketing Partner
Description

Multi-year global partnership providing the McLaren F1 Team with Groq-powered inference for decision-making, analysis, development, and real-time insights; co-marketed through a 'Partnership Spotlight' on groq.com and the McLaren newsroom announcement.

Strategic tierEcosystem Technology Partner — Primary Agent Framework For Groq-Powered Apps.TypeTechnology or Integration
Description

Groq is featured as a first-class inference backend in LangChain/LangGraph tutorials, with developers using Groq's OpenAI-compatible API and llama-3.3-70b-versatile as the reasoning backbone for agentic research assistants with tool calling, sub-agents, and persistent memory.

Strategic tierStrategic Ecosystem Alignment — Groqcloud Is Openai-Compatible For Seamless Migration.TypeTechnology or Integration
Description

Groq's API is OpenAI-compatible (https://api.groq.com/openai/v1) and supports a Responses API mirroring OpenAI's, plus day-zero support for OpenAI open models; reduces switching friction to two lines of code.

Recent move8 records

Each record includes

Date, Type, Title, Description, Source

Expansion highlight6 records

Each record includes

Type, Description

Peers10 records
TypeDirect peer
Description

Direct peer building custom AI silicon and an inference cloud targeting low-latency LLM serving. Competes with Groq in the same inference-chip category and the same neocloud distribution model.

TypeDirect peer
Description

Direct peer offering a custom AI accelerator (RDU) and an enterprise AI cloud for inference. Competes with Groq in custom-silicon inference for enterprise and government customers.

TypeDirect peer
Description

Neocloud peer providing GPU-based AI compute and inference at scale. Competes with GroqCloud for the same AI-native enterprise workloads, though CoreWeave runs Nvidia GPUs rather than custom silicon.

TypeDirect peer
Description

Neocloud peer offering GPU clusters and on-demand inference for AI developers. Competes head-to-head with GroqCloud on pay-per-token inference pricing and developer self-service.

TypeDirect peer
Description

Neocloud peer providing hosted open-source model inference and fine-tuning infrastructure. Competes with GroqCloud for the same AI-native developer and enterprise inference workloads.

TypeDirect peer
Description

Inference-focused neocloud peer offering fast, low-cost open-model serving to developers. Directly comparable product proposition ("fast, low-cost inference") to GroqCloud's value proposition.

TypeBroad incumbent
Description

Broad incumbent in AI compute and now a strategic partner of Groq via the ~$20B LPU licensing/acqui-hire. Nvidia productized Groq's LPU IP inside its Vera Rubin platform, making it both licensor/partner and direct inference competitor.

TypeBroad incumbent
Description

Broad incumbent in generative AI and the API whose compatibility Groq emulates. Indirectly comparable as the dominant inference provider that Groq's OpenAI-compatible API is designed to siphon workloads from.

TypeBroad incumbent
Description

Frontier AI lab serving its own Claude models via cloud inference at scale. Comparable as a major hosted AI inference provider for enterprise and developer customers.

TypeEmerging player
Description

AI platform offering open-model hosting, inference endpoints, and an enterprise catalog. Comparable as an emerging inference cloud alternative for developers; competes with GroqCloud on open-model availability and developer mindshare.

Market position
Strengths5 records

Each record includes

Headline, Details, Source

Weaknesses5 records

Each record includes

Headline, Details, Source

Competitive moat6 records

Each record includes

Type, Details

Key risks7 records

Each record includes

Headline, Details, Source

Key highlights7 records

Each record includes

Headline, Details, Source

Customer concentration

Classification, Details

Named customers13 records

Each record includes

Name, Industry, Type, Use case, Source, UUID

Ideal customer profile2 records

Each record includes

Profile, Firmographic size, Sales motion, Sales cycle length, Buying structure, Purchase trigger, Buyer persona, Geography, Industry vertical, Primary use case, Description, Pain points, Evidence proof points, Target buyer

Technology focused
Yes
API detail
Has APIbool
Yes

Docs URL, Description

Integration9 records

Each record includes

Title, Type, Description, Source

AI capability14 records

Each record includes

Type, Description, Source

AI maturity
App detail

Has app

Feature10 records

Each record includes

Title, Differentiator, Description, Source

Core technology
Revenue estimate
Valuation estimate
Number of profiles
Profiles8 records

Each record includes

Name, Designation, Designation category, Overview, Profile commentary, Source

Subsidiaries3 records

Each record includes

Name, Acquired on, Relationship type, Type, Business focus

Compliance6 records

Each record includes

Name, Class, Description

Funding overview

Funding stage, Last funding date, Total funding USD

Funding rounds11 records

Each record includes

Round, Amount USD, Date, Pre money valuation, Total investors, Investors, News

Investors69 records

Each record includes

Name, Type, Date of entry, Rounds participated, Website

Funding detail is available on the Subscription and Enterprise plan.Contact sales →

M&A2 records

Each record includes

Name, Acquisition type, Announced date, Completed date, Status, Website, News

Investment

Each record includes

Name, Round, Announced date, Lead investor, Website, News

M&A and investment is available on the Subscription and Enterprise plan.Contact sales →

Groq

AI Inference Cloud Infrastructuregroq.com

Groq designs the LPU (Language Processing Unit), a purpose-built inference accelerator, and operates GroqCloud, a hyperscale AI inference cloud across 13 global data centers serving 3M+ developers, thousands of AI-native companies, and enterprise customers including Dropbox, Chevron, Volkswagen, and Workday.

What Groq does

Groq, Inc. is a privately held AI infrastructure company founded in 2016 that designs the LPU (Language Processing Unit), a deterministic, SRAM-based, statically scheduled inference accelerator, and operates GroqCloud, a hyperscale AI inference cloud spanning 13 data centers across North America, Europe, the Middle East, and Asia-Pacific. The company's technical core is custom silicon purpose-built for low-latency, high-throughput LLM serving — featuring an on-chip SRAM memory architecture with up to ~7x the memory bandwidth of leading GPUs and a synchronous low-diameter interconnect that scales to thousands of chips. On top of the LPU, GroqCloud offers an OpenAI-compatible REST API, a developer/enterprise tiered product with a Free, Developer (pay-per-token), and Enterprise (custom commitment) plan, plus add-on modules including Compound agentic AI, Tool Use with built-in and MCP-compatible remote tools, Speech-to-Text (Whisper), Text-to-Speech (Orpheus), OCR/Vision, Reasoning, Content Moderation, Structured Outputs, Prompt Caching, and LoRA Inference, served across Performance, Flex, and Batch processing tiers.

In December 2025, Groq entered a ~$20 billion non-exclusive LPU technology licensing and acqui-hire arrangement with Nvidia: founder/CEO Jonathan Ross and president Sunny Madra departed to Nvidia while Groq retained full ownership of its LPU IP and continued independently under interim CEO Adam Winter and CFO Matt Eng. The company subsequently rebranded its strategy as 'Groq 2.0', pivoting from standalone chip sales to an AI inference neocloud model, closing a $650 million growth-capital round in June 2026 (led by Disruptive and Infinitum) to scale to 200MW of compute capacity by end-2027. Groq monetizes through usage-based developer pricing, prepaid Service Credits, enterprise subscription contracts, and — historically — chip and IP licensing. Its customer base spans 3M+ developers and thousands of AI-native companies, alongside named enterprise customers such as Dropbox, Vercel, Chevron, Volkswagen, Canva, Robinhood, Riot Games, Workday, Ramp, McLaren F1, and the PGA of America; Forbes reported revenue near $100 million at the time of the December 2025 Nvidia transaction.

Groq firmographics

Firmographics
Name
Groq
Legal name
Groq, Inc.
Website
https://groq.com
Company type
Private
Founded year
2016
Operating status
Operating
Headcount range
101–250 employees
Short description
Groq designs the LPU (Language Processing Unit), a purpose-built inference accelerator, and operates GroqCloud, a hyperscale AI inference cloud across 13 global data centers serving 3M+ developers, thousands of AI-native companies, and enterprise customers including Dropbox, Chevron, Volkswagen, and Workday.
Ownership category
akta.pro rank

Groq industry classification

Industry
Product category
AI Inference Cloud Infrastructure
NAICS
Software Publishers (513210), Computer Systems Design and Related Services (54151), Computer Systems Design Services (541512)
SIC
Services-Prepackaged Software (7372), Services-Computer Integrated Systems Design (7373)
akta.pro primary industry
Model Governance, Risk & Compliance (GRC) Platforms (HDAAAKAA)
akta.pro secondary industries
AI Governance, Risk & Compliance (GRC) Platforms (HDAAAMAA), Enterprise AI Governance, Risk & Compliance Platforms (Model Risk, Audit, Policies) (HDAEANAE), Cloud Compliance, Audit & Continuous Controls Monitoring (CCM/GRC) (HDABAHAI)

Keywords

  • AI inference cloud
  • Custom inference silicon
  • LLM serving platform
  • Language processing unit
  • Neocloud infrastructure

Where Groq is headquartered

Location

Headquarters

HQ city
San Jose
HQ country
United States
HQ region
North America

Offices5 records

Markets served

Groq business model

Business model
GTM type
B2B
Offering type
Software
Cost components
Technology or R&D, Infrastructure, Personnel, Supply Chain, Operations, Marketing or Sales

Revenue model

  1. Developer tier — pay per token: Self-serve Developer plan charges customers per token consumed for LPU-powered inference, with higher token limits, chat support, and access to Flex and Batch processing tiers.
  2. Enterprise tier — committed contracts: Enterprise plan sold via Order Form with custom commitments, scalable capacity, dedicated support, LoRA Inference, SSO/SCIM, 90-day audit logs, and prepaid Service Credits; supports multi-year, committed-spend contracts.
  3. Prepaid Service Credits: Customers can prepay for Cloud Services by purchasing Service Credits (with optional auto-reload Balance Maintenance), which can be redeemed against usage; Service Credits expire 1 year after purchase.
  4. Free tier: Free tier for developers to build and test on Groq APIs with community support; functions as a top-of-funnel acquisition channel.
  5. Historical hardware / IP licensing (transitioning): Prior revenue included chip sales and a $20B non-exclusive licensing deal with Nvidia for Groq's LPU technology (December 2025); the company is now pivoting to AI inference cloud and infrastructure services rather than standalone chip sales.
  6. Neocloud / AI inference cloud services: Strategic shift to operating 13 LPU-powered data centers across North America, Europe, the Middle East, and APAC, processing trillions of AI tokens weekly for developers and AI-native companies; aiming to scale to 200MW of capacity by end of 2027.

Pricing tiers

ModelBillingPrice
FreemiumPay-as-you-goFree — build and test on Groq APIs with community support
Usage-basedPay-as-you-goDeveloper — pay per token, higher limits, chat support
SubscriptionMulti-year contractEnterprise — custom contract with dedicated support and SSO/SCIM
Usage-basedPay-as-you-goBatch processing — 50% discount, 24h-7d turnaround
OtherPay-as-you-goPrepaid Service Credits (with optional auto-reload)

Go-to-market motion5 records

Groq product offering

Product offering

Core offering

Groq develops and operates the LPU (Language Processing Unit), a custom inference accelerator silicon pioneered in 2016, and delivers it as a vertically integrated AI inference cloud called GroqCloud. GroqCloud exposes an OpenAI-compatible API plus developer and enterprise consoles that serve open and proprietary LLMs, multimodal, speech, and agentic AI workloads across 13 global data centers with deterministic low latency.

Product overview

Groq is a single-vendor, vertically integrated AI inference platform: its purpose-built LPU (Language Processing Unit) silicon powers the GroqCloud developer and enterprise platform, which together form the core of its offering. The portfolio combines a hardware layer (LPU architecture, deployed across 13 data centers globally targeting 200MW by 2027) with a software/platform layer (GroqConsole, GroqCloud, GroqChat, Groq Playground) and a set of add-on modules — Tool Use (with built-in tools like Web Search, Code Execution, Wolfram Alpha, and Browser Search), Remote Tools and MCP, Compound (agentic AI), Speech to Text (Whisper), Text to Speech (Orpheus), OCR/Image Recognition, Reasoning, Content Moderation, Structured Outputs, Prompt Caching, and LoRA Inference. Service is offered in three billing tiers (Free, Developer, Enterprise) with Performance, Flex, and Batch processing tiers, and an OpenAI-compatible REST API plus Python/JS SDKs and a partner catalog of coding-agent integrations (Factory Droid, OpenCode, Kilo Code, Roo Code, Cline).

Differentiator

Problem solved

Functional benefit

Brands

  • GroqCloud: Cloud-based AI inference platform / console and APIs (including Groq Playground, GroqChat and the OpenAI-compatible Groq API) that runs on Groq's LPU-based stack.
  • LPU (Language Processing Unit)
  • GroqCloud Console
  • Compound

Products and services

  • GroqCloud Cloud-based AI inference platform that runs LPU-powered inference across 13 data centers, serving developers and enterprises with low-latency, low-cost access to a catalog of open and proprietary LLMs, vision, speech, and agentic AI models via an OpenAI-compatible API.
  • Groq LPU (Language Processing Unit) Custom-built inference accelerator chip with a deterministic, statically scheduled, SRAM-based architecture optimized for low-latency, high-throughput LLM and multimodal inference at scale.
  • GroqCloud Console Online developer and administrator console for managing GroqCloud accounts, organizations, API keys, billing, rate limits, usage, logs, and audit logs; includes GroqChat chat interface and Groq Playground prompt experimentation interface.

Quantifiable outcome

  • 7.41x faster chat speed and 89% cost reduction, enabling a 3x increase in token consumption (Fintool)
  • +4 more outcomes

Companies that use Groq

Customer profile

Named customers13 records

Ideal customer profiles2 records

Groq technology and API

Technology

Technology focussed Yes

API detail

Has API
Yes
API docs
API detail

Core technology

AI maturity

App detail

Integration9 records

AI capability14 records

Feature10 records

Groq partnerships and signals

Strategic signal

Partnerships

Six partnerships are on record, tiered strategic supplier — exclusive rack-scale system integrator for groq lpu-based products., flagship strategic partner — entered a ~$20b non-exclusive licensing and acqui-hire deal for groq's lpu technology., core manufacturing partner — fabricates groq's lpu silicon on samsung's 4nm process node., flagship partnership — official partner of the mclaren formula 1 team., ecosystem technology partner — primary agent framework for groq-powered apps. and strategic ecosystem alignment — groqcloud is openai-compatible for seamless migration..

  • Foxconnstrategic supplier — exclusive rack-scale system integrator for groq lpu-based products.Technology or Integration · 1 April 2026Selected as the exclusive rack-scale supplier for the Groq 3 LPX rack, which packs 256 language processing units per rack with 640 Tb/s scale-up bandwidth. Initial shipments of approximately 6,000 racks are scheduled for Q3 2026, followed by 10,000 racks in 2027. Foxconn plans to double its data center cabinet production capacity to 2,000 per week to support this and other demand.
  • Nvidiaflagship strategic partner — entered a ~$20b non-exclusive licensing and acqui-hire deal for groq's lpu technology.Strategic or Co-development Partner · 1 December 2025Signed a non-exclusive licensing agreement for Groq's LPU chip technology in December 2025 (reportedly $17B–$20B), and hired away Groq's founder/CEO Jonathan Ross, president Sunny Madra, and other key engineering talent. Groq retained full ownership of its LPU IP. Nvidia productized the licensed technology as the Groq 3 LPX / LP30, manufactured by Samsung on 4nm, and integrated it into its Vera Rubin rack-scale platform. The deal is under U.S. Senate antitrust scrutiny (Senators Warren and Blumenthal).
  • Samsung Foundrycore manufacturing partner — fabricates groq's lpu silicon on samsung's 4nm process node.OEM/ Whitelabel/ Licensing Partner · 1 September 2025Manufactures Groq's third-generation LPU chips at its Pyeongtaek campus using the 4nm (SF4) process. Samsung secured a deal to produce Groq's AI6-equivalent LPU for AI inference; Groq is one of the foundry customers (alongside Tesla) that has shifted to Samsung from TSMC for advanced-node AI chip production.
  • McLaren Racing (McLaren F1 Team)flagship partnership — official partner of the mclaren formula 1 team.GTM or Marketing PartnerMulti-year global partnership providing the McLaren F1 Team with Groq-powered inference for decision-making, analysis, development, and real-time insights; co-marketed through a 'Partnership Spotlight' on groq.com and the McLaren newsroom announcement.
  • LangChain (LangGraph)ecosystem technology partner — primary agent framework for groq-powered apps.Technology or IntegrationGroq is featured as a first-class inference backend in LangChain/LangGraph tutorials, with developers using Groq's OpenAI-compatible API and llama-3.3-70b-versatile as the reasoning backbone for agentic research assistants with tool calling, sub-agents, and persistent memory.
  • OpenAI (API compatibility)strategic ecosystem alignment — groqcloud is openai-compatible for seamless migration.Technology or IntegrationGroq's API is OpenAI-compatible (https://api.groq.com/openai/v1) and supports a Responses API mirroring OpenAI's, plus day-zero support for OpenAI open models; reduces switching friction to two lines of code.

Scale indicators12 records

Recent moves8 records

Expansion highlights6 records

Groq competitors and assessment

Company assessment

Direct peers

  • Cerebras Systems: Direct peer building custom AI silicon and an inference cloud targeting low-latency LLM serving. Competes with Groq in the same inference-chip category and the same neocloud distribution model.
  • SambaNova Systems: Direct peer offering a custom AI accelerator (RDU) and an enterprise AI cloud for inference. Competes with Groq in custom-silicon inference for enterprise and government customers.
  • CoreWeave: Neocloud peer providing GPU-based AI compute and inference at scale. Competes with GroqCloud for the same AI-native enterprise workloads, though CoreWeave runs Nvidia GPUs rather than custom silicon.
  • Lambda Labs: Neocloud peer offering GPU clusters and on-demand inference for AI developers. Competes head-to-head with GroqCloud on pay-per-token inference pricing and developer self-service.
  • Together AI: Neocloud peer providing hosted open-source model inference and fine-tuning infrastructure. Competes with GroqCloud for the same AI-native developer and enterprise inference workloads.
  • Fireworks AI: Inference-focused neocloud peer offering fast, low-cost open-model serving to developers. Directly comparable product proposition ("fast, low-cost inference") to GroqCloud's value proposition.

Broad incumbents

  • Nvidia: Broad incumbent in AI compute and now a strategic partner of Groq via the ~$20B LPU licensing/acqui-hire. Nvidia productized Groq's LPU IP inside its Vera Rubin platform, making it both licensor/partner and direct inference competitor.
  • OpenAI: Broad incumbent in generative AI and the API whose compatibility Groq emulates. Indirectly comparable as the dominant inference provider that Groq's OpenAI-compatible API is designed to siphon workloads from.
  • Anthropic: Frontier AI lab serving its own Claude models via cloud inference at scale. Comparable as a major hosted AI inference provider for enterprise and developer customers.

Emerging players

  • Hugging Face: AI platform offering open-model hosting, inference endpoints, and an enterprise catalog. Comparable as an emerging inference cloud alternative for developers; competes with GroqCloud on open-model availability and developer mindshare.

Market position

Strengths5 records

Weaknesses5 records

Competitive moat6 records

Key risks7 records

Key highlights7 records

Customer concentration

Groq social profiles

Digital presence

Groq compliance and trust

Trust signal

Compliance6 records

Groq financial estimates

Financial estimate

Revenue estimate

Valuation estimate

Groq leadership team

Management profile

Number of profiles

Profiles8 records

Groq subsidiaries and ownership

Company hierarchy

Subsidiaries3 records

Groq funding detail

Funding detail

Funding overview

Funding rounds11 records

Investors69 records

Funding detail is available on the Subscription and Enterprise plan.Contact sales →

Groq M&A and investment

M&A and investment

M&A2 records

Investments

M&A and investment is available on the Subscription and Enterprise plan.Contact sales →

Frequently asked questions about Groq

What does Groq do?

Groq develops and operates the LPU (Language Processing Unit), a custom inference accelerator silicon pioneered in 2016, and delivers it as a vertically integrated AI inference cloud called GroqCloud. GroqCloud exposes an OpenAI-compatible API plus developer and enterprise consoles that serve open and proprietary LLMs, multimodal, speech, and agentic AI workloads across 13 global data centers with deterministic low latency.

Is Groq a public or private company?

Groq is a private company. It is classified as venture growth investor backed and is currently operating.

When was Groq founded?

Groq was founded in 2016. It employs 101 to 250 people.

Where is Groq based?

Groq is headquartered in San Jose, United States, in the North America region.

How does Groq make money?

Six revenue lines are on record. Developer tier — pay per token is the primary driver. The others are enterprise tier — committed contracts, prepaid Service Credits, free tier, historical hardware / IP licensing (transitioning) and neocloud / AI inference cloud services.

Who are Groq's main competitors?

Direct peers on record are Cerebras Systems, SambaNova Systems, CoreWeave, Lambda Labs, Together AI and Fireworks AI. Broad incumbents are Nvidia, OpenAI and Anthropic. Hugging Face is listed as an emerging player.

Does Groq have an API?

Yes. Groq offers a public, OpenAI-compatible REST API at https://api.groq.com/openai/v1, enabling developers to integrate fast LPU-accelerated inference for text generation, speech-to-text, text-to-speech (Orpheus), vision/OCR, reasoning, and agentic AI into their applications. The API is OpenAI-compatible, allowing integration with just two lines of code, and is accessible via a free API key, a Developer pay-per-token tier, and an Enterprise tier with scalable capacity, dedicated support, LoRA inference, SSO, and SCIM. It supports tool use, structured outputs, prompt caching, and remote tools/MCP. Official SDKs and code samples are available for Python (pip install groq) and JavaScript. Developer documentation is at console.groq.com/docs/overview.

What industry is Groq in?

Groq's product category is AI Inference Cloud Infrastructure. Its primary akta.pro industry code is HDAAAKAA, Model Governance, Risk & Compliance (GRC) Platforms, with a secondary code of HDAAAMAA, AI Governance, Risk & Compliance (GRC) Platforms. Its NAICS code is 513210 and its SIC code is 7372.

Unlock the full company data

50 free credits on sign-up, no credit card required.

Contact sales
Live signals
The National Law ReviewNvidia-Groq’s Business Deal Draws Delaware Fiduciary-Duty SuitTwo former Groq engineers sued Groq's directors and officers in Delaware, challenging Nvidia's $20 billion licensing-and-hiring deal. They allege the arrangement transferred core technology and workforce without a shareholder vote, and that insiders received benefits unavailable to outside shareholders. The suit follows a reported DOJ antitrust probe.Law360Groq Investors Sue Over Nvidia's $20B 'Reverse Acqui-Hire'Two former Groq Inc. stockholders sued Groq's directors and a former officer in Delaware Chancery Court, alleging an improper $20 billion reverse acqui-hire to Nvidia Corp. The complaint claims the deal transferred Groq's technology and engineering workforce without a stockholder vote or a process aimed at securing the best price.Seeking AlphaGroq-Nvidia $20B deal faces suit alleging stockholders were shortchanged: report (NVDA:NASDAQ)Two former Groq engineers sued Groq's board in Delaware, alleging the $20B Nvidia deal undervalued the company and bypassed a shareholder vote. The lawsuit claims the board failed to maximize asset value. The deal also faces U.S. Justice Department antitrust scrutiny.MarketScreenerNvidia's $20bn deal with Groq challenged in courtTwo engineers and shareholders sued Groq over its $20bn licensing deal with Nvidia, alleging improper surrender of technology assets and failure to seek best terms. Groq rejected the claims, calling the lawsuit without merit and defending the deal's value.The Times of IndiaNvidia Groq deal lawsuit: Nvidia’s $20 billion Groq deal under fire: Why shareholders are suing - The Economic TimesTwo former Groq engineers sued Nvidia in Delaware, alleging the company's 2025 acqui-hiring deal transferred key technology and employees, leaving remaining shareholders with a weaker company. The lawsuit claims shareholders were inadequately compensated while select employees received substantial payouts, with ongoing antitrust investigations.CoinMarketCapGroq shareholders sue Nvidia over the structure of their $20 billion deal: Guest Post by Cryptopolitan_NewsTwo ex-Groq engineers filed a class action against Groq's board, alleging the $20 billion Nvidia licensing deal was inequitable. The deal split $17 billion in licensing fees and $3 billion in stock bonuses, leaving remaining Groq valued at $3.5 billion. The case adds shareholder litigation to existing antitrust scrutiny.CryptopolitanGroq shareholders sue Nvidia over the structure of their $20 billion dealTwo former Groq engineers sued the startup's board and Nvidia over the structure of their roughly $20 billion licensing-and-hiring deal, alleging remaining shareholders were short-changed. The suit alleges the board was conflicted and denied some investors a vote, with the deal split into a $17 billion licensing fee and a $3 billion stock bonus pool. The outcome could set a precedent for AI giants acquiring rivals' technology and talent.EntrepreneurEveryone Wants to Build AI Apps — I’m Betting on What Powers Them. Here’s Why.The author argues that inference-focused AI chipmakers like Cerebras, Groq, and Tenstorrent hold the real value, not the app layer. Cerebras' IPO raised $5.55 billion, and Groq doubled down on inference after a $20 billion Nvidia deal. The author holds stakes in these companies.QuartzEx-Groq engineers sue board over $20B Nvidia asset dealTwo former Groq engineers sued the board over the $20 billion Nvidia deal, alleging a low price and bypassed stockholder vote. The complaint says $17 billion went to a non-exclusive license and $3 billion to restricted stock units, with 150-200 engineers transitioning. Groq defended the deal and will vigorously defend the lawsuit.CNBCNvidia's $20 billion Groq deal faces lawsuit alleging startup's stockholders were shortchangedA lawsuit filed in Delaware alleges Nvidia's $20 billion acquisition of Groq's assets squeezed out stockholders, offering a lowball price. The suit claims the board sold the company without a required stockholder vote and conflicted, costing Groq stockholders billions. Groq denies the claims, calling the lawsuit meritless.