Developer docs
API playgroundTry for free, no card

Search company profiles

Inception

Full company profile

uuid00007t4

Namestring
Inception
Legal namestring
Inception AI, Inc.
Company typeenum
Private
Founded yearint
2024
Descriptiontext

Inception AI, Inc. is a Palo Alto-based AI company founded in 2024 by Stanford professor Stefano Ermon that develops diffusion-based large language models (dLLMs), an alternative architecture to traditional autoregressive LLMs. The company's Mercury product family generates text and code through parallel token refinement rather than sequential decoding, achieving approximately 1,000 tokens per second on NVIDIA Blackwell GPUs—roughly 10x faster than leading autoregressive models—while maintaining competitive benchmark quality (e.g., AIME 2025: 91.1%, GPQA: 73.6%) at substantially lower inference cost. The portfolio spans Mercury 2 (reasoning model with 128K context), Mercury Edit 2 (next-edit prediction for code), Mercury Coder (code generation), and the general Mercury chat model, with a research-stage multimodal extension called LaViDa. The team includes researchers from Stanford, UCLA, Cornell, Google DeepMind, Meta AI, Microsoft AI, and OpenAI, and the founders are credited with inventing foundational techniques including diffusion models, Flash Attention, and Direct Preference Optimization.

The product is distributed through an OpenAI-compatible REST API, the Mercury Chat playground, AWS Bedrock, Azure AI Foundry (with SOC2 and HIPAA compliance), OpenRouter, and Models.dev, with deep integrations into developer tooling (Zed, Continue, ProxyAI, Augment Code) and the Microsoft NLWeb open project. Inception generates revenue primarily through usage-based API pricing ($0.25 per 1M input tokens, $0.75 per 1M output tokens) supplemented by enterprise licensing on cloud marketplaces and custom contracts with dedicated capacity and SLAs. A 10-million-token free tier supports a self-serve developer acquisition motion running in parallel with a dedicated enterprise sales motion. Customer logos span coding tools (Zed, Augment Code, Buildglare, ProxyAI), enterprise search (SearchBlox), AI agent platforms (Radient, Skyvern), and real-time voice AI (Happyverse AI, OpenCall), with strategic partnerships extending into financial services (Bain & Company, Kensho), Arabic-language AI (Cerebras, MBZUAI, Jais 2), and government/security (Brain Co., Mirror Security/G42).

In November 2025, Inception raised a $50 million seed round led by Menlo Ventures with participation from Microsoft M12, NVentures, Snowflake Ventures, Databricks Ventures, Mayfield, and Innovation Endeavors, plus angels Andrew Ng and Andrej Karpathy. As of mid-2026, the company was reportedly in acquisition discussions with Microsoft at a valuation exceeding $1 billion.

Short descriptiontext

Inception develops diffusion-based large language models (Mercury family) that generate tokens in parallel for ~10x faster inference than autoregressive LLMs, serving developers, voice AI platforms, enterprise search, and agent platforms via API and cloud marketplace channels.

Operating statusenum
Operating
Ownership categoryenum
Headcount rangeband
11–50
akta.pro rankint
HeadquartersPalo Alto, United States
HQ citystring
Palo Alto
HQ countrystring
United States
HQ regionstring
North America
Markets served

Serves global market

Offices1 record

Each record includes

City, Country, Type, Description, Source

Keyword5 values
diffusion language models, generative AI foundation models, AI inference API, code generation AI, reasoning LLM platform
Industry5 codes
1Multimodal Generation (Text-Image-Video-Audio) & Creative Toolchains
CodeHDAAACAGPrimaryYes
2LLMOps & Generative AI Platforms (Prompt/Agent Orchestration, RAG)
CodeHDAEANADPrimaryNo
3Fine-Tuning, Adaptation & Custom Model Training (PEFT/LoRA/RLHF)
CodeHDAAACACPrimaryNo
4Machine Translation & Multilingual NLP
CodeHDAAADADPrimaryNo
5MLOps/LLMOps & Model Lifecycle Management Services
CodeBPAEAHAHPrimaryNo
NAICS code3 codes
  • Software Publishers5132
  • Software Publishers513210
  • Computer Systems Design and Related Services5415
SIC code1 code
  • Services-Prepackaged Software7372
Product category
AI Foundation Models
GTM motion3 records

Each record includes

Type, Description, Source

Revenue model3 records
1API Usage-Based Pricing
TypeUsage Based
Description

Pay-per-token pricing model for Mercury and Mercury Coder models through Inception API platform, with differentiated pricing for input tokens, output tokens, and cached inputs

inceptionlabs.ai
2Enterprise Licensing
TypeSubscription Recurring
Description

Enterprise deployment licensing through AWS Bedrock and Azure AI Foundry with software license costs (e.g., $0.78/hour for Mercury on Azure) plus cloud compute costs

inceptionlabs.ai
3Custom Enterprise Contracts
TypeSubscription Recurring
Description

Custom rate limits, SLA guarantees, volume-based pricing, dedicated capacity, and custom terms for large enterprise deployments

inceptionlabs.ai
Marketing channels9 records

Each record includes

Title, Type, Stage, Description, Source

Distribution channels5 records

Each record includes

Title, Type, Scope, Target buyer, Description, Source

Pricing details4 tiers
1Free tier with 10 million tokens
ModelFreemiumBilling cadencePay-as-you-go
Notes

New users receive 10 million free tokens to try all models (Mercury 2, Mercury Edit 2)

inceptionlabs.ai
2Developer tier with usage-based pricing
ModelUsage-basedBilling cadencePay-as-you-go
Notes

Usage-based pricing with generous rate limits and priority support

inceptionlabs.ai
3Enterprise tier with custom pricing and SLAs
ModelSubscriptionBilling cadenceAnnual
Notes

Custom rate limits, SLA guarantees (99.5%+ uptime), security and privacy features, volume-based pricing, private networking options, dedicated capacity

inceptionlabs.ai
4Azure AI Foundry enterprise deployment
ModelUsage-basedBilling cadencePay-as-you-go
Notes

$0.78/hour Mercury software license, with compute costs billed separately through Azure account

inceptionlabs.ai
GTM typeB2B
B2B
Offering typeSoftware
Software
Core offering1 text field

Inception builds and deploys diffusion-based large language models (dLLMs) under the Mercury family that generate text and code through parallel iterative refinement rather than sequential autoregressive decoding. Their models deliver 5-10x faster inference at roughly half the cost of conventional LLMs while maintaining frontier-class quality, with distribution through a self-serve OpenAI-compatible API and enterprise deployments on AWS Bedrock and Azure AI Foundry.

Differentiator
Functional benefit
Problem solved
Quantifiable outcome1 of 6 values shown
  • 1,000 tokens/second throughput vs ~71-89 tokens/second for GPT-5 Mini and Claude 4.5 Haiku
+5 more records
Product overview1 text field

Inception is an AI research and product company developing diffusion-based large language models (dLLMs) that generate text and code through parallel refinement rather than sequential token-by-token generation. The product portfolio centers on the Mercury family of diffusion LLMs: Mercury 2 serves as the flagship reasoning model with 128K context, Mercury Edit 2 provides next-edit prediction for coding workflows, Mercury Coder powers code generation and apply-edit capabilities, and the general chat Mercury model enables conversational AI. The Inception API (OpenAI-compatible) and Mercury Chat playground provide developer access, while enterprise deployment options include AWS Bedrock, Azure AI Foundry, and model routers like OpenRouter. The company also partners with Bain & Company, Kensho, and Brain Co. for enterprise solutions, and collaborates on specialty models like Jais 2 (Arabic LLM with Cerebras/MBZUAI). Mercury's diffusion approach delivers 5-10x faster inference than traditional autoregressive models while maintaining competitive quality, enabling use cases in coding assistants, real-time voice agents, agentic loops, and enterprise search.

Product and service8 records
1Mercury 2
CategoryGenerative AI Model
Description

The flagship diffusion-based reasoning LLM from Inception achieving approximately 1,009 tokens per second on NVIDIA Blackwell GPUs through parallel token generation. Features 128K context window, tunable reasoning, native tool use, and schema-aligned JSON output for developers and enterprises building latency-sensitive AI applications.

2Mercury Edit 2
CategoryGenerative AI Model
Description

A small, coding-focused diffusion LLM optimized for next-edit prediction in coding workflows. Uses recent edits and codebase context to predict what developers will change next, delivering fast autocomplete and apply-edit suggestions.

3Mercury Coder
CategoryGenerative AI Model
Description

The first commercial-scale diffusion LLM optimized for code generation, achieving 1000+ tokens per second on NVIDIA H100s. Supports fill-in-the-middle (FIM) and apply-edit capabilities for coding workflows.

4Mercury Coder Small
CategoryGenerative AI Model
Description

A smaller coding-focused diffusion model variant available via API, running more than 5x faster than speed-optimized frontier models like GPT-4o Mini and Claude 3.5 Haiku while matching them in quality. Supports 32K context window and fill-in-the-middle workflows.

5Mercury (General Chat)
CategoryGenerative AI Model
Description

First general chat diffusion LLM that matches the performance of speed-optimized frontier models like GPT-4.1 Nano and Claude 3.5 Haiku while running over 7x faster. Powers conversational AI applications, enterprise search, and voice interfaces.

6Inception API Platform
CategoryDeveloper Platform
Description

REST API providing programmatic access to Mercury diffusion LLMs. OpenAI API compatible with support for chat completions, fill-in-the-middle, and apply-edit endpoints. New accounts receive 10 million free tokens.

7Mercury Chat Playground
CategoryDeveloper Platform
Description

Interactive web-based playground for testing Mercury models. Allows developers to explore model capabilities and experiment with diffusion-based text generation.

8Mercury on Azure AI Foundry
Scale indicator6 records

Each record includes

Type, Value, Description, Source

Partnership17 partners
Strategic tierFlagshipTypeTechnology or IntegrationAnnounced on2026-05-12
Description

Production deployment of Mercury 2 for context compaction, model routing, and tool search; achieved 82% latency reduction and 90% cost savings

Strategic tierCoreTypeTechnology or IntegrationAnnounced on2026-03-30
Description

Mercury Edit 2 integrated into Zed's edit prediction system for fast code editing

Strategic tierFlagshipTypeTechnology or IntegrationAnnounced on2026-01-12
Description

Mercury-powered SearchAI for sub-second GenAI search with 60-90% lower inference costs across enterprise workloads

Strategic tierFlagshipTypeStrategic or Co-development PartnerAnnounced on2025-12-09
Description

Co-developed Jais 2 Arabic LLM (70B parameters) trained on largest Arabic-first dataset; Cerebras provides wafer-scale computing infrastructure for training

Strategic tierFlagshipTypeStrategic or Co-development PartnerAnnounced on2025-12-09
Description

Co-developed Jais 2 Arabic LLM for high fluency, cultural depth, and safety in Arabic language AI

Strategic tierStrategicTypeStrategic or Co-development PartnerAnnounced on2025-11-24
Description

Strategic partnership with G42's Inception to co-develop AI security products for government and enterprise clients

Strategic tierFlagshipTypeTechnology or IntegrationAnnounced on2025-11-18
Description

Mercury available on Azure AI Foundry with enterprise-grade infrastructure, compliance standards (SOC2, HIPAA), and $0.78/hour software license

Strategic tierCoreTypeTechnology or IntegrationAnnounced on2025-10-27
Description

Partnership to bring faster, more consistent code edits to ProxyAI platform

Strategic tierStrategicTypeImplementation/ SI/ Consulting PartnerAnnounced on2025-10-20
Description

Strategic collaboration to develop AI solutions for financial institutions at GITEX Global 2025

Strategic tierStrategicTypeImplementation/ SI/ Consulting PartnerAnnounced on2025-10-14
Description

Strategic collaboration to develop and deploy AI solutions for financial services, starting with (In)Alpha for investment workflows

Strategic tierStrategicTypeStrategic or Co-development PartnerAnnounced on2025-10-14
Description

Strategic partnership to develop next-generation AI enterprise applications for public services, healthcare, and energy sectors

Strategic tierCoreTypeOthersAnnounced on2025-10-13
Description

Inception is a G42 company with access to G42's AI infrastructure and regional presence in Middle East

Strategic tierFlagshipTypeTechnology or IntegrationAnnounced on2025-08-27
Description

Mercury and Mercury Coder available on Amazon Bedrock Marketplace and SageMaker JumpStart for enterprise deployment

Strategic tierCoreTypeTechnology or IntegrationAnnounced on2025-08-27
Description

IDE integration partnership with Continue extension for Mercury Coder autocomplete and next-edit features

Strategic tierCoreTypeTechnology or IntegrationAnnounced on2025-08-19
Description

Mercury powering model routing for Radient Automatic, enabling sub-second classification and 70% cost savings

16Buildglare
Strategic tierCoreTypeTechnology or IntegrationAnnounced on2025-07-06
Description

Mercury Coder powering Apply-Edit functionality for low-code web development platform

inceptionlabs.ai
Strategic tierFlagshipTypeTechnology or IntegrationAnnounced on2025-05-15
Description

Founding LLM partner for Microsoft's NLWeb open project, enabling natural language interfaces for websites with ultra-fast Mercury dLLM

Recent move6 records

Each record includes

Date, Type, Title, Description, Source

Expansion highlight6 records

Each record includes

Type, Description

Peers10 records
TypeDirect peer
Description

Frontier foundation model provider (GPT-4o, GPT-5 family) competing directly with Mercury's reasoning and coding capabilities. Inception positions Mercury 2 as '2x faster than GPT-5.2' per customer testimonials, and offers an OpenAI-compatible API as a drop-in alternative.

TypeDirect peer
Description

Enterprise-focused foundation model provider (Claude family). Anthropic's Claude Haiku 4.5 is benchmarked head-to-head with Mercury 2 on speed, and Inception's enterprise GTM via AWS Bedrock and Azure Foundry competes directly with Anthropic's distribution footprint.

TypeBroad incumbent
Description

Develops Gemini family of foundation models with broad modality support (text, image, video, audio). A broad incumbent whose research output (including diffusion-based image/video models) competes architecturally with Inception's diffusion-based language approach.

TypeDirect peer
Description

European foundation model lab offering open and commercial LLMs with emphasis on efficiency and inference cost — a similar value proposition to Inception's Mercury dLLMs targeting latency-sensitive and cost-sensitive enterprise workloads.

TypeDirect peer
Description

Enterprise-focused foundation model provider offering API-accessible LLMs with private deployment and RAG tooling. Competes with Inception for enterprise search, retrieval, and agentic use cases through similar cloud marketplace channels.

TypeEmerging player
Description

Provides ultra-low-latency LLM inference through custom LPU hardware. Competes with Inception on the same 'speed as a differentiator' thesis — Groq at the hardware/serving layer, Inception at the model architecture layer.

TypeOthers
Description

Wafer-scale AI compute provider and co-developer of Inception's Jais 2 Arabic LLM. While not a direct model competitor, Cerebras is an infrastructure partner and adjacent player in the speed-optimized AI stack that underpins Mercury's value proposition.

TypeEmerging player
Description

Open-source-focused AI inference and fine-tuning platform hosting multiple foundation models. Competes with Inception's API platform for developer mindshare in serving fast, cost-efficient LLM inference.

TypeDirect peer
Description

Foundation model provider offering task-specific and general LLMs (Jamba family) with enterprise focus. Similar stage and enterprise GTM motion to Inception, with comparable emphasis on efficiency and developer-friendly APIs.

TypeEmerging player
Description

Multimodal foundation model startup offering Reka Core, Flash, and Edge models. Comparable in stage and ambition to Inception, with overlapping multimodal and enterprise positioning.

Market position
Strengths4 records

Each record includes

Headline, Details, Source

Weaknesses5 records

Each record includes

Headline, Details, Source

Competitive moat5 records

Each record includes

Type, Details

Key risks6 records

Each record includes

Headline, Details, Source

Key highlights7 records

Each record includes

Headline, Details, Source

Customer concentration

Classification, Details

Named customers9 records

Each record includes

Name, Industry, Type, Use case, Source, UUID

Segment5 records

Each record includes

Title, Type, Primary, Description, Pain point addressed, Use case, Source

Technology focused
Yes
API detail
Has APIbool
Yes

Docs URL, Description

Integration12 records

Each record includes

Title, Type, Description, Source

AI capability9 records

Each record includes

Type, Description, Source

AI maturity
App detail

Has app

Feature5 records

Each record includes

Title, Differentiator, Description, Source

Core technology
Revenue estimate
Valuation estimate
Number of profiles
No data
Compliance2 records

Each record includes

Name, Class, Description

Funding overview

Funding stage, Last funding date, Total funding USD

Funding rounds3 records

Each record includes

Round, Amount USD, Date, Pre money valuation, Total investors, Investors, News

Investors10 records

Each record includes

Name, Type, Date of entry, Rounds participated, Website

Funding detail is available on the Subscription and Enterprise plan.Contact sales →

M&A

Each record includes

Name, Acquisition type, Announced date, Completed date, Status, Website, News

Investment

Each record includes

Name, Round, Announced date, Lead investor, Website, News

M&A and investment is available on the Subscription and Enterprise plan.Contact sales →

Inception

AI Foundation Modelsinceptionlabs.ai

Inception develops diffusion-based large language models (Mercury family) that generate tokens in parallel for ~10x faster inference than autoregressive LLMs, serving developers, voice AI platforms, enterprise search, and agent platforms via API and cloud marketplace channels.

What Inception does

Inception AI, Inc. is a Palo Alto-based AI company founded in 2024 by Stanford professor Stefano Ermon that develops diffusion-based large language models (dLLMs), an alternative architecture to traditional autoregressive LLMs. The company's Mercury product family generates text and code through parallel token refinement rather than sequential decoding, achieving approximately 1,000 tokens per second on NVIDIA Blackwell GPUs—roughly 10x faster than leading autoregressive models—while maintaining competitive benchmark quality (e.g., AIME 2025: 91.1%, GPQA: 73.6%) at substantially lower inference cost. The portfolio spans Mercury 2 (reasoning model with 128K context), Mercury Edit 2 (next-edit prediction for code), Mercury Coder (code generation), and the general Mercury chat model, with a research-stage multimodal extension called LaViDa. The team includes researchers from Stanford, UCLA, Cornell, Google DeepMind, Meta AI, Microsoft AI, and OpenAI, and the founders are credited with inventing foundational techniques including diffusion models, Flash Attention, and Direct Preference Optimization.

The product is distributed through an OpenAI-compatible REST API, the Mercury Chat playground, AWS Bedrock, Azure AI Foundry (with SOC2 and HIPAA compliance), OpenRouter, and Models.dev, with deep integrations into developer tooling (Zed, Continue, ProxyAI, Augment Code) and the Microsoft NLWeb open project. Inception generates revenue primarily through usage-based API pricing ($0.25 per 1M input tokens, $0.75 per 1M output tokens) supplemented by enterprise licensing on cloud marketplaces and custom contracts with dedicated capacity and SLAs. A 10-million-token free tier supports a self-serve developer acquisition motion running in parallel with a dedicated enterprise sales motion. Customer logos span coding tools (Zed, Augment Code, Buildglare, ProxyAI), enterprise search (SearchBlox), AI agent platforms (Radient, Skyvern), and real-time voice AI (Happyverse AI, OpenCall), with strategic partnerships extending into financial services (Bain & Company, Kensho), Arabic-language AI (Cerebras, MBZUAI, Jais 2), and government/security (Brain Co., Mirror Security/G42).

In November 2025, Inception raised a $50 million seed round led by Menlo Ventures with participation from Microsoft M12, NVentures, Snowflake Ventures, Databricks Ventures, Mayfield, and Innovation Endeavors, plus angels Andrew Ng and Andrej Karpathy. As of mid-2026, the company was reportedly in acquisition discussions with Microsoft at a valuation exceeding $1 billion.

Inception firmographics

Firmographics
Name
Inception
Legal name
Inception AI, Inc.
Website
https://inceptionlabs.ai
Company type
Private
Founded year
2024
Operating status
Operating
Headcount range
11–50 employees
Short description
Inception develops diffusion-based large language models (Mercury family) that generate tokens in parallel for ~10x faster inference than autoregressive LLMs, serving developers, voice AI platforms, enterprise search, and agent platforms via API and cloud marketplace channels.
Ownership category
akta.pro rank

Inception industry classification

Industry
Product category
AI Foundation Models
NAICS
Software Publishers (5132), Software Publishers (513210), Computer Systems Design and Related Services (5415)
SIC
Services-Prepackaged Software (7372)
akta.pro primary industry
Multimodal Generation (Text-Image-Video-Audio) & Creative Toolchains (HDAAACAG)
akta.pro secondary industries
LLMOps & Generative AI Platforms (Prompt/Agent Orchestration, RAG) (HDAEANAD), Fine-Tuning, Adaptation & Custom Model Training (PEFT/LoRA/RLHF) (HDAAACAC), Machine Translation & Multilingual NLP (HDAAADAD), MLOps/LLMOps & Model Lifecycle Management Services (BPAEAHAH)

Keywords

  • Diffusion language models
  • Generative AI foundation models
  • AI inference API
  • Code generation AI
  • Reasoning LLM platform

Where Inception is headquartered

Location

Headquarters

HQ city
Palo Alto
HQ country
United States
HQ region
North America

Offices1 record

Markets served

Inception business model

Business model
GTM type
B2B
Offering type
Software

Revenue model

  1. API Usage-Based Pricing: Pay-per-token pricing model for Mercury and Mercury Coder models through Inception API platform, with differentiated pricing for input tokens, output tokens, and cached inputs
  2. Enterprise Licensing: Enterprise deployment licensing through AWS Bedrock and Azure AI Foundry with software license costs (e.g., $0.78/hour for Mercury on Azure) plus cloud compute costs
  3. Custom Enterprise Contracts: Custom rate limits, SLA guarantees, volume-based pricing, dedicated capacity, and custom terms for large enterprise deployments

Pricing tiers

ModelBillingPrice
FreemiumPay-as-you-goFree tier with 10 million tokens
Usage-basedPay-as-you-goDeveloper tier with usage-based pricing
SubscriptionAnnualEnterprise tier with custom pricing and SLAs
Usage-basedPay-as-you-goAzure AI Foundry enterprise deployment

Go-to-market motion3 records

Distribution channels5 records

Marketing channels9 records

Inception product offering

Product offering

Core offering

Inception builds and deploys diffusion-based large language models (dLLMs) under the Mercury family that generate text and code through parallel iterative refinement rather than sequential autoregressive decoding. Their models deliver 5-10x faster inference at roughly half the cost of conventional LLMs while maintaining frontier-class quality, with distribution through a self-serve OpenAI-compatible API and enterprise deployments on AWS Bedrock and Azure AI Foundry.

Product overview

Inception is an AI research and product company developing diffusion-based large language models (dLLMs) that generate text and code through parallel refinement rather than sequential token-by-token generation. The product portfolio centers on the Mercury family of diffusion LLMs: Mercury 2 serves as the flagship reasoning model with 128K context, Mercury Edit 2 provides next-edit prediction for coding workflows, Mercury Coder powers code generation and apply-edit capabilities, and the general chat Mercury model enables conversational AI. The Inception API (OpenAI-compatible) and Mercury Chat playground provide developer access, while enterprise deployment options include AWS Bedrock, Azure AI Foundry, and model routers like OpenRouter. The company also partners with Bain & Company, Kensho, and Brain Co. for enterprise solutions, and collaborates on specialty models like Jais 2 (Arabic LLM with Cerebras/MBZUAI). Mercury's diffusion approach delivers 5-10x faster inference than traditional autoregressive models while maintaining competitive quality, enabling use cases in coding assistants, real-time voice agents, agentic loops, and enterprise search.

Differentiator

Problem solved

Functional benefit

Products and services

  • Mercury 2 The flagship diffusion-based reasoning LLM from Inception achieving approximately 1,009 tokens per second on NVIDIA Blackwell GPUs through parallel token generation. Features 128K context window, tunable reasoning, native tool use, and schema-aligned JSON output for developers and enterprises building latency-sensitive AI applications.
  • Mercury Edit 2 A small, coding-focused diffusion LLM optimized for next-edit prediction in coding workflows. Uses recent edits and codebase context to predict what developers will change next, delivering fast autocomplete and apply-edit suggestions.
  • Mercury Coder The first commercial-scale diffusion LLM optimized for code generation, achieving 1000+ tokens per second on NVIDIA H100s. Supports fill-in-the-middle (FIM) and apply-edit capabilities for coding workflows.
  • Mercury Coder Small A smaller coding-focused diffusion model variant available via API, running more than 5x faster than speed-optimized frontier models like GPT-4o Mini and Claude 3.5 Haiku while matching them in quality. Supports 32K context window and fill-in-the-middle workflows.
  • Mercury (General Chat) First general chat diffusion LLM that matches the performance of speed-optimized frontier models like GPT-4.1 Nano and Claude 3.5 Haiku while running over 7x faster. Powers conversational AI applications, enterprise search, and voice interfaces.
  • Inception API Platform REST API providing programmatic access to Mercury diffusion LLMs. OpenAI API compatible with support for chat completions, fill-in-the-middle, and apply-edit endpoints. New accounts receive 10 million free tokens.
  • Mercury Chat Playground Interactive web-based playground for testing Mercury models. Allows developers to explore model capabilities and experiment with diffusion-based text generation.
  • Mercury on Azure AI Foundry

Quantifiable outcome

  • 1,000 tokens/second throughput vs ~71-89 tokens/second for GPT-5 Mini and Claude 4.5 Haiku
  • +5 more outcomes

Companies that use Inception

Customer profile

Named customers9 records

Segments5 records

Inception technology and API

Technology

Technology focussed Yes

API detail

Has API
Yes
API docs
API detail

Core technology

AI maturity

App detail

Integration12 records

AI capability9 records

Feature5 records

Inception partnerships and signals

Strategic signal

Partnerships

17 partnerships are on record, tiered flagship, core and strategic.

  • Augment CodeflagshipTechnology or Integration · 12 May 2026Production deployment of Mercury 2 for context compaction, model routing, and tool search; achieved 82% latency reduction and 90% cost savings
  • ZedcoreTechnology or Integration · 30 March 2026Mercury Edit 2 integrated into Zed's edit prediction system for fast code editing
  • SearchBloxflagshipTechnology or Integration · 12 January 2026Mercury-powered SearchAI for sub-second GenAI search with 60-90% lower inference costs across enterprise workloads
  • Cerebras SystemsflagshipStrategic or Co-development Partner · 9 December 2025Co-developed Jais 2 Arabic LLM (70B parameters) trained on largest Arabic-first dataset; Cerebras provides wafer-scale computing infrastructure for training
  • MBZUAI (Mohamed bin Zayed University of Artificial Intelligence)flagshipStrategic or Co-development Partner · 9 December 2025Co-developed Jais 2 Arabic LLM for high fluency, cultural depth, and safety in Arabic language AI
  • Mirror SecuritystrategicStrategic or Co-development Partner · 24 November 2025Strategic partnership with G42's Inception to co-develop AI security products for government and enterprise clients
  • Microsoft Azure AI FoundryflagshipTechnology or Integration · 18 November 2025Mercury available on Azure AI Foundry with enterprise-grade infrastructure, compliance standards (SOC2, HIPAA), and $0.78/hour software license
  • ProxyAIcoreTechnology or Integration · 27 October 2025Partnership to bring faster, more consistent code edits to ProxyAI platform
  • KenshostrategicImplementation/ SI/ Consulting Partner · 20 October 2025Strategic collaboration to develop AI solutions for financial institutions at GITEX Global 2025
  • Bain & CompanystrategicImplementation/ SI/ Consulting Partner · 14 October 2025Strategic collaboration to develop and deploy AI solutions for financial services, starting with (In)Alpha for investment workflows
  • Brain Co.strategicStrategic or Co-development Partner · 14 October 2025Strategic partnership to develop next-generation AI enterprise applications for public services, healthcare, and energy sectors
  • G42 (Inception subsidiary context)coreOthers · 13 October 2025Inception is a G42 company with access to G42's AI infrastructure and regional presence in Middle East
  • Amazon (AWS Bedrock & SageMaker JumpStart)flagshipTechnology or Integration · 27 August 2025Mercury and Mercury Coder available on Amazon Bedrock Marketplace and SageMaker JumpStart for enterprise deployment
  • Continue (VS Code Extension)coreTechnology or Integration · 27 August 2025IDE integration partnership with Continue extension for Mercury Coder autocomplete and next-edit features
  • RadientcoreTechnology or Integration · 19 August 2025Mercury powering model routing for Radient Automatic, enabling sub-second classification and 70% cost savings
  • BuildglarecoreTechnology or Integration · 6 July 2025Mercury Coder powering Apply-Edit functionality for low-code web development platform
  • Microsoft NLWebflagshipTechnology or Integration · 15 May 2025Founding LLM partner for Microsoft's NLWeb open project, enabling natural language interfaces for websites with ultra-fast Mercury dLLM

Scale indicators6 records

Recent moves6 records

Expansion highlights6 records

Inception competitors and assessment

Company assessment

Direct peers

  • OpenAI: Frontier foundation model provider (GPT-4o, GPT-5 family) competing directly with Mercury's reasoning and coding capabilities. Inception positions Mercury 2 as '2x faster than GPT-5.2' per customer testimonials, and offers an OpenAI-compatible API as a drop-in alternative.
  • Anthropic: Enterprise-focused foundation model provider (Claude family). Anthropic's Claude Haiku 4.5 is benchmarked head-to-head with Mercury 2 on speed, and Inception's enterprise GTM via AWS Bedrock and Azure Foundry competes directly with Anthropic's distribution footprint.
  • Mistral AI: European foundation model lab offering open and commercial LLMs with emphasis on efficiency and inference cost — a similar value proposition to Inception's Mercury dLLMs targeting latency-sensitive and cost-sensitive enterprise workloads.
  • Cohere: Enterprise-focused foundation model provider offering API-accessible LLMs with private deployment and RAG tooling. Competes with Inception for enterprise search, retrieval, and agentic use cases through similar cloud marketplace channels.
  • AI21 Labs: Foundation model provider offering task-specific and general LLMs (Jamba family) with enterprise focus. Similar stage and enterprise GTM motion to Inception, with comparable emphasis on efficiency and developer-friendly APIs.

Broad incumbents

  • Google DeepMind: Develops Gemini family of foundation models with broad modality support (text, image, video, audio). A broad incumbent whose research output (including diffusion-based image/video models) competes architecturally with Inception's diffusion-based language approach.

Emerging players

  • Groq: Provides ultra-low-latency LLM inference through custom LPU hardware. Competes with Inception on the same 'speed as a differentiator' thesis — Groq at the hardware/serving layer, Inception at the model architecture layer.
  • Together AI: Open-source-focused AI inference and fine-tuning platform hosting multiple foundation models. Competes with Inception's API platform for developer mindshare in serving fast, cost-efficient LLM inference.
  • Reka AI: Multimodal foundation model startup offering Reka Core, Flash, and Edge models. Comparable in stage and ambition to Inception, with overlapping multimodal and enterprise positioning.

Others

  • Cerebras Systems: Wafer-scale AI compute provider and co-developer of Inception's Jais 2 Arabic LLM. While not a direct model competitor, Cerebras is an infrastructure partner and adjacent player in the speed-optimized AI stack that underpins Mercury's value proposition.

Market position

Strengths4 records

Weaknesses5 records

Competitive moat5 records

Key risks6 records

Key highlights7 records

Customer concentration

Inception social profiles

Digital presence

Inception compliance and trust

Trust signal

Compliance2 records

Inception financial estimates

Financial estimate

Revenue estimate

Valuation estimate

Inception leadership team

Management profile

Number of profiles

Inception funding detail

Funding detail

Funding overview

Funding rounds3 records

Investors10 records

Funding detail is available on the Subscription and Enterprise plan.Contact sales →

Inception M&A and investment

M&A and investment

M&A

Investments

M&A and investment is available on the Subscription and Enterprise plan.Contact sales →

Frequently asked questions about Inception

What does Inception do?

Inception builds and deploys diffusion-based large language models (dLLMs) under the Mercury family that generate text and code through parallel iterative refinement rather than sequential autoregressive decoding. Their models deliver 5-10x faster inference at roughly half the cost of conventional LLMs while maintaining frontier-class quality, with distribution through a self-serve OpenAI-compatible API and enterprise deployments on AWS Bedrock and Azure AI Foundry.

Is Inception a public or private company?

Inception is a private company. It is classified as venture growth investor backed and is currently operating.

When was Inception founded?

Inception was founded in 2024. It employs 11 to 50 people.

Where is Inception based?

Inception is headquartered in Palo Alto, United States, in the North America region.

How does Inception make money?

Three revenue lines are on record. API Usage-Based Pricing is the primary driver. The others are enterprise Licensing and custom Enterprise Contracts.

Who are Inception's main competitors?

Direct peers on record are OpenAI, Anthropic, Mistral AI, Cohere and AI21 Labs. Google DeepMind is listed as a broad incumbent. Emerging players are Groq, Together AI and Reka AI. Cerebras Systems is listed as an others.

Does Inception have an API?

Yes. Inception offers a REST API compatible with OpenAI API standards, providing programmatic access to Mercury diffusion large language models (dLLMs). The API supports standard chat completions, fill-in-the-middle (FIM) completions, and apply-edit completions endpoints. Available libraries include AISuite, LiteLLM, and LangChain. New accounts receive 10 million free tokens. The API is a drop-in replacement for traditional LLMs. Developer documentation is at docs.inceptionlabs.ai/get-started/get-started.

What industry is Inception in?

Inception's product category is AI Foundation Models. Its primary akta.pro industry code is HDAAACAG, Multimodal Generation (Text-Image-Video-Audio) & Creative Toolchains, with a secondary code of HDAEANAD, LLMOps & Generative AI Platforms (Prompt/Agent Orchestration, RAG). Its NAICS code is 5132 and its SIC code is 7372.

Unlock the full company data

50 free credits on sign-up, no credit card required.

Contact sales
Live signals
RuntimewireInception launches Mercury Voice for faster AI phone agentsInception launched Mercury Voice for enterprise customers on September 29th, claiming a 320-millisecond median time to first answer token on production prompts. The diffusion-based model outperformed several competitors on benchmark tests, with pricing at $0.40 per million input tokens and $1.50 per million output tokens.FinancialContent Business PageInception Launches Mercury 2.5, the Next Tier of Intelligence for Diffusion LLMsInception launched Mercury 2.5, a diffusion-based LLM running over 1,100 tokens per second. It offers 10 points more intelligence than Mercury 2, with pricing at $0.04 per 1M tokens at launch. The model extends context to 260K tokens and adds tunable reasoning and tool use.Business Wire BlogInception Launches Mercury 2.5, the Next Tier of Intelligence for Diffusion LLMsInception launched Mercury 2.5, its fastest diffusion LLM, running over 1,100 tokens per second. The model offers 10-point intelligence gains over Mercury 2, with pricing at $0.20/$0.75 per million tokens and 80% off at $0.04/$0.15. It extends context to 260K tokens and adds tunable reasoning and tool use.MIT Technology ReviewThese startups are chasing the next big thing in LLMsMultiple AI startups are developing alternative technologies to overcome the computational inefficiencies and architectural limitations of transformer-based large language models, which are becoming a bottleneck as AI demands grow. Four approaches are profiled: sparse attention (Subquadratic's SubQ), power retention (Manifest AI's PowerCoder), liquid neural networks (Liquid AI's hybrid LFMs), and diffusion-based text generation (Inception's Mercury 2), each offering different trade-offs between speed, efficiency, and capability. Industry observers note that transformers, while foundational, represent just one phase in AI development, with significant innovation still needed to achieve more general intelligence at lower computational cost.ZawyaBanco Santander and G42 sign MoU to explore strategic cooperation in artificial intelligenceBanco Santander and G42 signed an MOU to explore strategic cooperation in artificial intelligence, establishing a framework for joint initiatives. The partnership will focus on AI-enabled advisory and savings solutions and a banking intelligence layer, with G42's Inception and Presight expected to contribute. The collaboration aims to build a durable, compliant AI platform for Santander's customers.AI CERTsMicrosoft–Inception Talks Signal AI M&A ShiftMicrosoft is in active acquisition talks to buy AI startup Inception, with the deal potentially exceeding $1 billion, according to Reuters. The startup's diffusion-based LLM technology reportedly achieves 1,109 tokens per second on NVIDIA H100 hardware—ten times faster than optimized GPT-class models—while SpaceX's xAI unit had also circled Inception earlier, adding competitive pressure to the negotiations. The acquisition would inject alternative AI capabilities into Microsoft, reducing its dependency on OpenAI (in which it holds a 49% stake) and lowering latency costs across Azure and Office products.Analytics InsightMicrosoft Explores New AI Partnerships Amid OpenAI Dependence ConcernsMicrosoft is exploring AI startup acquisitions and partnerships to reduce long-term dependence on OpenAI. The company is in talks with Inception, a diffusion-based language model startup, and previously considered buying Cursor but walked away. Microsoft has spent over $100 billion on OpenAI-related investments.Channel NewsAsiaExclusive-Microsoft eyeing startup deals for life after OpenAIMicrosoft is pursuing AI startups, including Inception, to build an independent AI capability after OpenAI. It backed away from acquiring Cursor due to regulatory concerns. Discussions are ongoing, with Inception seeking over $1 billion.DoNews微软探索收购AI初创公司应对后OpenAI时代- DoNews快讯Microsoft is negotiating with AI startup Inception for a merger or strategic partnership valued at no less than $1 billion. The deal aims to strengthen Microsoft's technological autonomy in generative AI and reduce dependence on a single external partner. Earlier, Microsoft shelved an acquisition of code generation tool Cursor due to antitrust concerns.MarketScreenerMicrosoft eyeing startup deals for life after OpenAIMicrosoft is pursuing AI startups like Inception and Cursor as it plans to move beyond its OpenAI partnership. Microsoft backed away from acquiring Cursor due to regulatory concerns, while SpaceX also courted Inception, which seeks over $1 billion. The discussions are ongoing and may not result in a deal.