General Compute
General Compute operates an ASIC-native AI inference neocloud using SambaNova dataflow accelerators at renewable-powered LATAM data centers, serving developers building autonomous coding agents and real-time AI workloads via an OpenAI-compatible API with usage-based and dedicated capacity tiers.
- Company typePrivate
- Founded2025
- HeadquartersWest Covina, United States
- Headcount11–50
- GTM typeB2B
- OfferingSoftware
What General Compute does
General Compute is an AI inference neocloud startup operating an OpenAI-compatible API platform purpose-built for autonomous AI coding agents and other sequential inference workloads. Rather than running inference on repurposed GPUs, the company deploys SambaNova SN40L and SN50 dataflow ASIC accelerators with on-chip weight residency in SRAM, eliminating per-token memory fetches that bottleneck GPU decode performance at batch size one. The company serves three customer tiers: a self-serve usage-based API at $0.45 per million input tokens and $0.60 per million output tokens (with $100 signup credit), quote-based dedicated capacity deployments with SLAs and reserved infrastructure, and a Bring Your Own Model offering for private model weight hosting. Primary target customers are developers building autonomous coding agents (OpenClaw, Codex, Claude Code, OpenCode, Cursor, Aider) and voice AI applications requiring sub-second latency.
The company's infrastructure footprint is concentrated in South America — an Asunción, Paraguay site leveraging that country's near-100% hydroelectric generation, plus five Brazilian campuses through a partnership with Elea Data Centers running on certified renewable power. This produces a 3.7x energy cost advantage ($0.035/kWh versus $0.13/kWh US average) and a 7x efficiency advantage at the accelerator level (17 kW versus 120 kW per comparable unit). A Q4 2026 roadmap introduces a disaggregated prefill/decode architecture pairing AMD MI300X GPUs with SambaNova SN50 silicon, targeting 20x throughput improvement on 600B+ parameter models. The company also exposes agent-native onboarding that allows autonomous coding agents to sign up, claim credits, and retrieve API keys without human intervention, and distributes through the OpenRouter marketplace alongside direct sales.
General Compute raised a $15 million seed round at a $60 million post-money valuation in May 2026, led by FUSE VC with participation from Carya Venture Partners and Village Global Ventures. The capital is committed to deploying SambaNova inference silicon under a $300 million chip purchase LOI. The company operates as General Compute Inc., is privately held, and is reported to be founder-led with Finn P. as CEO and Jason Goodison as CTO. Revenue is not yet disclosed and the platform was launched into general availability on May 15, 2026.
General Compute firmographics
Firmographics- Name
- General Compute
- Legal name
- General Compute Inc.
- Website
- https://generalcompute.com
- Company type
- Private
- Founded year
- 2025
- Operating status
- Operating
- Headcount range
- 11–50 employees
- Short description
- General Compute operates an ASIC-native AI inference neocloud using SambaNova dataflow accelerators at renewable-powered LATAM data centers, serving developers building autonomous coding agents and real-time AI workloads via an OpenAI-compatible API with usage-based and dedicated capacity tiers.
- Ownership category
- akta.pro rank
General Compute industry classification
Industry- Product category
- AI Inference Cloud Infrastructure
- NAICS
- Computing Infrastructure Providers, Data Processing, Web Hosting, and Related Services (518210), Computing Infrastructure Providers, Data Processing, Web Hosting, and Related Services (5182), Computing Infrastructure Providers, Data Processing, Web Hosting, and Related Services (518), Computer Systems Design and Related Services (5415)
- SIC
- Services-Computer Programming, Data Processing, Etc. (7370), Services-Computer Integrated Systems Design (7373), Services-Computer Processing & Data Preparation (7374)
- akta.pro primary industry
- AI Compute Cloud & GPU-as-a-Service (HDAAAAAK)
- akta.pro secondary industries
- AI Accelerators (GPUs/TPUs/NPUs/ASICs) (HDAAAAAA), AI Compiler, Runtime & Kernel Optimization Software (CUDA/ROCm/XLA, graph compilers) (HDAAAAAI), AI Datacenter Power, Cooling & Racks (liquid cooling, power delivery) (HDAAAAAH), AI Compute Virtualization & Scheduling (GPU virtualization, cluster schedulers) (HDAAAAAG)
Keywords
Where General Compute is headquartered
LocationHeadquarters
- HQ city
- West Covina
- HQ country
- United States
- HQ region
- North America
Offices7 records
Markets served
General Compute business model
Business model- GTM type
- B2B
- Offering type
- Software
- Cost components
- Technology or R&D, Infrastructure, Supply Chain, Personnel, Operations, Marketing or Sales
Revenue model
- API Access (Usage-Based): Usage-based inference API charging per token. Self-serve model with $100 free credit for new accounts. Customers pay based on input and output token volume consumed through the OpenAI-compatible API.
- Dedicated Capacity: Custom deployments with dedicated infrastructure, guaranteed throughput, custom SLAs, and reserved capacity for production workloads. Target for customers with predictable high-volume needs or latency-sensitive applications.
- Bring Your Own Model: Customers can deploy private model weights on General Compute's optimized ASIC infrastructure while maintaining OpenAI-compatible API surface. Custom pricing for private weights and deployment-specific requirements.
Pricing tiers
| Model | Billing | Price |
|---|---|---|
| Usage-based | Pay-as-you-go | Self-serve API with $100 free credit for new accounts |
| Subscription | Multi-year contract | Custom deployments with dedicated infrastructure and SLAs |
| Subscription | Multi-year contract | Bring your own model weights on optimized infrastructure |
Go-to-market motion1 record
Distribution channels4 records
Marketing channels5 records
General Compute product offering
Product offeringCore offering
General Compute operates an ASIC-native AI inference cloud (General Compute Cloud) running on SambaNova dataflow accelerators and exposed via an OpenAI-compatible REST API. It sells three commercial tiers — self-serve API Access, Custom Deployments with dedicated capacity and SLAs, and a Bring Your Own Model program — targeted at AI coding agents, voice AI, and enterprise production workloads. The platform is designed for sub-second, high-throughput token generation at lower cost than GPU-based hyperscalers.
Product overview
General Compute is an AI inference neocloud startup offering three product tiers through a unified ASIC-native infrastructure: API Access (self-serve OpenAI-compatible REST endpoints with $100 free credit), Custom Deployments (dedicated capacity with SLAs), and Bring Your Own Model (deploy private weights on optimized hardware). The flagship General Compute Cloud platform runs on purpose-built SambaNova SN40/SN50 dataflow ASICs rather than GPUs, delivering claimed speeds up to 7x faster and 1,000 tokens/second for AI agent workloads. The architecture separates prefill and decode stages for independent scaling, and enables autonomous agent signup without human intervention.
Differentiator
Problem solved
Functional benefit
Products and services
- API Access Self-serve REST API access with OpenAI-compatible endpoints that allows developers to obtain an API key and start making inference requests immediately. Includes $100 free credit for new accounts and serves individual developers, startups, and AI coding agents.
- Custom Deployments Dedicated inference capacity with SLAs and guaranteed throughput, custom scaling, and deployment support for production workloads and latency-sensitive applications. Targeted at enterprises with predictable high-volume needs.
- Bring Your Own Model Deploy private model weights on General Compute's optimized ASIC infrastructure while maintaining the OpenAI-compatible API surface. Custom pricing for enterprises with specific model, compliance, or deployment requirements.
- General Compute Cloud ASIC-native neocloud platform for autonomous AI development, featuring SambaNova SN40 and SN50 dataflow silicon with agent-native onboarding and independently benchmarked speeds on the MiniMax M2 model family. The unified platform sits behind the three commercial tiers.
Quantifiable outcome
- 2.6x faster time-to-first-token vs Together AI on GPT-OSS-120B (738ms vs 1,899ms)
- +4 more outcomes
Companies that use General Compute
Customer profileSegments3 records
Ideal customer profiles3 records
General Compute technology and API
TechnologyTechnology focussed Yes
API detail
- Has API
- Yes
- API docs
- API detail
Core technology
AI maturity
App detail
Integration8 records
AI capability3 records
Feature5 records
General Compute partnerships and signals
Strategic signalPartnerships
Six partnerships are on record, tiered core and flagship.
- OpenRoutercoreDistribution partnership going live at May 1, 2026 launch. OpenRouter provides marketplace distribution for General Compute's inference API alongside other AI providers.
- Elea Data CentersflagshipBrazil's largest data center platform with nine campuses across São Paulo, Rio de Janeiro, Brasília, Porto Alegre, and Curitiba, all running on certified renewable energy. Anchor partner for Brazilian deployments, purpose-built for high-density AI workloads including ingestion, training, fine-tuning, and inference.
- Modular Data CenterscoreProvides containerized data center products for fast deployment at powered sites. ISO 9001 certified, Brazil-based, internationally deployable. Enables rapid infrastructure deployment when cheap energy sites are identified.
- Cloud2GroundcoreSite sourcing and strategy partner. Finds sites, aligns capital, and handles execution in markets most infrastructure advisors don't operate. Team has experience scaling CyrusOne from private equity asset to public company.
- SambaNovaflagshipSambaNova provides SN40L and SN50 dataflow accelerators purpose-built for AI inference. General Compute has $300M of SN50 chips on order. Executed LOI covering dedicated SN40L capacity today with path to SN50 as it comes online.
- AMDcoreActive partnership track for AMD MI300X GPUs to handle prefill phase in Q4 2026 disaggregated architecture. AMD provides compute for compute-bound prefill operations while SambaNova handles decode.
Scale indicators10 records
Recent moves6 records
Expansion highlights6 records
General Compute competitors and assessment
Company assessmentOthers
- OpenRouter: Marketplace aggregator distributing inference APIs from multiple providers including General Compute. Functions as an indirect channel partner and ecosystem enabler for cross-provider model routing.
Direct peers
- Fireworks AI: Inference-as-a-service platform focused on low-latency LLM serving for production applications. Targets similar developer and enterprise use cases with custom optimization layers and OpenAI-compatible APIs.
- Groq: AI inference cloud built on proprietary LPU silicon, offering low-latency OpenAI-compatible APIs. Closest architectural analog to General Compute's ASIC-native inference approach targeting the same decode-heavy workload profile.
- Anyscale: AI compute platform built on Ray, offering both training and inference services for production AI workloads. Competes for the same developer and enterprise AI infrastructure budgets with OpenAI-compatible endpoints.
- Cerebras Systems: AI inference and training cloud built on wafer-scale ASICs. Offers OpenAI-compatible APIs optimized for ultra-low latency, directly competing in the purpose-built silicon inference category.
- Together AI: AI inference cloud platform offering OpenAI-compatible APIs across open-source models. Directly benchmarked competitor in General Compute's own performance comparisons, serving the same developer and agent customer base.
- Modal Labs: Serverless AI infrastructure platform for running inference and batch workloads on GPUs. Targets the same developer and startup customer base with API-first distribution and rapid provisioning.
Broad incumbents
- CoreWeave: Large-scale GPU cloud provider offering both training and inference services. Broader incumbent in the AI neocloud category with significantly more capital and infrastructure scale than General Compute.
- Lambda Labs: GPU cloud provider offering reserved instances and on-demand compute for AI training and inference. Established alternative to hyperscalers with a similar developer-first go-to-market motion.
Emerging players
- SambaNova: SambaNova is both General Compute's primary silicon supplier and an emerging AI cloud provider offering its own inference-as-a-service on dataflow chips. The partnership creates both alignment and competitive overlap.
Market position
Strengths5 records
Weaknesses5 records
Competitive moat5 records
Key risks7 records
Key highlights7 records
Customer concentration
General Compute social profiles
Digital presenceGeneral Compute financial estimates
Financial estimateRevenue estimate
Valuation estimate
General Compute leadership team
Management profileNumber of profiles
Profiles2 records
General Compute funding detail
Funding detailFunding overview
Funding rounds2 records
Investors4 records
Funding detail is available on the Subscription and Enterprise plan.Contact sales →
General Compute M&A and investment
M&A and investmentM&A
Investments
M&A and investment is available on the Subscription and Enterprise plan.Contact sales →
Frequently asked questions about General Compute
What does General Compute do?
General Compute operates an ASIC-native AI inference cloud (General Compute Cloud) running on SambaNova dataflow accelerators and exposed via an OpenAI-compatible REST API. It sells three commercial tiers — self-serve API Access, Custom Deployments with dedicated capacity and SLAs, and a Bring Your Own Model program — targeted at AI coding agents, voice AI, and enterprise production workloads. The platform is designed for sub-second, high-throughput token generation at lower cost than GPU-based hyperscalers.
Is General Compute a public or private company?
General Compute is a private company. It is classified as venture growth investor backed and is currently operating.
When was General Compute founded?
General Compute was founded in 2025. It employs 11 to 50 people.
Where is General Compute based?
General Compute is headquartered in West Covina, United States, in the North America region.
How does General Compute make money?
Three revenue lines are on record. API Access (Usage-Based) is the primary driver. The others are dedicated Capacity and bring Your Own Model.
Who are General Compute's main competitors?
OpenRouter is listed as an others. Direct peers are Fireworks AI, Groq, Anyscale, Cerebras Systems, Together AI and Modal Labs. Broad incumbents are CoreWeave and Lambda Labs. SambaNova is listed as an emerging player.
Does General Compute have an API?
Yes. General Compute offers an OpenAI-compatible REST API for AI inference at https://api.generalcompute.com/v1. The API supports chat completions with streaming, tool calling, JSON mode, and embeddings. Authentication uses Bearer tokens. Official SDKs are provided for Python (pip install generalcompute) and Node.js (npm install generalcompute). The platform also supports webhooks for account, billing, and inference events with HMAC-signed deliveries, and provides an MCP (Model Context Protocol) server for MCP-aware agents at https://mcp.generalcompute.com. Developer documentation is at docs.generalcompute.com.
What industry is General Compute in?
General Compute's product category is AI Inference Cloud Infrastructure. Its primary akta.pro industry code is HDAAAAAK, AI Compute Cloud & GPU-as-a-Service, with a secondary code of HDAAAAAA, AI Accelerators (GPUs/TPUs/NPUs/ASICs). Its NAICS code is 518210 and its SIC code is 7370.