Developer docs
API playgroundTry for free, no card

Search company profiles

VLM Run

Full company profile

uuid00078y8

Namestring
VLM Run
Legal namestring
Autonomi AI Inc.
Websiteurl
vlm.run
Company typeenum
Private
Founded yearint
2022
Descriptiontext

VLM Run (legal entity Autonomi AI Inc.) is a Palo Alto-based private company founded in 2022 that operates an API-first visual intelligence platform built around its flagship product, Orion. Orion is described as a "visual agent" that combines the reasoning capability of large Vision-Language Models with the accuracy of specialized computer-vision tools, exposed through a single OpenAI-compatible REST API at api.vlm.run/v1. The platform ingests images, documents, video, and audio and outputs structured JSON, with support for custom schemas via Pydantic/Zod and visual grounding bounding boxes for auditability. VLM Run complements Orion with open-source projects (vlmrun-hub for domain schemas, vlmbench for VLM benchmarking, and mm for multi-modal context) and a Skills/Agents system for reusable, versioned extraction workflows.

The company targets two overlapping customer sets: developers and AI-native companies that want a drop-in OpenAI SDK replacement for visual tasks (logos include Voxel51, Langflow, MCP.run, York IE, MongoDB, Aurochs), and regulated enterprises that need SOC 2 Type II, HIPAA, BAA, and In-VPC deployment (logos include Valerie Health, Scan.com, basata, Peak Health in healthcare; Zapier, Make, n8n, Activepieces in automation; Google Cloud as a broader ecosystem signal). Vertical solutions are pre-packaged for Healthcare (faxes, scans, clinical paperwork) and Construction (blueprints, schedules, specs).

VLM Run monetizes on a credit-based usage model (100 credits = $1 USD), structured into three tiers: a free Starter tier with 1,000 sign-up credits and 10 requests/min, a Pro tier at $799/month including 100,000 credits and 100 requests/min, and a custom-priced Enterprise tier with unlimited credits, in-VPC deployment, and dedicated SLAs. Go-to-market is product-led growth for developers and direct enterprise sales for regulated deployments, supported by native integrations with Zapier (8,000+ apps) and MongoDB Atlas. The company has 1–10 employees, is pre-seed funded (South Park Commons, 2023), and discloses no revenue or ARR.

Short descriptiontext

VLM Run (Autonomi AI Inc.) operates an API-first visual intelligence platform whose Orion agent combines Vision-Language Models with specialized computer-vision tools to extract structured JSON from images, documents, video, and audio for developers and regulated enterprises.

Operating statusenum
Operating
Ownership categoryenum
Headcount rangeband
1–10
akta.pro rankint
HeadquartersPalo Alto, United States
HQ citystring
Palo Alto
HQ countrystring
United States
HQ regionstring
North America
Markets served

Serves global market

Offices1 record

Each record includes

City, Country, Type, Description, Source

Keyword5 values
visual intelligence platform, vision language models, document AI processing, multimodal AI API, structured data extraction
Industry2 codes
1End-to-End Enterprise AI Platforms (MLOps & Model Lifecycle Management)
CodeHDAEANAAPrimaryYes
2Model Development & Training Platforms (AutoML, Notebooks, Feature Stores)
CodeHDAEANABPrimaryNo
NAICS code1 code
  • Computer Systems Design and Related Services54151
SIC code1 code
  • Services-Prepackaged Software7372
Product category
Visual AI Platform
Social media profiles3 records
GTM motion1 record

Each record includes

Type, Description, Source

Revenue model3 records
1Credit-based API usage
TypeUsage Based
Description

Consumption-based model where customers purchase credits (100 credits = $1 USD) to make API calls. Different domains charge different rates per page/image/duration segment with processing levels (lo, auto, hi) affecting credit consumption.

docs.vlm.run
2Subscription Plans
TypeSubscription Recurring
Description

Pro plan at $799/month includes 100,000 credits plus usage-based overages. Enterprise plans offer custom invoiced billing with volume discounts and tier-based pricing.

docs.vlm.run
3Service Tier Multipliers
TypeUsage Based
Description

Flex tier (0.5x credits) for batch workloads and Priority tier (1.8x credits) for latency-sensitive workloads, applied uniformly on top of domain costs.

docs.vlm.run
Marketing channels7 records

Each record includes

Title, Type, Stage, Description, Source

Distribution channels4 records

Each record includes

Title, Type, Scope, Target buyer, Description, Source

Cost components5 values
Technology or R&D, Personnel, Infrastructure, Marketing or Sales, Operations
Pricing details3 tiers
1Starter - Free entry tier with pay-as-you-go scaling
ModelUsage-basedBilling cadencePay-as-you-go
Notes

Pay-as-you-go at 100 credits = $1 USD. From 1 credit per page/image. 1,000 free sign-up credits included. Up to 10 requests/min. Basic usage logs. Community Discord support.

docs.vlm.run
2Pro - $799/month with 100K included credits
ModelSubscriptionBilling cadenceMonthly
Notes

$799/month base subscription plus usage-based overages. 100,000 credits included. Up to 100 requests/min. Dedicated Slack support. Zero-Data Retention (ZDR). Business Associate Agreement (BAA). Up to 5 custom models.

docs.vlm.run
3Enterprise - Custom pricing for large-scale deployments
ModelSubscriptionBilling cadenceMulti-year contract
Notes

Custom invoiced billing with tier-based pricing and volume discounts. Unlimited credits. Custom rate limits. In-VPC deployments. SOC 2 Type II, HIPAA, BAA. Custom SLAs. Dedicated Slack + priority support.

docs.vlm.run
GTM typeB2B
B2B
Offering typeSoftware
Software
Core offering1 text field

VLM Run provides a unified visual intelligence platform, centered on its Orion visual agent, that processes images, video, documents, and audio into structured JSON outputs via an OpenAI-compatible REST API. The platform supports multi-step agentic visual workflows, structured outputs with Pydantic/Zod schemas, visual grounding with bounding boxes, and SOC 2 Type II / HIPAA-compliant deployment options including In-VPC installations.

Differentiator
Functional benefit
Problem solved
Quantifiable outcome1 of 2 values shown
  • 100 credits = $1 USD conversion rate
+1 more record
Product overview1 text field

VLM Run is a unified visual intelligence platform that provides AI infrastructure for processing images, documents, video, and audio. The flagship product is Orion, a visual agent that combines Vision-Language Model reasoning with specialized computer-vision tools through a unified API. The platform includes open-source tools (vlmrun-hub for structured schemas, vlmbench for benchmarking, mm for multi-modal context), industry-specific solutions for Healthcare and Construction, and a Skills system for reusable visual extraction workflows. Orion powers the core Chat Playground interface and Agents API, while Open Source tools enable custom VLM deployment and evaluation.

Product and service1 record
1Orion Visual Agent Platform
Scale indicator2 records

Each record includes

Type, Value, Description, Source

Partnership2 partners
Strategic tierCoreTypeTechnology or Integration
Description

VLM Run has a native integration with MongoDB for extracting structured JSON from visual content and storing it in MongoDB Atlas. The integration enables enterprises to seamlessly process and index unstructured visual data, transforming raw multi-modal information into valuable, queryable insights stored in MongoDB's document-oriented database.

Strategic tierCoreTypeChannel Partner/ Reseller/ Distributor
Description

VLM Run integration for Zapier enables connecting visual AI capabilities with over 8,000 apps. Automates visual data processing workflows including document parsing, image analysis, and content extraction into existing business processes. Pre-built templates available for Markdown Extraction and Invoice Extraction.

Recent move6 records

Each record includes

Date, Type, Title, Description, Source

Expansion highlight5 records

Each record includes

Type, Description

Peers10 records
TypeDirect peer
Description

Document AI for invoice and supply-chain workflows, with strong enterprise compliance certifications. Comparable to VLM Run's healthcare-adjacent and document-extraction positioning.

TypeDirect peer
Description

Enterprise visual AI platform offering image, video, and text understanding via API with custom model training and an AI workflow orchestrator. Closely comparable to VLM Run's API-first, multi-modal, enterprise-focused positioning.

TypeDirect peer
Description

API-first document parsing and ETL platform that extracts structured data from PDFs, images, and documents. Directly comparable to VLM Run's document intelligence and visual ETL positioning, with similar OpenAI-compatible API patterns.

TypeDirect peer
Description

Developer-focused computer vision platform offering training, deployment, and inference APIs for image and video models. Directly comparable as a PLG, API-first visual AI platform targeting developers and enterprises, though it leans more toward custom model training versus VLM-based reasoning.

5Azure AI Document Intelligence
TypeBroad incumbent
Description

Microsoft's document AI service for forms, invoices, and custom document extraction within the Azure stack. Competes with VLM Run's document intelligence offering on enterprise procurements already standardized on Azure.

TypeDirect peer
Description

Document AI platform extracting structured data from invoices, receipts, and forms with workflow automation. Comparable as a verticalized visual/document AI enterprise platform with usage-based pricing and SOC 2/HIPAA posture.

TypeDirect peer
Description

Enterprise computer vision platform (founded by Andrew Ng) targeting manufacturing, healthcare, and other verticals with visual inspection and document AI. Directly comparable as a verticalized visual AI enterprise platform.

TypeBroad incumbent
Description

Hyperscaler offering broad image, document (Document AI), and video AI APIs inside the Google Cloud ecosystem. Competes with VLM Run across the full multi-modal stack, with the advantage of bundled cloud consumption.

TypeOthers
Description

Data infrastructure for AI, providing labeling, evaluation, and testing services for vision and language models. Adjacent rather than directly competing — Scale is upstream of model training while VLM Run is downstream for inference and extraction.

TypeBroad incumbent
Description

Hyperscaler visual AI services covering image (Rekognition) and document text extraction (Textract). Comparable at the capability layer but without the agentic/structured-reasoning wedge Orion emphasizes.

Market position
Strengths5 records

Each record includes

Headline, Details, Source

Weaknesses5 records

Each record includes

Headline, Details, Source

Competitive moat5 records

Each record includes

Type, Details

Key risks6 records

Each record includes

Headline, Details, Source

Key highlights7 records

Each record includes

Headline, Details, Source

Customer concentration

Classification, Details

Named customers15 records

Each record includes

Name, Industry, Type, Use case, Source, UUID

Segment4 records

Each record includes

Title, Type, Primary, Description, Pain point addressed, Use case, Source

Ideal customer profile4 records

Each record includes

Profile, Firmographic size, Sales motion, Sales cycle length, Buying structure, Purchase trigger, Buyer persona, Geography, Industry vertical, Primary use case, Description, Pain points, Evidence proof points, Target buyer

Technology focused
Yes
API detail
Has APIbool
Yes

Docs URL, Description

Integration3 records

Each record includes

Title, Type, Description, Source

AI capability13 records

Each record includes

Type, Description, Source

AI maturity
App detail

Has app

Feature10 records

Each record includes

Title, Differentiator, Description, Source

Core technology
Revenue estimate
Valuation estimate
Number of profiles
Profiles2 records

Each record includes

Name, Designation, Designation category, Overview, Profile commentary, Source

No data
Compliance3 records

Each record includes

Name, Class, Description

Funding overview

Funding stage, Last funding date, Total funding USD

Funding rounds1 record

Each record includes

Round, Amount USD, Date, Pre money valuation, Total investors, Investors, News

Investors1 record

Each record includes

Name, Type, Date of entry, Rounds participated, Website

Funding detail is available on the Subscription and Enterprise plan.Contact sales →

M&A

Each record includes

Name, Acquisition type, Announced date, Completed date, Status, Website, News

Investment

Each record includes

Name, Round, Announced date, Lead investor, Website, News

M&A and investment is available on the Subscription and Enterprise plan.Contact sales →

VLM Run

Visual AI Platformvlm.run

VLM Run (Autonomi AI Inc.) operates an API-first visual intelligence platform whose Orion agent combines Vision-Language Models with specialized computer-vision tools to extract structured JSON from images, documents, video, and audio for developers and regulated enterprises.

What VLM Run does

VLM Run (legal entity Autonomi AI Inc.) is a Palo Alto-based private company founded in 2022 that operates an API-first visual intelligence platform built around its flagship product, Orion. Orion is described as a "visual agent" that combines the reasoning capability of large Vision-Language Models with the accuracy of specialized computer-vision tools, exposed through a single OpenAI-compatible REST API at api.vlm.run/v1. The platform ingests images, documents, video, and audio and outputs structured JSON, with support for custom schemas via Pydantic/Zod and visual grounding bounding boxes for auditability. VLM Run complements Orion with open-source projects (vlmrun-hub for domain schemas, vlmbench for VLM benchmarking, and mm for multi-modal context) and a Skills/Agents system for reusable, versioned extraction workflows.

The company targets two overlapping customer sets: developers and AI-native companies that want a drop-in OpenAI SDK replacement for visual tasks (logos include Voxel51, Langflow, MCP.run, York IE, MongoDB, Aurochs), and regulated enterprises that need SOC 2 Type II, HIPAA, BAA, and In-VPC deployment (logos include Valerie Health, Scan.com, basata, Peak Health in healthcare; Zapier, Make, n8n, Activepieces in automation; Google Cloud as a broader ecosystem signal). Vertical solutions are pre-packaged for Healthcare (faxes, scans, clinical paperwork) and Construction (blueprints, schedules, specs).

VLM Run monetizes on a credit-based usage model (100 credits = $1 USD), structured into three tiers: a free Starter tier with 1,000 sign-up credits and 10 requests/min, a Pro tier at $799/month including 100,000 credits and 100 requests/min, and a custom-priced Enterprise tier with unlimited credits, in-VPC deployment, and dedicated SLAs. Go-to-market is product-led growth for developers and direct enterprise sales for regulated deployments, supported by native integrations with Zapier (8,000+ apps) and MongoDB Atlas. The company has 1–10 employees, is pre-seed funded (South Park Commons, 2023), and discloses no revenue or ARR.

VLM Run firmographics

Firmographics
Name
VLM Run
Legal name
Autonomi AI Inc.
Website
https://vlm.run
Company type
Private
Founded year
2022
Operating status
Operating
Headcount range
1–10 employees
Short description
VLM Run (Autonomi AI Inc.) operates an API-first visual intelligence platform whose Orion agent combines Vision-Language Models with specialized computer-vision tools to extract structured JSON from images, documents, video, and audio for developers and regulated enterprises.
Ownership category
akta.pro rank

VLM Run industry classification

Industry
Product category
Visual AI Platform
NAICS
Computer Systems Design and Related Services (54151)
SIC
Services-Prepackaged Software (7372)
akta.pro primary industry
End-to-End Enterprise AI Platforms (MLOps & Model Lifecycle Management) (HDAEANAA)
akta.pro secondary industry
Model Development & Training Platforms (AutoML, Notebooks, Feature Stores) (HDAEANAB)

Keywords

  • Visual intelligence platform
  • Vision language models
  • Document AI processing
  • Multimodal AI API
  • Structured data extraction

Where VLM Run is headquartered

Location

Headquarters

HQ city
Palo Alto
HQ country
United States
HQ region
North America

Offices1 record

Markets served

VLM Run business model

Business model
GTM type
B2B
Offering type
Software
Cost components
Technology or R&D, Personnel, Infrastructure, Marketing or Sales, Operations

Revenue model

  1. Credit-based API usage: Consumption-based model where customers purchase credits (100 credits = $1 USD) to make API calls. Different domains charge different rates per page/image/duration segment with processing levels (lo, auto, hi) affecting credit consumption.
  2. Subscription Plans: Pro plan at $799/month includes 100,000 credits plus usage-based overages. Enterprise plans offer custom invoiced billing with volume discounts and tier-based pricing.
  3. Service Tier Multipliers: Flex tier (0.5x credits) for batch workloads and Priority tier (1.8x credits) for latency-sensitive workloads, applied uniformly on top of domain costs.

Pricing tiers

ModelBillingPrice
Usage-basedPay-as-you-goStarter - Free entry tier with pay-as-you-go scaling
SubscriptionMonthlyPro - $799/month with 100K included credits
SubscriptionMulti-year contractEnterprise - Custom pricing for large-scale deployments

Go-to-market motion1 record

Distribution channels4 records

Marketing channels7 records

VLM Run product offering

Product offering

Core offering

VLM Run provides a unified visual intelligence platform, centered on its Orion visual agent, that processes images, video, documents, and audio into structured JSON outputs via an OpenAI-compatible REST API. The platform supports multi-step agentic visual workflows, structured outputs with Pydantic/Zod schemas, visual grounding with bounding boxes, and SOC 2 Type II / HIPAA-compliant deployment options including In-VPC installations.

Product overview

VLM Run is a unified visual intelligence platform that provides AI infrastructure for processing images, documents, video, and audio. The flagship product is Orion, a visual agent that combines Vision-Language Model reasoning with specialized computer-vision tools through a unified API. The platform includes open-source tools (vlmrun-hub for structured schemas, vlmbench for benchmarking, mm for multi-modal context), industry-specific solutions for Healthcare and Construction, and a Skills system for reusable visual extraction workflows. Orion powers the core Chat Playground interface and Agents API, while Open Source tools enable custom VLM deployment and evaluation.

Differentiator

Problem solved

Functional benefit

Products and services

  • Orion Visual Agent Platform

Quantifiable outcome

  • 100 credits = $1 USD conversion rate
  • +1 more outcomes

Companies that use VLM Run

Customer profile

Named customers15 records

Segments4 records

Ideal customer profiles4 records

VLM Run technology and API

Technology

Technology focussed Yes

API detail

Has API
Yes
API docs
API detail

Core technology

AI maturity

App detail

Integration3 records

AI capability13 records

Feature10 records

VLM Run partnerships and signals

Strategic signal

Partnerships

Two partnerships are on record, tiered core.

  • MongoDBcoreTechnology or IntegrationVLM Run has a native integration with MongoDB for extracting structured JSON from visual content and storing it in MongoDB Atlas. The integration enables enterprises to seamlessly process and index unstructured visual data, transforming raw multi-modal information into valuable, queryable insights stored in MongoDB's document-oriented database.
  • ZapiercoreChannel Partner/ Reseller/ DistributorVLM Run integration for Zapier enables connecting visual AI capabilities with over 8,000 apps. Automates visual data processing workflows including document parsing, image analysis, and content extraction into existing business processes. Pre-built templates available for Markdown Extraction and Invoice Extraction.

Scale indicators2 records

Recent moves6 records

Expansion highlights5 records

VLM Run competitors and assessment

Company assessment

Direct peers

  • Rossum: Document AI for invoice and supply-chain workflows, with strong enterprise compliance certifications. Comparable to VLM Run's healthcare-adjacent and document-extraction positioning.
  • Clarifai: Enterprise visual AI platform offering image, video, and text understanding via API with custom model training and an AI workflow orchestrator. Closely comparable to VLM Run's API-first, multi-modal, enterprise-focused positioning.
  • Unstructured: API-first document parsing and ETL platform that extracts structured data from PDFs, images, and documents. Directly comparable to VLM Run's document intelligence and visual ETL positioning, with similar OpenAI-compatible API patterns.
  • Roboflow: Developer-focused computer vision platform offering training, deployment, and inference APIs for image and video models. Directly comparable as a PLG, API-first visual AI platform targeting developers and enterprises, though it leans more toward custom model training versus VLM-based reasoning.
  • Nanonets: Document AI platform extracting structured data from invoices, receipts, and forms with workflow automation. Comparable as a verticalized visual/document AI enterprise platform with usage-based pricing and SOC 2/HIPAA posture.
  • Landing AI: Enterprise computer vision platform (founded by Andrew Ng) targeting manufacturing, healthcare, and other verticals with visual inspection and document AI. Directly comparable as a verticalized visual AI enterprise platform.

Broad incumbents

  • Azure AI Document Intelligence: Microsoft's document AI service for forms, invoices, and custom document extraction within the Azure stack. Competes with VLM Run's document intelligence offering on enterprise procurements already standardized on Azure.
  • Google Cloud Vision AI: Hyperscaler offering broad image, document (Document AI), and video AI APIs inside the Google Cloud ecosystem. Competes with VLM Run across the full multi-modal stack, with the advantage of bundled cloud consumption.
  • AWS Rekognition and Textract: Hyperscaler visual AI services covering image (Rekognition) and document text extraction (Textract). Comparable at the capability layer but without the agentic/structured-reasoning wedge Orion emphasizes.

Others

  • Scale AI: Data infrastructure for AI, providing labeling, evaluation, and testing services for vision and language models. Adjacent rather than directly competing — Scale is upstream of model training while VLM Run is downstream for inference and extraction.

Market position

Strengths5 records

Weaknesses5 records

Competitive moat5 records

Key risks6 records

Key highlights7 records

Customer concentration

VLM Run social profiles

Digital presence

VLM Run compliance and trust

Trust signal

Compliance3 records

VLM Run financial estimates

Financial estimate

Revenue estimate

Valuation estimate

VLM Run leadership team

Management profile

Number of profiles

Profiles2 records

VLM Run funding detail

Funding detail

Funding overview

Funding rounds1 record

Investors1 record

Funding detail is available on the Subscription and Enterprise plan.Contact sales →

VLM Run M&A and investment

M&A and investment

M&A

Investments

M&A and investment is available on the Subscription and Enterprise plan.Contact sales →

Frequently asked questions about VLM Run

What does VLM Run do?

VLM Run provides a unified visual intelligence platform, centered on its Orion visual agent, that processes images, video, documents, and audio into structured JSON outputs via an OpenAI-compatible REST API. The platform supports multi-step agentic visual workflows, structured outputs with Pydantic/Zod schemas, visual grounding with bounding boxes, and SOC 2 Type II / HIPAA-compliant deployment options including In-VPC installations.

Is VLM Run a public or private company?

VLM Run is a private company. It is classified as founder individual operated bootstrapped and is currently operating.

When was VLM Run founded?

VLM Run was founded in 2022. It employs 1 to 10 people.

Where is VLM Run based?

VLM Run is headquartered in Palo Alto, United States, in the North America region.

How does VLM Run make money?

Three revenue lines are on record. Credit-based API usage is the primary driver. The others are subscription Plans and service Tier Multipliers.

Who are VLM Run's main competitors?

Direct peers on record are Rossum, Clarifai, Unstructured, Roboflow, Nanonets and Landing AI. Broad incumbents are Azure AI Document Intelligence, Google Cloud Vision AI and AWS Rekognition and Textract. Scale AI is listed as an others.

Does VLM Run have an API?

Yes. VLM Run offers a REST API (base URL: https://api.vlm.run/v1) with OpenAI-compatible endpoints for visual AI processing. The API supports structured extraction from images, documents, audio, and video, as well as chat completions and agent executions. Authentication uses Bearer tokens (X-Api-Key header). Available SDKs in Python and Node.js. Supports streaming, webhooks, and structured outputs via Pydantic/Zod. Three service tiers: standard (baseline rates), flex (50% discount, higher latency), and priority (1.8x premium). Rate limits vary by plan: 10 requests/min (Standard), 100 requests/min (Pro), unlimited (Enterprise). Developer documentation is at docs.vlm.run/introduction.

What industry is VLM Run in?

VLM Run's product category is Visual AI Platform. Its primary akta.pro industry code is HDAEANAA, End-to-End Enterprise AI Platforms (MLOps & Model Lifecycle Management), with a secondary code of HDAEANAB, Model Development & Training Platforms (AutoML, Notebooks, Feature Stores). Its NAICS code is 54151 and its SIC code is 7372.

Unlock the full company data

50 free credits on sign-up, no credit card required.

Contact sales
Live signals
RuntimewireVLM Run opens a vision-model gateway for OCR, documents and videoVLM Run opened its Gateway API to the public on September 1, providing an OpenAI-compatible interface for open-weight OCR, vision-language, and video models. The service handles PDF rasterization, retries, and normalized outputs, with a free alpha requiring no sign-up. The company bets that document and video plumbing will become the durable product as models change.HuggingfaceVLM Run Gateway: Run GLM-OCR, DeepSeek-OCR-2, dots.mocr with an OpenAI Compatible APIVLM Run has launched the VLM Run Gateway, an OpenAI-compatible API endpoint designed to facilitate the deployment of various open-weight OCR and Vision-Language Models. The platform allows developers to switch between models such as DeepSeek-OCR-2, GLM-OCR, and dots.mocr to achieve significant cost reductions compared to frontier VLMs while maintaining accuracy for document parsing tasks.YorkThe Next AI Wave: Why Vertical AI Will DominateThe article argues that the second wave of artificial intelligence will be dominated by vertical AI, which utilizes proprietary industry-specific data to deliver higher accuracy and compliance compared to general-purpose horizontal models. It highlights that this shift offers stronger ROI, lower churn, and better regulatory alignment for sectors like healthcare, finance, and legal services. Key examples of emerging vertical AI companies mentioned include VLM Run, Alivo, and Givzey.