Webrecorder
Webrecorder builds open-source web archiving tools that capture, store, and replay interactive websites, serving libraries, universities, government agencies, and cultural institutions worldwide through its Browsertrix SaaS, ArchiveWeb.page extension, ReplayWeb.page player, and the WACZ open standard.
- Company typePrivate
- Founded2014
- HeadquartersSan Francisco, United States
- Headcount1–10
- GTM typeB2B and B2C
- OfferingSoftware
What Webrecorder does
Webrecorder Software LLC is a US-based, privately held software company that builds open-source web archiving tools for capturing, replaying, and preserving interactive web content. Its technology stack centers on browser-based archiving that uses real browser rendering to capture dynamic, JavaScript-heavy content — including pages behind logins and paywalls — and exports archives in the open WACZ (Web Archive Collection Zipped) format developed by the company, which is listed as a sustainable digital format by the Library of Congress. Its core products are Browsertrix (a hosted SaaS for automated browser-based crawling at scale with real-time monitoring, crawl deduplication, and machine-assisted quality assurance), ArchiveWeb.page (a free Chrome extension and desktop app for manual capture), and ReplayWeb.page (serverless browser-side playback of WARC/WACZ files), supported by developer tools (Browsertrix Crawler, pywb, py-wacz) under permissive open-source licenses. The company, spun out of Rhizome in 2020 after founding in 2014 by Ilya Kreymer, runs on a hybrid go-to-market combining product-led growth (self-serve sign-up, Chrome Web Store distribution, free open-source tooling) with enterprise field sales for Pro and Enterprise tiers that include on-premise deployment and dedicated support; revenue is generated primarily from tiered subscriptions (individual plans from $30–$120/month, Pro/Enterprise custom annual contracts) and professional services, with the rest of the user base served through free and open-source tools. Webrecorder's customers are predominantly libraries, archives, universities, government agencies, and arts and human-rights organizations worldwide — including Internet Archive, UK Web Archive, Arquivo.pt, IIPC member organizations, Stanford University Press, and Perma.cc — placing the company in the digital preservation and compliance market where the Pew Research Center cited that 38% of webpages from 2013 are no longer accessible.
Webrecorder firmographics
Firmographics- Name
- Webrecorder
- Legal name
- Webrecorder Software LLC
- Website
- https://webrecorder.net
- Company type
- Private
- Founded year
- 2014
- Operating status
- Operating
- Headcount range
- 1–10 employees
- Short description
- Webrecorder builds open-source web archiving tools that capture, store, and replay interactive websites, serving libraries, universities, government agencies, and cultural institutions worldwide through its Browsertrix SaaS, ArchiveWeb.page extension, ReplayWeb.page player, and the WACZ open standard.
- Ownership category
- akta.pro rank
Webrecorder industry classification
Industry- Product category
- Web Archiving Software
- NAICS
- Web Search Portals and All Other Information Services (51929), Libraries and Archives (51921), Computing Infrastructure Providers, Data Processing, Web Hosting, and Related Services (5182)
- SIC
- Services-Computer Programming, Data Processing, Etc. (7370), Services-Prepackaged Software (7372)
- akta.pro primary industry
- Web Archiving, Cache & Snapshot Services (BPAMAAAL)
- akta.pro secondary industries
- Browser & Web Navigation Utilities (toolbars, new-tab, bookmarks) (BPAMAAAH), News & Content Aggregation Portals (BPAMAAAI)
Keywords
Where Webrecorder is headquartered
LocationHeadquarters
- HQ city
- San Francisco
- HQ country
- United States
- HQ region
- North America
Offices1 record
Markets served
Webrecorder business model
Business model- GTM type
- B2B and B2C
- Offering type
- Software
Revenue model
- Browsertrix Cloud Subscriptions: Tiered SaaS subscription model for hosted Browsertrix platform. Individual plans range from $30-120/month for varying execution time, storage, concurrency, and page limits. Pro and Enterprise tiers offer custom pricing with dedicated support and advanced features.
- Professional Services: Expert web archiving services for Enterprise tier including dedicated account managers, expert archivists, custom crawl reports, and on-premise deployment support.
- Open Source Tools (Free): Free browser extension (ArchiveWeb.page) and ReplayWeb.page for individual users. pywb Python toolkit and other developer tools are freely available under open source licenses.
Pricing tiers
| Model | Billing | Price |
|---|---|---|
| Subscription | Monthly | Starter: Free 7-day trial, then $30/month for individuals |
| Subscription | Monthly | Standard: $60/month for individuals with moderate needs |
| Subscription | Monthly | Plus: $120/month for power users |
| Subscription | Annual | Pro: Custom pricing for web archiving teams |
| Subscription | Annual | Enterprise: Full-service custom experience |
Go-to-market motion2 records
Distribution channels4 records
Marketing channels10 records
Webrecorder product offering
Product offeringCore offering
Webrecorder develops open-source and SaaS web archiving tools that allow users to capture, store, and replay interactive websites and web sessions. Its flagship product, Browsertrix, is a hosted cloud platform that uses real browsers to crawl dynamic, JavaScript-heavy websites, including pages behind logins, and to perform quality assurance on archived content. Complementary tools include ArchiveWeb.page (browser extension/desktop app for manual capture) and ReplayWeb.page (serverless playback of WARC/WACZ archives).
Product overview
Webrecorder provides a comprehensive suite of open source web archiving tools designed to make digital preservation accessible to everyone. The core product portfolio consists of Browsertrix (a cloud-hosted SaaS archiving platform with automated crawling and quality assurance), ArchiveWeb.page (a free browser extension for manual capture), and ReplayWeb.page (a serverless playback system for web archives). The products work together within the Webrecorder ecosystem: Browsertrix provides automated enterprise-grade archiving with collaborative features, ArchiveWeb.page enables manual capture directly from the browser, and ReplayWeb.page delivers high-fidelity replay of archived content without server requirements. All tools use and support the open WACZ (Web Archive Collection Zipped) format, a portable standard developed by Webrecorder. The developer tools include Browsertrix Crawler (command-line application), pywb (Python replay toolkit), and py-wacz (WARC-to-WACZ converter). The organization, founded in 2014 and spun off from Rhizome in 2020, maintains its commitment to open source principles while offering commercial hosted services.
Differentiator
Problem solved
Functional benefit
Brands
- Browsertrix: Automated browser-based web crawling platform at scale, including a hosted SaaS service.
- ArchiveWeb.page
- ReplayWeb.page
Products and services
- Browsertrix Cloud-hosted SaaS platform providing automated browser-based web crawling at scale, with high-fidelity capture of dynamic websites, real-time monitoring, quality assurance tools, deduplication, and collaborative workspace features.
- ArchiveWeb.page Free Chrome extension and desktop application that records complex web interactions while browsing; archives are stored locally and can be downloaded as WACZ files, with experimental IPFS peer-to-peer sharing.
- ReplayWeb.page Serverless web archive playback system that enables viewing and embedding WARC/WACZ files directly in the browser without specialized server infrastructure.
- Browsertrix Crawler Open-source command-line application responsible for all crawling operations in Browsertrix; available for self-hosted deployments and programmatic integration via API.
- pywb Open-source, production-grade Python-based web archive replay system; recommended for adoption by the International Internet Preservation Consortium (IIPC).
- py-wacz Command-line application for creating and validating WACZ archive files, providing lossless WARC-to-WACZ conversion.
- Webrecorder Player Webrecorder's first desktop application for recording and replaying web archives; legacy product superseded by newer tools.
Quantifiable outcome
- Captures 100% of dynamically loaded content that renders in a browser
- +1 more outcomes
Companies that use Webrecorder
Customer profileNamed customers9 records
Segments5 records
Webrecorder technology and API
TechnologyTechnology focussed Yes
API detail
- Has API
- Yes
- API docs
- API detail
Core technology
AI maturity
App detail
Integration1 record
Feature7 records
Webrecorder partnerships and signals
Strategic signalPartnerships
Five partnerships are on record, tiered minor and core.
- Digital OceanminorInfrastructure provider for Browsertrix hosted service. Listed as approved subcontractor for data processing. Digital Ocean hosts cloud infrastructure for the SaaS platform.
- WasabiminorObject storage provider for Browsertrix. Listed as approved subcontractor for storing archived web content and WACZ files.
- RhizomecoreOriginal home of Webrecorder project. Arts organization focused on digital art and culture. Webrecorder spun off as independent organization from Rhizome in 2020.
- IIPC (International Internet Preservation Consortium)coreGlobal consortium of libraries and archives working on web preservation. Recommended pywb to member organizations. Webrecorder tools built to IIPC standards.
- StripeminorPayment processor for Browsertrix subscriptions. Handles payment processing so Webrecorder never sees customer payment information directly.
Scale indicators4 records
Recent moves7 records
Expansion highlights5 records
Webrecorder competitors and assessment
Company assessmentDirect peers
- Internet Archive: The largest nonprofit digital preservation organization and a direct customer of Webrecorder. Operates its own web archiving tools (Heritrix, Wayback Machine) and is both a peer and ecosystem partner, competing in open-source crawling while endorsing Webrecorder's replay and capture formats.
- Hanzo: Enterprise web archiving vendor focused on compliance, legal hold, and eDiscovery use cases. Directly comparable to Browsertrix Enterprise tier and competes for the same institutional/professional archiving budgets.
- PageFreezer: Commercial web and social media archiving platform serving compliance, legal, and regulated industries. Competes head-to-head with Browsertrix Enterprise in the same buyer segment (large institutions with compliance mandates).
- Archive.today: Independent web archiving service that lets users capture and replay snapshots of webpages. Directly comparable to ArchiveWeb.page and ReplayWeb.page as a consumer-facing archiving tool with overlap in format support and target audience.
- Conifer (Rhizome): The original Webrecorder.io product, retained by Rhizome after Webrecorder spun off in 2020. Direct historical and functional peer offering browser-based web archiving with a focus on arts and cultural heritage content.
- Perma.cc: Service that creates permanent archives of web references for legal and academic citations. Operates in the same web archiving space, serves a similar audience (law schools, courts, journals), and is itself a Webrecorder customer.
Emerging players
- Visualping: Website change detection and monitoring platform. Lightweight competitor for users who primarily need to track changes to webpages rather than capture fully interactive archives, representing a partial-overlap use case.
- Stillio: Automated website screenshot and change-monitoring service. Adjacent rather than directly competing — captures visual snapshots rather than full interactive archives — but serves overlapping use cases (compliance, change tracking, website backup).
Broad incumbents
- Smarsh: Large compliance archiving and communications intelligence platform. Competes indirectly in the broader digital archiving market, especially in regulated industries where web archiving is one component of a wider retention solution.
- Proofpoint Archive: Enterprise compliance archiving solution covering email, social, and web content. A broader incumbent whose web archiving capabilities overlap with Browsertrix Enterprise in regulated verticals.
Market position
Strengths4 records
Weaknesses4 records
Competitive moat5 records
Key risks6 records
Key highlights7 records
Customer concentration
Webrecorder social profiles
Digital presenceWebrecorder financial estimates
Financial estimateRevenue estimate
Valuation estimate
Webrecorder leadership team
Management profileNumber of profiles
Webrecorder funding detail
Funding detailFunding overview
Funding rounds
Investors
Funding detail is available on the Subscription and Enterprise plan.Contact sales →
Webrecorder M&A and investment
M&A and investmentM&A
Investments
M&A and investment is available on the Subscription and Enterprise plan.Contact sales →
Frequently asked questions about Webrecorder
What does Webrecorder do?
Webrecorder develops open-source and SaaS web archiving tools that allow users to capture, store, and replay interactive websites and web sessions. Its flagship product, Browsertrix, is a hosted cloud platform that uses real browsers to crawl dynamic, JavaScript-heavy websites, including pages behind logins, and to perform quality assurance on archived content. Complementary tools include ArchiveWeb.page (browser extension/desktop app for manual capture) and ReplayWeb.page (serverless playback of WARC/WACZ archives).
Is Webrecorder a public or private company?
Webrecorder is a private company. It is classified as founder individual operated bootstrapped and is currently operating.
When was Webrecorder founded?
Webrecorder was founded in 2014. It employs 1 to 10 people.
Where is Webrecorder based?
Webrecorder is headquartered in San Francisco, United States, in the North America region.
How does Webrecorder make money?
Three revenue lines are on record. Browsertrix Cloud Subscriptions are the primary driver. The others are professional Services and open Source Tools (Free).
Who are Webrecorder's main competitors?
Direct peers on record are Internet Archive, Hanzo, PageFreezer, Archive.today, Conifer (Rhizome) and Perma.cc. Emerging players are Visualping and Stillio. Broad incumbents are Smarsh and Proofpoint Archive.
Does Webrecorder have an API?
Yes. Browsertrix offers an API that allows developers to integrate crawling capabilities into their own scripts and applications. The command-line application Browsertrix Crawler can be integrated programmatically via the API. Developer documentation is at docs.browsertrix.com/api.
What industry is Webrecorder in?
Webrecorder's product category is Web Archiving Software. Its primary akta.pro industry code is BPAMAAAL, Web Archiving, Cache & Snapshot Services, with a secondary code of BPAMAAAH, Browser & Web Navigation Utilities (toolbars, new-tab, bookmarks). Its NAICS code is 51929 and its SIC code is 7370.