The Unicode Consortium
The Unicode Consortium is a 501(c)(3) non-profit standards body that develops and maintains the Unicode Standard, the universal character encoding system deployed across 20+ billion devices. It serves global technology companies, developers, and the localization industry through open-source specifications, libraries, and locale data.
- Company typePrivate
- Founded1988
- HeadquartersMountain View, United States
- Headcount1–10
- GTM typeB2B
- OfferingServices
What The Unicode Consortium does
The Unicode Consortium is a 501(c)(3) non-profit, open-standards body that develops, maintains, and promotes the Unicode Standard — the universal character encoding system underlying text and emoji on every major operating system and more than 20 billion devices worldwide. Its core technology portfolio includes the Unicode Standard itself (Version 17.0 with 159,801 characters across 150+ scripts), the Unicode Character Database (UCD), Unicode Code Charts, ICU (C/C++ and Java internationalization libraries), the Common Locale Data Repository (CLDR) for locale-specific formatting, and the Unicode Emoji Standards (UTS #51). The organization is synchronized with ISO/IEC 10646 and coordinates CJK character research through the hosted Ideographic Research Group (IRG). Public review of proposed changes follows a formal Public Review Issues (PRI) process documented in the UTC Document Registry.
The Consortium's product architecture is entirely open source: all standards, specifications, code charts, and libraries are distributed free of charge under the Unicode Terms of Use and the OSI-approved Unicode License v3, with no commercial products or paid APIs. Its revenue model is non-commercial, consisting of tiered Organizational Membership dues (Full, Supporting, Associate) paid by technology companies, plus Individual Memberships, donations, gifts of stock, and the Adopt-a-Character sponsorship program. Named organizational members include Apple, Google, Microsoft, Adobe, Meta, Airbnb, Salesforce, and Translated, whose senior representatives sit on the Board of Directors and participate in the Unicode Technical Committee (UTC), CLDR technical committee, and ICU technical committee. Go-to-market is community-led and indirect: adoption is driven by the necessity of Unicode compliance in any software product that handles text.
The organization is governed by a Board of Directors chaired by Cathy Wissink (also Interim CTO), with CEO Toral Cowieson leading a lean core staff of 1–10 employees that orchestrates contributions from hundreds of expert volunteers, linguists, and language professionals worldwide. Headquartered in South San Francisco, California, the Consortium has operated continuously since its founding in 1988 and incorporation in 1991, releasing new major versions of the Unicode Standard on an annual cadence.
The Unicode Consortium firmographics
Firmographics- Name
- The Unicode Consortium
- Legal name
- The Unicode Consortium
- Website
- https://unicode.org
- Company type
- Private
- Founded year
- 1988
- Operating status
- Operating
- Headcount range
- 1–10 employees
- Short description
- The Unicode Consortium is a 501(c)(3) non-profit standards body that develops and maintains the Unicode Standard, the universal character encoding system deployed across 20+ billion devices. It serves global technology companies, developers, and the localization industry through open-source specifications, libraries, and locale data.
- Ownership category
- akta.pro rank
The Unicode Consortium industry classification
Industry- Product category
- Character Encoding & Internationalization Standards
- NAICS
- Religious, Grantmaking, Civic, Professional, and Similar Organizations (813), Professional Organizations (813920), Business Associations (813910)
- SIC
- Services-Membership Organizations (8600), Services-Computer Programming, Data Processing, Etc. (7370)
- akta.pro primary industry
- International Standards Development & Standardization Bodies (BPADANAA)
Keywords
Where The Unicode Consortium is headquartered
LocationHeadquarters
- HQ city
- Mountain View
- HQ country
- United States
- HQ region
- North America
Offices1 record
Markets served
The Unicode Consortium business model
Business model- GTM type
- B2B
- Offering type
- Services
- Cost components
- Personnel, Operations, Technology or R&D, Marketing or Sales, Others
Revenue model
- Organizational Membership Fees: The Consortium is funded by membership fees from organizational members (Full, Supporting, Associate tiers). Members include technology companies, academic institutions, and governmental organizations that participate in Unicode's governance and technical work.
- Individual Membership and Donations: Individual members and donors contribute to support Unicode's mission. The Consortium accepts donations, gifts of stock, and has an Adopt-a-Character sponsorship program.
Go-to-market motion1 record
Distribution channels3 records
Marketing channels8 records
The Unicode Consortium product offering
Product offeringCore offering
The Unicode Consortium is a non-profit standards body that develops, maintains, and promotes the Unicode Standard, a universal character encoding system covering 159,801 characters across hundreds of scripts and languages. It distributes its standards, specifications, code charts, the Unicode Character Database (UCD), Common Locale Data Repository (CLDR), and International Components for Unicode (ICU) software libraries freely to operating systems, browsers, applications, and devices worldwide.
Product overview
The Unicode Consortium is a non-profit open source, open standards body that maintains the Unicode Standard—the worldwide foundation for character encoding. The organization offers a portfolio of interconnected standards, specifications, and tools including: The Unicode Standard (core character encoding with 159,801 characters), Unicode Code Charts (PDF reference for all characters), International Components for Unicode (ICU - C/C++/Java libraries), Common Locale Data Repository (CLDR - locale data), Unicode Character Database (UCD - character properties), Unicode Emoji Standards (emoji specifications), Unicode Standard Annexes (technical specifications), UTC Document Registry (proposals and meeting records), and Public Review Issues (formal review process). These products work together as an integrated ecosystem enabling global software internationalization.
Differentiator
Problem solved
Functional benefit
Brands
- Common Locale Data Repository (CLDR): Unicode's locale data repository project for localization support
- ICU (International Components for Unicode)
- Unicode Emoji
Products and services
- The Unicode Standard The core character encoding standard providing a unique number for every character regardless of platform, program, or language. Version 17.0 contains 159,801 characters and supports over 150 scripts, serving software developers, operating system vendors, and platform implementers.
- Unicode Code Charts PDF charts showing representative glyphs for all Unicode characters organized by scripts and blocks, including delta charts for new additions and archival charts for each version release, for software developers and implementers.
- International Components for Unicode (ICU) A mature, widely-used set of C/C++ and Java libraries providing Unicode and globalization support for software applications, including collation, formatting, and conversion facilities for developers building internationalized software.
- Common Locale Data Repository (CLDR) A repository of locale data providing key building blocks for software to support the languages, countries, and scripts of the world. Used by major operating systems and applications for date, time, number, and currency formatting.
- Unicode Character Database (UCD) Machine-readable data file containing all the character properties and relationships for Unicode characters, serving as the comprehensive reference for implementing the Unicode Standard in software and platforms.
- Unicode Emoji Standards Technical specifications and data for emoji characters including emoji sequences, ZWJ sequences, and emoji variation selectors, governed by UTS #51, for platform vendors implementing consistent emoji rendering across devices.
- Unicode Standard Annexes (UAX/UTS) Detailed technical specifications covering specific aspects of Unicode including bidirectional text (UAX #9), normalization (UAX #15), line breaking (UAX #14), and emoji specifications (UTS #51) for software developers and implementers.
- UTC Document Registry The official document registry for the Unicode Technical Committee containing all proposals, meeting minutes, and technical documents dating back to 1991, for committee members and standards observers.
- Public Review Issues (PRI) Formal public review process for proposed changes to the Unicode Standard, where each PRI has a title, deadline, and allows industry feedback before technical decisions.
- Unicode Emoji Proposals Formal submission process for proposing new emoji characters, open for submissions from April to July each year with detailed guidelines for proposal preparation, used by the public to suggest new emoji.
Quantifiable outcome
- Reduced engineering costs: Organizations using Unicode libraries avoid duplicating millions of hours of development work
- +2 more outcomes
Companies that use The Unicode Consortium
Customer profileNamed customers1 record
Segments4 records
Ideal customer profiles3 records
The Unicode Consortium technology and API
TechnologyTechnology focussed Yes
API detail
- Has API
- No
- API docs
- API detail
Core technology
AI maturity
App detail
Feature5 records
The Unicode Consortium partnerships and signals
Strategic signalPartnerships
Ten partnerships are on record, tiered core.
- ApplecoreApple is an organizational member of the Unicode Consortium. Apple implements Unicode standards in iOS, macOS, and all Apple platforms. Apple participates in the Unicode Technical Committee through representatives like Kulpreet Chilana (Director of Software Localization at Apple) who manages teams evangelizing localization across Apple platforms.
- GooglecoreGoogle is an organizational member represented by Bob Jung (Director of Engineering for Internationalization). Google develops and maintains internationalization technologies and infrastructure across Google products using Unicode standards.
- MicrosoftcoreMicrosoft is an organizational member represented on the Board of Directors by Vishal Chowdhary (VP of Science, Office AI Science). Microsoft has been a long-time participant in Unicode's technical committees, including UTC participation from 2000-2005 by Cathy Wissink.
- AdobecoreAdobe is an organizational member represented by Brent Getlin (Director of Product Development and General Manager for Adobe Fonts and Type). Adobe contributes to typography and font-related Unicode initiatives.
- MetacoreMeta is an organizational member represented by Cori Alcorn (Director of Program Management for Internationalization and Product Quality). Meta contributes to global-ready software and localization solutions.
- AirbnbcoreAirbnb is an organizational member represented by Salvatore Giammarresi (Head of Localization). Airbnb contributes to localization and internationalization standards development.
- SalesforcecoreSalesforce is an organizational member represented by Teresa Marshall (VP of Globalization). Salesforce contributes to product globalization and localization initiatives.
- TranslatedcoreTranslated is an organizational member represented by John Tinsley (VP of AI Solutions). Translated contributes expertise in language technology and machine translation to Unicode initiatives.
- ISO/IEC JTC1/SC2/WG2coreUnicode is synchronized with ISO/IEC 10646, the International Standard for character encoding. WG2 maintains roadmaps and allocations of standardized characters jointly with Unicode.
- Ideographic Research Group (IRG)coreIRG is hosted at Unicode and coordinates CJK (Chinese, Japanese, Korean) character encoding research and proposals.
Scale indicators4 records
Recent moves6 records
Expansion highlights6 records
The Unicode Consortium competitors and assessment
Company assessmentDirect peers
- World Wide Web Consortium (W3C): International non-profit standards body that develops open web standards (HTML, CSS, WebAssembly) under a member-supported governance model with Member-funded staff. Directly comparable as a peer non-profit consortium serving the global software industry with similar member-fee economics and standards-development processes.
- Internet Engineering Task Force (IETF): Open-standards organization that develops internet protocols (TCP/IP, HTTP, TLS) under a community-driven, volunteer-heavy model with minimal paid staff. Closely comparable to Unicode in operating philosophy and volunteer-driven open standards production.
- Ecma International: Industry association that standardizes information and communication systems (ECMAScript/JavaScript, C#, Office Open XML). Directly comparable as a member-funded non-profit producing widely adopted tech standards with similar corporate-member composition.
- WHATWG: Community that maintains the HTML and DOM Living Standards, with members drawn from Apple, Google, Microsoft, and Mozilla. Comparable to Unicode in being a non-accredited web/tech standards community funded by the same large browser vendors.
- OASIS (Organization for the Advancement of Structured Information Standards): Non-profit standards consortium that develops open standards for security, cloud, IoT, and blockchain (e.g., SAML, KMIP). Comparable as a member-funded standards body with overlapping member rosters in enterprise software and similar open-process governance.
Broad incumbents
- ISO (International Organization for Standardization): Umbrella international standards body of which ISO/IEC 10646 (the standard to which Unicode is synchronized) is a part. Comparable as a standards-setting organization and the de jure counterpart to Unicode's de facto character-encoding work.
- IEEE Standards Association: Large established standards organization producing widely adopted technical standards (Wi-Fi, Ethernet, 802.x family). Comparable as a broader, more diversified standards body that overlaps with Unicode's technology-standards role but on a much wider portfolio.
Others
- The Internet Society (ISOC): Non-profit that supports and promotes the open internet, including organizational and financial backing for the IETF. Comparable as a non-profit standards-adjacent organization in the same ecosystem, and a former employer of Unicode CEO Toral Cowieson.
- Linux Foundation: Non-profit that hosts critical open-source infrastructure projects (Linux, Kubernetes, Node.js) under a vendor-neutral governance model. Comparable as a non-profit steward of foundational open-source infrastructure with similar corporate-member support.
Emerging players
- Internet Corporation for Assigned Names and Numbers (ICANN): Non-profit that coordinates the DNS, IP addresses, and unique internet identifiers under a multi-stakeholder governance model. Comparable as a non-profit coordinator of unique global identifiers, paralleling Unicode's role in coordinating unique character code points.
Market position
Strengths4 records
Weaknesses4 records
Competitive moat6 records
Key risks6 records
Key highlights7 records
Customer concentration
The Unicode Consortium social profiles
Digital presenceThe Unicode Consortium financial estimates
Financial estimateRevenue estimate
Valuation estimate
The Unicode Consortium leadership team
Management profileNumber of profiles
Profiles13 records
The Unicode Consortium funding detail
Funding detailFunding overview
Funding rounds
Investors
Funding detail is available on the Subscription and Enterprise plan.Contact sales →
The Unicode Consortium M&A and investment
M&A and investmentM&A
Investments
M&A and investment is available on the Subscription and Enterprise plan.Contact sales →
Frequently asked questions about The Unicode Consortium
What does The Unicode Consortium do?
The Unicode Consortium is a non-profit standards body that develops, maintains, and promotes the Unicode Standard, a universal character encoding system covering 159,801 characters across hundreds of scripts and languages. It distributes its standards, specifications, code charts, the Unicode Character Database (UCD), Common Locale Data Repository (CLDR), and International Components for Unicode (ICU) software libraries freely to operating systems, browsers, applications, and devices worldwide.
Is The Unicode Consortium a public or private company?
The Unicode Consortium is a private company. It is classified as nonprofit foundation owned and is currently operating.
When was The Unicode Consortium founded?
The Unicode Consortium was founded in 1988. It employs 1 to 10 people.
Where is The Unicode Consortium based?
The Unicode Consortium is headquartered in Mountain View, United States, in the North America region.
How does The Unicode Consortium make money?
Two revenue lines are on record. Organizational Membership Fees are the primary driver. The others are individual Membership and Donations.
Who are The Unicode Consortium's main competitors?
Direct peers on record are World Wide Web Consortium (W3C), Internet Engineering Task Force (IETF), Ecma International, WHATWG and OASIS (Organization for the Advancement of Structured Information Standards). Broad incumbents are ISO (International Organization for Standardization) and IEEE Standards Association. Others are The Internet Society (ISOC) and Linux Foundation. Internet Corporation for Assigned Names and Numbers (ICANN) is listed as an emerging player.
Does The Unicode Consortium have an API?
No public API is recorded for The Unicode Consortium.
What industry is The Unicode Consortium in?
The Unicode Consortium's product category is Character Encoding & Internationalization Standards. Its primary akta.pro industry code is BPADANAA, International Standards Development & Standardization Bodies. Its NAICS code is 813 and its SIC code is 8600.