Developer docs
API playgroundTry for free, no card

Search company profiles

Stanford Data Science

Full company profile

uuid0048kt3

Namestring
Stanford Data Science
Legal namestring
Stanford University
Company typeenum
Private
Founded yearint
2022
Descriptiontext

Stanford Data Science (SDS) is a university-wide academic research institute within Stanford University, founded in 2018 and formally launched with its inaugural campus-wide conference on April 5, 2022. SDS operates as a horizontal institute that enables data-driven discovery across all Stanford schools (Humanities & Sciences, Engineering, Medicine, Business, Law, and Sustainability) through three primary lines of activity: (1) early-career talent programs including the Data Science Scholars program (16 new scholars in 2024–2025) and the Postdoctoral Fellows program (4 fellows per year); (2) faculty-led research centers covering causal science, open and reproducible science (CORES), sustainability data science, health data science, the Center for Decoding the Universe (C4DU, jointly with KIPAC), computational market design, and neural data science; and (3) community and convening programs including the annual Stanford Data Science Conference, Women in Data-Driven Discovery (WiD3), Rising Stars in Data Science, the Sustainability Data Science (SuDS) Conference, and the Data Science for Social Good (DSSG) summer fellowship.

The institute's core technical asset is Marlowe, Stanford's GPU-based computational instrument — an NVIDIA DGX H100 SuperPod comprising 31 H100 nodes (248 H100 80GB GPUs total), 2.5 PB of DDN Lustre high-performance storage, 2 TB RAM per node, 30 TB node-local NVMe, 900 GB/s GPU-to-GPU bandwidth via NVSwitch, and 3.2 Tbps InfiniBand connectivity. Marlowe is offered free to approved Stanford investigators with introductory GPU-hour awards, and is supported by a dedicated research data science team. As of May 2026, Stanford merged its AI and data science efforts under a single institute retaining the Stanford HAI name, with James Landay as head and Fei-Fei Li as Special Advisor on AI.

SDS is a nonprofit academic unit within Stanford University, funded through philanthropic giving (individuals, foundations, and corporations) and a Corporate Affiliates Program. It generates no commercial revenue; all programs, compute access, and events are provided free of charge to Stanford affiliates, and external programs such as DSSG, DSURP, and Rising Stars are fellowship-supported. Leadership includes Executive Director Chris Mentzel, Faculty Director Emmanuel Candès, Director Guido W. Imbens (2021 Nobel laureate in Economic Sciences), and Associate Director Chiara Sabatti.

Short descriptiontext

Stanford Data Science is a nonprofit academic research institute within Stanford University that advances data-driven discovery through PhD and postdoctoral fellowship programs, seven faculty-led research centers, and the Marlowe NVIDIA DGX H100 GPU SuperPod, serving Stanford faculty, students, and external collaborators.

Operating statusenum
Operating
Ownership categoryenum
Headcount rangeband
1–10
akta.pro rankint
HeadquartersStanford, United States
HQ citystring
Stanford
HQ countrystring
United States
HQ regionstring
North America
Markets served

Serves global market

Offices1 record

Each record includes

City, Country, Type, Description, Source

Keyword5 values
data science research, academic research institute, GPU computing cluster, postdoctoral fellowship programs, interdisciplinary AI research
Industry2 codes
1Technical Universities with Computing, Data & AI Focus
CodeEDAHAEAKPrimaryYes
2Data Science & Analytics (Data Literacy, SQL, Visualization)
CodeEDAMACAFPrimaryNo
NAICS code2 codes
  • Other Scientific and Technical Consulting Services54169
  • Computer Systems Design and Related Services54151
SIC code2 codes
  • Services-Computer Programming, Data Processing, Etc.7370
  • Services-Commercial Physical & Biological Research8731
Product category
Higher Education Research Institute
GTM motion3 records

Each record includes

Type, Description, Source

Revenue model2 records
1Philanthropic Giving
TypeProfessional Services
Description

Stanford Data Science relies on philanthropic support from individuals, foundations, and corporations. Endowed gifts provide a lasting foundation for SDS and faculty hired in partnership with Stanford's schools and institutes. Annual, unrestricted gifts support core activities including the Data Science Scholars program, SDS Postdoctoral Fellows, faculty-led research centers, and curriculum development. Gifts are tax-deductible under Stanford University's 501(c)(3) status.

datascience.stanford.edu
2Industry Affiliates Program
TypeLicensing Royalties
Description

Corporate partners join the SDS Corporate Affiliates Program, providing financial and strategic support in exchange for engagement with Stanford's data science research community, talent pipeline, and collaborative opportunities.

datascience.stanford.edu
Marketing channels9 records

Each record includes

Title, Type, Stage, Description, Source

Distribution channels5 records

Each record includes

Title, Type, Scope, Target buyer, Description, Source

Cost components5 values
Personnel, Technology or R&D, Infrastructure, Operations, Marketing or Sales
Pricing details1 tier
1Free access for Stanford affiliates (students, postdocs, faculty)
ModelFreemiumBilling cadenceAnnual
Notes

All SDS programs, events, and seminars are free for Stanford University affiliates. Marlowe offers 5,000 free introductory GPU hours for PI groups new to the platform.

datascience.stanford.edu
GTM typeB2B
B2B
Offering typeServices
Services
Brand1 of 7 records shown
1Marlowe
Description

Stanford's GPU-based computational instrument, a high-performance compute cluster for creating, analyzing, and using large-scale models.

datascience.stanford.edu
+6 more records
Core offering1 text field

Stanford Data Science operates an interdisciplinary research and education institute that funds PhD scholarships, postdoctoral fellowships, faculty-led research centers, and a GPU supercomputing platform (Marlowe) for Stanford researchers. It convenes the data science community through conferences (Annual Conference, Sustainability Data Science, Women in Data-Driven Discovery, Rising Stars), Distinguished Lectures, and seminars, and runs social-impact summer programs (Data Science for Social Good) that pair students with external project partners.

Differentiator
Functional benefit
Problem solved
Quantifiable outcome1 of 5 values shown
  • 16 new Data Science Scholars awarded in 2024-2025 cohort from five Stanford schools
+4 more records
Product overview1 text field

Stanford Data Science is a university-wide institute at Stanford University that enables data-driven discovery and expands data science education. The institute operates as a portfolio of interconnected programs and research centers, including the Marlowe GPU-based computational cluster for high-performance AI research, PhD Data Science Scholars and Postdoctoral Fellows programs for training the next generation of researchers, the Data Science for Social Good summer fellowship program, and multiple faculty-led research centers covering Open and Reproducible Science (CORES), Causal Science (SC2), Health Data Science, Sustainability Data Science, and the Center for Decoding the Universe. Additional programs include Women in Data Science (WiDS) conferences and the Rising Stars in Data Science workshop.

Product and service12 records
1Marlowe GPU-Based Computational Instrument
CategoryHigh-performance GPU compute infrastructure
Description

Stanford's high-performance compute cluster (NVIDIA DGX H100 Superpod with 31 H100 nodes, 248 H100 GPUs, 2.5 PB DDN Lustre storage, 900 GB/s GPU-to-GPU bandwidth via NVSwitch, 3.2Tbps InfiniBand) available to Stanford investigators for creating, analyzing, and using large-scale AI models.

2Data Science Scholars Program
CategoryPhD fellowship program
Description

PhD student fellowship program supporting exceptional graduate students advancing data science methods in their research; awarded 50% compensation over two years and drawn from five Stanford schools.

3Data Science Postdoctoral Fellows Program
CategoryPostdoctoral fellowship program
Description

Postdoctoral fellowship program hiring recent PhDs of exceptional promise for interdisciplinary research combining data science with domains like physical, life, social sciences, humanities, medicine, and engineering.

4Data Science for Social Good (DSSG) Summer Program
CategorySummer fellowship program
Description

Summer fellowship program inviting undergraduate and graduate students to work on data science projects with social impact, mentored by faculty and advanced researchers; pairs fellows with external project partners (e.g., Code for Africa, Stanford RegLab, Human Trafficking Data Lab).

5Women in Data Science (WiDS) Program
CategoryCommunity program and conference
Description

Program celebrating and accelerating careers of women in data science, including the annual Women in Data-Driven Discovery (WiD3) conference and workshops.

6Rising Stars in Data Science Workshop
CategoryCareer development workshop
Description

Workshop supporting transition to roles such as postdoctoral scholar, research scientist, industry researcher, or tenure-track faculty for graduate students near PhD completion; co-hosted with UC San Diego and University of Chicago.

7SDS Corporate Affiliates Program
CategoryIndustry affiliate membership program
Description

Industry engagement program where corporate partners provide financial and strategic support in exchange for access to Stanford's data science research community, talent pipeline, and collaborative opportunities.

8CORES (Center for Open and REproducible Science)
CategoryFaculty-led research center
Description

Multi-school center developing resources and supporting activities that promote open science practices and methodological innovations for transparent and reproducible research, including the Open Science Champion and Innovator awards.

9Stanford Causal Science Center (SC2)
CategoryFaculty-led research center
Description

Faculty-led research center focusing on causal inference methodologies and their applications across various domains of scholarship; led by Nobel laureate Guido Imbens.

10Center for Decoding the Universe (C4DU)
CategoryFaculty-led research center
Description

Research center focused on AI and data science applications in astrophysics and fundamental physics research; joint initiative of Stanford Data Science and KIPAC.

11Sustainability Data Science Center
CategoryFaculty-led research center
Description

Research center focused on applying data science to sustainability and climate change challenges; co-organizer of the annual Sustainability Data Science (SuDS) Conference.

12Health Data Science Center
CategoryFaculty-led research center
Description

Research center applying data science methods to health and medical research challenges.

Scale indicator8 records

Each record includes

Type, Value, Description, Source

Partnership11 partners
Strategic tierCoreTypeStrategic or Co-development PartnerAnnounced on2026-01-01
Description

As of May 2026, Stanford merged its AI and data science efforts under a single institute retaining the Stanford HAI name. James Landay serves as head of the combined institute. Co-founder Fei-Fei Li serves as Special Advisor on AI and co-chair of the advisory council alongside John Hennessy. The merged institute encompasses both HAI and Stanford Data Science.

Strategic tierCoreTypeTechnology or IntegrationAnnounced on2025-01-01
Description

Marlowe is built on the NVIDIA DGX H100 Superpod reference architecture. NVIDIA provides the hardware platform (31 H100 nodes, 248 GPUs) and a joint initiative connecting Stanford researchers with NVIDIA solutions architects and research teams for early-stage software support, optimization guidance, and access to advanced NVIDIA software libraries. Zoe Ryan (Solutions Architect) and Bruce McGowan (Senior Account Manager) at NVIDIA support the collaboration.

Strategic tierCoreTypeStrategic or Co-development PartnerAnnounced on2025-01-01
Description

The Stanford Doerr School of Sustainability co-organizes the annual Sustainability Data Science (SuDS) Conference with Stanford Data Science. The 2025 conference was held at the new Computing and Data Science (CoDa) building. The partnership bridges data science methods with sustainability science and climate change research.

Strategic tierCoreTypeStrategic or Co-development PartnerAnnounced on2020-05-14
Description

The R Consortium is a co-sponsor of the COVID-19 Data Forum, a multidisciplinary online meeting series organized by Stanford Data Science to discuss data-related aspects of the pandemic. The forum brought together topic experts to focus on data access, sharing, essential data resources for modeling, and decision-making support. The R Consortium Board of Directors chair Joseph Rickert serves on the Organizing Committee.

Strategic tierCoreTypeStrategic or Co-development Partner
Description

The Center for Decoding the Universe (C4DU) is a joint initiative of Stanford Data Science and KIPAC. KIPAC scientists collaborate with SDS researchers on fundamental AI and data science applied to astrophysics problems. The center is actively recruiting a Research Scientist with a background in AI/astrophysics.

6Stanford Wu Tsai Neurosciences Institute
Strategic tierCoreTypeStrategic or Co-development Partner
Description

The Wu Tsai Neurosciences Institute collaborated with Stanford Data Science and the Statistics Department on a new faculty position in Data Science and Neuroscience, seeking candidates for theoretical and computational neuroscience at the tenure-track level.

login.stanford.edu
Strategic tierMinorTypeStrategic or Co-development Partner
Description

RegLab partnered with Stanford Data Science through the Data Science for Social Good program to use machine learning, AI, and causal inference to modernize government regulation of CAFO (concentrated animal feeding operation) wastewater polluters.

Strategic tierMinorTypeStrategic or Co-development Partner
Description

Code for Africa partnered with Stanford Data Science through the 2020 Data Science for Social Good program to build a database of networks of corporations and persons involved in land transfer and ownership in Kenya, supporting journalists fighting corruption and promoting good governance.

9Stanford Human Trafficking Data Lab
Strategic tierMinorTypeStrategic or Co-development Partner
Description

The Human Trafficking Data Lab at Stanford partnered with SDS through the Data Science for Social Good program to help the Brazilian Federal Labor Prosecution Office target firms involved in human trafficking, using the Intuition Engine — an ensemble predictive model combining regression, NLP, deep learning, and network analysis.

datascience.stanford.edu
10COVID-19 Host Genetics Initiative / Harvard Medical School
Strategic tierMinorTypeOthers
Description

Andrea Ganna from the COVID-19 Host Genetics Initiative and Harvard Medical School participated as a speaker in the COVID-19 Data Forum webinar on August 13, 2020, focused on making COVID-19 clinical data available and useful.

login.stanford.edu
11EndPandemic National Data Consortium / Saama Technologies
Strategic tierMinorTypeOthers
Description

Ken Massey from EndPandemic National Data Consortium and Saama Technologies participated as a speaker in the COVID-19 Data Forum webinar on clinical data, discussing efforts to make COVID-19 clinical data available and useful.

login.stanford.edu
Recent move7 records

Each record includes

Date, Type, Title, Description, Source

Expansion highlight5 records

Each record includes

Type, Description

Peers10 records
1Oxford Department of Statistics
TypeRegional player
Description

Oxford's longstanding statistics and data science hub, with parallel emphasis on causal inference, methodological research, and cross-disciplinary collaboration. Comparable in academic rigor but operating in a different geographic market.

TypeEmerging player
Description

A non-university research institute focused on fundamental AI research with a strong open-science ethos. Comparable to SDS in its research-only, mission-driven posture and CORES-aligned open science philosophy.

TypeDirect peer
Description

Berkeley's campus-wide data science and computing unit, founded around the same time, with parallel mission of interdisciplinary data science education and research. Most direct structural peer to Stanford Data Science.

TypeBroad incumbent
Description

CMU's broad computer science school with deep expertise in AI, machine learning, and data science at scale. A larger, more established incumbent providing overlapping capabilities across research, education, and industry partnerships.

TypeDirect peer
Description

Harvard's university-wide data science initiative, focused on methodological research and cross-school collaboration. Operates with similar postdoctoral and faculty programs and a comparable philanthropy-driven funding model.

TypeDirect peer
Description

Berkeley's data science research institute with a similar emphasis on open science, reproducibility, and cross-disciplinary methods. Comparable to Stanford's CORES in scope and mission.

TypeBroad incumbent
Description

Stanford HAI, the household-name AI institute at Stanford that merged with SDS in May 2026. Operating as a broader AI research and policy organization, it is the dominant institutional AI brand at Stanford and SDS's parent structure going forward.

TypeDirect peer
Description

One of the earliest university-wide data science institutes (founded 2013), with an explicit PhD training program in data science. Comparable to Stanford Data Science in its emphasis on interdisciplinary methods and external student pathways.

TypeDirect peer
Description

MIT's flagship institute for computing and AI cross-disciplinary research and education, with similar faculty hiring, fellowship, and research-center model. Directly comparable in scale and ambition to Stanford Data Science.

TypeDirect peer
Description

UChicago's data science institute and co-host of the Rising Stars in Data Science workshop with Stanford. Direct peer in mission, talent programs, and career-development workshops for early-career researchers.

Market position
Strengths4 records

Each record includes

Headline, Details, Source

Weaknesses4 records

Each record includes

Headline, Details, Source

Competitive moat5 records

Each record includes

Type, Details

Key risks5 records

Each record includes

Headline, Details, Source

Key highlights6 records

Each record includes

Headline, Details, Source

Customer concentration

Classification, Details

Named customers7 records

Each record includes

Name, Industry, Type, Use case, Source, UUID

Segment4 records

Each record includes

Title, Type, Primary, Description, Pain point addressed, Use case, Source

Ideal customer profile4 records

Each record includes

Profile, Firmographic size, Sales motion, Sales cycle length, Buying structure, Purchase trigger, Buyer persona, Geography, Industry vertical, Primary use case, Description, Pain points, Evidence proof points, Target buyer

Technology focused
No
API detail
Has APIbool
No

Docs URL, Description

AI capability6 records

Each record includes

Type, Description, Source

AI maturity
App detail

Has app

Feature3 records

Each record includes

Title, Differentiator, Description, Source

Core technology
Revenue estimate
Valuation estimate
Number of profiles
Profiles3 records

Each record includes

Name, Designation, Designation category, Overview, Profile commentary, Source

No data
No data
Funding overview

Funding stage, Last funding date, Total funding USD

Funding rounds

Each record includes

Round, Amount USD, Date, Pre money valuation, Total investors, Investors, News

Investors

Each record includes

Name, Type, Date of entry, Rounds participated, Website

Funding detail is available on the Subscription and Enterprise plan.Contact sales →

M&A

Each record includes

Name, Acquisition type, Announced date, Completed date, Status, Website, News

Investment

Each record includes

Name, Round, Announced date, Lead investor, Website, News

M&A and investment is available on the Subscription and Enterprise plan.Contact sales →

Stanford Data Science

Higher Education Research Institutedatascience.stanford.edu

Stanford Data Science is a nonprofit academic research institute within Stanford University that advances data-driven discovery through PhD and postdoctoral fellowship programs, seven faculty-led research centers, and the Marlowe NVIDIA DGX H100 GPU SuperPod, serving Stanford faculty, students, and external collaborators.

What Stanford Data Science does

Stanford Data Science (SDS) is a university-wide academic research institute within Stanford University, founded in 2018 and formally launched with its inaugural campus-wide conference on April 5, 2022. SDS operates as a horizontal institute that enables data-driven discovery across all Stanford schools (Humanities & Sciences, Engineering, Medicine, Business, Law, and Sustainability) through three primary lines of activity: (1) early-career talent programs including the Data Science Scholars program (16 new scholars in 2024–2025) and the Postdoctoral Fellows program (4 fellows per year); (2) faculty-led research centers covering causal science, open and reproducible science (CORES), sustainability data science, health data science, the Center for Decoding the Universe (C4DU, jointly with KIPAC), computational market design, and neural data science; and (3) community and convening programs including the annual Stanford Data Science Conference, Women in Data-Driven Discovery (WiD3), Rising Stars in Data Science, the Sustainability Data Science (SuDS) Conference, and the Data Science for Social Good (DSSG) summer fellowship.

The institute's core technical asset is Marlowe, Stanford's GPU-based computational instrument — an NVIDIA DGX H100 SuperPod comprising 31 H100 nodes (248 H100 80GB GPUs total), 2.5 PB of DDN Lustre high-performance storage, 2 TB RAM per node, 30 TB node-local NVMe, 900 GB/s GPU-to-GPU bandwidth via NVSwitch, and 3.2 Tbps InfiniBand connectivity. Marlowe is offered free to approved Stanford investigators with introductory GPU-hour awards, and is supported by a dedicated research data science team. As of May 2026, Stanford merged its AI and data science efforts under a single institute retaining the Stanford HAI name, with James Landay as head and Fei-Fei Li as Special Advisor on AI.

SDS is a nonprofit academic unit within Stanford University, funded through philanthropic giving (individuals, foundations, and corporations) and a Corporate Affiliates Program. It generates no commercial revenue; all programs, compute access, and events are provided free of charge to Stanford affiliates, and external programs such as DSSG, DSURP, and Rising Stars are fellowship-supported. Leadership includes Executive Director Chris Mentzel, Faculty Director Emmanuel Candès, Director Guido W. Imbens (2021 Nobel laureate in Economic Sciences), and Associate Director Chiara Sabatti.

Stanford Data Science firmographics

Firmographics
Name
Stanford Data Science
Legal name
Stanford University
Website
https://datascience.stanford.edu
Company type
Private
Founded year
2022
Operating status
Operating
Headcount range
1–10 employees
Short description
Stanford Data Science is a nonprofit academic research institute within Stanford University that advances data-driven discovery through PhD and postdoctoral fellowship programs, seven faculty-led research centers, and the Marlowe NVIDIA DGX H100 GPU SuperPod, serving Stanford faculty, students, and external collaborators.
Ownership category
akta.pro rank

Stanford Data Science industry classification

Industry
Product category
Higher Education Research Institute
NAICS
Other Scientific and Technical Consulting Services (54169), Computer Systems Design and Related Services (54151)
SIC
Services-Computer Programming, Data Processing, Etc. (7370), Services-Commercial Physical & Biological Research (8731)
akta.pro primary industry
Technical Universities with Computing, Data & AI Focus (EDAHAEAK)
akta.pro secondary industry
Data Science & Analytics (Data Literacy, SQL, Visualization) (EDAMACAF)

Keywords

  • Data science research
  • Academic research institute
  • GPU computing cluster
  • Postdoctoral fellowship programs
  • Interdisciplinary AI research

Where Stanford Data Science is headquartered

Location

Headquarters

HQ city
Stanford
HQ country
United States
HQ region
North America

Offices1 record

Markets served

Stanford Data Science business model

Business model
GTM type
B2B
Offering type
Services
Cost components
Personnel, Technology or R&D, Infrastructure, Operations, Marketing or Sales

Revenue model

  1. Philanthropic Giving: Stanford Data Science relies on philanthropic support from individuals, foundations, and corporations. Endowed gifts provide a lasting foundation for SDS and faculty hired in partnership with Stanford's schools and institutes. Annual, unrestricted gifts support core activities including the Data Science Scholars program, SDS Postdoctoral Fellows, faculty-led research centers, and curriculum development. Gifts are tax-deductible under Stanford University's 501(c)(3) status.
  2. Industry Affiliates Program: Corporate partners join the SDS Corporate Affiliates Program, providing financial and strategic support in exchange for engagement with Stanford's data science research community, talent pipeline, and collaborative opportunities.

Pricing tiers

ModelBillingPrice
FreemiumAnnualFree access for Stanford affiliates (students, postdocs, faculty)

Go-to-market motion3 records

Distribution channels5 records

Marketing channels9 records

Stanford Data Science product offering

Product offering

Core offering

Stanford Data Science operates an interdisciplinary research and education institute that funds PhD scholarships, postdoctoral fellowships, faculty-led research centers, and a GPU supercomputing platform (Marlowe) for Stanford researchers. It convenes the data science community through conferences (Annual Conference, Sustainability Data Science, Women in Data-Driven Discovery, Rising Stars), Distinguished Lectures, and seminars, and runs social-impact summer programs (Data Science for Social Good) that pair students with external project partners.

Product overview

Stanford Data Science is a university-wide institute at Stanford University that enables data-driven discovery and expands data science education. The institute operates as a portfolio of interconnected programs and research centers, including the Marlowe GPU-based computational cluster for high-performance AI research, PhD Data Science Scholars and Postdoctoral Fellows programs for training the next generation of researchers, the Data Science for Social Good summer fellowship program, and multiple faculty-led research centers covering Open and Reproducible Science (CORES), Causal Science (SC2), Health Data Science, Sustainability Data Science, and the Center for Decoding the Universe. Additional programs include Women in Data Science (WiDS) conferences and the Rising Stars in Data Science workshop.

Differentiator

Problem solved

Functional benefit

Brands

  • Marlowe: Stanford's GPU-based computational instrument, a high-performance compute cluster for creating, analyzing, and using large-scale models.
  • Data Science Scholars Program
  • Data Science Fellows Program
  • Data Science for Social Good
  • Rising Stars in Data Science
  • Women in Data-Driven Discovery (WiD3)
  • Ram and Vijay Shriram Data Science Fellows

Products and services

  • Marlowe GPU-Based Computational Instrument Stanford's high-performance compute cluster (NVIDIA DGX H100 Superpod with 31 H100 nodes, 248 H100 GPUs, 2.5 PB DDN Lustre storage, 900 GB/s GPU-to-GPU bandwidth via NVSwitch, 3.2Tbps InfiniBand) available to Stanford investigators for creating, analyzing, and using large-scale AI models.
  • Data Science Scholars Program PhD student fellowship program supporting exceptional graduate students advancing data science methods in their research; awarded 50% compensation over two years and drawn from five Stanford schools.
  • Data Science Postdoctoral Fellows Program Postdoctoral fellowship program hiring recent PhDs of exceptional promise for interdisciplinary research combining data science with domains like physical, life, social sciences, humanities, medicine, and engineering.
  • Data Science for Social Good (DSSG) Summer Program Summer fellowship program inviting undergraduate and graduate students to work on data science projects with social impact, mentored by faculty and advanced researchers; pairs fellows with external project partners (e.g., Code for Africa, Stanford RegLab, Human Trafficking Data Lab).
  • Women in Data Science (WiDS) Program Program celebrating and accelerating careers of women in data science, including the annual Women in Data-Driven Discovery (WiD3) conference and workshops.
  • Rising Stars in Data Science Workshop Workshop supporting transition to roles such as postdoctoral scholar, research scientist, industry researcher, or tenure-track faculty for graduate students near PhD completion; co-hosted with UC San Diego and University of Chicago.
  • SDS Corporate Affiliates Program Industry engagement program where corporate partners provide financial and strategic support in exchange for access to Stanford's data science research community, talent pipeline, and collaborative opportunities.
  • CORES (Center for Open and REproducible Science) Multi-school center developing resources and supporting activities that promote open science practices and methodological innovations for transparent and reproducible research, including the Open Science Champion and Innovator awards.
  • Stanford Causal Science Center (SC2) Faculty-led research center focusing on causal inference methodologies and their applications across various domains of scholarship; led by Nobel laureate Guido Imbens.
  • Center for Decoding the Universe (C4DU) Research center focused on AI and data science applications in astrophysics and fundamental physics research; joint initiative of Stanford Data Science and KIPAC.
  • Sustainability Data Science Center Research center focused on applying data science to sustainability and climate change challenges; co-organizer of the annual Sustainability Data Science (SuDS) Conference.
  • Health Data Science Center Research center applying data science methods to health and medical research challenges.

Quantifiable outcome

  • 16 new Data Science Scholars awarded in 2024-2025 cohort from five Stanford schools
  • +4 more outcomes

Companies that use Stanford Data Science

Customer profile

Named customers7 records

Segments4 records

Ideal customer profiles4 records

Stanford Data Science technology and API

Technology

Technology focussed No

API detail

Has API
No
API docs
API detail

Core technology

AI maturity

App detail

AI capability6 records

Feature3 records

Stanford Data Science partnerships and signals

Strategic signal

Partnerships

Eleven partnerships are on record, tiered core and minor.

  • Stanford Institute for Human-Centered AI (HAI)coreStrategic or Co-development Partner · 1 January 2026As of May 2026, Stanford merged its AI and data science efforts under a single institute retaining the Stanford HAI name. James Landay serves as head of the combined institute. Co-founder Fei-Fei Li serves as Special Advisor on AI and co-chair of the advisory council alongside John Hennessy. The merged institute encompasses both HAI and Stanford Data Science.
  • NVIDIAcoreTechnology or Integration · 1 January 2025Marlowe is built on the NVIDIA DGX H100 Superpod reference architecture. NVIDIA provides the hardware platform (31 H100 nodes, 248 GPUs) and a joint initiative connecting Stanford researchers with NVIDIA solutions architects and research teams for early-stage software support, optimization guidance, and access to advanced NVIDIA software libraries. Zoe Ryan (Solutions Architect) and Bruce McGowan (Senior Account Manager) at NVIDIA support the collaboration.
  • Stanford Doerr School of SustainabilitycoreStrategic or Co-development Partner · 1 January 2025The Stanford Doerr School of Sustainability co-organizes the annual Sustainability Data Science (SuDS) Conference with Stanford Data Science. The 2025 conference was held at the new Computing and Data Science (CoDa) building. The partnership bridges data science methods with sustainability science and climate change research.
  • R ConsortiumcoreStrategic or Co-development Partner · 14 May 2020The R Consortium is a co-sponsor of the COVID-19 Data Forum, a multidisciplinary online meeting series organized by Stanford Data Science to discuss data-related aspects of the pandemic. The forum brought together topic experts to focus on data access, sharing, essential data resources for modeling, and decision-making support. The R Consortium Board of Directors chair Joseph Rickert serves on the Organizing Committee.
  • KIPAC (Kavli Institute for Particle Astrophysics and Cosmology)coreStrategic or Co-development PartnerThe Center for Decoding the Universe (C4DU) is a joint initiative of Stanford Data Science and KIPAC. KIPAC scientists collaborate with SDS researchers on fundamental AI and data science applied to astrophysics problems. The center is actively recruiting a Research Scientist with a background in AI/astrophysics.
  • Stanford Wu Tsai Neurosciences InstitutecoreStrategic or Co-development PartnerThe Wu Tsai Neurosciences Institute collaborated with Stanford Data Science and the Statistics Department on a new faculty position in Data Science and Neuroscience, seeking candidates for theoretical and computational neuroscience at the tenure-track level.
  • Stanford Law School RegLabminorStrategic or Co-development PartnerRegLab partnered with Stanford Data Science through the Data Science for Social Good program to use machine learning, AI, and causal inference to modernize government regulation of CAFO (concentrated animal feeding operation) wastewater polluters.
  • Code for AfricaminorStrategic or Co-development PartnerCode for Africa partnered with Stanford Data Science through the 2020 Data Science for Social Good program to build a database of networks of corporations and persons involved in land transfer and ownership in Kenya, supporting journalists fighting corruption and promoting good governance.
  • Stanford Human Trafficking Data LabminorStrategic or Co-development PartnerThe Human Trafficking Data Lab at Stanford partnered with SDS through the Data Science for Social Good program to help the Brazilian Federal Labor Prosecution Office target firms involved in human trafficking, using the Intuition Engine — an ensemble predictive model combining regression, NLP, deep learning, and network analysis.
  • COVID-19 Host Genetics Initiative / Harvard Medical SchoolminorOthersAndrea Ganna from the COVID-19 Host Genetics Initiative and Harvard Medical School participated as a speaker in the COVID-19 Data Forum webinar on August 13, 2020, focused on making COVID-19 clinical data available and useful.
  • EndPandemic National Data Consortium / Saama TechnologiesminorOthersKen Massey from EndPandemic National Data Consortium and Saama Technologies participated as a speaker in the COVID-19 Data Forum webinar on clinical data, discussing efforts to make COVID-19 clinical data available and useful.

Scale indicators8 records

Recent moves7 records

Expansion highlights5 records

Stanford Data Science competitors and assessment

Company assessment

Regional players

  • Oxford Department of Statistics: Oxford's longstanding statistics and data science hub, with parallel emphasis on causal inference, methodological research, and cross-disciplinary collaboration. Comparable in academic rigor but operating in a different geographic market.

Emerging players

  • Allen Institute for AI (AI2): A non-university research institute focused on fundamental AI research with a strong open-science ethos. Comparable to SDS in its research-only, mission-driven posture and CORES-aligned open science philosophy.

Direct peers

  • UC Berkeley Division of Computing, Data Science, and Society (CDSS): Berkeley's campus-wide data science and computing unit, founded around the same time, with parallel mission of interdisciplinary data science education and research. Most direct structural peer to Stanford Data Science.
  • Harvard Data Science Initiative: Harvard's university-wide data science initiative, focused on methodological research and cross-school collaboration. Operates with similar postdoctoral and faculty programs and a comparable philanthropy-driven funding model.
  • Berkeley Institute for Data Science (BIDS): Berkeley's data science research institute with a similar emphasis on open science, reproducibility, and cross-disciplinary methods. Comparable to Stanford's CORES in scope and mission.
  • NYU Center for Data Science: One of the earliest university-wide data science institutes (founded 2013), with an explicit PhD training program in data science. Comparable to Stanford Data Science in its emphasis on interdisciplinary methods and external student pathways.
  • MIT Schwarzman College of Computing: MIT's flagship institute for computing and AI cross-disciplinary research and education, with similar faculty hiring, fellowship, and research-center model. Directly comparable in scale and ambition to Stanford Data Science.
  • University of Chicago Data Science Institute: UChicago's data science institute and co-host of the Rising Stars in Data Science workshop with Stanford. Direct peer in mission, talent programs, and career-development workshops for early-career researchers.

Broad incumbents

  • Carnegie Mellon School of Computer Science: CMU's broad computer science school with deep expertise in AI, machine learning, and data science at scale. A larger, more established incumbent providing overlapping capabilities across research, education, and industry partnerships.
  • Stanford Institute for Human-Centered AI (HAI): Stanford HAI, the household-name AI institute at Stanford that merged with SDS in May 2026. Operating as a broader AI research and policy organization, it is the dominant institutional AI brand at Stanford and SDS's parent structure going forward.

Market position

Strengths4 records

Weaknesses4 records

Competitive moat5 records

Key risks5 records

Key highlights6 records

Customer concentration

Stanford Data Science social profiles

Digital presence

Stanford Data Science financial estimates

Financial estimate

Revenue estimate

Valuation estimate

Stanford Data Science leadership team

Management profile

Number of profiles

Profiles3 records

Stanford Data Science funding detail

Funding detail

Funding overview

Funding rounds

Investors

Funding detail is available on the Subscription and Enterprise plan.Contact sales →

Stanford Data Science M&A and investment

M&A and investment

M&A

Investments

M&A and investment is available on the Subscription and Enterprise plan.Contact sales →

Frequently asked questions about Stanford Data Science

What does Stanford Data Science do?

Stanford Data Science operates an interdisciplinary research and education institute that funds PhD scholarships, postdoctoral fellowships, faculty-led research centers, and a GPU supercomputing platform (Marlowe) for Stanford researchers. It convenes the data science community through conferences (Annual Conference, Sustainability Data Science, Women in Data-Driven Discovery, Rising Stars), Distinguished Lectures, and seminars, and runs social-impact summer programs (Data Science for Social Good) that pair students with external project partners.

Is Stanford Data Science a public or private company?

Stanford Data Science is a private company. It is classified as nonprofit foundation owned and is currently operating.

When was Stanford Data Science founded?

Stanford Data Science was founded in 2022. It employs 1 to 10 people.

Where is Stanford Data Science based?

Stanford Data Science is headquartered in Stanford, United States, in the North America region.

How does Stanford Data Science make money?

Two revenue lines are on record. Philanthropic Giving is the primary driver. The others are industry Affiliates Program.

Who are Stanford Data Science's main competitors?

Oxford Department of Statistics is listed as a regional player. Allen Institute for AI (AI2) is listed as an emerging player. Direct peers are UC Berkeley Division of Computing, Data Science, and Society (CDSS), Harvard Data Science Initiative, Berkeley Institute for Data Science (BIDS), NYU Center for Data Science, MIT Schwarzman College of Computing and University of Chicago Data Science Institute. Broad incumbents are Carnegie Mellon School of Computer Science and Stanford Institute for Human-Centered AI (HAI).

Does Stanford Data Science have an API?

No public API is recorded for Stanford Data Science.

What industry is Stanford Data Science in?

Stanford Data Science's product category is Higher Education Research Institute. Its primary akta.pro industry code is EDAHAEAK, Technical Universities with Computing, Data & AI Focus, with a secondary code of EDAMACAF, Data Science & Analytics (Data Literacy, SQL, Visualization). Its NAICS code is 54169 and its SIC code is 7370.

Unlock the full company data

50 free credits on sign-up, no credit card required.

Contact sales
Live signals
Stanford University School of EngineeringResponse to OSTP's Request for Information on Accelerating the American Scientific EnterpriseStanford HAI and Stanford Data Science scholars submitted a response to the White House Office of Science and Technology Policy's Request for Information on strengthening the American scientific enterprise, recommending a new "team science" academic research model for AI-enabled discovery. The scholars proposed two key policy measures: establishing an interdisciplinary consortium of academic, government, and industry experts, and requiring that federally funded AI research make software and datasets openly available. These recommendations aim to shape future federal science and AI funding policies.SacbeeGov. Gavin Newsom announces council to promote ‘responsible’ AIGovernor Gavin Newsom announced the creation of the California Innovation Council to advise on artificial intelligence policies within state government. The 30-member council includes executives and leaders from various organizations, including the Mozilla Foundation and Stanford University, aiming to promote responsible AI use. This initiative follows Newsom's push for AI solutions to enhance government efficiency, alongside the implementation of an AI-powered digital assistant called Poppy.AI.Snorkel AIChat With the Terminal-Bench TeamResearchers from Stanford University and Laude Institute discuss the origins and design philosophy of Terminal-Bench, an open-source benchmark for evaluating AI agents that control computers through terminal commands, as opposed to traditional GUI-based approaches. The team also introduces Harbor, a new execution framework that simplifies scaling container-based rollouts for agent evaluation and reinforcement learning across cloud platforms like Daytona, E2B, and Modal. Terminal-Bench 2.0 launches with 89 manually verified tasks that are harder and more reproducible than its predecessor, reflecting lessons learned from the open-source community-building process.