Skip to content
OpenAIOpenAISan Francisco, CA

Data Scientist, Inference Capacity Optimization

Data Scientist optimizing inference capacity on OpenAI's global GPU fleet. Build statistical/ML forecasting models, analyze workloads for bottlenecks, design experiments for scheduling/serving tradeoffs, and partner with engineering to guide infrastructure investments and efficiency improvements. Requires MS/PhD and 5+ years in infrastructure data science.

293k – 325k/yr
Hybrid5+ YOEData Science

About the role

Key Responsibilities

  • Build statistical and machine learning models to profile and improve GPU utilization, latency, throughput, and overall fleet efficiency.
  • Develop forecasting models for inference demand across products, regions, and model families.
  • Analyze production workloads to identify latency bottlenecks and capacity constraints, highlighting optimization opportunities.
  • Partner with Capacity Systems Engineering to inform infrastructure planning and long-term GPU investment strategies.
  • Design experiments and simulations to evaluate scheduling policies, serving strategies, and infrastructure tradeoffs.
  • Build dashboards and operational metrics that enable leadership to make data-driven capacity decisions.
  • Collaborate with Product, Research, Finance, and Infrastructure teams to align compute planning with business growth and model roadmaps.
  • Communicate technical findings clearly to both engineering teams and executive leadership.

Qualifications

  • MS or PhD in Statistics, Computer Science, Operations Research, Applied Mathematics, Economics, or related quantitative discipline (or equivalent industry experience).
  • 5+ years of experience working in the infrastructure data science space.
  • Strong expertise in Python and SQL.
  • Experience building forecasting, optimization, or predictive models.
  • Strong understanding of experimentation, statistical inference, and causal analysis.
  • Experience communicating analytical insights to executive stakeholders.

Preferred Skills

  • Capacity planning
  • Distributed systems
  • AI infrastructure
  • Datacenter design and buildout
  • Queueing theory
  • Time-series forecasting
  • Operations research
  • Supply-demand modeling
  • Reinforcement learning for resource allocation
  • Cost optimization

Skills

PythonSQLForecastingMachine LearningStatistical Modelingoptimizationtime-series analysisoperations researchReinforcement Learningqueueing theorycausal analysisExperimentationDistributed SystemsCapacity Planning

Similar roles

Data Science jobs
OpenAI

Data Scientist, SMB Ads Growth

OpenAISan Francisco, CA

Lead Data Scientist for SMB Ads Growth at OpenAI, architecting end-to-end analytics for targeting, funnel optimization, campaign measurement, and revenue forecasting. Requires 5+ years in ads/growth analytics, ML model deployment (propensity/LTV), strong SQL/Python/R, and experience in 0-to-1 environments.

293k – 515k/yrOn-site5+ YOEData Science
OpenAI

Data Scientist, Core Experimentation

OpenAIBellevue, WA +1

Leads evolution of OpenAI's core experimentation platform, driving statistical strategy, designing methodologies, and building scalable Python/Spark pipelines to ensure reliable, trustworthy experiments at massive scale. Requires deep stats expertise, causal inference, and production experimentation experience.

293k – 325k/yrHybridData Science
OpenAI

Data Scientist, GTM Intelligence

OpenAISan Francisco, CA

Data Scientist building GTM intelligence systems at OpenAI. Own roadmap, feature datasets, decision models (heuristic/ML/ranking), production SQL/Python pipelines, monitoring, and stakeholder alignment to drive account prioritization, risk detection, interventions, and outcome measurement for customer-facing teams.

290k – 340k/yrHybrid5+ YOEData Science
Anthropic

People Research Scientist, Recruiting

AnthropicSan Francisco, CA +2

Serve as the research expert for Anthropic's Recruiting organization on the People Data Solutions team. Design and execute studies on recruiting funnels, interview validation, psychometrics, and candidate experience to drive evidence-based hiring decisions using SQL, Python/R, and advanced analytics.

285k – 380k/yrHybrid5+ YOEData Science
Anthropic

Data Scientist, GTM

AnthropicNew York, NY +1

Drive data-informed decisions across Anthropic's commercial customer lifecycle by owning GTM metrics, causal inference experiments, statistical modeling, and actionable insights for acquisition, activation, expansion, and retention of a consumption-based AI platform.

285k – 380k/yrHybrid5+ YOEData Science