Data Scientist, Inference Capacity Optimization

OpenAI Inc.
San Francisco, CA, United States
10 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$293,000.0 - $325,000.0
Working hours
Regular working hours

Tech stack

Artificial Intelligence Systems Engineering Big Data Distributed Systems Python (Programming Language) Machine Learning SQL Databases AI Infrastructure Reinforcement Learning Information Technology Data Analytics

Job description

OpenAI’s Industrial Compute organization is responsible for ensuring our compute infrastructure scales efficiently to support millions of users and increasingly sophisticated AI models., We’re looking for a Data Scientist to partner closely with Capacity Systems Engineering, Infrastructure, Product, and Research to optimize inference capacity across our global GPU fleet. This role combines statistical modeling, large-scale data analysis, forecasting, and systems thinking to drive critical decisions around infrastructure investments, performance-efficiency trade-offs, and customer experience., * Build statistical and machine learning models to profile and improve GPU utilization, latency, throughput, and overall fleet efficiency.

  • Develop forecasting models for inference demand across products, regions, and model families.
  • Analyze production workloads to identify latency bottlenecks and capacity constraints, highlighting optimization opportunities.
  • Partner with Capacity Systems Engineering to inform infrastructure planning and long-term GPU investment strategies.
  • Design experiments and simulations to evaluate scheduling policies, serving strategies, and infrastructure tradeoffs.
  • Build dashboards and operational metrics that enable leadership to make data-driven capacity decisions.
  • Collaborate with Product, Research, Finance, and Infrastructure teams to align compute planning with business growth and model roadmaps.
  • Communicate technical findings clearly to both engineering teams and executive leadership.

Requirements

  • MS or PhD in Statistics, Computer Science, Operations Research, Applied Mathematics, Economics, or related quantitative discipline (or equivalent industry experience).
  • 5+ years of experience working in the infrastructure data science space.
  • Strong expertise in Python and SQL.
  • Experience building forecasting, optimization, or predictive models.
  • Strong understanding of experimentation, statistical inference, and causal analysis.
  • Experience communicating analytical insights to executive stakeholders.

Preferred Skills

  • Capacity planning
  • Distributed systems
  • AI infrastructure
  • Datacenter design and buildout
  • Queueing theory
  • Time-series forecasting
  • Operations research
  • Supply-demand modeling
  • Reinforcement learning for resource allocation
  • Cost optimization

Benefits & conditions

$293K - $325K medical insurance, dental insurance, vision insurance, parental leave, paid time off, paid holidays, 401(k), retirement plan United States, California, San Francisco Aug 01, 2026

About the company

You’ll transform complex operational data into actionable insights that directly influence how OpenAI allocates and scales one of the world’s largest AI compute environments., OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity., At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on diversityjobs.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

1:14 min

Evolution of distributed SQL database architectures

Wei Hu Wei Hu · WWC 2024

1:32 min

Structuring platforms for new services and data analytics

Nevelina Aleksandrova · LIVE

2:54 min

Optimizing infrastructure for agentic flows and inference

Michael Kagan Michael Kagan +1 · WWC Europe 2026

2:10 min

Why organizations combine big data and machine learning

Ayon Roy · LIVE

1:10 min

Introduction to Microsoft Fabric and data agents

Dr. Alexander Wachtel Dr. Alexander Wachtel +1 · WWC 2025

Videos

See all

Related articles

See all