> Markdown version of [/jobs/ext/1890713-data-scientist-inference-capacity-optimization](https://www.wearedevelopers.com/jobs/ext/1890713-data-scientist-inference-capacity-optimization). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Scientist, Inference Capacity Optimization - **Company:** OpenAI Inc. - **Location:** San Francisco, CA, United States - **Experience:** Expert - **Salary:** $293,000.0 - $325,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Systems Engineering, Big Data, Distributed Systems, Python (Programming Language), Machine Learning, SQL Databases, AI Infrastructure, Reinforcement Learning, Information Technology, Data Analytics - **Published:** August 2, 2026 - **Apply:** https://diversityjobs.com/main/sendform/8/8/28176/1/17792560?backUrl=%2Fcareer%2F17792560%2FData-Scientist-Inference-Capacity-Optimization-California-San-Francisco ## About the Role * MS or PhD in Statistics, Computer Science, Operations Research, Applied Mathematics, Economics, or related quantitative discipline (or equivalent industry experience). * 5+ years of experience working in the infrastructure data science space. * Strong expertise in Python and SQL. * Experience building forecasting, optimization, or predictive models. * Strong understanding of experimentation, statistical inference, and causal analysis. * Experience communicating analytical insights to executive stakeholders. Preferred Skills * Capacity planning * Distributed systems * AI infrastructure * Datacenter design and buildout * Queueing theory * Time-series forecasting * Operations research * Supply-demand modeling * Reinforcement learning for resource allocation * Cost optimization ## Description OpenAI's Industrial Compute organization is responsible for ensuring our compute infrastructure scales efficiently to support millions of users and increasingly sophisticated AI models., We're looking for a Data Scientist to partner closely with Capacity Systems Engineering, Infrastructure, Product, and Research to optimize inference capacity across our global GPU fleet. This role combines statistical modeling, large-scale data analysis, forecasting, and systems thinking to drive critical decisions around infrastructure investments, performance-efficiency trade-offs, and customer experience., * Build statistical and machine learning models to profile and improve GPU utilization, latency, throughput, and overall fleet efficiency. * Develop forecasting models for inference demand across products, regions, and model families. * Analyze production workloads to identify latency bottlenecks and capacity constraints, highlighting optimization opportunities. * Partner with Capacity Systems Engineering to inform infrastructure planning and long-term GPU investment strategies. * Design experiments and simulations to evaluate scheduling policies, serving strategies, and infrastructure tradeoffs. * Build dashboards and operational metrics that enable leadership to make data-driven capacity decisions. * Collaborate with Product, Research, Finance, and Infrastructure teams to align compute planning with business growth and model roadmaps. * Communicate technical findings clearly to both engineering teams and executive leadership. ## Related Videos - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [AI beyond the code: Master your organisational AI implementation.](https://www.wearedevelopers.com/videos/1248-ai-beyond-the-code-master-your-organisational-ai-implementation) - [Fault Tolerance and Consistency at Scale: Harnessing the Power of Distributed SQL Databases](https://www.wearedevelopers.com/videos/1146-fault-tolerance-and-consistency-at-scale-harnessing-the-power-of-distributed-sql-databases) - [Data Analytics with Microsoft Fabric: End-to-End Use Case with Data Agents](https://www.wearedevelopers.com/videos/1547-data-analytics-with-microsoft-fabric-end-to-end-use-case-with-data-agents) - [How Data is Shaping our Games](https://www.wearedevelopers.com/videos/176-how-data-is-shaping-our-games) ## Related Articles - [What Industries Outside of AI Are Hiring The Most AI Experts?](https://www.wearedevelopers.com/magazine/98-what-industries-outside-of-ai-are-hiring-the-most-ai-experts) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [Graph and AI Trends 2026: Why Is AI Running but Not Yet Delivering?](https://www.wearedevelopers.com/magazine/680-graph-and-ai-trends-2026-why-is-ai-running-but-not-yet-delivering)