> Markdown version of [/jobs/ext/2511653-research-scientist-infrastructure-modeling-and-reliability](https://www.wearedevelopers.com/jobs/ext/2511653-research-scientist-infrastructure-modeling-and-reliability). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Research Scientist, Infrastructure Modeling and Reliability - **Company:** Facebook Inc. - **Location:** Menlo Park, CA, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Data Analysis, C++ (Programming Language), Data Centers, Distributed Systems, R (Programming Language), Monitoring of Systems, Python (Programming Language), Machine Learning, Reliability Engineering - **Published:** August 11, 2026 - **Apply:** https://www.jobmonkeyjobs.com/career/27923230/Research-Scientist-Infrastructure-Modeling-Reliability-California-Menlo-Park-7418 ## About the Role Meta builds technologies that help people connect, find communities, and grow businesses. Meta's infrastructure supports services used by billions of people, and operating that infrastructure efficiently requires increasingly sophisticated modeling of demand, utilization, reliability, and physical resource constraints.We are seeking an industry-leading Research Scientist or Applied Scientist to define and build new modeling approaches for power utilization across Meta's infrastructure. This role will lead the development of statistical and machine learning models that monitor power consumption, project peak demand, quantify uncertainty, and inform how Meta maximizes usable power within failure domains while maintaining target reliability levels. The ideal candidate has deep experience modeling high-dimensional, noisy, and interdependent systems, and has demonstrated the ability to translate scientific advances into production systems that influence large-scale infrastructure strategy., * 10+ years of experience developing statistical, machine learning, simulation, forecasting, optimization, or other quantitative modeling systems * Experience leading ambiguous, cross-functional technical programs from problem definition through model development, evaluation, deployment, and business impact * Experience coding in Python, R, C++, Java, or similar languages for data analysis, modeling, simulation, or production systems * Experience communicating complex technical concepts, assumptions, uncertainty, and tradeoffs to technical and non-technical audiences * Experience influencing technical strategy across multiple teams or organizations, * Experience modeling high-dimensional, sparse, noisy, or strongly correlated data in production environments * Experience with time-series forecasting, probabilistic forecasting, Bayesian modeling, extreme-value modeling, causal inference, stochastic processes, simulation, or uncertainty quantification * Experience with infrastructure, capacity planning, power systems, energy systems, data centers, reliability engineering, distributed systems, supply-chain optimization, or resource allocation * Experience building models that support operational decisions under explicit reliability, safety, cost, or utilization constraints * Experience developing peak-demand forecasts, confidence intervals, risk estimates, anomaly detection, or backtesting frameworks * Experience applying optimization, operations research, or decision science to large-scale resource planning * Demonstrated record of industry-level technical leadership, such as defining new research directions, influencing company strategy, publishing in leading venues, or shaping external technical standards * Experience mentoring senior technical contributors and building scientific communities across organizations ## Description * Define the scientific and technical strategy for modeling power consumption, peak risk, and reliability tradeoffs across large-scale infrastructure systems. * Develop statistical, machine learning, and/or optimization models that forecast power demand, estimate peak distributions, quantify uncertainty, and support operational decision-making. * Build approaches that reason about high-dimensional signals, correlated demand, failure-domain constraints, reserve margins, and reliability targets. * Partner with engineering, capacity planning, data center, energy, hardware, operations, and finance teams to translate model outputs into infrastructure planning and utilization decisions. * Establish evaluation frameworks, backtesting methods, confidence intervals, and monitoring systems to measure model quality and operational risk. * Identify opportunities to safely increase power utilization, reduce stranded capacity, improve cost efficiency, and guide long-term infrastructure investment. * Lead ambiguous, company-critical technical initiatives across organizations, influencing strategy and aligning stakeholders around scientifically grounded decisions. * Mentor senior scientists and engineers, raise the technical bar for modeling and forecasting systems, and represent Meta's work through appropriate external publications, talks, or industry engagement. ## Related Videos - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [How building an industry DBMS differs from building a research one](https://www.wearedevelopers.com/videos/768-how-building-an-industry-dbms-differs-from-building-a-research-one) - [The Sustainability Race: AI's Promises, Pitfalls and Potential](https://www.wearedevelopers.com/videos/100155-the-sustainability-race-ai-s-promises-pitfalls-and-potential) - [Introduction to Azure Machine Learning](https://www.wearedevelopers.com/videos/368-introduction-to-azure-machine-learning) - [Your organization as a Graph](https://www.wearedevelopers.com/videos/2051-your-organization-as-a-graph) - [Building the Nervous System of AI - Michael Kagan (NVIDIA)](https://www.wearedevelopers.com/videos/2133-building-the-nervous-system-of-ai-michael-kagan-nvidia) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [How Much FAANG Companies Actually Pay Software Engineers in 2025](https://www.wearedevelopers.com/magazine/230-how-much-faang-companies-actually-pay-software-engineers-in-2025) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Résumé-Driven Development: How IT trends affect the job market for software developers](https://www.wearedevelopers.com/magazine/59-resume-driven-development-how-it-trends-affect-the-job-market-for-software-developers) - [How Much Does a Software Engineer Make? Realistic Software Engineering Salaries](https://www.wearedevelopers.com/magazine/425-how-much-does-a-software-engineer-make-realistic-software-engineering-salaries) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer)