> Markdown version of [/jobs/ext/1920023-software-engineer-core-machine-learning](https://www.wearedevelopers.com/jobs/ext/1920023-software-engineer-core-machine-learning). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Software Engineer, Core Machine Learning - **Company:** Facebook Inc. - **Location:** Sunnyvale, CA, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Automation of Tests, Code Generation, Profiling, Computer Engineering, Software Debugging, Software Design Documents, Distributed Computing Environment, Distributed Systems, Machine Learning, Open Source Technology, Tensorflow, Azure Machine Learning, Software Engineering, Strategies of Testing, Feature Engineering, Pytorch, Information Technology, Machine Learning Operations - **Published:** August 4, 2026 - **Apply:** https://www.jobmonkeyjobs.com/career/27901501/Software-Engineer-Core-Machine-Learning-California-Sunnyvale-7418 ## About the Role * Bachelor's degree in Computer Science, Computer Engineering, relevant technical field, or equivalent practical experience * 8+ years of experience in software engineering with a focus on machine learning systems, ML infrastructure, or large-scale distributed systems * Experience designing and implementing production ML systems such as distributed training frameworks, model serving infrastructure, or large-scale feature engineering pipelines * Experience leading major technical initiatives from design through production, including cross-team coordination and phased rollout management * Experience with performance analysis and optimization of ML training or inference workloads, including profiling, instrumentation, and bottleneck resolution * Experience communicating technical decisions and trade-offs in writing to both engineering and non-engineering stakeholders through design documents, architectural proposals, or postmortems, * Experience defining and operating ML platform reliability programs, including resiliency testing, SLO frameworks, and incident retrospective processes * Experience building or contributing to open-source ML frameworks or platform tooling such as PyTorch, TensorFlow, Ray, or similar systems * Demonstrated ongoing AI skill development (e.g., prompt/context engineering, agent orchestration) and staying current with emerging AI technologies * Demonstrated ability to integrate AI tools to optimize/redesign workflows and drive measurable impact (e.g., efficiency gains, quality improvements) * Experience applying AI-assisted development tools to accelerate engineering workflows, including code generation, automated testing, or intelligent debugging * Experience with hardware-software co-design for ML workloads, including quantization, model compression, or resource-efficient AI techniques * Experience adhering to and implementing responsible, ethical AI practices (e.g., risk assessment, bias mitigation, quality and accuracy reviews) ## Description Meta is seeking a Staff Software Engineer to join the Core Machine Learning team, focused on building and scaling the foundational ML infrastructure and systems that power Meta's family of products. In this role, you will architect and deliver high-impact ML platform capabilities - spanning training infrastructure, model serving, feature engineering pipelines, and AI-accelerated developer tooling - that enable thousands of engineers and researchers across Meta to build and ship state-of-the-art machine learning models at scale. Software Engineer, Core Machine Learning Responsibilities: * Architect and own large-scale ML infrastructure systems, including distributed training frameworks, model serving platforms, and feature computation pipelines that support production workloads across Meta's product surface * Lead the technical design and implementation of foundational ML platform components, evaluating trade-offs across performance, reliability, and developer experience * Drive end-to-end delivery of major ML infrastructure initiatives, coordinating across teams and disciplines to align on priorities, manage dependencies, and execute phased rollouts * Identify and resolve performance bottlenecks in ML training and inference systems through instrumentation, profiling, and targeted optimization * Define and enforce service level objectives for core ML platform services, building dashboards, alerting systems, and runbooks to reduce mean time to mitigation during incidents * Establish and advocate for engineering best practices in ML systems development, including testing strategies, safe rollout patterns, and AI-accelerated development workflows * Collaborate with research, product engineering, and infrastructure teams as a credible technical co-owner, independently driving design reviews, data analyses, and architectural decisions * Mentor other engineers on ML systems design, debugging complex distributed system issues, and applying AI tools to accelerate development velocity * Contribute to the team's technical roadmap by identifying opportunities to improve ML platform capabilities and obtaining buy-in from key stakeholders across the organization ## Related Videos - [Introduction to Azure Machine Learning](https://www.wearedevelopers.com/videos/368-introduction-to-azure-machine-learning) - [Machine learning in the browser with TensorFlowjs](https://www.wearedevelopers.com/videos/155-machine-learning-in-the-browser-with-tensorflowjs) - [Profiling Symfony & PHP apps with Blackfire](https://www.wearedevelopers.com/videos/265-profiling-symfony-php-apps-with-blackfire) - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [Getting Started with Machine Learning](https://www.wearedevelopers.com/videos/260-getting-started-with-machine-learning) - [Developing an AI.SDK](https://www.wearedevelopers.com/videos/198-developing-an-ai-sdk) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it)