> Markdown version of [/jobs/ext/3097211-research-scientist-artificial-intelligence](https://www.wearedevelopers.com/jobs/ext/3097211-research-scientist-artificial-intelligence). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Research Scientist, Artificial Intelligence - **Company:** Facebook Inc. - **Location:** Menlo Park, CA, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Program Optimization, Profiling, Computer Engineering, Software Design Documents, Memory Management, Linux Kernel, Machine Learning, Performance Tuning, High Performance Computing, Pytorch, Delivery Pipeline, Reliability of Systems, Information Technology, Data Analytics, Machine Learning Operations - **Published:** September 26, 2026 - **Apply:** https://www.jobmonkeyjobs.com/career/28057470/Research-Scientist-Artificial-Intelligence-California-Menlo-Park-7418 ## About the Role * Bachelor's degree in Computer Science, Computer Engineering, relevant technical field, or equivalent practical experience * 8+ years of experience in machine learning systems, model optimization, or high-performance computing research * Experience with TPU architecture and performance optimization, including profiling, kernel development, and memory management * Experience with XLA compilation, graph optimization, and low-level performance tuning for accelerator hardware * Experience developing and optimizing large-scale distributed training systems, including parallelism strategies such as data, tensor, and pipeline parallelism * Experience with PyTorch and its integration with accelerator backends * Experience communicating complex technical findings in writing, including technical reports, design documents, or peer-reviewed publications, * Experience developing custom kernels using Pallas or similar kernel authoring frameworks for TPU or GPU * Demonstrated track record of transitioning performance research into deployed systems used at significant scale * PhD in Computer Science, Machine Learning, Computer Architecture, or a related technical field, or equivalent depth of research experience * Publication record in systems for ML venues such as MLSys, OSDI, SOSP, or related AI conferences such as NeurIPS, ICML, or ICLR * Experience with Mixture of Experts (MoE) architectures and their optimization for efficient training and inference * Experience optimizing production-scale models with billions of parameters ## Description * Lead the design and execution of TPU performance optimization research, including kernel development, memory optimization, and compute efficiency improvements * Develop and optimize Pallas kernels for large-scale model training and inference on TPU architectures * Drive model optimization techniques including Mixture of Experts (MoE), tensor parallelism, pipeline parallelism, and other distributed training strategies * Optimize first party models within Meta's native PyTorch stack, ensuring efficient integration with XLA compilation and TPU execution * Identify and resolve complex technical challenges in model training efficiency, inference latency, and system reliability that require novel approaches * Define and drive multi-quarter research roadmaps for TPU optimization, aligning project milestones with broader organizational goals * Establish rigorous experimentation frameworks for performance benchmarking, including metric selection, profiling methodology, and data-driven optimization decisions * Translate research findings into production-ready optimizations by collaborating with engineering teams on deployment pipelines and reliability at scale * Communicate research findings and technical trade-offs clearly through publications, design documents, and presentations to both technical and non-technical audiences * Mentor other researchers and engineers on TPU optimization techniques, providing structured feedback on technical direction and experimental rigor ## Related Videos - [How AI Models Get Smarter](https://www.wearedevelopers.com/videos/1374-how-ai-models-get-smarter) - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [An Applied Introduction to eBPF with Go](https://www.wearedevelopers.com/videos/1075-an-applied-introduction-to-ebpf-with-go) - [Profiling Symfony & PHP apps with Blackfire](https://www.wearedevelopers.com/videos/265-profiling-symfony-php-apps-with-blackfire) - [Into the hive of eBPF!](https://www.wearedevelopers.com/videos/1199-into-the-hive-of-ebpf) - [Enhancing Workload Security in Kubernetes](https://www.wearedevelopers.com/videos/356-enhancing-workload-security-in-kubernetes) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [Stephan Gillich - Bringing AI Everywhere](https://www.wearedevelopers.com/magazine/489-stephan-gillich-bringing-ai-everywhere) - [The Best X (Twitter) Accounts for Developers](https://www.wearedevelopers.com/magazine/294-the-best-x-twitter-accounts-for-developers) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it)