> Markdown version of [/jobs/ext/2734544-reinforcement-learning-engineer-policy-optimus](https://www.wearedevelopers.com/jobs/ext/2734544-reinforcement-learning-engineer-policy-optimus). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Reinforcement Learning Engineer, Policy, Optimus - **Company:** Tesla Motors - **Location:** Palo Alto, CA, United States - **Salary:** $176,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Artificial Neural Networks, Python (Programming Language), NumPy, Reinforcement Learning, Pytorch, Deep Learning - **Published:** September 5, 2026 - **Apply:** https://diversityjobs.com/main/sendform/8/8/28176/1/8966866?backUrl=%2Fcareer%2F8966866%2FSoftware-Engineer-Reinforcement-Learning-Tesla-Bot-California-Palo-Alto ## About the Role * Experience in end-to-end robotic learning, with either imitation or reinforcement learning * Experience writing production-level Python (including Numpy and Pytorch) * Experience with distributed deep learning systems * Exposure to robot learning through tactile and/or vision-based sensors is a plus * Proven track record of training and deploying real world neural networks ## Description Tesla AI is solving robust embodied intelligence through humanoid robots.The goal ofourreinforcement learningteamis to build anddemonstratea general robot learning system that canleverageAI to perform complex physical tasks, ranging from full body locomotion, precise manipulation, and more.Our reinforcement and imitation learningengineersare responsible forend-to-end robotic learningand own this stack frominceptionto deployment. Most importantly, you will see your work repeatedly shippedto andutilizedby thousands of humanoid robots in real world applications. What You'll Do * Develop end-to-end robotic learning with either reinforcement or imitation learning * Reinforcing correct set of actions, rewarding correct behavior and negating incorrect behavior (with real-time action/reward feedback loops) * Perform a large number of instructions and generalize new tasks with different objects and environments * Learn to perform dexterous tasks using high degree of freedom hands * Learn different robot policies to solve language-conditioned tasks from vision * Ship production quality, safety-critical software ## Related Videos - [Getting Started with Machine Learning](https://www.wearedevelopers.com/videos/260-getting-started-with-machine-learning) - [On the straight and narrow path - How to get cars to drive themselves using reinforcement learning and trajectory optimization](https://www.wearedevelopers.com/videos/205-on-the-straight-and-narrow-path-how-to-get-cars-to-drive-themselves-using-reinforcement-learning-and-trajectory-optimization) - [Vectorize all the things! Using linear algebra and NumPy to make your Python code lightning fast.](https://www.wearedevelopers.com/videos/562-vectorize-all-the-things-using-linear-algebra-and-numpy-to-make-your-python-code-lightning-fast) - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [Geometric deep learning for drug discovery](https://www.wearedevelopers.com/videos/264-geometric-deep-learning-for-drug-discovery) - [How to implement convenient Python bindings to C++](https://www.wearedevelopers.com/videos/618-how-to-implement-convenient-python-bindings-to-c) ## Related Articles - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Dev Digest 137 - AI'm not sure about this](https://www.wearedevelopers.com/magazine/485-dev-digest-137-ai-m-not-sure-about-this) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [Graph and AI Trends 2026: Why Is AI Running but Not Yet Delivering?](https://www.wearedevelopers.com/magazine/680-graph-and-ai-trends-2026-why-is-ai-running-but-not-yet-delivering)