> Markdown version of [/jobs/ext/3009388-senior-ai-reinforcement-learning-engineer](https://www.wearedevelopers.com/jobs/ext/3009388-senior-ai-reinforcement-learning-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # (Senior) AI / Reinforcement Learning Engineer - **Company:** Agile Robots Ag - **Location:** München, Germany - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Agile Methodology, Artificial Intelligence, C++ (Programming Language), Python (Programming Language), Kinematics, Motion Planning, Robotic Automation Software, Software Engineering, Reinforcement Learning, Pytorch, Information Technology, ONNX (Open Neural Network Exchange) Format - **Published:** September 20, 2026 - **Apply:** https://www.adzuna.de/details/5890816023 ## About the Role * Master's or PhD in robotics, computer science, engineering, or a related field * Proven experience (5+ years) applying RL or IL in physical systems, preferably on high DOF robots * Familiarity with common RL algorithms (PPO, SAC, ...) * Proficiency in Python and/or C++, PyTorch, ONNX, and common RL frameworks (e.g, Stable Baselines 3, RSL-RL) * Experience with physics-based simulators (IsaacSim, MuJoCo or equivalent) for large-scale policy training * Strong understanding of control theory, motion planning, and robot kinematics/dynamics * Demonstrated success in sim-to-real transfer and deploying learning-based policies on physical robots * Strong problem-solving skills and the ability to work independently in uncertain scenarios Beneficial Skills * Experience in training locomotion and locomanipulation policies * Experience in training and deploying on humanoid robots * Track record of publishing at CoRL, ICRA, RSS, NeurIPS, or ICLR ## Description Agile Robots SE is a high-tech startup based in Munich. Our mission is to bridge the gap between AI and robotics by developing robotic systems that offer state-of-the-art full-body force sensitivity and world-leading vision intelligence. This unique combination of technologies enables us to provide intelligent, easy-to-use and affordable robotic solutions with safe human-robot interaction. We are a dynamic and innovative software development company dedicated to pushing the boundaries of technology. We specialize in creating cutting-edge solutions that transform industries and redefine user experiences. We are building humanoid robots that work reliably, autonomously, in the real world. As a Reinforcement Learning Engineer, you will own the development of end-to-end control policies for Agile One: from training in simulation to deployment on hardware, across locomotion, manipulation, and whole-body coordination in unstructured environments. This is a high-ownership role. You will work at the frontier of what's possible with humanoid robots, make decisions with incomplete information, and iterate directly on physical systems. Success in this role requires both technical excellence and the ability to adapt and make decisions under uncertainty. Your Responsibilities * Design, maintain and deploy scalable, end-to-end training pipelines, including hyperparameter optimization * Transfer control policies from simulation to real robotic hardware * Integrate learned policies into a full-stack robotic system, including perception, planning, and actuation * Analyze robot behavior and learning performance,e and iterate quickly based on real-world results * Evaluate and benchmark state-of-the-art algorithms * Stay updated with the latest research in RL and IL * Work autonomously on open-ended, evolving tasks, * A dynamic high-tech company, combined with financial soundness and world-class investors. * Join an interdisciplinary, international team with 60+ different nationalities in a collaborative work environment. * Lots of development opportunities as we continue to grow. * Challenging tasks and impactful projects alongside experts that enable professional and personal growth. * Corporate Benefits Program covering health, mobility, and learning for 100€ net per month. * Modern office facilities with a rooftop terrace overlooking Munich, free drinks & fruits, and regular company events contribute to a good working environment ## Related Videos - [Robots are coming into the wild! Full-Stack Robotics Engineers, be ready!](https://www.wearedevelopers.com/videos/479-robots-are-coming-into-the-wild-full-stack-robotics-engineers-be-ready) - [How I built my own intelligent Robot Arm from Scratch](https://www.wearedevelopers.com/videos/100097-how-i-built-my-own-intelligent-robot-arm-from-scratch) - [Shift Left On Accessibility - Geri Reid](https://www.wearedevelopers.com/videos/1712-shift-left-on-accessibility-geri-reid) - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [How Robots Learn to be Robots](https://www.wearedevelopers.com/videos/1632-how-robots-learn-to-be-robots) - [Robots Among us: Advances in AI for Everyday Androids](https://www.wearedevelopers.com/videos/100230-robots-among-us-advances-in-ai-for-everyday-androids) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [What is Agentic Programming and Why Should Developers Care?](https://www.wearedevelopers.com/magazine/625-what-is-agentic-programming-and-why-should-developers-care) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [From Prototype to Production: Build AI Agents with This Free 4-Course Learning Path](https://www.wearedevelopers.com/magazine/655-from-prototype-to-production-build-ai-agents-with-this-free-4-course-learning-path)