> Markdown version of [/jobs/ext/1356636-senior-ai-engineer-reinforcement-learning](https://www.wearedevelopers.com/jobs/ext/1356636-senior-ai-engineer-reinforcement-learning). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior AI Engineer - Reinforcement Learning - **Company:** Resaro AI - **Location:** München, Germany - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Systems Engineering, Machine Learning, Reinforcement Learning, Deep Learning, Deployment Automation - **Published:** July 20, 2026 - **Apply:** https://www.adzuna.de/details/5651692694 ## About the Role * Master / Ph.D. in Robot Reinforcement Learning or a closely related field. + Proven track record in developing and implementing novel RL and ML algorithms, e.g. research or commercial implementation. + Demonstrated deep theoretical understanding of and practical experience with the RL framework, including bandit setting, (in-)finite horizon setting, on- and off-policy RL, and trust-region RL approaches. * Experience in Bayesian Machine Learning and probabilistic models. * Understanding of AI/ML/RL lifecycle and the state-of-the-art approaches and limitations of testing and validating complex use cases. * Strong skills in requirements gathering, stakeholder communication, and solution scoping. Nice-to-Have * Experience with fully differentiable deep learning for highly unstable systems. * Experience with Active Learning and RLHF. * Background in model compression and pruning for deploying large RL models onto edge devices. * Hands-on experience with Bayesian Meta-Learning to reduce training time and absolute error in complex models. * A strong portfolio of innovation, including multiple successful paper submissions at conferences like NeurIPS, ICML, ICLR, IROS, ICRA, CoRL, and a strong patent history. * Experience spearheading global AI initiatives and delivering AI solutions for both B2G (Unmanned Systems) and B2B (IoT) sectors. * Demonstrated success in leading cross-functional teams to deliver technical solutions. * Knowledge of deployment constraints in high-security or classified environments. * Prior exposure or experience with directly engaging senior stakeholders from Director to C-suite level. ## Description As the Senior AI Engineer - Reinforcement Learning, you will be the primary architect of our AI Test, Evaluation, Verification, and Validation (TEVV) product suite for reinforcement learning systems. You will lead the development of next-generation AI testing and assurance frameworks with applications in Autonomous Driving and Robotics. Your mission is to scale our capabilities in Reinforcement Learning, to ensure autonomous agents are safe, robust, and explainable in the field., * Independently implement Resaro's RL validation prototype to expose agent instability and vulnerability in a mission-critical and complex environment. * Build, lead and mentor a global, cross-functional, high-performing team of AI researchers and engineers as the RL practice scales. * Define the long-term vision and technical roadmap for RL TEVV, focusing on validating RL algorithms and learned policies in complex environments with mission-critical applications across system control, autonomous vehicles, and robotics. * Advance methods for learning probabilistic reward functions from human feedback (RLHF) to align AI behavior with mission goals. * Partner with Product Management to translate product vision, customer problems, and market opportunities into end-to-end solution architecture and technical roadmaps., * Help define the future of AI testing and assurance in real-world environments. * Collaborate with a tight-knit, expert team working at the intersection of AI, systems engineering, and policy. * Shape product direction while being close to the operational reality of AI deployments. Resaro is an Equal Opportunity Employer. We respect each individual and support the diverse cultures, perspectives, skills and experiences within our teams. ## Related Videos - [Introduction to Azure Machine Learning](https://www.wearedevelopers.com/videos/368-introduction-to-azure-machine-learning) - [Robots are coming into the wild! Full-Stack Robotics Engineers, be ready!](https://www.wearedevelopers.com/videos/479-robots-are-coming-into-the-wild-full-stack-robotics-engineers-be-ready) - [Getting Started with Machine Learning](https://www.wearedevelopers.com/videos/260-getting-started-with-machine-learning) - [Model Based Systems Engineering in an Agile Product Development Process](https://www.wearedevelopers.com/videos/68-model-based-systems-engineering-in-an-agile-product-development-process) - [What non-automotive Machine Learning projects can learn from automotive Machine Learning projects](https://www.wearedevelopers.com/videos/397-what-non-automotive-machine-learning-projects-can-learn-from-automotive-machine-learning-projects) - [Rethinking Recruiting: What you didn’t know about Responsible AI](https://www.wearedevelopers.com/videos/1090-rethinking-recruiting-what-you-didn-t-know-about-responsible-ai) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it) - [What Industries Outside of AI Are Hiring The Most AI Experts?](https://www.wearedevelopers.com/magazine/98-what-industries-outside-of-ai-are-hiring-the-most-ai-experts) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift)