> Markdown version of [/jobs/ext/2300841-ai-research-engineer-reinforcement-learning](https://www.wearedevelopers.com/jobs/ext/2300841-ai-research-engineer-reinforcement-learning). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # AI Research Engineer - Reinforcement Learning - **Company:** Helsing - **Location:** Barcelona, Spain (Remote available) - **Contract:** Temporary contract - **Skills:** Artificial Intelligence, Formal Verification, Python (Programming Language), Software Engineering, Reinforcement Learning, Multi-Agent Systems, Low Latency, Operational Systems, C++14 - **Published:** August 30, 2026 - **Apply:** https://www.jobleads.com/es/job/ea88182877829ebd81563142b3ef20cd7 ## About the Role * Hold a MSc in Reinforcement Learning, Robotics, Automation and Control, or a closely related field, with a strong focus on sequential decision-making and autonomous systems. * Have hands-on experience building, training, and deploying reinforcement learning agents and iterated on a policy beyond simulation, understanding what it takes to make learned behaviour reliable in a real operational system. * Are deeply familiar with modern RL and multi-agent RL techniques, including but not limited to model-free methods (e.g. PPO, SAC), population-based training, handling of partial observability and long horizons. * Have experience integrating RL policies into high-performance runtime systems, with a solid understanding of the latency and throughput constraints that come with real-time autonomous decision-making. * Possess solid software engineering skills, writing clean and well-structured code in Python and/or languages like Rust or modern C++, and have experience deploying AI software to production including testing, QA, and monitoring. * Have excellent communication skills and the ability to report and present research findings clearly and efficiently, both internally and externally. * Are passionate about keeping up to date with current research and enjoy reimplementing and extending state-of-the-art approaches in deep reinforcement learning. Note: We operate at an intersection where women, as well as other minority groups, are systematically under-represented. We encourage you to apply even if you don't meet all the listed qualifications; ability and impact cannot be summarised in a few bullet points., * PhD in Reinforcement Learning, Multi-Agent Systems, Automation and Control, Robotics, or a related field, with publications in top-tier venues. * Experience with large-scale distributed RL training frameworks, the infrastructure challenges of running thousands of parallel simulation environments, and GPU-based simulators. * Experience modelling and training multi-agent controllers using state-of-the-art techniques, including emergent coordination, competitive self-play, or decentralised execution with centralised training. * Familiarity with flight dynamics, aerospace systems, or guidance, navigation, and control (GNC) concepts. * Experience deploying AI software to safety-critical production systems, including formal verification, testing pipelines, and runtime monitoring. ## Related Videos - [Swapping Low Latency Data Storage Under High Load](https://www.wearedevelopers.com/videos/746-swapping-low-latency-data-storage-under-high-load) - [When testing just doesn’t cut it](https://www.wearedevelopers.com/videos/720-when-testing-just-doesn-t-cut-it) - [Designing UX for SRE Agents in High-Stakes Incidents](https://www.wearedevelopers.com/videos/100003-designing-ux-for-sre-agents-in-high-stakes-incidents) - [Unleash the power of 5G in your code: transform your apps](https://www.wearedevelopers.com/videos/1567-unleash-the-power-of-5g-in-your-code-transform-your-apps) - [Agentic employees in world's most downloaded FinTech app](https://www.wearedevelopers.com/videos/100123-agentic-employees-in-world-s-most-downloaded-fintech-app) - [Developer’s Perspective: Overview of the Tezos Blockchain Ecosystem](https://www.wearedevelopers.com/videos/237-developer-s-perspective-overview-of-the-tezos-blockchain-ecosystem) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [Dev Digest 137 - AI'm not sure about this](https://www.wearedevelopers.com/magazine/485-dev-digest-137-ai-m-not-sure-about-this) - [Graph and AI Trends 2026: Why Is AI Running but Not Yet Delivering?](https://www.wearedevelopers.com/magazine/680-graph-and-ai-trends-2026-why-is-ai-running-but-not-yet-delivering) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift)