> Markdown version of [/jobs/ext/2260959-rl-research-engineer-safe-scalable-ai-systems](https://www.wearedevelopers.com/jobs/ext/2260959-rl-research-engineer-safe-scalable-ai-systems). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # RL Research Engineer: Safe, Scalable AI Systems - **Company:** Anthropic - **Location:** Greater London, UK - **Salary:** £260,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Python (Programming Language), Machine Learning, Software Safety, Reinforcement Learning - **Published:** August 26, 2026 - **Apply:** https://www.collegerecruiter.com/job/2815051963-rl-research-engineer-safe-scalable-ai-systems ## About the Role Anthropic is seeking a Research Engineer specializing in Reinforcement Learning to advance large language model capabilities. This role involves collaborative research and engineering, focused on optimizing core reinforcement learning infrastructure and driving performance through novel methodologies. We're looking for candidates proficient in Python, with strong experience in machine-learning frameworks and systems design. The position offers a salary range of £260,000-£630,000 GBP and requires a Bachelor's degree or equivalent, along with a commitment to AI safety and benefits. #J-18808-Ljbffr ## Description Anthropic is seeking a Research Engineer specializing in Reinforcement Learning to advance large language model capabilities. This role involves collaborative research and engineering, focused on optimizing core reinforcement learning infrastructure and driving performance through novel methodologies. We're looking for candidates proficient in Python, with strong experience in machine-learning frameworks and systems design. The position offers a salary range of £260,000-£630,000 GBP and requires a Bachelor's degree or equivalent, along with a commitment to AI safety and benefits. ## Related Videos - [How AI Models Get Smarter](https://www.wearedevelopers.com/videos/1374-how-ai-models-get-smarter) - [Psychological Safety in Software Engineering - Jenny-Margrethe Vej & Alexandra Hou Aldershaab](https://www.wearedevelopers.com/videos/2142-psychological-safety-in-software-engineering-jenny-margrethe-vej-alexandra-hou-aldershaab) - [On the straight and narrow path - How to get cars to drive themselves using reinforcement learning and trajectory optimization](https://www.wearedevelopers.com/videos/205-on-the-straight-and-narrow-path-how-to-get-cars-to-drive-themselves-using-reinforcement-learning-and-trajectory-optimization) - [Introduction to Azure Machine Learning](https://www.wearedevelopers.com/videos/368-introduction-to-azure-machine-learning) - [Getting Started with Machine Learning](https://www.wearedevelopers.com/videos/260-getting-started-with-machine-learning) - [What non-automotive Machine Learning projects can learn from automotive Machine Learning projects](https://www.wearedevelopers.com/videos/397-what-non-automotive-machine-learning-projects-can-learn-from-automotive-machine-learning-projects) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [What Industries Outside of AI Are Hiring The Most AI Experts?](https://www.wearedevelopers.com/magazine/98-what-industries-outside-of-ai-are-hiring-the-most-ai-experts) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [Should AI be Regulated? The Arguments For and Against](https://www.wearedevelopers.com/magazine/271-should-ai-be-regulated-the-arguments-for-and-against)