RL Research Engineer: Safe, Scalable AI Systems

Anthropic
Greater London, UK
2 days ago
Apply on www.collegerecruiter.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Compensation
£260,000.0
Working hours
Regular working hours

Tech stack

Artificial Intelligence Python (Programming Language) Machine Learning Software Safety Reinforcement Learning

Job description

Anthropic is seeking a Research Engineer specializing in Reinforcement Learning to advance large language model capabilities. This role involves collaborative research and engineering, focused on optimizing core reinforcement learning infrastructure and driving performance through novel methodologies. We’re looking for candidates proficient in Python, with strong experience in machine-learning frameworks and systems design. The position offers a salary range of £260,000-£630,000 GBP and requires a Bachelor’s degree or equivalent, along with a commitment to AI safety and benefits.

Requirements

Anthropic is seeking a Research Engineer specializing in Reinforcement Learning to advance large language model capabilities. This role involves collaborative research and engineering, focused on optimizing core reinforcement learning infrastructure and driving performance through novel methodologies. We’re looking for candidates proficient in Python, with strong experience in machine-learning frameworks and systems design. The position offers a salary range of £260,000-£630,000 GBP and requires a Bachelor’s degree or equivalent, along with a commitment to AI safety and benefits. #J-18808-Ljbffr

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.collegerecruiter.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:01 min

Transitioning into the automotive artificial intelligence safety field

Tillman Radmer +2 · World Congress 2021

1:14 min

Addressing automotive mission-critical safety in embedded software development

David Romić · World Congress 2023

2:36 min

Applying supervised machine learning for practical rule extraction

Katja Träumner

1:45 min

Supervised, unsupervised, and reinforcement learning paradigms explained

Alexandra Waldherr · LIVE

2:14 min

Exploring internal AI product initiatives and global engineering roles

Maria Apazoglou · Coffee With Developers

2:44 min

Improving software safety through frequent production deployments

Steve Upton Steve Upton · World Congress 2024

Videos

See all

Related articles

See all