Reinforcement Learning Engineer

Bright Vision Technologies
Novi, MI, United States
14 days ago
Apply on www.careerjet.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
6 years minimum
Compensation
$96,000.0 - $120,000.0
Working hours
Regular working hours

Tech stack

Artificial Neural Networks Computer Clusters Python (Programming Language) Machine Learning Reinforcement Learning Supervised Learning Large Language Models Multi-Agent Systems Deep Learning Information Technology Free and Open-Source Software

Job description

We are looking for a Reinforcement Learning Engineer to design, train, and deploy RL-based systems for high-impact decision-making problems where supervised learning alone is insufficient. The role requires deep familiarity with modern reinforcement learning algorithms, simulation environments, reward modeling, and the engineering complexity of training and evaluating policies at scale. The ideal candidate has both research depth and engineering pragmatism, with experience taking RL solutions out of the lab and into production where stability, safety, and ongoing improvement are critical.

Requirements

  • Master’s or PhD in Computer Science, Machine Learning, or a related field; or equivalent applied experience.
  • Six or more years of combined RL research and engineering experience.
  • Strong proficiency in Python and modern deep learning frameworks.
  • Hands-on experience with at least one major RL library or in-house RL stack.
  • Solid understanding of probability, optimization, and the theoretical foundations of RL.
  • Experience designing and tuning reward functions in non-trivial environments.
  • Familiarity with simulation environments and large-scale experience collection.
  • Experience training neural network policies on GPU clusters.
  • Strong written and verbal communication skills.
  • Track record of shipping or publishing impactful RL work.

Preferred Qualifications

  • Experience with RLHF for large language models.
  • Familiarity with multi-agent RL or hierarchical RL.
  • Exposure to robotics, control systems, or autonomous driving.
  • Publications in RL or related research venues.
  • Open-source contributions to RL libraries or environments.

About the company

Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential., Aquent Talent

  • Ann Arbor, MI
  • $42.06-46.73 per hour Partnering with Aquent, we are thrilled to represent a leading organization at the forefront of financial innovation. This institution is dedicated to building secure, reliable, an…

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:25 min

Distinguishing artificial intelligence from deep learning

Sam Witteveen · Coffee With Developers

2:36 min

Applying supervised machine learning for practical rule extraction

Katja Träumner

6:12 min

Provisioning a free SingleStore workspace and computing cluster

Akmal Chaudhri Akmal Chaudhri · LIVE

1:02 min

Training models with reward functions and reinforcement learning

Carl Lapierre Carl Lapierre · World Congress 2024

3:51 min

Overcoming hardware configuration barriers in machine learning

Jose Luis Latorre Millas · LIVE

2:17 min

Distinguishing between AI, machine learning, and deep learning

Mary Grygleski Mary Grygleski · LIVE

Videos

See all

Related articles

See all