Machine Learning Engineer for AI Model Evaluation

Johns Hopkins Applied Physics Laboratory
Washington, DC, United States
8 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
2 years minimum
Compensation
$85,000.0 - $195,000.0
Working hours
Regular working hours

Tech stack

Artificial Intelligence Cursor (Graphical User Interface Elements) Machine Learning Large Language Models Model Validation Machine Learning Operations

Job description

Contribute to a frontier code agents project run with a leading AI research lab, evaluating and improving state of the art AI coding models via structured technical assessments. The work centers on realistic machine learning engineering workflows, model evaluation, and comparing outputs from multiple frontier models. Spots are limited and filling quickly on a first come, first serve basis. Key Responsibilities

  • Use frontier AI coding agents to complete and evaluate complex machine learning and AI engineering tasks.
  • Review model-generated implementations that involve model training, inference systems, MLOps, and large language model applications.
  • Identify bugs, edge cases, performance regressions, and failure modes in model outputs and implementations.
  • Compare and contrast outputs from multiple frontier models, assessing strengths, weaknesses, and tradeoffs.
  • Apply professional engineering judgment to realistic ML engineering scenarios and provide clear, actionable evaluations.

Requirements

  • Minimum 2 years of professional machine learning engineering experience.
  • Experience building production ML systems, model deployment infrastructure, LLM applications, or AI-powered products.
  • Regular, hands-on use of AI coding agents such as Cursor, Claude Code, Codex, Windsurf, Gemini CLI, or similar tools.
  • Ability to evaluate model-generated machine learning implementations and reason about technical tradeoffs.
  • Experience deploying ML systems to production is preferred., * You must meet the qualifications above, including the minimum 2 years of ML engineering experience.
  • Regular use of AI coding agents is required, and experience evaluating model-generated implementations is essential.

Benefits & conditions

  • Pay details provided in source materials: $85.00 per hour.
  • Project compensation also specified as $400 per accepted task, with typical tasks taking approximately 2 to 3 hours after ramp-up.
  • Payment is tied to accepted work, and accepted tasks are compensated at the stated per-task amount., + $85,000-195,000 per year Description Are you interested in developing advanced radar systems to make an impact on our nation’s defense? Do you want to be part of an enthusiastic team dedicated to advan…

  • 2 days ago +

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:36 min

Applying supervised machine learning for practical rule extraction

Katja Träumner

5:08 min

Validating requests and responses using data transfer objects

Roman Alexis Anastasini · WWC 2021

3:32 min

Fundamentals and limitations of large language models

Krzystof Czieslak · LIVE

3:22 min

Evaluating advanced artificial intelligence platforms for daily recruitment

Rudi Bauer Rudi Bauer +1 · Cappuccino with HR

7:10 min

Exploring pathways into the machine learning engineering field

Jose Luis Latorre Millas · LIVE

1:57 min

Evolution of machine learning algorithms and computing hardware

Alexandra Waldherr · LIVE

Videos

See all

Related articles

See all