AI QA Evaluation Engineer: Scale Metrics for ML

Appnovation Technologies
Greater London, UK
15 days ago
Apply on www.collegerecruiter.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours

Tech stack

Artificial Intelligence Continuous Integration Python (Programming Language) Large Language Models

Job description

Appnovation Technologies is seeking a QA / AI Evaluation Engineer to join a forward-thinking data/ML focused team. You will run large-scale evals, measure factual grounding and accuracy lift, and build metrics frameworks to demonstrate quality improvements across AI outputs.

Requirements

Applicants should have a strong background in QA/test engineering, Python, and experience with LLM evaluation frameworks, plus the ability to design automated evaluation harnesses and integrate them with CI/CD. #J-18808-Ljbffr

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.collegerecruiter.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:32 min

Fundamentals and limitations of large language models

Krzystof Czieslak · LIVE

51 sec

Continuous integration and continuous deployment core definitions

Chris Ayers · LIVE

3:46 min

Core terminology and audiences for interpretable artificial intelligence

Karol Przystalski · LIVE

2:06 min

Elevating the QA engineering role for complex challenges

Ondřej Gróf Ondřej Gróf · World Congress 2026 Europe

3:21 min

Automating complete quality assurance pipelines with artificial intelligence

Evelyn Haslinger · LIVE

2:25 min

Understanding the evolution and nature of large language models

Krzysztof Cieślak Krzysztof Cieślak · World Congress 2024

Videos

See all

Related articles

See all