Remote Domain Expert Reviewer - Sports

Cornerstone Barricades
San Francisco, CA, United States
1 day ago
Apply on turing.betterteam.com
Prepare application

Role details

Contract type
Contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Compensation
$83,200.0 - $104,000.0
Working hours
Regular working hours
Languages
English

Tech stack

Artificial Intelligence Large Language Models Prompt Engineering

Job description

We are seeking a highly qualified Sports Domain Reviewer to support the quality assurance of Large Language Model (LLM) evaluation projects. In this role, you will review domain-specific prompts, evaluate completed tasks for accuracy and quality, ensure adherence to project guidelines, and provide actionable feedback to maintain high annotation standards. Your expertise in sports will help ensure the reliability, consistency, and quality of AI evaluation data across a wide range of sporting disciplines., * Review and validate domain-specific prompts covering global sports, leagues, athletes, tournaments, rules, statistics, sports history, analytics, and related topics.

  • Evaluate completed tasks to ensure factual accuracy, reasoning quality, completeness, and compliance with project guidelines.
  • Identify factual inaccuracies, logical inconsistencies, hallucinations, outdated information, and low-quality annotations.
  • Ensure prompts are challenging, relevant, and aligned with project objectives.
  • Provide clear, constructive, and evidence-based feedback to contributors to improve task quality.
  • Maintain consistency across reviews by following established quality standards and review guidelines.
  • Escalate ambiguous or complex cases when necessary and document review findings.
  • Collaborate with project managers and AI teams to continuously improve evaluation quality and review processes.

Requirements

  • Master’s degree or higher in any field; degrees in Sports Management, Sports Science, Journalism, Communications, Media Studies, or other sports-related disciplines are preferred.
  • Strong knowledge of sports, including familiarity with major leagues, teams, athletes, tournaments, rules, statistics, and key developments across one or more sports.
  • 3+ years of relevant professional experience, preferably in sports journalism, sports media, sports content, research, analysis, reporting, or a related field.
  • Excellent written English, research, and analytical skills.
  • Strong attention to detail and ability to assess sports-related information accurately., * Experience reviewing AI-generated content, LLM evaluations, prompt engineering, or annotation quality.
  • Familiarity with quality assurance, editorial review, or technical content evaluation workflows.
  • Strong understanding of scientific principles, technological advancements, and emerging technologies.
  • Ability to provide objective, evidence-based feedback while maintaining consistent quality standards.
  • Experience working independently in a fast-paced, quality-focused environment.

About the company

Based in San Francisco, California, Turing is the world’s leading research accelerator for frontier AI labs and a trusted partner for global enterprises deploying advanced AI systems. Turing supports customers in two ways: first, by accelerating frontier research with high-quality data, advanced training pipelines, plus top AI researchers who specialize in coding, reasoning, STEM, multilinguality, multimodality, and agents; and second, by applying that expertise to help enterprises transform AI from proof of concept into proprietary intelligence with systems that perform reliably, deliver measurable impact, and drive lasting results on the P&L.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on turing.betterteam.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:57 min

Evaluating large language model responses to sports data

Celeste Horgan Celeste Horgan · World Congress 2026 Europe

1:46 min

Refining complex software instructions using automated prompt engineers

Markus Walker Markus Walker · World Congress 2023

3:32 min

Fundamentals and limitations of large language models

Krzystof Czieslak · LIVE

3:46 min

Core terminology and audiences for interpretable artificial intelligence

Karol Przystalski · LIVE

1:38 min

Evaluating agent code via previews and critic models

Guillaume Moigneu Guillaume Moigneu · World Congress 2026 Europe

3:29 min

Transitioning from basic prompt engineering to context engineering

Himanshu Vasishth Himanshu Vasishth +3 · World Congress 2025

Videos

See all

Related articles

See all