Remote Mathematics Expert

Mercor
Madrid, Spain
22 days ago
Apply on www.buscojobs.com.es
Prepare application

Role details

Contract type
Permanent contract
Employment type
Part-time / full-time
Compensation
€151,840.0
Working hours
Regular working hours

Tech stack

Artificial Intelligence Catalyst (Software) Large Language Models Model Validation

Job description

About the jobMercorconnects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors includeBenchmark,General Catalyst,Peter Thiel,Adam D’Angelo,Larry Summers, andJack Dorsey.Position:Mathematics AI EvaluatorType:Full-time or Part-time Contract WorkCompensation:$73/hourLocation:Geography restricted to USA, UK, Canada, EURole ResponsibilitiesWrite and refinepromptsto guide model behavior in mathematical contexts.EvaluateLLM-generated responsesto mathematics-related queries for correctness, rigor, and logical coherence.Verify mathematical claims, derivations, and proofs using domain expertise.Conduct fact-checking using authoritative public sources and domain knowledge.Annotate model responses by identifying strengths, areas of improvement, and factual or conceptual inaccuracies.Assess clarity, structure, and appropriateness of explanations for different audiences.QualificationsMust-HavePhD in Mathematicsor a closely related field, or demonstrated exceptional achievement in mathematics.Strong experience across core areas of mathematics: Algebra & Number Theory, Calculus & Analysis, Geometry & Topology, Discrete Mathematics, Logic & Computation, Probability & Statistics.Significant experience using large language models(LLMs).Excellent writing skillsand ability to explain complex mathematical concepts.Strong attention to detailand ability to notice subtle issues.Experience reviewing or editing technical or academic writing.PreferredPrior experience withRLHF, model evaluation, or data annotation work.Experience teaching, mentoring, or explaining mathematical concepts to non-expert audiences.Familiarity with evaluation rubrics, benchmarks, or structured review frameworks.Application Process (Takes *** mins to complete)Upload resumeAI interview based on your resumeSubmit formResources & SupportFor details about the interview process and platform information, please check: For any help or support, reach out to: ****PS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this opportunity.

Requirements

Must-Have PhD in Mathematics or a closely related field, or demonstrated exceptional achievement in mathematics. Strong experience across core areas of mathematics: Algebra & Number Theory, Calculus & Analysis, Geometry & Topology, Discrete Mathematics, Logic & Computation, Probability & Statistics. Significant experience using large language models (LLMs). Excellent writing skills and ability to explain complex mathematical concepts. Strong attention to detail and ability to notice subtle issues. Experience reviewing or editing technical or academic writing. Preferred Prior experience with RLHF , model evaluation, or data annotation work. Experience teaching, mentoring, or explaining mathematical concepts to non-expert audiences. Familiarity with evaluation rubrics, benchmarks, or structured review frameworks.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.buscojobs.com.es
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:42 min

Combining technical skills with strong business communication

Lukas Kölbl · LIVE

5:08 min

Validating requests and responses using data transfer objects

Roman Alexis Anastasini · World Congress 2021

1:13 min

Establishing shared vocabulary with Green Software Foundation principles

Pierre-Luc Noel +1 · World Congress 2023

3:32 min

Fundamentals and limitations of large language models

Krzystof Czieslak · LIVE

2:23 min

Preparing system prompts and evaluating models before launch

Julia Kasper · Coffee With Developers

2:49 min

Comparing software testing and software verification

Onur Kasimlar Onur Kasimlar · World Congress 2025

Videos

See all

Related articles

See all