Ai Evaluation Data Scientist (9-Month Contract With Bonus Incentives)

Multiverse Computing
Madrid, Spain
1 day ago
Apply on www.buscojobs.com.es
Prepare application

Role details

Contract type
Temporary contract
Employment type
Full-time (> 32 hours)
Experience required
2 years minimum
Working hours
Regular working hours

Tech stack

Artificial Intelligence Data Analysis Big Data Data Visualization Python (Programming Language) Machine Learning NumPy Quantum Computing Large Language Models Model Validation Pandas Scikit Learn
+4 more
Information Technology Data Analytics Machine Learning Operations Software Version Control

Job description

We are looking to fill this roleimmediatelyand are reviewing applications daily. Expect a fast, transparent process with quick feedback.Why join us?We are a European deep-tech leader in quantum and AI, backed by major global strategic investors and strong EU support. Our groundbreaking technology is already transforming how AI is deployed worldwide - compressing large language models by up to 95% without losing accuracy and cutting inference costs by **%.Joining us means working on cutting-edge solutions that make AI faster, greener, and more accessible - and being part of a company often described as a “quantum-AI unicorn in the making.”We offerCompetitive annual salaryTwo unique bonuses: signing bonus at incorporation and retention bonus at contract completion.Relocation package (if applicable).Up to 9-month contract, ending on June ****.Hybrid role and flexible working hours.Be part of a fast-scaling Series B company at the forefront of deep tech.Equal pay guaranteed.International exposure in a multicultural, cutting-edge environment.As an AI Evaluation Data Scientist, you will:Build robust evaluation pipelines that accurately mirror real production use cases for AI agentic systems and SML components.Define and track success metrics for both agentic workflows and key model components, focusing on measurement aligned with practical production performance rather than pure benchmarks.Design annotation frameworks and automated evaluation tools to conduct rigorous, repeatable analysis of model/system output.Execute deep-dive data investigations, monitor success and failure cases, and deliver actionable feedback loops to guide model/system improvements.Collaborate with engineers and researchers to shape experiments, validate hypotheses, and iterate toward higher performing model deployments.Communicate evaluation methodologies, results, and insights to technical and non-technical stakeholders, supporting ongoing improvement and transparency.Required QualificationsMaster’s degree in Computer Science, Data Science, Mathematics, Engineering, or related field.2+ years of experience in AI/ML model evaluation or large-scale data analysis.Strong proficiency in Python, and evaluation/data analysis libraries (NumPy, Pandas, scikit-learn, etc.).Direct experience creating evaluation pipelines for AI or data-driven systems, with a track record in metric definition and performance analysis.Familiarity with ML infrastructure (experiment tracking, model versioning, continuous evaluation).Experience working with large unstructured datasets and data visualization/reporting to communicate findings.Preferred QualificationsExperience evaluating generative AI agents, LLMs, or advanced SML components in production or R&D settings.Exposure to human annotation workflows and automated labeling tools.Knowledge of statistical analysis, fairness, bias, and robust metric development for model evaluation.Familiarity with ML experimentation platforms, cloud-based AI infrastructure, and advanced model validation techniques.About Multiverse ComputingFounded in ****, we are a well-funded, fast-growing deep-tech company with a team of 180+ employees worldwide. Recognized by CB Insights ** & **) as one of theTop 100 most promising AI companies globally, we are also the largest quantum software company in the EU.Our flagship products address critical industry needs:CompactifAI ? a groundbreaking compression tool for foundational AI models, reducing their size by up to 95% while maintaining accuracy, enabling portability across devices from cloud to mobile and beyond.Singularity ? a quantum and quantum-inspired optimization platform used by blue-chip companies in finance, energy, and manufacturing to solve complex challenges with immediate performance gains.You’ll be working alongside world-leading experts in quantum computing and AI, developing solutions that deliver real-world impact for global clients. We are committed to an inclusive, ethics-driven culture that values sustainability, diversity, and collaboration - a place where passionate people can grow and thrive. Come and join us!As an equal opportunity employer, Multiverse Computing is committed to building an inclusive workplace. The company welcomes people from alldifferent backgrounds, including age, citizenship, ethnic and racial origins, gender identities, individuals with disabilities, marital status, religions and ideologies, and sexual orientations to apply.

Requirements

Master’s degree in Computer Science, Data Science, Mathematics, Engineering, or related field. 2+ years of experience in AI/ML model evaluation or large-scale data analysis. Strong proficiency in Python, and evaluation/data analysis libraries (NumPy, Pandas, scikit-learn, etc.). Direct experience creating evaluation pipelines for AI or data-driven systems, with a track record in metric definition and performance analysis. Familiarity with ML infrastructure (experiment tracking, model versioning, continuous evaluation). Experience working with large unstructured datasets and data visualization/reporting to communicate findings. Preferred Qualifications Experience evaluating generative AI agents, LLMs, or advanced SML components in production or R&D settings. Exposure to human annotation workflows and automated labeling tools. Knowledge of statistical analysis, fairness, bias, and robust metric development for model evaluation. Familiarity with ML experimentation platforms, cloud-based AI infrastructure, and advanced model validation techniques.

Benefits & conditions

Competitive annual salary Two unique bonuses: signing bonus at incorporation and retention bonus at contract completion. Relocation package (if applicable). Up to 9-month contract, ending on June **. Hybrid role and flexible working hours. Be part of a fast-scaling Series B company at the forefront of deep tech. Equal pay guaranteed. International exposure in a multicultural, cutting-edge environment.

About the company

Founded in **, we are a well-funded, fast-growing deep-tech company with a team of 180+ employees worldwide. Recognized by CB Insights *** & **) as one of the Top 100 most promising AI companies globally , we are also the largest quantum software company in the EU. Our flagship products address critical industry needs: CompactifAI ? a groundbreaking compression tool for foundational AI models, reducing their size by up to 95% while maintaining accuracy, enabling portability across devices from cloud to mobile and beyond. Singularity ? a quantum and quantum-inspired optimization platform used by blue-chip companies in finance, energy, and manufacturing to solve complex challenges with immediate performance gains. You’ll be working alongside world-leading experts in quantum computing and AI, developing solutions that deliver real-world impact for global clients. We are committed to an inclusive, ethics-driven culture that values sustainability, diversity, and collaboration - a place where passionate people can grow and thrive. Come and join us! As an equal opportunity employer, Multiverse Computing is committed to building an inclusive workplace. The company welcomes people from all different backgrounds, including age, citizenship, ethnic and racial origins, gender identities, individuals with disabilities, marital status, religions and ideologies, and sexual orientations to apply.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.buscojobs.com.es
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:34 min

Maximizing execution memory effectively via python numpy broadcasting

Jodie Burchell · LIVE

2:03 min

Accelerating pandas dataframes using cudf module plugins

Ankit Patel Ankit Patel · World Congress 2024

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

3:22 min

Evaluating advanced artificial intelligence platforms for daily recruitment

Rudi Bauer Rudi Bauer +1 · Cappuccino with HR

1:25 min

Replacing NumPy with cuPy for straightforward GPU acceleration

Paul Graham Paul Graham · World Congress 2025

3:09 min

Evaluating artificial intelligence maturation in talent acquisition

Rudi Bauer Rudi Bauer +1 · Cappuccino with HR

Videos

See all

Related articles

See all