Senior Research Scientist, Gemini Release Evaluations, DeepMind

Google LLC
New York, NY, United States
3 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
1 year minimum
Working hours
Regular working hours

Tech stack

Artificial Intelligence Computer Programming Data Structures Machine Learning Large Language Models Information Technology

Job description

Experteer Overview In this role you advance AI research at DeepMind, shaping impactful scientific work and contributing to product innovations for billions of users. You will lead research agendas, publish findings, and help align academic excellence with real-world deployment. You will drive data-driven evaluation, expand datasets, and bridge research results with production systems. This opportunity enables you to influence frontier AI responsibly within a collaborative, interdisciplinary environment. Compensation / Benefits * Author and disseminate research findings to the community and across functions * Define data structures, frameworks, design, and evaluation metrics; manage timelines and resources * Develop SOTA datasets to push frontier models and close gaps in release evaluations * Clarify the relationship between static evaluations and live production traffic Tasks * PhD in Computer Science or related field (or equivalent practical experience) * 2 years leading a research agenda * 1 year of experience in a data science field * Experience with AI model training, testing, evaluation, and tuning; familiarity with LLM release cycles and timeline management * 2 years coding experience (preferred) Key requirements * bonus target (15%) * equity * benefits

Requirements

Experteer Overview In this role you advance AI research at DeepMind, shaping impactful scientific work and contributing to product innovations for billions of users. You will lead research agendas, publish findings, and help align academic excellence with real-world deployment. You will drive data-driven evaluation, expand datasets, and bridge research results with production systems. This opportunity enables you to influence frontier AI responsibly within a collaborative, interdisciplinary environment. Compensation / Benefits * Author and disseminate research findings to the community and across functions * Define data structures, frameworks, design, and evaluation metrics; manage timelines and resources * Develop SOTA datasets to push frontier models and close gaps in release evaluations * Clarify the relationship between static evaluations and live production traffic Tasks * PhD in Computer Science or related field (or equivalent practical experience) * 2 years leading a aaaaaau _ agenda * 1 year of experience in a data science field * Experience with AI model training, testing, evaluation, and tuning; familiarity with LLM release cycles and timeline management * 2 years coding experience (preferred) Key requirements * bonus target (15%) * equity * benefits

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on us.experteer.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

4:04 min

Building practical AI agents using Google Gemini

Philipp Schmid Philipp Schmid · WWC 2025

4:42 min

Building robust data structures with structs and bound functions

Rainer Stropek Rainer Stropek · WWC 2021

2:36 min

Applying supervised machine learning for practical rule extraction

Katja Träumner

3:31 min

Revolutionizing computer programming through natural language code generation

Demetris Cheatham Demetris Cheatham +1 · WWC 2024

3:02 min

Understanding Google Gemini history and available context models

4:22 min

Representing domain entities as plain data structures

Dan Lebrero · LIVE

Videos

See all

Related articles

See all