Research Engineer, Multimodal Reasoning For Information Literacy

Google LLC
Mountain View, CA, United States
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
2 years minimum
Compensation
$174,000.0 - $252,000.0
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Data Analysis Computer Vision Python (Programming Language) Machine Learning Language Modeling Rapid Prototyping Process Tensorflow Software Engineering Pytorch Large Language Models Prompt Engineering
+2 more
Deep Learning Information Technology

Job description

To succeed in this role, you will need to be passionate about advancing information literacy using machine learning and other computational techniques. You’ll join an interdisciplinary team of domain experts, ML researchers, and engineers to research and build multimodal reasoning systems and Vision-Language Models (VLMs) to assess the trustworthiness of media (images, audio, and videos) on the internet., * Plan and perform rapid prototyping of computer vision and multimodal machine learning techniques applied to determining authenticity of media information.

  • Design and train multimodal models capable of complex visual reasoning.
  • Undertake exploratory analysis to inform experimentation and research directions.
  • Engage with product teams to drive the development of our research.
  • Implement tools, libraries, and frameworks to speed up and enable new research.
  • Report and present research findings, software developments, experimental results, and data analysis clearly and efficiently.
  • Collaborate with internal and external scientific domain experts.

Requirements

In order to set you up for success as a Research Engineer at Google DeepMind, we look for the following skills and experience:

  • PhD/Master’s degree in Computer Science, AI, ML, or equivalent practical experience.
  • At least 2 years of relevant experience developing computer vision techniques or multimodal machine learning models.
  • Experience in software development using Python and deep learning frameworks (e.g., Jax, TensorFlow, PyTorch), with a proven track record of building high-quality research prototypes and systems.
  • Quantitative skills in math and statistics.
  • Experience exploring, analysing and visualising data.

In addition, the following would be an advantage:

  • Experience in training and deployment of large-scale models.
  • Experience with Video Understanding
  • Experience with Large Language Models, prompt engineering, few-shot learning, post-training techniques, and evaluations.
  • A proven track record of research or engineering achievements, such as publications in peer-reviewed conferences or journals.

When assessing technical background we will take a holistic view of the mix of scientific, ML and computational experience. We do not expect you to be an expert in all fields simultaneously.

At Google DeepMind, we value diversity of experience, knowledge, backgrounds and perspectives and harness these qualities to create extraordinary impact. We are committed to equal employment opportunity regardless of sex, race, religion or belief, ethnic or national origin, disability, age, citizenship, marital, domestic or civil partnership status, sexual orientation, gender identity, pregnancy, or related condition (including breastfeeding) or any other basis as protected by applicable law. If you have a disability or additional need that requires accommodation, please do not hesitate to let us know.

Benefits & conditions

The US base salary range for this full-time position is between 174,000 USD - 252,000 USD + bonus + equity + benefits. Your recruiter can share more about the specific salary range for your targeted location during the hiring process.

About the company

At Google DeepMind, our research team is dedicated to tackling the most complex challenges in online information quality. We strive to advance the state of the art by developing innovative solutions to detect manipulated media and misleading narratives, ensuring the integrity of digital discourse. A prominent example of our scientific discovery is Backstory (https://deepmind.google/discover/blog/exploring-the-context-of-online-images-with-backstory/) . Our interdisciplinary work spans provenance analysis and the creation of tools for AI-assisted information literacy, leveraging our technologies for the widespread public benefit of a safer online environment. We thrive in a supportive environment that encourages rapid prototyping and iteration, driving our research achievements directly into Google’s flagship models, including Gemini., Artificial Intelligence could be one of humanity’s most useful inventions. At Google DeepMind, we’re a team of scientists, engineers, machine learning experts and more, working together to advance the state of the art in artificial intelligence. We use our technologies for widespread public benefit and scientific discovery, and collaborate with others on critical challenges, ensuring safety and ethics are the highest priority.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on dejobs.org

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · WWC 2023

1:39 min

Fundamentals of tensors and the TensorFlow library

Håkan Silfvernagel · LIVE

1:48 min

Automating exploratory data analysis within training pipelines

Dora Petrella · WWC 2023

3:17 min

Balancing AI regulation with technological innovation in human resources

Rudi Bauer Rudi Bauer +1 · Cappuccino with HR

1:06 min

Compiling PyTorch environments for advanced time forecasting

Christoph Lohrmann Christoph Lohrmann +1 · WWC Europe 2026

1:31 min

Essential AI and human skills for future teams

Alexander Weißhaupt Alexander Weißhaupt +1 · WWC 2025

Videos

See all

Related articles

See all