> Markdown version of [/jobs/ext/1578555-research-engineer-multimodal-reasoning-for-information-literacy](https://www.wearedevelopers.com/jobs/ext/1578555-research-engineer-multimodal-reasoning-for-information-literacy). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Research Engineer, Multimodal Reasoning For Information Literacy - **Company:** Google LLC - **Location:** Mountain View, CA, United States - **Experience:** Experienced - **Salary:** $174,000.0 - $252,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Data Analysis, Computer Vision, Python (Programming Language), Machine Learning, Language Modeling, Rapid Prototyping Process, Tensorflow, Software Engineering, Pytorch, Large Language Models, Prompt Engineering, Deep Learning, Information Technology - **Published:** July 9, 2026 - **Apply:** https://dejobs.org/x/x/E1D07701DC744118AB0A64B7D9C3EAB9/job/ ## About the Role In order to set you up for success as a Research Engineer at Google DeepMind, we look for the following skills and experience: * PhD/Master's degree in Computer Science, AI, ML, or equivalent practical experience. * At least 2 years of relevant experience developing computer vision techniques or multimodal machine learning models. * Experience in software development using Python and deep learning frameworks (e.g., Jax, TensorFlow, PyTorch), with a proven track record of building high-quality research prototypes and systems. * Quantitative skills in math and statistics. * Experience exploring, analysing and visualising data. In addition, the following would be an advantage: * Experience in training and deployment of large-scale models. * Experience with Video Understanding * Experience with Large Language Models, prompt engineering, few-shot learning, post-training techniques, and evaluations. * A proven track record of research or engineering achievements, such as publications in peer-reviewed conferences or journals. When assessing technical background we will take a holistic view of the mix of scientific, ML and computational experience. We do not expect you to be an expert in all fields simultaneously. At Google DeepMind, we value diversity of experience, knowledge, backgrounds and perspectives and harness these qualities to create extraordinary impact. We are committed to equal employment opportunity regardless of sex, race, religion or belief, ethnic or national origin, disability, age, citizenship, marital, domestic or civil partnership status, sexual orientation, gender identity, pregnancy, or related condition (including breastfeeding) or any other basis as protected by applicable law. If you have a disability or additional need that requires accommodation, please do not hesitate to let us know. ## Description To succeed in this role, you will need to be passionate about advancing information literacy using machine learning and other computational techniques. You'll join an interdisciplinary team of domain experts, ML researchers, and engineers to research and build multimodal reasoning systems and Vision-Language Models (VLMs) to assess the trustworthiness of media (images, audio, and videos) on the internet., * Plan and perform rapid prototyping of computer vision and multimodal machine learning techniques applied to determining authenticity of media information. * Design and train multimodal models capable of complex visual reasoning. * Undertake exploratory analysis to inform experimentation and research directions. * Engage with product teams to drive the development of our research. * Implement tools, libraries, and frameworks to speed up and enable new research. * Report and present research findings, software developments, experimental results, and data analysis clearly and efficiently. * Collaborate with internal and external scientific domain experts. ## Related Videos - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [Machine learning in the browser with TensorFlowjs](https://www.wearedevelopers.com/videos/155-machine-learning-in-the-browser-with-tensorflowjs) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Google Gemini: Open Source and Deep Thinking Models - Sam Witteveen](https://www.wearedevelopers.com/videos/1332-google-gemini-open-source-and-deep-thinking-models-sam-witteveen) - [How We Built a Machine Learning-Based Recommendation System (And Survived to Tell the Tale)](https://www.wearedevelopers.com/videos/752-how-we-built-a-machine-learning-based-recommendation-system-and-survived-to-tell-the-tale) - [Geometric deep learning for drug discovery](https://www.wearedevelopers.com/videos/264-geometric-deep-learning-for-drug-discovery) ## Related Articles - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [DeepSeek R1 vs ChatGPT o1: How Do They Compare?](https://www.wearedevelopers.com/magazine/542-deepseek-r1-vs-chatgpt-o1-how-do-they-compare) - [The Prompt Engineer ✍️](https://www.wearedevelopers.com/magazine/216-the-prompt-engineer)