Research Engineer, Multimodal Reasoning For Information Literacy
Job description
About the role
At Google DeepMind, the Research Engineer, Multimodal Reasoning For Information Literacy role centers on tackling the most complex challenges in online information quality. The research team is dedicated to advancing the state of the art by developing innovative solutions to detect manipulated media and misleading narratives, ensuring the integrity of digital discourse. A prominent example of this scientific discovery is the Backstory project, which explores the context of online images. The work is interdisciplinary, spanning provenance analysis and the creation of tools for AI-assisted information literacy, leveraging technologies for the widespread public benefit of a safer online environment. The team thrives in a supportive environment that encourages rapid prototyping and iteration, driving research achievements directly into Google's flagship models, including Gemini.
You will own the end to end research and delivery of multimodal reasoning systems that assess the trustworthiness of online images, audio, and video. You will design and train Vision-Language Models capable of complex visual reasoning to support information literacy at scale. Your work will directly underpin public facing tools that analyze media context and detect manipulated content. You will collaborate closely with interdisciplinary domain experts to translate scientific insight into robust prototypes. This role emphasizes rapid experimentation and iteration to advance the state of the art in online information quality. Your contributions will be integrated into Google's flagship models including Gemini, driving real world impact for public benefit.
What you'll do
Investigate and prototype computer vision and multimodal machine learning techniques to determine the authenticity of media information. Design and train multimodal models capable of complex visual reasoning across diverse data sources. Conduct exploratory analysis to shape experimentation directions and research priorities. Engage with product teams to translate research findings into practical development roadmaps. Implement tools, libraries, and frameworks that accelerate research workflows and enable scalable solutions. Report and present research findings, software developments, experimental results, and data analysis clearly and efficiently to varied audiences. Collaborate with internal and external scientific domain experts to validate methodologies and align on standards. Evaluate emerging techniques in video understanding and large language models within the context of information integrity. Integrate provenance analysis and AI assisted information literacy tools into coherent system demonstrations. Champion ethical considerations and safety practices throughout the research and deployment lifecycle. Contribute to high impact publications and external collaborations that define best practices for media trustworthiness. Drive innovation by iterating quickly on prototypes and demonstrating concrete progress in real world scenarios.
Requirements
Hold a PhD or Master's degree in Computer Science, Artificial Intelligence, Machine Learning, or equivalent practical experience. Possess at least 2 years of relevant experience developing computer vision techniques or multimodal machine learning models. Demonstrate strong software development skills using Python and deep learning frameworks such as Jax, TensorFlow, or PyTorch. Show a proven track record of building high quality research prototypes and production ready systems. Bring quantitative skills in mathematics and statistics to analyze model behavior and experimental outcomes. Have hands on experience exploring, analyzing, and visualizing complex datasets to inform research decisions. Commit to working in a collaborative environment where interdisciplinary teamwork is essential. Embrace a culture of learning and rapid iteration, supporting peers and incorporating feedback to improve results. Align with the mission of using technology for widespread public benefit and maintaining high standards of safety and ethics. Demonstrate ownership of complex problems and the initiative to drive solutions from conception to deployment.
About us
Artificial Intelligence could be one of humanity's most useful inventions. At Google DeepMind, we are a team of scientists, engineers, machine learning experts and more, working together to advance the state of the art in artificial intelligence. We use our technologies for widespread public benefit and scientific discovery, and collaborate with others on critical challenges, ensuring safety and ethics are the highest priority. We are a dedicated scientific community, committed to "solving intelligence" and ensuring our technology is used for widespread public benefit. We have built a supportive and inclusive environment where collaboration is encouraged and learning is shared freely. We do not set limits based on what others think is possible or impossible. We drive ourselves and inspire each other to push boundaries and achieve ambitious goals.
Compensation and location
Location: USA