Deepmind Logo

Deepmind

Research Engineer, Human Understanding

Reposted One Month Ago
Be an Early Applicant
In-Office
Mountain View, CA, USA
174K-252K Annually
Mid level
In-Office
Mountain View, CA, USA
174K-252K Annually
Mid level
The Research Engineer will develop and deploy multimodal AI models, conduct applied research, and contribute to scalable infrastructure for understanding human likeness across modalities.
The summary above was generated by AI
Snapshot

We are seeking a highly motivated Research Engineer (L5) with a strong background in multi-modal modelling for humans and a focus on speech & audio/visual to join the effort within Google DeepMind's Frontier AI unit. This role is pivotal in developing foundational multimodal AI capabilities to understand, generate, and protect human likeness. As a key contributor, you will design and implement cutting-edge models and frameworks, pushing the boundaries of AI to enable foundational capabilities for human-centric understanding and generation. This is a unique opportunity to contribute to impactful research and advance Google DeepMind's mission towards Artificial General Intelligence (AGI).

About us

Artificial Intelligence could be one of humanity’s most useful inventions. At Google DeepMind, we’re a team of scientists, engineers, machine learning experts and more, working together to advance the state of the art in artificial intelligence and ultimately achieve Artificial General Intelligence. We use our technologies for widespread public benefit and scientific discovery, and collaborate with others on critical challenges, ensuring safety and ethics are the highest priority.

The effort is a part of Google DeepMind's Frontier AI unit. The team aims to build holistic representation encompassing a full spectrum of human understanding. We develop systems to provide perception skills critical for person-centric applications, which is crucial for enabling AI to interact naturally & seamlessly, depict humans accurately &  responsibly in generative AI, and build trustworthy & resilient systems that can detect and prevent misuse like deepfakes and impersonation.

The role

You will drive outcomes for critical technical components aimed at advancing our capabilities in multimodal human understanding. You will play a critical role in developing and deploying models that can provide accurate human understanding across multiple modalities (e.g., visual appearance, voice, dynamics, etc), while also building robust defenses against sophisticated AI-driven manipulation and impersonation.

This role involves tackling complex, ambiguous problems with no obvious "best" solution, requiring independent judgment and a proactive approach to exploring multiple technical avenues. You will be instrumental in shaping the technical direction for core components of the effort. Your contribution will lead to key breakthrough and impactful landings within GDM and across Google products, ensuring our technologies are both groundbreaking and responsibly deployed.

Key responsibilities
  • Advance multimodal human representations & understanding : Research and implement novel models and other multimodal techniques for a more holistic understanding of humans across visual, audio, and textual data.
  • Conduct applied research: Conduct experimental research cycles from hypothesis to deployment.
  • Drive technical projects: Take ownership of substantial technical projects within the effort, from ideation and design to implementation and evaluation, often involving cross-functional collaboration.
  • Contribute to Infrastructure: Inform and contribute to the development of scalable and efficient research infrastructure for multimodal human understanding models and datasets.
  • Design and execute strategies for tuning and adapting VLMs and other foundation models for specific tasks
About you

In order to set you up for success as a Research Engineer at Google DeepMind, we look for the following skills and experience:

Requirements:

  • PhD degree in Computer Science, Machine Learning, or a related technical field with 3+ years of relevant experience.
  • Experience in developing machine learning models, such as audio & speech-visual models.
  • Experience in working with and tuning large-scale vision language models.
  • Strong programming skills in Python and experience with at least one major deep learning framework (e.g., JAX)
  • Experience conducting independent research and development, including experimental design, implementation, and analysis.

In addition, the following would be an advantage:

  • Experience with Generative AI techniques and architectures.
  • Familiarity with Reinforcement Learning or alignment methods.
  • A track record of publications in top-tier AI/ML conferences (e.g., NeurIPS, ICML, ICLR, CVPR, ICCV, ECCV).
  • Experience with multimodal learning, integrating information from different data types (e.g., vision, audio, text).
  • Understanding of privacy-preserving machine learning or responsible AI practices.

The US base salary range for this full-time position is between 174,000 USD - 252,000 USD + bonus + equity + benefits. Your recruiter can share more about the specific salary range for your targeted location during the hiring process.

Note: In the event your application is successful and an offer of employment is made to you, any offer of employment will be conditional on the results of a background check, performed by a third party acting on our behalf. For more information on how we handle your data, please see our Applicant and Candidate Privacy Policy.

At Google DeepMind, we value diversity of experience, knowledge, backgrounds and perspectives and harness these qualities to create extraordinary impact. We are committed to equal employment opportunities regardless of sex, race, religion or belief, ethnic or national origin, disability, age, citizenship, marital, domestic or civil partnership status, sexual orientation, gender identity, pregnancy, or related condition (including breastfeeding) or any other basis as protected by applicable law. If you have a disability or additional need that requires accommodation, please do not hesitate to let us know.

  

Deepmind Mountain View, California, USA Office

Ampitheatre Pkwy, Mountain View, CA, United States, 94034

Deepmind San Francisco, California, USA Office

San Francisco, United States

Similar Jobs

An Hour Ago
Remote or Hybrid
2 Locations
173K-312K Annually
Senior level
173K-312K Annually
Senior level
Fintech • Payments • Software • Financial Services
Lead Sales Development across North America and the UK, owning top-of-funnel pipeline, regional alignment, performance reporting, upmarket expansion, cross-functional partnerships, talent development, and technology-enabled productivity. The role manages teams and managers, establishes scalable inbound and outbound motions, partners with Marketing, Sales, Finance, and Strategy, and drives revenue-focused planning, qualification, and conversion.
Top Skills: AIAnalyticsAutomationReporting Tools
2 Hours Ago
Hybrid
San Mateo, CA, USA
220K-290K Annually
Senior level
220K-290K Annually
Senior level
Artificial Intelligence • Legal Tech
Leads and scales product marketing for a legal AI platform, owning positioning, narrative, launches, pricing, packaging, enablement, competitive strategy, customer research, market intelligence, and category leadership. Partners cross-functionally with Product, Sales, Customer Success, Finance, and Marketing to drive adoption, revenue, expansion, and win rates. Builds the product marketing team and operating systems while combining strategic leadership with hands-on execution.
Top Skills: Artificial IntelligenceSaaS
3 Hours Ago
In-Office
137K-239K Annually
Senior level
137K-239K Annually
Senior level
Aerospace • Information Technology • Software • Cybersecurity • Design • Defense • Manufacturing
Leads flight dynamics architecture, requirements, design, integration, and mission systems engineering for space programs. Develops and evaluates astrodynamics and guidance, navigation, and control solutions; supports space operations, ground architecture, technical roadmaps, risk mitigation, and technology demonstrations. The role requires onsite work, an active U.S. Top Secret clearance, U.S. citizenship, and at least nine years of related experience.
Top Skills: AstrodynamicsConfiguration Management ToolsGuidance Navigation And Control (Gnc)Model-Based Systems Engineering (Mbse)Orbital Analysis SoftwarePythonScaled Agile Framework (Safe)Systems Engineering ToolsetsTcl/Tk

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account