Deepmind Logo

Deepmind

Software Engineer, Large Scale Pre-Training Performance

Sorry, this job was removed at 08:08 p.m. (PST) on Tuesday, Jul 22, 2025
Be an Early Applicant
In-Office
Mountain View, CA, USA
In-Office
Mountain View, CA, USA

Similar Jobs

2 Hours Ago
Remote or Hybrid
USA
Senior level
Senior level
Machine Learning • Payments • Security • Software • Financial Services
The Senior Software Engineer designs, develops, tests, and deploys software solutions. They maintain and debug software while ensuring alignment with customer needs and managing business risks.
Top Skills: Application DevelopmentSoftware SolutionsSystem Development Life Cycle
2 Hours Ago
Hybrid
39K-69K Annually
Mid level
39K-69K Annually
Mid level
Machine Learning • Payments • Security • Software • Financial Services
The Technology Solution Center Analyst Lead provides technical support for clients, handles trouble tickets, performs software installations, and assists with performance analysis. Requires strong customer service skills and the ability to work independently.
Top Skills: MS OfficeOutlookServicenow
6 Hours Ago
Remote or Hybrid
California, USA
186K-248K Annually
Expert/Leader
186K-248K Annually
Expert/Leader
AdTech • Digital Media • Marketing Tech
The Principal Technical Program Manager leads the strategic planning and execution of complex technical programs, aligning them with business goals and managing cross-functional teams.
Top Skills: Advertising Technologies
Snapshot
We are seeking a software engineer to define, drive, and critically contribute to the next generation of the state-of-the-art ML models on TPU. As part of the Pre-Training team you will co-design the model, and implement critical components across Model architecture, ML frameworks, custom kernels and platform, to deliver frontier models with maximum efficiency.
 
About Us
 
Artificial Intelligence could be one of humanity’s most useful inventions. At Google DeepMind, we’re a team of scientists, engineers, machine learning experts and more, working together to advance the state of the art in artificial intelligence. We use our technologies for widespread public benefit and scientific discovery, and collaborate with others on critical challenges, ensuring safety and ethics are the highest priority.
 
The Role

We’re looking for a Software Engineer to re-define efficient training of frontier LLMs at massive scale. This role offers an opportunity to influence the design of frontier LLM models, and drive an effort to ensure efficient training and inference.

Key responsibilities:

  • Being responsible for Pre-Training efficiency and optimising the performance of the latest models on Google’s fleet of hardware accelerators - throughout the entire LLM research, training and deployment lifecycle.

  • Being responsible for guiding model design to ensure inference-efficiency.

  • Greatly improving the performance of LLM models on hardware accelerators by optimizing at all levels, including developing custom kernels when necessary.

  • Collaborating with the compiler, framework, and platform teams. And ensure efficient training at industry-largest scale.

  • Profile models to identify performance bottlenecks and opportunities for optimization.

  • Develop low-level custom kernels for maximum performance of the most critical operators.

  • Collaborating with research teams by enabling new critical operators in advance of their availability in frameworks and compilers.

About You

You're an engineer looking to re-define efficient training of frontier LLMs at massive scale and have:

  • A proven track record of critical contributions to the distributed training of LLMs at 1e25 FLOPs scale on modern GPU/TPU clusters  

  • Experience in programming hardware accelerators GPU/TPUs via ML frameworks (e.g. JAX, PyTorch) and low-level programming models (e.g. CUDA, OpenCL)

  • Experience in leveraging custom kernels and compiler infrastructure to improve performance on hardware

  • Experience with Python and neural network training (publications, open-source projects, relevant work experience, etc.)

The US base salary range for this full-time position is between $235,000 - $350,000 + bonus + equity + benefits. Your recruiter can share more about the specific salary range for your targeted location during the hiring process.

Application deadline: July 31st, 2025

Note: In the event your application is successful and an offer of employment is made to you, any offer of employment will be conditional on the results of a background check, performed by a third party acting on our behalf. For more information on how we handle your data, please see our Applicant and Candidate Privacy Policyopen_in_new.

At Google DeepMind, we value diversity of experience, knowledge, backgrounds and perspectives and harness these qualities to create extraordinary impact. We are committed to equal employment opportunity regardless of sex, race, religion or belief, ethnic or national origin, disability, age, citizenship, marital, domestic or civil partnership status, sexual orientation, gender identity, pregnancy, or related condition (including breastfeeding) or any other basis as protected by applicable law. If you have a disability or additional need that requires accommodation, please do not hesitate to let us know.

Deepmind Mountain View, California, USA Office

Ampitheatre Pkwy, Mountain View, CA, United States, 94034

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account