NVIDIA Jobs

Senior AI Performance and Efficiency Engineer

NVIDIA

Senior AI Performance and Efficiency Engineer

Reposted 21 Days Ago

Be an Early Applicant

In-Office or Remote

2 Locations

152K-288K Annually

Senior level

In-Office or Remote

2 Locations

152K-288K Annually

Senior level

Collaborate with AI/ML researchers to identify and fix infrastructure and application inefficiencies on GPU clusters; build tools and ML-based analyzers; monitor fleet utilization; optimize hardware, software, and distributed training/inference performance; and promote adoption of efficient AI/ML practices.

The summary above was generated by AI

We are seeking a Senior AI/ML Performance and Efficiency Engineer, GPU Clusters at NVIDIA to join our AI Efficiency efforts. As an Engineer, you will have a pivotal role in enhancing efficiency for our researchers by implementing progressions throughout the entire stack. Your main task will revolve around collaborating closely with customers to pinpoint and address infrastructure and application deficiencies, facilitating groundbreaking AI and ML research on GPU Clusters. Together, we can craft potent, effective, and scalable solutions as we mold the future of AI/ML technology!

What you will be doing:

Collaborate closely with our AI/ML researchers to make their ML models more efficient leading to significant productivity improvements and cost savings
Build tools, frameworks, and apply ML techniques to detect & analyze efficiency bottlenecks and deliver productivity improvements for our researchers
Work with researchers working on a variety of innovative ML workloads across Robotics, Autonomous vehicles, LLM’s, Videos and more
Collaborate across the engineering organizations to deliver efficiency in our usage of hardware, software, and infrastructure
Proactively monitor fleet wide utilization patterns, analyze existing inefficiency patterns, or discover new patterns, and deliver scalable solutions to solve them
Keep up to date with the most recent developments in AI/ML technologies, frameworks, and successful strategies, and advocate for their integration within the organization.

What we need to see:

BS or similar background in Computer Science or related area (or equivalent experience)
Minimum 5+ years of experience designing and operating large scale compute infrastructure
Strong understanding of modern ML techniques and tools
Experience investigating, and resolving, training & inference performance end to end
Debugging and optimization experience with NSight Systems and NSight Compute
Experience with debugging large-scale distributed training using NCCL
Proficiency in programming & scripting languages such as Python, Go, Bash, as well as familiarity with cloud computing platforms (e.g., AWS, GCP, Azure) in addition to experience with parallel computing frameworks and paradigms.
Dedication to ongoing learning and staying updated on new technologies and innovative methods in the AI/ML infrastructure sector.
Excellent communication and collaboration skills, with the ability to work effectively with teams and individuals of different backgrounds

Ways to stand out from the crowd:

Background with NVIDIA GPUs, CUDA Programming, NCCL and MLPerf benchmarking
Experience with Machine Learning and Deep Learning concepts, algorithms and models
Familiarity with InfiniBand with IBOP and RDMA
Understanding of fast, distributed storage systems like Lustre and GPFS for AI/HPC workloads
Familiarity with deep learning frameworks like PyTorch and TensorFlow

NVIDIA offers competitive salaries and a comprehensive benefits package. Our engineering teams are growing rapidly due to outstanding expansion. If you're a passionate and independent engineer with a love for technology, we want to hear from you.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 152,000 USD - 241,500 USD for Level 3, and 184,000 USD - 287,500 USD for Level 4.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until March 23, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

2701 San Tomas Expressway, Santa Clara, CA, United States, Santa Clara

Similar Jobs

Trail of Bits

Security Engineer

2 Hours Ago

Remote

United States

200K-250K Annually

Expert/Leader

200K-250K Annually

Expert/Leader

Artificial Intelligence • Blockchain • Professional Services • Security • Consulting • Cybersecurity • Defense

The Principal Security Engineer leads projects, drives technical vision, mentors engineers, engages in business development, and oversees security software development.

Top Skills: C++GoJavaPythonRust

CSC

Technical Writer

5 Hours Ago

Remote or Hybrid

70K-84K Annually

Mid level

70K-84K Annually

Mid level

Fintech • Legal Tech • Software • Financial Services • Cybersecurity • Data Privacy

Create, edit, and manage a library of training materials and eLearning assets for tax and licensing software. Collaborate with SMEs to produce guides, courses, videos, SCORM packages, in-app tours, and customer communications, ensuring brand and quality standards across publishing platforms.

Top Skills: Adobe AcrobatAdobe FramemakerAdobe PremiereArticulate Rise 360Articulate StorylineBoxCamtasiaGitHTMLMadcap FlareMarkdownMicrosoft WordNetSuiteScormSharepointSnagitThought Industries LmsXML

Onebrief

Designer

5 Hours Ago

Remote

United States

180K-220K Annually

Mid level

180K-220K Annually

Mid level

Software • Defense

Design and productize AI decision-making systems for a game engine, utilizing HTNs and GOAP to create intelligent behaviors for entities.

Top Skills: Goal-Oriented Action Planning (Goap)Hierarchical Task Networks (Htns)State MachinesVisual Scripting Editor

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
Major Tech Employers: Google, Apple, Salesforce, Meta
Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

NVIDIA

Senior AI Performance and Efficiency Engineer

NVIDIA Santa Clara, California, USA Office

Similar Jobs

Security Engineer

Technical Writer

Designer

What you need to know about the San Francisco Tech Scene

Key Facts About San Francisco Tech