NVIDIA Logo

NVIDIA

Senior Manager, DGX Cloud Technical Program Management

Reposted Yesterday
Be an Early Applicant
In-Office
Santa Clara, CA, USA
240K-380K Annually
Senior level
In-Office
Santa Clara, CA, USA
240K-380K Annually
Senior level
Lead and manage a team of Technical Program Managers to deliver DGX Cloud core infrastructure programs (network, storage, security, trust services, telemetry, break/fix). Drive cross-functional alignment with engineering, product, operations, and cloud providers; establish operating rhythms, risk tracking, metrics and dashboards; coordinate operational readiness and incident response; and standardize TPM practices to improve reliability, scalability, and customer readiness.
The summary above was generated by AI

For over 25 years, NVIDIA has led the world in visual computing and accelerated computing. Today, we’re crafting the future of AI by driving breakthroughs in generative models, autonomous systems, and large-scale research. The DGX Cloud organization builds and operates the AI infrastructure that makes this innovation possible. We are seeking a Technical Program Management Manager to lead core infrastructure programs across DGX Cloud, including network, storage, trust services, security, break/fix operations, and telemetry. This role manages a team of TPMs responsible for bringing structure, operational rigor, and cross-functional alignment to infrastructure programs that keep DGX Cloud resilient, scalable, and customer-ready.

What you’ll be doing:

  • Lead and nurture a team of Technical Program Managers engaged in DGX Cloud core infrastructure projects.

  • Propel progress across network, storage, trust services, security programs, telemetry, and break/fix operational workstreams.

  • Partner with engineering, product, operations, security, and cloud provider teams to define priorities, achievements, dependencies, and delivery plans.

  • Build clear operating rhythms for infrastructure planning, managing blocking issues, risk tracking, and cross-functional decision-making.

  • Improve access to infrastructure health, delivery status, blockers, and program risks through practical metrics, dashboards, and reporting.

  • Coordinate break/fix and operational readiness programs that improve reliability, response time, and customer impact management.

  • Support continuous improvement across TPM practices, helping the team standardize planning, execution, and communication across DGX Cloud infrastructure.

What we need to see:

  • More than 12 overall years in technical program management, infrastructure program management, or similar roles, including upwards of 3 years directing or supervising TPMs.

  • Experience managing infrastructure programs in domains such as networking, storage, security, trust services, observability, telemetry, or cloud operations.

  • Strong ability to manage priorities, dependencies, risks, and execution plans across multiple engineering teams.

  • Experience building TPM operating rhythms, including status reviews, paths for handling blocking issues, tracking critical achievements, and leadership-ready updates.

  • Working knowledge of cloud infrastructure, distributed systems, or large-scale platform operations.

  • Strong communication skills with the ability to translate complex infrastructure work into clear program status, risks, and decisions.

  • Bachelor’s or Master’s degree in Computer Science, Engineering, or related field, or equivalent experience.

Ways to stand out from the crowd:

  • Experience supporting infrastructure for AI/ML platforms, GPU clusters, or large-scale cloud services.

  • Background with observability and telemetry tools such as Grafana, Prometheus, or similar platforms.

  • Experience with security, trust, compliance, or reliability programs in cloud infrastructure environments.

  • Track record improving operational processes for break/fix, incident response, or infrastructure readiness.

  • Strong technical judgment and ability to partner closely with engineering leaders while developing TPM talent.

Join NVIDIA and help us build the future of AI infrastructure with your expertise and passion!

#Li-Hybrid

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 240,000 USD - 379,500 USD.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until August 9, 2026.

This posting is for an existing vacancy. 

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

HQ

NVIDIA Santa Clara, California, USA Office

2701 San Tomas Expressway, Santa Clara, CA, United States, Santa Clara

NVIDIA San Francisco, California, USA Office

San Francisco, United States

NVIDIA San Jose, California, USA Office

San Jose, United States

Similar Jobs

13 Hours Ago
In-Office
201K-251K Annually
Senior level
201K-251K Annually
Senior level
Artificial Intelligence • Cloud • Software • Infrastructure as a Service (IaaS)
Lead and mentor a multidisciplinary engineering team to build and deliver scalable security products (IAM, threat detection, data protection). Drive technical roadmap, manage delivery timelines and inter-team dependencies, prioritize technical debt, ensure code quality and security best practices, collaborate with product and security architects, and handle recruitment and career development.
Top Skills: Apache FlinkAWSAzureChefCryptographyGCPGoHelmIds/IpsJavaScriptKafkaKubernetesNode.jsSIEMTerraformVulnerability ScannersWaf
13 Hours Ago
In-Office
159K-199K Annually
Senior level
159K-199K Annually
Senior level
Artificial Intelligence • Cloud • Software • Infrastructure as a Service (IaaS)
Manage the execution of Security Products by translating customer needs into requirements, collaborating with engineering and UX, and ensuring consistent communication for product launches and incidents.
Top Skills: CloudIamSaaSSecurity
13 Hours Ago
Remote or Hybrid
United States
Entry level
Entry level
Fintech • Machine Learning • Software • Financial Services
This position is an opportunity to fill a form for future job matching after the ASPLOS Conference with IMC Trading.

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account