NVIDIA Logo

NVIDIA

Senior Engineering Manager, Infrastructure Security Engineering - DGX Cloud

Posted One Month Ago
Be an Early Applicant
Remote
Hiring Remotely in Canada
245K-295K Annually
Senior level
Remote
Hiring Remotely in Canada
245K-295K Annually
Senior level
Leads a distributed team of security and infrastructure engineers responsible for securing NVIDIA’s large-scale DGX Cloud GPU fleet. Sets technical direction for security control planes and foundational services, oversees hiring and talent development, drives measurable risk reduction, and ensures operationally ready engineering delivery. The role requires cross-functional influence across platform, SRE, networking, and security teams, along with expertise in cloud-native architecture, Kubernetes, identity and access, policy enforcement, and vulnerability management.
The summary above was generated by AI

NVIDIA DGX Cloud is the AI supercomputing-as-a-service substrate designed to power the next generation of AI and industrial-scale breakthroughs. As an Engineering Manager within our Infrastructure Security Engineering organization, you will enable the engineers who secure a GPU fleet numbering in the hundreds of thousands. Your team does not write reports about risk. It engineers whole classes of risk out of existence, and delivers the result as paved roads the rest of DGX Cloud will happily take.

What You Will Be Doing:

  • Grow the Team: Manage, coach, and develop a distributed team of senior security and infrastructure engineers. You own hiring, growth, and performance, and you build a bench deep enough that our capability never rests on one person. Heroics signal a system that needs fixing, so you protect sustainable pace.

  • Set Technical Direction: Partner with your engineers on the roadmap for our security control plane, our foundational security services, and the path toward autonomous security operations. Leadership owns the mission and the strategy; the team owns the execution.

  • Deliver Outcomes, Not Tickets: Run a lightweight planning rhythm where the team commits to a result and the reason it matters, then reflects openly on what to change next. We are measured by the risks we eliminate, not the tickets we close.

  • Assemble Teams Around the Work: Staff each effort with the smallest group who can deliver it, and a clear owner accountable for landing it. Teams form around the work and dissolve when it's done. Make that motion fast, fair, and a real growth opportunity.

  • Raise the Engineering Bar: Hold a high standard for what "done" means: tested code, observability, operational readiness, and documentation someone who didn't build the system can operate from. AI-assisted work ships with real human ownership behind it.

  • Communicate Risk Upward: Translate a complex, fast-moving risk picture into something executives can act on. Declare the unknowns honestly, turn risk assessment into measurable themes, and make trade-offs legible to the people funding them.

  • Multi-Functional Collaboration: Build the partnerships that make the work land. We operate horizontally across vertically organized teams, with influence rather than authority. Your credibility with DGX Cloud platform, SRE, networking, and broader NVIDIA security partners is a delivery mechanism, not a nicety.

What We Need to See:

We are looking for high-caliber engineering leaders with deep spikes of expertise in a few of these areas and the intellectual curiosity to dive into the rest. If your experience aligns with the core of this role—building and growing teams that engineer risk out of existence—and you can show us how, we want to hear from you!

  • Engineering Management: Experience (typically 7+ years) managing engineers, within a broader career (typically 14+ years) across software engineering, SRE, infrastructure, or security. You have grown people, not only shipped projects.

  • Technical Credibility: Enough hands-on depth in infrastructure, distributed systems, or security engineering to earn the trust of senior engineers. You can review a design honestly and recognize when an estimate doesn't add up. You don't need to be the best engineer in the room; you need to build the room.

  • Hiring and Developing Talent: A history of attracting, hiring, and growing excellent engineers, including people far more expert than you in their domain. You build interview and calibration practices that make hiring repeatable and fair.

  • Delivery Under Ambiguity: You can take a high-stakes, loosely defined mandate and turn it into sequenced, measurable delivery. You are comfortable proving a model on a few systems first, then using that success to expand.

  • Cloud-Native and Security Fluency: Working understanding of cloud-native architecture, container orchestration (Kubernetes), identity and access, policy enforcement, and vulnerability management—enough to build it and ask the questions that matter.

  • Influence Without Authority: Experience delivering through partner teams you do not own, in environments where relationships, not org charts, determine whether the work ships.

  • Foundation: Bachelor's degree in Computer Science, Engineering, or a related technical field (or equivalent experience).

Ways To Stand Out from the Crowd:

  • HPC/AI Infrastructure: Experience running teams that secure or operate high-performance computing environments, large GPU fleets, or multi-tenant platforms at cloud scale.

  • Security as an Internal Product: You have delivered security or infrastructure capabilities with real adoption metrics, where teams took the paved road because it was the easiest path.

  • Building a Function from Scratch: Experience standing up a team, a practice, or an engineering culture where none existed—charter, agreements, rituals, and all.

  • Scaling Without Dilution: A history of growing a team quickly while the bar went up rather than down. How have you scaled a team and raised the standard at the same time?

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 245,000 CAD - 295,000 CAD.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until September 13, 2026.

This posting is for an existing vacancy. 

NVIDIA uses AI tools in its recruiting processes.

HQ

NVIDIA Santa Clara, California, USA Office

2701 San Tomas Expressway, Santa Clara, CA, United States, Santa Clara

NVIDIA San Francisco, California, USA Office

San Francisco, United States

NVIDIA San Jose, California, USA Office

San Jose, United States

Similar Jobs

An Hour Ago
Remote or Hybrid
7 Locations
153K-245K Annually
Senior level
153K-245K Annually
Senior level
eCommerce • Fintech • Hardware • Payments • Software • Financial Services
Own Square’s organic social strategy, audience definition, content calendar, community management, creator and influencer programs, social listening, crisis support, and channel performance. Partner with brand, creative, paid media, communications, legal, partnerships, and agencies to connect social activity to brand positioning and perception goals. This individual contributor role spans strategy through publication and requires strong judgment, writing, execution, and cross-functional influence.
Top Skills: Sprinklr
10 Hours Ago
Easy Apply
Remote
United States
Easy Apply
139K-235K Annually
Senior level
139K-235K Annually
Senior level
Cloud • Security • Software • Cybersecurity • Automation
Build GitLab code security capabilities spanning dependency analysis, SBOM and lockfile parsing, vulnerability detection, AI-assisted analyzers, and automated dependency remediation. Design evaluation harnesses, package analyzers for CI and AI workflows, improve security report accuracy, and extend ecosystem coverage. Own initiatives from design through delivery, mentor engineers, collaborate across product and engineering teams, document systems, and participate in on-call support.
Top Skills: AIAstsBundlerCargoCi/CdCweDockerGoLlmMavenNpmOwasp Top 10PipPythonRubyRustSbom
10 Hours Ago
Easy Apply
Remote
United States
Easy Apply
153K-259K Annually
Expert/Leader
153K-259K Annually
Expert/Leader
Cloud • Security • Software • Cybersecurity • Automation
Leads the architecture and delivery of GitLab’s static analysis engine and evaluation tooling. Owns program modeling, vulnerability detection quality, SAST rules, AI-assisted development guardrails, and benchmark validation. Sets technical direction, mentors engineers, coordinates complex initiatives, contributes to research and open source, and participates in on-call rotations. The role requires deep program analysis, application security, systems programming, performance optimization, containerized CI/CD, and experience building trustworthy LLM tooling.
Top Skills: Ai AgentsAstsCall GraphsCi/CdControl-Flow GraphsCweData-Flow AnalysisDockerFuzzingGoIntermediate RepresentationsLlm ToolingMutation TestingOwasp Top 10RubyRustSastSsaStatic AnalysisTaint AnalysisType Inference

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account