NVIDIA Logo

NVIDIA

Senior Software Engineer, Network Visibility Platform - DGX Cloud

Posted Yesterday
Be an Early Applicant
In-Office or Remote
Hiring Remotely in Santa Clara, CA, USA
168K-322K Annually
Senior level
In-Office or Remote
Hiring Remotely in Santa Clara, CA, USA
168K-322K Annually
Senior level
Leads architecture and development of NVIDIA’s network visibility platform, integrating topology, configuration, telemetry, active measurement, change intelligence, and self-service tools. Builds authoritative network models, defines observability and measurement strategies, establishes data contracts and quality standards, and delivers health, path, and blast-radius insights. Drives cross-functional roadmaps, build-versus-buy decisions, operational tradeoffs, and technical mentorship across large-scale infrastructure systems.
The summary above was generated by AI

As a Senior Software Engineer on NVIDIA’s Global Network Visibility (GNV) team within NVIDIA's Global Network Infrastructure (GNI) organization, you will lead the platform that turns network topology, configuration, telemetry, and direct measurement into trusted, self-service insight. Your work will help teams understand network health and customer impact, accelerate incident triage, and bring AI, storage, backbone, and edge infrastructure online with confidence.


You will set direction across authoritative network data, scalable measurement, and customer-facing visibility products. You will define durable interfaces, lead cross-domain decisions, and guide build-versus-buy choices. Success means the platform is credible, adopted, and trusted.


What you'll be doing:

  • Set the architecture for a unified platform spanning topology, configuration, telemetry, active measurement, change intelligence, and self-service experiences
  • Build an authoritative network model connecting physical topology, logical overlays, configurations, and service dependencies across systems of record, and use it to catch physical and logical configuration errors during cluster bring-up
  • Define a telemetry & measurement strategy combining reachability tests, streaming signals, traffic & capacity data, and change & maintenance events
  • Deliver intuitive health, path, blast-radius, and self-service exploration products that help customers answer network questions independently
  • Establish interfaces, data contracts, quality standards, and operating mechanisms that let teams contribute without fragmenting the experience
  • Drive cross-functional roadmaps; make clear tradeoffs among speed, cost, reliability, and maintainability; and mentor engineers through critical designs

What we need to see:

  • Bachelor’s degree or equivalent experience, plus 10+ years of relevant industry experience
  • Record of architecting and operating large-scale software or infrastructure platforms, with strong proficiency in Python, Go, or a comparable systems language
  • Deep expertise in one area—with breadth across the others: network topology and sources of truth, telemetry and active measurement, or observability products
  • Strong data center and backbone networking fundamentals across physical connectivity, routing, overlays, services, and customer-visible failures
  • Hands-on experience across an observability stack such as Prometheus, Grafana, and OpenTelemetry, including active or synthetic monitoring and containerized deployment on Kubernetes
  • Experience with large-scale graph, relational, time-series, streaming, API, or event-driven systems that reconcile multiple sources
  • Technical leadership across teams, with strong product and operational judgment and clear tradeoffs among performance, resiliency, usability, and cost

Ways to stand out from the crowd:

  • Led a network visibility or infrastructure intelligence platform spanning thousands of devices or hosts
  • Experience with AI or HPC networks, including RDMA, RoCE over Spectrum-X, or InfiniBand
  • Experience with network sources of truth such as NetBox or Nautobot, streaming telemetry (gRPC/gNMI), graph databases, or front-end development in TypeScript and React
  • SRE or network operations experience, plus a record of setting standards, making build-versus-buy decisions, or contributing to open source

NVIDIA is widely considered one of the world's most desirable employers in technology. We have some of the world's most forward-thinking and passionate people working for us. If you're creative and autonomous, we want to hear from you!


NVIDIA has been redefining computer graphics, PC gaming, and accelerated computing for over 25 years. It’s a unique legacy of innovation fueled by great technology—and dynamic people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 168,000 USD - 270,250 USD for Level 4, and 200,000 USD - 322,000 USD for Level 5.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until September 27, 2026.

This posting is for an existing vacancy. 

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

HQ

NVIDIA Santa Clara, California, USA Office

2701 San Tomas Expressway, Santa Clara, CA, United States, Santa Clara

NVIDIA San Francisco, California, USA Office

San Francisco, United States

NVIDIA San Jose, California, USA Office

San Jose, United States

Similar Jobs

41 Minutes Ago
Remote
United States
165K-220K Annually
Senior level
165K-220K Annually
Senior level
Artificial Intelligence • Blockchain • Professional Services • Security • Consulting • Cybersecurity • Defense
Lead cryptographic security assessments of protocols, libraries, applications, and high-assurance systems. Analyze advanced cryptographic constructions, identify design and implementation weaknesses, validate exploitability, and recommend remediation. Build testing and analysis tools, manage client delivery, write precise technical reports, mentor engineers, contribute to methodologies and services, and support open-source research and technical publications.
Top Skills: CC++CryptographyGitGoMulti-Party ComputationPost-Quantum CryptographyRustZero-Knowledge Proofs
An Hour Ago
Easy Apply
Remote or Hybrid
US
Easy Apply
146K-240K Annually
Senior level
146K-240K Annually
Senior level
Marketing Tech • Social Media • Software • Analytics • Business Intelligence
Leads and develops a team of Strategic Customer Success Managers responsible for enterprise accounts. Drives retention, expansion, account health, renewals, QBRs, forecasting, and operational consistency. Serves as an executive escalation point, coaches strategic customer engagement, incorporates generative AI into workflows, and partners cross-functionally with Sales, Renewals, Onboarding, Product, and Professional Services. Tracks portfolio performance and supports customer value realization across high-value SaaS accounts.
Top Skills: GainsightGenerative AiSalesforceTableau
An Hour Ago
Remote
United States
105K-168K Annually
Senior level
105K-168K Annually
Senior level
Cloud • Fintech • Food • Information Technology • Software • Hospitality
Lead credit risk strategy for lending products by analyzing portfolio performance, identifying risk drivers, refining underwriting and pricing policies, and conducting A/B tests. Partner with Data Science to develop and monitor credit models, while building AI-enabled workflows that automate analysis and surface risk signals. Collaborate cross-functionally with Risk, Product, Engineering, Operations, and Data Science to support scalable fintech growth and compliant automated credit decisioning.
Top Skills: A/B TestingAIHexLlmsLookerPythonRSQLStatistical ModelingTableau

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account