NVIDIA Logo

NVIDIA

Senior Cloud Software Engineer, DGXC Data Services

Reposted 4 Days Ago
In-Office or Remote
Hiring Remotely in Santa Clara, CA, USA
152K-288K Annually
Senior level
In-Office or Remote
Hiring Remotely in Santa Clara, CA, USA
152K-288K Annually
Senior level
Design and implement cloud-native data management services (catalog, metadata, dataset/checkpoint lifecycle) for exabyte-scale GPU workloads. Build backend systems on Kubernetes and major cloud providers, collaborate with research and cross-functional teams, document architecture, and drive integration with storage/compute innovations (GPU Direct Storage, DPU).
The summary above was generated by AI

The NVIDIA DGXC Data Services team builds cloud-native systems, frameworks, and services for managing data across hybrid and multi-cloud infrastructure. We are building the next-generation data and storage infrastructure to solve some of the hardest problems in AI: storage, access, ingestion, governance, observability, and data management for exabyte-scale, high-performance GPU-based training and inference jobs. Our work gives NVIDIA teams the foundational capabilities they need to build, train, deploy, and operate AI products at scale without reinventing critical data infrastructure for every workload.

What you will be doing:

  • Build cloud-native data and storage services for hybrid and multi-cloud infrastructure, including dataset discovery, ingestion, governance, checkpointing, observability, and low-latency access.

  • Develop scalable cloud-native services and APIs that support exabyte-scale, high-performance GPU training and inference workflows.

  • Work closely with product managers, internal AI teams, platform teams, and partner engineering teams to understand requirements and turn them into reliable production systems.

  • Collaborate with SRE, operations, and support teams to improve service reliability, performance, observability, on-call readiness, and operational scale.

  • Use modern software engineering practices, including AI-assisted and agentic development workflows, while maintaining high standards for design, testing, security, and verification.

What we need to see:

  • BS in Computer Science, Information Systems, Computer Engineering, or equivalent experience, with 5+ years of software engineering experience.

  • Strong foundation in algorithms, data structures, distributed systems, and practical software design.

  • Experience building, shipping, and operating backend or cloud-native services using Kubernetes, cloud providers such as AWS, GCP, or Azure, and languages such as Go, Python, Rust, C/C++, or Java.

  • Ability to design APIs, document systems, reason through tradeoffs, communicate clearly, and break ambiguous problems into practical execution plans.

  • Experience working across engineering, product, platform, and operations teams to deliver reliable production software.

  • Curiosity and practical judgment around AI-assisted or agentic engineering workflows, including using clear intent, specifications, acceptance criteria, tests, and verification to guide development.

Ways to stand out from the crowd:

  • Hands-on experience building, scaling, or operating large-scale data, storage, or ML infrastructure services.

  • Experience solving enterprise-grade data management, governance, analytics, or AI workflow problems with modern data and ML infrastructure technologies.

  • Strong background in distributed systems, storage systems, cloud infrastructure, performance engineering, observability, or agentic engineering practices.

NVIDIA is leading the way in groundbreaking developments in Artificial Intelligence, High-Performance Computing, and Visualization. Our invention, the GPU, serves as the visual cortex of modern computers and is at the heart of our products and services. Our work opens up new universes to explore, enables amazing creativity and discovery, and powers what were once science fiction inventions, from artificial intelligence to autonomous cars. NVIDIA is looking for great people like you to help us accelerate the next wave of artificial intelligence.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 152,000 USD - 241,500 USD for Level 3, and 184,000 USD - 287,500 USD for Level 4.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until July 26, 2026.

This posting is for an existing vacancy. 

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

HQ

NVIDIA Santa Clara, California, USA Office

2701 San Tomas Expressway, Santa Clara, CA, United States, Santa Clara

NVIDIA San Francisco, California, USA Office

San Francisco, United States

NVIDIA San Jose, California, USA Office

San Jose, United States

Similar Jobs

36 Minutes Ago
Remote or Hybrid
110K-140K Annually
Senior level
110K-140K Annually
Senior level
AdTech • Cloud • Digital Media • Information Technology • News + Entertainment • App development
Lead ownership and evolution of NBCUniversal's CLM platform (Malbek). Configure workflows, manage roadmap and vendor relationships, gather requirements, design solutions, drive CLM expansion, implement AI-enabled contracting features, support testing and change adoption, and measure adoption and efficiency through KPIs and reporting.
Top Skills: AgiloftAi-Enabled Contracting ToolsAribaClmIcertisIntegrationsIroncladMalbekMetadata ManagementReportingSemantic SearchSirionWorkflow Automation
2 Hours Ago
Remote or Hybrid
United States
20-30 Hourly
Mid level
20-30 Hourly
Mid level
Artificial Intelligence • Automotive • Greentech • Information Technology • Machine Learning • Software • Cybersecurity
Perform preliminary website compliance audits and independently manage monthly OEM compliance submissions to third-party auditors. Troubleshoot issues with agencies, communicate remediation steps to clients and internal teams, log and track compliance cases, collaborate with cross-functional teams, and develop brand-specific subject-matter expertise.
2 Hours Ago
Remote
United States
205K-255K Annually
Senior level
205K-255K Annually
Senior level
Software • Defense
Design, build, and run secure production infrastructure across cloud-native and air-gapped deployments. Own end-to-end platform outcomes, harden the artifact pipeline for signed releases, embed with teams to advise on security and deployment best practices, and partner with GRC on STIGs, CVE remediation, and audit readiness for classified environments.
Top Skills: Air-Gapped DeploymentsAmazon AwsApplication StreamingArtifact Pipeline (Ci/Cd)Aws GovcloudAzure GovernmentContainer/Image SigningCve RemediationFedrampGoJwicsKubernetesAzureNetwork SegmentationPythonRustSecrets ManagementSoc 2StigTypescript

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account