NVIDIA Logo

NVIDIA

Solutions Architect, Infrastructure

Reposted One Month Ago
Be an Early Applicant
In-Office
Santa Clara, CA, USA
152K-288K Annually
Mid level
In-Office
Santa Clara, CA, USA
152K-288K Annually
Mid level
Lead deployment and scaling of NVIDIA data center GPU and networking platforms for hyperscalers. Drive cross-functional alignment with product and engineering, analyze performance and deployment data, troubleshoot GPU/network/driver/container issues, and enable platform improvements and large‑scale operationalization.
The summary above was generated by AI

Do you thrive on taking a strategic product from launch to go‑to‑market at scale across the world’s largest customers? NVIDIA is looking for an Infrastructure Solutions Architect to lead deployment and bring‑up of our next‑generation Data Center GPUs and networking platforms.

As part of the NVIDIA Solutions Architecture team, we navigate uncharted technical and organizational spaces — serving as the bridge between early platform readiness, cloud engineering teams, product strategy, and large‑scale customer deployments. We are looking for Solution Architects to combine hands‑on infrastructure expertise with multi-functional leadership to accelerate adoption of NVIDIA technologies across worldwide cloud hosting providers and large enterprise environments.

What You’ll Be Doing:

  • Lead end‑to‑end execution for Hyperscaler customers to rapidly bring NVIDIA Data Center GPU and networking platforms to market at scale.

  • Drive strategic partnership and alignment with Product teams to understand roadmap intent, co‑define critical metrics, and ensure unified direction across technical, sales, and leadership organizations.

  • Influence without authority across Product, Engineering, Sales, Operations, and CSP customers, driving clarity, alignment, and unblock paths for scale‑up.

  • Analyze deployment and performance data, identifying product health trends, system bottlenecks, and operational risks.

  • Solve challenging technical problems involving GPUs, networking, drivers, containers, firmware, and distributed system interactions.

  • Deliver streamlined executive‑level communication on status, risks, progress, and required decisions.

  • Collaborate with Product and Engineering, enabling future improvements in platform design, validation, and operational workflows.

What We Need to See:

  • BS/MS/PhD in Electrical/Computer Engineering, Computer Science, Physics, or similar, or equivalent experience.

  • 4+ years experience in Solutions Architecture, Infrastructure Engineering, or similar technical roles.

  • Hands‑on experience with bring‑up and validation of large‑scale NVIDIA GPU platforms, including multi‑GPU and multi‑node architectures.

  • Understanding of high‑performance networking technologies (e.g., RDMA, congestion control, high‑bandwidth interconnects) and their role in distributed AI workloads.

  • Familiarity with NVIDIA system software stacks: CUDA, NCCL, NVSwitch/NVLink, driver behavior, and performance tuning.

  • Proficiency with Linux systems tools for identifying issues and evaluating system performance, such as: dmesg, journalctl, lspci, numactl, ethtool, iostat, perf, nvidia-smi, top/htop, ipmitool, container‑level tooling, and related utilities.

  • Understanding of server hardware architecture, including PCIe topologies, system firmware, NUMA, BIOS/UEFI configuration, power/thermal envelopes, and memory/subsystem behavior.

  • Understanding of BMC/IPMI/Redfish for remote management, hardware health monitoring, and out‑of‑band debugging during early‑stage bring‑up.

  • Strong Linux fundamentals across drivers, kernel subsystems, cgroups, containers, and node‑level performance analysis.

  • Ability to identify performance bottlenecks at the cluster, node, accelerator, network, or application layer.

Ways to Stand Out from the Crowd:

  • Outstanding interpersonal skills and the ability to build clarity and direction across diverse, fast paced technical teams.

  • Knowledge of Compute and networking infrastructure (e.g., Instance types, networking primitives, high‑performance communication paths etc) at Hyperscalers or Cloud Service Providers.

  • Demonstrated leadership resolving multi‑team infrastructure challenges across engineering, product, and customer groups.

  • A consistent record of taking GPU or infrastructure products from pilot to high‑volume deployment in large data center environments.

  • Familiarity with modern deep learning, LLM architectures, and distributed training/inference challenges at scale.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 152,000 USD - 241,500 USD for Level 3, and 184,000 USD - 287,500 USD for Level 4.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until August 30, 2026.

This posting is for an existing vacancy. 

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

HQ

NVIDIA Santa Clara, California, USA Office

2701 San Tomas Expressway, Santa Clara, CA, United States, Santa Clara

NVIDIA San Francisco, California, USA Office

San Francisco, United States

NVIDIA San Jose, California, USA Office

San Jose, United States

Similar Jobs

One Month Ago
In-Office or Remote
4 Locations
184K-288K Annually
Senior level
184K-288K Annually
Senior level
Artificial Intelligence • Computer Vision • Hardware • Robotics • Metaverse
Advise ISVs on designing, deploying, and optimizing large-scale accelerated AI infrastructure. Lead architecture reviews, POCs, benchmarks, and production deployment guidance across compute, networking, storage, containers, orchestration, observability, and CI/CD. Create reference architectures, sizing guidance, technical playbooks, demos, and whitepapers. Support cluster monitoring, reliability, and performance improvements. Up to 20% travel for customer engagements.
Top Skills: Ci/CdContainersDistributed Training FrameworksEthernetGpuHplInfinibandMlperfNcclObservabilityOpenmpiOrchestrationRdmaSchedulingTelemetry
An Hour Ago
Remote or Hybrid
United States
71K-88K Annually
Junior
71K-88K Annually
Junior
Digital Media • Gaming • Information Technology • Software • Sports • Esports • Big Data Analytics
Administer and enhance Salesforce for VIP teams by configuring flows, reports, dashboards, access, and data management. Gather stakeholder requirements, troubleshoot user issues, support adoption, and collaborate with development, analytics, and data engineering teams on integrations and technical solutions. Query and validate data using SQL and Snowflake, test system enhancements, and identify process improvements while maintaining Salesforce data accuracy and integrity.
Top Skills: ApexExcelGoogle SheetsSalesforceSalesforce Flow BuilderSnowflakeSQL
3 Hours Ago
Hybrid
146K-234K Annually
Senior level
146K-234K Annually
Senior level
AdTech • eCommerce • Information Technology • Software • Travel • Generative AI
Design, build, test, and operate scalable software services for Expedia's lodging inventory platform. Own complex components, contribute to architecture and API design, troubleshoot distributed systems, improve reliability through monitoring and automation, and collaborate across engineering and product teams. Apply Kotlin, functional programming, SQL, NoSQL databases, cloud technologies, CI/CD, and AI/ML-enabled solutions to deliver secure, high-availability systems.
Top Skills: Amazon AuroraAmazon DocumentdbAmazon DynamodbAPIsCi/CdJvmKotlinMongoDBNoSQLSQL

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account