Nuro Logo

Nuro

Software Engineer, ML Infrastructure Platform

Posted 2 Days Ago
In-Office
Mountain View, CA, USA
160K-241K Annually
Junior
In-Office
Mountain View, CA, USA
160K-241K Annually
Junior
Build and operate large-scale ML training infrastructure: distributed GPU training, multi-cluster scheduling, data pipelines (batch and streaming), ML workflows, observability, reliability, and on-call incident response to keep autonomy model training running efficiently.
The summary above was generated by AI

Who We Are 

Nuro believes self-driving vehicles are the most immediate and profound opportunity for AI to drive positive change in the physical world. Safer streets, more time for what matters, and easier access to the world around us, that’s why we’re building a universal autonomy platform: self-driving for all roads and all rides.
Founded in 2016, Nuro is a physical AI company developing Level 4 autonomous driving technology for a wide range of vehicles, use cases, and markets. Powered by the Nuro Driver™, our universal autonomy platform enables the global mobility ecosystem to deploy autonomy at scale, from robotaxis and logistics fleets to personal vehicles.
With years of real-world deployment experience and a flexible, partner-led business model, Nuro is working toward a future where millions of autonomous vehicles powered by our technology help make everyday life safer, easier, and more connected.
Nuro has raised over $2B in capital from Uber, NVIDIA, Google, Softbank, Fidelity, T. Rowe Price, and other leading investors

About the Role

Nuro takes a machine-learning-first approach to autonomous driving, and the ML Infrastructure team builds and operates the infrastructure that makes that possible. We own the systems that train the models at the core of the Nuro Driver™ - from distributed GPU training and closed-loop reinforcement learning, to the workflows, orchestration, observability, and cost management that keep the fleet running efficiently.

Our work sits directly on the critical path of autonomy development. When a training run stalls, when a pipeline silently regresses, or when GPU utilization slips, it shows up in how fast the rest of the company can ship. We care as much about reliability and operational maturity as we do about raw scale.

About the Work

  • Contribute to Nuro’s training infrastructure, spanning multi-generation accelerators, and multi-cluster scheduling and orchestration.
  • Design and operate large-scale data pipelines - batch and streaming ingestion, storage layout, and high-throughput data generation and storage.
  • Design and develop agentic-first ML workflows - data-to-training-to-evaluation pipelines that are introspectable, reproducible, and easy for autonomy teams to run and extend.
  • Own reliability for critical training and release pipelines: instrument them, define meaningful alerting, and build the on-call and incident-response practices that let the team catch regressions.

About You

  • BS, MS, or PhD in Computer Science, Electrical Engineering, or a closely related field, plus 1+ years of relevant work experience.
  • Willingness to deep-dive into implementation and to raise the technical and operational standards of the broader engineering organization.
  • A demonstrated ownership mindset: you drive systems to operational maturity e.g. through monitoring, alerting, runbooks.
  • Strong proficiency in Python (and comfort with C++, Go or a similar systems language).
  • Hands-on experience running production infrastructure on Kubernetes.
  • Solid distributed-systems fundamentals and the ability to reason about performance, failure modes, and reliability across a complex system.

Bonus Points

  • Strong working knowledge of GCP.
  • Experience with building large scale data generation pipelines.
  • Experience with Kubernetes-native orchestration for ML workloads.
  • Depth in GPU / distributed training internals, including NCCL and collective communication.
  • Familiarity with GPU and training observability tooling and using it to diagnose real bottlenecks.
  • A track record of driving down infrastructure cost while improving reliability.

At Nuro, your base pay is one part of your total compensation package. For this position, the reasonably expected base pay range is between $160,360 and $240,540 for the level at which this job has been scoped. Your base pay will depend on several factors, including your experience, qualifications, education, location, and skills. In the event that you are considered for a different level, a higher or lower pay range would apply. This position is also eligible for an annual performance bonus, equity, and a competitive benefits package.

At Nuro, we celebrate differences and are committed to a diverse workplace that fosters inclusion and psychological safety for all employees. Nuro is proud to be an equal opportunity employer and expressly prohibits any form of workplace discrimination based on race, color, religion, gender, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other legally protected characteristics.

HQ

Nuro Mountain View, California, USA Office

Mountain View, CA, United States

Nuro San Francisco, California, USA Office

San Francisco, United States

Similar Jobs

2 Days Ago
In-Office
Mountain View, CA, USA
194K-291K Annually
Senior level
194K-291K Annually
Senior level
Artificial Intelligence • Automotive • Information Technology • Robotics
Build and operate large-scale ML training infrastructure: distributed GPU training, multi-cluster scheduling, data pipelines (batch & streaming), ML workflows, observability, alerting, and on-call/incident response to ensure reliable, cost-effective model training and releases.
Top Skills: Batch Data PipelinesC++Distributed Gpu TrainingGCPGoGpusKubernetesMulti-Cluster SchedulingNcclObservability/MonitoringPythonReinforcement LearningStreaming Data Pipelines
11 Minutes Ago
Remote or Hybrid
Santa Clara, CA, USA
229K-412K Annually
Senior level
229K-412K Annually
Senior level
Artificial Intelligence • Cloud • HR Tech • Information Technology • Productivity • Software • Automation
Lead delivery transformation for AMS by scaling partner ecosystems, embedding AI-driven delivery tools, managing customer health and risk, and driving time-to-value reduction. Oversee partner delivery, AI adoption, solution architecture for complex deals, and delivery quality governance while mentoring cross-functional teams and collaborating with Sales, A&M, and global partners to improve CSAT and operational metrics.
Top Skills: AIAuctorAxis AgentsConfig AgentEngage CentralIntelligent AgentsServicenowWorkflow Automation
Expert/Leader
Financial Services
Lead regional consumer banking operations to grow deposits and banking business, coach Market Directors, drive financial metrics, integrate cross-functional partners, ensure compliance and strong customer experience.

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account