Intuitive Logo

Intuitive

Sr MLOps Engineer

Posted 6 Days Ago
Be an Early Applicant
In-Office
Sunnyvale, CA, USA
Senior level
In-Office
Sunnyvale, CA, USA
Senior level
Design, build, and maintain production MLOps infrastructure: Kubernetes clusters, ML orchestration, GPU validation, storage and CI/CD integration, migrations, runbooks, on-call incident response, and security/compliance collaboration to support reproducible ML workflows at scale.
The summary above was generated by AI
Company Description

It started with a simple idea: what if surgery could be less invasive and recovery less painful? Nearly 30 years later, that question still fuels everything we do at Intuitive. As a global leader in robotic-assisted surgery and minimally invasive care, our technologies—like the da Vinci surgical system and Ion—have transformed how care is delivered for millions of patients worldwide.

We’re a team of engineers, clinicians, and innovators united by one purpose: to make surgery smarter, safer, and more human. Every day, our work helps care teams perform with greater precision and patients recover faster, improving outcomes around the world.

The problems we solve demand creativity, rigor, and collaboration. The work is challenging, but deeply meaningful—because every improvement we make has the potential to change a life.

If you’re ready to contribute to something bigger than yourself and help transform the future of healthcare, you’ll find your purpose here.

Job Description

Primary Function of Position
In this role, you will be responsible for designing, building, and maintaining the infrastructure and tools necessary to support the entire machine learning lifecycle, from development to deployment. You will work closely with ML engineers and software developers across Intuitive to ensure that machine learning models are seamlessly integrated into our systems and deliver value at scale. The ideal candidate is an independent and fast-paced engineer with excellent problem-solving skills and practical working knowledge of modern ML development techniques.

Essential Job Duties

  • Bootstrap and maintain a production-grade Kubernetes cluster, including CNI networking and storage integration
  • Deploy and configure ML orchestration tooling (e.g., Metaflow) and artifact/dataset storage solutions to support reproducible ML workflows
  • Validate GPU node health and configuration across heterogeneous hardware (B200, L40S, A6000, V100), including driver/CUDA standardization and topology checks
  • Design and execute team migration playbooks, working directly with engineering teams to port workflows, migrate datasets/artifacts, and roll out tool updates
  • Write and maintain runbooks, architecture documentation, and disaster recovery procedures
  • Participate in on-call rotation and incident response for platform-level issues
  • Collaborate with IT/Security on identity integration, access control, and compliance requirements
  • Continuously evaluate and adopt infrastructure best practices for reliability, cost, and developer experience

Qualifications

Required Skills and Experience

  • 3+ years of experience in infrastructure, DevOps, or MLOps roles, or equivalent practical experience
  • Demonstrated experience operating Kubernetes in production (networking, storage, RBAC, troubleshooting)
  • Strong scripting/automation skills in Python and/or Bash; comfort with Infrastructure-as-Code tools (Ansible, Helm, Terraform, or similar)
  • Hands-on experience with at least one distributed storage system (S3, MinIO, NetApp, or similar)
  • Experience building or maintaining CI/CD pipelines (GitLab CI, ArgoCD, or equivalent)
  • Solid understanding of Linux systems administration and networking fundamentals
  • Excellent communication and documentation skills, with the ability to write clear runbooks and migration guides
  • High degree of autonomy and comfort working across the full stack, iteratively building solutions

Required Education and Training

  • Bachelor’s or Master’s degree in Computer Science, Engineering, or a related field; or equivalent experience

Preferred Skills and Experience

  • Experience with ML orchestration frameworks (Metaflow, MLflow, Kubeflow, or similar)
  • Familiarity with GPU infrastructure (NVIDIA drivers, CUDA, NVLink/NUMA topology, MIG partitioning)
  • Prior experience in a regulated industry (healthcare, finance, or similar) where auditability and access control are critical
  • Experience leading or supporting large-scale infrastructure migrations with multiple stakeholder teams

Additional Information

Due to the nature of our business and the role, please note that Intuitive and/or your customer(s) may require that you show current proof of vaccination against certain diseases including COVID-19.  Details can vary by role.

Intuitive is an Equal Opportunity Employer. We provide equal employment opportunities to all qualified applicants and employees, and prohibit discrimination and harassment of any type, without regard to race, sex, pregnancy, sexual orientation, gender identity, national origin, color, age, religion, protected veteran or disability status, genetic information or any other status protected under federal, state, or local applicable laws.

Mandatory Notices

U.S. Export Controls Disclaimer:  In accordance with the U.S. Export Administration Regulations (15 CFR §743.13(b)), some roles at Intuitive Surgical may be subject to U.S. export controls for prospective employees
who are nationals from countries currently on embargo or sanctions status.

Certain information you provide as part of the application will be used for purposes of determining whether Intuitive Surgical will need to (i) obtain an export license from the U.S. Government on your behalf (note: the government’s licensing process can take 3 to 6+ months) or (ii) implement a Technology Control Plan (“TCP”) (note: typically adds 2 weeks to the hiring process).  

For any Intuitive role subject to export controls, final offers are contingent upon obtaining an approved export license and/or an executed TCP prior to the prospective employee’s
start date, which may or may not be flexible, and within a timeframe that does not unreasonably impede the hiring need. If applicable, candidates will be notified and instructed on any requirements for these purposes. 

We will consider for employment qualified applicants with arrest and conviction records in accordance with fair chance laws.

Preference will be given to qualified candidates who do not reside, or plan to reside, in Alabama, Arkansas, Delaware, Florida, Indiana, Iowa, Louisiana, Maryland, Mississippi, Missouri, Oklahoma, Pennsylvania, South Carolina, or Tennessee.

This position may be filled at a different job level than listed here depending on
business need and/or on the selected candidate’s experience, knowledge and skills.
Compensation will be based primarily on the job level at which the role is filled and the
candidate’s qualifications, consistent with applicable law.

We provide market-competitive compensation packages, inclusive of base pay, incentives, benefits, and equity. It would not be typical for someone to be hired at the top end of range for the role, as actual pay will be determined based on several factors, including experience, skills, and qualifications. The target compensation ranges are listed.

HQ

Intuitive Sunnyvale, California, USA Office

1020 Kifer Road, Sunnyvale, CA, United States, 94086

Similar Jobs

5 Days Ago
In-Office
Santa Clara, CA, USA
184K-357K Annually
Senior level
184K-357K Annually
Senior level
Artificial Intelligence • Computer Vision • Hardware • Robotics • Metaverse
Design, build, and operate end-to-end cloud data and ML pipelines that ingest, validate, process, label, and evaluate multimodal sensor data for autonomous driving. Own architecture, reliability, observability, and operational metrics; collaborate with perception, ML, labeling, and product teams; provide technical leadership, code contributions, and mentorship to deliver AV-scale systems.
Top Skills: 3D GeometryC++CameraCi/CdCloud InfrastructureComputer VisionData PlatformsDeep LearningDistributed SystemsGpu-Accelerated ComputingLidarMlopsPerception PipelinesPythonRadarWorkflow Orchestration
4 Hours Ago
In-Office
Palo Alto, CA, USA
Senior level
Senior level
Artificial Intelligence • Enterprise Web • Software • Generative AI
Own and operate end-to-end ML infrastructure for training, serving, and evaluation of LLM/SLM models. Build scalable low-latency inference (vLLM, batching, autoscaling), multi-GPU training/serving clusters, observability and monitoring, inference-time optimizations (quantization, distillation), reproducibility, versioning, and enterprise-grade auditability. Set MLOps best practices and standards.
Top Skills: AirflowAwqAWSAzureDockerFp8GCPGgufGptqKubeflowKubernetesMlflowMulti-Gpu/Gpu Cluster ManagementPythonRaySglangSparkTerraformTgiTrtVllmWeights & Biases
16 Days Ago
In-Office
San Jose, CA, USA
149K-216K Annually
Senior level
149K-216K Annually
Senior level
Artificial Intelligence • Internet of Things • Machine Learning
Design, build, and operate scalable ML pipelines and infrastructure across cloud and on‑prem HPC. Implement MLOps tooling (tracking, registries, feature stores), CI/CD/CT, containerized GPU orchestration, model optimization (LLMs, GNNs, RL), data/versioning pipelines, monitoring, and mentor engineers to productionize ML for EDA and simulation workloads.
Top Skills: AirflowArizeAutogenAws SagemakerAzure MlBashCadenceCloudFormationDeepspeedDockerDvcElk StackEvidently AiFeastFsdpGcp Vertex AiGoGrafanaHugging FaceJaxKubeflowKubernetesLangchainLsfMlflowPrometheusPythonPyTorchScikit-LearnSiemens EdaSlurmSQLSynopsysTensorFlowTerraformWeights & BiasesXgboost

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account