Claryo, Inc. Logo

Claryo, Inc.

Senior Software Engineer - ML Infrastructure

Posted 7 Hours Ago
Be an Early Applicant
In-Office
San Francisco, CA, USA
170K-190K Annually
Senior level
In-Office
San Francisco, CA, USA
170K-190K Annually
Senior level
Design, build, and operate large-scale ML infrastructure and GPU compute clusters for computer vision and multi-modal models. Own end-to-end pipelines from data ingestion and training to low-latency cloud deployment, orchestration, monitoring, and model performance evaluation. Collaborate with research and product teams to productionize models across warehouse environments.
The summary above was generated by AI

We're looking for a Senior Software Engineer - ML Infrastructure to build and scale the infrastructure that powers our AI-driven warehouse intelligence platform. You'll own the end-to-end lifecycle of computer vision models — from training pipelines through optimized cloud deployment — ensuring our cutting-edge computer vision and multi-modal AI systems run reliably and efficiently in production. Your work will directly enable the real-time perception and autonomous decision-making capabilities at the core of our platform.

This is a deeply technical role at the intersection of machine learning, distributed systems, and cloud infrastructure. You'll design scalable GPU compute clusters, build robust orchestration pipelines, and optimize model serving for low-latency inference at scale. You'll work closely with our research scientists, computer vision engineers, and product teams to bridge the gap between experimental models and production-ready systems that operate across diverse warehouse environments. We've found tremendous value in collaborative problem-solving, thus our team works from our SF office three days a week.

Responsibilities

  • Develop and maintain distributed cloud GPU infrastructure for large-scale world model training and low-latency inference.

  • Build end-to-end computer vision pipelines — from data ingestion and preprocessing through model training, evaluation, and deployment — and integrate them into core product workflows.

  • Deploy and optimize state-of-the-art machine learning models in the cloud using model serving platforms and inference optimization techniques, including VLMs and VLAs.

  • Design and operate orchestration systems that enable both engineers and non-engineers to build and manage data and ML pipelines.

  • Establish monitoring, benchmarking, and evaluation frameworks to ensure model performance and reliability in production environments.

Required Experience

  • B.S. / M.S. in Computer Science, Robotics, or similar technical field, or equivalent practical experience.

  • 7+ years of professional software engineering experience, with at least 3 years in machine learning infrastructure — developing, scaling, training, deploying, and optimizing large-scale ML systems from data to model.

  • Track record of deploying machine learning models in production environments with real-world constraints.

  • Experience with distributed messaging and compute systems (Kafka, gRPC, ROS2, or similar).

  • Strong programming skills in Python with solid software engineering practices.

Preferred Experience

  • Experience with training and/or deployment of machine learning models in the computer vision domain.

  • Experience developing, running, and managing orchestration systems (Flyte, Temporal, Airflow, or similar) for ML and data pipelines.

  • Proficiency with ML frameworks (PyTorch, TensorFlow, DeepSpeed) and model serving platforms (TorchServe, TensorFlow Serving, NVIDIA Triton Inference Server, or similar).

  • Deep understanding of state-of-the-art machine learning models such as auto-regressive transformers and familiarity with inference optimization techniques (TensorRT, quantization, custom kernels).

  • Experience with C++ or CUDA programming for GPU acceleration.

  • Prior experience working at autonomous vehicles or robotics companies.

Equal Opportunity Statement

We’re an equal opportunity employer that values diversity and inclusion. We welcome teammates of all backgrounds and don’t discriminate based on race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or veteran status.

Benefits

At Claryo, we offer a competitive benefits package that supports your health and well-being, including — top-tier medical, dental, and vision coverage, 401k with employer matching, parental leave, and unlimited vacation.

Similar Jobs

21 Days Ago
In-Office
San Francisco, CA, USA
176K-220K Annually
Senior level
176K-220K Annually
Senior level
Edtech • Enterprise Web • HR Tech • Software
Build and operate shared ML/AI infrastructure: data pipelines, feature stores, training, model serving and LLM platform work. Scale inference and training (GPU, batching, autoscaling), enable evaluation and post-training workflows, and partner with AI, data science, and product teams to productionize models and improve platform reliability and developer experience.
Top Skills: AirflowApache BeamAutoscalingAWSBatchingBigQueryCi/CdDataflowDockerEmbeddingsFeature StoresGCPGoGpu InferenceKubernetesLlmsModel ServingPythonSparkStreaming PipelinesTerraformTypescript
10 Days Ago
In-Office
Sunnyvale, CA, USA
153K-222K Annually
Senior level
153K-222K Annually
Senior level
Hardware • Industrial
Build and operate end-to-end ML infrastructure: distributed cloud GPU training, datasets, training frameworks, evaluation, deployment, and integrate pipelines into product workflows while collaborating with modeling teams.
Top Skills: Apache AirflowCloud GpuDistributed Gpu TrainingFlyteMl PipelinesNvidia TritonPyTorchTensorFlowTensorflow ServingTorchserve
19 Days Ago
In-Office
Mountain View, CA, USA
194K-291K Annually
Senior level
194K-291K Annually
Senior level
Artificial Intelligence • Automotive • Information Technology • Robotics
Build and evolve ML infrastructure platform for resource provisioning, workload scheduling, petabyte-scale ETL, feature management, and platform abstractions to support large-scale model development and distributed training.
Top Skills: Apache BeamSparkAWSAzureCephCrossplaneFeastGCPHopsworksKuberayKubernetesLustreNvmePulumiRayRedisSlurmTerraformVolcano

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account