Apptronik Logo

Apptronik

Senior Software Engineer, ML Infrastructure

Posted 19 Days Ago
Easy Apply
Hybrid
Austin, TX
Senior level
Easy Apply
Hybrid
Austin, TX
Senior level
Build production ML infrastructure for humanoid robotics, including multimodal data curation, annotation, versioning, large-scale pipelines, simulation and evaluation harnesses, model registries, deployment promotion, experiment tracking, and developer tooling. Partner with autonomy, data platform, and teleoperation teams on lifecycle contracts and technical direction while mentoring engineers through code and design reviews.
The summary above was generated by AI

Apptronik is a human-centered robotics company developing AI-powered robots to support humanity in every facet of life. Our flagship humanoid robot, Apollo, is built to collaborate thoughtfully with people, starting with critical industries such as manufacturing and logistics, with future applications in healthcare, the home, and beyond.
We operate at the cutting edge of Applied AI, applying our expertise across the full robotics stack to solve some of society's most important problems. You will join a team dedicated to bringing Apollo to market at scale, tackling the complex challenges like safety, commercialization, and mass production to change the world for the better.

JOB SUMMARY

Apptronik is building Apollo, a general-purpose humanoid robot, and the physical AI that drives it. Scale is the name of the game: every robot and teleoperator we field produces synchronized video, proprioceptive, tactile, and force-torque streams, and the fleet's output grows with every deployment. Turning that volume of data into shipped autonomy — routinely, at multi-terabyte scale — is what this role is about.

We are looking for a Senior Software Engineer, ML Infrastructure to build that platform: the self-serve services and pipelines that carry data from collection through curation, training, and evaluation to a qualified model running on real hardware. Much of it is being created ground-up for the long term — humanoid robotics has few off-the-shelf answers — so the team builds first-party platform services alongside the open-source and commercial tooling we adopt where it genuinely fits.

This is a hands-on role on a small team whose platform is depended on daily by researchers and engineers across MLOps, Autonomy, Data Platform, and TeleOp.

ESSENTIAL DUTIES AND RESPONSIBILITIES

You will build the ML platform — the APIs, workers, and control planes that let researchers and robot teams move data and models through the system in a self-serve manner, with the testing and observability that being a dependency implies. The platform's responsibilities include:

  • Data Curation & Annotation: Turn raw robot and simulation data into training-ready datasets — selection and filtering of manipulation episodes with synchronized sensor streams; annotation workflows that combine automatic labeling with human-in-the-loop review at throughput; and dataset versioning and lineage strong enough that any model traces back to the exact data that produced it.
  • Data Pipelines at Scale: Make multi-terabyte dataset operations routine — transformation and assembly, coverage and quality statistics that tell us a training set is good before we spend a cluster-week on it, and read paths that keep GPUs fed.
  • Simulation & Evaluation: Build the rollout harnesses that evaluate policies in simulation on our GPU cluster; the benchmarks and metrics captured consistently across simulation, real-robot, and teleoperation sources; and the qualification gates a model must pass before it reaches Apollo — automatic, not manual review.
  • Model Promotion: Build the model store — versioning, metadata, attached evaluation results, lineage — and the promotion path from trained to qualified to deployed on robot, including packaging (ONNX, TensorRT) in partnership with Autonomy.
  • Developer Experience: Provide the tooling researchers use daily — experiment tracking, training job submission, sweeps, and reproducible container environments. Reduce time from idea to running training job; win adoption by being the fastest path, not by mandate.

Alongside the technical work, you will partner with Autonomy, Data Platform, and TeleOp on dataset and model lifecycle contracts, contribute to the technical direction of these layers, and mentor the engineers around you through code and design review.

SKILLS AND REQUIREMENTS

No single person will have depth in everything below. We are looking for someone who has built platform services in production at scale with real depth in at least one of three areas — large-scale data pipelines, annotation and labeling, or evaluation and simulation — plus solid cloud and Python across the board:

  • A builder at scale: a track record of designing and shipping production systems and services that other teams depend on daily.
  • Deep hands-on experience with large-scale data pipelines for ML: multi-terabyte transformation and dataset assembly of multimodal sensor data — video and image streams, time-synchronized robot telemetry, the kind of data that trains vision-language-action and computer-vision models — with columnar and time-series formats (Parquet, Arrow), dataset versioning and lineage (lakeFS, DVC, Iceberg, or equivalent), and object storage (S3, MinIO).
  • Experience with ML annotation and labeling at scale: automatic annotation of data combined with human-in-the-loop workflows — the tooling, quality control, and throughput management.
  • Experience building large-scale evaluation or simulation harnesses: many parallel jobs on GPU infrastructure, aggregated into decision-grade results.
  • Strong Python and general software engineering ability (testing, API design, code review), plus cloud infrastructure, Kubernetes, Docker, and modern CI/CD.
EDUCATION and/or EXPERIENCE
  • 5+ years of professional software engineering experience in ML platforms, data infrastructure, or related fields, OR 3+ years of direct, hands-on experience owning the data and evaluation infrastructure behind models shipped to production.
  • Bachelor's or Master's degree in Computer Science, Machine Learning, or a related technical field, or equivalent experience.

Bonus Qualifications:

  • Robotics data formats and fleet-scale telemetry (MCAP, ROS, LeRobot, or equivalent).
  • Simulation-in-the-loop evaluation with Isaac Sim, IsaacLab, MuJoCo, or equivalent.
  • Reinforcement or imitation learning infrastructure for embodied agents (rollout workers, sim-eval harnesses).
  • Deploying ML models to edge targets (ONNX Runtime, TensorRT, robot fleets).
PHYSICAL REQUIREMENTS
  • Prolonged periods of sitting at a desk and working on a computer
  • Must be able to lift 15 pounds at times
  • Vision to read printed materials and a computer screen
  • Hearing and speech to communicate


*This is a direct hire.  Please, no outside Agency solicitations. 

Apptronik provides equal employment opportunities to all employees and applicants for employment and prohibits discrimination and harassment of any type without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state or local laws.

Similar Jobs at Apptronik

Yesterday
Easy Apply
Hybrid
Easy Apply
Senior level
Senior level
Computer Vision • Hardware • Machine Learning • Robotics • Software
Owns compensation, equity, benefits, compliance, and total rewards strategy for a high-growth robotics company. Leads job architecture, salary structures, compensation cycles, market benchmarking, equity programs, benefits renewals, leave administration, pay equity analysis, dashboards, and governance. Partners with Finance, Legal, People Partners, Talent Acquisition, managers, executives, and the board to support competitive, compliant, and data-informed rewards programs.
Top Skills: Radford
4 Days Ago
Easy Apply
Hybrid
Easy Apply
Expert/Leader
Expert/Leader
Computer Vision • Hardware • Machine Learning • Robotics • Software
Lead the end-to-end architecture and implementation of Apollo humanoid robot’s compute and sensing hardware. Design NVIDIA Jetson carrier boards, high-speed interfaces, power and thermal systems, sensor integration, synchronization pipelines, diagnostics, redundancy, and functional-safety mechanisms. Drive manufacturability, reliability, roadmap development, technical reviews, and mentorship of electrical and embedded engineers.
Top Skills: 10GbeAltium DesignerAnsysAsilDdr4Ddr5Depth CamerasFpgasGigeGmsl2HyperlynxImusIso 26262LidarMcusMipiNvidia JetsonPcb DesignPciePower Delivery NetworksSignal IntegritySilSocsSpiceTactile SensorsThermal ManagementUsb 3.0
8 Days Ago
Easy Apply
Hybrid
Easy Apply
Mid level
Mid level
Computer Vision • Hardware • Machine Learning • Robotics • Software
Develop and deploy actuator models and low-level controls for humanoid robots. Responsibilities include motor and actuator modeling, torque and current control, system identification, hardware characterization, controller tuning, sim-to-real validation, experiment automation, simulator integration, and hardware bring-up. The role requires strong C++ and Python skills, motor electromechanics expertise, classical control knowledge, and hands-on experience with real robotic hardware.
Top Skills: BraxC++DockerGitGpu/Cloud InfrastructureIsaac Sim/LabMujocoPython

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account