Atomic Machines Logo

Atomic Machines

MLOps Engineer

Posted 19 Days Ago
Be an Early Applicant
In-Office
Emeryville, CA, USA
200K-250K Annually
Senior level
In-Office
Emeryville, CA, USA
200K-250K Annually
Senior level
Build and operate production MLOps infrastructure spanning experimentation, model training, deployment, serving, monitoring, retraining, and rollback. Develop ML data and feature pipelines using multimodal manufacturing data, maintain lakehouse and feature-store capabilities, and establish observability and human-in-the-loop feedback systems. Collaborate with data, AI, software, process, materials, and simulation engineers to deliver reliable, scalable ML platforms and model feedback loops.
The summary above was generated by AI
Atomic Machines is ushering in a new era of micromanufacturing with its Matter Compiler™ technology platform. This platform enables new classes of micromachines to be designed and built by providing manufacturing processes and a materials library that are inaccessible to semiconductor manufacturing methods. It unlocks MEMS manufacturing not only for device classes that could never be produced by semiconductor methods, but also for entirely new categories. Furthermore, this digital platform is fully programmable in the way 3D printing is digital—but whereas 3D printing produces parts of a single material using a single process, the Matter Compiler™ technology platform is a multi-process, multi-material system: bits and raw materials go in, and complete, functional micromachines come out. The Atomic Machines team has also created an exciting first device—made possible only through the Matter Compiler™ technology platform—that we will be unveiling to the world soon.
 
Our offices are in Emeryville and Santa Clara, California.
About The Role:

We are seeking an MLOps Engineer to join our AI and Modeling & Simulation org within the Data Engineering and Analytics team.

You will build and operate the infrastructure that takes AI and machine learning models from experimentation to reliable production - covering training, deployment, serving, monitoring, and continuous improvement. This is a DevOps-leaning MLOps role centered on the model feedback loop: connecting production signals and expert feedback back to training so models improve as the system operates.

We are looking for senior-level candidates who can take meaningful ownership of production ML infrastructure. The scope and seniority of the role will be shaped by the candidate's experience, technical depth, and demonstrated impact.

You will work closely with Data, AI, Process, Design, and Software engineers in a highly cross-functional environment.

What You’ll Do:
  • Build and evolve the MLOps platform and CI/CD: Own the path from experiment to production, including experiment tracking, model registry, packaging, automated training and retraining, deployment, and safe rollout and rollback.
  • Operate model serving infrastructure: Build reliable, scalable batch, streaming, and real-time inference for models and digital twins supporting design, process control, scheduling, and inspection.
  • Build ML data and feature pipelines: Turn machine telemetry, process and knowledge graphs, images, time-series, agentic conversations, and other production data into contextualized, model-ready datasets and features.
  • Maintain ML data infrastructure: Support feature-store capabilities and a lakehouse foundation using Apache Iceberg on S3, with strong data quality, lineage, versioning, and reproducibility.
  • Close the model feedback loop: Build model observability and human-in-the-loop systems that capture production signals and expert corrections, version them as ground truth, and feed them into evaluation and retraining workflows.
  • Create paved roads for ML development: Develop standardized tooling and workflows that enable Data and AI engineers to move quickly while maintaining production reliability and reproducibility.
  • Drive technical ownership: Identify infrastructure, reliability, and scalability challenges and drive solutions from design through production. More senior candidates will have opportunities to shape architecture, technical direction, and engineering practices across the ML platform.
  • Collaborate across disciplines: Work with Process, Chemical, Materials, Simulation, Software, Data, and AI engineers to define deployment, serving, and data-collection requirements.
What You’ll Need:
  • 5+ years of relevant industry experience building production software, infrastructure, data, or machine learning systems. We value demonstrated technical depth, ownership, and impact over a specific number of years.
  • Proven experience building and operating machine learning systems in production, with a strong MLOps/DevOps orientation.
  • Strong DevOps fundamentals, including CI/CD, containers, Kubernetes, cloud infrastructure, and infrastructure-as-code.
  • Proficiency in Python and SQL.
  • Hands-on experience with MLflow or similar tooling for experiment tracking, model registry, and model lifecycle management.
  • Experience with S3, lakehouse technologies such as Apache Iceberg, and workflow orchestration tools such as Airflow or Dagster.
  • Experience building pipelines for multimodal ML data, including images, time-series, structured, and semi-structured data.
  • Familiarity with manufacturing systems, sensors, process automation, or other physical-world data systems.
  • Strong problem-solving skills, attention to data quality and reliability, and clear technical communication.
  • Bachelor's or Master's degree in Computer Science, Data Engineering, Data Science, or a related STEM field, or equivalent practical experience.
  • This role is open across multiple levels, from early in career though Staff (L4 through L6). We'll determine the appropriate level and compensation based on your experience, skills, and the scope of the role through the interview process.
Bonus Points For:
  • Experience with feature stores, human-in-the-loop systems, active learning, or data-labeling infrastructure.
  • Robotics or robotic automation experience, including sensors, vision systems, or robotics data.
  • Experience operating ML systems in manufacturing or other physical-world environments.
  • Experience building internal tools for expert feedback, labeling, model evaluation, or model interaction.
  • Experience designing shared ML infrastructure or platforms used across multiple teams or applications.

The compensation for this position also includes equity and benefits.

Salary Range
$200,000$250,000 USD
HQ

Atomic Machines Berkeley, California, USA Office

950 Gilman Street , Suite 800, , Berkeley, CA, United States, 94710

Similar Jobs

2 Days Ago
In-Office or Remote
California, USA
152K-230K Annually
Senior level
152K-230K Annually
Senior level
Software • Travel
Build and maintain internal platforms and tooling for AI engineers, including React and Next.js dashboards, Python backend services, REST APIs, model deployment automation, observability integrations, and ML pipeline testing. Partner with engineering and ML teams to improve developer experience, production AI reliability, and self-service workflows while documenting systems and runbooks.
Top Skills: AWSAzureDatabasesDatadogDockerGCPGrafanaInfrastructure As Code (Iac)Job Queuing SystemsKubernetesLangsmithNext.JsPrometheusPythonReactRest ApisTypescript
6 Days Ago
Hybrid
106K-189K Annually
Mid level
106K-189K Annually
Mid level
Insurance
Designs, deploys, and operates scalable machine learning infrastructure and pipelines on AWS. Builds Python-based automation, model deployment frameworks, CI/CD workflows, and infrastructure-as-code solutions. Monitors reliability, security, observability, and cost efficiency of production ML systems while troubleshooting complex issues. Collaborates with data scientists and engineering teams, contributes to architectural standards, conducts code reviews, and mentors junior engineers.
Top Skills: Amazon EcsAmazon EksAmazon S3Amazon SagemakerAWSAws CdkAws CloudformationAws IamAws LambdaAws Step FunctionsCi/CdDockerGitPythonTerraform
One Month Ago
Hybrid
Palo Alto, CA, USA
Expert/Leader
Expert/Leader
Financial Services
Lead MLOps engineering for recommendation systems: build distributed GPU training pipelines, real-time and batch serving, deploy quantized LLMs, manage vector databases, implement monitoring/observability, optimize performance and reliability, and collaborate with product and architecture teams to scale AI infrastructure in AWS.
Top Skills: AwqAWSCudaDapoDockerEcsGpuGrpoKubernetesLlmsPtqPythonRayTransformer ModelsTrlVector DatabasesVerlVllm

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account