Matter Logo

Matter

AI Infrastructure Engineer

Reposted 22 Days Ago
Be an Early Applicant
In-Office
Menlo Park, CA, USA
Mid level
In-Office
Menlo Park, CA, USA
Mid level
The AI Infrastructure Engineer will design and operate data and compute systems, manage GPU infrastructure, and develop data pipelines for machine learning models.
The summary above was generated by AI

ABOUT MATTER

Matter is building the AI-native autonomy stack for physical manufacturing in the United States. We operate our own factories, deploy our own software, and collect data from every stage of production — from CAD intake to finished goods.

Our platform, MatterOS, is the unified software layer for factory operations, process orchestration, and autonomy deployment. The data pipeline that feeds it — from machine telemetry on the floor to model training in the cloud — is the infrastructure you will build and own.

 

THE ROLE

We are hiring an AI Infrastructure Engineer to design and operate the data and compute systems that power MatterOS and our Sim2Real training pipeline. You will work across edge computing, cloud training infrastructure, and the data pipelines that make our “Smart Data” strategy real.

Your job is to ensure that every data point — from a torque sensor reading to a camera frame — is tagged with the machine ID, process state, and production context that makes it trainable.

 

WHAT YOU’LL DO

•      Design and maintain the edge-to-cloud data pipeline with semantic context preserved end-to-end

•      Build and manage GPU compute infrastructure for VLA model training, experiment tracking, and distributed training workflows

•      Implement the data collection layer for 100% capture from modular assembly workcells, including camera feeds, sensor streams, machine state, and process metadata

•      Develop feature engineering pipelines that transform raw operational data into structured training inputs for AI models

•      Manage model deployment to edge hardware in the factory: latency, versioning, rollback, and monitoring in production

•      Build observability systems that surface model performance degradation, data drift, and equipment anomalies in real time

•      Collaborate with AI researchers to translate model requirements into infrastructure specifications and vice versa

 

WHAT WE’RE LOOKING FOR

•      3+ years of experience in ML infrastructure, MLOps, or data engineering in a production environment

•      Strong command of distributed data systems: Kafka, Flink, or equivalent; time-series databases (InfluxDB, TimescaleDB, or similar)

•      Experience with GPU cluster management and distributed training (SLURM, Ray, or Kubernetes-based)

•      Familiarity with industrial protocols: OPC UA, MQTT, Modbus (or willingness to learn quickly)

•      Proficiency in Python; comfort with C++ or Rust for performance-critical edge components is a plus

•      Systems thinking: you understand that data quality, not data volume, is what makes AI work in constrained physical environments

 

NICE TO HAVE

•      Experience with NVIDIA Isaac Sim, ROS2, or edge AI deployment (Jetson, FPGA, or similar)

•      Background in industrial IoT or factory automation systems

•      Familiarity with model serving frameworks (Triton, TorchServe, or ONNX Runtime)

 

WHY MATTER

Most AI infrastructure roles are about keeping existing systems running. At Matter, you are building the infrastructure from scratch for a category that doesn’t fully exist yet: autonomous physical manufacturing.

HQ

Matter San Francisco, California, USA Office

San Francisco, CA, United States

Similar Jobs

6 Days Ago
Hybrid
Sunnyvale, CA, USA
147K-224K Annually
Expert/Leader
147K-224K Annually
Expert/Leader
Artificial Intelligence • Big Data • Healthtech • Machine Learning • Software • Biotech
Leads the design, development, deployment, and governance of an enterprise AI platform on AWS and Kubernetes. Builds agentic and multi-agent workflows, secure MCP and tool integrations, identity-scoped authorization, policy controls, isolated container runtimes, audit trails, and observability. Partners with engineering, security, regulatory, product, and business teams to deliver compliant AI infrastructure, troubleshoot complex distributed systems, shape technology strategy, and mentor technical teams.
Top Skills: Amazon EksAmazon VpcAuth0AutogenAWSAws BedrockAws CdkAws CloudtrailAws IamAws KmsAws PrivatelinkCedarCi/CdClaude Agent SdkFaissGoHelmJwtKubernetesLangchainLanggraphMlopsModel Context Protocol (Mcp)OauthOktaOpensearch ServerlessOpentelemetryPgvectorPythonRagTerraformTypescript
12 Days Ago
In-Office
Sunnyvale, CA, USA
182K-242K Annually
Senior level
182K-242K Annually
Senior level
Cloud • Information Technology • Machine Learning
Senior Software Engineer focused on AI infrastructure performance insights and observability. Responsible for designing and building monitoring, instrumentation, metrics, and tooling to measure and optimize compute and platform performance while collaborating with infrastructure and ML teams.
4 Days Ago
In-Office
255K-375K Annually
Expert/Leader
255K-375K Annually
Expert/Leader
Aerospace • Artificial Intelligence • Hardware • Machine Learning • Software • Defense • Manufacturing
Lead technical vision for AI and platform infrastructure across the company. Solve ambiguous, high-stakes problems end-to-end, set architecture and paved-road standards, guide AI provider selection, ensure compliance for government environments, mentor engineers, and drive adoption of best practices in DevOps, CI/CD, developer experience, and compute infrastructure.
Top Skills: Agent Frameworks (CrewaiArtifact/Registry ManagementAWSAws GovcloudAzureAzure GovernmentCi/CdContainersEmbedding PipelinesFedrampGCPIl4/Il5Infrastructure-As-Code (Terraform)Large Language Model ApisLlm Eval FrameworksNist 800-53OpentelemetryOrchestrationPrompt EngineeringPydantic Ai)Retrieval-Augmented Generation (Rag)Service Mesh

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account