Onos Health Logo

Onos Health

Lead Infrastructure Engineer

Posted 9 Days Ago
Be an Early Applicant
Hybrid
San Francisco, CA, USA
200K-275K Annually
Expert/Leader
Hybrid
San Francisco, CA, USA
200K-275K Annually
Expert/Leader
Own Onos Health’s production infrastructure, including AWS architecture, multi-region disaster recovery, backups, monitoring, SLOs, incident response, SOC 2 and HIPAA controls, Terraform, CI/CD, preview environments, and AI-agent deployment guardrails. Establish platform strategy, reliability practices, security automation, audit evidence, and on-call operations as the company’s first dedicated infrastructure hire. The role also requires technical leadership, delegation, and communication with non-technical stakeholders.
The summary above was generated by AI
About Onos Health

Onos Health’s mission is simple but ambitious: ensure every healthcare dollar goes toward delivering the highest quality care. Today, 30% of total U.S. healthcare spending is wasted due to ineffective care and administrative burden caused by misalignment between providers and payers.

Onos is addressing this by building the largest AI-driven healthcare data platform. Our models enables payers to make faster, more accurate decisions across their populations. By guiding members to the right care, Onos is channeling more dollars to high-quality care that drives better outcomes while making healthcare more affordable.

Onos is well-funded by some of the best healthcare investors and is working with the nation’s largest health plans. Come join a category-defining company and help reimagine healthcare for the better.

Why Onos?
  • Meaningful impact: Help fix what is fundamentally broken in healthcare

  • Direct collaboration: Work alongside experienced founders with deep healthcare and data expertise

  • Culture: Join a high-performing, transparent, and results-oriented team

  • Ownership: Significant responsibility and autonomy from day one

  • Opportunity: Play a pivotal role in building a fast-growing, category-defining healthcare AI company

The Role

We're seeking an experienced infrastructure engineer to become our first dedicated platform hire and the owner of the infrastructure the Onos platform runs on. Onos is in production with the nation's largest health plans, which comes with contractual uptime SLAs, disaster recovery commitments, and a security bar (SOC 2, HIPAA) our clients audit. Until now this has been carried collectively by our product engineers and founders — you'll own it end to end. We build heavily with AI coding agents, so much of your leverage will come from specifying work well and directing agents rather than typing every line yourself; prior tech lead or engineering management experience translates directly. As an early team member, you'll set the patterns every future platform engineer at Onos inherits. This role is a hybrid role based in San Francisco, where you'll be expected to work at our office in person 3 times a week.

What you'll be doing at Onos:
  • Own our availability, disaster recovery, and backup commitments to enterprise clients — multi-region failover architecture and the recovery exercises that prove it

  • Stand up production monitoring, alerting, SLOs, and our on-call rotation and incident response process

  • Own the technical controls behind SOC 2 and HIPAA: cloud security posture (AWS org guardrails, IAM least-privilege, KMS/encryption), vulnerability remediation, and continuous audit evidence through Vanta, with a path toward HITRUST

  • Build the CI/CD pipelines, Terraform/IaC foundations, preview environments, and test infrastructure the whole team ships on

  • Build the guardrails that let AI coding agents ship safely — policy-as-code, deploy verification, and agent-operated operations tooling

  • Set the strategy and operating rhythm for platform work: priorities, status, and what we deliberately defer

Technical Challenges At Onos:
  • Architect AI SRE agents to ensure up-to-date compliance and reliability, enabling engineers to work more effectively and strategically

  • Right-size enterprise-grade reliability: SLOs and alerting you can trust without drowning a small team in pager noise

  • Turn compliance into continuously verified infrastructure — security controls and audit evidence as code

  • Scale a CI/CD and environments platform where AI agents, not just humans, are the primary users

  • Support multi-region disaster recovery with defined RTO/RPO targets and immutable, restore-tested backups for a multi-tenant healthcare platform

Tech Stack:

At Onos, we work with a modern tech stack where we continuously evaluate and adopt cutting-edge technologies as we scale.

  • Infrastructure/Systems: AWS (ECS, Bedrock, Glue, etc.), Langfuse, Terraform

  • Languages/Frameworks

    • Backend: Python, Django, Celery / Celery Beat, django-ninja, django-tenants

    • Frontend: NextJS, Typescript, Tanstack Query, Shadcn UI, Zod, Nuqs

  • Database/Storage: PostgreSQL (AWS RDS), S3, Clickhouse

  • Development Tools: Github, Linear, Claude Code, Codex, CoderabbitAI

What we're looking for:
  • Deep AWS experience — you've owned production cloud infrastructure end to end (IAM, networking, KMS, containers, managed databases)

  • Strong Terraform/IaC and CI/CD expertise; you treat pipelines and environments as products with users

  • Taken a company through at least one SOC 2 (or HITRUST/ISO 27001) audit cycle with your hands on the technical controls

  • SRE fundamentals — SLOs, incident management, DR design — with the judgment to right-size reliability for a startup with enterprise contracts

  • Prior experience leading engineering teams as a tech lead or engineering manager — you can break down ambiguous goals, delegate (to humans or AI agents), and communicate crisply with non-technical stakeholders

  • Customer obsessed and motivated to make an impact in the healthcare space

Bonus points if you have:
  • Significant experience in healthcare or another regulated industry, with HIPAA fluency

  • Policy-as-code (OPA, Kyverno) or compliance automation (e.g., Vanta, Drata) experience

  • Been the first infrastructure hire, or helped found or lead a platform team

  • Built internal tooling or infrastructure for LLM/agent systems

Benefits and Perks
  • Hybrid arrangement: 3 days/week at San Francisco office (Financial District)

  • Unlimited vacation policy

  • Paid parental leave

  • Medical, dental, and vision insurance

  • Pre-tax commuter benefits

  • 401(k)

  • Significant equity as an early employee

  • Direct mentorship from experienced founders

  • Ground-floor opportunity to help build a team and culture

  • Regular team events and offsites

  • Company-provided equipment and home office setup

We are an equal opportunity employer and value diversity at our company. We do not discriminate on the basis of race, religion, color, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status.

HQ

Onos Health San Francisco, California, USA Office

San Francisco, CA, United States

Similar Jobs

8 Days Ago
Hybrid
San Leandro, CA, USA
214K-224K Annually
Senior level
214K-224K Annually
Senior level
Fintech • Financial Services
Leads infrastructure engineering initiatives for business applications, including architecture, design, deployment, maintenance, production support, outage analysis, and infrastructure upgrades. Builds and automates solutions using cloud platforms, CI/CD, scripting, monitoring, and configuration-management tools. Evaluates software solutions, leads implementations, manages risks and controls, documents technical processes, and collaborates with internal teams, customers, and vendors.
Top Skills: AgileAnsibleApacheAppdynamicsAviAWSAzureBashCi/CdF5 LtmGitGit SaasGCPHarnessJenkinsLinuxOpenshift Container PlatformPythonSplunkTasTomcatUdeployUnix
9 Days Ago
Hybrid
San Francisco, CA, USA
Senior level
Senior level
Financial Services
Leads infrastructure engineering projects involving network architecture, migration, performance testing, troubleshooting, automation, resiliency, and cloud operations. Designs and implements infrastructure changes across platforms, performs pre- and post-migration validation and cutovers, and applies Cisco, Arista, Fortinet, BGP, OSPF, switching, scripting, and AI-assisted workflows. Ensures recommendations and changes meet security, sensitivity, auditability, and resiliency standards while collaborating cross-functionally.
Top Skills: Ai-Assisted Infrastructure ToolsAristaBgpCiscoFortinetLacpOspfPrivate CloudPublic CloudPythonStpVlans
24 Days Ago
Hybrid
San Francisco, CA, USA
Senior level
Senior level
Financial Services
Lead design, build, and delivery of a next-generation SD‑WAN ecosystem. Produce HLD/LLD artifacts, define requirements, lead testing and automation, apply security and resiliency controls, maintain audit-ready documentation, and coordinate stakeholders to mitigate operational and technology risks while leveraging enterprise AI-assisted practices.
Top Skills: Api-Driven OperationsEnterprise Ai CapabilitiesItil/CabNetwork AutomationNetwork MonitoringNetwork VirtualizationOperational AnalyticsOverlay/Underlay NetworkingRouting (Layer 2/Layer 3)ScriptingSd‑WanSd‑Wan PlatformsTelemetry

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account