Airbyte Logo

Airbyte

Senior Site Reliability Engineer

Posted 25 Days Ago
Be an Early Applicant
Hybrid
San Francisco, CA, USA
196K-255K Annually
Senior level
Hybrid
San Francisco, CA, USA
196K-255K Annually
Senior level
Own and improve infrastructure for the Data Replication platform: Kubernetes, CI/CD, secrets, networking, cloud (AWS/GCP). Drive reliability, observability, AI-augmented tooling, canary rollouts, incident reduction, runbooks, and partner with product engineers.
The summary above was generated by AI

Airbyte is the data and action layer for AI agents. We give agents fast, accurate, authenticated access to business data across hundreds of sources, so they can discover the entities that matter, reason over real-time context, and take action in the systems they read from, not just observe them.

We started as the open-source standard for data movement and proved the economics of data integration at scale: hundreds of connectors, thousands of companies, and, since 2020, have raised $181M from leading investors including Benchmark, Accel, Altimeter, Coatue, and Y Combinator. As our CEO Michel Tricot puts it, "the last ten years were all about structured data. The future is all about context." We're now building that context infrastructure for production-grade agents on the same open foundation, as agents become the primary consumers of enterprise data.

Our mission is unchanged: make data available and actionable to everyone, everywhere. That everyone now includes AI agents.

 
The Role:

You'll be the infrastructure and reliability engineer on the Data Replication team - a full-stack product team running over 3 million sync jobs a week powering thousands of data use cases across multiple regions and clouds. You’ll build and maintain the infrastructure, set reliability standards, drive down incidents, and make it easier and safer for engineers to ship through tooling. You're equally comfortable in a Terraform file, a Kubernetes cluster, and a postmortem doc.


We expect engineers here to actively use AI as a force multiplier - agentic tools to automate toil, augment incident response, and build smarter internal tooling. If you're not already doing this, you should be excited to start. We care as much about how you work as what you build. Trust, directness, and craftsmanship matter here.

 
What You’ll Do:
  • Own the infrastructure underpinning the Data Replication platform - Kubernetes clusters, CI/CD pipelines, secrets management, networking, and cloud resource configuration across AWS and GCP.

  • Partner with product engineers to reliably integrate product features with infrastructure.

  • Maintain and enhance observability, alerting, and anomaly detection with an eye towards LLM automation.

  • Maintain and enhance AI-augmented release and internal tooling: canary deployments, progressive rollouts, automated release qualification, and rollback automation - with an eye towards LLM automation.

  • Set the infrastructure bar for the team - build self-serve tooling, write runbooks, and coach engineers to own more of their stack.

 
What You’ll Need:
  • 7+ years in infrastructure, platform engineering, SRE, or DevOps.

  • Hands-on ownership of Kubernetes, Helm, and Terraform in production environments.

  • Deep experience with observability stacks (Prometheus, Grafana, Datadog) and on-call operations.

  • Experience with CI/CD pipeline ownership and developer tooling.

  • Ability & willingness to read backend code to understand how systems break and instrument them correctly.

  • Fluency with AI tools - LLMs and agentic frameworks to automate, debug faster, and reduce toil.

  • A startup-ready mindset: comfortable with ambiguity, moving fast, and owning problems end-to-end.

 
Nice To Have:
  • Data pipelines, replication systems, or ETL/ELT platforms.

  • Control plane / data plane architectures or internal developer platforms.

  • Experience with Airbyte, CDKs, or connector-based architectures.

 
Location:
  • Onsite 4 days/week in San Francisco, CA

Why You'll Love Working at Airbyte:

At Airbyte, we believe great work happens when people feel supported, trusted, and empowered to grow. Our market-leading Total Rewards package is designed to help you thrive professionally and personally. Our benefits and perks include:

  • Flexible PTO with a culture that encourages at least 25 days off annually

  • 16 weeks fully paid parental leave for all parents

  • Comprehensive medical, dental, and vision coverage for employees and dependents

  • 401(k) retirement plan

  • Professional development budget, conference sponsorship, and book reimbursement

  • Commuter benefits and monthly internet reimbursement

  • Breakfast and lunch in our San Francisco office

  • A collaborative, in-person culture focused on learning, growth, and impact

If you find this role exciting, we encourage you to apply even if you think you don’t meet all of the requirements!

We are not accepting agency submissions or recruiting firm support for this role. Unsolicited resumes will not be considered.

Airbyte is an equal opportunity employer that does not discriminate on the basis of actual or perceived race, creed, color, religion, national origin, ancestry, age, physical or mental disability, pregnancy, genetic information, sex, sexual orientation, gender identity or expression, marital status, familial status, domestic violence victim status, veteran or military status, or any other legally recognized protected basis under federal, state or local laws. Pursuant to the San Francisco Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.

Airbyte is committed to providing reasonable accommodations for qualified individuals with disabilities in our job application procedures. Please let us know if you need assistance or accommodations due to a disability.

Similar Jobs

Yesterday
Remote or Hybrid
United States
Senior level
Senior level
Fintech • Software
Lead SRE efforts for DFIN SaaS: ensure availability, performance, scalability, and automation. Implement monitoring, CI/CD, IaC, container orchestration, AI-enhanced observability, incident response, RCA, and runbook automation while collaborating across engineering teams.
Top Skills: .NetAiopsAksAnsibleAppdynamicsAWSAzureAzure DevopsBashC#Ci/CdCloud Ai ServicesContainersCosmosDatadogDynatraceEksFirewallHarnessIdera Sql Diagnostic ManagerInfrastructure As Code (Iac)JavaJenkinsKubernetesLinuxLoad BalancingNew RelicPowershellPythonRedgate Sql MonitorSolarwinds Database Performance AnalyzerSQLTerraformWindows
8 Days Ago
Easy Apply
Hybrid
Easy Apply
170K-190K Annually
Senior level
170K-190K Annually
Senior level
AdTech • Big Data • Cloud • Marketing Tech • Software • Analytics
Lead SRE efforts to improve security, reliability, cost efficiency, and observability. Build automation, CI/CD, agentic AI platforms (MCPs), and tooling for capacity planning, incident response, and self-service. Evangelize SecDevOps and zero-trust designs across product and platform teams.
Top Skills: Ai Agentic FrameworksArgocdAWSBashCi/CdDevsecopsDockerEksGoGrafanaKubernetesLinuxLokiMcpNew RelicPrometheusPythonSamTerraformZero Trust
2 Days Ago
Hybrid
Palo Alto, CA, USA
Senior level
Senior level
Financial Services
Lead design and implementation of SRE practices, observability, reliability and AI-assisted operational workflows. Mentor engineers, define NFRs/SLOs, build automation, logging/metrics/tracing pipelines, containerized CI/CD/GitOps, and integrate AI agents for production-grade reliability.
Top Skills: AutogenCassandraChaos MonkeyChromaClaudeCrewaiDatadogDockerDynamoDBDynatraceFlinkFluentdGithub CopilotGitopsGo (Golang)GrafanaGremlinHadoopInfluxdbJavaKafkaKubernetesLangchainLanggraphLitmuschaosLogstashModel Context Protocol (Mcp)MongoDBNeo4JPineconePrometheusPythonPyTorchRabbitMQScikit-LearnSparkSplunkSqsTensorFlowTerraformTigergraphTimescaledbVectorWeaviate

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account