MeridianLink Logo

MeridianLink

Senior Site Reliability Engineer (AWS)

Posted One Month Ago
Remote
Hiring Remotely in US
104K-163K Annually
Senior level
Remote
Hiring Remotely in US
104K-163K Annually
Senior level
Own day-to-day AWS and database operations for a serverless production platform, manage backups and disaster recovery, monitor and debug production, lead incident response and on-call, maintain infrastructure-as-code (SST/Pulumi), optimize costs, and mentor the team on operational best practices.
The summary above was generated by AI

As a Senior Site Reliability Engineer on our cloud engineering team, you'll keep our production environment healthy, secure, and running smoothly. This is an operations-focused role: you'll own the day-to-day administration of our AWS accounts and databases, backup posture across our data stores, and production monitoring and debugging for a fully serverless platform. Your work will span the operational side of the software development life cycle — from deployment to maintenance and updates — always striving for continuous improvement. You'll keep our infrastructure clean, easily deployable, and scalable, creating a stable operating environment for the whole team.

Responsibilities

  • Own day-to-day administration across AWS services, accounts, and access, as well as database administration across PostgreSQL and our other data stores.

  • Own backup posture across databases, S3 buckets, and queues; verify restores regularly and maintain a tested disaster recovery plan.

  • Proactively monitor production — CloudWatch dashboards, metric alarms, log-based metrics, and Slack alerting — addressing operational issues before they impact users.

  • Lead production debugging and incident response: build and maintain runbooks, participate in the on-call rotation, and resolve queue and dead-letter-queue failures through retry, redrive, and recovery.

  • Continuously refine our infrastructure to ensure it is easily deployable and scalable: keep infrastructure as code (SST/Pulumi) accurate, retire unused infrastructure, and keep cost visible and justified.

  • Share your knowledge of production operations with the team, fostering a culture of learning and growth.

Qualifications: Knowledge, Skills, & Abilities

  • Bachelor's degree and 4-6 years of related experience or equivalent work experience.

  • 5+ years of experience in DevOps, site reliability, or platform operations, with significant responsibility for production systems.

  • 3+ years of hands-on experience with AWS, with an emphasis on serverless services (Lambda, SQS, EventBridge, CloudWatch, S3).

  • Strong database administration experience: PostgreSQL operations, backup and recovery, and query performance; comfort administering other data stores.

  • Proficiency in scripting languages such as TypeScript, Python, and bash for production automation and operational tooling.

  • Strong understanding of Linux, DNS, TLS, Docker, GitHub Actions, and infrastructure as code (SST, Pulumi, or Terraform).

  • Experience with production monitoring and alerting, incident response, and on-call ownership.

Similar Jobs

2 Minutes Ago
In-Office or Remote
115K-160K Annually
Senior level
115K-160K Annually
Senior level
Fintech
Design, develop, test, deploy, troubleshoot, and maintain scalable web applications and APIs using React, TypeScript, Node.js, GraphQL, and cloud technologies. Collaborate with product, design, architecture, and engineering teams; contribute to software standards, automation, documentation, security, performance, and reliability. Mentor engineers, participate in code reviews, work within Agile processes, and complete a technical assessment during hiring.
Top Skills: Apollo GraphqlCi/Cd PipelinesDockerGitGCPGraphQLJavaScriptJestNode.jsPostgresReactTypescript
2 Minutes Ago
In-Office or Remote
66K-83K Annually
Mid level
66K-83K Annually
Mid level
Fintech
Delivers accurate, timely reporting and actionable analysis for internal and external stakeholders. Handles ad hoc requests, validates and reconciles data, performs quality checks, documents report logic, manages request queues and SLAs, and identifies automation opportunities. Uses SQL and Python for data extraction, transformation, validation, analysis, and recurring process automation. Maintains report templates and collaborates with data, visualization, product, operations, compliance, and account management teams while protecting sensitive data.
Top Skills: BigQueryClaude CodeGitGithub CopilotLookerPower BIPythonSnowflakeSQLTableau
2 Minutes Ago
In-Office or Remote
25-30 Hourly
Senior level
25-30 Hourly
Senior level
Fintech
Manages real-time workforce operations for multi-department contact centers by monitoring performance, adjusting staffing and schedules, balancing skills and volumes, and administering overtime and time off. Produces reports, compares results against forecasts, analyzes workforce metrics, resolves issues, coordinates with operations and vendors, and recommends process improvements. The role also mentors team members and supports service-level and financial objectives.
Top Skills: CxoneIexMS Office

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account