Senior Site Reliability Engineer

| Hybrid

Sorry, this job was removed at 12:50 p.m. (PST) on Tuesday, February 11, 2020

View 2012 Jobs

Find out who's hiring in San Francisco.

See all Developer + Engineer jobs in San Francisco

View 2012 Jobs

Apply

By clicking Apply Now you agree to share your profile information with the hiring company.

Save job

“The front page of the internet," Reddit brings over 330 million people together each month through their common interests, inviting them to share, vote, comment, and create across thousands of communities. Come for the cats, stay for the empathy.

Reddit is poised to rapidly innovate and grow like no other time in its history. This is a unique opportunity to leave your mark on one of the most influential and trafficked corners of the internet.

As a Site Reliability Engineer on Reddit’s Infrastructure team, you’ll use your knowledge of operating distributed systems to improve the consistency, reliability, and performance of our growing ecosystem of services. You’ll also use your development experience to contribute to the internal Infrastructure Product that all of Reddit Engineering uses to develop, deploy, and operate their services.

Join us and help build the future of Reddit!

Responsibilities

Advise: Work with engineering teams in designing and developing systems that are resilient and highly performant at tremendous scale
Amplify: Contribute to the development our internal Infrastructure Product, which is used by Reddit engineering teams to build, deploy, and operate their services
Automate: Build tools and systems to support the operation of our infrastructure and services
Diagnose: Draw on your knowledge of distributed systems to identify and fix network, system, and service-level issues
Optimize: Observe and improve performance, reduce cost, and improve the experience for millions of users

Qualifications

3+ years of Infrastructure, Operations, or Site Reliability Engineering experience
Experience with the development and operation of high-traffic backend systems
A demonstrated ability to debug, fix, and optimize code
Troubleshooting skills that span applications, networking (TCP/IP), and systems
Strong working knowledge of Linux (or UNIX) and TCP/IP.
Excellent communication and collaborative skills

Nice-to-haves

While not required, familiarity with any of these is a big plus!

Experience working in an environment that applies Infrastructure-as-code principles
Exposure to a Configuration Management System (Puppet, Chef, Salt, etc)
Experience with Infrastructure-as-code processes (via Terraform, CloudFormation, etc)
Docker or Kubernetes in a production setting
Working knowledge of Amazon Web Services or Google Cloud Platform
Operational knowledge of Postgres or Cassandra

Read Full Job Description

Senior Site Reliability Engineer

Location

Similar Jobs