Early Warning Logo

Early Warning

Principal Site Reliability Engineer - Paze

Posted 7 Days Ago
Be an Early Applicant
In-Office
San Francisco, CA, USA
194K-284K Annually
Expert/Leader
In-Office
San Francisco, CA, USA
194K-284K Annually
Expert/Leader
Provides enterprise-level technical leadership for site reliability engineering across production services. Improves reliability, scalability, observability, deployment, automation, incident response, resilience, capacity management, and operational readiness. Establishes SLIs, SLOs, error budgets, and technical direction while partnering with software engineering teams. Leads critical incident response, reduces operational toil, and develops reusable engineering practices, tooling, and platforms that improve organizational capability.
The summary above was generated by AI

At Early Warning, we’ve powered and protected the U.S. financial system for over thirty years with cutting-edge solutions like Zelle®, Paze℠, and so much more. As a trusted name in payments, we partner with thousands of institutions to increase access to financial services and protect transactions for hundreds of millions of consumers and small businesses.

Positions located in Scottsdale, San Francisco, Chicago, or New York follow a hybrid work model to allow for a more collaborative working environment.

Candidates responding to this posting must independently possess the eligibility to work in the United States, for any employer, at the date of hire. This position is ineligible for employment Visa sponsorship.

Role Summary 

The Principal Site Reliability Engineer applies software engineering and systems engineering practices to improve the reliability, resilience, scalability, and operational health of production services. The role partners with Software Engineering and other technology teams to ensure reliability, observability, recoverability, performance, and operational readiness are engineered into systems throughout their lifecycle. 

The role operates at enterprise scope, establishing technical direction and applying evidence-driven engineering, technical rigor, sound judgment, automation, and broad systems expertise across organizational boundaries. 

Core Responsibilities 

  • Use software engineering, automation, and DevOps principles and practices to continually improve how services are built, tested, deployed, observed, operated, and recovered. 

  • Use data, evidence, experimentation, and rigorous engineering analysis appropriate to the level to identify reliability risks, test assumptions, and guide technical decisions. 

  • Define, implement, or improve SLIs, SLOs, error budgets, and other service-health measures appropriate to the scope of responsibility. 

  • Improve observability through metrics, logging, tracing, monitoring, alerting, dashboards, and service-health instrumentation. 

  • Drive continuous improvement across CI/CD, observability, deployment practices, Infrastructure as Code, automation, testing, incident response, capacity management, resilience, and operational readiness. 

  • Identify recurring or systemic production issues and translate operational experience into improvements in code, architecture, automation, tooling, and engineering practices. 

  • Partner with Software Engineering teams to incorporate reliability, resiliency, scalability, performance, observability, recoverability, and operational readiness throughout the development lifecycle. 

  • Participate in or lead incident response, troubleshooting, service restoration, and blameless post-incident learning appropriate to the level. 

  • Provides enterprise-level technical leadership for critical production incidents and establishes or influences engineering practices that improve incident response, escalation, service restoration and sustainable on-call operations across the organization. 

  • Reduce operational toil and unnecessary manual intervention through software, automation, reusable patterns, and better engineering practices. 

Leveling Intent 

Principal represents domain-level technical leadership and organizational impact. Deep individual expertise is expected, but Principal-level impact comes from identifying systemic risk, establishing technical direction, influencing engineering practices and architecture, and multiplying the capability of the broader engineering organization. 

Level Expectations 

  • Acts as an enterprise force multiplier, raising the effectiveness and technical capability of engineers and teams across the organization while building sustainable organizational capability rather than individual dependency. 

  • Demonstrates software engineering, systems thinking, troubleshooting, and production reliability capabilities appropriate to the level. 

  • Applies evidence-driven reasoning and technical rigor to distinguish observed facts from assumptions and make defensible engineering recommendations. 

  • Shares knowledge and contributes to sustainable engineering capability rather than creating dependency on individual expertise. 

  • Operates with significant autonomy across the organization's most consequential reliability challenges. 

  • Establishes enterprise technical direction, develops senior technical leaders, and demonstrates impact well beyond systems personally touched. 

Minimum Qualifications 

  • Typically 15+ years of relevant professional experience in Software Engineering, Site Reliability Engineering, Systems Engineering, Cloud/Platform Engineering, DevOps, Infrastructure Engineering, Architecture where applicable, or a comparable technical discipline. 

  • Experience with software development or scripting using one or more modern programming languages. 

  • Experience with software engineering principles, distributed systems, production troubleshooting, automation, and observability appropriate to the level. 

  • Experience with public cloud technologies and architectures, preferably AWS, along with infrastructure, networking, Linux/Unix, and modern application architectures appropriate to the level. 

  • Demonstrated analytical, problem-solving, communication, and collaboration skills appropriate to the scope of the role. 

Preferred Qualifications 

  • Hands-on experience with AWS is preferred, or comparable experience with another major cloud platform such as Microsoft Azure, Google Cloud Platform (GCP), or Oracle Cloud Infrastructure (OCI). 

  • Experience developing, deploying, operating, or improving highly available production software or distributed systems. 

  • Experience with CI/CD, Infrastructure as Code, containers or orchestration, observability, monitoring, alerting, and software-delivery automation. 

  • Experience with SLIs, SLOs, error budgets, incident management, performance analysis, capacity management, resilience testing, disaster recovery, or operational readiness appropriate to the level. 

  • Experience creating reusable automation, tooling, platforms, patterns, or practices that improve engineering effectiveness. 

  • Bachelor's degree in Computer Science, Software Engineering, Computer Engineering, Information Systems, or a related technical field, or equivalent practical experience. 

The base pay scale for this position in:
Phoenix, AZ/ Chicago, IL / Washington, DC in USD per year is: $194,000 - $237,000.
New York, NY/ San Francisco, CA in USD per year is: $232,000 - $284,000.

Additionally, candidates are eligible for a discretionary incentive plan and benefits.

This pay scale is subject to change and is not necessarily reflective of actual compensation that may be earned, nor a promise of any specific pay for any specific candidate, which is always dependent on legitimate factors considered at the time of job offer. Early Warning Services takes into consideration a variety of factors when determining a competitive salary offer, including, but not limited to, the job scope, market rates and geographic location of a position, candidate’s education, experience, training, and specialized skills or certification(s) in relation to the job requirements and compared with internal equity (peers). The business actively supports and reviews wage equity to ensure that pay decisions are not based on gender, race, national origin, or any other protected classes.

Physical Requirements

Early Warning works together in a highly collaborative office environment. Working conditions consist of a normal office environment. Work is primarily sedentary and requires extensive use of a computer and involves sitting for periods of approximately four hours. Work may require occasional standing, walking, kneeling, and reaching. Must be able to lift 10 pounds occasionally and/or negligible amount of force frequently. Requires visual acuity and dexterity to view, prepare, and manipulate documents and office equipment including personal computers. Requires the ability to communicate with internal and/or external customers.

Employee must be able to perform essential functions and physical requirements of position with or without reasonable accommodation.

Candidates responding to this posting must independently possess the eligibility to work in the United States at the date of hire.

Some of the Ways We Prioritize Your Health and Happiness 


  • Healthcare Coverage – Competitive medical (PPO/HDHP), dental, and vision plans as well as company contributions to your Health Savings Account (HSA) or pre-tax savings through flexible spending accounts (FSA) for commuting, health & dependent care expenses.

  • 401(k) Retirement Plan – Featuring a 100% Company Safe Harbor Match on your first 6% deferral immediately upon eligibility.

  • Paid Time Off – Flexible Time Off for Exempt (salaried) employees, as well as generous PTO for Non-Exempt (hourly) employees, plus 11 paid company holidays and a paid volunteer day.

  • 12 weeks of Paid Parental Leave

  • Maven Family Planning – provides support through your Parenting journey including egg freezing, fertility, adoption, surrogacy, pregnancy, postpartum, early pediatrics, and returning to work.


And SO much more! We continue to enhance our program, so be sure to check our Benefits page here for the latest. Our team can share more during the interview process!

Pursuant to the San Francisco Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.

Early Warning Services, LLC (“Early Warning”) considers for employment, hires, retains and promotes qualified candidates on the basis of ability, potential, and valid qualifications without regard to race, religious creed, religion, color, sex, sexual orientation, genetic information, gender, gender identity, gender expression, age, national origin, ancestry, citizenship, protected veteran or disability status or any factor prohibited by law, and as such affirms in policy and practice to support and promote equal employment opportunity and affirmative action, in accordance with all applicable federal, state, and municipal laws. The company also prohibits discrimination on other bases such as medical condition, marital status or any other factor that is irrelevant to the performance of our employees. 

Early Warning San Francisco, California, USA Office

275 Sacramento St, San Francisco, CA, United States, 94111

Similar Jobs

An Hour Ago
Hybrid
111K-178K Annually
Mid level
111K-178K Annually
Mid level
AdTech • eCommerce • Information Technology • Software • Travel • Generative AI
Build and scale backend services, APIs, and advertiser-facing features for Expedia Group’s unified advertising platform. The role involves system and API design, data modeling, deployment, operational support, performance improvement, observability, troubleshooting, and secure testing. Engineers will primarily develop backend systems with some React frontend work, collaborate across distributed teams, and contribute to scalable, high-throughput advertising technology.
Top Skills: ClaudeFlinkGithub CopilotGraphQLJavaKafkaKotlinReactRest
An Hour Ago
Hybrid
172K-275K Annually
Senior level
172K-275K Annually
Senior level
AdTech • eCommerce • Information Technology • Software • Travel • Generative AI
Lead design, build, and operate scalable GenAI backend services and APIs (Python/FastAPI). Integrate LLM providers, vector stores, and Aurora PostgreSQL for RAG workflows. Implement guardrails, observability, CI/CD, testing, and on-call support. Partner with cross-functional teams and produce documentation to enable safe, multi-tenant GenAI platform adoption.
Top Skills: Aurora PostgresqlAWSCi/CdDjangoElasticsearchFastapiFlaskFlowiseLangchainLangfuseLanggraphLangsmithLlamaindexLlmN8NPostgresPythonVector Stores
An Hour Ago
Hybrid
75K-85K Annually
Junior
75K-85K Annually
Junior
Automotive • Professional Services • Software • Consulting • Energy • Chemical • Renewable Energy
Determines project scopes, test programs, specifications, and schedules for fire protection product evaluations. Communicates technical requirements, costs, timelines, and results to clients; examines samples for compliance with UL requirements; coordinates laboratory technicians and engineering assistants; prepares reports and follow-up procedures; resolves engineering issues; supports standards, test-method, and equipment development; and may travel to client sites or industry events.
Top Skills: Laboratory Testing EquipmentMS OfficeUl Product Safety Standards

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account