Summit (summithq.com) Logo

Summit (summithq.com)

Site Reliability Architect

Posted 8 Days Ago
Remote
Hiring Remotely in United States
136K-175K Annually
Senior level
Remote
Hiring Remotely in United States
136K-175K Annually
Senior level
Lead the vision, strategy, and architecture for observability and reliability platforms. Define SLOs/SLIs/SLAs, implement monitoring/alerting automation, lead incident response and post-incident analysis, evaluate tools, and mentor engineering teams to improve resilience, scalability, and reliability across multi-cloud environments.
The summary above was generated by AI

At Summit, we're on the lookout for talent that doesn't just think "outside the box," but brings their own unique perspective to the table. With our relentless pursuit of excellence and curiosity, we lead innovation in our industry. We humanize technology by actively listening to our clients, crafting tailored proposals, and delivering on the promise of technology with precision and purpose.

 

Summit is a leading provider of enterprise-class Application Hosting, Managed Services, and Cloud Solutions for regulated industries, with deep experience supporting compliance, security, and performance in complex IT environments. Our mission is to simplify the complex - ensuring our clients’ technology environments are secure, performant, and purpose-built for their most critical applications.

 

As our Site Reliability Architect, you will own the vision, strategy, and architecture of Summit’s observability and reliability platforms. You will ensure that the systems supporting our most critical applications are designed with resilience, scalability, and efficiency in mind. Beyond managing platforms, you will shape the standards and practices that make Summit’s infrastructure visible, measurable, and reliable.

 

You will collaborate across engineering, operations, and product teams to embed observability into the fabric of our technology, mentor engineers on best practices, and influence architecture decisions that improve performance and reliability. The right candidate will thrive in a fast-moving environment, combining deep technical expertise with the ability to see the bigger picture and guide the evolution of Summit’s reliability ecosystem.

 

What You’ll Do:

  • Define the observability and reliability architecture strategy across Summit’s platforms and services.
  • Implement site reliability concepts based around SLOs, SLIs, and SLAs, across teams and platforms
  • Partner with engineering and operations leadership to ensure system design aligns with resilience and scalability goals.
  • Lead the design, implementation, and governance of observability frameworks and standards across teams.
  • Oversee the automation of monitoring, alerting, and incident response processes to maximize efficiency and reduce human error.
  • Serve as the escalation point and lead for complex, cross-platform incidents, guiding resolution and post-incident analysis.
  • Evaluate and introduce emerging tools, frameworks, and practices to strengthen Summit’s reliability and observability posture.
  • Mentor and coach engineering teams, fostering a culture of reliability, automation, and continuous improvement.


What You’ll Deliver:

  • A scalable, standards-driven observability framework embedded across Summit’s technology stack.
  • Clear alignment of monitoring, alerting, and automation with business-critical definitions of health.
  • Reduced time-to-detect and time-to-resolve for incidents through automation and well-designed response frameworks.
  • Improved reliability practices across teams through coaching, knowledge-sharing, and collaboration.
  • Strategic roadmaps for observability and monitoring platforms that support Summit’s growth and evolving client needs.


You’ll Thrive in This Role If You:

  • Have a proven track record in designing and implementing observability and reliability platforms at scale.
  • Are comfortable influencing architecture and strategy across engineering, operations, and product teams.
  • Enjoy mentoring engineers and shaping organizational practices, not just managing tools.
  • Think strategically but aren’t afraid to get hands-on when solving complex problems.
  • Thrive in environments where you can innovate, set standards, and bring structure to complex ecosystems.

 

Bonus Points:

  • Expertise with observability stacks such as ELK, Grafana/Graphite/InfluxDB, LogicMonitor, Prometheus, or related platforms.
  • Advanced automation experience using Ansible, Terraform, or equivalent tooling.
  • Experience designing reliability frameworks in multi-cloud environments (Azure, AWS, or hybrid).
  • Knowledge of scripting and development languages (Python, Go, Ruby, or Javascript).
  • Familiarity with compliance-heavy industries where reliability, security, and auditability are paramount.


How We Work at Summit

At Summit, our culture and core values are important to us. As a diverse team of passionate pathfinders, we deliver on the promise of technology. If this sparks your interest, we'd love to chat with you!


How you approach the work matters just as much as what you do. At Summit, our core values show up in how we make decisions, work together, and support our customers. 


We build the infrastructure that keeps businesses running. We need people who take that seriously, and have fun doing it.


Our core values:

  • Empower our people
  • Constant elevation
  • Customer first
  • Focus on outcomes
  • Embrace curiosity


Benefits:

Summit offers a total rewards package designed to support you at work, at home, and everywhere in between. Here’s a snapshot of what that looks like:

  • Flexible Time Off (yes, really) – take what you need, we trust you to manage it
  • Medical, Dental & Vision – comprehensive coverage with HSA/HRA options
  • 401(k) with 4% match – we invest alongside you
  • Parental Leave – paid time for the moments that matter most
  • Life & Disability Insurance – built-in peace of mind
  • Wellness Support – resources for your mental and physical health
  • Free Colocation & Cloud Access – build and experiment in real environments
  • Work From Anywhere – remote-friendly by design
  • A low-ego, get-it-done culture – come as you are, do great work


Compensation:

The salary range for this role is $136,000 – $175,000 annually, depending on your skills, experience, and location. We aim to make offers that feel fair, forward-looking, and reflective of what you bring to the table — and where you want to grow. 


Internal candidates may see variations based on current role, compensation, and progression at Summit. Same thoughtful approach, just with additional context.

 

Summit is committed to a diverse and inclusive workplace. Summit is an equal opportunity employer and does not discriminate on the basis of race, national origin, gender, gender identity, sexual orientation, protected veteran status, disability, age, or other legally protected status. 


As part of this commitment, Summit will ensure that persons with disabilities are provided with reasonable accommodations. If reasonable accommodation is needed to participate in the job application or interview process, to perform essential job functions, and/or to receive other benefits and privileges of employment, please email [email protected].

Similar Jobs

3 Days Ago
Remote
United States
Senior level
Senior level
Agency • Information Technology
Lead SRE role designing and maintaining CI/CD pipelines (GitHub Actions), containerized deployments (Docker, Kubernetes, AKS, Helm), web/mobile app releases, observability, automated testing, and DevOps best practices across cloud environments with cross-functional collaboration and regulatory compliance.
Top Skills: AksAndroidAzure Application InsightsAzure Log AnalyticsAzure MonitorBashBranchingDockerDocker ComposeGitGit HooksGithub ActionsGoogle PlayHelmHerokuiOSIos App StoreJavaKubernetesNpmPowershellPull RequestsPythonSonarqubeVeracodeVercel
An Hour Ago
Easy Apply
Remote or Hybrid
United States
Easy Apply
90K-176K Annually
Entry level
90K-176K Annually
Entry level
Big Data • Cloud • Software • Database
Provide expert technical support to enterprise customers by troubleshooting complex MongoDB, database, infrastructure, networking, storage, security, performance, recovery, and scalability issues. Analyze problems across application, database, and storage layers; advise on architecture and production operations; collaborate with product and development teams; contribute to diagnostic and performance tools; and mentor new engineers. This remote role follows a weekend shift after a six-to-nine-month weekday ramp period.
Top Skills: Active DirectoryAWSAzureCC#C++Cloud ManagerDnsFedrampGCPGitGoJavaJavaScriptKerberosKubernetesLdapLinuxMongoDBMongodb AtlasNasNode.jsNoSQLPythonRdbmsRubySanSsdSsl/TlsTcp/Ip
An Hour Ago
Remote or Hybrid
Mountain View, CA, USA
116K-192K Annually
Expert/Leader
116K-192K Annually
Expert/Leader
Artificial Intelligence • Cloud • HR Tech • Information Technology • Productivity • Software • Automation
Owns and operates a scaled digital customer success program for mid-market SaaS customers, spanning onboarding through renewal. Responsibilities include designing lifecycle campaigns, driving product adoption, protecting gross revenue retention, creating automated risk interventions, building customer health scores and segmentation, monitoring engagement metrics, and developing executive dashboards. The role partners with Customer Success, Product, Marketing, and Operations teams to use AI, automation, and customer signals to personalize engagement, improve retention, and increase CSM capacity.
Top Skills: AIAutomationCrm SystemsCustomer Success PlatformsMarketing Automation ToolsSaaS

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account