Niche Logo

Niche

Tech Lead Manager, Site Reliability

Posted 2 Days Ago
Remote
Hiring Remotely in USA
146K-183K Annually
Mid level
Remote
Hiring Remotely in USA
146K-183K Annually
Mid level
Leads a 3–4 person Site Reliability team while remaining hands-on with software engineering, incident response, infrastructure automation, platform tooling, observability, disaster recovery, and modernization. Owns technical direction, reliability outcomes, technical debt, architecture, code reviews, hiring, performance management, career development, and team processes. The role participates in on-call and drives platform adoption, self-service, delivery quality, and effective AI-assisted development.
The summary above was generated by AI

About Niche

Niche is the leader in school search. Our mission is to make researching and enrolling in schools easy, transparent, and free. With in-depth profiles on every school and college in America, 140 million reviews and ratings, and powerful search tools, we help millions of people find the right school for them. We also help thousands of schools recruit more best-fit students, by highlighting what makes them great and making it easier to visit and apply.

Niche is all about finding where you belong, and that mission inspires how we operate every day. We want Niche to be a place where people truly enjoy working and can thrive professionally.


About The Role


The Tech Lead Manager (TLM), Site Reliability is the single technical and people leader for a small, high-ownership Site Reliability team (3–4 engineers) focused on building reliable, scalable, and secure environments that power the applications our students and schools rely on. This is a pivotal role in executing our strategy to accelerate product development and reduce barriers to software delivery without sacrificing quality and reliability.

This is not a traditional Engineering Manager role with technical fluency as a nice-to-have — the TLM is expected to be hands-on with code, architecture, and technical decision-making on a regular, ongoing basis, in addition to owning the full scope of people leadership: 1:1s, growth, performance, and hiring. Where a typical EM posting emphasizes strategy, stakeholder management, and guiding technical direction through others, the TLM personally takes on inbound coding work, personally owns technical debt, and personally carries enough depth in every project their engineers are running to challenge and improve the technical approach — not just track progress against it.

There is no separate Tech Lead on this team. The TLM holds both halves of the job — technical decision making and people management — as one accountable owner.

Success for the team is measured by reliability and performance metrics for the overall product and delivery systems, reduced incident detection and recovery times, adoption rates of platform tools and services, increased product development team self-service, modernization efforts across a portfolio of legacy services and infrastructure, and reduction of recurring toilsome or interrupt activities.


What You Will Do

  • Take on inbound SRE and platform work directly rather than delegating it by default — you are a working contributor to incident response, infrastructure automation, and tooling, not just their reviewer
  • Own technical debt and infrastructure modernization within your team's scope, prioritizing and resolving it directly rather than routing it to someone else
  • Maintain deep enough understanding of every project your engineers are running — observability, disaster recovery, infrastructure work — to push on it, not just check in on it
  • Own delivery outcomes and the team's quality bar: reliability and performance metrics, incident response effectiveness, and platform tool adoption
  • Guide technical direction through system design, architectural consult, and hands-on code review, ensuring engineering excellence
  • Own people leadership for your team: career development, performance management, hiring, and overall team health for 3–4 Site Reliability Engineers
  • Own — jointly with engineering leadership — how the team works. Today that's Agile Scrum or Kanban; you're expected to give feedback and help shape and implement our development process
  • Ensure the team balances urgency over perfection, while maintaining high standards for quality and reliability
  • Model effective AI-assisted development in your own hands-on work and set the standard for your team's usage

During the First Month:

  • Learn about Niche by meeting with various team members to learn more about our company through our Onboarding meetings
  • Get hands-on in the codebase and infrastructure your team owns — not just reading it, but shipping something
  • Build rapport with your direct reports and understand each ongoing project (observability, disaster recovery, infrastructure work) deeply enough to have a real technical opinion on it
  • Work alongside the team to learn the tech stack, the products you support, and the team's current process (ceremonies, planning, review, oncall, incident response) as it exists today.  

Within 3 Months:

  • Be a regular, direct contributor to inbound technical work — incident response, platform tooling, technical debt, infrastructure — alongside your team, not just reviewing their output
  • Foster a culture of quality by providing thoughtful, constructive feedback through code reviews and pairing, and promoting adoption of platform tools and practices
  • Make your first joint call with engineering leadership on a process adjustment and technical improvement, however small
  • Have established a working rhythm for 1:1s, performance conversations, and career development with each direct report
  • Join the on-call rotation suggesting improvements to improve response time and repeated notifications

Within 6 Months:

  • Demonstrate ownership of a nontrivial technical decision or architectural direction for your team's systems (e.g., observability platform adoption, disaster recovery improvements, infrastructure)
  • Show a track record of catching and correcting technical or platform issues before they became delivery problems, including participating in disaster recovery and incident response exercises
  • Complete a full performance/calibration cycle for your directs

Within 12 Months:

  • Have led — jointly with engineering leadership — a deliberate evolution of your team's process, grounded in what's actually worked for your team rather than an inherited default
  • Be a recognized technical authority for Site Reliability, trusted with architectural calls without needing to route them elsewhere
  • Have full confidence navigating the Niche codebase and triaging technical work, getting directly involved as needed
  • Have a demonstrated record of balancing hands-on technical contribution with sustained people leadership — neither crowding out the other

What We Are Looking For

  • Strong, current hands-on software engineering experience — this role requires writing and reviewing code regularly, not occasionally
  • 4+ years of professional software engineering experience
  • 3+ years of professional platform engineering experience incorporating DevOps and Site Reliability Engineering principles
  • Prior experience owning technical direction, or people management, and a clear readiness to take on both at once
  • Comfort operating without a separate Tech Lead to lean on — you are the final technical word for your team
  • Experience with distributed systems and technologies such as AWS, Kubernetes (EKS), kops, HashiCorp Vault, Grafana, GitHub Actions, Postgres, Kafka/MSK, Redis/Valkey, or equivalent platforms (Azure/GCP also welcome)
  • Understanding of modern microservice-based architectures and methodologies
  • Experience with Golang, TypeScript, and React is a plus
  • Experience with or strong aptitude for AI-assisted development tooling and workflows
  • A track record of making pragmatic technical trade-offs under real delivery pressure
  • Experience with, or openness to evolving, agile/Scrum-based team processes
  • Excellent collaboration and communication skills, both verbal and written

Compensation 

Our national target base salary range is $146,000-$183,000, plus participation in our Annual Bonus and Stock Option Program. Base compensation will be commensurate with experience and skills.

At Niche, our Total Rewards Philosophy is centered around creating a workplace environment that attracts, motivates, and retains top talent by providing a comprehensive and competitive rewards package.  This philosophy is built on the principles of performance-based compensation, best-in-class benefits and work-life balance, and employee well-being.

Interview Process

Candidate experience is a top priority for our talent and hiring teams.  We believe in providing a transparent, authentic and comprehensive interview process where you have the opportunity to learn about us while we get to know you and your experience.  The interview process is outlined here:

  • Phone Screen with Talent Acquisition Partner - 30 Minutes

  • Video Interview with Hiring Manager - 45 Minutes

  • Technical Assessment - 45 Minutes
  • Team Interview - 45 Minutes

  • Leadership Interview - 45 Minutes

Why Niche?

  • We are a fully flexible workforce empowering our employees to choose to work remotely, in our Pittsburgh office or whatever combination suits you
  • Full time, salaried position with competitive compensation in a fast-growing company
  • Best-in-class 100% paid employee health plan, including vision and dental and supplemental coverage
  • Flexible Paid Time Off Policy
  • Stipend that allows you to build your work from home office in a style and function that suits your personal preferences
  • Parental leave for all employees (12 weeks fully paid) in addition to short term disability for birthing parents
  • Meaningful 401(k) with employer match
  • Your ideas and work will make an immediate impact on our company and millions of users
  • You will join a team that cares about you, our mission, our work - and celebrates our wins together!

Niche will only employ those who are legally authorized to work in the United States without sponsorship now or in the future for this opening.  

We are currently hiring in states where we currently have employees: AZ, CO, CT, DE, FL, GA, IL, IN, KY, LA, ME, MD, MA, MI, MO, NE, NV, NH, NJ, NY, NC, OH, OK, OR, PA, SC, TN, TX, VA, WA, DC, WV.

Candidates only.  No recruiters or agencies, please. Sorry, we do not offer relocation assistance.

Niche is an equal opportunity employer committed to fostering an inclusive, innovative environment with the best employees. Therefore, we provide employment opportunities without regard to age, race, creed, color, national origin, ancestry,  marital status, affectional or sexual orientation, gender identity or expression, disability, nationality, sex, or any other protected status in accordance with applicable law.

All interviews are being held remotely. If there are preparations we can make to help ensure you have a comfortable and positive interview experience, please let us know.

Similar Jobs

7 Minutes Ago
Remote
USA
100K-140K Annually
Senior level
100K-140K Annually
Senior level
Healthtech • Pet
Own clinic communications, customer case studies, engagement programs, pet-parent educational content, partner resources, and scalable content processes. Collaborate with Product, Customer Success, Growth, sales, veterinary experts, and external partners to create and distribute content across products and marketing channels. Manage video production, vendors, budgets, editorial workflows, and performance measurement while strengthening Vetcove’s brand presence.
Top Skills: CRMMobile PlatformsVideo Production
18 Minutes Ago
Easy Apply
Remote
United States
Easy Apply
71K-121K Annually
Entry level
71K-121K Annually
Entry level
Cloud • Security • Software • Cybersecurity • Automation
Supports revenue operations by maintaining go-to-market policies, rules of engagement, documentation, and data. Responds to policy questions, assists with dispute resolution, applies established policies, identifies process gaps, and coordinates with sales, compensation, strategy, and systems teams. This remote entry-level role requires careful data review, organization, clear communication, and consistent policy application.
Top Skills: SalesforceSalesforce ChatterSlack
26 Minutes Ago
Remote
US
Senior level
Senior level
Information Technology • Consulting
Owns projects that build and standardize revenue accounting processes, controls, documentation, and system requirements. The role investigates data and process issues, translates accounting needs for Data, Engineering, Finance, and Product teams, and drives cross-functional initiatives to completion. It requires strong US GAAP and ASC 606 knowledge, project management, process design, data analysis, and ERP or revenue recognition systems experience.
Top Skills: ChatgptClaudeExcelNetSuiteRilletSQLZuora

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account