Form Energy Logo

Form Energy

Manager, Site Reliability Engineering

Posted 9 Days Ago
Be an Early Applicant
In-Office
Berkeley, CA, USA
170K-223K Annually
Expert/Leader
In-Office
Berkeley, CA, USA
170K-223K Annually
Expert/Leader
Leads the operational function supporting deployed energy storage systems, including fleet monitoring, incident response, field triage, engineering escalation, and production support. Builds automated processes, diagnostic tooling, on-call operations, and readiness requirements for new releases. Partners across software, firmware, controls, hardware, data, cloud, field service, and customer teams to improve reliability, availability, observability, and recovery while scaling support for a growing commercial fleet.
The summary above was generated by AI

Are you ready to build America’s energy future? Form Energy is an American manufacturing and energy technology company. We’re revolutionizing energy storage with cost-effective, multi-day technology designed to keep the electric grid secure and reliable, even during extended periods of stress. By strengthening the electric system and reimagining what’s possible, we’re giving clean energy a whole new form! 

In recent years, Form Energy has earned a number of accolades, including being named by TIME as a “Best Invention”, MIT Technology Review as a “Top Climate Tech Company To Watch”, and Fast Company as “One of the Next Big Things In Tech”. We are making rapid progress on our mission of delivering energy storage for a better world, and our team is growing just as rapidly to meet demand. We have signed contracts with leading electric utilities across the United States and production of our iron-air batteries is underway at our first high-volume manufacturing facility in West Virginia.

Working for Form Energy is more than just a job, it’s a chance to be part of something extraordinary. And now - right as we significantly scale up battery manufacturing -  might be the most exciting moment in the company’s history to join. We are assembling a team of highly talented and driven individuals across the country. Driven by our core values of humanity, excellence, and creativity, our team is determined to deliver on our mission and transform the energy landscape for the better.

Feeling energized to make a meaningful impact on the world? Then keep reading - you’ve come to the right place.

Role Description

Form Energy is hiring a Manager, Site Reliability Engineer to lead the operational function responsible for maintaining the reliability, availability, and supportability of our deployed energy storage systems. This role will own the mechanisms used to monitor fleet health, respond to incidents, triage field issues, coordinate engineering escalation, and ensure new products and releases are operationally ready. As Form Energy moves from early deployments to a growing commercial fleet, a key objective of this role is to scale operational capacity through automation, tooling, and product design improvements. The successful candidate will work across software, firmware, controls, hardware, field service, and customer-facing teams to build the scalable processes and capabilities required to enable reliable fleet operations.

Relocation assistance is available.

What you'll do:
  • Lead Product Operations Assurance with a DevOps first principals approach to fleet monitoring, incident response, field triage, engineering escalation, and production support.

  • Build a highly automated operating model that enables the team to support a rapidly growing deployed fleet, managing operational work through automation and driving product design improvements that reduce sustaining engineering overhead.

  • Establish incident management processes, including severity definitions, escalation paths, incident command, communications, and post-incident reviews.

  • Coordinate cross-functional engineering response to complex issues spanning software, firmware, controls, networking, hardware, and site infrastructure.

  • Develop diagnostic playbooks, troubleshooting procedures, and operational tooling that improve first-line response and reduce dependence on individual experts.

  • Partner with Data, Analytics, and Cloud Applications teams to define the telemetry, dashboards, alerts, and workflows required to effectively monitor and support the fleet.

  • Define operational readiness requirements for new product releases and deployments, including monitoring, diagnostics, recovery procedures, escalation paths, and support documentation.

  • Analyze incidents and fleet data to identify recurring failure modes and drive reliability, diagnosability, and serviceability improvements back into the product.

  • Establish and track operational metrics such as fleet availability, incident frequency, time to detection, time to containment, time to recovery, and recurrence.

  • Build the team, processes, on-call model, and automation required to scale fleet operations as the installed base grows.

What you'll bring:
  • 12+ years of experience supporting complex production, industrial, energy, infrastructure, automotive, robotics, or other cyber-physical systems, including technical leadership or people management experience.

  • Demonstrated experience leading production incident response, technical troubleshooting, escalation management, and root-cause investigation in deployed systems.

  • Experience leading SRE, DevOps, or production operations teams in highly automated environments, with a track record of scaling operational capacity through software, tooling, and product improvements.

  • Strong ability to diagnose and coordinate resolution of problems spanning software, firmware, controls, networking, hardware, and field operations.

  • Experience with operational monitoring, observability, alerting, remote diagnostics, reliability practices, and the operational processes needed to support deployed products.

  • Proven ability to lead cross-functional teams through high-priority technical issues, communicate risk and recovery plans clearly, and translate operational learnings into improvements in product reliability, diagnosability, and serviceability.

  • Bachelor’s degree in engineering, computer science, or a related technical discipline, or equivalent practical experience.


#LI-TR1

Humanity is a cornerstone of Form Energy’s culture, and we make sure our compensation and benefits reflect that. Form Energy offers competitive salaries, stock options, and a holistic benefits package to ensure all employees have what they need to thrive while working here. 

When it comes to you and your family’s health, we cover 100% of medical, dental, and vision premiums for full-time employees - and 80% of healthcare premiums for dependents. This starts from day one. We also offer at least 12 weeks of paid leave for new parents (up to 20 weeks for birthing parents), and generous vacation policies to give employees time to recharge when needed. 

To build America’s energy future, we need everyone at the table. We are proud to be an equal opportunity employer, and encourage candidates from all backgrounds to apply to our open jobs.


If you may require reasonable accommodations to participate in our interview process, please contact [email protected]. Requests for accommodations will be treated with discretion.
Form Energy is committed to maintaining the privacy of our applicants. Please be aware that we will never solicit sensitive personal information such as Social Security numbers or bank account details during the recruiting or hiring process.

Form Energy Berkeley, California, USA Office

2850 7th Street , Berkeley, CA, United States, 94710

Similar Jobs

Yesterday
Remote or Hybrid
United States
205K-255K Annually
Senior level
205K-255K Annually
Senior level
Software • Defense
Lead and develop the SRE team responsible for reliable, secure deployments across on-premises DoD and AWS environments. Own capacity planning, prioritization, reliability roadmaps, operational readiness, incident response, observability, and toil reduction. Coordinate delivery across engineering, security, and customer success teams while guiding infrastructure, automation, Kubernetes, CI/CD, networking, and application reliability decisions. Manage team performance, hiring, career development, on-call practices, postmortems, and stakeholder communication.
Top Skills: AnsibleAWSAws GovcloudBashCi/CdDatadogElkGitopsGoGrafanaHyper-VIcd 503Infrastructure As CodeKubernetesNode.jsNutanixProxmoxPythonRmfSlisSlosStigsTerraformTypescriptVMware
8 Days Ago
In-Office or Remote
United States
254K-313K Annually
Senior level
254K-313K Annually
Senior level
Cloud • Security • Software • Cybersecurity
Leads and mentors site reliability engineering teams supporting products across thousands of clusters and global enterprises. Drives reliability, scalability, security, usability, product lifecycle requirements, technical innovation, system design, and solutions development. Partners with stakeholders on product roadmaps and operates in an Agile environment to deliver SRE initiatives while building accountable, collaborative teams.
Top Skills: AgileAnsibleArgo CdAWSAws CdkAzureCi/Cd AutomationConfluenceDockerGitGradleInfrastructure As CodeJ2EeJavaJbossJenkinsJIRAKubernetesMavenObject-Oriented ProgrammingOracleRestful ServicesServicenowSpringSpring BootTerraformWeb ServicesWebsphereXldeployXlrelease
Yesterday
Hybrid
Mountain View, CA, USA
298K-368K Annually
Expert/Leader
298K-368K Annually
Expert/Leader
Automotive
Lead the Site Reliability Engineering team responsible for Waymo’s release and pipeline infrastructure. Own observability, incident response, capacity planning, on-call operations, software hardening, playbooks, and continuous reliability improvements. Investigate novel production events, contribute to system architecture, and partner across engineering organizations to improve large-scale software delivery systems supporting autonomous vehicles. Manage organizational growth, including hiring and developing additional engineers.
Top Skills: Deep LearningMachine Learning

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account