Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Manager, Site Reliability Engineering

Form Energy, Inc.

Are you ready to build America's energy future? Form Energy is an American manufacturing and energy technology company. We're revolutionizing energy storage with cost-effective, multi-day technology designed to keep the electric grid secure and reliable, even during extended periods of stress. By strengthening the electric system and reimagining what's possible, we're giving clean energy a whole new form!

In recent years, Form Energy has earned a number of accolades, including being named by TIME as a "Best Invention", MIT Technology Review as a "Top Climate Tech Company To Watch", and Fast Company as "One of the Next Big Things In Tech". We are making rapid progress on our mission of delivering energy storage for a better world, and our team is growing just as rapidly to meet demand. We have signed contracts with leading electric utilities across the United States and production of our iron-air batteries is underway at our first high-volume manufacturing facility in West Virginia.

Working for Form Energy is more than just a job, it's a chance to be part of something extraordinary. And now - right as we significantly scale up battery manufacturing - might be the most exciting moment in the company's history to join. We are assembling a team of highly talented and driven individuals across the country. Driven by our core values of humanity, excellence, and creativity, our team is determined to deliver on our mission and transform the energy landscape for the better.

Feeling energized to make a meaningful impact on the world? Then keep reading - you've come to the right place.

Role Description

Form Energy is hiring a Manager, Site Reliability Engineer to lead the operational function responsible for maintaining the reliability, availability, and supportability of our deployed energy storage systems. This role will own the mechanisms used to monitor fleet health, respond to incidents, triage field issues, coordinate engineering escalation, and ensure new products and releases are operationally ready. As Form Energy moves from early deployments to a growing commercial fleet, a key objective of this role is to scale operational capacity through automation, tooling, and product design improvements. The successful candidate will work across software, firmware, controls, hardware, field service, and customer-facing teams to build the scalable processes and capabilities required to enable reliable fleet operations.

Relocation assistance is available.

What you'll do:
  • Lead Product Operations Assurance with a DevOps first principals approach to fleet monitoring, incident response, field triage, engineering escalation, and production support.

  • Build a highly automated operating model that enables the team to support a rapidly growing deployed fleet, managing operational work through automation and driving product design improvements that reduce sustaining engineering overhead.

  • Establish incident management processes, including severity definitions, escalation paths, incident command, communications, and post-incident reviews.

  • Coordinate cross-functional engineering response to complex issues spanning software, firmware, controls, networking, hardware, and site infrastructure.

  • Develop diagnostic playbooks, troubleshooting procedures, and operational tooling that improve first-line response and reduce dependence on individual experts.

  • Partner with Data, Analytics, and Cloud Applications teams to define the telemetry, dashboards, alerts, and workflows required to effectively monitor and support the fleet.

  • Define operational readiness requirements for new product releases and deployments, including monitoring, diagnostics, recovery procedures, escalation paths, and support documentation.

  • Analyze incidents and fleet data to identify recurring failure modes and drive reliability, diagnosability, and serviceability improvements back into the product.

  • Establish and track operational metrics such as fleet availability, incident frequency, time to detection, time to containment, time to recovery, and recurrence.

  • Build the team, processes, on-call model, and automation required to scale fleet operations as the installed base grows.

What you'll bring:
  • 12+ years of experience supporting complex production, industrial, energy, infrastructure, automotive, robotics, or other cyber-physical systems, including technical leadership or people management experience.

  • Demonstrated experience leading production incident response, technical troubleshooting, escalation management, and root-cause investigation in deployed systems.

  • Experience leading SRE, DevOps, or production operations teams in highly automated environments, with a track record of scaling operational capacity through software, tooling, and product improvements.

  • Strong ability to diagnose and coordinate resolution of problems spanning software, firmware, controls, networking, hardware, and field operations.

  • Experience with operational monitoring, observability, alerting, remote diagnostics, reliability practices, and the operational processes needed to support deployed products.

  • Proven ability to lead cross-functional teams through high-priority technical issues, communicate risk and recovery plans clearly, and translate operational learnings into improvements in product reliability, diagnosability, and serviceability.

  • Bachelor's degree in engineering, computer science, or a related technical discipline, or equivalent practical experience.

#LI-TR1

Humanity is a cornerstone of Form Energy's culture, and we make sure our compensation and benefits reflect that. Form Energy offers competitive salaries, stock options, and a holistic benefits package to ensure all employees have what they need to thrive while working here.

When it comes to you and your family's health, we cover 100% of medical, dental, and vision premiums for full-time employees - and 80% of healthcare premiums for dependents. This starts from day one. We also offer at least 12 weeks of paid leave for new parents (up to 20 weeks for birthing parents), and generous vacation policies to give employees time to recharge when needed.

To build America's energy future, we need everyone at the table. We are proud to be an equal opportunity employer, and encourage candidates from all backgrounds to apply to our open jobs.

If you may require reasonable accommodations to participate in our interview process, please contact View email address on click.appcast.io. Requests for accommodations will be treated with discretion.
Form Energy is committed to maintaining the privacy of our applicants. Please be aware that we will never solicit sensitive personal information such as Social Security numbers or bank account details during the recruiting or hiring process.

#J-18808-Ljbffr
Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Manager, Site Reliability Engineering in Berkeley, CA vacancy
  • $210.38k - $243.21k

    Manager, Site Reliability Engineer (Hybrid in South San Francisco)About the RoleWe are seeking an experienced and hands-on Site Reliability Engineering (SRE) Manager to lead our Site Operations and infrastructure initiatives. This role is responsible for ensuring the reliability... 
    Suggested

    Twist Bioscience

    San Francisco, CA
    4 days ago
  • $150k - $220k

     ...innovators in this way. The Role: As an engineering organization, we pride ourselves on...  ...engineering as a creative activity. Engineering managers enable engineers to do their best work...  ..., mastery, and purpose. The Manager, Site Reliability Engineering will lead Forge’s SRE team... 
    Suggested
    Local area

    Forge Global

    San Francisco, CA
    1 day ago
  • $15k

     ...and worked at the frontier of applying AI/ML to investment management. We have become a multibillion-dollar asset manager, and...  ...office, daily catered lunches, and more.As a Senior Cluster Site Reliability Engineer (SRE), you will help scale our research compute cluster to... 
    Suggested
    Work at office
    Local area
    Remote work

    The Voleon Group

    Berkeley, CA
    4 days ago
  • $80 per hour

     ...possible, and we need a sharp, self-motivated SRE to help keep that engine running without interruption. If you love solving real problems...  ..., Prometheus/VictoriaMetrics, Alertmanager, and building management/cooling systems Network security fundamentals (ACLs, firewalls... 
    Suggested
    Contract work
    Shift work

    LTD Global, LLC

    Berkeley, CA
    1 day ago
  •  ...Center (NERSC) is inviting applications for the position of Site Reliability Engineer. NERSC's mission is to accelerate scientific discovery...  ...advanced data collection and monitoring systems to proactively manage the health of our environment. Ultimately, your work ensures... 
    Suggested
    Work at office
    Night shift

    Bay Systems Inc

    Berkeley, CA
    1 day ago
  •  ...org/vision. Position Summary We are looking for a Site Reliability Engineer to own the digital infrastructure that powers our research...  ...automatic scaling of compute resources based on demand. Access Management: Ensure the right people have access to the right... 
    Visa sponsorship

    Astera

    Emeryville, CA
    1 day ago
  • $232.34k - $290.42k

     ...access to data as simple and reliable as electricity. With Fivetran...  ...and ready to query, with no engineering or maintenance required. We’re...  ...our teams, systems, and career sites. About the Role...  ...with engineering teams, product managers, as well as support and sales... 
    Full time
    Work at office
    Remote work
    Flexible hours

    Fivetran

    Oakland, CA
    2 days ago
  •  ...technology designed to keep the electric grid secure and reliable, even during extended periods of stress. By...  ...right place. Role Description Form Energy is hiring a Manager, Site Reliability Engineer to lead the operational function responsible for maintaining... 
    Full time
    Remote work
    Relocation package

    Form Energy, Inc.

    Berkeley, CA
    2 days ago
  • $75 per hour

     .... For this role, you’ll join the team supporting NERSC, where reliable computing systems help scientists advance research in energy,...  ...and operational workflows. Use ServiceNow to support service management workflows and customized solutions. Requirements... 
    Hourly pay
    Contract work
    Night shift

    Essnova Solutions, Inc.

    Berkeley, CA
    8 days ago
  •  ...builds the platforms and tooling that help engineering teams develop, deploy, and operate...  ...default for every product team.As a Staff Site Reliability Engineer on Release Engineering, you'll...  ...habits and tooling.Architect and manage the SLO and error-budget framework, empowering... 
    Permanent employment
    Work experience placement
    Work at office
    Local area

    Plaid Financial

    San Francisco, CA
    17 hours ago
  • About the RoleWe’re looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability, observability, and...  ...cross-functionallyNice-to-HaveExperience with cloud and managed services (e.g. AWS)Experience supporting data-intensive platforms... 

    Alembic

    San Francisco, CA
    1 day ago
  • $148.5k - $223.9k

     ...future of Salesforce.Salesforce is seeking a senior engineering candidate to join the Site Reliability organization in San Francisco. Working closely with...  ...eliminate toil and improve operational efficiency.Incident Management: Lead the coordinated response to incidents as an... 
    Full time
    Worldwide
    Weekend work

    Salesforce

    San Francisco, CA
    17 hours ago
  •  ...the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Enterprise Technology,...  ...exposure (training can be provided)Cloud/SaaS experienceMemory management and dump analysis (Java heap dump analysis preferred)ITSM/... 

    JP Morgan Chase

    San Francisco, CA
    3 days ago
  • $152.5k - $205k

     ...a stakeholder.What you’ll be responsible for:As a Senior Site Reliability Engineer on Circle’s platform team, you’ll design, build, and operate...  ...experience, including authoring reusable modules, managing state and environments, and delivering infrastructure changes... 
    Flexible hours

    Circle

    San Francisco, CA
    1 day ago
  •  ...principles to see it in full.About the teamThe Engineering team at Airwallex is a diverse group of...  ..., working together to build scalable, reliable, and secure products that empower...  ...Global services.What you’ll doAs a Senior Site Reliability Engineer, you’ll work closely... 
    Temporary work
    Local area

    Airwallex

    San Francisco, CA
    1 day ago
  • $190.8k - $267.1k

     ...while helping Reddit grow its business. The reliability of our Ads systems directly impacts...  ...Reliability team partners closely with Ads Engineering to improve reliability, scalability,...  ...advertiser trust. We’re looking for a Senior Site Reliability Engineer to build, operate,... 
    For contractors
    Work experience placement

    Reddit

    San Francisco, CA
    1 day ago
  • $152.5k - $205k

     ...everyone is a stakeholder.What you’ll be responsible forThe Site Reliability Engineer builds and maintains shared platform capabilities, common...  ..., observability, access controls, auditability, and cost management. You will troubleshoot production issues, document operational... 
    Flexible hours

    Circle

    San Francisco, CA
    1 day ago
  • $106k - $130k

     ...ineligible for employment Visa sponsorship.Role Summary The Senior Site Reliability Engineer applies software engineering and systems engineering...  ...as Code, automation, testing, incident response, capacity management, resilience, and operational readiness. Identify recurring... 
    Hourly pay
    Full time
    Immediate start
    Visa sponsorship
    Work visa
    Flexible hours

    Early Warning

    San Francisco, CA
    2 days ago
  • $139.76k - $287.75k

     ...business.We are seeking a Senior Site ReliabilityEngineer to help...  ...in advancing the reliability, scalability, automation, observability...  ...candidate is a highly hands-on engineer with strong production...  ...infrastructure provisioning and change management through Terraform/... 
    Work at office
    Local area
    Relocation
    Relocation package

    Pinterest

    San Francisco, CA
    17 hours ago
  • $127k - $249k

    Platform Engineering is the department within SRE that is responsible for a range of critical...  ...and alerting systems.The Fleet Management team provides the core runtime environment...  ...critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager... 
    Work at office
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    San Francisco, CA
    3 days ago
  • $127k - $249k

    The TeamPlatform Engineering sits within SRE and builds the core infrastructure...  ...Engineering, the Fabric team manages the global network substrate...  ...role in engineering the reliable, globally connected, multi-...  ...seeking a talented Senior Site Reliability Engineer (SRE) with... 
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    San Francisco, CA
    2 days ago
  • $113.4k - $162k

     ...conversation for people everywhere.TextNow is looking for motivated Site Reliability Engineer to own infrastructure, monitoring, logging, ci/cd,...  ...GitHub, Terraform, Ansible, or similar tools to build and manage cloud infrastructure efficiently.Incident Management Expert... 
    Temporary work

    TextNow

    San Francisco, CA
    4 days ago
  • $117k - $209.33k

     ...Position OverviewWant to help make a better world? As a Senior Site Reliability Engineer at Autodesk, you can help us build and operate reliable,...  ...such as SLOs/SLIs, production readiness, incident management, observability, resilience testing, and toil reduction. Success... 
    Full time
    For contractors

    Autodesk

    San Francisco, CA
    2 days ago
  • $194k - $267k

     ...Overview:We are seeking a highly technical StaffObservabilitySite Reliability Engineer with a specialty in Splunk to own and evolve our Splunk...  ....Required Skills & Experience (The Essentials)Log Management: Minimum 5+ Experience scaling and managing Splunk Cloud at... 
    Permanent employment
    Work at office
    Local area
    Worldwide
    Flexible hours

    Okta

    San Francisco, CA
    17 hours ago
  • $195k - $257.5k

     ...is a stakeholder.What you’ll be responsible for:As a Staff Site Reliability Engineer on Circle’s Platform team, you’ll design, build, and...  ...on:Operate and scale production blockchain infrastructure, managing full nodes across networks such as Arc, Ethereum, Solana,... 
    Flexible hours

    Circle

    San Francisco, CA
    4 days ago
  • $204k - $306k

     ...We're all in on this mission. If you are too, let's talk.Manager, Site Reliability EngineeringSan Francisco, CaliforniaSecure Every Identity,...  ...week in our San Francisco Office. The IDaaS Site Reliability Engineering GroupOkta authenticates, authorizes and provisions... 
    Permanent employment
    Work at office
    Local area
    Worldwide
    Flexible hours
    2 days per week

    Okta

    San Francisco, CA
    1 day ago
  • $181k - $263k

     ...line operational support. We are looking for a Senior Staff Site Reliability Engineer who will set the technical direction for reliability...  ...expertise: internals, autoscaling, multi-tenant workload management, and rightsizingAdvanced experience with real-time and NoSQL... 
    Worldwide

    LiveRamp

    San Francisco, CA
    1 day ago
  • $55k - $151.47k

     ...LevelSenior AssociateJob Description & SummaryThe OpportunityAs a Site Reliability Engineer - Senior Associate, you will play a pivotal role in...  ...data integrity and accessibility- Leading incident management and resolution efforts to maintain operational continuityWhat... 
    Full time
    H1b

    PwC

    San Francisco, CA
    1 day ago
  • $194k - $267k

     ..., automate it” and who can rapidly self-educate on new concepts and tools.Position Overview:The Site Reliability Engineer (SRE) will play a key role in building and managing Kubernetes platforms that support cloud-native applications and services. This position focuses... 
    Permanent employment
    Work at office
    Local area
    Worldwide
    Flexible hours

    Okta

    San Francisco, CA
    2 days ago
  • $153k - $191.3k

     ...manufacturing, data processing, and software engineering, our office is a truly inspiring mix of...  ...environments, to guarantee the reliability, scalability, and availability of our services...  ..., particularly resource optimization, management, and cluster tuning in a constrained... 
    Full time
    Temporary work
    For contractors
    Work at office
    Local area
    Remote work
    Home office
    3 days per week

    Planet Labs PBC

    San Francisco, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Manager, Site Reliability Engineering. Be the first to apply!