Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer

$230k - $250k

Forward Networks Inc

Site Reliability Engineer

Forward is transforming how the world's most complex networks are managed and secured. Founded in 2013 by four Stanford Ph.D.s, we built the industry's first network digital twin — a mathematically precise model of the production network that gives IT teams unmatched visibility, verification, and agility across every major cloud and vendor environment.

Our customers include global leaders such as Goldman Sachs, PayPal, S&P Global, IBM, and Dell, as well as fast-growing enterprises and government agencies. According to IDC, Forward customers realize an average of $14.2 million in annual benefits through improved efficiency and security.

Backed by world-class investors including Andreessen Horowitz, Goldman Sachs, MSD Partners, and Threshold Ventures, Forward offers a people-centric, innovative culture where brilliant minds are shaping the future of network reliability, security, and AI-ready operations. Forward is looking for a Site Reliability Engineer.

About the Role

This is not a "keep the lights on" SRE role. As our first or early SRE hire you will be building the reliability engineering function at Forward — defining how we think about availability, observability, incident response, and operational excellence across a complex, distributed SaaS platform. You will work closely with engineering, infrastructure, and product to ensure our platform meets the reliability bar our enterprise customers demand.

If you thrive in environments where you're handed a problem rather than a playbook this role is for you.

What You'll Own
  • Define and drive SRE practices from the ground up — SLOs, SLIs, error budgets, and the frameworks the engineering org will actually use
  • Drive the reliability and operational excellence of the Forward SaaS platform
  • Build and maintain observability infrastructure — logging, metrics, tracing, and alerting — so the team always knows what's happening before customers do
  • Lead incident response: on-call rotations, runbooks, post-mortems, and the follow-through to make sure the same incident doesn't happen twice
  • Partner with engineering teams to embed reliability thinking into the SDLC — capacity planning, load testing, chaos engineering, and production readiness reviews
  • Help define and build the SRE team as the company scales — this is a foundational hire with a path to leadership
What We're Looking For
  • 6+ years of experience in site reliability engineering, DevOps, or infrastructure engineering in a SaaS or cloud environment
  • Proven experience building or significantly maturing an SRE function — not just operating within one someone else built
  • Strong fundamentals in networking — TCP/IP, DNS, routing, switching, firewalls, and load balancing. Experience with network management or observability platforms is a significant plus
  • Hands-on experience with Kubernetes and container orchestration in production environments
  • Deep proficiency with observability tooling — Prometheus, Grafana, Datadog, Splunk, or similar
  • Strong scripting and automation skills in Python, Bash, or similar
  • Experience with cloud platforms — AWS, GCP, or Azure — including infrastructure as code (Terraform, Ansible, or equivalent)
  • Track record of owning and improving incident response processes including blameless post-mortems and SLO-driven reliability improvements
  • Ability to communicate clearly with both engineering teams and non-technical stakeholders — you can explain an outage to a customer-facing team without jargon and explain an SLO to an executive without losing them
Nice to Have
  • Experience supporting enterprise or federal government customers with high availability requirements
  • Experience in a foundational or early SRE hire capacity at a growth stage company
What This Role Is Not
  • A pure ops or NOC role — you are building and engineering, not just monitoring
  • A siloed function — you will be deeply embedded with product and engineering teams
  • A ticket-taker — you will be proactively identifying and solving reliability problems before they become incidents
Why Forward
  • You'll be building something from scratch at a company with real enterprise traction and world-class investors behind it
  • Our customers include some of the most complex network environments on the planet — the reliability bar is high and the work is genuinely interesting
  • People-centric culture built by Stanford Ph.D.s who care deeply about doing things the right way
  • Competitive compensation, equity, and the opportunity to grow into a leadership role as the SRE function scales

The base pay range for this role is between $230,000 and $250,000. Base pay will depend on your skills, qualifications, experience, and location

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in Santa Clara, CA vacancy
  •  ...that keep the world running. Location: 5 on-site days a week in Sunnyvale, CA Headquarters. Our Team's Vision: Our Engineering team is shaping the future of cybersecurity...  ...are looking for an experienced Senior Site Reliability Engineer (SRE) with a strong background in... 
    Suggested
    Full time
    Work experience placement
    Immediate start

    Illumio

    Sunnyvale, CA
    2 days ago
  • $145k - $165k

     ...Your Ego : Selflessly collaborate towards our shared purpose. About the role Bolt Graphics is seeking a highly experienced Site Reliability Engineer (SRE) to design, build, and operate highly reliable developer and production systems. This role is mission-critical to... 
    Suggested
    Work at office
    Immediate start

    Bolt Graphics, Inc.

    Sunnyvale, CA
    4 days ago
  • $126.5k - $182k

     ...collaborate with a global team of software engineers and SREs responsible for delivering...  ...product, and operations partners to ensure reliability and performance. Webex is powering...  ...technology leader Your impact As a Site Reliability Engineer (SRE) supporting backend... 
    Suggested
    Permanent employment
    Full time
    Temporary work
    Local area
    Worldwide
    Flexible hours
    Shift work

    Cisco

    San Jose, CA
    2 days ago
  •  ...Site Reliability Engineer (SRE) The successful applicant may be performing work in FedRAMP High or IL-5 environments, and therefore, must be a U.S. Person (i.e. U.S. citizen, U.S. national, lawful permanent resident, asylee, or refugee). This position may also perform... 
    Suggested
    Permanent employment
    Worldwide
    Shift work

    Webex Events (formerly Socio)

    Santa Clara, CA
    1 day ago
  •  ...Role: Site Reliability Engineer (SRE) Location: Santa Clara Valley (Cupertino), California, Hybrid. Duration: 6+ Months Job Description: Deploy, support and monitor new and existing services, platforms, and application stacks.... 
    Suggested

    Zortech Solutions

    Santa Clara, CA
    1 day ago
  • $159.2k - $301.6k

     ...running Graphs on the cloud. In this reliability-focused role, you will own the availability...  .... You'll partner with the backend engineers building these APIs to make sure the system...  ...Science. ~5-10 years of experience in site reliability engineering, infrastructure,... 
    Temporary work
    Local area
    Worldwide

    Adobe

    San Jose, CA
    10 hours ago
  • $150k - $195k

     ...customers worldwide. Our team is growing, and we are looking for engineers with passion for automation. You will help support the...  ...alongside engineering/operations teams to improve the scalability and reliability of internal processes. Participate in an on‑call rotation.... 
    Full time
    Worldwide

    Fortinet

    Sunnyvale, CA
    2 days ago
  • $65 - $85 per hour

     ...Site Reliability Engineer Sustainable Talent is partnering with a global leader who's been transforming computer graphics, PC gaming, and accelerated computing for over 25 years. We are looking for a Site Reliability Engineer to support our client's team based out of... 
    Full time
    Contract work
    Worldwide

    Sustainable Talent

    Santa Clara, CA
    10 hours ago
  • $150k

     ...S23 - Telecommunications Role - Site Reliability Engineer Location: Santa Clara, CA or Wall Township, NJ (Hybrid) Salary: $150,000 + Bonus + Comprehensive Benefits About the Role Join a team building and operating a modern Kubernetes-based platform... 

    S23 Recruitment

    Santa Clara, CA
    1 day ago
  •  ...Qualifications: 8+ years of software engineering experience, or equivalent demonstrated through...  ...implement and maintain scalable and reliable infrastructure on Google Cloud Platform...  ...vendor resources. Willingness to work on-site at stated location in the job opening.... 
    For contractors
    Work experience placement

    Cedent Life Talent

    San Jose, CA
    2 days ago
  •  ...Site Reliability Engineer Foxconn Industrial Internet (Fii), is a world leading professional design and manufacturing service provider of communication network equipment, cloud service equipment, precision tools and industrial robots. FII provides customers with intelligent... 
    Permanent employment
    Full time
    Work at office
    Local area

    Foxconn Industrial Internet

    San Jose, CA
    4 days ago
  •  ...AWS Infra SRE/DevOps Engineer AWS Infra SRE/DevOps engineer with proven work experience ensuring reliability, availability and performance of cloud infra and platform. Specialist on Cisco Cloud run-on for infrastructure management, who can install, run, and maintain... 
    Work experience placement

    The Dignify Solutions, LLC

    San Jose, CA
    1 day ago
  •  ...Senior Site Reliability Engineer LeanData helps the world's fastest-growing companies automate, simplify, and accelerate revenue. We are looking for a Senior Site Reliability Engineer to lead the strategic evolution of our cloud infrastructure. Reporting directly... 
    Full time
    Work at office
    Flexible hours
    2 days per week

    LeanData

    Santa Clara, CA
    2 days ago
  •  ...Site Reliability Engineer Sunnyvale, CA Site Reliability Eng. Must have LinkedIn profile. • t least 8+ years in a Reliability Engineering, DevOps or infrastructure focused role • Strong AWS and Linux Operating System, standard networking protocols, component... 

    Q1 Technologies

    Sunnyvale, CA
    4 days ago
  •  ...work from home day is currently Tuesday.Engineering at Lambda is responsible for building and...  ...and networking teams to improve service reliability and deployment workflowsDeploy and...  ...rotationYouHave 5+ years of experience in Site Reliability Engineering, Production Engineering... 
    Part time
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    2 hours ago
  •  ...Senior Site Reliability Engineer ILocationSan Jose, Costa Rica - RemoteSummary of roleOwn availability, the most important product feature, by continually striving for sustained operational excellence of Sumo’s planet-scale observability and security products. Work with... 
    Part time
    Flexible hours

    Sumo Logic

    San Jose, CA
    2 hours ago
  • $185k - $278k

     ...transforming them into scalable solutions. * Debugging OS and engineering issues within our provided Linux environment. *...  ...efficiently. The Impact You Will Have: * Enhancing the reliability and performance of our engineering environment. * Streamlining... 
    Remote work

    Synopsys

    Sunnyvale, CA
    2 days ago
  • $180k - $260k

     ...effortless integration into customers' logistics operations. About the role We are seeking an experienced Senior/Staff Site Reliability Engineer to support the operation, monitoring, and scaling of our growing fleet of autonomous vehicles. In this role, you will work... 
    Odd job
    Work at office
    Remote work

    Gatik AI

    Santa Clara, CA
    10 hours ago
  • $174k - $253k

    Bachelor’s degree in Computer Science, Engineering, a related field, or equivalent practical experience. 5 years of experience with...  ...'s degree in Computer Science or Engineering. About The Job Site Reliability Engineering (SRE) is what you get when you treat operations... 

    Google

    Sunnyvale, CA
    1 day ago
  • $210.6k - $305.1k

     ...Minimum Qualifications:  You have led a distributed team of 5+ engineers, can demonstrate strong technical vision for your team, and ensure...  ..., and basic life insurance. Please see the Cisco careers site to discover more benefits and perks. Employees may be eligible... 
    Full time
    Temporary work
    Part time
    Local area
    Flexible hours

    CISCO Systems

    San Jose, CA
    2 hours ago
  • $207.4k - $259.2k

     ...differences, and supports and celebrates all of our team members.We are seeking a highly experienced and passionate Sr. Staff Site Reliability Engineer (SRE) to join our growing team. In this critical role, you will be responsible for the reliability, scalability,... 
    Permanent employment
    Part time
    Local area

    Archer Aviation

    San Jose, CA
    2 hours ago
  • $184k - $287.5k

     ...At NVIDIA, Site Reliability Engineering provides a rare chance to define, develop, and support large-scale production systems with high efficiency and availability. This demanding position merges software and systems engineering efforts to guarantee flawless service operation... 
    Full time
    Part time

    Nvidia

    Santa Clara, CA
    2 hours ago
  • $100k - $200k

     ...OPPO US Research Center is seeking a skilled and proactive Site Reliability Engineer (SRE) to join our team. In this role, you will be responsible for ensuring the stability, scalability, and performance of our application systems. The ideal candidate is passionate about... 
    Full time

    OPPO

    Palo Alto, CA
    4 days ago
  • $189k - $232k

     ...Site Reliability Engineer Mountain View, US About EarnIn As one of the first pioneers of earned wage access, our passion at EarnIn is building products that deliver real-time financial flexibility for those with the unique needs of living paycheck to paycheck.... 
    Full time
    Work at office
    2 days per week

    Earnin

    Mountain View, CA
    4 days ago
  •  ...Site Reliability Engineer III There's nothing more exciting than being at the center of a rapidly growing field in technology and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. As a Site Reliability... 
    Work at office

    Chase

    Palo Alto, CA
    2 days ago
  •  ...SRE Engineer Location: Sunnyvale CA Rate: DOE Duration: 12+ Months What You Will Do: Identify, develop and execute opportunities to raise the bar on engineering & operational excellence. Providing thought leadership and defining strategy on developer... 

    Redolent

    Sunnyvale, CA
    2 days ago
  •  ...Job Description Job Description About the Role Senior Site Reliability Engineer (Payments Infrastructure) Kody is seeking a Senior Site Reliability Engineer to ensure the reliability, availability, scalability, and operational excellence of our global payment platform... 

    Kody

    Sunnyvale, CA
    11 days ago
  •  ...design by customizing MES tool per business needs Education Requirements, Ideal Experience: Associate’s degree in Industrial Engineering or IT related field Minimum of 0-3 years’ relevant experience Experience in C#, Delphi desired Knowledge of the... 
    Work at office

    Foxconn Industrial Internet - FII

    Sunnyvale, CA
    27 days ago
  •  ...of Huobi globe spanning infrastructure. •       Work with engineering teams to make sure new features and changes are deployed quickly...  .... •       Constantly improve our system performance and reliability through better tools, process and monitoring system. •... 
    Worldwide

    Cryptoware Technologies Inc

    Santa Clara, CA
    6 days ago
  • $180k - $200k

     ...Holmdel, NJ. Join us and be part of a team that's shaping the future of payments—one experience at a time. As our Site Reliability Engineer, you will design, build, and maintain the systems and infrastructure that power our applications, ensuring their... 
    For contractors
    Work at office
    Work from home
    Flexible hours

    PayNearMe, Inc.

    Santa Clara, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!