Site Reliability Engineer
$230k - $250kForward
Forward is transforming how the world's most complex networks are managed and secured. Founded in 2013 by four Stanford Ph.D.s, we built the industry's first network digital twin - a mathematically precise model of the production network that gives IT teams unmatched visibility, verification, and agility across every major cloud and vendor environment.
Our customers include global leaders such as Goldman Sachs, PayPal, S&P Global, IBM, and Dell, as well as fast-growing enterprises and government agencies. According to IDC, Forward customers realize an average of $14.2 million in annual benefits through improved efficiency and security. Backed by world-class investors including Andreessen Horowitz, Goldman Sachs, MSD Partners, and Threshold Ventures, Forward offers a people-centric, innovative culture where brilliant minds are shaping the future of network reliability, security, and AI-ready operations. Forward is looking for a Site Reliability Engineer About the Role This is not a "keep the lights on" SRE role. As our first or early SRE hire you will be building the reliability engineering function at Forward - defining how we think about availability, observability, incident response, and operational excellence across a complex, distributed SaaS platform. You will work closely with engineering, infrastructure, and product to ensure our platform meets the reliability bar our enterprise customers demand. If you thrive in environments where you're handed a problem rather than a playbook this role is for you. What You'll Own- Define and drive SRE practices from the ground up - SLOs, SLIs, error budgets, and the frameworks the engineering org will actually use
- Drive the reliability and operational excellence of the Forward SaaS platform
- Build and maintain observability infrastructure - logging, metrics, tracing, and alerting - so the team always knows what's happening before customers do
- Lead incident response: on-call rotations, runbooks, post-mortems, and the follow-through to make sure the same incident doesn't happen twice
- Partner with engineering teams to embed reliability thinking into the SDLC - capacity planning, load testing, chaos engineering, and production readiness reviews
- Help define and build the SRE team as the company scales - this is a foundational hire with a path to leadership
- 6+ years of experience in site reliability engineering, DevOps, or infrastructure engineering in a SaaS or cloud environment
- Proven experience building or significantly maturing an SRE function - not just operating within one someone else built
- Strong fundamentals in networking - TCP/IP, DNS, routing, switching, firewalls, and load balancing. Experience with network management or observability platforms is a significant plus
- Hands-on experience with Kubernetes and container orchestration in production environments
- Deep proficiency with observability tooling - Prometheus, Grafana, Datadog, Splunk, or similar
- Strong scripting and automation skills in Python, Bash, or similar
- Experience with cloud platforms - AWS, GCP, or Azure - including infrastructure as code (Terraform, Ansible, or equivalent)
- Track record of owning and improving incident response processes including blameless post-mortems and SLO-driven reliability improvements
- Ability to communicate clearly with both engineering teams and non-technical stakeholders - you can explain an outage to a customer-facing team without jargon and explain an SLO to an executive without losing them
- Experience supporting enterprise or federal government customers with high availability requirements
- Experience in a foundational or early SRE hire capacity at a growth stage company
- A pure ops or NOC role - you are building and engineering, not just monitoring
- A siloed function - you will be deeply embedded with product and engineering teams
- A ticket-taker - you will be proactively identifying and solving reliability problems before they become incidents
- You'll be building something from scratch at a company with real enterprise traction and world-class investors behind it
- Our customers include some of the most complex network environments on the planet - the reliability bar is high and the work is genuinely interesting
- People-centric culture built by Stanford Ph.D.s who care deeply about doing things the right way
- Competitive compensation, equity, and the opportunity to grow into a leadership role as the SRE function scales
The base pay range for this role is between $230,000 and $250,000. Base pay will depend on your skills, qualifications, experience, and location
Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in Santa Clara, CA vacancy
- ...that keep the world running. Location: 5 on-site days a week in Sunnyvale, CA Headquarters. Our Team's Vision: Our Engineering team is shaping the future of cybersecurity... ...are looking for an experienced Senior Site Reliability Engineer (SRE) with a strong background in...SuggestedFull timeWork experience placementImmediate start
$150k - $195k
...customers worldwide. Our team is growing, and we are looking for engineers with passion for automation. You will help support the... ...alongside engineering/operations teams to improve the scalability and reliability of internal processes. Participate in an on‑call rotation....SuggestedFull timeWorldwide$145k - $165k
...Your Ego : Selflessly collaborate towards our shared purpose. About the role Bolt Graphics is seeking a highly experienced Site Reliability Engineer (SRE) to design, build, and operate highly reliable developer and production systems. This role is mission-critical to...SuggestedWork at officeImmediate start- ...Site Reliability Engineer (SRE) Location: Santa Clara Valley (Cupertino), California, Hybrid. Duration: 6+ Months Job Description Deploy, support and monitor new and existing services, platforms, and application stacks. Use scale testing to measure, tune...Suggested
- ...Senior Site Reliability Engineer Location: Remote Duration: 12 month contract to start IV Process: 1-3 Round IV process International Tech Top Skills: Java Python NodeJS -DevOps Engineer should work here too Main Responsibilities:...SuggestedContract workLocal areaRemote work
- ...Site Reliability Engineer (SRE) The successful applicant may be performing work in FedRAMP High or IL-5 environments, and therefore, must be a U.S. Person (i.e. U.S. citizen, U.S. national, lawful permanent resident, asylee, or refugee). This position may also perform...Permanent employmentWorldwideShift work
- ...Senior Site Reliability Engineer LeanData helps the world's fastest-growing companies automate, simplify, and accelerate revenue. We are looking for a Senior Site Reliability Engineer to lead the strategic evolution of our cloud infrastructure. Reporting directly...Full timeWork at officeFlexible hours2 days per week
- ...Site Reliability Engineer Sunnyvale, CA Site Reliability Eng. Must have LinkedIn profile. • t least 8+ years in a Reliability Engineering, DevOps or infrastructure focused role • Strong AWS and Linux Operating System, standard networking protocols, component...
$65 - $85 per hour
...Site Reliability Engineer Sustainable Talent is partnering with a global leader who's been transforming computer graphics, PC gaming, and accelerated computing for over 25 years. We are looking for a Site Reliability Engineer to support our client's team based out of...Full timeContract workWorldwide$150k
...S23 - Telecommunications Role - Site Reliability Engineer Location: Santa Clara, CA or Wall Township, NJ (Hybrid) Salary: $150,000 + Bonus + Comprehensive Benefits About the Role Join a team building and operating a modern Kubernetes-based platform...- ...Qualifications: 8+ years of software engineering experience, or equivalent demonstrated through... ...implement and maintain scalable and reliable infrastructure on Google Cloud Platform... ...vendor resources. Willingness to work on-site at stated location in the job opening....For contractorsWork experience placement
$230k - $250k
...Site Reliability Engineer Forward is transforming how the world's most complex networks are managed and secured. Founded in 2013 by four Stanford Ph.D.s, we built the industry's first network digital twin — a mathematically precise model of the production network that...Night shift$159.2k - $301.6k
...running Graphs on the cloud. In this reliability-focused role, you will own the availability... .... You'll partner with the backend engineers building these APIs to make sure the system... ...Science. ~5-10 years of experience in site reliability engineering, infrastructure,...Temporary workLocal areaWorldwide$152.5k - $219.2k
...global cloud platform. As a team of six engineers distributed across the US, Canada, and the... ...with a strong focus on automation, reliability, and operational excellence. We are one... ...Qualifications ~2+ years of experience in Site Reliability Engineering, DevOps, Infrastructure...Permanent employmentFull timeTemporary workLocal areaWorldwideFlexible hours- ...Overview: *Must have Apple experience* • At least 8+ years in a Reliability Engineering, DevOps or infrastructure focused role • Advanced experience with programming languages (Python, Java) • Passion for designing and building reliable systems • Strong sense...
- ...AWS Infra SRE/DevOps Engineer AWS Infra SRE/DevOps engineer with proven work experience ensuring reliability, availability and performance of cloud infra and platform. Specialist on Cisco Cloud run-on for infrastructure management, who can install, run, and maintain...Work experience placement
- ...Site Reliability Engineer Foxconn Industrial Internet (Fii), is a world leading professional design and manufacturing service provider of communication network equipment, cloud service equipment, precision tools and industrial robots. FII provides customers with intelligent...Permanent employmentFull timeWork at officeLocal area
$170k - $200k
...Site Reliability Engineer We are seeking a talented and motivated Site Reliability Engineer to join our engineering team. You will be responsible for building, maintaining, and troubleshooting cloud service/cluster, infrastructure, and monitoring systems to ensure high...Full timeWorldwide$110k - $120k
...interested in working with the World's leading AI-powered Quality Engineering Company? Ready to advance your career, team up with global... ...every day? Join us at QualityAI! We are looking for a Site Reliability Engineer (SRE)) to join our growing team in the United...Local area2 days per week3 days per week- ...Senior Site Reliability Engineer ILocationSan Jose, Costa Rica - RemoteSummary of roleOwn availability, the most important product feature, by continually striving for sustained operational excellence of Sumo’s planet-scale observability and security products. Work with...Part timeFlexible hours
- ...work from home day is currently Tuesday.Engineering at Lambda is responsible for building and... ...and networking teams to improve service reliability and deployment workflowsDeploy and... ...rotationYouHave 5+ years of experience in Site Reliability Engineering, Production Engineering...Part timeWork at officeLocal areaWork from homeFlexible hours
$185k - $278k
...transforming them into scalable solutions. * Debugging OS and engineering issues within our provided Linux environment. *... ...efficiently. The Impact You Will Have: * Enhancing the reliability and performance of our engineering environment. * Streamlining...Remote work$180k - $260k
...effortless integration into customers' logistics operations. About the role We are seeking an experienced Senior/Staff Site Reliability Engineer to support the operation, monitoring, and scaling of our growing fleet of autonomous vehicles. In this role, you will work...Odd jobWork at officeRemote work$151.6k - $245.3k
...Summary Your Career Palo Alto Networks runs a large hybrid infrastructure and is one of the largest GCP customers. As a Site Reliability Engineer, you will be part of a team supporting the services running on this infrastructure. This includes automation, architecture...Full timeWork at office$174k - $253k
Bachelor’s degree in Computer Science, Engineering, a related field, or equivalent practical experience. 5 years of experience with... ...'s degree in Computer Science or Engineering. About The Job Site Reliability Engineering (SRE) is what you get when you treat operations...$210.6k - $305.1k
...Minimum Qualifications: You have led a distributed team of 5+ engineers, can demonstrate strong technical vision for your team, and ensure... ..., and basic life insurance. Please see the Cisco careers site to discover more benefits and perks. Employees may be eligible...Full timeTemporary workPart timeLocal areaFlexible hours$207.4k - $259.2k
...differences, and supports and celebrates all of our team members.We are seeking a highly experienced and passionate Sr. Staff Site Reliability Engineer (SRE) to join our growing team. In this critical role, you will be responsible for the reliability, scalability,...Permanent employmentPart timeLocal area$184k - $287.5k
...At NVIDIA, Site Reliability Engineering provides a rare chance to define, develop, and support large-scale production systems with high efficiency and availability. This demanding position merges software and systems engineering efforts to guarantee flawless service operation...Full timePart time$189k - $232k
...Site Reliability Engineer Mountain View, US About EarnIn As one of the first pioneers of earned wage access, our passion at EarnIn is building products that deliver real-time financial flexibility for those with the unique needs of living paycheck to paycheck....Full timeWork at office2 days per week$100k - $200k
...OPPO US Research Center is seeking a skilled and proactive Site Reliability Engineer (SRE) to join our team. In this role, you will be responsible for ensuring the stability, scalability, and performance of our application systems. The ideal candidate is passionate about...Full time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
Related searches
- site reliability engineer Santa Clara, CA
- construction site safety Santa Clara, CA
- on-site clinical research associate (traveling/remote) Santa Clara, CA
- site safety Santa Clara, CA
- historic site Santa Clara, CA
- junior website developer Santa Clara, CA
- official site Santa Clara, CA
- site services specialist Santa Clara, CA
- site reliability engineering manager
- junior site reliability engineer






