Site Reliability Engineer
F5 Networks
At F5, we strive to bring a better digital world to life. Our teams empower organizations across the globe to create, secure, and run applications that enhance how we experience our evolving digital world. We are passionate about cybersecurity, from protecting consumers from fraud to enabling companies to focus on innovation. Everything we do centers around people. That means we obsess over how to make the lives of our customers, and their customers, better. And it means we prioritize a diverse F5 community where each individual can thrive.
The Role
This hybrid role combines the hands‑on responsibilities of a Technical Support Engineer within a SaaS (Software as a Service) environment with a growing focus on Site Reliability Engineering (SRE). The ideal candidate has a strong technical foundation, thrives in a fast‑paced troubleshooting environment, and is passionate about automation and delivering an exceptional customer experience. This role provides the opportunity to run, support, and scale an AI Security Public SaaS platform, operating AI inference workloads at scale.
Key Responsibilities
- Proactive Monitoring & Uptime Assurance Monitor key SaaS application metrics, logs, and alerts to proactively identify and prevent service disruptions
- Support a 24/7 operational model to ensure high availability and performance of the SaaS environment
- Customer-Centric Incident Response Serve as a primary point of contact for technical customer inquiries and issues
- Partner with customers to understand requirements, troubleshoot effectively, and provide clear, concise communication
- Triage, elevate, and own technical incidents through resolution
- Development Team Collaboration Analyze metrics, logs, and incident reports to provide actionable insights to engineering teams
- Identify recurring patterns and opportunities for automation and continuous improvement
- Partner cross-functionally to improve platform reliability and performance
- Site Reliability Engineering (SRE) Development Design and implement automation to streamline incident response and improve system reliability
- Introduce and mature SRE best practices, including enhanced monitoring, alerting, and self‑healing capabilities
- Contribute to building scalable, resilient infrastructure to support AI inference workloads
Required Qualifications
- Bachelor’s degree in Computer Science, Information Technology, or a related field (or equivalent practical experience)
- 1–3+ years of experience in technical support, systems administration, or a similar role
- Strong understanding of SaaS environments and cloud-based architectures (preferably AWS)
- Proficiency in at least one scripting language (e.g., Python)
- Solid understanding of web technologies ( REST APIs, JSON, etc.)
- Experience working with ticketing systems
- Strong problem‑solving and analytical skills
- Excellent written and verbal communication skills
- Ability to work independently and collaboratively in a team environment
- Willingness to learn new technologies and adapt to evolving requirements
Preferred Qualifications
- Experience with monitoring and observability tools (e.g., Prometheus, Grafana)
- Familiarity with configuration management tools (e.g., Terraform)
- Experience working with cloud infrastructure technologies
- Exposure to SRE principles and reliability engineering practices
- Strong understanding of networking fundamentals
- Experience with databases (PostgreSQL), operating systems (Linux), and Kubernetes #LI-ZB1
The Job Description is intended to be a general representation of the responsibilities and requirements of the job. However, the description may not be all-inclusive, and responsibilities and requirements are subject to change.
The annual base pay for this position is: $149,800.00 - $224,600.00 F5 maintains broad salary ranges for its roles in order to account for variations in knowledge, skills, experience, geographic locations, and market conditions, as well as to reflect F5’s differing products, industries, and lines of business. The pay range referenced is as of the time of the job posting and is subject to change. You may also be offered incentive compensation, bonus, restricted stock units, and benefits. More details about F5’s benefits can be found at the following link: F5 reserves the right to change or terminate any benefit plan without notice.
Please note that F5 only contacts candidates through F5 email address (ending with @f5.com) or auto email notification from Workday (ending with f5.com or @myworkday.com).
Equal Employment Opportunity It is the policy of F5 to provide equal employment opportunities to all employees and employment applicants without regard to unlawful considerations of race, religion, color, national origin, sex, sexual orientation, gender identity or expression, age, sensory, physical, or mental disability, marital status, veteran or military status, genetic information, or any other classification protected by applicable local, state, or federal laws. This policy applies to all aspects of employment, including, but not limited to, hiring, job assignment, compensation, promotion, benefits, training, discipline, and termination.
F5 offers a variety of reasonable accommodations for candidates. Requesting an accommodation is completely voluntary. F5 will assess the need for accommodations in the application process separately from those that may be needed to perform the job. Request by contacting View email address on click.appcast.io.
Hybrid: Employees within 30 commutable miles of an F5 office are required to work from the office a minimum of 30 business days per quarter. Remote: Primarily work from designated home location but can come into an F5 office to work or travel to an offsite location as needed.
Together, we’re building a better digital world. Founded in 1996, F5 is a global leader in application delivery and security. Our premier platform helps customers secure and deliver every app, API, and piece of infrastructure across all environments. Backed by over three decades of expertise and 553 patents, our solutions protect against threats while ensuring fast, reliable digital experiences. With over 6,400 employees, we serve more than 23,000 customers in over 170 countries. To continue this work, we need people like you—the best minds in the industry. We’re committed to a unique, human-first culture that encourages authenticity, prioritizes diversity and inclusion, and fosters the growth and success of our employees.
#J-18808-Ljbffr$150.4k - $277.6k
...Services The Media Platforms SRE team under the Apple Service Engineering division is one of the most exciting examples of Apple’s long... ...field with 4+ years experience At least 6 years in a Reliability Engineering, DevOps or infrastructure focused role Advanced...SuggestedRelocationDay shift$132.6k - $214.5k
...As part of this role, you will collaborate closely with our engineering teams to develop innovative solutions that provide clear and... ...team to influence the operability of the product and ensure the reliability and availability of our services. Qualifications...SuggestedFull timeWork at officeVisa sponsorshipWork visa$230k - $250k
...minds are shaping the future of network reliability, security, and AI‑ready operations. About... ...you will be building the reliability engineering function at Forward — defining how we... ...Looking For ~6+ years of experience in site reliability engineering, DevOps, or...SuggestedNight shift- ...Site Reliability Engineer (SRE) Location: Santa Clara Valley (Cupertino), California, Hybrid. Duration: 6+ Months Job Description Deploy, support and monitor new and existing services, platforms, and application stacks. Use scale testing to measure, tune...Suggested
$110k - $130k
...interested in working with the World's leading AI-first Quality Engineering Company? Ready to advance your career, team up with global... ...every day? Join us at QualityAI! We are looking for a Site Reliability Engineer to join our growing team in Riverwoods, IL United States...SuggestedCasual workLocal areaFlexible hours$230k - $250k
...network. It's the foundation for autonomous networking, giving engineers and AI agents the ability to know the impact of every change... ...how things have always been done.Forward is looking for a Site Reliability EngineerAbout the Role This is not a "keep the lights on"...Night shift$149.8k - $224.6k
...This hybrid role combines the hands-on responsibilities of a Technical Support Engineer within a SaaS (Software as a Service) environment with a growing focus on Site Reliability Engineering (SRE). The ideal candidate has a strong technical foundation, thrives...Local area- ...Site Reliability Engineer Location – San Jose, CA What You'll Do - Responsibilities Engage in and improve the whole lifecycle of services—from inception and design, through automated deployment, operation and refinement. Work with all relative teams to make...
- ...AWS Infra SRE/DevOps Engineer AWS Infra SRE/DevOps engineer with proven work experience ensuring reliability, availability and performance of cloud infra and platform. Specialist on Cisco Cloud run-on for infrastructure management, who can install, run, and maintain...Work experience placement
$104.9k - $174.7k
...Site Reliability Engineer The Site Reliability Engineer role is responsible for improving the reliability, availability, performance, and operational quality of production systems. This role provides technical input into project plans, schedules, methodologies, and...Temporary workLocal area- ...Site Reliability Engineer Foxconn Industrial Internet (Fii), is a world leading professional design and manufacturing service provider of communication network equipment, cloud service equipment, precision tools and industrial robots. FII provides customers with intelligent...Permanent employmentFull timeWork at officeLocal area
$60 - $62 per hour
...and improving existing processes to enhance overall system reliability. Key Responsibilities: Deploying software to cloud... ...Computer Science or a related field. 3+ years of experience in Site Reliability Engineering. Proficiency with Kubernetes, Helm, Linux, AWS networking...Hourly payContract workRemote work- ...Job Title : Senior Site Reliability Engineer Location : Santa Clara, CA Contract ENGAGEMENT SUMMARY The Candidate will provide SRE services for AI platforms and supporting infrastructure with emphasis on reliability engineering, incident response...Contract work
- ...Qualifications: 8+ years of software engineering experience, or equivalent demonstrated through... ...implement and maintain scalable and reliable infrastructure on Google Cloud Platform... ...vendor resources. Willingness to work on-site at stated location in the job opening....For contractorsWork experience placement
- ...Job Title : Site Reliability Engineer Location: San Jose, CA Duration: Contract Job Description: Extensive experience working with linux flavors like rhel/centos os, shells, filesystems and utilities Knowledge of distributed computing...Contract workImmediate start
$65 - $85 per hour
...Site Reliability Engineer Sustainable Talent is partnering with a global leader who's been transforming computer graphics, PC gaming, and accelerated computing for over 25 years. We are looking for a Site Reliability Engineer to support our client's team based out of...Full timeContract workWorldwide- ...Position: Site Reliability Engineering (SRE) Location: Santa Clara, CA (Onsite) Duration: W2 / C2C Contract Experience: 10+ Years Job Description: • WS application and CI/CD pipelines, Microsoft Server admin and workload support (Data Center and AWS...Contract workImmediate start
$122.5k - $175k
...impact at the company pioneering security transformation in the AI era? Join us at Zscaler.RoleWe are looking for a Staff Site Reliability Engineer to join our team. This is a hybrid role going into the San Jose, CA office 3 days a week, reporting to the Chief Architect...Full timeWork at officeLocal area3 days per week$248k - $396.75k
Site Reliability Engineering (SRE) at NVIDIA is an engineering discipline focused on designing, building, and operating large-scale production systems with exceptional efficiency, resilience, and availability. It combines software and systems engineering practices with...Full time- ...precision that drives great outcomes. Job Summary Key Responsibilities Lead, mentor, and develop a team of Site Reliability/Production Engineers, providing technical direction, coaching, and career development. Own the reliability, availability, and...Full timeWork at officeVisa sponsorshipWork visa
$192.4k - $275.8k
...CloudOps— the team that keeps Splunk Cloud running for some of the world's most demanding enterprise customers, blending Site Reliability Engineering, Systems Engineering, and Service Engineering disciplines at a scale very few teams ever get to operate at. When the...Full timeTemporary workLocal areaFlexible hours$145k - $160k
...We are seeking a specialized Observability & Infrastructure Engineer to lead the monitoring, telemetry, and platform-as-code initiatives critical to our multi-region disaster recovery roadmap. You will architect and implement robust observability pipelines, ensure deep...Temporary workRemote workFlexible hours$260k - $275k
...Saviynt • Work on a mission-critical SaaS platform used by global enterprises • Solve complex reliability challenges at scale • Influence architecture and engineering culture at a company level • Competitive compensation, benefits, and growth opportunities...$122.5k - $175k
...at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer to join our team. This is a hybrid role going into the San Jose, CA office 3 days a week, reporting to the Chief Architect...Full timeWork at officeLocal area3 days per week$152k - $228k
...Up to 25% Job ID: 1874 The Role: The Platform Engineering team builds, secures, and operates scalable infrastructure... ...with on-premises components deployed at customer sites. The Site Reliability Engineering discipline keeps the platform stable and reliable...Permanent employmentContract workWork at officeRemote work$195k - $285k
...purpose-built AI inference silicon, and the infrastructure underpinning our engineering organization must be as reliable and scalable as the chips we build. This role builds and leads d-Matrix's Site Reliability Engineering function from the ground up, owning the...Full timeRemote work$175k - $265k
...d-Matrix's SRE team owns the infrastructure layer that every engineering team and customer depends on — colocation facilities, on-premises... .... This role is a core member of that team, responsible for reliability, automation, and observability across colo, on-premises lab,...Full time$145k - $165k
...: Selflessly collaborate towards our shared purpose. About the role Bolt Graphics is seeking a highly experienced Site Reliability Engineer (SRE) to design, build, and operate highly reliable developer and production systems. This role is mission-critical to maintaining...Work at officeImmediate start$170k - $200k
...We are seeking a talented and motivated Site Reliability Engineer to join our engineering team. You will be responsible for building, maintaining, and troubleshooting cloud service/cluster, infrastructure, and monitoring systems to ensure high availability, performance...Full time- ...Site Reliability Engineer Onsite- Bay Area, CA Skills Relevant Skills and Experience What You’ll Do (Day-to-Day) Own and manage our cloud infrastructure (GCP or AWS, on-prem). Build, maintain, and optimize Kubernetes clusters (including GPU-backed clusters...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- site reliability engineer San Jose, CA
- site reliability engineer sre San Jose, CA
- site safety San Jose, CA
- website coordinator San Jose, CA
- on-site clinical research associate (traveling/remote) San Jose, CA
- site services specialist San Jose, CA
- on site coordinator San Jose, CA
- construction site safety San Jose, CA
- junior website developer San Jose, CA
- site recruiter San Jose, CA

