Sr. Site Reliability Engineer
Illumio
Location: 5 on-site days a week in Sunnyvale, CA Headquarters.
Responsibilities
- Monitor system performance, application health, and infrastructure metrics using monitoring and logging services, and implement proactive measures to optimize performance and availability
- Oncall duty for production uptime and support for customer escalations
- Release upgrades and maintenance activities including hotfixes and infrastructure updates
- Lead incident response and resolution efforts, conducting root cause analysis, implementing corrective actions, and documenting post‑incident reviews
- Implement security best practices and controls in the cloud environments to protect data, applications, and infrastructure, and ensure compliance with regulatory requirements
- Drive continuous improvement initiatives to enhance reliability, scalability, and efficiency of infrastructure and services, leveraging automation and emerging technologies
Qualifications
- Bachelor’s degree in computer science, Engineering, or related field; or equivalent work experience
- 5+ years of experience working as a Site Reliability Engineer (SRE) or similar role, with a focus on AWS and/or Azure cloud platform
- Hands‑on experience in designing, deploying, and managing AWS and/or Azure infrastructure, including compute, storage, networking, and security services
- Proficiency in scripting and programming languages such as PowerShell, Python, or Go for automation and infrastructure management tasks
- Strong understanding of CI/CD principles and experience with tools such as Azure DevOps, Jenkins, or GitLab CI/CD
- Experience with containerization technologies (e.g., Docker, Kubernetes) and microservices architecture in AWS and Azure environments is a plus
- Excellent analytical, problem‑solving, and communication skills, with the ability to collaborate effectively with cross‑functional teams
- AWS or Azure certifications such as AWS/Azure Solutions Architect, Azure DevOps Engineer, or Azure Security Engineer are preferred
For roles in San Francisco and Los Angeles: Pursuant to the San Francisco Fair Chance Ordinance and the Los Angeles Fair Chance Initiative for Hiring, Illumio will consider for employment qualified applicants with arrest and conviction records.
#J-18808-Ljbffr$175k - $265k
...Overviewd-Matrix's SRE team owns the infrastructure layer that every engineering team and customer depends on — colocation facilities, on-... .... This role is a core member of that team, responsible for reliability, automation, and observability across colo, on-premises lab,...Senior$104.9k - $174.7k
...SRE role is responsible for improving the reliability, availability, performance, and... ...actions through completion.Follow up with engineering, development, security, support, and business... ...Qualifications5+ years of experience in Site Reliability Engineering, Systems Engineering...SeniorFull timeLocal area$132.6k - $214.5k
...As part of this role, you will collaborate closely with our engineering teams to develop innovative solutions that provide clear and... ...team to influence the operability of the product and ensure the reliability and availability of our services. Qualifications...SeniorFull timeWork at officeVisa sponsorshipWork visa$150.4k - $277.6k
...Services The Media Platforms SRE team under the Apple Service Engineering division is one of the most exciting examples of Apple’s long... ...field with 4+ years experience At least 6 years in a Reliability Engineering, DevOps or infrastructure focused role Advanced...SeniorRelocationDay shift- ...CloudOps— the team that keeps Splunk Cloud running for some of the world's most demanding enterprise customers, blending Site Reliability Engineering, Systems Engineering, and Service Engineering disciplines at a scale very few teams ever get to operate at. When the...Senior
- ...logs-store, traces-store, profiles-store, analytics-lake, enrichment-service, collection-monitor. Alert, Correlation & SLO: alert-engine-framework, alert-correlation, slo-framework, default M-series alert rules. Topology, Cluster-Health & Cluster Platform Services:...SeniorFull timeContract workLocal area
$114.4k - $124.8k
...Salary: $114,400 - 124,800 per year Requirements: At least 5 years of experience in Site Reliability Engineering, Systems Engineering, or Infrastructure Operations in large-scale enterprise environments Deep expertise in Chef or Cinc cookbook development, serverless...SeniorHourly payFull timeTemporary work$260k - $275k
...Saviynt Work on a mission-critical SaaS platform used by global enterprises Solve complex reliability challenges at scale Influence architecture and engineering culture at a company level Competitive compensation, benefits, and growth opportunities...$152k - $287.5k
...infrastructure platforms for automated host lifecycle management, fleet reliability/auto-healing, E2E observability or data-driven operations (... ...such as Python, Go, Perl, or Ruby. ~ Mentored other engineers and influenced technical direction through design reviews, architecture...SeniorFull time$248k - $396.75k
...Site Reliability Engineering (SRE) at NVIDIA is an engineering discipline focused on designing, building, and operating large-scale production systems with exceptional efficiency, resilience, and availability. It combines software and systems engineering practices with...$207k - $300k
...areas within SU SRE, mentoring team members to enhance system reliability and efficiency.Initiate, own, and lead large-scale,... ...Design for Reliability techniques.3 years of experience as a Site Reliability Engineer.3 years of experience leading projects.3 years of experience...$230k - $250k
...network. It's the foundation for autonomous networking, giving engineers and AI agents the ability to know the impact of every change... ...how things have always been done.Forward is looking for a Site Reliability EngineerAbout the Role This is not a "keep the lights on"...Night shift$230k - $250k
...minds are shaping the future of network reliability, security, and AI‑ready operations. About... ...you will be building the reliability engineering function at Forward — defining how we... ...Looking For ~6+ years of experience in site reliability engineering, DevOps, or...Night shift$65 - $85 per hour
...Talent is partnering with Nvidia, a global leader in computer graphics, PC gaming, and accelerated computing, to bring a Site Reliability Engineer (Contract) to the team based in Santa Clara, CA. This is a full‑time (W‑2) contract role. Pay ranges from $65/hr to $85...Full timeContract workWorldwide- ...Job Description Job Description Site Reliability Engineer Foxconn Industrial Internet (Fii), is a world leading professional design and manufacturing service provider of communication network equipment, cloud service equipment, precision tools and industrial robots...Permanent employmentFull timeWork at officeLocal area
- ...of Huobi globe spanning infrastructure. • Work with engineering teams to make sure new features and changes are deployed quickly... .... • Constantly improve our system performance and reliability through better tools, process and monitoring system. •...Worldwide
- ...infrastructure, DevOps, SRE, and platform engineering. You will test AI-generated commands,... ...and deployment workflows for accuracy and reliability. Work with AWS, Azure, GCP,... ...Azure DevOps Cloud Infrastructure Site Reliability Engineering (SRE) Platform...For contractorsRemote work
- ...Role This hybrid role combines the hands‑on responsibilities of a Technical Support Engineer within a SaaS (Software as a Service) environment with a growing focus on Site Reliability Engineering (SRE). The ideal candidate has a strong technical foundation, thrives in...Work at officeLocal areaRemote workWork from home
$195k - $285k
...purpose-built AI inference silicon, and the infrastructure underpinning our engineering organization must be as reliable and scalable as the chips we build. This role builds and leads d-Matrix's Site Reliability Engineering function from the ground up, owning the...Remote work- ...Up to 25% Job ID: 1874 The Role The Platform Engineering team builds, secures, and operates scalable infrastructure... ...products with on‑premises components deployed at customer sites. The Site Reliability Engineering discipline keeps the platform stable and reliable...Work at officeRemote work
$120k - $150k
...interested in working with the World's leading AI-first Quality Engineering Company? Ready to advance your career, team up with global... ...every day? Join us at QualityAI! We are looking for a Site Reliability Engineer to join our growing team in Riverwoods, IL United States...Casual workLocal areaFlexible hours$161k - $299k
...Overview We are seeking a talented Software Engineer to join our Research & Development (R&D)... ...document software solutions to ensure reliability and accuracy. Solve complex... ...the company and want to make sure our job site is accessible to all. If you experience...SeniorWork experience placement$122.5k - $175k
...at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer to join our team. This is a hybrid role going into the San Jose, CA office 3 days a week, reporting to the Chief Architect...Full timeWork at officeLocal area3 days per week$174k - $252k
...systems by pushing for changes that improve reliability and velocity.Practice sustainable... ...:Bachelor’s degree in Computer Science, Engineering, a related field, or equivalent practical... ...degree in Computer Science or Engineering.Site Reliability Engineering (SRE) is what you...Senior$160k - $240k
...one another millions of times a day - quickly, reliably, and securely. Any time you swipe your credit... ...come make a difference at Fiserv.Job TitleSenior Site Reliability EngineerWhat does a successful Site Reliability Engineer do at Fiserv?You will join our global team in...SeniorFull time$161k - $299k
...the Innovus R&D team, you will build high‑performance software engines and algorithms behind next‑generation physical synthesis, helping... ...welcome your interest in the company and want to make sure our job site is accessible to all. If you experience difficulty using this...Senior- ...Own the architecture and design of reliable, scalable, cost-effective, and performant AI... ...master's degree in computer science or engineering is preferred. Key Skills Software... ...Machine Learning, Artificial Intelligence, Site Reliability Engineering, AI Infrastructure...Senior
$262k - $364k
...and training AI infrastructure from SRE side, ensuring it is reliable, scalable, cost effective and performant, while working closely... ...qualifications:Master's degree in Computer Science or Engineering.Site Reliability Engineering (SRE) combines software and systems engineering...Senior$160k - $200k
...positions to the technology community. We seek talented, passionate, and committed engineers, technologists, and business leaders to join us.Job Summary:Supermicro is seeking a top-notch hands-on Sr. Software Engineer to work on PCIe, SAS/SATA, USB and other HW related areas...SeniorWorldwide$160k - $200k
...positions to the technology community. We seek talented, passionate, and committed engineers, technologists, and business leaders to join us.Job Summary:Supermicro is seeking a top-notch hands-on Sr. Software Engineer to work on PCIe, SAS/SATA, USB and other HW related areas...SeniorWorldwide
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Sr. Site Reliability Engineer. Be the first to apply!
- site reliability engineer San Jose, CA
- senior computer engineer San Jose, CA
- senior manager customer operations San Jose, CA
- senior software engineer ruby on rails San Jose, CA
- sr finance manager San Jose, CA
- sr marketing manager San Jose, CA
- senior customer service San Jose, CA
- senior business manager San Jose, CA
- senior account executive San Jose, CA
- senior accounts receivable analyst San Jose, CA




