Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Sr. Site Reliability Engineer

Illumio

Location: 5 on-site days a week in Sunnyvale, CA Headquarters.

Responsibilities
  • Monitor system performance, application health, and infrastructure metrics using monitoring and logging services, and implement proactive measures to optimize performance and availability
  • Oncall duty for production uptime and support for customer escalations
  • Release upgrades and maintenance activities including hotfixes and infrastructure updates
  • Lead incident response and resolution efforts, conducting root cause analysis, implementing corrective actions, and documenting post‑incident reviews
  • Implement security best practices and controls in the cloud environments to protect data, applications, and infrastructure, and ensure compliance with regulatory requirements
  • Drive continuous improvement initiatives to enhance reliability, scalability, and efficiency of infrastructure and services, leveraging automation and emerging technologies
Qualifications
  • Bachelor’s degree in computer science, Engineering, or related field; or equivalent work experience
  • 5+ years of experience working as a Site Reliability Engineer (SRE) or similar role, with a focus on AWS and/or Azure cloud platform
  • Hands‑on experience in designing, deploying, and managing AWS and/or Azure infrastructure, including compute, storage, networking, and security services
  • Proficiency in scripting and programming languages such as PowerShell, Python, or Go for automation and infrastructure management tasks
  • Strong understanding of CI/CD principles and experience with tools such as Azure DevOps, Jenkins, or GitLab CI/CD
  • Experience with containerization technologies (e.g., Docker, Kubernetes) and microservices architecture in AWS and Azure environments is a plus
  • Excellent analytical, problem‑solving, and communication skills, with the ability to collaborate effectively with cross‑functional teams
  • AWS or Azure certifications such as AWS/Azure Solutions Architect, Azure DevOps Engineer, or Azure Security Engineer are preferred

For roles in San Francisco and Los Angeles: Pursuant to the San Francisco Fair Chance Ordinance and the Los Angeles Fair Chance Initiative for Hiring, Illumio will consider for employment qualified applicants with arrest and conviction records.

#J-18808-Ljbffr
Vacancy posted 4 hours ago
Similar jobs that could be interesting for youBased on the Sr. Site Reliability Engineer in San Jose, CA vacancy
  • $175k - $265k

     ...Overviewd-Matrix's SRE team owns the infrastructure layer that every engineering team and customer depends on — colocation facilities, on-...  .... This role is a core member of that team, responsible for reliability, automation, and observability across colo, on-premises lab,... 
    Senior

    d-Matrix

    Santa Clara, CA
    1 day ago
  • $104.9k - $174.7k

     ...SRE role is responsible for improving the reliability, availability, performance, and...  ...actions through completion.Follow up with engineering, development, security, support, and business...  ...Qualifications5+ years of experience in Site Reliability Engineering, Systems Engineering... 
    Senior
    Full time
    Local area

    LexisNexis Risk Solutions Group

    San Jose, CA
    4 days ago
  • $132.6k - $214.5k

     ...As part of this role, you will collaborate closely with our engineering teams to develop innovative solutions that provide clear and...  ...team to influence the operability of the product and ensure the reliability and availability of our services. Qualifications... 
    Senior
    Full time
    Work at office
    Visa sponsorship
    Work visa

    Palo Alto Networks

    Santa Clara, CA
    5 days ago
  • $150.4k - $277.6k

     ...Services The Media Platforms SRE team under the Apple Service Engineering division is one of the most exciting examples of Apple’s long...  ...field with 4+ years experience At least 6 years in a Reliability Engineering, DevOps or infrastructure focused role Advanced... 
    Senior
    Relocation
    Day shift

    Apple

    Cupertino, CA
    1 day ago
  •  ...CloudOps— the team that keeps Splunk Cloud running for some of the world's most demanding enterprise customers, blending Site Reliability Engineering, Systems Engineering, and Service Engineering disciplines at a scale very few teams ever get to operate at. When the... 
    Senior

    Webex Events (formerly Socio)

    San Jose, CA
    1 day ago
  •  ...logs-store, traces-store, profiles-store, analytics-lake, enrichment-service, collection-monitor. Alert, Correlation & SLO: alert-engine-framework, alert-correlation, slo-framework, default M-series alert rules. Topology, Cluster-Health & Cluster Platform Services:... 
    Senior
    Full time
    Contract work
    Local area

    Bitdeer

    San Jose, CA
    6 days ago
  • $114.4k - $124.8k

     ...Salary: $114,400 - 124,800 per year Requirements: At least 5 years of experience in Site Reliability Engineering, Systems Engineering, or Infrastructure Operations in large-scale enterprise environments Deep expertise in Chef or Cinc cookbook development, serverless... 
    Senior
    Hourly pay
    Full time
    Temporary work

    CYNET SYSTEMS

    Santa Clara, CA
    4 days ago
  • $260k - $275k

     ...Saviynt Work on a mission-critical SaaS platform used by global enterprises Solve complex reliability challenges at scale Influence architecture and engineering culture at a company level Competitive compensation, benefits, and growth opportunities... 

    Saviynt

    Milpitas, CA
    4 hours ago
  • $152k - $287.5k

     ...infrastructure platforms for automated host lifecycle management, fleet reliability/auto-healing, E2E observability or data-driven operations (...  ...such as Python, Go, Perl, or Ruby. ~ Mentored other engineers and influenced technical direction through design reviews, architecture... 
    Senior
    Full time

    NVIDIA

    Santa Clara, CA
    3 days ago
  • $248k - $396.75k

     ...Site Reliability Engineering (SRE) at NVIDIA is an engineering discipline focused on designing, building, and operating large-scale production systems with exceptional efficiency, resilience, and availability. It combines software and systems engineering practices with... 

    NVIDIA Gruppe

    Santa Clara, CA
    4 hours ago
  • $207k - $300k

     ...areas within SU SRE, mentoring team members to enhance system reliability and efficiency.Initiate, own, and lead large-scale,...  ...Design for Reliability techniques.3 years of experience as a Site Reliability Engineer.3 years of experience leading projects.3 years of experience... 

    Google

    San Jose, CA
    21 hours ago
  • $230k - $250k

     ...network. It's the foundation for autonomous networking, giving engineers and AI agents the ability to know the impact of every change...  ...how things have always been done.Forward is looking for a Site Reliability EngineerAbout the Role This is not a "keep the lights on"... 
    Night shift

    Forward Networks Inc

    Santa Clara, CA
    5 days ago
  • $230k - $250k

     ...minds are shaping the future of network reliability, security, and AI‑ready operations. About...  ...you will be building the reliability engineering function at Forward — defining how we...  ...Looking For ~6+ years of experience in site reliability engineering, DevOps, or... 
    Night shift

    Forward

    Santa Clara, CA
    3 days ago
  • $65 - $85 per hour

     ...Talent is partnering with Nvidia, a global leader in computer graphics, PC gaming, and accelerated computing, to bring a Site Reliability Engineer (Contract) to the team based in Santa Clara, CA. This is a full‑time (W‑2) contract role. Pay ranges from $65/hr to $85... 
    Full time
    Contract work
    Worldwide

    Sustainable Talent

    Santa Clara, CA
    4 hours ago
  •  ...Job Description Job Description Site Reliability Engineer Foxconn Industrial Internet (Fii), is a world leading professional design and manufacturing service provider of communication network equipment, cloud service equipment, precision tools and industrial robots... 
    Permanent employment
    Full time
    Work at office
    Local area

    Foxconn Industrial Internet - FII

    San Jose, CA
    a month ago
  •  ...of Huobi globe spanning infrastructure. •       Work with engineering teams to make sure new features and changes are deployed quickly...  .... •       Constantly improve our system performance and reliability through better tools, process and monitoring system. •... 
    Worldwide

    Cryptoware Technologies Inc

    Santa Clara, CA
    a month ago
  •  ...infrastructure, DevOps, SRE, and platform engineering. You will test AI-generated commands,...  ...and deployment workflows for accuracy and reliability. Work with AWS, Azure, GCP,...  ...Azure DevOps Cloud Infrastructure Site Reliability Engineering (SRE) Platform... 
    For contractors
    Remote work

    YO AI Labs

    San Jose, CA
    a month ago
  •  ...Role This hybrid role combines the hands‑on responsibilities of a Technical Support Engineer within a SaaS (Software as a Service) environment with a growing focus on Site Reliability Engineering (SRE). The ideal candidate has a strong technical foundation, thrives in... 
    Work at office
    Local area
    Remote work
    Work from home

    F5 Networks

    San Jose, CA
    5 days ago
  • $195k - $285k

     ...purpose-built AI inference silicon, and the infrastructure underpinning our engineering organization must be as reliable and scalable as the chips we build. This role builds and leads d-Matrix's Site Reliability Engineering function from the ground up, owning the... 
    Remote work

    d-Matrix

    Santa Clara, CA
    4 hours ago
  •  ...Up to 25% Job ID: 1874 The Role The Platform Engineering team builds, secures, and operates scalable infrastructure...  ...products with on‑premises components deployed at customer sites. The Site Reliability Engineering discipline keeps the platform stable and reliable... 
    Work at office
    Remote work

    Physics World

    Santa Clara, CA
    4 hours ago
  • $120k - $150k

     ...interested in working with the World's leading AI-first Quality Engineering Company? Ready to advance your career, team up with global...  ...every day? Join us at QualityAI! We are looking for a Site Reliability Engineer to join our growing team in Riverwoods, IL United States... 
    Casual work
    Local area
    Flexible hours

    QualiTest Group

    Santa Clara, CA
    1 day ago
  • $161k - $299k

     ...Overview We are seeking a talented Software Engineer to join our Research & Development (R&D)...  ...document software solutions to ensure reliability and accuracy. Solve complex...  ...the company and want to make sure our job site is accessible to all. If you experience... 
    Senior
    Work experience placement

    Cadence Design Systems

    San Jose, CA
    4 days ago
  • $122.5k - $175k

     ...at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer to join our team. This is a hybrid role going into the San Jose, CA office 3 days a week, reporting to the Chief Architect... 
    Full time
    Work at office
    Local area
    3 days per week

    Zscaler

    San Jose, CA
    2 days ago
  • $174k - $252k

     ...systems by pushing for changes that improve reliability and velocity.Practice sustainable...  ...:Bachelor’s degree in Computer Science, Engineering, a related field, or equivalent practical...  ...degree in Computer Science or Engineering.Site Reliability Engineering (SRE) is what you... 
    Senior

    Google

    Sunnyvale, CA
    1 day ago
  • $160k - $240k

     ...one another millions of times a day - quickly, reliably, and securely. Any time you swipe your credit...  ...come make a difference at Fiserv.Job TitleSenior Site Reliability EngineerWhat does a successful Site Reliability Engineer do at Fiserv?You will join our global team in... 
    Senior
    Full time

    Fiserv

    Sunnyvale, CA
    1 day ago
  • $161k - $299k

     ...the Innovus R&D team, you will build high‑performance software engines and algorithms behind next‑generation physical synthesis, helping...  ...welcome your interest in the company and want to make sure our job site is accessible to all. If you experience difficulty using this... 
    Senior

    Cadence Design Systems

    San Jose, CA
    1 day ago
  •  ...Own the architecture and design of reliable, scalable, cost-effective, and performant AI...  ...master's degree in computer science or engineering is preferred. Key Skills Software...  ...Machine Learning, Artificial Intelligence, Site Reliability Engineering, AI Infrastructure... 
    Senior

    engineeringjobs.net, Inc.

    Sunnyvale, CA
    18 hours ago
  • $262k - $364k

     ...and training AI infrastructure from SRE side, ensuring it is reliable, scalable, cost effective and performant, while working closely...  ...qualifications:Master's degree in Computer Science or Engineering.Site Reliability Engineering (SRE) combines software and systems engineering... 
    Senior

    Google

    Sunnyvale, CA
    2 days ago
  • $160k - $200k

     ...positions to the technology community. We seek talented, passionate, and committed engineers, technologists, and business leaders to join us.Job Summary:Supermicro is seeking a top-notch hands-on Sr. Software Engineer to work on PCIe, SAS/SATA, USB and other HW related areas... 
    Senior
    Worldwide

    Supermicro

    San Jose, CA
    2 days ago
  • $160k - $200k

     ...positions to the technology community. We seek talented, passionate, and committed engineers, technologists, and business leaders to join us.Job Summary:Supermicro is seeking a top-notch hands-on Sr. Software Engineer to work on PCIe, SAS/SATA, USB and other HW related areas... 
    Senior
    Worldwide

    Super Micro Computer

    San Jose, CA
    2 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Sr. Site Reliability Engineer. Be the first to apply!