Staff Site Reliability Engineer - Observability GCP (Chicago)
$194k - $267kOkta
Secure Every Identity, from AI to HumanIdentity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence.This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk.We are seeking a highly technical ObservabilitySite Reliability Engineer with a specialty in Google Cloud, to own and expand our Observability ecosystem into GCP. In this role, you will move beyond simple monitoring to delivering a world class, comprehensive, scalable Observability Platform that enables our SRE teams and business partners. You will treat infrastructure as code—utilizing Terraform and strong coding proficiency in Go, Python, or Ruby—to automate the deployment of agents and collectors across complex distributed systems.Key ResponsibilitiesAutomated Infrastructure: Design, build, and maintain scalable observability infrastructure using tools like Terraform.GCP Observabilty Engineering: Optimize the collection, processing, and storage of Observabilty data to ensure high reliability and low latency of our Splunk and Grafana servicesIncident Response: Participate in on-call rotations and lead post-incident reviews to drive systemic improvements and observability-driven development.Automation: Eliminate toil by automating the deployment and scaling of observability agents and collectors.Required Skills & Experience (The Essentials)GKE: Minimum 5+ Experience scaling and managing observability in a Google Cloud platform. Visualization: Expertise in creating intuitive, actionable Splunk or Grafana dashboards that correlate data across multiple sources.SRE Mindset: Minimum 3+ years of experience in an SRE, DevOps, or Systems Engineering role with a focus on high-availability systems.Programming Proficiency: Strong coding skills in Python, Go for building internal tools and automating workflows.Distributed Systems: Deep understanding of Linux internals, networking (TCP/IP, DNS, Load Balancing), and container orchestration (Kubernetes/GKE).Problem Solving: A data-driven approach to debugging complex, cross-service performance bottlenecks.Bonus Skills (The Nice-to-Haves)Telemetry Standards: Hands-on experience with OpenTelemetry (OTel), Vector, or similar frameworks for instrumenting applications.Grafana Loki: Experience in migrating Splunk to Grafana LokiOther Cloud Platforms: Experience managing observability native tools within AWS.Additional requirements:This position requires the ability to access federal environments and/or have access to protected federal data. As a condition of employment for this position, the successful candidate must be able to submit documentation establishing U.S. Person status (e.g. a U.S. Citizen, National, Lawful Permanent Resident, Refugee, or Asylee. 22 CFR 120.15) upon hire.#LI-MM#LI-HybridP24517_3387022The annual base salary range for this position for candidates located in the San Francisco Bay area is between: $194,000—$267,000 USDBelow is the annual base salary range for candidates located in California (excluding San Francisco Bay Area), Colorado, Illinois, New York and Washington. Your actual base salary will depend on factors such as your skills, qualifications, experience, and work location. In addition, Okta offers equity (where applicable), bonus, and benefits, including health, dental and vision insurance, 401(k), flexible spending account, and paid leave (including PTO and parental leave) in accordance with our applicable plans and policies. To learn more about our Total Rewards program please visit: . The annual base salary range for this position for candidates located in California (excluding San Francisco Bay Area), Colorado, Illinois, New York, and Washington is between:$174,000—$239,000 USDThe Okta ExperienceSupporting Your Well-BeingDriving Social Impact Developing Talent and Fostering Connection + CommunityWe are intentional about connection. Our global community, spanning over 20 offices worldwide, is united by a drive to innovate. Your journey begins with an immersive, in-person onboarding experience designed to accelerate your impact and connect you to our mission and team from day one.Okta is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, ancestry, marital status, age, physical or mental disability, or status as a protected veteran. We also consider for employment qualified applicants with arrest and convictions records, consistent with applicable laws.If reasonable accommodation is needed to complete any part of the job application, interview process, or onboarding please use this Form to request an accommodation.Notice for New York City Applicants & Employees: Okta may use Automated Employment Decision Tools (AEDT), as defined by New York City Local Law 144, that use artificial intelligence, machine learning, or other automated processes to assist in our recruitment and hiring process. In accordance with NYC Local Law 144, if you are an applicant or employee residing in New York City, please click here to view our full NYC AEDT Notice.
$194k - $267k
...concepts and tools.Position Overview:The Site Reliability Engineer (SRE) will play a key role in... ...service communication, security, and observability within the Kubernetes clusters. Enable... ...our San Francisco, CA HQ office or our Chicago office during the first week of employment...SuggestedPermanent employmentPart timeWork at officeLocal areaWorldwideFlexible hours$125.04k - $187.56k
...Technology and more. Overview The Site Reliability Engineer (SRE) III is responsible... ...through automation, observability, incident response, and infrastructure... ...3 in-person days at our Chicago office and 2 remote days.... ...(AWS, Azure, or GCP). Strong analytical, debugging...SuggestedFull timeWork at officeRemote workFlexible hours$86k - $105k
...Scottsdale, San Francisco, Chicago, or New York follow a... ...to be responsible for reliability, automation and... ...Implement and evangelize Observability and monitoring systems... ...prior DevOps, software engineering or related experience.... ...providers – AWS, GCP and Azure Background...SuggestedHourly payWork at officeImmediate startVisa sponsorshipWork visaFlexible hours- ...Site Reliability Engineer As a Site Reliability Engineer, you will build and... ...infrastructure primarily on GCP, but to include multi-cloud... ...Response: Implement comprehensive observability strategies using Prometheus... ...Engineer salary in the Chicago market with company-paid...Suggested
- ...landscape. Role Overview We are seeking an experienced Observability / Site Reliability Engineer (SRE) to design, scale, and maintain our enterprise... ...-leading cloud-native tools. Key Responsibilities GCP & Cloud Management: Architect, optimize, and maintain observability...SuggestedRemote jobContract work
$172.5k - $260.1k
...Salesforce.Lead Software Engineer – Platform ServicesAt... ...design, testing, observability, scalability, and operational... ...building scalable, reliable, and secure customer-... ...microservices on AWS (preferred), GCP, or other public cloud... ...Palo Alto; Illinois - Chicago; New York - New York;...Full timePart time$197.3k - $313.7k
...As the Principal Engineer focused on architecture... ..., performance, observability, and user... ...and role based sites such as the developer... ...operate accurately and reliably.Critically evaluate... ...AWS (preferred), GCP, or other public cloud... ...Alto; Illinois - Chicago; New York - New York...Full timePart time$112.5k - $187.5k
...to a DevOps Director. The Site Reliability Engineering team drives reliability strategy... ...on the platform. As a Staff Site Reliability Engineer... ...bring deep expertise across GCP, Kubernetes, CI/CD... ...level command of monitoring, observability, and alerting platforms (e....Full timeTemporary workWork experience placementWork at officeFlexible hours2 days per week$157.9k - $282.1k
...Principal Full Stack Engineer, AI Platform & AgentsBuild... ...you ship (latency, reliability, hallucination reduction... ...(primary), Azure, GCP; Docker, Terraform, GitHub... ...tooling, CI/CD, and observability for safe, fast iteration... ...: USA - Chicago, IL, West Adams St; USA...Full timePart timeWork at officeRemote work2 days per week$150k - $200k
...Financial Technology We are seeking a Site Reliability Engineer to join our team and assist with the design... ...in cloud technology (AWS and GCP experience preferred) Experience with... ...skills This role must sit in the firms Chicago office. Seniority level Seniority level...Full timeWork at office- ...message the job poster from Algo Capital Group Senior Site Reliability Engineer - Observability and Automation A leading high-frequency trading firm is... ...Kubernetes experience to join their Infrastructure team in Chicago focusing on observability and automation. The position...Full timeWork at officeFlexible hours
- ...Overview: Senior Site Reliability Engineer (SRE) Location: Chicago, IL (Onsite) Type: Contract Role Overview: We are seeking a Senior... ...expertise in AWS infrastructure, automation, observability, and production support . The ideal candidate will...Contract work
- ...Site Reliability Engineer (SRE) Immediate need for a talented Site Reliability Engineer (SRE). This... ...with long-term potential and is in Chicago, IL (Hybrid). Key Requirements and... ...Apps - Must ~8+ years Monitoring & Observability tools of which 3+ years with Grafana,...Contract workLocal areaImmediate start
$190.8k - $267.1k
...corners of the internet. As a Senior Site Reliability Engineer on Reddit’s Infrastructure SRE team,... ...closely with the Compute, Traffic, and Observability infrastructure teams. They will own a... ...portal. Locations: OneTrust Chicago, Illinois; Toronto, Ontario. #J-18808...Work experience placementHome officeFlexible hours$85k - $130k
...Site Reliability Engineer Passionate about precision medicine and advancing the healthcare industry... ...managing cloud infrastructure in AWS, GCP, or Azure You have built and deployed... ...either in person or via Slack Chicago: $85,000 - $130,000 The expected salary...$150k - $155k
...Site Reliability Engineer Hybrid (3 days onsite, 2 days remote) full‑time. No visa sponsorship. Base... ...Site Reliability Engineer focused on observation, logging, and capacity planning. The role... ...environments like AWS (preferred), Azure or GCP Experience with AIOps and predictive...Full timeWork experience placementRemote workVisa sponsorship$127k - $249k
THE TEAM Platform Engineering is the department within SRE that is... ...internal service mesh), and observability and alerting systems. The... ...components that ensure cluster reliability and security (e.g., CoreDNS,... ...platforms, including AWS, GCP, or Azure * Proficiency in...Work at officeLocal areaRemote workWorldwideFlexible hours- ...Job Title: Site Reliability Engineer Location: Chicago, IL FTE Only Job Description Must Have Technical/Functional... ...deep experience in AWS infrastructure, automation, observability, and production support. As an SRE, you will ensure...
$171.5k
...infrastructure that power secure, observable GenAI capabilities... ...product and domain engineering teams, you will... ...as part of scalable, reliable RAG and retrieval workflows... ....Partner closely with staff and principal... ...for this position in Chicago is $171,500.00 to $240...Full timePart time- ...the Role:As a Senior Software Engineer II at Confluent, you will... ...decisions that thoughtfully balance reliability, scalability, performance,... ...SLA running across 100+ AWS, GCP, and Azure regions.Partner... ...KLocationBoston, Massachusetts; Chicago, Illinois; Dallas, Texas;...Part timeImmediate startRemote work
$165k - $247k
...OpportunityAs a Senior Software Engineer, Applied AI, you’ll play a... ...excellence standards (SLOs, observability, incident response frameworks... ...teams to ensure services scale reliably across the org.What You’ll... ...EKS), and cloud platforms (AWS/GCP/Azure).Familiarity with secure...Part timeLocal areaWorldwideFlexible hours$171.5k
...growing our Advertiser Experience engineering team and are looking for a... ...with strong performance, reliability, trust, and transparency.Own... ...code quality, testing, observability, system reliability, and model... ...this role is only available in Chicago, IL, in alignment with our...Full timePart timeWork at officeRelocation packageFlexible hours3 days per week$170k - $250k
...Chicago, IllinoisOnsiteDirect Hire$170k - $250k A growth stage investment firm with a strong engineering culture is hiring a senior, hands-on technical leader to join its internal group... ...experience with cloud platforms (AWS, GCP, or Azure) Ability to operate...Full timePart time- ...8+ years of Software Engineering experience, or equivalent... ...and maintain scalable and reliable infrastructure on Google Cloud Platform (“GCP”) for Snowflake data warehousing... ..., IT management and staff, and other groups in... ...Willingness to work on-site at stated location in the...Contract workFor contractorsWork experience placement
$231.4k - $331.8k
...the Team The Platform Engineering organization is responsible... ..., automation, observability, security, and self-service... ...and operate services reliably at scale. We are a... ...infrastructure across AWS, Azure, GCP, or hybrid cloud... ...see the Cisco careers site to discover more...Permanent employmentFull timeTemporary workPart timeLocal areaFlexible hours$194k - $267k
...in on this mission. If you are too, let's talk.The TeamThe Site Reliability team is dedicated to architecting and owning the foundational... ...durable, automated systems that maximize platform reliability and engineering velocity.The ideal candidate is someone who enjoys analyzing...Part timeLocal areaWorldwideFlexible hours- ...Front-of-House and Heart-of-House support staff, managers, and chefs Properly sets-up... ...departmental meetings Learn by listening, observing other team members, and sharing knowledge... ...time Locations 632 N Dearborn St, Chicago, IL, 60610, US Minimum Salary 17.87...Part timeWork at officeImmediate startRemote workAll shiftsShift workAfternoon shift
- ...Chicago, IllinoisHybridDirect Hire$200k - $220kAn innovative... ...a Senior DevOps Engineer to help scale the... ...helping deliver highly reliable and secure healthcare... ...experience in DevOps, Site Reliability Engineering... ...Experience with monitoring and observability platforms Experience...Full timePart time
$182k - $273k
...AVP Software Engineering - IE05FEWe’re determined to make a difference... ..., built, and scaled safely, reliably, and in alignment with... ...Hartford CT, Charlotte, NC or Chicago IL) will have the expectation... ...readiness through robust testing, observability, monitoring, and incident...Full timeTemporary workPart timeWork at officeRemote work3 days per week- ...operating across Chicago, San Francisco, Bangalore... ...Forward-deployed engineering. AI coding tools... ...Cloud Platform (GCP), leveraging managed services for reliable, scalable backend... ...services ~ Improve observability, monitoring, and... ...our team on-site, in our beautiful...Work at officeVisa sponsorship3 days per week
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff Site Reliability Engineer - Observability GCP (Chicago). Be the first to apply!
- engineering aide Chicago, IL
- software engineer staff Chicago, IL
- assistant engineer Chicago, IL
- technology administrator Chicago, IL
- senior staff systems engineer Chicago, IL
- staff engineer Chicago, IL
- site reliability engineer Chicago, IL
- site reliability engineer sre Chicago, IL
- construction site safety Chicago, IL
- site recruiter Chicago, IL











