Observability Engineer / Site Reliability Engineer
Ontrac Solutions
About Ontrac Solutions
Ontrac Solutions is a leading technology consulting firm, specializing in cutting-edge solutions that drive business transformation. We partner with organizations to modernize their infrastructure, streamline processes, and deliver tangible results. By creating value beyond the hype, we help businesses modernize technology and build new strategies that fuel growth. Our team is committed to innovation, collaboration, and excellence, empowering our clients to succeed in an evolving digital landscape.
Role Overview
We are seeking an experienced Observability / Site Reliability Engineer (SRE) to design, scale, and maintain our enterprise monitoring and alerting ecosystems. In this role, you will bridge the gap between development and operations by ensuring high availability, performance tuning, and deep visibility across distributed multi-cloud and native systems. You will play a critical role in automating infrastructure and building robust observability pipelines using industry-leading cloud-native tools.
Key Responsibilities
- GCP & Cloud Management: Architect, optimize, and maintain observability frameworks across cloud environments, with a specific focus on implementing Google Cloud Platform (GCP) observability tools (Cloud Logging, Cloud Monitoring, Trace, and Profiler).
- Platform Management: Design, deploy, and maintain robust observability stacks across hybrid ecosystems, utilizing Prometheus, Grafana, and cloud-native integrations.
- Automation & IaC: Drive infrastructure-as-code (IaC) initiatives using Terraform and Ansible to ensure consistent, automated deployments of infrastructure and observability tooling.
- CI/CD Integration: Build, maintain, and optimize deployment workflows within Kubernetes and Google Kubernetes Engine (GKE) / OpenShift environments using GitHub, Harness, and other CI/CD pipelines.
- System Performance: Deeply analyze Linux/Unix system administration architectures, optimizing compute resource metrics and performance tuning across complex, distributed environments.
- SRE Evangelism: Implement SRE best practices, establishing meaningful Service Level Indicators (SLIs), Service Level Objectives (SLOs), and Error Budgets to ensure platform reliability.
Required Skills & Qualifications
- Cloud Infrastructure: Proven engineering experience within Google Cloud Platform (GCP) environments, particularly managing cloud-native monitoring and compute resources.
- Observability Tooling: Hands-on experience with Grafana, Prometheus, and Google Cloud Observability suites. Direct experience with GEM (Grafana Enterprise Metrics) is highly desirable.
- OS & Scripting: Expert-level knowledge of Linux/Unix operating systems paired with strong shell scripting skills for automation and systems management.
- Programming: Professional coding proficiency in at least one modern language (Python, Go, Java, Perl, or advanced Shell).
- Containers & Orchestration: Hands-on experience managing containerized applications on Kubernetes, GKE, and/or Red Hat OpenShift.
__________________________________
(Ontrac Solutions has partnered with PinpointVerify to help genuine applicants rise above the noise. Today, qualified candidates are too often overshadowed by fake and fraudulent applications. PinpointVerify gives our recruiters confidence that you are exactly who you say you are — and gives you a portable verification credential you can share with any employer.
Applicants who complete verification are prioritized over non-verified candidates with comparable experience. And if you're hired, Ontrac reimburses the full cost of your verification.
Get verified →
- Play a key role in ensuring system reliability at one of the world’s most iconic and... ...largest financial institutions.As a Site Reliability Engineer II at JPMorgan Chase within the Commercial... ...application codeUnderstands observability patterns and strives to implement and...Suggested
$130k - $180k
...belonging, collaboration, and accomplishment.Being a Senior Site Reliability Engineer at iManage Means… You are an engineer, a builder, and a... ...in on-call rotations. You’ll be a key voice in observability, change management, and service scalability, providing guidance...SuggestedWork at officeLocal areaRemote workWorldwideMonday to FridayFlexible hours$158.5k - $172k
...deserve.About The OpportunityAs a Senior Engineer on the Runtime Automation team, you... ...system engineering, while ensuring top-tier observability and strict security across our mission-... ...-impact position driving continuous reliability, deep system optimization, and automation...SuggestedFull timeTemporary workWork at officeFlexible hours3 days per week$100.7k - $167.8k
Job SummaryThe Site Reliability Engineer III is a pivotal architect of stability for CME Clearing & Risk. You will engineer secure, scalable... ...defending SLIs, SLOs, and SLAs to maintain system health.Observability Mindset: Experience navigating the telemetry landscape using...SuggestedFull timeWorldwide$108.08k - $172.5k
Work with development and platform engineering teams to migrate and maintain applications in Google Cloud. Apply Observability concepts and applications to maintain services. Monitor metrics, system health and analyze reports. Provide on-call rotation support for production...SuggestedFull timeRemote workWorldwide- ...the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Commercial and... ...discipline (e.g., Cloud, AI, Android, etc.)Experience in observability such as white and black box monitoring, service level objective...
$100k - $120k
OverviewThe Site Reliability Engineer is a key force behind improving Origami’s time to resolution and advancing overall site reliability and... ...delivery.Cross trains colleagues on how to best leverage observability tools during incident and performance investigations....Full timeTemporary workWork experience placementFlexible hours- ...Edward Jones Site Reliability Engineer 100% remote Initial contract is 6 months, but will be a multi year engagement. Position Overview... ...with development teams to integrate reliability and observability best practices into the software development lifecycle....Contract workRemote work
$130k - $225k
...consensus.The Algorithmic Trading Team is looking for a Site Reliability Engineer for our Chicago office. The SRE team is critical to the success... ..., trading and infrastructure - someone who creates observability that surfaces issues before they ever cause a problem, owns...Temporary workWork at officeFlexible hours$127k - $249k
The TeamPlatform Engineering is the department within SRE that is responsible for a range... ...edge and internal service mesh), and observability and alerting systems.The Fleet Management... ...components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager...Work at officeLocal areaRemote workWorldwideFlexible hours$130k - $150k
...technologies is essential for this role. Position OverviewThe Site Reliability Engineer (SRE) helps ensure CRA’s critical business services are... ...reduce manual toil through automation, improve service observability, and strengthen incident response. The SRE partners...Work at officeWork from home3 days per week$98.18k - $115.5k
...communities, and each other. Job DescriptionResponsibilitiesThe Reliability Observability Engineer 3 is responsible for enabling reliable, measurable, and... ....Partner with Product Owners, Application Engineering, Site Reliability Engineering (SRE), and Operations Teams to...Full timeWork experience placementLocal area3 days per week$204k - $306k
...If you are too, let's talk.Manager, Site Reliability EngineeringSan Francisco, CaliforniaSecure... ...Office. The IDaaS Site Reliability Engineering GroupOkta authenticates, authorizes... ...Edge networking, K8s platform, CI/CD, Observability, automation platform & tooling. What...Permanent employmentWork at officeLocal areaWorldwideFlexible hours2 days per week$194k - $267k
...-educate on new concepts and tools.Position Overview:The Site Reliability Engineer (SRE) will play a key role in building and managing Kubernetes... ...provide service-to-service communication, security, and observability within the Kubernetes clusters. Enable fine-grained...Permanent employmentWork at officeLocal areaWorldwideFlexible hours- ...world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the Corporate Technology team... ...tools such as Python, Ansible and Terraform.Experience in observability including white and black box monitoring, service level...
$194k - $267k
...let's talk.The TeamWe are looking for an experienced Staff Site Reliability Engineer to join Okta's Emerging Products Group (EPG). Our mission... ...mindset and continuously invest in platform engineering, observability, and operational excellence to enable our engineering...Local areaWorldwideFlexible hours$112.5k - $187.5k
...TransUnion, this role will report to a DevOps Director. The Site Reliability Engineering team drives reliability strategy, elevates engineering... ...and built for scale.Expert-level command of monitoring, observability, and alerting platforms (e.g., Datadog, Prometheus,...Full timeTemporary workWork experience placementWork at officeFlexible hours2 days per week$194k - $267k
...:We are seeking a highly technical StaffObservabilitySite Reliability Engineer with a specialty in Splunk to own and evolve our Splunk ecosystem... ...to delivering a world class, comprehensive, scalable Observability Platform that enables our SRE teams and business partners....Permanent employmentWork at officeLocal areaWorldwideFlexible hours$160k - $210k
...you'll do:Join our Platform Engineering team, where you'll ensure the... ...mentoring engineers across reliability initiativesAnalyze, troubleshoot... ...of experience in DevOps, Site Reliability Engineering, or... ...environmentsKnowledge of monitoring and observability tools such as Prometheus,...Work at officeWorldwideMonday to FridayFlexible hours- ...and companies, alikeKlover’s engineering team powers one of the... ...grade systems that prioritize reliability, security, and performance,... ...candidateAbout the RoleAs a Senior/Staff Site Reliability Engineer, you... ..., a strong dedication to observability, and a laser-focus on...Work at officeImmediate startRemote work
$61k - $101k
...training or certification in software engineering concepts, plus at least 2 years of applied... .... We look for familiarity with site reliability concepts, principles, and practices. We look for familiarity with observability practices, including white-box and black...Full timeLocal area$250k - $350k
...where quantitative researchers, engineers, traders, and operational... ...stability, throughput, and reliability Qualifications Minimum of 3... ...experience in production support, site reliability, or... ...production setting Familiarity with observability tools (e.g., Prometheus), and...Full time$190.8k - $267.1k
...influential and trafficked corners of the internet. As a Senior Site Reliability Engineer on Reddit’s Infrastructure SRE team, you’ll use your... ...will work very closely with the Compute, Traffic, and Observability infrastructure teams. They will own a suite of tools for...Work experience placementHome officeFlexible hours- ...powered advice on this job and more exclusive features. Direct message the job poster from Algo Capital Group Senior Site Reliability Engineer - Observability and Automation A leading high-frequency trading firm is seeking a mid to senior-level Site Reliability Engineer...Full timeWork at officeFlexible hours
$232k - $319k
...scale the service with great people and reliable, cost-effective, and efficient... ...focused on Edge networking, K8s platform, Observability, automation platform & tooling. What you... ...serviceAccelerate the velocity of SRE and product engineering by developing robust platforms,...Permanent employmentLocal areaWorldwideFlexible hours$139.23k - $163.8k
...One.Job DescriptionJob SummaryThe Lead Engineer (Generative AI) is a senior technical role... ...and release managementMonitoring, observability, and continuous optimizationImplement GenAIOps... ...best practices to ensure scalability, reliability, and cost efficiencyEstablish logging,...Full timeWork experience placementLocal area3 days per week- ...IL, United StatesIndustry: Trading FirmPosted: 2026-08-17Contact: Mike LaTulipEmail: ****@*****.*** Title: Site Reliability Engineer (Infrastructure & Systems)Location: Chicago, IL (Greater Metro Area)About the OpportunityJoin a premier financial technology...Local area
- Qualifications: 8+ years of Software Engineering experience, or equivalent demonstrated through... ...implement and maintain scalable and reliable infrastructure on Google Cloud Platform... ...vendor resources Willingness to work on-site at stated location in the job openingDepartment...Contract workFor contractorsWork experience placement
$132.1k - $220.1k
Staff Site Reliability Engineer (SRE) - Platform EngineeringNote: This position follows a hybrid work model, requiring 2 days per week on-site at our corporate office 20 S Wacker Dr, Chicago, IL 60606The first preference for this role is given to local candidates in the...Full timeWork at officeLocal areaWorldwide2 days per week$140k - $170k
We are looking for a Senior Site Reliability Engineer to work as part of a lean, product‑focused engineering organization. This role is about... ...monitoring, logging, and alerting that make system behavior observable and actionable Use AI‑assisted tools to accelerate...Full timeWork experience placementFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Observability Engineer / Site Reliability Engineer. Be the first to apply!
- site reliability engineer Chicago, IL
- site reliability engineer remote Chicago, IL
- site reliability engineer sre Chicago, IL
- site recruiter Chicago, IL
- junior website developer Chicago, IL
- on site coordinator Chicago, IL
- construction site safety Chicago, IL
- site services specialist Chicago, IL
- website content developer Chicago, IL
- website coordinator Chicago, IL


