Staff Site Reliability Engineer
Informatic Technologies
Staff Site Reliability Engineer (SRE) – Platform Engineering
We are seeking a Staff Site Reliability Engineer to serve as the foundational Technical Lead for our Platform Engineering SRE organization. In this role, you will be the primary architect and visionary for the core technology foundations that underpin the world's leading derivatives marketplace. As the technical lead for all SRE sub-teams, you will bridge the gap between high-level business goals and deep technical implementation, ensuring our GCP-native stack provides the mission-critical intersection of ultra-low latency and absolute reliability required for high-volume financial ecosystems. As the Staff SRE, your goal is to evolve our platform from "infrastructure as a service" to "reliability as a product." You will be responsible for the technical roadmap of our entire SRE domain, mentoring senior engineers and setting the global standard for operational excellence across our Python, Kafka, and Kubernetes shop.
What You Will Do
- Technical Vision & Roadmap: Define the 12–18 month technical strategy for the Platform SRE teams, focusing on the evolution of our global footprint and self-service capabilities.
- Architectural Authority: Act as the final technical authority for major infrastructure changes involving GCP, GKE, and our mission-critical Kafka event bus.
- Incident Command & Systemic Resilience: Lead the response for complex, cross-functional outages and drive a "blameless" culture that prioritizes systemic, code-based fixes over manual intervention.
- Internal Development Platform (IDP): Architect and oversee the building of high-level abstractions in Python to mask underlying complexity, providing a seamless "Golden Path" for our global technology stack.
- Reliability Governance: Standardize SLIs, SLOs, and Error Budgets across all platform teams, ensuring they are technically rigorous and directly tied to market integrity.
- Engineering Mentorship: Level up the entire SRE organization through design reviews, architectural "office hours," and fostering an environment of continuous technical evolution.
What We're Looking For
- Strategic AI Integration: A mastery of leveraging Generative AI and Agentic workflows (e.g., Gemini) to build self-healing infrastructure and sophisticated automated troubleshooting frameworks.
- Software Engineering Mastery: Expert-level proficiency in Python (and ideally Go) to build production-grade distributed systems and custom Kubernetes operators.
- Cloud-Native Leadership: Deep-seated expertise in GCP (Networking, IAM, GKE) and the ability to scale Kafka clusters for high-throughput, low-latency financial environments.
- Advanced IaC & GitOps: Mastery of Terraform module design and ArgoCD for managing immutable infrastructure at an enterprise scale.
- Distributed Systems Theory: A rigorous understanding of non-linear system behaviors, distributed consensus, and the nuances of high-concurrency architectures.
- Executive Communication: The ability to translate sophisticated technical debt and architectural risks into clear business outcomes for senior leadership.
Experience:
- 10+ years in SRE, Systems Engineering, or Software Engineering roles within high-pressure environments.
- 3+ years in a Staff, Principal, or Tech Lead capacity overseeing multiple teams or complex platform domains.
- Proven Track Record: Experience leading large-scale cloud migrations or re-architecting core messaging/compute platforms in a regulated environment.
- Certifications: GCP Professional Cloud Architect or Kubernetes (CKA/CKAD).
- Full-Stack Exposure: Proficiency in Node.js or modern front-end frameworks.
- Domain Expertise: Experience in Financial Markets or highly regulated, high-concurrency environments.
- Agile Integration: Comfort working within Agile frameworks and highly collaborative software development lifecycles.
- ...applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the Commercial and Investment Banking and Payment Technology Team, you will solve complex...SuggestedFull timeLocal area
$158.5k - $172k
...exceptional value they deserve. About The Opportunity As a Senior Engineer on the Runtime Automation team, you will design, automate, and... .... This is a high-impact position driving continuous reliability, deep system optimization, and automation across our entire technology...SuggestedFull timeWork at office3 days per week- Play a key role in ensuring system reliability at one of the world’s most iconic and largest financial institutions. As a Site Reliability Engineer II at JPMorgan Chase within the Commercial and Investment Banking and Payment Technology Team, you will use technology...SuggestedFull timeLocal area
$125.04k - $187.56k
...services, including Finance, Legal, Sustainability, Commercial, Digital and E-commerce, Technology and more. Overview The Site Reliability Engineer (SRE) III is responsible for ensuring the scalability, reliability, and performance of production systems through automation...SuggestedFull timeWork at officeRemote workFlexible hours- ...building and running systems that must perform reliably under real-time market conditions. The culture is highly collaborative, engineering-driven, and focused on continuous... ...a related field 3+ years of experience in site reliability, systems engineering, or technical...Suggested
$190.8k - $267.1k
...unique opportunity to leave your mark on one of the most influential and trafficked corners of the internet. As a Senior Site Reliability Engineer on Reddit’s Infrastructure SRE team, you’ll use your knowledge of distributed systems and architecture to improve the reliability...Work experience placementHome officeFlexible hours- ...Okta is seeking a Staff Site Reliability Engineer to join the TCore team in Chicago. You will design scalable network solutions, maintain a highly available cloud edge for the identity platform, and automate infrastructure with Terraform and Chef. You will analyze data...
$164.6k - $288k
...of Practice (CoP) Senior Implementation Lead is responsible for driving the adoption, standardization, and maturity of Site Reliability Engineering (SRE) practices across the organization. This role serves as a key enabler in scaling SRE principles, fostering collaboration...Visa sponsorshipWork visa$150k - $155k
...Site Reliability Engineer Hybrid (3 days onsite, 2 days remote) full‑time. No visa sponsorship. Base pay: $150,000 – $155,000 per year, subject to skills and experience. A prestigious company seeks a Site Reliability Engineer focused on observation, logging, and capacity...Full timeWork experience placementRemote workVisa sponsorship- ...Chicago. This pivotal role involves leading edge infrastructure operations, collaborating across teams to ensure high availability and reliability of the platform. Candidates should have 10+ years in infrastructure roles, strong leadership skills, and knowledge in cloud...
- ...upon to keep lives moving forward when it matters most. Learn more about CCC at **The Role**We are seeking a talented Sr. Site Reliability Engineering Developer to be part of the fast moving, innovative CCC Site Reliability Team. We build enterprise class, hosted...Night shift
- ...Get AI-powered advice on this job and more exclusive features. Direct message the job poster from Algo Capital Group Senior Site Reliability Engineer - Observability and Automation A leading high-frequency trading firm is seeking a mid to senior-level Site Reliability...Full timeWork at officeFlexible hours
- ...Site Reliability Engineer (SRE) Immediate need for a talented Site Reliability Engineer (SRE). This is a 12+ months contract opportunity with long-term potential and is in Chicago, IL (Hybrid). Key Requirements and Technology Experience: ~ Must have skills:...Contract workLocal areaImmediate start
- ...Overview: Senior Site Reliability Engineer (SRE) Location: Chicago, IL (Onsite) Type: Contract Role Overview: We are seeking a Senior Site Reliability Engineer (SRE) with strong expertise in AWS infrastructure, automation, observability, and production...Contract work
$152.5k - $219.2k
...global cloud platform. As a team of six engineers distributed across the US, Canada, and the... ...with a strong focus on automation, reliability, and operational excellence. We are one... ...Qualifications ~2+ years of experience in Site Reliability Engineering, DevOps, Infrastructure...Permanent employmentFull timeTemporary workLocal areaWorldwideFlexible hours$85k - $130k
...Site Reliability Engineer Passionate about precision medicine and advancing the healthcare industry? Recent advancements in underlying technology have finally made it possible for AI to impact clinical care in a meaningful way. Tempus' proprietary platform connects...- ...Site Reliability Engineer As a Site Reliability Engineer, you will build and secure infrastructure supporting our AI platform with special attention to safeguarding US customer data and supporting the Aerospace and Defense Industrial Base. You'll have strong ownership...
- ...with the investigation, Always validate if the team is following the SOPs or the process defined for an alerts/ issue. Contacting all the external vendors if in case their integrations fail, Measure the front-end metrics for the site with various tools available...
$250k - $350k
...and India, where quantitative researchers, engineers, traders, and operational teams work together... ...that boost stability, throughput, and reliability Qualifications Minimum of 3 years’ experience in production support, site reliability, or infrastructure operations in...Full time- ...Qualifications: 8+ years of software engineering experience, or equivalent... ...and maintain scalable and reliable infrastructure on Google... ...the client, IT management and staff, and other groups in Information... .... Willingness to work on-site at stated location in the job...For contractorsWork experience placement
- ...Edward Jones Site Reliability Engineer 100% remote Initial contract is 6 months, but will be a multi year engagement. Position Overview: As a Senior Site Reliability Engineer, you will play a critical role in ensuring the reliability, availability, and performance...Contract workRemote work
- ...Job Title: Site Reliability Engineer Location: Chicago, IL FTE Only Job Description Must Have Technical/Functional Skills ~ We are looking for a Senior Site Reliability Engineer (SRE) with deep experience in AWS infrastructure...
$130k - $170k
...Senior Site Reliability Engineer About Us Founded in 2014, we offer the industry’s first and only cloud‑based, fully‑customisable, end‑to‑end software solution to automate securities‑based lending from origination through the life of the loan. By combining thought leadership...Full timeFlexible hoursShift work- ...Senior Site Reliability Engineer We are looking for a Senior Reliability Engineer to join our Platform team. In this position, you will be responsible for maintaining, designing, implementing and upgrading our cloud infrastructure to support our microservices platforms...Temporary workFlexible hours
$97.5k - $130k
...Overview: Site Reliability Engineer Salary: $97,500-$130,000 Role Summary The SRE & Cloud Engineer will support the client by accelerating remediation of security issues and contributing to cloud modernization initiatives. This role is hands-on and execution...$86k - $105k
...generation of application infrastructure and to be responsible for reliability, automation and scalability using and the latest best... ...certifications. Minimum of 2 years prior DevOps, software engineering or related experience. Must be able to work different schedules...Hourly payWork at officeImmediate startVisa sponsorshipWork visaFlexible hours$114k - $155k
...applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the Commercial and Investment Banking and Payment Technology Team, you will solve complex and...Local area$150k - $200k
...message the job poster from Selby Jennings Recruitment Consultant @ Selby Jennings | Financial Technology We are seeking a Site Reliability Engineer to join our team and assist with the design, development, and administration of our trading and research systems. This...Full timeWork at office$112.5k - $187.5k
...TransUnion, this role will report to a DevOps Director. The Site Reliability Engineering team drives reliability strategy, elevates engineering... ...most complex and consequential work on the platform. As a Staff Site Reliability Engineer at TransUnion, you will serve as...Full timeTemporary workWork experience placementWork at officeFlexible hours2 days per week$102.6k - $193.43k
...Software Engineer Chamberlain Group (CG) is a global leader in intelligent access and Blackstone portfolio company. Powered by our... ...Establish working relationships with business members, technical staff, and subject matter experts to help create and execute the infrastructure...Temporary workSecond jobWorldwide
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff Site Reliability Engineer. Be the first to apply!
- engineering aide Chicago, IL
- software engineer staff Chicago, IL
- assistant engineer Chicago, IL
- technology administrator Chicago, IL
- senior staff systems engineer Chicago, IL
- staff engineer Chicago, IL
- site reliability engineer Chicago, IL
- site reliability engineer sre Chicago, IL
- construction site safety Chicago, IL
- site recruiter Chicago, IL

