Site Reliability Engineering Manager
ClifyX
Site Reliability Engineering Manager MIDWEST IL - CHICAGO
The Performance Engineering practice within Technology is focused on optimizing the performance and scalability of enterprise applications through the combination of testing, diagnostics & monitoring, performance analytics, and business optimization. As a Site Reliability Engineering Manager, some of your key responsibilities may include defining the strategy for enabling performance diagnostics and monitoring through the use of an Application Performance Management (APM) tool, other monitoring tools, and diagnostic techniques. You will also identify, evaluate, and recommend monitoring tools and diagnostic techniques relevant to the application architecture. Additionally, you will interact with client and/or development, operations, and infrastructure resources to recommend solutions to remediate performance issues. Participating in re-architecture, redesign, and refactoring decisions to satisfy performance requirements is also part of your role. You will develop dashboards and reports to provide ongoing visibility into the performance of client applications.
Qualifications Basic Qualifications: 7 years hands-on design/development/engineering experience (e.g. Java,.Net, etc.), 1 year hands-on experience performance monitoring & diagnostic tools (e.g. AppDynamics, Dynatrace, New Relic, CA APM (previously Wily Introscope), etc.), minimum of 2 years of Team Lead experience leading a team with experience in Project Planning, Estimating or Project Management. Bachelor's degree or equivalent (minimum 12 years work experience). If Associate's Degree, must have equivalent minimum 6 years work experience.
Preferred Qualifications: Previous Consulting experience, experience with Agile and DevOps, experience with logging solutions, including ELK and Splunk, experience with open source monitoring and visualization systems and tools, i.e. Prometheus (monitoring + tracing), Grafana/Kibana (dashboards), Zipkin (distributed tracing), etc., experience with stream-processing open source frameworks/systems, i.e. Kafka, Spark, etc. Knowledge of defining and monitoring system quality measures, including SLO and SLA, understanding or exposure to Chaos Engineering Tools (Chaos Toolkit, Gremlin, Simian Army, Etc.), experience with distributed computing, Web Services, SOA, and JEE design concepts, experience delivering software designed for high concurrency, scalability, or availability, hands-on experience collecting performance data, analyzing, troubleshooting, and tuning, experience with different flavors of Linux, i.e. RedHat, Ubuntu, CentOS, etc. Built tooling to improve reliability of systems, automated remediation of issues, or improve scalability, systems often need to be reconfigured, so you should have experience with a configuration management system like Puppet, Chef or Salt. Experience with usage of common application protocols and messages (e.g. TCP/IP, SOAP, RESTful APIs, XML/JSON, JDBC, JMS/MQ), exposure to Cloud, SaaS, and virtualization concepts and performance concerns, exposure to application threading and concurrency concerns, working knowledge of operating system design, processes, and threading model, ability to work in other languages such as JavaScript, Ruby, PHP, Perl, Python, PowerShell, and Linux shell scripting, experience with Amazon Web Services, experience with Containers (kubernetes & docker).
$204k - $306k
...We're all in on this mission. If you are too, let's talk.Manager, Site Reliability EngineeringSan Francisco, CaliforniaSecure Every Identity,... ...week in our San Francisco Office. The IDaaS Site Reliability Engineering GroupOkta authenticates, authorizes and provisions...SuggestedPermanent employmentWork at officeLocal areaWorldwideFlexible hours2 days per week$232k - $319k
...scale the service with great people and reliable, cost-effective, and efficient infrastructure... ..., processes, and tooling. As the Sr. Manager of Infrastructure Platform and Shared... ...serviceAccelerate the velocity of SRE and product engineering by developing robust platforms, powerful...SuggestedPermanent employmentLocal areaWorldwideFlexible hours$158.5k - $172k
...deserve.About The OpportunityAs a Senior Engineer on the Runtime Automation team, you will... ...ecosystems. Our team is responsible for managing our centralized Enterprise Logging... ...high-impact position driving continuous reliability, deep system optimization, and automation...SuggestedFull timeTemporary workWork at officeFlexible hours3 days per week$100k - $120k
OverviewThe Site Reliability Engineer is a key force behind improving Origami’s time to resolution and advancing overall site reliability and... ...role.Strong knowledge of SRE best practices and incident management protocolsDeep experience using and/or configuring New Relic...SuggestedFull timeTemporary workWork experience placementFlexible hours$130k - $150k
...technologies is essential for this role. Position OverviewThe Site Reliability Engineer (SRE) helps ensure CRA’s critical business services are... ...including Windows Server, VMware vSphere, VMware Site Recovery Manager (SRM), SAN technologies, and the Rubrik ecosystem, with the...SuggestedWork at officeWork from home3 days per week- Play a key role in ensuring system reliability at one of the world’s most iconic and... ...largest financial institutions.As a Site Reliability Engineer II at JPMorgan Chase within the Commercial... ...transaction processing and asset management. We offer a competitive total...
$130k - $180k
...belonging, collaboration, and accomplishment.Being a Senior Site Reliability Engineer at iManage Means… You are an engineer, a builder, and a... ...rotations. You’ll be a key voice in observability, change management, and service scalability, providing guidance during complex...Work at officeLocal areaRemote workWorldwideMonday to FridayFlexible hours- ...the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Commercial and Investment... ...banking, financial transaction processing and asset management. We offer a competitive total rewards package including base...
$108.08k - $172.5k
Work with development and platform engineering teams to migrate and maintain applications in Google Cloud. Apply Observability concepts... ...rotation support for production systems, facilitate incident management and conduct post-incident reviews. Drive, contribute and...Full timeRemote workWorldwide- ...Contact: Mike LaTulipEmail: ****@*****.*** Title: Site Reliability Engineer (Infrastructure & Systems)Location: Chicago, IL (Greater... ...(AWS or Google Cloud Platform / GCP).Familiarity with managed container orchestrators such as Amazon EKS or Google GKE.Exposure...Local area
- Qualifications: 8+ years of Software Engineering experience, or equivalent... ...and maintain scalable and reliable infrastructure on Google... ...effectively with the client, IT management and staff, and other groups in... ...resources Willingness to work on-site at stated location in the job...Contract workFor contractorsWork experience placement
$130k - $225k
...expectations, integrity, innovation and a willingness to challenge consensus.The Algorithmic Trading Team is looking for a Site Reliability Engineer for our Chicago office. The SRE team is critical to the success of our trading - ensuring that our production trading...Temporary workWork at officeFlexible hours$127k - $249k
The TeamPlatform Engineering is the department within SRE that is responsible for a range... ...observability and alerting systems.The Fleet Management team provides the core runtime... ...critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager...Work at officeLocal areaRemote workWorldwideFlexible hours$100.7k - $167.8k
Job SummaryThe Site Reliability Engineer III is a pivotal architect of stability for CME Clearing & Risk. You will engineer secure, scalable,... ...gap between development and operations, you ensure our risk management services remain resilient and high-performing for...Full timeWorldwide$194k - $267k
..., automate it” and who can rapidly self-educate on new concepts and tools.Position Overview:The Site Reliability Engineer (SRE) will play a key role in building and managing Kubernetes platforms that support cloud-native applications and services. This position focuses...Permanent employmentWork at officeLocal areaWorldwideFlexible hours$127k - $249k
...Central time zones. We are looking for an experienced Senior Engineer for our SRE, Atlas team to support, maintain and grow the Atlas... ...crucial workloads. Role OverviewWe are seeking a talented Site Reliability Engineer (SRE) with a strong infrastructure background. This...Local areaRemote workWorldwideFlexible hours- ...and companies, alikeKlover’s engineering team powers one of the fastest... ...systems that prioritize reliability, security, and performance, and... ...candidateAbout the RoleAs a Senior/Staff Site Reliability Engineer, you... ...metrics to our Google-managed Prometheus instance and build...Work at officeImmediate startRemote work
$160k - $210k
...you'll do:Join our Platform Engineering team, where you'll ensure the... ...mentoring engineers across reliability initiativesAnalyze, troubleshoot... ...provisioning, scaling, and management across all... ...years of experience in DevOps, Site Reliability Engineering, or...Work at officeWorldwideMonday to FridayFlexible hours$194k - $267k
...Overview:We are seeking a highly technical StaffObservabilitySite Reliability Engineer with a specialty in Splunk to own and evolve our Splunk... ....Required Skills & Experience (The Essentials)Log Management: Minimum 5+ Experience scaling and managing Splunk Cloud at...Permanent employmentWork at officeLocal areaWorldwideFlexible hours- ...the world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the Corporate Technology team... ...hands-on experience in supporting SRE practices for Data management/migration platforms and products, with familiarity in on-...
$132.1k - $220.1k
Staff Site Reliability Engineer (SRE) - Platform EngineeringNote: This position follows a hybrid work model, requiring 2 days per week on-site... ...GitOps: Mastery of Terraform module design and ArgoCD for managing immutable infrastructure at an enterprise scale.Distributed...Full timeWork at officeLocal areaWorldwide2 days per week$112.5k - $187.5k
...NoticePersonal Information We CollectYour Privacy ChoicesTeam OverviewAt TransUnion, this role will report to a DevOps Director. The Site Reliability Engineering team drives reliability strategy, elevates engineering standards, and owns some of the most complex and consequential...Full timeTemporary workWork experience placementWork at officeFlexible hours2 days per week$194k - $267k
...let's talk.The TeamWe are looking for an experienced Staff Site Reliability Engineer to join Okta's Emerging Products Group (EPG). Our mission... ...balancing, ingress, TLS, service networking, and traffic management.Strategic experience designing comprehensive observability...Local areaWorldwideFlexible hours$160k - $200k
...our efforts. Chicago, IL (Hybrid 3X a week) Senior Site Reliability Engineer (SRE) Chicago, IL (Hybrid) Opportunity Overview... ...AWS infrastructure supporting production applications. Manage infrastructure as code using Terraform . Build and maintain...Full timeWork at office- ...Edward Jones Site Reliability Engineer 100% remote Initial contract is 6 months, but will be a multi year engagement. Position Overview... ...of our systems. You will be responsible for incident management, root cause analysis, and implementing postmortem processes...Contract workRemote work
- ...Site Reliability Engineer in Wealth Management Chicago (IL) / Tempe (AZ) Onsite Job ROLE: This role will be Responsible for application observability, maintenance, and support, identifying and implementing preventive measures proactively, evaluates and...Flexible hours
- ...Job Description As a Senior DevOps / SRE Engineer on contract, you will be embedded with the Central Technology AI enablement team, working alongside engineers from Direct, PitchBook, Retirement, and other business units. Your initial focus will be on the SRE and hosting...Contract workImmediate start
$50 - $53 per hour
...Immediate need for a talented Site Reliability Engineer (SRE) This is a 12+ Months contract opportunity with long-term potential and is in... ...resolution status (written and verbal) to project team and management ~ Provide reactive, break-fix support Our client...Contract workLocal areaImmediate start- ...Senior Site Reliability Engineer We are looking for a Senior Reliability Engineer to join our Platform team. In this position, you will be... ...engineering organization to build automated processes and tools for managing application and service deployments Own and support...Temporary workFlexible hours
$118.3k - $219.8k
Are you excited to lead Site Reliability Engineering teams that keep mission-critical, 24/7 services running reliably and securely?Do you enjoy... ...teams to deliver proactive monitoring, effective incident management, and continuous optimisation. Through strong alignment...Full timeLocal area
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineering Manager. Be the first to apply!
- site reliability engineer Chicago, IL
- site reliability engineer remote Chicago, IL
- site reliability engineer sre Chicago, IL
- site agent Chicago, IL
- site manager Chicago, IL
- site superintendent Chicago, IL
- site director Chicago, IL
- site supervisor Chicago, IL
- site construction manager Chicago, IL
- hvac site manager Chicago, IL

