Senior Site Reliability Engineer
$106k - $150.18kCDW
Senior Site Reliability Engineer (SRE)
As a Senior Site Reliability Engineer (SRE) at CDW, you will improve the reliability, scalability, performance, security, and operational excellence of applications and platforms supporting Managed Services. This includes customer-connectivity platforms that enable CDW service delivery teams to monitor, troubleshoot, and manage customer environments. You will serve as a senior technical escalation point, resolving complex operational issues, reducing operational demands on development teams, and partnering with software engineering, infrastructure, security, and business stakeholders to improve service resilience. This role combines software engineering, systems engineering, automation, and operational leadership to reduce risk, improve the customer experience, and accelerate delivery. The Senior SRE will also mentor engineers, establish reliability engineering practices, and influence architectural and operational decisions across the organization.
What You Will Do:
- Drive service reliability, scalability, performance, security, and operational excellence across Managed Services applications and customer-connectivity platforms.
- Serve as a senior escalation point for complex operational issues, resolving incidents that exceed the current SRE team's expertise and reducing operational interruptions for development teams.
- Design, develop, and maintain automation using Ansible and Python to reduce operational toil, improve consistency, and increase platform reliability.
- Establish and mature reliability and observability practices, including SLIs, SLOs, Error Budgets, metrics, logs, traces, alerting, and dashboards.
- Lead major incident response, Problem Management, root cause analysis, post-incident reviews, and corrective actions that address systemic issues and prevent recurrence.
- Perform operational readiness and resiliency reviews while identifying reliability risks, operational gaps, technical debt, and opportunities for continuous improvement.
- Troubleshoot complex issues across applications, Kubernetes environments, infrastructure, networking, identity services, databases, certificates, cloud services, and third-party integrations.
- Mentor engineers and partner with development, infrastructure, security, and business teams to improve technical standards, operational practices, documentation, release quality, and production readiness.
- Participate in a scheduled primary and secondary on-call rotation after completing training and demonstrating readiness to independently support the environment.
What We Expect of You:
- Bachelor's degree in Computer Science, Software Engineering, Information Technology, or a related field and 7+ years of experience in Software Engineering, Site Reliability Engineering, DevOps, Platform Engineering, or a related discipline; or 10+ years of equivalent professional experience.
- 5+ years of experience administering Linux-based systems in enterprise environments.
- 5+ years of experience developing automation solutions using Ansible.
- 3+ years of experience developing automation and operational tooling using Python.
- 3+ years of experience supporting Kubernetes or other container orchestration platforms in production environments.
- Experience supporting business-critical production applications and distributed systems.
- Experience leading major incident response, Problem Management, root cause analysis, and corrective action initiatives.
- Experience working within ITIL-aligned Incident, Problem, and Change Management processes.
- Experience implementing observability solutions using metrics, logs, traces, alerting, and dashboards.
- Experience defining and measuring service reliability using SLIs, SLOs, and Error Budgets.
- Strong knowledge of networking, CI/CD pipelines, source control, certificates, secrets management, and modern software delivery practices.
- Demonstrated ability to troubleshoot complex technical issues, influence technical direction, mentor engineers, and communicate effectively with technical and non-technical stakeholders.
- Experience reviewing, troubleshooting, and making minor enhancements to existing Java-based applications. This is not primarily a Java application-development role, is a plus.
- Experience with Spring Boot, REST APIs, messaging technologies, and microservices architectures, is a plus.
- Experience with OpenTelemetry, Dynatrace, Prometheus, Grafana, or similar observability platforms, is a plus.
- Experience supporting Azure, AWS, or GCP cloud services and cloud-native architectures, is a plus.
- Experience troubleshooting and supporting PostgresSQL, MySQL, MongoDB, DB2, IBMi (AS/400), or other enterprise database platforms, is a plus.
- Experience implementing OAuth 2.0, JWT, RBAC, and secure application practices, is a plus.
- Familiarity with Agile, Scrum, or SAFe delivery methodologies, is a plus.
- Experience supporting enterprise-scale Managed Services environments, is a plus.
- Kubernetes, cloud, security, or automation-related certifications, is a plus.
- Pay range: $106,000 - $150,180 depending on experience and skill set
- Annual bonus target of 10% subject to terms and conditions of plan
- Benefits overview: ranges may be subject to geographic differentials
- ...and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Site Reliability Engineer Overview-The ProCOM team is looking for a Site Reliability Engineering (SRE) who can help us solve problems, build...SeniorFull timePart timeImmediate startWorldwideFlexible hours
$96k - $163k
...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Site Reliability Engineer, Performance Engineering Senior Site Reliability Engineer, Performance Engineering Payment Optimization unifies...SeniorFull timePart timeWorldwideFlexible hours$96k - $163k
...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Site Reliability Engineer Who is Mastercard? At Mastercard technology, we work to connect and power an inclusive, digital economy that...SeniorFull timePart timeWorldwideFlexible hours- ...Evaluate applications, platforms, and vendors to assess resiliency, reliability, and operational risk.Design and implement processes that... ...and reliability tooling.Actively participate in reliability engineering and resilience communities of practice, contributing to...SeniorFull time
- About the RoleWe’re looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability, observability, and operational excellence at the core. You’ll partner with engineers and data scientists to build, automate, and maintain the infrastructure...Senior
$170k - $220k
Who We're Looking ForWe’re looking for a hands-on, high-agency Site Reliability Engineer to help shape and scale the reliability layer of our stack. You'll own the release pipeline end-to-end — managing daily releases, weekly deploys, and hotfixes — while also automating...Senior- ...TechMContact: Meghana GorusuCompany: SRI Tech SolutionsJob Title: Senior Site Reliability EngineerLocation: Plano , TX (remote)Years of Experience: 8... ...are seeking a highly skilled Senior Site Reliability Engineer (SRE) to join our dynamic team. The ideal candidate will...SeniorRemote work
- Inspire Brands is hiring two Senior Site Reliability Engineers to help build and scale reliable, resilient, and observable systems supporting high-traffic, customer-facing digital platforms. These role blends software engineering, systems thinking, and operational excellence...SeniorWorldwide
$104.9k - $174.7k
About the role:A FinOps Site Reliability Engineer (SRE) bridges the gap between engineering, operations, and financial governance by embedding cost optimization into infrastructure design, automation, monitoring, and operational processes. A FinOps SRE proactively identifies...SeniorFull timeLocal area- Senior Site Reliability Engineer ILocationSan Jose, Costa Rica - RemoteSummary of roleOwn availability, the most important product feature, by continually striving for sustained operational excellence of Sumo’s planet-scale observability and security products. Work with...SeniorFlexible hours
- IXL Learning, developer of personalized learning products used by millions of people globally, is seeking a Senior Site Reliability Engineer to join our team, and help maintain the reliability and optimal performance of our products. We are seeking engineers with a passion...SeniorWork at officeImmediate start
- ...and foster a dynamic work environment where new ideas thrive. Are you ready to join our team and make an impact?As a Senior Site Reliability Engineer at TeamViewer, you’ll be a key player in ensuring the reliability, scalability, and performance of our Azure-based SaaS...SeniorTemporary workCasual workWorldwide
$210k - $230k
GovCIO is currently hiring for a Senior Site Reliability Engineer (SRE) to design, implement, and maintain highly available, scalable, and resilient infrastructure systems. The ideal candidate will bridge the gap between development and operations, focusing on automation...SeniorCurrently hiringRemote work- ...work from home day is currently Tuesday.Engineering at Lambda is responsible for building and... ...and networking teams to improve service reliability and deployment workflowsDeploy and... ...rotationYouHave 5+ years of experience in Site Reliability Engineering, Production Engineering...SeniorWork at officeLocal areaWork from homeFlexible hours
$168k - $270.25k
...phenomenal people like you to help us accelerate the next wave of artificial intelligence.Join our team at NVIDIA as a Senior Site reliability engineer focused on HPC storage and play a crucial role in designing, implementing, and optimizing on-prem High-Performance...SeniorFull time- ...professionalism. We are seeking an experienced AWS solution design engineer/architect to join our infrastructure cloud team. The... ...product features efficiently and confidently them into production.As Senior SRE, you will be responsible for providing leadership, design and...Senior
$174k - $252k
...systems by pushing for changes that improve reliability and velocity.Practice sustainable... ...:Bachelor’s degree in Computer Science, Engineering, a related field, or equivalent practical... ...degree in Computer Science or Engineering.Site Reliability Engineering (SRE) is what you...Senior$101k - $161k
...several prestigious awards, such as Best Engineering Team, Best Company for Diversity,... ...DescriptionWho You'll Work WithWe’re looking for Site Reliability Engineers to join our growing Arista’s... ...: EngineeringExperience level: Mid-Senior LevelIndustry: Computer NetworkingSenior- The Senior Site Reliability Engineer is responsible for improving the reliability, availability, scalability, and operational excellence of our critical infrastructure platforms and services. This role partners closely with Engineering, Security, and Infrastructure teams...SeniorFull timeWork at officeLocal area
$152.5k - $205k
...flexible work environment where new ideas are encouraged and everyone is a stakeholder.What you’ll be responsible for:As a Senior Site Reliability Engineer on Circle’s platform team, you’ll design, build, and operate the secure, scalable platform infrastructure behind...SeniorFlexible hours- We are looking for a Senior or Staff level Site Reliability Engineer to strengthen the reliability, scalability, and operational maturity of our platform in San Francisco, California. This role will focus on improving service health, refining observability, and partnering...Senior
- LeanData helps the world’s fastest-growing companies automate, simplify, and accelerate revenue.We are looking for a Senior Site Reliability Engineer to lead the strategic evolution of our cloud infrastructure. Reporting directly to the SVP of Engineering, this role is...SeniorFull timeWork at office2 days per week
$160k - $240k
...millions of times a day - quickly, reliably, and securely. Any time you... ...at Fiserv.Job TitleSenior Site Reliability EngineerWhat does a successful Site Reliability Engineer do at Fiserv?You will join our... ...operations or DevOps at a mid-to-senior level.Strong shell scripting...SeniorFull time- ...Lambda’s designated work from home day is currently Tuesday.Engineering at Lambda is responsible for building and scaling our cloud offering... ...and SLIs for Kubernetes services, workloads, and platform reliability.You6+ years of experience in a SRE, operations engineer, or...SeniorWork at officeLocal areaWork from homeFlexible hours
$80k - $140k
Job DescriptionRBC Wealth Management Technology is seeking a Senior Site Reliability Engineer to join its Wealth Management SRE Team. This team is responsible for ensuring the performance, availability, resilience, and operational excellence of critical applications and...SeniorFull timeFlexible hoursShift work$158.5k - $172k
...the exceptional value they deserve.About The OpportunityAs a Senior Engineer on the Runtime Automation team, you will design, automate,... ...environment. This is a high-impact position driving continuous reliability, deep system optimization, and automation across our entire...SeniorFull timeTemporary workWork at officeFlexible hours3 days per week$104.9k - $174.7k
Are you passionate about improving reliability, scalability, and resilience in complex database... ....Own prioritization of reliability engineering tasks within team backlogs.Lead incident... ...a Service (IaaS).Background in DevOps, site reliability engineering practices, or related...SeniorFull timeLocal area$91.7k - $163.7k
...potential to change lives. Ready to build the next breakthrough? Join us to start Caring. Connecting. Growing together. The Site Reliability Engineer will architect, develop, and maintain Optum Serve's cloud environment in both the commercial and government cloud. The...SeniorMinimum wageFull timeWork experience placementWork at officeLocal areaRemote work$267k - $356k
...day is currently Tuesday.Lambda's Storage Engineering team is the backbone behind our world-... ...workloads in the industry, which means reliability and performance aren't just goals—they're... ...defined storage across new and existing sites using tools such as Ansible, Jenkins etc...SeniorWork experience placementWork at officeLocal areaWork from homeFlexible hours$119.8k - $234.7k
...yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole type: Individual... ...EngineeringDiscipline: Site Reliability EngineeringCompany: MicrosoftOverviewMicrosoft... ...’s most demanding workloads. As a Senior Site Reliability Engineer, you will lead reliability...SeniorOngoing contractLocal area3 days per week
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- site reliability engineer United States
- site reliability engineering manager United States
- site reliability engineer sre United States
- site reliability engineer remote United States
- lead site reliability engineer United States
- senior human resources associate United States
- senior network engineer remote United States
- senior education consultant United States
- senior benefits manager United States
- senior app developer United States

