Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff Site Reliability Engineer

$96k - $192k

Carrier

About Carrier

Carrier Global Corporation, global leader in intelligent climate and energy solutions, is committed to creating innovations that bring comfort, safety and sustainability to life. Through cutting-edge advancements in climate solutions such as temperature control, air quality and transportation, we improve lives, empower critical industries and ensure safe transport of food, lifesaving medicines and more. Since inventing modern air conditioning in 1902, we lead with purpose: enhancing the lives we live and the world we share. We continue to lead because of our world-class, inclusive workforce that puts the customer at the center of everything we do. For more information, visit corporate.carrier.com or follow on Carrier social media at @Carrier.

About the Role

Carrier's Building Automation Systems (BAS) Cloud organization is seeking a highly skilled Senior Site Reliability Engineer (SRE) to help build, scale, and continuously improve our cloud-native SaaS platform. This role combines software engineering, cloud infrastructure, platform engineering, observability, automation, and operational excellence to ensure world-class reliability, security, and performance for business-critical customer solutions.

As a Senior SRE, you will play a key role in defining reliability standards, building self-service platform capabilities, driving automation, and enabling engineering teams to operate highly available services at scale. You will partner closely with software engineering, architecture, cybersecurity, DevOps, cloud operations, and product teams to improve system resilience, accelerate delivery, and reduce operational toil.

This is an opportunity to influence the technical direction of a rapidly evolving cloud platform while helping establish engineering best practices across a global organization.

What You'll Do

Platform Engineering & Reliability
  • Design, implement, and maintain highly available, scalable, and secure cloud infrastructure supporting Carrier's SaaS platforms.
  • Define and drive Service Level Objectives (SLOs), Service Level Indicators (SLIs), and Service Level Agreements (SLAs) across critical services.
  • Build self-service platform capabilities that enable development teams to deploy and operate services efficiently and consistently.
  • Develop reliability frameworks, standards, and operational best practices that improve platform resiliency and customer experience.
  • Identify reliability bottlenecks and architect solutions to eliminate single points of failure.
Automation & Infrastructure as Code
  • Develop automation solutions that eliminate repetitive operational tasks and reduce manual intervention.
  • Design and maintain Infrastructure as Code (IaC) using Terraform, AWS CloudFormation, AWS CDK, or similar technologies.
  • Build and enhance CI/CD pipelines to improve deployment velocity, consistency, and security.
  • Implement automated remediation, self-healing capabilities, and operational workflows.
Observability & Incident Response
  • Design and enhance modern observability solutions using metrics, logs, traces, and distributed monitoring technologies.
  • Develop dashboards, monitoring standards, alerting frameworks, and reliability reporting mechanisms.
  • Lead incident response efforts for critical production events, including root cause analysis and long-term corrective actions.
  • Drive post-incident reviews focused on systemic improvements rather than individual fault.
Cloud Operations & Resilience
  • Improve platform performance, availability, scalability, disaster recovery, and business continuity capabilities.
  • Validate backup, restoration, and recovery processes through regular testing and continuous improvement efforts.
  • Collaborate with engineering teams to implement resilient architectures capable of meeting defined reliability and compliance objectives.
  • Support operational readiness reviews for new services and platform capabilities.
Engineering Excellence & Leadership
  • Mentor engineers and provide technical leadership in reliability engineering practices.
  • Partner with development teams to improve application reliability throughout the software development lifecycle.
  • Establish and promote a culture of automation, ownership, continuous learning, and operational excellence.
  • Stay current on emerging technologies and industry best practices within cloud infrastructure, platform engineering, AI-assisted operations, and Site Reliability Engineering.
On-Call & Operational Excellence
  • Participate in a shared on-call rotation supporting production environments.
  • Drive initiatives focused on reducing operational toil through automation, self-healing systems, and platform improvements.
  • Continuously improve incident response processes and platform reliability practices.
Required Qualifications
  • Bachelor's degree in Computer Science, Software Engineering, Information Technology, or a technical field with 7+ years of experience in Site Reliability Engineering, Platform Engineering, DevOps, or Cloud Engineering -OR- Master's degree in Computer Science, Software Engineering, Information Technology, or a technical field with 5+ years of experience in Site Reliability Engineering, Platform Engineering, DevOps, or Cloud Engineering.
  • 3+ years of hands-on experience designing and supporting large-scale cloud environments.
  • 3+ years of experience with Infrastructure as Code (such as Terraform, CloudFormation, AWS CDK).
  • 3+ years of experience building and maintaining CI/CD pipelines and deployment automation.
  • 3+ years of extensive experience with Amazon Web Services (AWS), including services such as EKS, EC2, VPC, RDS, Lambda, CloudWatch, and IAM.
  • 3+ years of experience implementing observability platforms utilizing technologies such as Prometheus, Grafana, Open Telemetry, Datadog, Splunk, or New Relic.
Preferred Qualifications
  • Strong understanding of modern SRE principles, including reliability engineering, observability, toil reduction, incident management, and error budgets.
  • Strong scripting or programming experience using Python, Go, PowerShell, Bash, or similar languages.
  • Strong knowledge of networking, security, systems architecture, and distributed systems concepts.
  • Proven experience leading production incident response and root cause analysis activities.
  • Proven experience with cloud-native platforms and container technologies, including Kubernetes and Docker.
  • Experience operating SaaS platforms serving large-scale customer environments.
  • Experience supporting compliance frameworks such as SOC 2, ISO 27001, NIST, or similar standards.
  • Experience implementing AI-assisted engineering solutions to improve operational efficiency, troubleshooting, automation, and service reliability.
  • Knowledge of platform engineering concepts including Internal Developer Platforms (IDP), developer self-service capabilities, and engineering enablement practices.
  • AWS certifications or other cloud certifications.
  • Excellent communication, leadership, and collaboration skills.
Pay Range
The annual salary for this position is between $96,000.00 - $192,000.00 annually. Factors which may affect pay within this range include, but are not limited to, skills, education, experience, and other unique qualifications of the successful candidate.

Other Compensation
This position is entitled to short-term cash incentives, subject to plan requirements.

Benefits

Employees are eligible for benefits, including:
  • Health Care Benefits : Medical, Dental, Vision; Wellness incentives
  • Retirement Benefits
  • Time off and Leave : Paid vacation days, up to 15 days; paid sick days, up to 5 days; paid personal leave, up to 5 days; paid holidays, up to 13 days; birth and adoption leave; parental leave; family and medical leave; bereavement leave; jury duty leave; military leave; purchased vacation
  • Disability : Short-term and long-term disability
  • Life Insurance and Accidental Death and Dismemberment
  • Tax-Advantaged Accounts: Health Savings Account; Health Care Spending Account; Dependent Care Spending Account
  • Tuition Assistance


To learn more about our benefits offering, please click here Work with us | Carrier Corporate The specific benefits available to any employee may vary depending on state and local laws and eligibility factors, such as date of hire and the applicability of collective bargaining agreements.

Carrier EEO Statement and Accommodations Process

Carrier is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, age, disability or veteran status or any other applicable state or federal protected class. Carrier provides affirmative action in employment for qualified individuals with a Disability and Protected Veterans in compliance with section 503 of Rehabilitation Act and the Vietnam Era Veterans' Readjustment Assistance Act.


If you require a reasonable accommodation to complete the application process, participate in an interview, or otherwise engage in the hiring process, please contact us at View email address on click.appcast.io. We will make every effort to meet your needs in accordance with applicable laws.

Application Deadline
Applications will be accepted for at least 3 days from Job Posting Date: 26 August 2026

Job Applicant's Privacy Notice

Please click on the link to review the Job Applicant Privacy Notice.

Use of AI

Technology-enabled tools may support parts of the recruitment process, with oversight by people.
Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Staff Site Reliability Engineer in United States vacancy
  •  ...Infrastructure team builds the platforms and tooling that help engineering teams develop, deploy, and operate production systems...  ...make safe shipping the default for every product team.As a Staff Site Reliability Engineer on Release Engineering, you'll define and scale... 
    Suggested
    Permanent employment
    Work experience placement
    Work at office
    Local area

    Plaid Financial

    San Francisco, CA
    2 days ago
  • $117k - $209.33k

    Job Requisition ID #26WD99276Position OverviewWant to help make a better world? As a Senior Site Reliability Engineer at Autodesk, you can help us build and operate reliable, secure, and scalable cloud services for Autodesk GovCloud products.As part of a new SRE team supporting... 
    Suggested
    Full time
    For contractors
    Remote work

    Autodesk

    Plano, TX
    4 days ago
  •  ...Fluency: English (Required)Work Shift:1st shift (United States of America)Please review the following job description:The Site Reliability Engineering Lead role focuses on enhancing the reliability and operational excellence of enterprise platforms across hybrid cloud... 
    Suggested
    Permanent employment
    Full time
    Part time
    H1b
    Work at office
    Local area
    Immediate start
    Work visa
    Monday to Friday
    Shift work
    Day shift

    Truist

    Atlanta, GA
    3 days ago
  • $98.58k - $138.02k

     ...Northern California / Silicon Valley Region / Denver, COProduct Engineering - DevOps /Full Time /HybridRestaurant365 is a SaaS company...  ...office locations: Austin, TX; Irvine, CA; or Akron, OH. The Site Reliability Engineer II will be responsible for supporting, enhancing,... 
    Suggested
    Full time
    Work at office

    Restaurant 365

    Irvine, CA
    2 days ago
  •  ...communities.This is a Lead Software Production Management & Reliability Engineering position at Director level which is part of the job family responsible...  ...across the business.Job SummaryWe are looking for a Site Reliability Engineer with a minimum of 5 years of industry... 
    Suggested
    Flexible hours
    Weekend work

    Morgan Stanley

    Alpharetta, GA
    2 days ago
  • $104.9k - $174.7k

     ...you passionate about improving reliability, scalability, and resilience...  ...informal guidance to junior staff.ResponsibilitiesSupport Business...  ...of reliability engineering tasks within team backlogs.Lead...  ...(IaaS).Background in DevOps, site reliability engineering practices... 
    Full time
    Local area

    RELX Group

    Georgia
    8 hours ago
  •  ...Evaluate applications, platforms, and vendors to assess resiliency, reliability, and operational risk.Design and implement processes that...  ...and reliability tooling.Actively participate in reliability engineering and resilience communities of practice, contributing to... 
    Full time

    Vanguard

    Dallas, TX
    1 day ago
  • $145k - $195k

     ...SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SITE RELIABILITY ENGINEER (TOP SECRET CLEARANCE)As a member of the Classified IT Systems Engineering team, the Site Reliability Engineer is involved... 
    Permanent employment
    Temporary work
    Weekend work

    SpaceX

    Hawthorne, CA
    3 days ago
  • $148k - $235.75k

     ...see how you can make a lasting impact on the world.Join our team of innovative engineers who are building an AI Data Center AIOps platform that turns raw, high-volume telemetry into reliable, job-centric insights and automation for GPU fleets. We’re hiring a DevOps Engineer... 
    Full time

    Nvidia

    Santa Clara, CA
    8 hours ago
  • $125k - $185k

    Washington, D.C.Engineering /Full-time /HybridA World-Changing CompanyPalantir builds the world’s leading software for data-driven decisions...  ...locate missing children, and more.The RoleWe’re looking for Site Reliability Engineers who can help us build, operate, and maintain high-... 
    Full time
    Work experience placement
    Work at office
    Remote work
    Work from home
    Relocation package

    Palantir Technologies

    Washington DC
    8 hours ago
  • $209.7k - $266.8k

     ...inclusive work environment; we back each other to deliver impact. Make Wayve the experience that defines your career! Site Reliability Engineer - Vehicle Software The role As a Site Reliability Engineer at Wayve, you will work across the full reliability stack... 
    Full time
    Work at office
    Work from home
    3 days per week

    Wayve

    Sunnyvale, CA
    2 days ago
  • Role Profile:We are evolving our Site Reliability Engineering capabilities to strengthen reliability, observability, security, and operational excellence across our Markets and Risk Intelligence division.As a Senior SRE, you will be a senior hands‑on technical person help... 
    Full time
    Shift work

    London Stock Exchange Group

    Raleigh, NC
    4 days ago
  • $166k - $258k

     ...Seattle office a minimum of 4 days/week in order to be considered for this position.Nordstrom is looking for a Senior Engineer 2 to join our Site Reliability Engineering (SRE) team — and we think that person could be you.You'll help build the scalable, reliable, and... 
    Full time
    Work at office

    Nordstrom

    Seattle, WA
    2 days ago
  •  ...English (Required)Work Shift:1st Shift (United States of America)Please review the following job description:Lead Site Reliability & Environment Monitoring Engineer (Azure / Dynatrace / ServiceNow)We are seeking a Lead Site Reliability & Environment Monitoring Engineer to... 
    Full time
    Temporary work
    Shift work
    Day shift

    TIH

    Charlotte, NC
    3 days ago
  • $174k - $252k

     ...systems by pushing for changes that improve reliability and velocity.Practice sustainable...  ...:Bachelor’s degree in Computer Science, Engineering, a related field, or equivalent practical...  ...degree in Computer Science or Engineering.Site Reliability Engineering (SRE) is what you... 

    Google

    Cambridge, MA
    3 days ago
  • Job ID: 18719802Reference Number: 23-00164Title: site reliability engineerLocation: Iselin, NJ, 08830Posted Date: 2023-01-20Contact: Shyam MaramContact Email: ****@*****.*** Phone: (***) ***-****Company: HAN Staffing Devops/SRE/Python Role Malvern PA - hybrid... 

    HAN Staffing

    Iselin, NJ
    4 days ago
  • $128k - $216k

     ...millions of times a day - quickly, reliably, and securely. Any time you...  ...at Fiserv.Job TitleSr. Site Reliability EngineerAbout CloverClover...  ...processing to inventory and staff management. With over 15...  ...successful Senior Site Reliability Engineer do at Fiserv?As a Senior Site... 
    Full time
    Worldwide

    Fiserv

    Sunnyvale, CA
    3 days ago
  • $147k - $227k

     ...to do career-defining work. We're all in on this mission. If you are too, let's talk. The Engineering Opportunity We are looking for an experienced Senior Site Reliability Engineer to join Okta's Emerging Products Group (EPG). Our mission is to build highly reliable... 
    Full time
    Local area
    Worldwide
    Flexible hours

    Okta

    San Francisco, CA
    5 days ago
  • $138.4k - $173k

     ...infrastructure as well as help improve the reliability, quality of services and overall...  ...recovery. You’ll collaborate or embed with engineering teams, helping them to improve the reliability...  ...about our locations by visiting our site.Compensation & BenefitsThe base salary that... 
    Full time
    Flexible hours

    AppFolio

    San Diego, CA
    3 days ago
  •  ...candidates that are particularly strong in a few areas, and have some interest and capabilities in others.About the Role:As a Site Reliability Engineer, you’ll join the global Platform SRE team responsible for building, operating, and scaling Kong’s multi-region SaaS... 
    Temporary work

    Kong

    Washington DC
    3 days ago
  • The Senior Site Reliability Engineer is responsible for improving the reliability, availability, scalability, and operational excellence of our critical infrastructure platforms and services. This role partners closely with Engineering, Security, and Infrastructure teams... 
    Full time
    Work at office
    Local area

    Castleton Commodities International

    Houston, TX
    3 days ago
  • $112k - $179k

     ...delivery of system, network, software, and security solutions.About The RolePeraton is seeking a self-driven and resourceful Site Reliability Engineer to join our dynamic of Network and UC engineers in Washington, DC. This position combines software engineering and systems... 
    Contract work
    Worldwide
    Shift work

    Peraton Corporation

    Washington DC
    3 days ago
  • LeanData helps the world’s fastest-growing companies automate, simplify, and accelerate revenue.We are looking for a Senior Site Reliability Engineer to lead the strategic evolution of our cloud infrastructure. Reporting directly to the SVP of Engineering, this role is... 
    Full time
    Work at office
    2 days per week

    LeanData

    Santa Clara, CA
    3 days ago
  •  ...and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Commercial and Investment Bank, Healthcare Payments team, you will solve complex and broad... 

    JP Morgan Chase

    Irvine, CA
    1 day ago
  •  ...Georgia, and serves customers in more than 35 countries worldwide.Position OverviewWe are seeking a highly experienced Senior Site Reliability Engineer (Unified Observability) to lead the design, implementation, and operational maturity of the F1 Next Generation Customer... 
    Full time
    Worldwide
    Flexible hours

    NCR

    Atlanta, GA
    2 days ago
  • $102.1k - $202.2k

     ...per yearEmployment type: Full-TimeWork site: Fully on-siteRole type: Individual ContributorTravel...  ...: Software EngineeringDiscipline: Site Reliability EngineeringCompany:...  ...at the intersection of large-scale cloud engineering, service reliability, and operational excellence... 
    Ongoing contract
    Local area
    Worldwide

    Microsoft

    Reston, VA
    2 days ago
  • $152.5k - $205k

     ...flexible work environment where new ideas are encouraged and everyone is a stakeholder.What you’ll be responsible forThe Site Reliability Engineer builds and maintains shared platform capabilities, common libraries, and infrastructure that help Circle teams ship secure... 
    Flexible hours

    Circle

    San Francisco, CA
    3 days ago
  •  ...we are dedicated to connecting talented professionals with your ideal opportunities. We are currently seeking a qualified Site Reliability Engineer (AI & Agentic Systems) to join our client’s organization and contribute to their ongoing success. Job summaryThis role demands... 

    OpenArc

    Plano, TX
    8 hours ago
  • $69.8k - $148.3k

     ...of infrastructure and service to ensure reliability and functionality. Responds to infrastructure...  ...tools and develops working knowledge of site reliability trends.Only Oracle brings...  ...Science, Information Technology, Engineering, or a related field, or equivalent practical... 
    Temporary work
    Flexible hours

    Oracle Corporation

    Nashville, TN
    8 hours ago
  • $127k - $249k

    The TeamPlatform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions...  ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper).... 
    Work at office
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    Denver, CO
    8 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff Site Reliability Engineer. Be the first to apply!