Senior Site Reliability Engineer
Inspire Brands
Inspire Brands is hiring two Senior Site Reliability Engineers to help build and scale reliable, resilient, and observable systems supporting high-traffic, customer-facing digital platforms. These role blends software engineering, systems thinking, and operational excellence to reduce toil, prevent incidents, and improve system reliability at scale.The ideal candidate has hands-on experience applying and implementing SRE principles — not just supporting production systems, but engineering reliability into them.RESPONSIBILITIESReliability EngineeringDefine and manage SLIs, SLOs, and Error Budgets for critical servicesDrive production readiness reviews and reliability requirements into architecture and designPerform capacity planning, failure mode analysis, and dependency risk assessmentsIdentify systemic reliability risks and drive remediation before they cause customer impactObservabilityDesign monitoring, alerting, logging, and tracing solutions using modern observability toolingImprove signal-to-noise ratio and reduce alert fatigueBuild dashboards and telemetry that reflect true service health, not just infrastructure metricsIncident ManagementLead technical response for high-severity incidentsDrive blameless postmortems and root cause analysis focused on systemic fixesContinuously improve detection, response, and recovery processesParticipate in an on-call rotationAutomation & Toil ReductionIdentify and eliminate manual, repetitive operational work through automationBuild self-healing systems, tooling, and scripts to reduce human interventionImprove CI/CD pipelines and deployment safety (canary, rollback, blue-green)Support Infrastructure as Code (Terraform, Bicep, or similar)Performance & ScalabilityConduct load testing, performance benchmarking, and bottleneck analysisPartner with engineering to design systems for horizontal scalability and fault toleranceCollaboration & CulturePartner with engineering teams to implement resiliency patterns (circuit breakers, retries, graceful degradation, rate limiting)Mentor engineers on SRE best practicesPromote a culture of engineering-driven reliability over reactive operationsEDUCATION AND EXPERIENCE QUALIFICATIONSRequired Qualifications5+ years experience in Site Reliability Engineering, Software Engineering, or Platform Engineering2+ years experience with Kubernetes and containerized workloads4-year degree in Computer Science or related fieldPreferred QualificationsExperience with chaos engineering or resiliency testingExperience with high-volume, high-availability transactional systemsExperience with AI-assisted observability or operational automationExperience making meaningful contributions to internal SRE tooling, frameworks, or platformsREQUIRED KNOWLEDGE, SKILLS, OR ABILITIESStrong programming/scripting skills (Python, Go, Java, or Node.js)Demonstrated experience defining and operating against SLOs/Error BudgetsStrong skills in leading incident response and root cause analysis for production systemsSolid understanding of distributed systems and microservices architectureDeep knowledge and expertise in at least one major cloud platform (Azure, AWS, or GCP)Expertise with observability platforms and monitoring strategyThis position is based in our Atlanta Support Center, with an expected on-site presence of 80%.Inspire is a multi-brand restaurant company whose portfolio includes more than 33,300 Arby's, Baskin-Robbins, Buffalo Wild Wings, Dunkin', Jimmy John's, and SONIC restaurants worldwide. We're made up of some of the world's most iconic restaurant brands, but we're much more than just a restaurant company. We're a team of hundreds of thousands who individually and collectively are changing the way people eat, drink, and gather around the table. We know that food is much more than a staple—it's an experience. At Inspire, that's our purpose: to ignite and nourish flavorful experiences.
$104.9k - $174.7k
...Data Management. You can learn more about LexisNexis Risk at the link below, About the Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the reliability of our production systems. This is not a purely advisory...SeniorFull timeWork at officeLocal areaRemote workWork from home- ...Georgia, and serves customers in more than 35 countries worldwide.Position OverviewWe are seeking a highly experienced Senior Site Reliability Engineer (Unified Observability) to lead the design, implementation, and operational maturity of the F1 Next Generation...SeniorFull timeWorldwideFlexible hours
$120k - $175k
...Requirements: We need at least 5 years of experience as a reliability-focused engineer in a fast-moving, rapidly expanding enterprise setting.... ...to qualified remote applicants anywhere in the U.S. This Senior Site Reliability Engineer role offers a typical salary range...SeniorFull timeRemote workVisa sponsorshipFlexible hours- ...complex, distributed, cloud-native systems. As a Staff Platform Engineer, you will play a critical role in ensuring these systems... ...hands-on engineering and technical leadership role. You will own reliability for major platform domains, design scalable solutions on Kubernetes...Senior
- ...OneTrust is seeking a Senior Software Engineer in Atlanta, Georgia. The role involves designing and maintaining a reliable application platform, collaborating with engineering teams, and enhancing customer experiences through observability tools. The ideal candidate...Senior
- ...The Challenge We’re looking for a Senior Software Engineer that will report to the Development Manager / R&D Head. In this role you will... ...services. Manage and enforce error budgets to balance system reliability with product feature velocity. Improve alert quality by...SeniorLocal areaFlexible hours
- ...Onetrust is seeking a Senior Software Engineer in Atlanta, Georgia, to enhance application reliability and contribute to mission-critical R&D projects. The ideal candidate will have over 4 years of experience in application development, particularly with Java, and a solid...SeniorFlexible hours
- ...T-Mobile USA, Inc. is seeking an experienced Site Reliability Engineer to lead reliability across cloud platforms and enterprise applications. You will own monitoring, incident response, CI/CD, and automation while guiding AI-enabled development practices and cross-team...Senior
$100k - $120k
OverviewThe Site Reliability Engineer is a key force behind improving Origami’s time to resolution and advancing overall site reliability and scalability. This person participates in efforts to identify root causes during post-incident investigations, while also identifying...Full timeTemporary workWork experience placementFlexible hours$152.13k - $162.13k
...challenge the status-quo.Unum is changing, and we’re excited about what’s next. Join us.General Summary:Unum Group seeks Site Reliability Engineers in Atlanta, GA.Applicants who are interested in this position may apply at (Ref #66753) for consideration.Design, build,...Full timeTemporary workWork at officeRemote work$138.1k - $198.2k
...more intuitive with technology that simply works. The SRE Engineering Enablement Team supports our CI Platforms, Developer Environments... ...Our customers are all engineers at Cisco. Your Impact As a Site Reliability Engineer, you will be at the epicenter of our engineering...Permanent employmentFull timeTemporary workWork experience placementLocal areaRemote workFlexible hours- ...T-Mobile USA, Inc. seeks a Site Reliability Engineer lead to design, build, and operate secure, scalable software platforms using SRE and AI-native practices across cloud-native and distributed systems. You will mentor the SRE team, drive Terraform-based infrastructure...Senior
$35 - $45 per hour
DescriptionKforce has a client that is seeking a remote Site Reliability Engineer to join their team.Summary:The team consists of systems that can track lead management, job management and sales management. It is built on Salesforce but underpinned by a lot of Java/API'...Remote work$101.5k - $169.1k
...include an incentive program.Job DescriptionThe Release Train Engineer (RTE) has a primary purpose of supporting an Agile Release Train... ...organizational AI policies and standards. Monitor AI tool reliability across teams. Create backup plans for system failures. Maintain...SeniorFull timeWork at officeRemote workVisa sponsorshipFlexible hours$167.7k - $245.2k
...Cisco Meraki, we are responsible for building and growing the cloud that supports these customers and their networks. As a Site Reliability Engineer, you will be focused on supporting a specific, highly available, and very secure production environment. You will analyze...Permanent employmentFull timeTemporary workLocal areaFlexible hours- ...of America)Please review the following job description:The Site Reliability Engineering Lead role focuses on enhancing the reliability and... ...platforms across hybrid cloud and on-premises environments. This senior technical leader drives improvements in automation, observability...Permanent employmentFull timePart timeH1bWork at officeLocal areaImmediate startWork visaMonday to FridayShift workDay shift
- Job-ID31782898Reference26-00225 Job Title: ( Senior Software Configuration/Release Engineer ) About Kyyba: Founded in 1998 and headquartered in Farmington... ...combined with career development. Job Description On-Site Interviews Only HYBRID - IN THE OFFICE 2 DAYS PER WEEK...SeniorWork at officeVisa sponsorshipWork visa2 days per week
- OverviewJob PurposeIntercontinental Exchange (ICE) presents a unique opportunity for a full-time Engineer to join a team responsible for ICE OpenShift container platform. The Engineer will drive enablement and adoption to migrate high performant and critical applications...SeniorFull time
- ...speed capabilities our nation and its allies need to maintain a durable, asymmetric advantage.About The Role:The Mission Systems Engineering (MSE) Team develops the Mission Management System (MMS)—a software platform that integrates mission subsystems, autonomy services...SeniorWeekly payPermanent employmentFull timeWork at office
- ...and test of the company's first combined turbojet-ramjet engine and is now being scaled through its first flight vehicle... ...capabilities to the warfighter. Hermeus is seeking a Senior Software or Site Reliability Engineer to join the Information Team and take charge of...Full timeRemote work
- ...A technology company is seeking a skilled Site Reliability Engineer (SRE) with expertise in AEM to ensure application reliability, performance, and scalability. Responsibilities include implementing monitoring solutions, automating deployments, and optimizing cloud costs...
- ...Manager @ STAFFWORXS | US IT Recruitment Job Opening: AWS Site Reliability Engineer (SRE) We’re hiring a Site Reliability Engineer (SRE) to... ...interview is mandatory as part of the hiring process. Seniority level ~ Seniority level Mid-Senior level...Contract work
$207.4k - $298.1k
...enabled product roadmap.About the role:We are hiring a Senior Principal Software Engineer to lead UKG Ready SMB — Architecture & Design, with a heavy... ...-to-market for cross-team integrations, operational reliability and cost-efficiency as Ready scales to its revenue...SeniorWorldwide- Job Title: Senior Software EngineerWork Location:Atlanta, GAJob Summary:Seeking a Senior Java developer with 3 to 5 years of experience to design and build cloud native applications leveraging Java technologies.Job Description:Design develop and maintain high quality Java...Senior
- ...are we looking for? We’re looking for experienced Software Engineers who are passionate about building rock-solid payment products... ...You’ll know you’re the right candidate when you enjoy crafting reliable SaaS APIs and are also passionate about helping other developers...SeniorFull timeWork experience placementFlexible hours
$165k - $216.56k
...redefine the future of how work gets done.We are looking for a Senior Solution Engineer who is accustomed to solving customer’s most complex... ...States, please visit the job posting on the Snowflake Careers Site for salary and benefits information: careers.snowflake.comCompensation...Senior$143k - $191k
...autonomy, AI, computer vision, sensor fusion, and networking technology to the military in months, not years.ABOUT THE TEAMThe Reliability Engineering team partners across Anduril's engineering, manufacturing, and operations organizations to ensure our autonomous systems...SeniorFull timeWork experience placementImmediate start- ...native services, and infrastructure automation to improve reliability and speed to delivery. Our global team operates in a... ...improvement and operational excellence.About the Role:As a Senior Network Platform Engineer, you will be a key technical contributor responsible for...SeniorFull timeWork at officeFlexible hours
$228k - $342k
...built. We’re forming small, senior, cross-functional AI teams that... ...leaders, machine learning engineers, and full-stack builders to create... ...them into systems that are reliable, explainable, and built to... ...Careers. Please be aware of sites that may ask for you to input...SeniorFull timeWork at officeRemote workHome officeFlexible hours- ...Saviynt’s platform is mission-critical for our customers. As we scale globally, reliability, availability, and performance are not optional—they are core product features. As a Principal Engineer, you will define and drive the reliability strategy for our SaaS platform....
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- site reliability engineer Atlanta, GA
- site reliability engineer remote Atlanta, GA
- site reliability engineer sre Atlanta, GA
- senior maintenance supervisor Atlanta, GA
- senior lead project manager Atlanta, GA
- senior robotics software engineer Atlanta, GA
- senior devops engineer remote Atlanta, GA
- senior sas administrator Atlanta, GA
- senior IT manager Atlanta, GA
- senior director of client services Atlanta, GA



