Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Site Reliability Engineer

$149.8k - $241.5k

Camunda

Camunda is the enterprise platform for agentic orchestration, enabling organizations to coordinate AI agents, people, and systems across complex, end-to-end business processes. With built-in governance, auditability, and human oversight, Camunda gives enterprises the control they need to move AI from pilots to production — safely and at scale. Trusted by over 700 organizations worldwide, including 9 of top 10 US banks, Camunda helps enterprises boost operational efficiency, accelerate time-to-value, and deliver better customer experiences.

About The Role

We're looking for a Senior Site Reliability Engineer who's passionate about building reliable, scalable infrastructure that helps developers ship better software faster. You'll design and maintain our Kubernetes-based multi-cloud platform, improve our monitoring and observability tools, and collaborate with product and engineering teams to solve real problems in real time. This is a role where you'll own the systems that power Camunda, drive automation that raises the bar for everyone, and mentor others who want to do the same. If you thrive on building things that work well and don't break when they shouldn't, we'd love to talk.

What You’ll Be Doing
  • Design and maintain our infrastructure – You'll evolve our Kubernetes-based, multi-cloud platform architecture, ensuring it's available, scalable, and fault-tolerant. You'll establish configuration best practices and network services that our teams rely on.
  • Build observability that matters – Implement and improve monitoring and alerting tools that give both SREs and developers real visibility into system health and performance. Make it easy for teams to understand what's happening across our stack.
  • Own your systems end-to-end – You'll adopt a "you build it, you run it" mentality, which means participating in on‑call rotations and being the go‑to person when things need quick fixes. You'll create runbooks and automation that turn complex problems into manageable processes.
  • Ship features and improvements with product teams – Work cross‑functionally with product engineering, product management, and support to define and deliver features that move the needle. Bring your expertise to the table early and often.
  • Push automation to the next level – Identify repetitive work and automate it away. Share what you learn with your teammates so everyone gets better at their craft. We measure progress by the quality of our systems, not hours worked.
  • Be the expert others learn from – Help less experienced engineers tackle complex infrastructure challenges. Break down technical problems into clear steps and support others in growing their skills.
What You Bring
Must Haves:
  • Deep hands‑on experience with Kubernetes – You've built, deployed, and maintained Kubernetes clusters in production environments. You understand how to manage workloads, networking, and storage at scale.
  • Infrastructure as code expertise – You're fluent in tools like Terraform (or similar IaC tools) and know how to version, test, and safely deploy infrastructure changes.
  • Demonstrated experience in monitoring and observability – You've worked with tools like Prometheus, Grafana, or similar platforms to instrument systems and alert on what matters.
  • Strong 3rd-level support and incident response skills – You've responded to production incidents, diagnosed complex issues, and communicated clearly with stakeholders under pressure. You understand root cause analysis and how to prevent issues from happening again.
  • A passion for automation and raising the quality bar – You see manual work as a problem to be solved. You care deeply about building systems that are reliable, maintainable, and easy to understand.
  • Responsible use of AI tools – You leverage AI for research, code review, documentation, test generation, and automation to improve your effectiveness. You know how to validate AI outputs against requirements, never share confidential or personal data with AI systems without authorization, and always retain human accountability for decisions and deliverables.
Nice‑to‑haves
  • Experience with major cloud providers – You've worked with AWS (EKS), Google Cloud Platform (GKE), or similar managed Kubernetes services.
  • ArgoCD or GitOps workflows – You've used declarative, Git‑driven infrastructure management to keep systems in sync.
  • Proficiency in Python, Go, or similar languages – You write scripts and automation tools that solve real problems.
  • Experience with SLOs and alerting frameworks – You've helped teams define meaningful service level objectives and set up alerts that don't cry wolf.
Compensation
What We Have to Offer:

We offer competitive, fair, and transparent compensation. Salary ranges are location-based, with Standard and Major markets (global tech hubs) reflecting local competition.

  • United States: $149,800.00 to $241,500.00
  • United Kingdom: £94,100.00 to £154,700.00
  • Singapore: S$186,100.00 to S$279,100.00
  • Canada: C$159,100 to C$261,600

If you’re based elsewhere, you’ll be hired via Remote.com (our global employer partner), and your Talent Acquisition Partner will provide a personalized Total Rewards Calculator after your first interview.

Equity:

We also offer equity (where applicable) through our Virtual Stock Option Plan (VSOP).

Benefits & Perks

We invest in your wellbeing, growth, and ability to connect, along with perks that support you no matter where you’re based. Our benefits are globally designed and locally delivered where applicable.

  • Remote & Flexible: Work from anywhere with the setup that suits you, home office budget, co‑working space support, and flexible time off to recharge when you need it.
  • In Person Connection: We invest in meaningful face time through our Annual Kickoff (Vienna in 2025, Madrid in 2026!), team offsites, and Camundi Connection Budgets, including contributing to meetups while travelling,, and local gatherings with fellow Camundi.
  • Health & Wellbeing: Access locally tailored healthcare, Modern Health for global mental wellbeing, and our Live Well Lifestyle Spending Account (LSA), a flexible, global benefit that puts you in control of your whole life, not just work, from: staying active, to caring for family, exploring personal passions, meaningful experiences, and investing in your financial wellbeing. The Live Well program launches in 2026 and scales to €1,000 annually from 2027.
  • Financial Security: Retirement and pension plans (often with company contributions), plus life and disability insurance where relevant.
  • Professional Growth: Up to $/€/£1,000 per year for self-driven learning: courses, certifications, books, you decide!
AI in our hiring process

Camunda may use AI tools to aid the screening of applications and during the interview process. You can learn more here

"Everyone is welcome at Camunda" — it’s a celebrated component of our culture. We strive to create an inclusive environment that empowers our people. At Camunda, we honour diverse cultures and backgrounds and are proud to be an equal opportunity employer. All qualified applicants will receive consideration without regard to gender, race, ethnicity, religion, belief, sexual orientation, age, disability or any other protected characteristics under applicable law.

We are looking forward to your application!

#J-18808-Ljbffr
Vacancy posted 7 hours ago
Similar jobs that could be interesting for youBased on the Senior Site Reliability Engineer in Atlanta, GA vacancy
  • As the Senior Site Reliability Engineer, you will serve as a trusted technical resource responsible for deploying, validating, and operationalizing AI, HPC, Kubernetes, and enterprise infrastructure environments. This role transforms newly installed hardware into production... 
    Senior
    Work at office
    Immediate start
    Worldwide
    Shift work

    Wesco International

    Atlanta, GA
    2 hours ago
  •  ...Site Reliability Engineer We're looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability, observability, and operational excellence at the core. You'll partner with engineers and data scientists to build, automate, and... 
    Senior

    Alembic Technologies

    Atlanta, GA
    2 days ago
  •  ...Lead Engineer, Site Reliability Engineering Team As a lead engineer with Retail, Site Reliability Engineering team, you will be at the forefront of Cloud and Big Data technology. In this role you will establish yourself as a technical leader by exposing yourself to... 
    Senior

    Next Level Business Services, Inc.

    Atlanta, GA
    5 days ago
  • $120k - $175k

     ...level of sports fandom. Ready to reimagine the DFS industry together? We are seeking a highly skilled and experienced Senior Site Reliability Engineer to join our team. We are passionate about delivering cutting-edge solutions and pushing the boundaries of what's... 
    Senior
    Full time
    Remote work
    Work visa
    Flexible hours

    AEG Presents

    Atlanta, GA
    2 days ago
  •  ...Senior Site Reliability Engineer Inspire Brands is hiring two Senior Site Reliability Engineers to help build and scale reliable, resilient, and observable systems supporting high-traffic, customer-facing digital platforms. These role blends software engineering, systems... 
    Senior
    Worldwide

    Inspire Brands Inc

    Atlanta, GA
    4 days ago
  • $168k - $200k

     ...that is passionate about creating transformative change in healthcare. What We're Looking For We're looking for a Senior Site Reliability Engineer to join our Data & ML Platform team. You'll be at the forefront of building and operating a resilient, observable,... 
    Senior
    Remote work

    Datavant

    Atlanta, GA
    2 days ago
  •  ...professionalism. We are seeking an experienced AWS solution design engineer/architect to join our infrastructure cloud team. The...  ...features efficiently and confidently them into production. As Senior SRE, you will be responsible for providing leadership, design and... 
    Senior

    Intercontinental Exchange

    Atlanta, GA
    5 days ago
  • $136.2k - $214.01k

     ...outcomes Visionary in future focused problem-solving Exceptional in execution and impact The Role As a Senior Site Reliability Engineer at Proofpoint you will develop a deep understanding of the various services and applications that come together to... 
    Senior
    Full time
    Flexible hours

    Proofpoint

    Atlanta, GA
    4 days ago
  • $81.1k - $187k

     ...Job Description We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The role focuses on improving service reliability, reducing operational risk, automating repetitive tasks, and driving faster detection... 
    Senior
    Temporary work
    Immediate start
    Flexible hours
    Shift work

    Oracle

    Atlanta, GA
    2 days ago
  • $178.13k - $205.4k

     ...Bachelor's degree or foreign degree equivalent in Computer Engineering, Computer Science, Engineering, or related field plus five (5)...  ...websites that are not Workday Careers. Please be aware of sites that may ask for you to input your data in connection with a job... 
    Senior
    Work at office
    Remote work
    Flexible hours

    Workday

    Atlanta, GA
    1 day ago
  • $123.4k - $222.53k

     ...! Ready grow your career as part of the Uncarrier journey at T-Mobile? Our team is searching for our next Sr. Site Reliability Engineer to strengthen the reliability and resilience of the systems powering T-Mobile's payment platforms, enabling faster, safer... 
    Senior
    Full time
    Temporary work
    Part time
    Work experience placement
    Local area
    Flexible hours

    T-Mobile

    Atlanta, GA
    1 day ago
  • $169.3k - $304.7k

     ...in building and maintaining fast, efficient, scalable, and reliable routing software and infrastructure that is responsible...  ...growth and stability of our global platform. As a Principal Site Reliability Engineer - Network, you will be responsible for: Architecting,... 
    Work experience placement
    Work at office

    Akamai

    Atlanta, GA
    3 days ago
  •  ...of America)Please review the following job description:The Site Reliability Engineering Lead role focuses on enhancing the reliability and...  ...platforms across hybrid cloud and on-premises environments. This senior technical leader drives improvements in automation, observability... 
    Permanent employment
    Full time
    Part time
    H1b
    Work at office
    Local area
    Immediate start
    Work visa
    Monday to Friday
    Shift work
    Day shift

    Truist

    Atlanta, GA
    3 days ago
  •  ...We are seeking an experienced Site Reliability Engineer (SRE) to support and maintain production systems hosted on AWS. The role focuses on production support, incident management, monitoring, observability, troubleshooting, and improving system reliability and availability... 
    Contract work

    2T Consulting

    Atlanta, GA
    27 days ago
  •  ...complex, distributed, cloud-native systems. As a Staff Platform Engineer, you will play a critical role in ensuring these systems...  ...hands-on engineering and technical leadership role. You will own reliability for major platform domains, design scalable solutions on Kubernetes... 
    Senior

    Saviynt

    Atlanta, GA
    23 days ago
  •  ...configure the monitoring and alerting metrics so the support engineers can proactively and timely validate, troubleshoot and...  ...availability critical application components. • 1+ Years in Site Reliability Engineering organization preferred • Overall 4-6years of experience... 
    Work experience placement

    3B Staffing LLC

    Atlanta, GA
    1 day ago
  •  ...availability. • Automation Experience with Build/deployment, Software Configuration/Continuous Integration/Continuous Delivery/Release Engineering related tasks in JavaEE/C++ Environments. • Experience in automating manual processes using Python, Ruby, Unix Shell (bash,... 
    Immediate start

    Navtech

    Atlanta, GA
    1 day ago
  • $60 - $68 per hour

     ...Site Reliability Engineer Immediate need for a talented Site Reliability Engineer. This is a 12+ months contract opportunity with long-term potential and is located in Atlanta, GA (Onsite). Please review the job description below and contact me ASAP if you are interested... 
    Contract work
    Local area
    Immediate start

    Pyramid Corporation

    Atlanta, GA
    1 day ago
  •  ...Technical Support Specialist In Site Reliability Engineering (Sre) Mandatory skills: Scripting and programming languages like Python, Java, Ruby. Cloud and infrastructure management – AWS, Google cloud and Azure is a plus- CI/CD Automation, Database Management. The... 

    Omni Inclusive

    Atlanta, GA
    1 day ago
  • $75.7k - $136.3k

     ...solve complex challenges? Do you have a passion for automation and building systems that scale? Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages applications and infrastructure that support Akamai Cloud's products and... 
    Work experience placement
    Work at office

    Akamai

    Atlanta, GA
    1 day ago
  •  ...enterprise initiatives such as public cloud, data science, AI, engineering innovation, and IoT. Our customers include the world's...  ...is founder-led, profitable, and growing. We are hiring a Site Reliability Engineer Our goal is to perfect enterprise infrastructure DevOps... 
    Work at office
    Local area
    Remote work
    Work from home
    Worldwide

    Canonical

    Atlanta, GA
    28 days ago
  •  ...data, and human expertise. We deliver faster, smarter, more reliable insights to insurance carriers and single-family rental...  ...at scale, you're in the right place.  The Role   As a Site Reliability Engineer, you'll be responsible for the availability, scalability,... 
    Full time
    Flexible hours

    Seek Now

    Atlanta, GA
    8 days ago
  •  ...can create the conditions for educators to teach, students to thrive, and districts to shape the future of education. Site Reliability Engineer (SRE) Overview:   We are looking for a Site Reliability Engineer (SRE) to join our Engineering team. This is a build-it... 
    Full time
    Live in
    Work at office

    Incident IQ

    Atlanta, GA
    28 days ago
  • $100k - $120k

     ...OverviewThe Site Reliability Engineer is a key force behind improving Origami’s time to resolution and advancing overall site reliability and scalability. This person participates in efforts to identify root causes during post-incident investigations, while also identifying... 
    Full time
    Temporary work
    Work experience placement
    Flexible hours

    Origami Risk

    Atlanta, GA
    2 days ago
  •  ...Job Purpose At Intercontinental Exchange (NYSE:ICE), we engineer technology, exchanges and clearing houses that connect companies...  ...-oriented people to join our team. We are seeking a Site Reliability Engineer to bring 3+ years of hands-on experience to our SRE... 

    Intercontinental Exchange Holdings, Inc.

    Atlanta, GA
    1 day ago
  • $130k - $145k

     ...Back Site Reliability Engineer Cloud/Infrastructure Atlanta , GA Sep 2, 2026 Site Reliability Engineer Atlanta, GA / Hybrid Blu Omega is seeking a Site Reliability Engineer to support a federal program focused on enterprise cloud modernization. This role operates... 
    Temporary work

    Blu Omega LLC

    Atlanta, GA
    1 day ago
  •  ...Join to apply for the Site Reliability Engineer role at Motion Recruitment Join to apply for the Site Reliability Engineer role at...  ...and security enhancements. Posted By: VMS Sourcing Seniority level ~ Seniority level Mid-Senior level Employment... 
    Contract work
    Worldwide

    Motion Recruitment

    Atlanta, GA
    2 days ago
  • $130k - $150k

     ....00/yr Overview: We are seeking a highly skilled Site Reliability Engineer (SRE) to join our team and help build and maintain scalable...  ...software development lifecycle and CI/CD principles Seniority level ~ Seniority level Mid-Senior level Employment... 
    Full time
    Remote work

    Prestige Staffing

    Atlanta, GA
    3 days ago
  •  ...We have an immediate need for a Senior Release Train Engineer for a contract assignment located in Carmel, Indiana . The Release Train Engineer (RTE) has a primary purpose of supporting an Agile Release Train (ART) by steering it to success and navigating the complexity... 
    Senior
    Contract work
    Work at office
    Immediate start

    Spartan Technologies

    Atlanta, GA
    2 days ago
  • $145k - $160k

     ...We are seeking a specialized Observability & Infrastructure Engineer to lead the monitoring, telemetry, and platform-as-code initiatives critical to our multi-region disaster recovery roadmap. You will architect and implement robust observability pipelines, ensure deep... 
    Temporary work
    Remote work
    Flexible hours

    EPAM Systems Inc

    Atlanta, GA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!