Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Site Reliability Engineer

QGenda

Senior Site Reliability Engineer

Atlanta, Georgia

Who We Are

QGenda is redefining healthcare workforce management everywhere care is delivered. We're on a mission to empower the healthcare industry to better onboarding, deploy, and manage their workforce. Over 4,500 healthcare organizations have trusted us to help them make strategic workforce decisions through our unified software platform. With more than 800 employees across the US, we are united in our vision and culture to make a difference for our customers, while enjoying the day-to-day.

At QGenda, we value our employees and their contributions toward the success of the business. We strive to create a dynamic work environment that fosters growth, innovation, and collaboration, where employees can be proud of the work they do and the impact it has on the healthcare industry.

QGenda is headquartered in Atlanta.

About Your Role

As a Senior Site Reliability Engineer, you will work with our Infrastructure and Product Development Teams to increase the scalability, reliability, and performance of our systems and services. You will build and extend existing automation for configuration and monitoring of our AWS hosted applications. You will have the opportunity to evaluate new AWS services and tools to determine if they could be utilized in our environments. You'll bring a focus to platform health and monitoring to allow us to deliver the best possible experience for our customers. This is an excellent opportunity to have a significant impact on the stability of our systems and contribute to the evolution of our technology stack.

NOTE: This role is hybrid with one required day in our Buckhead (Atlanta, Georgia) or our Uniontown, Ohio office depending on your current location.

How You'll Make an Impact

As a Senior Site Reliability Engineer, you will work with our Infrastructure and Product Development Teams to increase the scalability, reliability, and performance of our systems and services. You will build and extend existing automation for configuration and monitoring of our AWS hosted applications. You will have the opportunity to evaluate new AWS services and tools to determine if they could be utilized in our environments. You'll bring a focus to platform health and monitoring to allow us to deliver the best possible experience for our customers. This is an excellent opportunity to have a significant impact on the stability of our systems and contribute to the evolution of our technology stack.

Responsibilities:

System Reliability & Performance:

  • Design, implement, and manage scalable systems that ensure high availability, fault tolerance, and optimal performance.
  • Continuously monitor and enhance system health and performance through data analysis and metrics.

Automation & Tooling:

  • Develop and advocate for automation tools to eliminate repetitive manual processes and improve efficiency.
  • Build and enhance CI/CD pipelines to streamline software delivery and deployments.

Incident Management & Troubleshooting:

  • Participate in on-call rotation to respond to incidents, troubleshoot problems, and minimize downtime.
  • Conduct root cause analyses and implement permanent solutions to recurring issues.

Infrastructure Management:

  • Manage our cloud-based infrastructure environment in AWS.
  • Optimize costs and resources while maintaining robust and scalable systems.

Collaboration & Culture:

  • Serve as a technical advisor to engineering teams on infrastructure and operations best practices.
  • Actively contribute to fostering an SRE culture within the organization by promoting observability, retrospectives, and continuous improvement.
Who You Are
  • Curiosity-driven mindset with a desire to continuously learn and improve systems
  • Strong sense of ownership — you see problems through to resolution, not just escalation
  • Comfortable navigating ambiguity and making pragmatic tradeoffs under pressure
  • Availability for off-hours deployment and upgrades of production systems during release and maintenance windows
  • Strong problem-solving skills and ability to work effectively under pressure.
  • Excellent communication skills for cross-functional collaboration as well as documentation creation.
Experience You Bring
  • B.S. in Computer Science, Computer Information Systems, or Computer Engineering from a major U.S. university or equivalent industry experience
  • 7+ years of experience as a DevOps, SRE or Systems Engineer
  • Advanced proficiency with at least one scripting or programming language
  • Experience with Docker and container orchestration tools such as AWS ECS and EKS/Kubernetes
  • Hands-on experience building infrastructure and supporting applications in AWS using services such as Lambda, EC2, ECS, S3, SNS, SQS, RDS, Redshift, and Elasticache
  • Strong understanding of networking and DNS
  • Strong experience with Terraform for infrastructure provisioning and module development, along with configuration management and infrastructure as code (IaC) practices
  • Firm understanding and experience with Agile and Scrum SDLC processes
  • Using distributed version control system experience (Git preferred) to check-in code, branching, merging, pull request, code review, etc
  • Knowledge of CI/CD best practices and tools such as AWS CodeBuild, Jenkins and/or TeamCity
  • Experience using AI-assisted coding tools (e.g., Claude, GitHub Copilot) to accelerate IaC development, scripting, and operational workflows
  • Familiarity with AI/ML-driven approaches to observability, anomaly detection, log analysis, or incident triage
  • Experience designing and delivering secure, high performance and highly available cloud services
  • Experience with observability platforms (e.g., Datadog, CloudWatch, PagerDuty) for monitoring, alerting, and incident response
  • Awareness of cloud security best practices including IAM policies, network segmentation, and secrets management

#LI-Hybrid

Applicants for this position must be authorized to work for any employer in the United States (U.S.), including being located in the US. We are unable to sponsor, take over sponsorship of, or hire candidates with an employment visa at this time.

What's In It For You

We offer a comprehensive total rewards package to support our full-time employees and their family's day-to-day needs, well-being and major life events, which includes:

  • Fully company-paid options for medical (both in-person and virtual), dental and vision insurance
  • Generous paid time off (PTO) policy to enjoy periods of uninterrupted rest and relaxation for a healthy work/life balance
  • Paid parental leave for birth, adoption or permanent placement
  • 401(k) with company match
  • Options to work in a hybrid-working model or remotely from home, depending on the position
  • Annual Costco membership, cell phone stipend, commuter benefits, in-office perks and more
  • QGenda delivers technology solutions to improve how healthcare is delivered and increase access — for everyone. We can only succeed by bringing together diverse minds, thoughts, ideas and team members to create better solutions for our customers and make us a better company as a whole. We are committed to creating a culture of embracing diversity, inclusion and equity for all.

    QGenda is an Equal Employment Opportunity employer and makes all employment decisions without regard to race, color, religion, creed, gender, sex (including pregnancy), sexual orientation, gender identity or expression, natural origin, ancestry, age, marital status, disability or genetic information, military status, status as a disabled or protected veteran or any other protected status under applicable law.

    If you require accommodations or assistance to complete the online application process, please contact View email address on click.appcast.io and identify the type of accommodation or assistance you are requesting. Do not include any medical or health information in this email. We will respond to your email promptly.

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Senior Site Reliability Engineer in Atlanta, GA vacancy
  • Inspire Brands is hiring two Senior Site Reliability Engineers to help build and scale reliable, resilient, and observable systems supporting high-traffic, customer-facing digital platforms. These role blends software engineering, systems thinking, and operational excellence... 
    Senior
    Worldwide

    Inspire Brands

    Atlanta, GA
    1 day ago
  • $104.9k - $174.7k

     ...Data Management. You can learn more about LexisNexis Risk at the link below, About the Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the reliability of our production systems. This is not a purely advisory... 
    Senior
    Full time
    Work at office
    Local area
    Remote work
    Work from home

    LexisNexis Risk Solutions Group

    Atlanta, GA
    1 day ago
  •  ...Georgia, and serves customers in more than 35 countries worldwide.Position OverviewWe are seeking a highly experienced Senior Site Reliability Engineer (Unified Observability) to lead the design, implementation, and operational maturity of the F1 Next Generation... 
    Senior
    Full time
    Worldwide
    Flexible hours

    NCR

    Atlanta, GA
    1 day ago
  • $121.4k - $218.6k

     ...will be responsible for ensuring best-in-class uptime and reliability of our AI hardware infrastructure offerings. Partner...  ...and defend them when they are breached. As a Senior Site Reliability Engineer, you will be responsible for: Developing and scaling... 
    Senior
    Work experience placement
    Work at office

    Akamai

    Atlanta, GA
    18 hours ago
  • $120k - $175k

     ...Senior Site Reliability Engineer (SRE) Atlanta, GA preferred, Remote At PrizePicks, we are the fastest-growing sports company in North America, as recognized by Inc. 5000. As the leading platform for Daily Fantasy Sports, we cover a diverse range of sports leagues... 
    Senior
    Remote work
    Work visa
    Flexible hours

    PrizePicks

    Atlanta, GA
    3 days ago
  •  ...development, testing, implementation, and operation of secure, scalable, resilient, and highly available software platforms using Site Reliability Engineering and AI-native engineering practices. The engineer collaborates across technical teams to build and support cloud-native,... 
    Senior
    Full time
    Temporary work
    Part time
    Work experience placement
    Local area
    Flexible hours

    T-Mobile

    Atlanta, GA
    3 days ago
  •  ...Senior Systems Reliability Engineer (SRE)At T-Mobile, we invest in YOU! Our Total Rewards Package ensures that employees get the same big love we give our customers. All team members receive a competitive base salary and compensation package - this is Total Rewards. Employees... 
    Senior
    Full time
    Temporary work
    Part time
    Work experience placement
    Flexible hours

    T Mobile US

    Atlanta, GA
    3 days ago
  •  ...ideal time and number for communication, and the expected pay rate for C2C/1099/W2. Job Description: Job Title : Sr. Site Reliability Engineer Location : Atlanta, GA - Hybrid Duration : 6+ Months Contract Visa : US Citizens/ Green Card Need Local to... 
    Senior
    Contract work
    Local area
    Immediate start

    Navtech

    Atlanta, GA
    1 day ago
  • OneTrust is seeking a Senior Software Engineer in Atlanta, Georgia. The role involves designing and maintaining a reliable application platform, collaborating with engineering teams, and enhancing customer experiences through observability tools. The ideal candidate will... 
    Senior

    OneTrust

    Atlanta, GA
    2 days ago
  • $35 - $45 per hour

    DescriptionKforce has a client that is seeking a remote Site Reliability Engineer to join their team.Summary:The team consists of systems that can track lead management, job management and sales management. It is built on Salesforce but underpinned by a lot of Java/API'... 
    Remote work

    KForce

    Atlanta, GA
    2 days ago
  • $100k - $120k

    OverviewThe Site Reliability Engineer is a key force behind improving Origami’s time to resolution and advancing overall site reliability and scalability. This person participates in efforts to identify root causes during post-incident investigations, while also identifying... 
    Full time
    Temporary work
    Work experience placement
    Flexible hours

    Origami Risk

    Atlanta, GA
    1 day ago
  • $101.5k - $169.1k

     ...include an incentive program.Job DescriptionThe Release Train Engineer (RTE) has a primary purpose of supporting an Agile Release Train...  ...organizational AI policies and standards. Monitor AI tool reliability across teams. Create backup plans for system failures. Maintain... 
    Senior
    Full time
    Work at office
    Remote work
    Visa sponsorship
    Flexible hours

    Cox Enterprises

    Atlanta, GA
    2 days ago
  • $60 - $68 per hour

     ...Site Reliability Engineer Immediate need for a talented Site Reliability Engineer. This is a 12+ months contract opportunity with long-term potential and is located in Atlanta, GA (Onsite). Please review the job description below and contact me ASAP if you are interested... 
    Contract work
    Local area
    Immediate start

    Pyramid Consulting

    Atlanta, GA
    4 days ago
  • $95k - $171k

     .... Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building and maintaining dashboards, alerts... 
    Permanent employment
    Work experience placement
    Work at office
    Remote work
    Work from home
    Worldwide
    Flexible hours

    Akamai

    Atlanta, GA
    19 hours ago
  • $71.6k - $119.4k

     ...support application teams. Our services provide applications with reliability, security, and better customer experiences. About the Job...  ...automation, troubleshoot issues, and work closely with senior engineers to learn and apply best practices. You'll gain exposure to a... 
    Full time
    Temporary work
    Internship
    Local area
    Work from home

    RELX

    Atlanta, GA
    1 day ago
  •  ...provisioning, monitoring, and troubleshooting staging and production cloud environments . Experienced in architectural design for reliability, scalability, and performance. Practical application of SRE principles : SLIs, SLOs, error budgets, automation, incident... 

    Purple Drive

    Atlanta, GA
    3 days ago
  • Job description Snowflake SRE JD Your Role Accountabilities Primarily responsible for administrating Snowflake environments on AWS Identify, tune, and fix the performance issues on priority. Diagnose and troubleshoot Snowflake related errors and work with team to raise...

    Rcinfotech

    Atlanta, GA
    2 days ago
  •  ...Site Reliability Engineering LeadThe Site Reliability Engineering Lead role focuses on enhancing the reliability and operational excellence of enterprise...  ...across hybrid cloud and on-premises environments. This senior technical leader drives improvements in automation,... 

    Truist Inc

    Atlanta, GA
    2 days ago
  •  ...OpenShift - Site Reliability Engineer Atlanta , GA / Onsite Qualifications: This position is 60 % SRE and 40% SDE. Required Skillset • Manage and optimize data streaming and API components in OpenShift Onpremise and AWS. • Proactively... 
    Work experience placement

    Fisec Global

    Atlanta, GA
    4 days ago
  •  ...We have an immediate need for a Senior Release Train Engineer for a contract assignment located in Carmel, Indiana . The Release Train Engineer (RTE) has a primary purpose of supporting an Agile Release Train (ART) by steering it to success and navigating the complexity... 
    Senior
    Contract work
    Work at office
    Immediate start

    Spartan Technologies

    Atlanta, GA
    1 day ago
  • OverviewJob PurposeIntercontinental Exchange (ICE) presents a unique opportunity for a full-time Engineer to join a team responsible for ICE OpenShift container platform. The Engineer will drive enablement and adoption to migrate high performant and critical applications... 
    Senior
    Full time

    Black Knight Financial Services

    Atlanta, GA
    2 days ago
  • Job-ID31782898Reference26-00225 Job Title: ( Senior Software Configuration/Release Engineer ) About Kyyba: Founded in 1998 and headquartered in Farmington...  ...combined with career development. Job Description On-Site Interviews Only HYBRID - IN THE OFFICE 2 DAYS PER WEEK... 
    Senior
    Work at office
    Visa sponsorship
    Work visa
    2 days per week

    Kyyba

    Atlanta, GA
    2 days ago
  •  ...speed capabilities our nation and its allies need to maintain a durable, asymmetric advantage.About The Role:The Mission Systems Engineering (MSE) Team develops the Mission Management System (MMS)—a software platform that integrates mission subsystems, autonomy services... 
    Senior
    Weekly pay
    Permanent employment
    Full time
    Work at office

    Hermeus

    Atlanta, GA
    2 days ago
  •  ...Role: Site Reliability Engineering (SRE) Architect Location: Atlanta, GA (Hybrid on-site) Contract Role Summary: As an...  ...Technical Leadership & Consultation: Act as a senior technical advisor and subject matter expert on reliability... 
    Contract work
    Early shift

    AceStack LLC

    Atlanta, GA
    1 day ago
  •  ...complex, distributed, cloud-native systems. As a Staff Platform Engineer, you will play a critical role in ensuring these systems...  ...hands-on engineering and technical leadership role. You will own reliability for major platform domains, design scalable solutions on Kubernetes... 
    Senior

    Saviynt

    Atlanta, GA
    9 days ago
  • $207.4k - $298.1k

     ...enabled product roadmap.About the role:We are hiring a Senior Principal Software Engineer to lead UKG Ready SMB — Architecture & Design, with a heavy...  ...-to-market for cross-team integrations, operational reliability and cost-efficiency as Ready scales to its revenue... 
    Senior
    Worldwide

    Ultimate Software

    Atlanta, GA
    3 days ago
  •  ...are we looking for? We’re looking for experienced Software Engineers who are passionate about building rock-solid payment products...  ...You’ll know you’re the right candidate when you enjoy crafting reliable SaaS APIs and are also passionate about helping other developers... 
    Senior
    Full time
    Work experience placement
    Flexible hours

    The Rainforest

    Atlanta, GA
    19 hours ago
  • Job Title: Senior Software EngineerWork Location:Atlanta, GAJob Summary:Seeking a Senior Java developer with 3 to 5 years of experience to design and build cloud native applications leveraging Java technologies.Job Description:Design develop and maintain high quality Java... 
    Senior

    LTM

    Atlanta, GA
    19 hours ago
  •  ...native services, and infrastructure automation to improve reliability and speed to delivery. Our global team operates in a...  ...improvement and operational excellence.About the Role:As a Senior Network Platform Engineer, you will be a key technical contributor responsible for... 
    Senior
    Full time
    Work at office
    Flexible hours

    Invesco

    Atlanta, GA
    3 days ago
  •  ...Manager @ STAFFWORXS | US IT Recruitment Job Opening: AWS Site Reliability Engineer (SRE) We’re hiring a Site Reliability Engineer (SRE) to join...  ...interview is mandatory as part of the hiring process. Seniority level Seniority level Mid-Senior level Employment type Employment... 
    Contract work

    STAFFWORXS

    Atlanta, GA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!