Senior Site Reliability Engineer
QGenda
Senior Site Reliability Engineer
Atlanta, Georgia
Who We Are
QGenda is redefining healthcare workforce management everywhere care is delivered. We're on a mission to empower the healthcare industry to better onboarding, deploy, and manage their workforce. Over 4,500 healthcare organizations have trusted us to help them make strategic workforce decisions through our unified software platform. With more than 800 employees across the US, we are united in our vision and culture to make a difference for our customers, while enjoying the day-to-day.
At QGenda, we value our employees and their contributions toward the success of the business. We strive to create a dynamic work environment that fosters growth, innovation, and collaboration, where employees can be proud of the work they do and the impact it has on the healthcare industry.
QGenda is headquartered in Atlanta.
About Your Role
As a Senior Site Reliability Engineer, you will work with our Infrastructure and Product Development Teams to increase the scalability, reliability, and performance of our systems and services. You will build and extend existing automation for configuration and monitoring of our AWS hosted applications. You will have the opportunity to evaluate new AWS services and tools to determine if they could be utilized in our environments. You'll bring a focus to platform health and monitoring to allow us to deliver the best possible experience for our customers. This is an excellent opportunity to have a significant impact on the stability of our systems and contribute to the evolution of our technology stack.
NOTE: This role is hybrid with one required day in our Buckhead (Atlanta, Georgia) or our Uniontown, Ohio office depending on your current location.
How You'll Make an Impact
As a Senior Site Reliability Engineer, you will work with our Infrastructure and Product Development Teams to increase the scalability, reliability, and performance of our systems and services. You will build and extend existing automation for configuration and monitoring of our AWS hosted applications. You will have the opportunity to evaluate new AWS services and tools to determine if they could be utilized in our environments. You'll bring a focus to platform health and monitoring to allow us to deliver the best possible experience for our customers. This is an excellent opportunity to have a significant impact on the stability of our systems and contribute to the evolution of our technology stack.
Responsibilities:
System Reliability & Performance:
- Design, implement, and manage scalable systems that ensure high availability, fault tolerance, and optimal performance.
- Continuously monitor and enhance system health and performance through data analysis and metrics.
Automation & Tooling:
- Develop and advocate for automation tools to eliminate repetitive manual processes and improve efficiency.
- Build and enhance CI/CD pipelines to streamline software delivery and deployments.
Incident Management & Troubleshooting:
- Participate in on-call rotation to respond to incidents, troubleshoot problems, and minimize downtime.
- Conduct root cause analyses and implement permanent solutions to recurring issues.
Infrastructure Management:
- Manage our cloud-based infrastructure environment in AWS.
- Optimize costs and resources while maintaining robust and scalable systems.
Collaboration & Culture:
- Serve as a technical advisor to engineering teams on infrastructure and operations best practices.
- Actively contribute to fostering an SRE culture within the organization by promoting observability, retrospectives, and continuous improvement.
Who You Are
- Curiosity-driven mindset with a desire to continuously learn and improve systems
- Strong sense of ownership — you see problems through to resolution, not just escalation
- Comfortable navigating ambiguity and making pragmatic tradeoffs under pressure
- Availability for off-hours deployment and upgrades of production systems during release and maintenance windows
- Strong problem-solving skills and ability to work effectively under pressure.
- Excellent communication skills for cross-functional collaboration as well as documentation creation.
Experience You Bring
- B.S. in Computer Science, Computer Information Systems, or Computer Engineering from a major U.S. university or equivalent industry experience
- 7+ years of experience as a DevOps, SRE or Systems Engineer
- Advanced proficiency with at least one scripting or programming language
- Experience with Docker and container orchestration tools such as AWS ECS and EKS/Kubernetes
- Hands-on experience building infrastructure and supporting applications in AWS using services such as Lambda, EC2, ECS, S3, SNS, SQS, RDS, Redshift, and Elasticache
- Strong understanding of networking and DNS
- Strong experience with Terraform for infrastructure provisioning and module development, along with configuration management and infrastructure as code (IaC) practices
- Firm understanding and experience with Agile and Scrum SDLC processes
- Using distributed version control system experience (Git preferred) to check-in code, branching, merging, pull request, code review, etc
- Knowledge of CI/CD best practices and tools such as AWS CodeBuild, Jenkins and/or TeamCity
- Experience using AI-assisted coding tools (e.g., Claude, GitHub Copilot) to accelerate IaC development, scripting, and operational workflows
- Familiarity with AI/ML-driven approaches to observability, anomaly detection, log analysis, or incident triage
- Experience designing and delivering secure, high performance and highly available cloud services
- Experience with observability platforms (e.g., Datadog, CloudWatch, PagerDuty) for monitoring, alerting, and incident response
- Awareness of cloud security best practices including IAM policies, network segmentation, and secrets management
#LI-Hybrid
Applicants for this position must be authorized to work for any employer in the United States (U.S.), including being located in the US. We are unable to sponsor, take over sponsorship of, or hire candidates with an employment visa at this time.
What's In It For You
We offer a comprehensive total rewards package to support our full-time employees and their family's day-to-day needs, well-being and major life events, which includes:
- Fully company-paid options for medical (both in-person and virtual), dental and vision insurance
- Generous paid time off (PTO) policy to enjoy periods of uninterrupted rest and relaxation for a healthy work/life balance
- Paid parental leave for birth, adoption or permanent placement
- 401(k) with company match
- Options to work in a hybrid-working model or remotely from home, depending on the position
- Annual Costco membership, cell phone stipend, commuter benefits, in-office perks and more
QGenda delivers technology solutions to improve how healthcare is delivered and increase access — for everyone. We can only succeed by bringing together diverse minds, thoughts, ideas and team members to create better solutions for our customers and make us a better company as a whole. We are committed to creating a culture of embracing diversity, inclusion and equity for all.
QGenda is an Equal Employment Opportunity employer and makes all employment decisions without regard to race, color, religion, creed, gender, sex (including pregnancy), sexual orientation, gender identity or expression, natural origin, ancestry, age, marital status, disability or genetic information, military status, status as a disabled or protected veteran or any other protected status under applicable law.
If you require accommodations or assistance to complete the online application process, please contact View email address on click.appcast.io and identify the type of accommodation or assistance you are requesting. Do not include any medical or health information in this email. We will respond to your email promptly.
- Inspire Brands is hiring two Senior Site Reliability Engineers to help build and scale reliable, resilient, and observable systems supporting high-traffic, customer-facing digital platforms. These role blends software engineering, systems thinking, and operational excellence...SeniorWorldwide
$104.9k - $174.7k
...Data Management. You can learn more about LexisNexis Risk at the link below, About the Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the reliability of our production systems. This is not a purely advisory...SeniorFull timeWork at officeLocal areaRemote workWork from home- ...Georgia, and serves customers in more than 35 countries worldwide.Position OverviewWe are seeking a highly experienced Senior Site Reliability Engineer (Unified Observability) to lead the design, implementation, and operational maturity of the F1 Next Generation...SeniorFull timeWorldwideFlexible hours
$121.4k - $218.6k
...will be responsible for ensuring best-in-class uptime and reliability of our AI hardware infrastructure offerings. Partner... ...and defend them when they are breached. As a Senior Site Reliability Engineer, you will be responsible for: Developing and scaling...SeniorWork experience placementWork at office$120k - $175k
...Senior Site Reliability Engineer (SRE) Atlanta, GA preferred, Remote At PrizePicks, we are the fastest-growing sports company in North America, as recognized by Inc. 5000. As the leading platform for Daily Fantasy Sports, we cover a diverse range of sports leagues...SeniorRemote workWork visaFlexible hours- ...development, testing, implementation, and operation of secure, scalable, resilient, and highly available software platforms using Site Reliability Engineering and AI-native engineering practices. The engineer collaborates across technical teams to build and support cloud-native,...SeniorFull timeTemporary workPart timeWork experience placementLocal areaFlexible hours
- ...Senior Systems Reliability Engineer (SRE)At T-Mobile, we invest in YOU! Our Total Rewards Package ensures that employees get the same big love we give our customers. All team members receive a competitive base salary and compensation package - this is Total Rewards. Employees...SeniorFull timeTemporary workPart timeWork experience placementFlexible hours
- ...ideal time and number for communication, and the expected pay rate for C2C/1099/W2. Job Description: Job Title : Sr. Site Reliability Engineer Location : Atlanta, GA - Hybrid Duration : 6+ Months Contract Visa : US Citizens/ Green Card Need Local to...SeniorContract workLocal areaImmediate start
- OneTrust is seeking a Senior Software Engineer in Atlanta, Georgia. The role involves designing and maintaining a reliable application platform, collaborating with engineering teams, and enhancing customer experiences through observability tools. The ideal candidate will...Senior
$35 - $45 per hour
DescriptionKforce has a client that is seeking a remote Site Reliability Engineer to join their team.Summary:The team consists of systems that can track lead management, job management and sales management. It is built on Salesforce but underpinned by a lot of Java/API'...Remote work$100k - $120k
OverviewThe Site Reliability Engineer is a key force behind improving Origami’s time to resolution and advancing overall site reliability and scalability. This person participates in efforts to identify root causes during post-incident investigations, while also identifying...Full timeTemporary workWork experience placementFlexible hours$101.5k - $169.1k
...include an incentive program.Job DescriptionThe Release Train Engineer (RTE) has a primary purpose of supporting an Agile Release Train... ...organizational AI policies and standards. Monitor AI tool reliability across teams. Create backup plans for system failures. Maintain...SeniorFull timeWork at officeRemote workVisa sponsorshipFlexible hours$60 - $68 per hour
...Site Reliability Engineer Immediate need for a talented Site Reliability Engineer. This is a 12+ months contract opportunity with long-term potential and is located in Atlanta, GA (Onsite). Please review the job description below and contact me ASAP if you are interested...Contract workLocal areaImmediate start$95k - $171k
.... Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building and maintaining dashboards, alerts...Permanent employmentWork experience placementWork at officeRemote workWork from homeWorldwideFlexible hours$71.6k - $119.4k
...support application teams. Our services provide applications with reliability, security, and better customer experiences. About the Job... ...automation, troubleshoot issues, and work closely with senior engineers to learn and apply best practices. You'll gain exposure to a...Full timeTemporary workInternshipLocal areaWork from home- ...provisioning, monitoring, and troubleshooting staging and production cloud environments . Experienced in architectural design for reliability, scalability, and performance. Practical application of SRE principles : SLIs, SLOs, error budgets, automation, incident...
- Job description Snowflake SRE JD Your Role Accountabilities Primarily responsible for administrating Snowflake environments on AWS Identify, tune, and fix the performance issues on priority. Diagnose and troubleshoot Snowflake related errors and work with team to raise...
- ...Site Reliability Engineering LeadThe Site Reliability Engineering Lead role focuses on enhancing the reliability and operational excellence of enterprise... ...across hybrid cloud and on-premises environments. This senior technical leader drives improvements in automation,...
- ...OpenShift - Site Reliability Engineer Atlanta , GA / Onsite Qualifications: This position is 60 % SRE and 40% SDE. Required Skillset • Manage and optimize data streaming and API components in OpenShift Onpremise and AWS. • Proactively...Work experience placement
- ...We have an immediate need for a Senior Release Train Engineer for a contract assignment located in Carmel, Indiana . The Release Train Engineer (RTE) has a primary purpose of supporting an Agile Release Train (ART) by steering it to success and navigating the complexity...SeniorContract workWork at officeImmediate start
- OverviewJob PurposeIntercontinental Exchange (ICE) presents a unique opportunity for a full-time Engineer to join a team responsible for ICE OpenShift container platform. The Engineer will drive enablement and adoption to migrate high performant and critical applications...SeniorFull time
- Job-ID31782898Reference26-00225 Job Title: ( Senior Software Configuration/Release Engineer ) About Kyyba: Founded in 1998 and headquartered in Farmington... ...combined with career development. Job Description On-Site Interviews Only HYBRID - IN THE OFFICE 2 DAYS PER WEEK...SeniorWork at officeVisa sponsorshipWork visa2 days per week
- ...speed capabilities our nation and its allies need to maintain a durable, asymmetric advantage.About The Role:The Mission Systems Engineering (MSE) Team develops the Mission Management System (MMS)—a software platform that integrates mission subsystems, autonomy services...SeniorWeekly payPermanent employmentFull timeWork at office
- ...Role: Site Reliability Engineering (SRE) Architect Location: Atlanta, GA (Hybrid on-site) Contract Role Summary: As an... ...Technical Leadership & Consultation: Act as a senior technical advisor and subject matter expert on reliability...Contract workEarly shift
- ...complex, distributed, cloud-native systems. As a Staff Platform Engineer, you will play a critical role in ensuring these systems... ...hands-on engineering and technical leadership role. You will own reliability for major platform domains, design scalable solutions on Kubernetes...Senior
$207.4k - $298.1k
...enabled product roadmap.About the role:We are hiring a Senior Principal Software Engineer to lead UKG Ready SMB — Architecture & Design, with a heavy... ...-to-market for cross-team integrations, operational reliability and cost-efficiency as Ready scales to its revenue...SeniorWorldwide- ...are we looking for? We’re looking for experienced Software Engineers who are passionate about building rock-solid payment products... ...You’ll know you’re the right candidate when you enjoy crafting reliable SaaS APIs and are also passionate about helping other developers...SeniorFull timeWork experience placementFlexible hours
- Job Title: Senior Software EngineerWork Location:Atlanta, GAJob Summary:Seeking a Senior Java developer with 3 to 5 years of experience to design and build cloud native applications leveraging Java technologies.Job Description:Design develop and maintain high quality Java...Senior
- ...native services, and infrastructure automation to improve reliability and speed to delivery. Our global team operates in a... ...improvement and operational excellence.About the Role:As a Senior Network Platform Engineer, you will be a key technical contributor responsible for...SeniorFull timeWork at officeFlexible hours
- ...Manager @ STAFFWORXS | US IT Recruitment Job Opening: AWS Site Reliability Engineer (SRE) We’re hiring a Site Reliability Engineer (SRE) to join... ...interview is mandatory as part of the hiring process. Seniority level Seniority level Mid-Senior level Employment type Employment...Contract work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- site reliability engineer Atlanta, GA
- site reliability engineer sre Atlanta, GA
- srs distribution Atlanta, GA
- senior operations coordinator Atlanta, GA
- senior associate architect Atlanta, GA
- senior dynamics crm developer Atlanta, GA
- senior application security Atlanta, GA
- senior account director Atlanta, GA
- sr hr business partner Atlanta, GA
- senior supervisor Atlanta, GA


