Senior Site Reliability Engineer
QGenda
QGenda is redefining healthcare workforce management everywhere care is delivered. We're on a mission to empower the healthcare industry to better onboard, deploy, and manage their workforce. Over 4,500 healthcare organizations have trusted us to help them make strategic workforce decisions through our unified software platform. With more than 800 employees across the US, we are united in our vision and culture to make a difference for our customers and enjoy the day-to-day. At QGenda, we value our employees and their contributions toward the success of the business. We strive to create a dynamic work environment that fosters growth, innovation, and collaboration, where employees can be proud of the work they do and the impact it has on the healthcare industry. QGenda is headquartered in Atlanta. To learn more about QGenda, visit us at qgenda.com or follow us on Instagram or LinkedIn. As a Senior Site Reliability Engineer, you will work with our Infrastructure and Product Development Teams to increase the scalability, reliability, and performance of our systems and services. You will build and extend existing automation for configuration and monitoring of our AWS hosted applications, evaluate new AWS services and tools, and focus on platform health and monitoring to deliver the best experience for our customers. This role is hybrid with one required day in our Buckhead (Atlanta, Georgia) or Uniontown, Ohio office depending on your current location. Responsibilities System Reliability & Performance Design, implement, and manage scalable systems that ensure high availability, fault tolerance, and optimal performance. Continuously monitor and enhance system health and performance through data analysis and metrics. Automation & Tooling Develop and advocate for automation tools to eliminate repetitive manual processes and improve efficiency. Build and enhance CI/CD pipelines to streamline software delivery and deployments. Incident Management & Troubleshooting Participate in on‑call rotation to respond to incidents, troubleshoot problems, and minimize downtime. Conduct root cause analyses and implement permanent solutions to recurring issues. Infrastructure Management Manage our cloud‑based infrastructure environment in AWS. Optimize costs and resources while maintaining robust and scalable systems. Collaboration & Culture Serve as a technical advisor to engineering teams on infrastructure and operations best practices. Actively contribute to fostering an SRE culture within the organization by promoting observability, retrospectives, and continuous improvement. Qualifications Who You Are Curiosity‑driven mindset with a desire to continuously learn and improve systems Strong sense of ownership — you see problems through to resolution, not just escalation Comfortable navigating ambiguity and making pragmatic tradeoffs under pressure Availability for off‑hours deployment and upgrades of production systems during release and maintenance windows Strong problem‑solving skills and ability to work effectively under pressure. Excellent communication skills for cross‑functional collaboration as well as documentation creation. Experience You Bring B.S. in Computer Science, Computer Information Systems, or Computer Engineering from a major U.S. university or equivalent industry experience 7+ years of experience as a DevOps, SRE or Systems Engineer Advanced proficiency with at least one scripting or programming language Experience with Docker and container orchestration tools such as AWS ECS and EKS/Kubernetes Hands‑on experience building infrastructure and supporting applications in AWS using services such as Lambda, EC2, ECS, S3, SNS, SQS, RDS, Redshift, and Elasticache Strong understanding of networking and DNS Strong experience with Terraform for infrastructure provisioning and module development, along with configuration management and infrastructure as code (IaC) practices Firm understanding and experience with Agile and Scrum SDLC processes Experience using distributed version control system (Git preferred) to check‑in code, branching, merging, pull request, code review, etc Knowledge of CI/CD best practices and tools such as AWS CodeBuild, Jenkins and/or TeamCity Experience using AI‑assisted coding tools (e.g., Claude, GitHub Copilot) to accelerate IaC development, scripting, and operational workflows Familiarity with AI/ML‑driven approaches to observability, anomaly detection, log analysis, or incident triage Experience designing and delivering secure, high‑performance and highly available cloud services Experience with observability platforms (e.g., Datadog, CloudWatch, PagerDuty) for monitoring, alerting, and incident response Awareness of cloud security best practices including IAM policies, network segmentation, and secrets management Benefits Fully company‑paid options for medical (both in‑person and virtual), dental and vision insurance Generous paid time off (PTO) policy to enjoy periods of uninterrupted rest and relaxation for a healthy work/life balance Paid parental leave for birth, adoption or permanent placement 401(k) with company match Options to work in a hybrid‑working model or remotely from home, depending on the position Annual Costco membership, cell phone stipend, commuter benefits, in‑office perks and more QGenda delivers technology solutions to improve how healthcare is delivered and increase access — for everyone. We can only succeed by bringing together diverse minds, thoughts, ideas and team members to create better solutions for our customers and make us a better company as a whole. We are committed to creating a culture of embracing diversity, inclusion and equity for all. QGenda is an Equal Employment Opportunity employer and makes all employment decisions without regard to race, color, religion, creed, gender, sex (including pregnancy), sexual orientation, gender identity or expression, natural origin, ancestry, age, marital status, disability or genetic information, military status, status as a disabled or protected veteran or any other protected status under applicable law. If you require accommodations or assistance to complete the online application process, please contact View email address on click.appcast.io and identify the type of accommodation or assistance you are requesting. We will respond to your email promptly. #J-18808-Ljbffr
$152k - $195k
...class investors including Silver Lake Waterman, Moody’s, Sequoia Capital, GV and Riverwood Capital. About the Team As a Senior Site Reliability Engineer, you will be a key technical leader driving the design and optimization of our Kubernetes‑based infrastructure and CI/...Senior$130k - $135k
...VARITE is looking for a qualified Senior Site Reliability Engineer (SRE) - 619374 in Atlanta, GA About the client: An American Software company that provides a suite of tools intended to support the development and deployment of large-scale service-oriented software installations...SeniorFull time$99.09k - $123.86k
...Site Reliability Engineer (SRE) – AI Systems We are seeking an experienced Site Reliability Engineer who thrives at the intersection of software engineering, infrastructure, and AI systems. The role focuses on ensuring our platforms are scalable, reliable, and secure,...SeniorLocal areaFlexible hours- ...Lead Engineer, Site Reliability Engineering Team As a lead engineer with Retail, Site Reliability Engineering team, you will be at the forefront of Cloud and Big Data technology. In this role you will establish yourself as a technical leader by exposing yourself to...Senior
$109.5k
...is looking for a highly motivated, diligent, and skillful Site Reliability Engineer to join the Cyber Security Engineering (CSE) Team. The CSE... ...This position can be remote anywhere in the U.S. The Senior Site Reliability Engineer will be responsible for ensuring...SeniorTemporary workLocal areaRemote work$178.13k - $205.4k
...telecommuting. Salary Range: $178,131 - $205,400 About You Basic Qualification Bachelor’s degree or foreign degree equivalent in Computer Engineering, Computer Science, Engineering, or related field plus five (5) years of progressive, post‑baccalaureate experience in job offered...SeniorWork at officeRemote workFlexible hours- ...complex, distributed, cloud-native systems. As a Staff Platform Engineer, you will play a critical role in ensuring these systems... ...hands-on engineering and technical leadership role. You will own reliability for major platform domains, design scalable solutions on Kubernetes...Senior
$104k - $187k
...Job Summary The Senior Release Train Engineer (RTE) is the chief facilitator for a newly forming Software Agile Release Train (ART). This role is accountable for establishing and maturing the ART operating model, enabling predictable delivery, improving flow, and orchestrating...Senior- ...for new team members who want to be a part of this journey! Who We’re Looking For We’re looking for a proactive, hands‑on Site Reliability Engineer who thrives in building and scaling cloud infrastructure in fast‑moving startup environments. You’re someone who enjoys owning...Work experience placementFlexible hours
$130k - $160k
...Role At Todyl, our Application Platform Engineering team is dedicated to building infrastructure... ...work will not only directly impact the reliability and security of our platform but also... ...space. Responsibilities As a Platform SRE (Site Reliability Engineer) at Todyl, you will...Temporary workLocal areaFlexible hours- ...configure the monitoring and alerting metrics so the support engineers can proactively and timely validate, troubleshoot and... ...availability critical application components. • 1+ Years in Site Reliability Engineering organization preferred • Overall 4-6years of experience...Work experience placement
- ...Job title: Site Reliability Engineer (SRE) Location: Atlanta, GA 30303 Duration of the project: 12 Months Strong expertise in Ansible with an SRE background Ability to review and test GitLab Duo generated code CICD pipelines (GitLab, GitHub Actions) Infrastructure...
- ...Technical Support Specialist In Site Reliability Engineering (Sre) Mandatory skills: Scripting and programming languages like Python, Java, Ruby. Cloud and infrastructure management – AWS, Google cloud and Azure is a plus- CI/CD Automation, Database Management. The...
- ...advances cures by helping the world's most important research sites do their best work. Our solutions are now used by over 30,00... ...What You'll Bring to the Team: We are seeking a Site Reliability Engineer (SRE) to join one of our Scrum teams and help ensure the...Work at office
- ...availability. • Automation Experience with Build/deployment, Software Configuration/Continuous Integration/Continuous Delivery/Release Engineering related tasks in JavaEE/C++ Environments. • Experience in automating manual processes using Python, Ruby, Unix Shell (bash,...Immediate start
$117k - $209.33k
...Job Requisition ID # 26WD98046 Position Overview An exciting new opportunity has opened for a Site Reliability Engineer within the Autodesk PDMS Platform SRE team. The successful candidate will wear multiple hats: first responder, performance analyst,...Permanent employmentFor contractors$101.5k - $169.1k
...The Release Train Engineer (RTE) has a primary purpose of supporting an Agile Release Train (ART) by steering it to success and navigating... ...organizational AI policies and standards. Monitor AI tool reliability across teams. Create backup plans for system failures....SeniorWork at officeVisa sponsorship- ...Job Description Job Overview Job ID: J52010 Job Title: Site Reliability Engineering (SRE) Architect Location: Atlanta, GA Duration: 12 Months +... ...operational efficiency Technical Leadership & Consultation Act as a senior technical advisor and subject matter expert on reliability...Hourly payPermanent employmentContract workH1bLocal areaEarly shift
- ...Site Reliability Engineer We are looking for a Site Reliability Engineer to ensure the reliability, security, and continuous operation of a multi‑cloud application security platform. This role combines platform engineering and security automation, focusing on Kubernetes...Work at officeRemote workVisa sponsorshipWork visaFlexible hours
$178k - $213k
...Splunk Ventures, and Vista Credit Partners of Vista Equity Partners 2022 Cybersecurity Excellence Award for MDR Manager, Site Reliability Engineering Reports to: VP, Product Engineering Location: While proximity to Tampa is preferred to support hybrid schedule in Tampa...Permanent employmentWork experience placementWork at officeRemote workWork from homeHome officeFlexible hours$51.9 per hour
...Company: Allegheny Health Network Job Title: Site Reliability Engineering – Clinical & Facility Services General Overview This role ensures the reliability, availability, and performance of critical healthcare IT systems in the Environment of Care (EOC), supporting patients...Local area- ...Turner Services Inc. is hiring a Senior Software Engineer for CNN's Growth Team in Atlanta, Georgia. You will play a vital role in developing features for CNN's digital platform, working closely with cross-functional teams in product, design, analytics, and marketing....Senior
- ...person. Looking for candidates local to Atlanta, Georgia or Mahwah, New Jersey Job Summary: We are seeking a Senior MLOps / AIOps Platform Engineer with deep DevSecOps expertise and hands-on experience managing enterprise-grade AI/ML platforms. This critical role...SeniorLocal area
- A leading AI research accelerator is seeking a software engineer to evaluate and refine AI-generated code. Candidates must have over... ...designing verification mechanisms, and ensuring code efficiency and reliability. This is a contractor position requiring flexible engagement...SeniorContract workFor contractorsRemote work10 hours per weekFlexible hours
- ...Warner Media, LLC. seeks a Senior Software Developer to join the Roku team. This role involves designing, building, and maintaining software applications for streaming platforms used by millions. Ideal candidates will have over 5 years of experience and proficiency in...Senior
- ...product, operations, security, and engineering partners to deliver outcomes that scale. As a Senior Lead Software Engineer at... ...standards that improve reliability and delivery velocity. You will... ...comprehensive health care coverage, on-site health and wellness centers, a...Senior
- ...build, and deliver top-notch technology products. As a Senior Lead Software Engineer - Python/Pyspark/Databricks/AI at JPMorganChase within the... ...These benefits include comprehensive health care coverage, on-site health and wellness centers, a retirement savings plan,...SeniorFor contractors
$124.5k - $168.08k
...Infrastructure Modernization business unit is seeking a Senior Principal Software Engineer to serve as a visionary and hands-on technical leader for... ...ensure our offerings deliver robust protection, highly reliable, performant, operationally simple, and highly secure....SeniorRemote workWorldwide- ...Senior Software Engineer Microservices & DevOps (AI-Enabled) Job Qualification Write high-quality, production-ready Java code using Spring Boot... ...and infrastructure, with a focus on scalability, reliability, and security. DevOps & Cloud Linux Administration Shell scripting...Senior
$84k - $126k
...as integration diagrams, best‑practice guides, and FAQs. Support continuous improvement by providing feedback to Product and Engineering based on client experiences. Qualifications Minimum of 5 years related experience with a Bachelor’s degree, or 3 years with...Senior
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- site reliability engineer Atlanta, GA
- site reliability engineer remote Atlanta, GA
- site reliability engineer sre Atlanta, GA
- senior trade analyst Atlanta, GA
- senior app developer Atlanta, GA
- senior customer service advisor Atlanta, GA
- senior international account manager Atlanta, GA
- senior product manager mobile Atlanta, GA
- senior magento developer Atlanta, GA
- senior quantitative risk analyst Atlanta, GA


