Site Reliability Engineer
$107.9k - $195.05kKoitecc Solutions
The Digital Sector at Leidos currently has an opening for a Site Reliability Engineer (SRE) / Senior Cloud Engineer to work in our Baltimore, Maryland office. This is an exciting opportunity to use your experience helping the Center for Medicare and Medicaid Services (CMS) modernize its legacy Contact Center CRM platform within the CMS AWS Enclave and Pega Cloud for Government.
Primary Responsibilities
- Design, build, and operate highly available AWS infrastructure within the CMS AWS Enclave (FedRAMP Moderate), applying AWS Well-Architected Framework best practices.
- Architect secure, scalable multi-account VPC interconnectivity between the CMS AWS Enclave and Pega Cloud for Government (PCFG) in AWS GovCloud US-West, including PrivateLink, API Gateway, and Direct Connect.
- Support container and serverless architectures (e.g., AWS Lambda, Glue) for data integration, batch processing, and API layers supporting the modernized CRM.
- Apply Site Reliability Engineering (SRE) practices, including defining and tracking Service Level Objectives (SLOs), error budgets, and reliability metrics aligned to contract SLAs (e.g., ≥99.9% availability).
- Build and maintain observability across the AWS and Pega environments, including centralized logging, monitoring, and alerting, to enable proactive detection of performance and availability issues.
- Lead incident response and root-cause analysis for production issues, and long-term reliability improvements.
- Automate infrastructure provisioning, configuration, and environment build-out using Infrastructure as Code (e.g., Terraform, CloudFormation, Ansible).
- Design and test Disaster Recovery capability for cloud-based workloads, including backup, failover, and Multi-AZ/Multi-Region resilience.
- Support performance testing and capacity planning to validate the platform's ability to scale to 20,000 concurrent CSR sessions and peak Open Enrollment Period (OEP) volumes.
- Support continuous security monitoring, vulnerability remediation, and Zero Trust alignment across the AWS and Pega environments.
- Partner with the DevOps Lead/Configuration Manager to build and maintain CI/CD pipelines, ensuring automated testing, security scanning, and deployment across all SDLC environments.
- Support cloud connectivity and data movement for the AWS-based data migration pipeline (e.g., AWS Glue, S3, RDS/Aurora PostgreSQL) between legacy Siebel and the modernized CRM.
- Coordinate with the CMS Hybrid Cloud Team on cloud environment provisioning, patching, and lifecycle management activities.
- Manage release coordination and change windows in support of OEP blackout periods and other critical operational periods, minimizing risk of service disruption.
- Collaborate with the Release Train Engineer, Solution Architect, and Agile delivery teams to align infrastructure readiness with sprint and PI planning.
- Support integration of Genesys Cloud CX infrastructure and telephony/chat channels with the modernized CRM environment.
- Continuously identify opportunities to reduce operational toil through automation of repetitive tasks, log analysis, and routine operational activities.
- Document cloud architecture, operational runbooks, and disaster recovery procedures to support the Transition-Out Plan and audit readiness.
- Provide technical mentoring and knowledge-sharing to other engineers on cloud architecture, automation, and reliability engineering best practices.
- Communicate technical status, risks, and dependencies to CMS leadership, and Leidos management.
- Support requirements traceability and technical documentation related to infrastructure and integration architecture.
- Actively participate in planning sessions, requirements gathering activities, design sessions, Agile sessions, and other events supporting the CRM modernization effort.
Required Qualifications:
- Bachelor’s degree and a minimum of 6-8 years of relevant experience in cloud engineering, site reliability engineering, or infrastructure, or an equivalent combination of education and experience
- Experience building highly available AWS infrastructure based on industry best practices and the AWS Well-Architected Framework
- Experience with Infrastructure as Code, automation, and configuration management of cloud-based resources
- Experience designing Disaster Recovery for cloud-based workloads, including Multi-AZ/Multi-Region resilience
- Ability to obtain Public Trust
Preferred Qualifications:
- AWS Solutions Architect or SysOps Administrator Certification
- Experience with AWS GovCloud and FedRAMP-authorized cloud environments
- Experience supporting Pega Cloud for Government (PCFG) or similar SaaS platform connectivity (e.g., AWS PrivateLink, VPC peering)
- Familiarity with container (Docker, ECS, EKS) and serverless (Lambda) architectures
- Experience with observability/monitoring tooling (e.g., Splunk, New Relic, CloudWatch) and incident response practices
- Experience supporting federal contact center or other 24x7 mission-critical government systems
- Agile delivery experience
If you're looking for comfort, keep scrolling. At Leidos, we outthink, outbuild, and outpace the status quo - because the mission demands it. We're not hiring followers. We're recruiting the ones who disrupt, provoke, and refuse to fail. Step 10 is ancient history. We're already at step 30 - and moving faster than anyone else dares.
Pay Range
Pay Range $107,900.00 - $195,050.00
The Leidos pay range for this job level is a general guideline only and not a guarantee of compensation or salary. Additional factors considered in extending an offer include (but are not limited to) responsibilities of the job, education, experience, knowledge, skills, and abilities, as well as internal equity, alignment with market data, applicable bargaining agreement (if any), or other law.
About Leidos
Leidos is an industry and technology leader serving government and commercial customers with smarter, more efficient digital and mission innovations. Headquartered in Reston, Virginia, with 47,000 global employees, Leidos reported annual revenues of approximately $16.7 billion for the fiscal year ended January 3, 2025. For more information, visit
Pay and Benefits
Pay and benefits are fundamental to any career decision. That's why we craft compensation packages that reflect the importance of the work we do for our customers. Employment benefits include competitive compensation, Health and Wellness programs, Income Protection, Paid Leave and Retirement. More details are available at
Commitment to Non-Discrimination
All qualified applicants will receive consideration for employment without regard to sex, race, ethnicity, age, national origin, citizenship, religion, physical or mental disability, medical condition, genetic information, pregnancy, family structure, marital status, ancestry, domestic partner status, sexual orientation, gender identity or expression, veteran or military status, or any other basis prohibited by law. Leidos will also consider for employment qualified applicants with criminal histories consistent with relevant laws.
#J-18808-Ljbffr$180k - $230k
...Acceleration Job Description We're looking for a Senior SRE to own the reliability, scalability, and observability of our production systems. You'll work closely with platform and data engineering to keep high-throughput, data-intensive services running at the...SuggestedWork at officeLocal areaImmediate startRemote work3 days per week- ...Engineering, Product, Design, and Marketing Engineering Compensation ~ Zone 1 Base Pay: $214K – $260K Superhuman offers... ...role will be responsible for building software to ensure the reliability of our back-end systems, working with engineers who develop them...SuggestedWorldwideHome officeFlexible hours
$150.4k - $277.6k
...Services The Media Platforms SRE team under the Apple Service Engineering division is one of the most exciting examples of Apple’s long... ...field with 4+ years experience At least 6 years in a Reliability Engineering, DevOps or infrastructure focused role Advanced...SuggestedRelocationDay shift$182.8k - $247.3k
...to develop education for our half a billion (and growing!) learners around the world. About the role... As a Senior Site Reliability Engineer, you will work closely with both product and platform engineering teams to ensure Duolingo’s sophisticated distributed systems...SuggestedWork experience placement- ...A senior Site Reliability Engineer will join an established infrastructure function responsible for highly available, security-conscious cloud systems supporting complex business-critical workloads. You’ll take significant ownership of reliability, scalability, and...SuggestedFull timeRemote work
$110k - $145k
...global, with headquarters in Denver, Colorado, and offices across the U.S., Canada, and India. We are seeking a Senior Site Reliability Engineer to own the reliability, scalability, performance, and operational integrity of critical production services. This role is...Flexible hours- ...personalized care faster. We are building AI agents to support the full arc of the patient journey. The Opportunity: Machine Learning Engineer Patients count on our platform 24/7. You'll build and maintain the tooling, alerts and incident-response playbooks that keep...
- ...profitable developer-tooling company whose product is used by engineering teams at thousands of software companies for application... ...well-resourced group of nine. As Senior SRE you will lead reliability initiatives across the platform — from defining and driving SLOs...
$114k - $148k
...Total compensation is based on experience, skills, and location using objective, job-related criteria. Summary As a Site Reliability Engineer, you will focus on ensuring the platform and services customers rely on are reliable, performant, and highly available. If...Work experience placement- ...Cloudflare, GitHub Actions, PostgreSQL, Redis/BullMQ, Node.js/NestJS, Datadog, TypeScript, React, SQL Position: Senior Site Reliability Engineer Engagement period: Ongoing Interview timeline: ASAP Interview process: 1) CV review 2) Interview with our CTO 3)...Contract workImmediate start
$115.5k - $164.8k
...matters at a company where you matter. Your Impact As an engineer on the APX SRE CloudOps team, you will spend a significant portion... ...that replace what previously required human intervention with reliable, tested automation. You will also participate in on‑call...Work experience placementWork at officeRemote work$150k - $220k
...Senior Site Reliability EngineerJob detailsDepartment / EngineeringRemoteFull-time$150,000 USD - $220,000 USD## About UsMetaRouter is a customer... ...architecture.## About The RoleAs a Senior Site Reliability Engineer, you own significant pieces of our infrastructure and...Full timeRemote work$140k - $195k
...Improve reliability, observability, service health, incident response, and operational readiness. CodeVertex works across data... ..., secure systems, and operational clarity matter. The Site Reliability Engineer role helps turn business needs into reliable execution, whether...Remote work- ...Quarterhill is seeking a Senior Site Reliability Engineer (SRE) to join our growing team. This role is an exciting opportunity to contribute to the reliability and performance of smart transportation systems, including a next-generation, cloud-native tolling platform that...Local area
- ...% uptime. You'll own SLOs, incident response, and production reliability for a system that processes millions of identity verifications... ...Sentry error tracking, structured logging Implement chaos engineering practices to proactively identify failure modes Optimize...Remote work
$148.5k - $223.9k
...Salesforce is seeking a senior engineering candidate to join the Site Reliability organization in San Francisco. Working closely with counterparts in the Infrastructure and R&D organizations, this organization provides a global team of engineers monitoring cloud service...WorldwideWeekend work- ...As a Senior Site Reliability Engineer on our cloud engineering team, you'll keep our production environment healthy, secure, and running smoothly. This is an operations-focused role: you'll own the day-to-day administration of our AWS accounts and databases, backup posture...Work experience placement
$110k - $145k
...operations. You will liaise with product and engineering teams to ensure applications and... ...feedback loop for platform and product reliability. The ideal candidate is a solutions-oriented... ...experience as a platform engineer, site reliability engineer, systems engineer...Work experience placement- ...Discover exciting DevOps job opportunities and connect with 28,396 DevOps professionals. The Senior Site Reliability Engineer role at Jobicy is designed for experienced professionals who are passionate about enhancing system reliability and operational efficiency. The...Remote workFlexible hours
- ...Zof AI is seeking a Site Reliability Engineer to run the infrastructure that lets fleets of sandboxed agents execute customer code safely and cheaply. This role owns the execution layer of our control plane: Kubernetes and container orchestration, CI/CD pipelines, hard...Full time
- ...75+ countries, including businesses, developers, IT professionals, and individuals. About the Role We are seeking a Site Reliability Engineer II (SRE II) to help ensure the stability, scalability, and reliability of our services and infrastructure. This role focuses...
$185 per hour
...important clinical and business workflows, so they must be available, secure, and easy to operate. We are hiring a Senior Site Reliability Engineer to join our Security and Site Reliability team. You will focus on a stable, scalable AWS and Kubernetes platform, reliable...Work at office- ...join us on our mission of providing humankind access to the galaxy beyond our planet. About the Role We are seeking a Site Reliability Engineer to join our Ground Software team. As a Site Reliability Engineer, you will design, build, and operate the ground and site...Full timeWork at office
- ...electronic production — all running on AWS with zero downtime tolerance for firms in active litigation. We're looking for a Site Reliability Engineer to help maintain the reliability, scalability, and security posture of that platform as we expand our AI capabilities (...
$95k - $171k
.... Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building and maintaining dashboards, alerts...Permanent employmentWork experience placementWork at officeRemote workWork from homeWorldwideFlexible hours- ...to grow, adapt and use your skills consistently. Our customers rely on us in the moments that matter. Engineering delivers on that promise.The Senior Site Reliability Engineer is responsible for ensuring our SaaS products are fast, stable and optimized for our customers...Work experience placementFlexible hours
$135.2k - $181.2k
...enhance electrical, mechanical, and sensor-based systems to ensure reliability and performance. Configure, calibrate, and validate... ...professional development, including an interest in emerging data engineering tools and methodologies. Preferred Qualifications: ~8+...Worldwide- ...match. The role We're looking for a Senior SRE to own the reliability, scalability, and operational posture of Satsuma's multi-... ...using AI-assisted development workflows Partner closely with engineering on reliability reviews and architecture decisions ~5-8...
$180k - $200k
...Come join tastytrade, part of IG Group, as we build the reliability practice behind the brokerage platform that active options... ...equities traders rely on every market day. As our first Senior Site Reliability Engineer, you'll define what reliability means at tastytrade, from...Work at office3 days per week$104k - $178k
## Sr. Site Reliability Engineer IApply: Hybrid: NYC Global HQ: Full time: Posted 12 Days Ago: JR00000779# ****Who We Are****DV is the leader in digital performance solutions, helping our advertiser and agency partners Verify the quality of their digital campaigns, Optimise...Full time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- site reliability engineer Eastern, KY
- site safety Eastern, KY
- on-site clinical research associate (traveling/remote) Eastern, KY
- construction site safety Eastern, KY
- junior website developer Eastern, KY
- historic site Eastern, KY
- IT site lead Eastern, KY
- site leader Eastern, KY
- official site Eastern, KY
- junior site reliability engineer

