Site Reliability Engineer
$158.5k - $230kMedallia
Overview Medallia is the pioneer and market leader in Experience Management. Our award-winning SaaS platform, Medallia Experience Cloud, leads the market in the management of experiences, insights, and actions for candidates, customers, employees, patients, and residents alike. We believe that every experience is a memory that can last a lifetime. Experiences shape the way people feel about a company. And they greatly influence how likely people are to advocate, contribute, and stay. At Medallia, we are committed to creating a world where organizations are loved by their customers and their employees. We empower exceptional people to create extraordinary experiences together. Bring your whole self. The Role and Team We are growing our GovCloud team and looking for a Staff Site Reliability Engineer to help scale how we operate Medallia's US public-sector cloud platform. You will support federal agencies and other regulated customers in a highly available, secure, and compliant environment built on AWS GovCloud and Kubernetes. This is a hybrid role based near Tysons, Virginia, with regular in-office collaboration and remote flexibility. As a Staff Engineer, you will provide technical leadership - improving how we run the platform, mentoring teammates, and driving reliability improvements across teams - while still being hands-on in production. Success in this role means delivering change safely in a regulated environment: strong automation, clear documentation, disciplined operations, and calm incident response under pressure. Responsibilities Design, build, and operate highly available, secure cloud infrastructure on AWS, including networking, identity/access management, Kubernetes clusters, DNS, certificates, and shared platform services. Design and operate AWS cloud networking end-to-end - VPC architecture, subnetting and routing, security groups/NACLs, VPC endpoints/PrivateLink, Transit Gateway, load balancing, and DNS - for secure, segmented, highly available connectivity. Operate and tune production PostgreSQL - high availability and replication, backups and recovery, query and performance optimization, version upgrades, and capacity planning - as part of the platform's data tier. Ensure the reliability and availability of Medallia applications and infrastructure by monitoring systems, responding to incidents, and eliminating recurring operational problems. Develop and maintain Infrastructure-as-Code (primarily Terraform) and Kubernetes deployment workflows using Git, CI/CD, and modern GitOps practices. Improve observability across metrics, logs, and uptime monitoring; tune alerting to reduce noise and speed up diagnosis. Partner with software engineering, security, release management, and customer-facing teams to deploy changes safely, resolve production issues, and improve operability. Lead or contribute to platform upgrades, security patching, and compliance-driven maintenance in a regulated cloud environment. Participate in an on-call rotation and help improve incident response, communication, and post-incident follow-through. Document systems and operational procedures clearly so others can run and improve the platform. Use AI-assisted tooling responsibly, with attention to security, privacy, and customer data boundaries. Mentor engineers and help raise engineering standards as the GovCloud platform and team grow. Candidates based in the Tysons vicinity will be prioritized as this role is Hybrid, 3 days per week onsite. Qualifications Minimum Qualifications Eligibility Requirement: Must reside in the U.S. and hold U.S. Citizenship or a Green Card (Lawful Permanent Resident) to meet AWS GovCloud compliance requirements. Bachelor's degree or equivalent experience in Computer Science or a related field. 8+ years of experience in Site Reliability Engineering, platform engineering, DevOps, or related production infrastructure roles - or 5+ years with demonstrated Staff-level scope (technical leadership, cross-team delivery, incident ownership, and platform/IaC ownership). Production experience with: Kubernetes AWS core services (IAM, compute, object storage, encryption/key management) and AWS cloud networking (VPC design, routing, security groups, load balancing, VPC endpoints/PrivateLink, Transit Gateway, DNS) Terraform or comparable infrastructure-as-code tools Git and CI/CD pipelines Linux and foundational systems concepts (networking, DNS, TLS/certificates) PostgreSQL (or comparable relational databases) in production - replication, backups, and performance tuning Programming and Automation: Proficiency in Python and/or Go experience to build automation scripts, operational tooling, and infrastructure services. Incident & Change Management: Experience troubleshooting production incidents, conducting root-cause analysis(RCA), and following change management processes. Experience participating in a production on-call rotation. Experience troubleshooting complex technical issues and writing clear documentation, runbooks and incident post-mortems. Preferred Qualifications Experience operating in FedRAMP, AWS GovCloud, or other regulated or compliance-heavy cloud environments. Familiarity with security and compliance practices such as FIPS, vulnerability management, and controlled production change processes. Experience with observability and logging platforms in enterprise production environments. Deep operational expertise with PostgreSQL (HA/replication, tuning, backup and recovery); familiarity with data/platform technologies such as Redis and Kafka. Experience supporting federal agencies or public-sector customers. Exposure to government networking and security requirements. Experience with tools such as Jenkins, Argo CD, and GitHub Enterprise. Excellent collaboration skills and a strong willingness to learn. Medallia is committed to equal pay and transparency. The annual base salary range for this position is $158,500 - $230,000. Please note that the salary range information provided is a general guideline and combines all of the distinct labor markets within the US. It is uncommon for an individual to be hired at or near the top of the range for their role and compensation decisions are dependent on a variety of factors. Medallia considers factors such as (but not limited to) scope and responsibilities of the position, candidate's work experience, candidate's work location, education/training, key skills, internal peer equity, external market data, as well as, market and business considerations when making compensation decisions. Medallia also offers competitive health and wellness benefits, including but not limited to medical, dental, vision, 401(k), short-term and long-term disability, life and AD&D insurance, statutory leaves, paid parental leave, and paid holidays. Benefits and eligibility may vary by location and role. At Medallia, we celebrate diversity and recognize the value it brings to our customers and employees. Medallia is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, age (40 and over), disability, genetic information, veteran status or military service, or any other status protected by state or local law. Individuals with a disability who need an accommodation to apply please contact us at . For information regarding how Medallia collects and uses personal information, please review our Privacy Policies. Applications will be accepted for 30 days from the date this role was posted or until the role has been filled.
$133k - $190k
Site Reliability Engineer needed for a full time opportunity with SOC's direct client based in Herndon, VA. Direct Hire Role **Due to federal requirements, candidates must hold and possess an Active DOW TS/SCI security clearance to be considered for this role.** SOC is...SuggestedFull time$210k - $230k
GovCIO is currently hiring for a Senior Site Reliability Engineer (SRE) to design, implement, and maintain highly available, scalable, and resilient infrastructure systems. The ideal candidate will bridge the gap between development and operations, focusing on automation...SuggestedCurrently hiringRemote work$80k - $133k
...degree, Four (4) years additional experience will be needed.Minimum Four (4) years of experience in IT administration, software engineering, or platform engineering, with a focus on AWS cloud infrastructure and enterprise systems.One(1)+ years of experience deploying and...SuggestedPermanent employmentFull timeContract workRemote workFlexible hours$230k - $250k
GovCIO is hiring a Site Reliability Engineer with an active Secret clearance to ensure reliability, scalability, performance, and availability of mission-critical systems by combining software engineering practices with infrastructure operations expertise. This role is...SuggestedRemote work$150k - $180k
...Umbra.About the JobWe are seeking an experienced SeniorSite Reliability Engineer to help design, build, operate, and scale the mission- and business... ...impact across the organization.This position is based on-site in either our Arlington, VA office, Reston, VA office or...SuggestedPermanent employmentFull timeWork at officeLocal areaRemote workWorldwide- We're seeking a skilled and proactive Site Reliability Engineer to join our team, ensuring the stability, security, and efficiency of our technological resources as we deliver cutting-edge AI solutions to the government. This is a fully remote position for candidates in...Remote work
$84.24k - $142.48k
OverviewJoin us to work collaboratively with our talented team of dynamic and passionate engineers to deliver capabilities that enable our customers to make a difference. You'll deploy and operate ArcGIS Velocity and ArcGIS Workflow Manager SaaS solutions. You will also...WorldwideFlexible hours$190k - $225k
...Job TitleSRE + Release Pipeline EngineerJob DescriptionSRE + Release Pipeline Engineer Clearance: Top Secret clearance Location: Remote Role Framing Owns the path from local development to deployed AWS cluster across DCSA's GovCloud (IL2/IL5) and classified (IL6/Secret...Full timePart timeWork experience placementLocal areaRemote work$128.5k - $190k
...We empower exceptional people to create extraordinary experiences together. Bring your whole self. The Role and Team The Site Reliability Engineering organization at Medallia brings together the infrastructure and applications that power a highly reliable global SaaS...Temporary workWork experience placementLocal area$87.1k - $157.45k
...throughout the entire USG arsenal. Our team of hackers, engineers, makers, and shakers brings deep experience across... ...to come in and help us build systems that stay reliable when things get complicated. We need a Site Reliability Engineer who has experience building, deploying...Local areaImmediate startWork from homeFlexible hours$81.1k - $187k
...infrastructure and/or service according to terms for reliability and functionality.- Assists team members... ...deployments.- Gains basic knowledge of site reliability trends and shares relevant... ...are seeking a skilled Site Reliability Engineer to design, build, operate, and automate...Temporary workImmediate startFlexible hoursShift work$166k - $220k
...requirements and customer expectations. Our systems integration engineers internalize the nuances of each deployment, ensuring the... ...-to-end solutions we ship.ABOUT THE JOBWe are looking for a Site Reliability Engineer (SRE) to join AGD, our rapidly growing team in Costa...Full timeWork experience placementImmediate start$185k - $230k
As a Sr. Site Reliability Engineer (SRE) III, you’ll work as part of a collaborative and high-performing team providing your expertise to deliver technical solutions within the highest levels of the federal government.We know that you can’t have great technology services...Full timeLocal areaImmediate start$112k - $179k
...delivery of system, network, software, and security solutions.About The RolePeraton is seeking a self-driven and resourceful Site Reliability Engineer to join our dynamic of Network and UC engineers in Washington, DC. This position combines software engineering and systems...Contract workWorldwideShift work- ...candidates that are particularly strong in a few areas, and have some interest and capabilities in others.About the Role:As a Site Reliability Engineer, you’ll join the global Platform SRE team responsible for building, operating, and scaling Kong’s multi-region SaaS...Temporary work
$119.8k - $234.7k
...per yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole type: Individual... ...: Software EngineeringDiscipline: Site Reliability EngineeringCompany:... ...opportunity for a Senior Site Reliability Engineer (SRE) to join the Azure Silver and Sovereign...Ongoing contractLocal area3 days per week$135.8k - $183.8k
...dynamic and flexible work environment with competitive benefits and the ability to grow your career.We are looking for a Site Reliability Engineer to support our team responsible for building, managing, maintaining, deploying, and securing mission-critical services to...Work at officeFlexible hours$165k - $230k
...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARSHIELD)Starshield leverages SpaceX’s Starlink technology and launch capability to support national security efforts....Permanent employmentTemporary workImmediate startWeekend work$125k - $185k
Washington, D.C.Engineering /Full-time /HybridA World-Changing CompanyPalantir builds the world’s leading software for data-driven decisions... ...locate missing children, and more.The RoleWe’re looking for Site Reliability Engineers who can help us build, operate, and maintain high-...Full timeWork experience placementWork at officeRemote workWork from homeRelocation package$146k - $194k
...focused on positioning Anduril as a lead provider of specialized engineering and products for Intelligence Community (IC) customers. We... ...pressing national security requirements.ABOUT THE JOBAs a Site Reliability Engineer, your primary mission is to ensure the health,...Full timeWork experience placementImmediate startRemote work$125k - $185k
...lifesaving drugs, forecast supply chain disruptions, locate missing children, and more.The RoleWe’re looking for Forward Deployed Site Reliability Engineers who can help us build, operate, and maintain high-performance, scalable, and reliable services for our production...Full timeWork experience placementWork at officeRemote workWork from homeRelocation package$109.18k - $163.77k
...channels. As a global company, we have offices in nine countries and can insert advertisements around the world.Job SummaryThe Site Reliability Engineering team is responsible for managing the critical infrastructure that powers FreeWheel's Streaming Hub platform. Streaming...Full time$91.4k - $187k
Work with Site Reliability Engineering (SRE) team on the shared full stack ownership of a collection of services and/or technology areas. Understand the end-to-end configuration, technical dependencies, and overall behavioral characteristics of production services. Responsible...Temporary workFlexible hoursShift workWeekend work$175k - $195k
...products need key features. They need to be reliable, scalable, performant, cost effective,... ...for thinking through these problems and engineering solutions to them. We hire excellent... ...completed or currently under development.Site Reliability EngineerAs a Site Reliability...Full timeTemporary workWork experience placementRemote work- ...to grow, adapt and use your skills consistently. Our customers rely on us in the moments that matter. Engineering delivers on that promise. The Senior Site Reliability Engineer is responsible for ensuring our SaaS products are fast, stable and optimized for our customers...Work experience placementRemote workFlexible hours
- ...A leading security infrastructure firm in Washington, D.C. is seeking a hands-on Site Reliability Engineer (SRE) with expertise in Kubernetes and cloud infrastructure. The role emphasizes total ownership of security infrastructure while defending against advanced threats...
$207k - $284.9k
...on this mission. If you are too, let's talk.Senior Manager, Site Reliability EngineeringSecure Every Identity, from AI to HumanIdentity is... ...mission. If you are too, let's talk.The Federal Operations Engineering GroupOkta's Federal Operations team supports government customers...Permanent employmentLocal areaWorldwideFlexible hoursDay shift$102k - $234.6k
...members in designing and architecting infrastructure and service for reliability and functionality. Provides day-to-day direction to help... ...to experiment with new technology, execute improvements, build site reliability knowledge, and provide clear data.Only Oracle brings...Temporary workImmediate startFlexible hours$174k - $239k
...From core infrastructure to enterprise platforms, we partner across functions to drive scale, reliability, and innovation through technology.The Staff Site Reliability Engineer OpportunityOkta Federal, Inc. is looking for an experienced Staff TDI Site Reliability...Local areaWorldwideFlexible hours$84.9k - $209.5k
.... You’ll partner with customer support, service owners, and engineering teams around the globe to ensure high-quality service for customers... ...posted.Career Level - IC4Escalation points for junior site reliability engineers during complex or high-impact incidents.Manage and...Temporary workMonday to FridayFlexible hoursShift workNight shift
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- website content developer McLean, VA
- site leader McLean, VA
- on-site clinical research associate (traveling/remote) McLean, VA
- official site McLean, VA
- site recruiter McLean, VA
- historic site McLean, VA
- IT site lead McLean, VA
- junior website developer McLean, VA
- site safety McLean, VA
- construction site safety McLean, VA

