Staff Site Reliability Engineer
$158.5k - $230kMedallia
Medallia is the pioneer and market leader in Experience Management. Our award-winning SaaS platform, Medallia Experience Cloud, leads the market in the management of experiences, insights, and actions for candidates, customers, employees, patients, and residents alike.
We believe that every experience is a memory that can last a lifetime. Experiences shape the way people feel about a company. And they greatly influence how likely people are to advocate, contribute, and stay. At Medallia, we are committed to creating a world where organizations are loved by their customers and their employees.
We empower exceptional people to create extraordinary experiences together.
Bring your whole self.
The Role and Team
We are growing our GovCloud team and looking for a Staff Site Reliability Engineer to help scale how we operate Medallia's US public-sector cloud platform. You will support federal agencies and other regulated customers in a highly available, secure, and compliant environment built on AWS GovCloud and Kubernetes.
This is a hybrid role based near Tysons, Virginia, with regular in-office collaboration and remote flexibility. As a Staff Engineer, you will provide technical leadership - improving how we run the platform, mentoring teammates, and driving reliability improvements across teams - while still being hands-on in production.
Success in this role means delivering change safely in a regulated environment: strong automation, clear documentation, disciplined operations, and calm incident response under pressure.
- Design, build, and operate highly available, secure cloud infrastructure on AWS, including networking, identity/access management, Kubernetes clusters, DNS, certificates, and shared platform services.
- Design and operate AWS cloud networking end-to-end - VPC architecture, subnetting and routing, security groups/NACLs, VPC endpoints/PrivateLink, Transit Gateway, load balancing, and DNS - for secure, segmented, highly available connectivity.
- Operate and tune production PostgreSQL - high availability and replication, backups and recovery, query and performance optimization, version upgrades, and capacity planning - as part of the platform's data tier.
- Ensure the reliability and availability of Medallia applications and infrastructure by monitoring systems, responding to incidents, and eliminating recurring operational problems.
- Develop and maintain Infrastructure-as-Code (primarily Terraform) and Kubernetes deployment workflows using Git, CI/CD, and modern GitOps practices.
- Improve observability across metrics, logs, and uptime monitoring; tune alerting to reduce noise and speed up diagnosis.
- Partner with software engineering, security, release management, and customer-facing teams to deploy changes safely, resolve production issues, and improve operability.
- Lead or contribute to platform upgrades, security patching, and compliance-driven maintenance in a regulated cloud environment.
- Participate in an on-call rotation and help improve incident response, communication, and post-incident follow-through.
- Document systems and operational procedures clearly so others can run and improve the platform.
- Use AI-assisted tooling responsibly, with attention to security, privacy, and customer data boundaries.
- Mentor engineers and help raise engineering standards as the GovCloud platform and team grow.
Candidates based in the Tysons vicinity will be prioritized as this role is Hybrid, 3 days per week onsite.
QualificationsMinimum Qualifications
- Eligibility Requirement: Must reside in the U.S. and hold U.S. Citizenship or a Green Card (Lawful Permanent Resident) to meet AWS GovCloud compliance requirements.
- Bachelor's degree or equivalent experience in Computer Science or a related field.
- 8+ years of experience in Site Reliability Engineering, platform engineering, DevOps, or related production infrastructure roles - or 5+ years with demonstrated Staff-level scope (technical leadership, cross-team delivery, incident ownership, and platform/IaC ownership).
- Production experience with:
- Kubernetes
- AWS core services (IAM, compute, object storage, encryption/key management) and AWS cloud networking (VPC design, routing, security groups, load balancing, VPC endpoints/PrivateLink, Transit Gateway, DNS)
- Terraform or comparable infrastructure-as-code tools
- Git and CI/CD pipelines
- Linux and foundational systems concepts (networking, DNS, TLS/certificates)
- PostgreSQL (or comparable relational databases) in production - replication, backups, and performance tuning
- Programming and Automation: Proficiency in Python and/or Go experience to build automation scripts, operational tooling, and infrastructure services.
- Incident & Change Management: Experience troubleshooting production incidents, conducting root-cause analysis(RCA), and following change management processes.
- Experience participating in a production on-call rotation.
- Experience troubleshooting complex technical issues and writing clear documentation, runbooks and incident post-mortems.
Preferred Qualifications
- Experience operating in FedRAMP, AWS GovCloud, or other regulated or compliance-heavy cloud environments.
- Familiarity with security and compliance practices such as FIPS, vulnerability management, and controlled production change processes.
- Experience with observability and logging platforms in enterprise production environments.
- Deep operational expertise with PostgreSQL (HA/replication, tuning, backup and recovery); familiarity with data/platform technologies such as Redis and Kafka.
- Experience supporting federal agencies or public-sector customers.
- Exposure to government networking and security requirements.
- Experience with tools such as Jenkins, Argo CD, and GitHub Enterprise.
- Excellent collaboration skills and a strong willingness to learn.
Medallia is committed to equal pay and transparency. The annual base salary range for this position is $158,500 - $230,000. Please note that the salary range information provided is a general guideline and combines all of the distinct labor markets within the US. It is uncommon for an individual to be hired at or near the top of the range for their role and compensation decisions are dependent on a variety of factors. Medallia considers factors such as (but not limited to) scope and responsibilities of the position, candidate's work experience, candidate's work location, education/training, key skills, internal peer equity, external market data, as well as, market and business considerations when making compensation decisions.
Medallia also offers competitive health and wellness benefits, including but not limited to medical, dental, vision, 401(k), short-term and long-term disability, life and AD&D insurance, statutory leaves, paid parental leave, and paid holidays. Benefits and eligibility may vary by location and role.
At Medallia, we celebrate diversity and recognize the value it brings to our customers and employees. Medallia is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, age (40 and over), disability, genetic information, veteran status or military service, or any other status protected by state or local law. Individuals with a disability who need an accommodation to apply please contact us at View email address on click.appcast.io. For information regarding how Medallia collects and uses personal information, please review our Privacy Policies. Applications will be accepted for 30 days from the date this role was posted or until the role has been filled.
$230k - $250k
...GovCIO is hiring a Site Reliability Engineer with an active Secret clearance to ensure reliability, scalability, performance, and availability of mission-critical systems by combining software engineering practices with infrastructure operations expertise. This role is...SuggestedRemote work$120k - $150k
...operates along the way. In This Role, You Will: Own reliability across a complex, global product portfolio. You'll be... ...Daily use of Claude Code, Cursor, or AI coding assistants as an engineering accelerant - not occasional dabbling Experience...SuggestedWork at officeWorldwide$100k - $160k
...all bring to the team; and empower our employees to create innovative and trusted results. We are looking for a dynamic Site Reliability Engineer (SRE) with a Top Secret clearance to join our team! The Site Reliability Engineer (SRE) will manage, monitor, and...SuggestedTemporary work- ...Catalyst, and In-Q-Tel. Mission | On Site | Full Time | Active TS/SCI with Full... ...government customer site, ensuring the reliability and performance of Twenty's mission-critical... ...technical ownership and customer-facing engineering: you'll define how we measure...SuggestedFull timeContract workRemote workFlexible hours
$128.5k - $190k
...exceptional people to create extraordinary experiences together. Bring your whole self. The Role and Team The Site Reliability Engineering organization at Medallia brings together the infrastructure and applications that power a highly reliable global SaaS...SuggestedTemporary workWork experience placementLocal area$107k - $220k
...The Site Reliability Engineer (SRE) will ensure the reliability, performance, and scalability of the WDP System. This person will define and track Key Performance Indicators (KPIs) and Service Level Objectives (SLOs), identify and resolve performance bottlenecks, and perform...Full timeContract workTemporary workWork at officeVisa sponsorshipWork visa- ..., to act first. Exiger is FedRAMP authorized and a 2x Leader in Gartner Magic Quadrant for Supplier Risk Management. Site Reliability Engineer Location: U.S. (Hybrid) This role requires U.S. citizenship and eligibility for a U.S. security clearance. Role...Work at officeWork from homeFlexible hours
$106.3k - $221.1k
...Senior Site Reliability Engineer At Accenture Federal Services, nothing matters more than helping the US federal government make the nation stronger and safer and life better for people. Our 13,000+ people are united in a shared purpose to pursue the limitless potential...$106.3k - $221.1k
...more. Join us to drive positive, lasting change that moves missions and the government forward! Job Description The Site Reliability Engineer will ensure the reliability, performance, and scalability of the Client System. The engineer will define and track Key...Live inWork at officeLocal area$150k - $180k
...what’s possible in remote sensing, you belong here at Umbra. About the Job We are seeking an experienced Senior Site Reliability Engineer to help design, build, operate, and scale the mission- and business-critical infrastructure that powers Umbra's systems....Permanent employmentWork at officeLocal areaRemote workWorldwideFlexible hours$210k - $230k
...GovCIO is currently hiring for a Senior Site Reliability Engineer (SRE) to design, implement, and maintain highly available, scalable, and resilient infrastructure systems. The ideal candidate will bridge the gap between development and operations, focusing on automation...Currently hiringRemote work$158.5k - $230k
...extraordinary experiences together. Bring your whole self. The Role and Team We are growing our GovCloud team and looking for a Staff Site Reliability Engineer to help scale how we operate Medallia's US public-sector cloud platform. You will support federal agencies and other...Permanent employmentTemporary workWork experience placementWork at officeLocal areaRemote work3 days per week$90k - $130k
...critical, world-changing federal challenges. Credence has an immediate opening for a Site Reliability SME who has hands-on experience working as a Cloud Operations Engineer with experience in IT operations to join our expanding Cloud Managed Services Provider team...Temporary workWork experience placementImmediate startWorldwide$114.4k - $125.4k
...much more. Job Category Software Engineering Job Details About Salesforce Salesforce... ...! Are you passionate about ensuring the reliability and performance of mission-critical... ...services? Salesforce is seeking a talented Site Reliability Engineer to join our dynamic...Full timeLocal areaShift workNight shift$145k - $160k
...We are seeking a specialized Observability & Infrastructure Engineer to lead the monitoring, telemetry, and platform-as-code initiatives critical to our multi-region disaster recovery roadmap. You will architect and implement robust observability pipelines, ensure deep...Temporary workRemote workFlexible hours$160k - $210k
...change and achieving remarkable growth in a rapidly evolving industry. Now, we're growing! We are looking for a Senior Site Reliability Engineer to strengthen our AWS infrastructure and improve service management across Cognitiv. Our immediate challenge is to scale...Work at officeImmediate startRemote workWork from home- ...Job Description Job Description Description: Onsite in Washington, DC our client seeks a Sr. Site Reliability Engineer III to design, automate, and operate mission-critical systems for federal environments. The role focuses on Kubernetes or VMWare platforms,...Hourly payPermanent employmentFull timeLocal areaImmediate start
- ...Job Description Job Description Role Overview We are seeking a high-caliber Site Reliability Engineer (SRE) to join our Forward Engineering team. You will be the guardian of our production ecosystems, ensuring that our complex, data-driven AI platforms remain resilient...Local area
$81.1k - $187k
.... You'll partner with customer support, service owners, and engineering teams around the globe to ensure high-quality service for customers... .... Responsibilities Escalation points for junior Site Reliability Engineers during complex or high-impact incidents. Manage...Temporary workWork experience placementMonday to FridayFlexible hoursShift workNight shift- ...solutions using a tailored Agile methodology. We are seeking a highly motivated and intellectually curious Senior Site Reliability Engineer to join our team working with a Federal client. The position will be a remote role open to US citizens residing in the...Remote work
- ...Site Reliability Engineer (SRE) Dexian is seeking a savvy Site Reliability Engineer (SRE) who will play a key role in building a sustainable... ...resolve technical issues and escalations with other technical staff as the need arises. Work to automate the detection and...Work experience placement
$136.2k - $214.01k
...outcomes Visionary in future focused problem-solving Exceptional in execution and impact The Role As a Senior Site Reliability Engineer at Proofpoint you will develop a deep understanding of the various services and applications that come together to...Full timeFlexible hours- ...Site Reliability Engineer (SRE) Randstad is seeking a skilled and proactive Site Reliability Engineer (SRE) to join our client in the Washington D.C. area, focusing on optimizing the availability, performance, and scalability of critical production services. The ideal...
$75.7k - $136.3k
...solve complex challenges? Do you have a passion for automation and building systems that scale? Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages applications and infrastructure that support Akamai Cloud's products and...Work experience placementWork at office- ...Site Reliability Engineer Location- Wilmington De, Washington DC, Dallas, TX (Onsite Position) Full time position Minimum Qualifications Bachelor’s degree in computer science, Engineering, or a related technical field. Minimum of 5 years of experience...Full time
$90k - $150k
...Top Workplaces honoree, is seeking a SRE Engineer to support our growing team. The SRE... ...This role is responsible for improving the reliability, availability, performance,... ...Government customers. Work Environment: On-site Key Responsibilities: Define,...Permanent employmentFull timeContract work- ...Site Reliability Engineer (SRE) Reston, VA Site Reliability Engineer (SRE) Position: Site Reliability Engineer (SRE) Work Authorization: All Work Authorizations Location: Reston, VA Contract: 24 months Description: Site Reliability Engineer (SRE) roles...Contract work
$121.4k - $218.6k
...solve complex challenges? Do you have a passion for automation and building systems that scale? Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages applications and infrastructure that support Akamai Cloud's products and...Work experience placementWork at office$135k - $154k
...where you matter.Your ImpactAs a contributor in the APX platform engineering organization on the CloudNet team, you are passionate about... .... You are also obsessed about achieving the high quality and reliability our customers demand. You will work closely with sovereign...Work experience placementWork at officeRemote work$102k - $234.6k
...members in designing and architecting infrastructure and service for reliability and functionality. Provides day-to-day direction to help... ...to experiment with new technology, execute improvements, build site reliability knowledge, and provide clear data. Responsibilities...Temporary workImmediate startFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff Site Reliability Engineer. Be the first to apply!
- senior staff systems engineer McLean, VA
- engineering aide McLean, VA
- assistant engineer McLean, VA
- technology administrator McLean, VA
- staff engineer McLean, VA
- IT site lead McLean, VA
- site safety McLean, VA
- website content developer McLean, VA
- site leader McLean, VA
- on-site clinical research associate (traveling/remote) McLean, VA


