Site Reliability Engineer II
$103.5k - $150kMedallia
Medallia is the pioneer and market leader in Experience Management. Our award-winning SaaS platform, Medallia Experience Cloud, leads the market in the management of experiences, insights, and actions for candidates, customers, employees, patients, and residents alike.
We believe that every experience is a memory that can last a lifetime. Experiences shape the way people feel about a company. And they greatly influence how likely people are to advocate, contribute, and stay. At Medallia, we are committed to creating a world where organizations are loved by their customers and their employees.
We empower exceptional people to create extraordinary experiences together.
Bring your whole self.
The Role and Team
The Site Reliability Engineering organization at Medallia brings together the infrastructure and applications that power a highly reliable global SaaS platform.
As an SRE II, you will help operate and improve the reliability, scalability, and performance of services running across Kubernetes-based environments in cloud and hybrid infrastructure. You will work closely with software engineering teams to build automation, improve operational excellence, and support production services used globally by Medallia customers.
We are looking for engineers who enjoy solving complex technical problems, automating repetitive tasks, improving system reliability, and learning modern cloud-native technologies in a fast-paced environment.
We value engineers who actively seek opportunities to improve scalability and operational efficiency through automation, AI-assisted engineering workflows, and continuous process improvement.
Please note this role participates in a rotating on-call schedule supporting production systems and services.
Engineering Leverage
At Medallia, we hire engineers who scale systems, teams, and outcomes through automation, platform thinking, and AI-assisted engineering.
We value engineers who challenge manual processes, reduce operational toil, and create reusable solutions that improve reliability and productivity for the broader engineering organization.
Successful engineers do not simply solve problems-they eliminate recurring problems through automation, simplification, and self-service capabilities.
Responsibilities- Collaborate with software engineering teams to improve application reliability, scalability, and operational maturity.
- Operate and support production services running in Kubernetes environments.
- Troubleshoot and resolve infrastructure and application issues across the full technology stack.
- Build automation and tooling to reduce operational overhead and eliminate manual work.
- Leverage AI-assisted engineering tools and automation platforms to accelerate troubleshooting, improve productivity, and reduce operational toil.
- Identify opportunities to streamline operational processes through automation, AI-enabled workflows, and self-service solutions.
- Create reusable solutions, tooling, and operational improvements that increase engineering leverage across the team.
- Support CI/CD and GitOps-based deployment workflows.
- Develop and maintain infrastructure-as-code configurations and operational tooling.
- Monitor system health, availability, and performance using observability and alerting platforms.
- Participate in incident response, root cause analysis, and operational improvements.
- Continuously improve reliability, deployment processes, and operational standards.
Candidates based in the Tysons vicinity will be prioritized as this role is Hybrid, 3 days per week onsite.
QualificationsMinimum Qualifications
- 2+ years of experience in Site Reliability Engineering, DevOps, Systems Engineering, Cloud Operations, or related roles.
- Demonstrated experience supporting production environments running on Kubernetes or other containerized platforms.
- Demonstrated experience with cloud infrastructure platforms such as AWS, OCI, or GCP.
- Demonstrated experience with Linux systems administration and troubleshooting.
- Demonstrated experience with scripting or programming languages such as Python, Bash, or Go.
- Familiarity with CI/CD pipelines and Git-based workflows.
- Demonstrated understanding of networking fundamentals including DNS, load balancing, TLS/SSL, and routing concepts.
- Demonstrated experience troubleshooting distributed systems and production incidents.
- Ability to participate in an on-call rotation supporting production systems.
- Fluency in English, both oral and written.
Preferred Qualifications
- Experience with GitOps and tools such as ArgoCD.
- Experience with infrastructure-as-code tools such as Terraform.
- Familiarity with observability platforms such as Prometheus, Grafana, Loki, or OpenTelemetry.
- Experience operating services in hybrid-cloud or multi-region environments.
- Understanding of release strategies such as rolling deployments, canary releases, or blue/green deployments.
- Familiarity with incident management and operational best practices.
- Exposure to security and compliance concepts in production environments.
- Experience using AI-assisted development, automation, or operational tooling to improve engineering productivity and service reliability.
- Demonstrated passion for automation, process improvement, and operational efficiency.
- Strong communication and collaboration skills.
Medallia is committed to equal pay and transparency. The annual base salary range for this position is $103,500 - $150,000. Please note that the salary range information provided is a general guideline and combines all of the distinct labor markets within the US. It is uncommon for an individual to be hired at or near the top of the range for their role and compensation decisions are dependent on a variety of factors. Medallia considers factors such as (but not limited to) scope and responsibilities of the position, candidate's work experience, candidate's work location, education/training, key skills, internal peer equity, external market data, as well as, market and business considerations when making compensation decisions.
Medallia also offers competitive health and wellness benefits, including but not limited to medical, dental, vision, 401(k), short-term and long-term disability, life and AD&D insurance, statutory leaves, paid parental leave, and paid holidays. Benefits and eligibility may vary by location and role.
At Medallia, we celebrate diversity and recognize the value it brings to our customers and employees. Medallia is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, age (40 and over), disability, genetic information, veteran status or military service, or any other status protected by state or local law. Individuals with a disability who need an accommodation to apply please contact us at View email address on click.appcast.io. For information regarding how Medallia collects and uses personal information, please review our Privacy Policies. Applications will be accepted for 30 days from the date this role was posted or until the role has been filled.
- ...Site Reliability Engineer II Join the leader in providing smarter solutions for a safer world. The property technology space is growing rapidly, and Kastle Systems is leading the way. Kastle Systems is the leader in managed security, with a track record of introducing...SuggestedRemote work
$103.5k - $150k
...experiences together. Bring your whole self. The Role and Team The Site Reliability Engineering organization at Medallia brings together the infrastructure... ...power a highly reliable global SaaS platform. As an SRE II, you will help operate and improve the reliability,...SuggestedTemporary workWork experience placementLocal area3 days per week$95k - $171k
...Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building and maintaining dashboards, alerts,...SuggestedPermanent employmentWork experience placementWork at officeRemote workWork from homeWorldwideFlexible hours$115.5k - $164.8k
...mission that matters at a company where you matter.Your ImpactAs an engineer on the APX SRE CloudOps team, you will spend a significant... ...that replace what previously required human intervention with reliable, tested automation. You will also participate in on-call rotations...SuggestedWork experience placementWork at officeRemote work$97.24k - $162.24k
...that power mission critical software.As a Software Development Engineer II focused on CI/CD and platform automation, you will design and... ...implement tools, pipelines, and automation frameworks that enable reliable, secure, and scalable software delivery across enterprise...SuggestedRelocationRelocation package$129.2k - $174.8k
...Do you have a deep passion and desire to engineer and operate the world's largest cloud computing... ...seeking a Systems Development Engineer II who can thing big and simplify solutions... ...-scale, distributed environments where reliability and performance are critical. You...InternshipFlexible hours$97.24k - $162.24k
...Defense and Intelligence operations.As a Software Development Engineer II, you will design and develop scalable backend services and APIs... .... You will collaborate with cross functional teams to build reliable, high performance systems that operate across secure, distributed...RelocationRelocation package$143.7k - $194.4k
Join our team as a Software Development Engineer (SDE II), where you'll build and enhance management console experiences that help AWS customers... ...of non-internship design or architecture (design patterns, reliability and scaling) of new and existing systems experience- 1+...InternshipLive inFlexible hours- ...bottlenecks, and improve system health—utilization, performance, and reliability—across our infrastructure Understand how systems fail and... ..., documentation, and code review—that automates reliability engineering work: Deployment tooling Fault-injection/chaos...Full timeCasual workLocal areaWorldwideFlexible hoursShift work
$128k - $252.5k
...scientists, operators, creatives, designers, engineers, and architects. Our team balances... ...Do As a Lead Agentic Software Engineer II, you will be responsible for:Lead the design... ...in software engineering, DevOps, site reliability engineering, quality engineering, or software...Local area- ...Software Engineer IIIt's a great time to join us at Airlines Reporting Corporation (ARC)! ARC accelerates the growth of global air travel... ...for the travel industry.We are looking for a Software Engineer II to join our team! In this role, you will provide product teams with...Work at officeWork from homeFlexible hours
$106k - $169k
...Software Engineer IIMastercard's Portfolio Intelligence team, part of the Services organization, is seeking a Software Engineer II to help build high-performance analytics platform that enable... ...roles; fitness reimbursement or on-site fitness facilities; eligibility for...Full timePart timeFlexible hours3 days per week$102.5k - $170.9k
Position Summary Our AI & Engineering team helps transform technology platforms, drive innovation, and deliver meaningful impact for... ...growth through innovation.Work you'll doAs a Software Engineer II on the Engineering as a Service team, you will be responsible...Temporary workLocal areaFlexible hours$106k - $169k
...Software Engineer II - Backend/Platform Agentic AIMastercard is a global technology company... ...ensuring correctness, performance, and reliability in a multi-tenant distributed environmentImplement... ...roles; fitness reimbursement or on-site fitness facilities; eligibility for...Full timePart timeWorldwideFlexible hours$92.5k - $146.3k
Platform - Engineering Productivity - Software Engineer II Elastic Cloud 21 July 2025 Elastic, the Search AI Company, enables everyone to find the answers... ...efficient and independent by providing realistic and reliable developer environments. We’re continuously evolving...Local areaWorldwideFlexible hours- A technology solutions provider is seeking an experienced individual for the position of Infrastructure Management - Level II in Arlington, VA. The role involves designing, configuring, and maintaining servers, as well as troubleshooting complex infrastructure issues....
$88.3k - $122.55k
...Job Description Job Description Senior Software Engineer As a Software Engineer II you will work with a team of full-stack developers that work on all server-side aspects of smart home security. Our mandate is very broad and includes but not limited to processing...Work experience placementCasual workWork at officeImmediate start- ...Job Description Job Description Overview Software Engineer II (Full-Stack) Location: Vienna, VA Job Type: Full-Time... ...including EC2, S3, RDS, and ECS. Monitor system performance and reliability using tools such as Prometheus and Grafana. Support...Full time
- ...Job Description Job Description Job Title: Software Engineer II Department: Software Development Reports To: SW Team Lead Location: Vienna, Virginia FLSA Status: Exempt Employment Type: Full-time Experience Level: Mid-level (3 years) Job Summary...Full time
$150k - $175k
As Sr. Cloud FinOps Engineer II, you’ll {main responsibility/task} with the goal to make an impact across the federal government. Our... ...lifecycle management while maintaining mission performance and reliability.Cloud Governance: Develop and enforce Azure tagging standards...Full timeLocal area$150k - $180k
...Umbra.About the JobWe are seeking an experienced SeniorSite Reliability Engineer to help design, build, operate, and scale the mission- and business... ...impact across the organization.This position is based on-site in either our Arlington, VA office, Reston, VA office or...Permanent employmentFull timeWork at officeLocal areaRemote workWorldwide$133k - $190k
Site Reliability Engineer needed for a full time opportunity with SOC's direct client based in Herndon, VA. Direct Hire Role **Due to federal requirements, candidates must hold and possess an Active DOW TS/SCI security clearance to be considered for this role.** SOC is...Full time$210k - $230k
GovCIO is currently hiring for a Senior Site Reliability Engineer (SRE) to design, implement, and maintain highly available, scalable, and resilient infrastructure systems. The ideal candidate will bridge the gap between development and operations, focusing on automation...Currently hiringRemote work$230k - $250k
GovCIO is hiring a Site Reliability Engineer with an active Secret clearance to ensure reliability, scalability, performance, and availability of mission-critical systems by combining software engineering practices with infrastructure operations expertise. This role is...Remote work$80k - $133k
...degree, Four (4) years additional experience will be needed.Minimum Four (4) years of experience in IT administration, software engineering, or platform engineering, with a focus on AWS cloud infrastructure and enterprise systems.One(1)+ years of experience deploying and...Permanent employmentFull timeContract workRemote workFlexible hours$165k - $230k
..., with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARSHIELD)Starshield leverages SpaceX’s Starlink technology... ..., applicant must be a (i) U.S. citizen or national, (ii) U.S. lawful, permanent resident (aka green card holder), (iii...Permanent employmentTemporary workImmediate startWeekend work- ...Title: Applications Developer II Description: Solutions 3 LLC is supporting a U.S. Government Prime Contractor and its customer... ...Required Education: Bachelor of Science in Computer Science, Engineering, or related field, or High School diploma and 4-6+ years of...Full timeFor contractorsRemote work
$129.2k - $174.8k
...Connectivity is seeking a Systems Development Engineer II to design and implement systems that ensure operational excellence, reliability, and scalability of mission-critical... ...systems integration, DevOps practices, and site reliability engineering will directly impact...Full timeInternshipWork at officeFlexible hours- ...Detail Description: The AWS Site Reliability Engineer (SRE) is responsible for the operational health, availability, and performance of the AWS and Databricks environments built by the Platform Engineering team. You prepare and take ownership of "day two" operations...
$81.1k - $187k
...Infrastructure Engineer Takes proactive steps to design and architect infrastructure and service to ensure reliability and functionality. Forecasts demands and responds to capacity... ...potential impact and develops knowledge of site reliability trends. Key...Temporary workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer II. Be the first to apply!
- junior website developer McLean, VA
- website content developer McLean, VA
- site leader McLean, VA
- site recruiter McLean, VA
- historic site McLean, VA
- on-site clinical research associate (traveling/remote) McLean, VA
- official site McLean, VA
- site safety McLean, VA
- construction site safety McLean, VA
- IT site lead McLean, VA


