Senior Site Reliability Engineer
Precisely US Jobs
At Precisely, we're not just building software — we're shaping the future of data integrity. As a global leader in data quality, data enrichment, and location intelligence, Precisely helps thousands of the world's most trusted brands make confident decisions with data they can rely on. We're an AI-first organization, which means artificial intelligence isn't a buzzword here — it's woven into how we build products, how we work, and how we think about solving complex problems for our customers. When you join Precisely, you join a team of curious, driven innovators who believe that better data makes the world run better. If you're ready to do meaningful work at the intersection of AI and data — and help define what's possible — we'd love for you to apply!
This position is 100% remote within the U.S., but applicants must reside in the Mountain or Pacific Time Zones to be considered.Overview:
The Site Reliability Engineer (SRE) is responsible for the reliability, performance, and scalability of Precisely's infrastructure platforms across CEDAR (CCX) — an on-premises, private cloud managed services environment; RapidCX (RCX) — an AWS cloud environment for SaaS-delivered customer communications management; and Hosted Managed Services (HMS) — an AWS cloud environment supporting managed client deployments.
This role bridges software engineering and systems operations, building automation, observability tooling, and reliability standards to ensure platform availability and operational excellence. SREs are enabling partners: they set reliability standards, define what 'reliable' looks like for each service, validate production readiness, and coach engineering teams on operational best practices. Engineering teams own the reliability outcomes of the services they build; the SRE ensures they have the standards, tooling, and guidance to meet them.
As a Senior SRE, this role is regarded as a platform expert across CCX, RCX, and HMS, taking on the majority of complex reliability engineering work, mentoring less experienced engineers, and contributing to release triage and coordination efforts.
What you will do:
- Define and maintain reliability standards across CCX, RCX, and HMS, including SLOs, SLIs, and error budgets.
- Define standards for actionable alerting, meaningful logging, and diagnostic tracing; partner with engineering teams to ensure those standards are met.
- Serve as a platform expert and lead complex reliability engineering work across on-premises, private cloud, and AWS environments.
- Build and maintain infrastructure-as-code, deployment automation, monitoring, alerting, and observability tooling using Terraform, Ansible, Datadog, and scripting languages.
- Partner with engineering teams to embed reliability, operability, scalability, backup, recovery, and failure-mode considerations into service design and delivery.
- Lead Operational Readiness Reviews and validate production and disaster recovery readiness for qualifying changes.
- Contribute to release triage, deployment coordination, maintenance windows, and change and security reviews.
- Lead response to complex P1/P2 incidents, serve as incident commander when required, and communicate status clearly to technical and non-technical stakeholders.
- Author root cause analyses and post-incident reports, identify recurring patterns, and implement preventive automation to reduce manual toil and improve MTTR.
- Maintain operational runbooks, reliability backlogs, and team knowledge resources covering monitoring gaps, recurring incidents, and operational risks.
- Use Precisely-provided AI tools for automation, incident analysis, troubleshooting, runbook creation, solution testing, and architecture documentation.
- Ensure infrastructure configurations meet security, compliance, vulnerability-remediation, and data-protection requirements.
- Mentor Associate SREs and SREs, coach engineering teams on production diagnostics and operational practices, and contribute to cross-team design reviews.
- Lead project-based work, coordinate on-call coverage, and model strong operational discipline and follow-through.
- Participate in a rotating on-call schedule covering after-hours, weekends, and holidays, including critical production changes and escalations within defined SLA windows.
What we are looking for:
Required:
- Educational requirements (equivalent work experience will be accepted in place of the education requirement): Bachelor's degree in Computer Science, Information Systems, Engineering, or equivalent practical experience.
- 5+ years of systems or infrastructure engineering experience in an enterprise production environment.
- High-level, developing subject-matter-expert competency in one or more complex infrastructure domains (e.g., cloud platform engineering, IaC automation, observability, or on-premises virtualization).
- Advanced proficiency with Linux (RHEL/Oracle Linux) across multiple environments (on-premises and cloud, not just multi-site).
- Proficient with Terraform for infrastructure-as-code; experienced with Ansible role development beyond basic hands-on use.
- Experience deploying and managing workloads in AWS at an intermediate level: EC2, ECS, S3, VPC, IAM, CloudWatch, Auto Scaling.
- Proficiency with at least one scripting language (Python, Bash) for automation development.
- Demonstrated experience designing monitoring system architecture and alerting strategy — not just operating existing dashboards (Datadog preferred).
- Solid understanding of TCP/IP networking, DNS, load balancing, and distributed systems.
- Experience designing and managing CI/CD pipelines and deployment automation standards.
- Strong analytical skills; demonstrated experience authoring root cause analyses for complex incidents and identifying systemic, preventive fixes from recurring incident patterns.
- Ability to define SLOs and lead Operational Readiness Reviews (ORRs); comfortable partnering with engineering teams on production readiness.
- Demonstrated ability to work cross-functionally with engineering teams on reliability standards and observability requirements.
- Experience participating in change advisory processes and contributing to capacity and reliability planning.
- Travel is required: No — approximately 0%.
AI Skills/Knowledge:
Active, proficient use of Precisely-provided AI tools (GitHub Copilot, Claude, or equivalent) for complex automation, solution testing, and architecture documentation is a required baseline for this role — not a differentiator — and includes mentoring others on AI-assisted engineering practices.
- Apply AI tools for complex automation, incident analysis, and runbook authoring.
- Mentor other engineers on AI-assisted engineering practices.
- Maintain fluency with Precisely-approved AI coding assistants as a baseline expectation.
Preferred Skills (a plus but not required):
- Experience with containerization and orchestration (Docker, ECS, Kubernetes).
- Familiarity with GitOps workflows and source control best practices (Git, GitLab).
- Knowledge of enterprise virtualization platforms in a hybrid cloud context.
- Understanding of change management and ITIL operational practices.
- Experience with enterprise security tooling (Qualys, CrowdStrike, Rapid7).
- AWS Solutions Architect, SysOps Administrator, or DevOps Engineer certification.
- Wireshark and protocol analysis experience.
- Prior experience mentoring engineers or leading on-call rotations for a team.
#LI-KM1
Application and Interview Impersonation Notice: Impersonating another individual when applying for employment, and/or participating in an interview process to assist another individual in obtaining employment, with Precisely Software Incorporated (“Precisely”) is unlawful. If Precisely identifies such fraudulent conduct, then as applicable and to the extent permitted by law, the application will be rejected, an offer (if made) will be rescinded, or the employment will be terminated, and legal action may be taken against the impersonators.
The personal data that you provide as a part of this job application will be handled in accordance with relevant laws. For more information about how Precisely handles the personal data of job applicants, please see thePrecisely Candidate Privacy Notice
- ...Job Summary We are seeking a Senior SRE / DevSecOps Engineer with strong experience in Kubernetes, AWS... ...troubleshooting. The role will focus on platform reliability, incident management, SLO/SLI... ...with SLO/SLI governance and site reliability practices. ~ Strong understanding...SeniorContract work
$174k - $252k
...systems by pushing for changes that improve reliability and velocity.Practice sustainable... ...:Bachelor’s degree in Computer Science, Engineering, a related field, or equivalent practical... ...degree in Computer Science or Engineering.Site Reliability Engineering (SRE) is what you...Senior$160k - $240k
...millions of times a day - quickly, reliably, and securely. Any time you... ...at Fiserv.Job TitleSenior Site Reliability EngineerWhat does a successful Site Reliability Engineer do at Fiserv?You will join our... ...operations or DevOps at a mid-to-senior level.Strong shell scripting...SeniorFull time$152.5k - $205k
...flexible work environment where new ideas are encouraged and everyone is a stakeholder.What you’ll be responsible for:As a Senior Site Reliability Engineer on Circle’s platform team, you’ll design, build, and operate the secure, scalable platform infrastructure behind...SeniorFlexible hours- ...importance of in-office collaboration and fully intend for the selected candidate for this role to work on site in the specified location(s).As a Site Reliability Engineer supporting the Cashiering organization, you will play a critical role in ensuring the stability,...SeniorFull timeWork at office
$80k - $140k
Job DescriptionRBC Wealth Management Technology is seeking a Senior Site Reliability Engineer to join its Wealth Management SRE Team. This team is responsible for ensuring the performance, availability, resilience, and operational excellence of critical applications and...SeniorFull timeFlexible hoursShift work$106.7k - $177.9k
OverviewJoin M&T Bank's Digital Banking organization and help drive the reliability, resiliency, and performance of the platforms our customers depend on every day. As a Senior Software Engineer, Site Reliability Engineering (SRE), you will play a key role in supporting...SeniorPermanent employmentFull timeWork experience placement- ...US Corp. is seeking a Lead Site Reliability Engineer to spearhead our mission of delivering highly available and performant systems. With an average of over 12 years of industry experience, the successful candidate will bridge the gap between software development and systems...Senior
$104.9k - $174.7k
...SRE role is responsible for improving the reliability, availability, performance, and... ...actions through completion.Follow up with engineering, development, security, support, and business... ...Qualifications5+ years of experience in Site Reliability Engineering, Systems Engineering...SeniorFull timeLocal area- The Senior Site Reliability Engineer is responsible for improving the reliability, availability, scalability, and operational excellence of our critical infrastructure platforms and services. This role partners closely with Engineering, Security, and Infrastructure teams...SeniorFull timeWork at officeLocal area
- ...Senior Site Reliability Engineer Company: CyberArk Work Type: Remote Employment: Full Time Location: US Seniority: Mid Level Technologies: AWS, Kubernetes, Terraform, CloudFormation, Ansible, CloudWatch, Grafana, Datadog, OpenSearch, PagerDuty Requirements: Senior SRE...SeniorFull timeRemote work
$95.3k - $158.8k
Senior Site Reliability Engineer Are you passionate about building resilient, scalable systems that power mission-critical applications?Do you thrive on automating operations, improving reliability, and ensuring exceptional system performance?About the team:Embedded Innovation...SeniorFull timeLocal areaRemote workWork from home- ...The Role:GIPHY is seeking a highly experienced Site Reliability Engineer to join our SRE team. You will help design, build, operate, and evolve the infrastructure that powers GIPHY, including our cloud environment, Kubernetes clusters, and CI/CD platforms.You will also...SeniorFull timeWork experience placementRemote work
- ...principles to see it in full.About the teamThe Engineering team at Airwallex is a diverse group of... ..., working together to build scalable, reliable, and secure products that empower... ...our Global services.What you’ll doAs a Senior Site Reliability Engineer, you’ll work...SeniorTemporary workLocal area
$134k - $170k
...independent system operator responsible for ensuring the safe and reliable flow of electricity in our region and planning for the... ...of New England’s ongoing transition to clean energy. The Senior Site Reliability Engineer (SRE) is a hands-on engineering role responsible for...SeniorPermanent employmentH1bRelocationVisa sponsorshipWork visaRelocation packageFlexible hours3 days per week- ...As a Senior Site Reliability Engineer, you will help design, deploy, maintain, and improve reliable, secure, and scalable infrastructure and services. You’ll proactively identify operational risks and potential failure points, troubleshoot system and application issues...Senior
- Site Reliability Engineers are responsible for ensuring the availability, reliability, scalability, and performance of the firm’s most critical customer-facing microservices that power all eCommerce channels. This role applies Google-inspired SRE principles to balance...SeniorLocal areaRemote workFlexible hoursShift work
$108k - $216k
...PermanentCompany: VizioBusiness Segment: Home OfficePosition: Senior Site Reliability EngineerJob Location: 39 Tesla, Irvine, CA 92618Duties:... ...to customer support experiences. Collaborate with engineering teams to embed reliability into the software development lifecycle...SeniorFull timeTemporary work- ...in Ausin, TX** Our Opportunity: We are looking for a skilled engineer with disciplines that incorporate aspects of software systems... ...applications — including AI/ML-driven approaches to observability and reliability. What you’ll do: • Evangelize SRE mindset and solve problems...Senior
- ...Senior Site Reliability Engineer Company: Sphera Work Type: Remote Employment: Full Time Location: US Seniority: Senior Level Technologies: Terraform, ARM templates, Kubernetes, Azure, SonarCloud, CheckPoint, Hadoop, Kafka, Presto, NewRelic, CI/CD, Linux, Windows, Redis...SeniorFull timeRemote work
- ...Discover exciting DevOps job opportunities and connect with 28,396 DevOps professionals. The Senior Site Reliability Engineer role at Jobicy is designed for experienced professionals who are passionate about enhancing system reliability and operational efficiency....SeniorRemote workFlexible hours
- ...developer-tooling company whose product is used by engineering teams at thousands of software companies for... ...commitments and the SRE team is a senior, well-resourced group of nine. As Senior SRE you will lead reliability initiatives across the platform — from defining...Senior
$139.3k - $203.6k
...FedRAMP team builds and operates secure, reliable cloud services for U.S. government... ...customers. We partner closely with application engineering, security, compliance, and... ...insurance. Please see the Cisco careers site to discover more benefits and perks. Employees...SeniorFull timeTemporary workWork at officeLocal areaFlexible hours- ...Job Title Location Remote - United States Job Category Information Technology, Platform Engineering, Site Reliability Engineering Industry Computer Software, SaaS, National Security Employee Type FT Exempt Manage Others No Minimum Experience 5 Years...SeniorRemote work
- ...infra has to match. The role We're looking for a Senior SRE to own the reliability, scalability, and operational posture of Satsuma's multi... ...-assisted development workflows Partner closely with engineering on reliability reviews and architecture decisions...Senior
$81.1k - $187k
.... You’ll partner with customer support, service owners, and engineering teams around the globe to ensure high-quality service for customers... ...posted.Career Level - IC3Escalation points for junior Site Reliability Engineers during complex or high-impact incidents.Manage and...SeniorTemporary workMonday to FridayFlexible hoursShift workNight shift$108k - $216k
...Senior Site Reliability Engineer Senior Site Reliability Engineer professional opening available at Wal-Mart in Irvine, CA. Qualifications Master's or equiv in CS, Comp Eng'g, Comp Info Systs, SW Eng'g, Electrical Eng'g, Info Systs Security, or rel....SeniorTemporary work- ...Senior Site Reliability Engineer We are looking for a Senior Site Reliability Engineer with Cloud platform experience. This individual will be part of a team responsible for operating and maintaining production clusters and developing our observability solutions; they...SeniorRemote work
$262k - $364k
...infrastructure from SRE side, ensuring it is reliable, scalable, cost effective and performant, while working closely with senior technical leads in the development teams.... ...:Master's degree in Computer Science or Engineering.Site Reliability Engineering (SRE) combines software...Senior- ...We need a Senior SRE to ensure Xident's verification platform runs at 99.99% uptime.... ...SLOs, incident response, and production reliability for a system that processes millions of... ...structured logging Implement chaos engineering practices to proactively identify failure...SeniorRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- site reliability engineering manager United States
- site reliability engineer United States
- lead site reliability engineer United States
- site reliability engineer remote United States
- site reliability engineer sre United States
- senior living director United States
- heritage senior living United States
- senior computer engineer United States
- senior manager customer operations United States
- senior support engineer United States



