Site Reliability Engineer
$145k - $160kGrabJobs
About Payliance Founded in 2007, Payliance is a trusted leader in payment processing — processing more than $63 billion annually, supporting 40,000+ merchant locations, and serving over 350 lending clients. We offer an all-in-one platform for real-time funding, payment processing, account verification, and recovery services, giving lenders the technology to operate efficiently and confidently. What sets Payliance apart is our blend of modern technology, deep industry expertise, and a highly collaborative, people-first culture. Backed by Serent Capital, we're expanding our capabilities and delivering measurable results for clients across lending, e-commerce, collections, and gaming. About the Role The Site Reliability Engineer (SRE) bridges software engineering and infrastructure operations, owning the reliability, scalability, and performance of Payliance's payment processing platform. You'll apply engineering discipline to operational problems — reducing toil, automating repetitive work, and building systems that are resilient by design. This role is ideal for an engineer who can read and debug production .NET code, has hands-on AWS compute and RDS SQL Server depth, and brings the observability mindset needed to keep a high-transaction payment platform running at scale. What You'll Do .NET Application Reliability · Read, debug, and contribute to production C#/.NET code to diagnose and fix app-level reliability issues. · Identify and resolve memory leaks, thread pool exhaustion, and GC pressure before they manifest as incidents. · Partner with application engineers to embed reliability into new feature design and deployment practices. · Instrument .NET services with distributed tracing and structured logging to surface runtime anomalies early. AWS Compute & Infrastructure · Operate and optimize EC2 Auto Scaling, ECS Fargate, and Lambda workloads — with clear judgment on when each is the right fit. · Build and maintain infrastructure-as-code using CloudFormation or CDK for consistent, reproducible environments. · Automate operational tasks, deployment pipelines, and disaster recovery procedures. · Continuously reduce toil through tooling and automation, freeing the team for higher-impact engineering work. RDS SQL Server Operations · Manage RDS SQL Server deployments including Multi-AZ failover configuration and read replica setup. · Operate backup and point-in-time recovery (PITR) processes and validate restore procedures regularly. · Diagnose and resolve performance issues: slow queries, missing indexes, and blocking chains. · Capacity plan and scale database infrastructure to support transaction volume growth. Observability & Monitoring · Build and maintain observability stacks using CloudWatch metrics, log insights, and alarms; AWS X-Ray for distributed tracing. · Own service health dashboards, SLOs/SLIs, and drive data-driven reliability improvements. · Design alerts that surface signal — not noise — and ensure on-call responders have the context to act quickly. · Conduct root cause analysis (RCA) on incidents and lead blameless post-mortems to capture lessons and prevent recurrence. Networking & Security · Design and maintain secure AWS network topologies: VPCs, subnets, security groups, and NACLs. · Configure and manage ALB/NLB routing, Route 53 DNS, and TLS certificate lifecycle via ACM. · Author and review least-privilege IAM policies; audit roles and resource-based policies for over-permissioning. · Support compliance and security controls relevant to a PCI-regulated payments environment. Incident Response & On-Call · Participate in on-call rotation to respond to production incidents and drive swift resolution. · Define and track error budgets; use them to balance velocity and reliability investment. · Communicate status updates clearly during incidents and coordinate cross-functional response. · Maintain and improve runbooks, escalation paths, and on-call health over time. Cross-Functional Partnership · Collaborate with platform engineering teams on architecture decisions and scalability requirements. · Share observability and reliability best practices with application teams. · Mentor engineers on SRE principles and operational excellence. Compensation & Benefits · Base Salary: $145,000 - $160,000 · Performance-based annual bonus. · Medical, Dental, and Vision insurance. · 401(k) with company match. · Generous PTO plus paid company holidays. · Company-paid life and long-term disability insurance. · Paid parental leave. Work Environment Remote-first with collaboration across U.S. time zones. On-call duties are part of this role with rotation schedules designed to be sustainable and fairly distributed. Equal Employment Opportunity Payliance is an equal opportunity employer. We value diversity and strive to create an inclusive workplace for everyone. Discrimination or harassment of any kind — based on race, color, sex, religion, sexual orientation, gender identity, national origin, age, disability, genetic information, or pregnancy — is not tolerated. Reasonable accommodations are available throughout the application and employment process. Requirements What You'll Bring Required Qualifications · 4+ years in SRE, DevOps, platform engineering, or a systems-focused software engineering role. · C#/.NET engineering ability — can read, debug, and contribute to production code; experience diagnosing memory leaks, thread exhaustion, and GC pressure. · AWS compute fluency: hands-on depth across EC2 Auto Scaling, ECS Fargate, and Lambda, with informed opinions on when to use each. · RDS SQL Server operational experience: Multi-AZ failover, read replicas, backup/PITR, slow query analysis, and blocking chain resolution. · Native AWS observability proficiency: CloudWatch (metrics, logs, alarms), X-Ray, and infrastructure-as-code via CloudFormation or CDK. · AWS networking and security competence: VPCs, security groups, ALB/NLB, Route 53, TLS/ACM, and least-privilege IAM. · SLO discipline: experience defining SLIs/SLOs against real metrics, running blameless postmortems, and carrying an on-call pager. · Strong scripting ability (PowerShell, Python, or Bash) for automation and operational tooling. · Excellent communication skills and a collaborative, blameless engineering mindset. · Genuine openness to adopting AI tools and a willingness to experiment with new technology to work smarter and faster. Preferred Qualifications · Experience in fintech, payments, or high-transaction-volume regulated environments (PCI-DSS, SOC 2). · Knowledge of ACH, card processing, or payment settlement workflows. · Familiarity with chaos engineering or resilience testing (e.g., AWS Fault Injection Simulator). · Experience with secrets management (AWS Secrets Manager, Parameter Store) and security scanning in CI/CD. · Exposure to GitOps workflows and modern CI/CD practices. · Experience working in a private equity-owned, venture-backed, or high-growth startup environment. · Demonstrated ability to leverage AI tools to accelerate diagnostics, automate runbook creation, or improve observability workflows. · Bachelor's degree in Computer Science, Engineering, or related field (or equivalent hands-on experience). How Your Success Will Be Measured · Platform Uptime & SLO Attainment: Meeting or exceeding agreed service level objectives. · MTTD & MTTR: Mean time to detection and resolution for production incidents. · Toil Reduction: Measurable decrease in manual operational tasks through automation. · Application Reliability: Reduction in .NET runtime issues (memory leaks, thread exhaustion) reaching production. · Deployment Frequency & Stability: Supporting rapid iteration with minimal risk. · Observability Coverage: Expansion of CloudWatch/X-Ray coverage and depth of incident insights. · On-Call Health: Sustainable rotation with clear runbooks and blameless post-mortems. · Team Impact: Mentorship and knowledge sharing that raises operational maturity across engineering.
- ...The Team Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational... ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper). As...SuggestedWork at officeLocal areaRemote workWorldwide
- ...Partner with software developers, platform engineers, and IT staff to improve system design,... ...requirements, service quality, reliability, security, and compliance needs. Drive... ...Required: ~8+ years of experience in Site Reliability Engineering, DevOps, Platform...SuggestedWork at officeRemote work
- ...Senior Site Reliability Engineer (Enterprise Platform) Location: Remote - US - Open to Europe if happy to overlap with EST Compensation: Competitive We are a high-growth software company supporting the development of a premier open-source, EVM-compatible public ledger...SuggestedContract workCurrently hiringRemote work
- ...Senior Principal Software Engineer, Ground System Software Architect Position Title: Sr. Principal Software Engineer, Ground System Software Architect Work Location: Albuquerque, NM (preferred) or Herndon, VA Clearance Requirement: Active Top-Secret security...SuggestedPermanent employmentFull timeContract workFlexible hours
- ...We hire smart Scientists and Software Engineers who love to create and maintain high quality, extensible code , and want to learn and... ...investigation, and perform some work at government and/or customer sites Desired: Advanced degree (M.S. or Ph.D.) in science,...SuggestedFull time
- ...Space Systems Integration (SSI) is a fast-growing engineering company that provides aerospace solutions to a variety of government and commercial... ..., and risk management. This position is full-time, on site at Kirtland AFB, NM and will require travel up to 25% of time...Full timeFor contractorsWork experience placementWork at officeLocal area
- ...founding in 1986, MILVETS Systems Technology, Inc. has been a reliable provider of quality services in the information and technology... ...Position Summary: MILVETS is currently seeking a full-time Software Engineer Mid-Level with an active SECRET clearance and CompTIASec+...Full timeImmediate start
- The DCS Air & Space Technology (AST) Sector is seeking a Modeling and Simulation Software Engineer to support extensive high visibility Modeling, Simulation, and Analysis (MS&A) efforts. Are you interested in working in a high-tech company on cutting edge technology to...Full time
- OverviewCarollo Engineers is a leading engineering firm dedicated exclusively to water. For over 90 years, we've specialized in the planning... ...entry and analysis, participate in field activities such as site investigations, and pilot testingQualificationsBachelor's...Full timeFlexible hours
- ...AI infrastructure, working with server, cloud, and platform engineering teams. Operationalize machine learning workflows and support... ...system enhancements to improve performance, scalability, reliability, and cost efficiency. Collaborate across divisions to support...Work at officeRemote work
- ...Senior Platform Engineer Radiance Technologies is an employee-owned company with benefits that are unmatched by most companies. Employee... ...tasks where possible. Ensure Security, Scalability, and Reliability: Proactively ensure the platform's security, scalability, and...Work experience placement
- ...Join to apply for the Technical Solutions Engineer role at Epic 2 weeks ago Be among the first 25 applicants Join to apply for the Technical Solutions Engineer role at Epic Get AI-powered advice on this job and more exclusive features. Please note that this position is...Full timeWork at officeRelocationVisa sponsorshipRelocation package
$84k - $126k
...Our Mission As the world’s number 1 job site*, our mission is to help people get jobs. We strive to cultivate an inclusive and accessible... ...delivery of Indeed’s recruiting solutions. As a Sr Solutions Engineer, you will be a client-facing technical expert working directly...Work experience placementLocal area- ...something you have written for work, for a class, or for fun. It should be long enough to help evaluate your programming and software engineering skills. Extremely flexible work schedule & generous benefits. US Citizenship required + willingness to undergo a background...Full timeSummer workRemote workFlexible hours
- ...Senior Software Engineer Location: Kirtland Air Force Base, Albuquerque, New Mexico, United States Job Tags: Software About The Role BlueHalo, an AV company, is seeking a skilled and driven Senior Software Engineer to join our team in supporting cutting-edge defense applications...
- ...the curve and compliant with the latest USPS® regulations. Job Description BCC Software is seeking an experienced Sr. Software Engineer with expertise in Delphi to join an Agile development team. The role involves contributing to both existing applications and new software...Full timeWork at officeRemote workMonday to FridayAfternoon shift
$98.4k - $147.6k
...they're making history. Northrop Grumman Defense Systems (NGDS), in Albuquerque, New Mexico, is seeking Full Stack Software Engineers to develop a cloud-based learning management ecosystem, consisting of multiple application modules. The system supports critical...Full timeRelocation packageShift work- ..., Mexico, Ciudad de Mexico, Ciudad de Mexico Senior Software Engineer - Full Stack Do you love building and pioneering in the technology... ..., educational tools or other information available through this site. Capital One Financial is made up of several different...InternshipLocal area
- ...boldest and most ambitious space missions. SENIOR SOFTWARE ENGINEER I – MES Based at Rocket Lab's site in Albuquerque, NM, the Senior Software Engineer I -... ..., UI/UX design, product ownership, DevOps, or Site Reliability Engineering (SRE). Develop world-class software using...Permanent employmentLocal areaImmediate start
- ...stability — on a system that demands high concurrency, high transaction volumes, and exceptional reliability. You will work cross-functionally with product, cloud, and engineering teams to deliver high- performance solutions that improve patient outcomes and modernize...Remote workFlexible hoursShift work
$114.5k - $148k
...health and take back their lives. Join us on our mission to reverse metabolic disease in one billion people. As a Backend Engineer, you’ll build reliable systems, improve data accuracy across member, client, and operational workflows, and automate processes that are...Currently hiringWork at officeRemote work$141k - $180k
...We're looking for talented and highly motivated software engineers to join our team. As a growth stage technology company, Clear Capital is seeking developers eager to be part of a team, working closely with every part of the company from operations to our executive team...Temporary workImmediate startRemote work- This position requires a proactive individual with strong programming skills and experience implementing successful applications. Additionally, the ideal candidate will also have strong communication and relationship capabilities. Responsibilities Plans, designs, develops...
- ...future is still being invented — and we want to be the ones building it.For more information, visit What we are building Our Sales Engineering team is at the heart of transforming curiosity into confidence. We help prospects envision the full potential of our platform,...For contractorsRemote work
$75k
...USA, Canada JobType: full-time This role involves being the internal voice of the customer, partnering with Sales, Product, and Engineering to solve complex customer problems. The ideal candidate will bring genuine finance domain credibility, capable of engaging in thoughtful...Full timeWork experience placementWeekday work$105k - $140k
...Inclusivity. Today, nearly 200 people around the globe work on Speechify in a 100% distributed setting. These include frontend and backend engineers, AI research scientists, and others from Amazon, Microsoft, and Google, leading PhD programs like Stanford, high growth startups...$80 - $100 per hour
...evaluation pipelines used to test frontier AI models on real software engineering work: Design coding benchmarks that evaluate frontier... ...workflows Analyze model-generated code for correctness, reliability, and edge-case failures Construct structured evaluation...Full timeContract workFor contractorsRemote work$169k - $208k
...fusion of former U.S. Special Operations cyber operators, startup engineers, and formerly frustrated cybersecurity practitioners. We're... ..., and evaluation that lets an LLM-driven agent do this work reliably, at scale, and unattended. You'd own and evolve the attack-agent...Full timeWork at officeRemote workFlexible hours$120k - $210k
...About the role: We’re looking for experienced, problem-solving engineers across the stack (Product, Full-Stack, Backend, Applied AI/LLM... ...on products, while also ensuring that they are performant and reliable. You’ll work in a deeply iterative, collaborative, fast-paced...Work at officeRemote workWork from homeWorldwideFlexible hours2 days per week- ...leader in the design and manufacture of high-reliability Connectivity, Power, and Control... ...asylees. Description As a Senior Software Engineer, you will architect, develop, verify, and... ...Vision insurance Work Location In person, on-site in Albuquerque, New Mexico #J-18808-...Permanent employmentWorldwideRelocation packageFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- site services specialist Albuquerque, NM
- construction site safety Albuquerque, NM
- site leader Albuquerque, NM
- official site Albuquerque, NM
- website content developer Albuquerque, NM
- on site coordinator Albuquerque, NM
- IT site lead Albuquerque, NM
- site safety Albuquerque, NM
- junior website developer Albuquerque, NM
- on-site clinical research associate (traveling/remote) Albuquerque, NM



