Sr. Site Reliability Engineer
FreedomPay
Job Description
Job Description
The FreedomPay Commerce Platform is the technology of choice for many of the largest companies across the globe in retail, hospitality, lodging, gaming, sports and entertainment, foodservice, education, healthcare and financial services. FreedomPay’s technology has been purposely built to deliver rock solid performance in the highly complex environment of global commerce. The company maintains a world-class security environment and was first to earn the coveted validation by the PCI Security Standards Council against Point-to-Point Encryption with EMV standard in North America. FreedomPay’s robust solutions across payments, security, identity and data analytics are available in-store, online and on-mobile and are supported by rapid API adoption. The award winning FreedomPay Commerce Platform operates on a single, unified technology stack across multiple continents allowing enterprises to deliver a consistent, repeatable experience on a global scale. FreedomPay is a fast paced, high growth company with a great culture with competitive benefits and compensation with a business casual atmosphere.
FreedomPay is seeking an experienced Senior Site Reliability Engineer to help ensure the highest possible availability and resiliency of a rapidly growing global payment platform. This full-time salaried position builds on a strong foundation of observability, incident response, and support experience across the development lifecycle — and pushes it forward with AI-driven operations and automation at its core. The right candidate finds real satisfaction in eliminating manual toil, treats every recurring task as an automation opportunity, and is eager to apply modern AI tooling to detect, diagnose, and resolve issues faster than ever before.
About the RoleYou’ll join a team of SREs who work closely with other teams of world-class engineers to tenaciously and creatively solve problems and reduce manual toil wherever possible. We expect AI and automation to be a force multiplier in everything you do — from accelerating root-cause analysis and enriching alerts, to generating runbooks and codifying remediation so that the platform increasingly heals itself.
Successful candidates are heavily results-driven, bring well-established expertise across both traditional and bleeding-edge technology, and have a strong desire to continuously grow and improve themselves and our platform. This is a global operation spanning multiple regions and time zones, and the role demands the flexibility and commitment that a 24/7 payment platform requires.
This position participates in an engineering on-call rotation and provides after-hours support for production issue escalations on a rotational basis.
This position is based in the Philadelphia area with a hybrid schedule. Remote arrangements may be considered for exceptional candidates, with occasional travel to Philadelphia required.
- Build and maintain a comprehensive understanding of the platform and custom application stack.
- Implement, maintain, and continuously improve observability strategies and metrics that ensure complete system health for numerous complex products throughout all stages of the development lifecycle, up to and including production.
- Continuously identify automation opportunities and follow through to successful implementation, applying AI-assisted tooling to accelerate development and reduce manual effort.
- Design, build, and maintain automated remediation and self-healing workflows that detect, triage, and resolve common failure modes with minimal human intervention.
- Leverage AI/ML-driven observability — anomaly detection, alert correlation, and intelligent noise reduction — to surface issues earlier and shorten time to detection.
- Use AI-assisted analysis to accelerate root-cause investigation, enrich incident context, and generate first-draft postmortems and runbooks for human review.
- Handle escalations and collaborate effectively with other team members to quickly determine the root cause of any type of service degradation.
- Implement, maintain, and continuously improve incident response procedures and other operational documentation, automating documentation generation and upkeep wherever practical.
- Assist with troubleshooting and remediation of failed scheduled jobs and data-related concerns.
- Champion responsible, secure adoption of AI tooling across the SRE function — sharing patterns, prompts, and automations that raise the productivity of the whole team
AI and automation are central to how this team operates. We are looking for someone who will not only use these tools but help define how the SRE function applies them. In this role you will:
- Apply AI-assisted development and operations tools — including Anthropic (Claude), OpenAI (Codex), and Azure AI services (Foundry, Azure SRE Agent) and the agentic workflows built on them — to write, review, and accelerate automation and infrastructure code.
- Build and integrate automation that turns repetitive operational work into codified, repeatable, and self-service workflows.
- Use AIOps and ML-driven observability capabilities within the APM stack for anomaly detection, predictive alerting, and alert correlation.
- Develop and refine prompts, agents, and integrations that connect monitoring, ticketing, and remediation systems into faster end-to-end response loops.
- Evaluate emerging AI tooling for reliability and operations use cases, and advocate for adoption where it delivers measurable improvements in toil reduction, MTTR, or availability.
- Ensure all AI and automation usage adheres to FreedomPay’s security, privacy, and PCI obligations — keeping sensitive data appropriately protected and human review in place for high-impact actions.
- BS degree in Computer Science or equivalent, or equivalent years of relevant experience.
- Minimum of 5 years of hands-on technical experience in highly available, high-throughput, web-based technology environments.
- Demonstrated history of self-directed learning — someone who independently seeks out knowledge, builds new skills without being told to, and doesn’t wait for formal training to close gaps.
- Next-level problem-solving abilities and a strong bias toward practical, proven solutions.
- A track record of identifying and eliminating manual toil through automation.
- Excellent communication and organizational skills, with a strong sense of ownership and service.
- Expert-level proficiency in an enterprise APM platform and its AI/ML-driven (AIOps) capabilities; Dynatrace experience strongly preferred, though deep expertise in comparable tools such as Datadog or New Relic where readily transferable.
- Hands-on experience with AI-assisted development and automation tools — such as Anthropic (Claude), OpenAI (Codex), and Azure AI services (Foundry, Azure SRE Agent) — and a demonstrated ability to apply them to real operational and engineering work.
- Proficiency in scripting and automation — PowerShell and/or Python — to build tooling and remediation workflows.
- Strong SQL / T-SQL skills.
- Solid understanding of core networking concepts: DNS, load balancing, and TCP/IP routing and switching.
- Working knowledge of modern technology infrastructure including container orchestration, IaaS/PaaS cloud services, Azure, and VMware.
- Working knowledge of application development processes.
- Proven track record of successfully implementing SLI/SLOs and fostering their adoption across an organization.
- Experience implementing enterprise incident management practices.
- Experience building AIOps or ML-driven automation into production observability and incident response.
- Azure Kubernetes Service (AKS) and broader container orchestration experience.
- Windows Server (IIS) administration.
- PagerDuty Process Automation (formerly Rundeck) or comparable runbook automation platforms.
- Comprehensive experience supporting real-time transaction processing applications.
- PCI policies and best practices.
AI/ML model deployment, evaluation, or operations (MLOps).
Documentation automation and self-service tooling / service catalog implementation.
Experience integrating QA test automation into CI/CD pipelines.
As the fastest growing commerce company in the industry, we offer the opportunity for tremendous upward mobility within the company as well as development and professional growth opportunities. FreedomPay's fulltime roles provide exceptional benefits including medical, prescription, dental and vision coverage, Life Insurance, Retirement Plans with company match, commission sharing plan, flexible hybrid working environment, and great parental and other leave programs. All positions must be able to successfully pass a background check as well as a credit check.
FreedomPay is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or status as a protected veteran.
$192.4k - $275.8k
...CloudOps— the team that keeps Splunk Cloud running for some of the world's most demanding enterprise customers, blending Site Reliability Engineering, Systems Engineering, and Service Engineering disciplines at a scale very few teams ever get to operate at. When the...SeniorFull timeTemporary workLocal areaFlexible hours$121.4k - $218.6k
...will be responsible for ensuring best-in-class uptime and reliability of our AI hardware infrastructure offerings. Partner with... ...and defend them when they are breached. As a Senior Site Reliability Engineer, you will be responsible for: Developing and scaling robust...SeniorWork experience placementWork at office$119.8k - $234.7k
...per yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole type: Individual... ..., Cloud Hardware, and Infrastructure Engineering (SCHIE) is the team behind Microsoft’s... ...organization. As a Senior Linux SRE (Site Reliability Engineer), your main focus will be to...SeniorOngoing contractPermanent employmentWork at officeLocal areaWorldwide3 days per week$130k - $180k
...building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference... ...and AI R&D. The role Nebius is looking for a Site Reliability Engineer in Hardware Infrastructure team. This is a remote position...SuggestedTemporary workImmediate startRemote work- ...Critical Security Technology We’re seeking a Senior Software Engineer to join a high-impact team supporting advanced, network-... ...that coordinate across multiple systems, designing secure and reliable inter-service communication to support them. Own and enhance...SeniorFull time
- ...platform: the APIs, libraries, services, and data stores other engineering teams integrate with to control who can do what in their part... ...of every dashboard, API, and piece of customer data, so reliability and security posture come first. Much of our roadmap is customer...SeniorFull timeLive inRemote workVisa sponsorshipWork visaFlexible hours
- ...trendsetters in global delivery practices with our Client-Centric Model for customer management and delivery. Job Description Title: Sr. Web Software Developer Job duration : 4.5 year Location: Salem, OR Required Skills: Experience with C#, MVC, . Microsoft...SeniorFull timeFor contractorsRelocation
$161.7k - $258.8k
...drives innovation and delivers better business results.Opportunity OverviewWe are seeking a Senior System Integration and HAL Software Engineer to join our Semiconductor Test Engineering team. In this role, you will take ownership of developing Hardware Abstraction Layer (...SeniorFlexible hours$119.77k - $140.9k
...including design, development, testing, implementation, and support.Develop and maintain program specifications, ensuring performance, reliability, and service-level objectives are achieved.Troubleshoot and resolve complex production and application issues across Mainframe...SeniorFull timeWork experience placementLocal area3 days per week- Java Developer Location: Salem, Oregon Duration: Long Term No Positions: 2 Key Responsibilities: Develop and deliver updates to eXPRS application. This includes software code changes and documentation. Complete and document required enhancements, defect...Senior
- ...settings Partner with electrical, mechanical, systems, and test engineers to define interfaces and validate system performance... ...root-cause analysis Drive improvements in software quality, reliability, and development practices What You Bring Required:...SeniorFull timeTemporary work
$110k
...and aerial map copy options make DataVerify Flood Services a premier provider for any flood zone determinations needs. TITLE: Sr. Software Developer REPORTS TO: Software Development Manager STATUS: Salaried LOCATION: Columbus, OH;Remote...SeniorTemporary workWork at officeLocal areaRemote workMonday to FridayNight shift$205k - $235k
...management. We have become a multibillion-dollar asset manager, and we have ambitious goals for the future. As a Senior Cluster Site Reliability Engineer (SRE), you will help scale our research compute cluster to meet our growing needs, and you will leverage engineering...SeniorRemote jobLocal area- ...collaborative, and obsessed with building the best product in the industry. Come disrupt an industry with us. About the Role: The engineering team is scaling to meet demand that is outpacing our ability to ship. You'll build across our core platform products that...SeniorFull timeWork at officeLocal areaImmediate startRemote workVisa sponsorshipWork visaDay shift
- Skills : 1. Strong Experience using C# and ASP.NET, MVC(4.5 or greater) and .Net Core 2. Strong Experience using HTML5, CSS, Bootstrap, JavaScript, jQuery 3. Experience working with MS SQL Server (2014 or greater) 4. Experience using Visual...Senior
- Required: C#, WCF, MVC, XML, SQL, NUnit, Building RESTFul web services, White automation framework. Preferred: Web API, Mock Framework, Dependency Injection, MSMQ, Windows Services,Agile Methodologies, Stash/Bitbucket. Continuous integration...Senior
$144.6k - $198.8k
...revolution in enterprise software? As a Principal Software Development Engineer for the Vista Platform team to lead the integration of AI... ...Helm), and Blazor web applications with a focus on high-scale reliability and security. Architectural Stewardship: Modernize...SeniorOngoing contractFull timeLocal area- ...Senior DevOps Engineer We're looking for an experienced Senior DevOps Engineer to design, build, and maintain scalable cloud... ...Kubernetes, Terraform, CI/CD, automation, and infrastructure reliability . The ideal candidate is self-driven, comfortable...SeniorContract work
- ...cultures and affords personal and professional growth opportunities. Learn more at . Overview of Job Function The Senior Software Engineer is responsible for all aspects of the development of platforms and applications. This is a highly skilled hands-on role requiring...SeniorLocal areaShift work
$170.6k - $261.3k
...resulting data from vehicle to cloud. We are seeking a Senior Software Engineer to develop the services, integrations, automation, and end-to-... ...and eyes-off driving. Your work will ensure that data flows reliably from vehicles into our AI/ML and analytics platforms, with...SeniorFull timeLocal areaRemote workWork from homeRelocation packageFlexible hours- Job Title Visa status: U.S. Citizens and those authorized to work in the U.S. are encouraged to apply. Tax Terms: W2, 1099 Corp-Corp or 3rd Parties: Yes Description Required experience & skills: HTML5/CSS/Bootstrap/SASS/LESS. Modern frontend frameworks, preferably React...Senior
$135k - $180k
...you and your unique viewpoint matter. Learn about the Danaher Business System which makes everything possible. The Sr. AI Engineer, Device Intelligence will be a key member of the Danaher Autonomous and Intelligent Labs team, reporting to its Vice President....SeniorRemote workWork from homeFlexible hours$209k - $238.5k
...Sr. Lead Software Engineer, Full Stack - Shopping Tech Do you love building and pioneering in the technology space? Do you enjoy solving complex... ...tools or other information available through this site. Capital One Financial is made up of several different entities...SeniorFull timePart timeInternshipLocal areaRemote work- ...DoorDash, Affirm, and Zillow.About the roleWe are looking for our Engineering Manager who has a solid track record of developing and... ...help drive efficiency and high velocity for the teamOwn the reliability of your product area in production. Our engineering teams are...SeniorFull timeRemote workVisa sponsorshipWork visa
$77k - $202k
...ApplicableSpecialismFunctional & Industry TechnologiesManagement LevelSenior AssociateJob Description & SummaryThe OpportunityAs an EAM IFS Cloud Consultant, Sr. Associate, you will engage with clients to optimize their operational efficiency through the analysis, implementation, and support...SeniorFull timeH1b$100k - $120k
...partnerships, while our focus on creativity and innovative solutions empowers our customer communities to thrive. The Senior Software Engineer on the New Ventures team designs, creates, maintains, audits, and improves software applications by performing coding, debugging,...SeniorFull timeTemporary workLocal areaRemote work$112.3k - $140k
...looking for a stimulating and challenging career where your engineering expertise can be leveraged to create, enhance, and maintain our... ...long-term partnership, uncompromising support, and proven reliability. Major product lines include fuel dispensers, tank gauges and...SeniorWork at officeLocal areaRemote workWorldwide- Description Strong Experience using C# and ASP.NET, MVC(4.5 or greater) and .Net Core Strong Experience using HTML5, CSS, Bootstrap, JavaScript, jQuery Experience working with MS SQL Server (2014 or greater) Experience using Visual Studio 2015 & 2017 Experience in writing...Senior
- Sr. Systems Engineer (Onsite in Wilsonville, OR)SummaryWe are seeking a Senior Systems Engineer to join our team in Portland, OR (PDX). In this role, you will lead the development, validation, and integration of cutting-edge automated laboratory instrumentation and bioprocess...Senior
- WHO WE’RE LOOKING FORWe are looking for a Senior Principal Software Engineer with deep expertise in FP&A, a strong understanding of P&L and financial drivers, and a proven track record delivering enterprise‑scale finance planning solutions using Anaplan. You are a recognized...SeniorFull time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Sr. Site Reliability Engineer. Be the first to apply!
- senior operations technician Oregon State
- senior cloud service delivery manager Oregon State
- senior director clinical development Oregon State
- senior financial analyst remote Oregon State
- senior designer Oregon State
- senior java full-stack developer Oregon State
- senior application support engineer Oregon State
- senior storage engineer Oregon State
- senior manager tax Oregon State
- senior manager product engineering Oregon State


