Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Sr. Site Reliability Engineer

FreedomPay

Job Description

Job Description

The FreedomPay Commerce Platform is the technology of choice for many of the largest companies across the globe in retail, hospitality, lodging, gaming, sports and entertainment, foodservice, education, healthcare and financial services.  FreedomPay’s technology has been purposely built to deliver rock solid performance in the highly complex environment of global commerce. The company maintains a world-class security environment and was first to earn the coveted validation by the PCI Security Standards Council against Point-to-Point Encryption with EMV standard in North America. FreedomPay’s robust solutions across payments, security, identity and data analytics are available in-store, online and on-mobile and are supported by rapid API adoption. The award winning FreedomPay Commerce Platform operates on a single, unified technology stack across multiple continents allowing enterprises to deliver a consistent, repeatable experience on a global scale.  FreedomPay is a fast paced, high growth company with a great culture with competitive benefits and compensation with a business casual atmosphere.

FreedomPay is seeking an experienced Senior Site Reliability Engineer to help ensure the highest possible availability and resiliency of a rapidly growing global payment platform. This full-time salaried position builds on a strong foundation of observability, incident response, and support experience across the development lifecycle — and pushes it forward with AI-driven operations and automation at its core. The right candidate finds real satisfaction in eliminating manual toil, treats every recurring task as an automation opportunity, and is eager to apply modern AI tooling to detect, diagnose, and resolve issues faster than ever before. 

About the Role

You’ll join a team of SREs who work closely with other teams of world-class engineers to tenaciously and creatively solve problems and reduce manual toil wherever possible. We expect AI and automation to be a force multiplier in everything you do — from accelerating root-cause analysis and enriching alerts, to generating runbooks and codifying remediation so that the platform increasingly heals itself. 

Successful candidates are heavily results-driven, bring well-established expertise across both traditional and bleeding-edge technology, and have a strong desire to continuously grow and improve themselves and our platform. This is a global operation spanning multiple regions and time zones, and the role demands the flexibility and commitment that a 24/7 payment platform requires. 

  • This position participates in an engineering on-call rotation and provides after-hours support for production issue escalations on a rotational basis. 

  • This position is based in the Philadelphia area with a hybrid schedule. Remote arrangements may be considered for exceptional candidates, with occasional travel to Philadelphia required.

Primary Responsibilities:

  • Build and maintain a comprehensive understanding of the platform and custom application stack.
  • Implement, maintain, and continuously improve observability strategies and metrics that ensure complete system health for numerous complex products throughout all stages of the development lifecycle, up to and including production.
  • Continuously identify automation opportunities and follow through to successful implementation, applying AI-assisted tooling to accelerate development and reduce manual effort.
  • Design, build, and maintain automated remediation and self-healing workflows that detect, triage, and resolve common failure modes with minimal human intervention.
  • Leverage AI/ML-driven observability — anomaly detection, alert correlation, and intelligent noise reduction — to surface issues earlier and shorten time to detection.
  • Use AI-assisted analysis to accelerate root-cause investigation, enrich incident context, and generate first-draft postmortems and runbooks for human review.
  • Handle escalations and collaborate effectively with other team members to quickly determine the root cause of any type of service degradation.
  • Implement, maintain, and continuously improve incident response procedures and other operational documentation, automating documentation generation and upkeep wherever practical.
  • Assist with troubleshooting and remediation of failed scheduled jobs and data-related concerns.
  • Champion responsible, secure adoption of AI tooling across the SRE function — sharing patterns, prompts, and automations that raise the productivity of the whole team
AI Enablement & Automation

AI and automation are central to how this team operates. We are looking for someone who will not only use these tools but help define how the SRE function applies them. In this role you will: 

  • Apply AI-assisted development and operations tools — including Anthropic (Claude), OpenAI (Codex), and Azure AI services (Foundry, Azure SRE Agent) and the agentic workflows built on them — to write, review, and accelerate automation and infrastructure code. 
  • Build and integrate automation that turns repetitive operational work into codified, repeatable, and self-service workflows. 

  • Use AIOps and ML-driven observability capabilities within the APM stack for anomaly detection, predictive alerting, and alert correlation. 

  • Develop and refine prompts, agents, and integrations that connect monitoring, ticketing, and remediation systems into faster end-to-end response loops. 

  • Evaluate emerging AI tooling for reliability and operations use cases, and advocate for adoption where it delivers measurable improvements in toil reduction, MTTR, or availability. 

  • Ensure all AI and automation usage adheres to FreedomPay’s security, privacy, and PCI obligations — keeping sensitive data appropriately protected and human review in place for high-impact actions. 
Required Background and Experience

  • BS degree in Computer Science or equivalent, or equivalent years of relevant experience.
  • Minimum of 5 years of hands-on technical experience in highly available, high-throughput, web-based technology environments.
  • Demonstrated history of self-directed learning — someone who independently seeks out knowledge, builds new skills without being told to, and doesn’t wait for formal training to close gaps.
  • Next-level problem-solving abilities and a strong bias toward practical, proven solutions.
  • A track record of identifying and eliminating manual toil through automation.
  • Excellent communication and organizational skills, with a strong sense of ownership and service. 
Required Technical Skills

  • Expert-level proficiency in an enterprise APM platform and its AI/ML-driven (AIOps) capabilities; Dynatrace experience strongly preferred, though deep expertise in comparable tools such as Datadog or New Relic where readily transferable.
  • Hands-on experience with AI-assisted development and automation tools — such as Anthropic (Claude), OpenAI (Codex), and Azure AI services (Foundry, Azure SRE Agent) — and a demonstrated ability to apply them to real operational and engineering work.
  • Proficiency in scripting and automation — PowerShell and/or Python — to build tooling and remediation workflows.
  • Strong SQL / T-SQL skills.
  • Solid understanding of core networking concepts: DNS, load balancing, and TCP/IP routing and switching.
  • Working knowledge of modern technology infrastructure including container orchestration, IaaS/PaaS cloud services, Azure, and VMware.
  • Working knowledge of application development processes. 
Preferred Technical Skills and Experience

  • Proven track record of successfully implementing SLI/SLOs and fostering their adoption across an organization.
  • Experience implementing enterprise incident management practices.
  • Experience building AIOps or ML-driven automation into production observability and incident response.
  • Azure Kubernetes Service (AKS) and broader container orchestration experience.
  • Windows Server (IIS) administration.
  • PagerDuty Process Automation (formerly Rundeck) or comparable runbook automation platforms.
  • Comprehensive experience supporting real-time transaction processing applications.
  • PCI policies and best practices. 
Additional Experience, a Plus

  • AI/ML model deployment, evaluation, or operations (MLOps). 

  • Documentation automation and self-service tooling / service catalog implementation. 

  • Experience integrating QA test automation into CI/CD pipelines. 

 

 

As the fastest growing commerce company in the industry, we offer the opportunity for tremendous upward mobility within the company as well as development and professional growth opportunities. FreedomPay's fulltime roles provide exceptional benefits including medical, prescription, dental and vision coverage, Life Insurance, Retirement Plans with company match, commission sharing plan, flexible hybrid working environment, and great parental and other leave programs. All positions must be able to successfully pass a background check as well as a credit check.

 

FreedomPay is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or status as a protected veteran.

Vacancy posted 12 days ago
Similar jobs that could be interesting for youBased on the Sr. Site Reliability Engineer in Oregon State vacancy
  •  ...alongside some of the most experienced and innovative leaders and engineers in the field. Where we work Headquartered in Amsterdam and...  ...vision, audio, and emerging multimodal architectures — fast, reliable, and effortless to deploy at massive scale. To deliver on that... 
    Senior

    GrabJobs

    Portland, OR
    3 days ago
  • $210k - $220k

     ...secure and private by design, it’s popular with security, IT, engineering, finance, and other security-focused teams. At Tines, we're...  ...we’re looking for others to join us on our journey. Senior Site Reliability Engineer - Government Cloud You'll join the team responsible... 
    Senior
    Work at office
    Remote work

    GrabJobs

    Portland, OR
    4 days ago
  • $81.1k - $187k

     ...Job Description We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The role focuses on improving service reliability, reducing operational risk, automating repetitive tasks, and driving faster detection... 
    Senior
    Temporary work
    Immediate start
    Flexible hours
    Shift work

    Oracle

    Salem, OR
    1 day ago
  • $121.4k - $218.6k

     ...will be responsible for ensuring best-in-class uptime and reliability of our AI hardware infrastructure offerings. Partner with...  ...and defend them when they are breached. As a Senior Site Reliability Engineer, you will be responsible for: Developing and scaling robust... 
    Senior
    Work experience placement
    Work at office

    Akamai

    Salem, OR
    5 days ago
  • $84.9k - $209.5k

     ...Job Description As a Principal Site Reliability Engineer (IC4), you will be responsible for designing, building, and operating highly available, scalable, secure, and resilient cloud services. You will combine software engineering with infrastructure expertise to improve... 
    Suggested
    Temporary work
    Flexible hours

    Oracle

    Salem, OR
    1 day ago
  • $100 per hour

     ...our New York City HQ (EST) Where you'll create impact Improve reliability of our systems Build & maintain our main infrastructure (...  ...learn, are highly curious about new frameworks and solutions to engineering problems Fast-moving: you deploy daily, iterate quickly, and... 
    Immediate start
    Remote work
    Work from home
    Relocation
    Home office
    Visa sponsorship
    Weekend work

    GrabJobs

    Portland, OR
    1 day ago
  • $119k - $170k

     ...the greater good, come make your next move with Zscaler. Our Engineering team built the world’s largest cloud security platform from...  ...cloud-first strategy. We’re looking for an experienced Staff Site Reliability Engineer (Federal) to join our Government Cloud team.... 
    Full time
    Work at office
    Local area
    Worldwide
    Night shift

    GrabJobs

    Portland, OR
    2 days ago
  • $1,500 per month

     ...game worlds they inhabit. Our approach is centered around World Engine, our state-of-the-art onchain game server framework. World...  ...architecture to keep our platform secure. Own delivery, scalability, and reliability of our backend infrastructure. Advise and collaborate with the... 
    Full time
    Flexible hours

    GrabJobs

    Portland, OR
    4 days ago
  • $74.1k - $148.3k

     ...systems. Facilitate service capacity planning and demand forecasting, software performance analysis, and system tuning. As a Site Reliability Engineer, you will solve interesting technical challenges by defining, designing, deploying, and solving key Oracle Cloud services,... 
    Temporary work
    Immediate start
    Flexible hours

    Oracle

    Salem, OR
    1 day ago
  • $114k - $148k

     ...Site Reliability Engineer Location: Remote, United States Employment Type: Full-Time Benefits Offered: Vision, Medical, Life, Dental, 401K Gross Annual Base Salary: USD 114,000-148,000 Additional variable compensation and benefits may apply. Total compensation is based... 
    Full time
    Temporary work
    Work experience placement
    Remote work

    GrabJobs

    Portland, OR
    3 days ago
  •  ...Site Reliability/DevOps Engineer As a Site Reliability Engineer, you will be a member of a cross-functional Engineering team building solutions to meet the needs of our customers, partners and internal operations staff. Your ability to lead your teammates towards robust... 

    1872 Consulting

    Beaverton, OR
    1 day ago
  • $101k - $152k

     ...regulatory and Medline standards throughout the software development life cycle.MINIMUM JOB REQUIREMENTSEducation: Bachelor’s degree in Engineering, Quality, Business, or Computer Science.Work Experience:5+ years of experience in Manufacturing, Quality or Engineering.3+ years... 
    Senior
    Minimum wage
    Full time
    Work experience placement
    Local area
    Worldwide

    Medline

    Redmond, OR
    1 day ago
  • $161.7k - $258.8k

     ...drives innovation and delivers better business results.Opportunity OverviewWe are seeking a Senior System Integration and HAL Software Engineer to join our Semiconductor Test Engineering team. In this role, you will take ownership of developing Hardware Abstraction Layer (... 
    Senior
    Flexible hours

    Universal Robots

    Tualatin, OR
    8 hours ago
  •  ...industrial growth markets that require advanced technology and high reliability. These markets include aerospace and defense, factory...  ...future.Job Summary:Teledyne FLIR is looking for a career Software Engineer to join our Surveillance team who has industry experience... 
    Senior
    Permanent employment
    Full time
    Local area
    Visa sponsorship
    Work visa

    Teledyne FLIR

    Wilsonville, OR
    8 hours ago
  • $209.1k - $275.1k

     ...remote within the USMeet the TeamJoin a team of Cisco Solutions Engineers! These technically skilled, solutions-focused individuals build...  ..., and basic life insurance. Please see the Cisco careers site to discover more benefits and perks. Employees may be eligible... 
    Senior
    Full time
    Temporary work
    Local area
    Remote work
    Flexible hours

    CISCO Systems

    Portland, OR
    6 days ago
  • $125k - $185k

     ...Pleasanton / New York / US - Remote / PortlandDivisions - Capture Engineering /Full-Time /HybridWho are we?Smarsh empowers its customers to...  ...incident response, driving root-cause analysis and long-term reliability and performance improvements across diverse runtime... 
    Senior
    Full time
    Local area
    Immediate start
    Remote work

    Smarsh

    Portland, OR
    3 days ago
  •  ...management and organizational skills ~ Excellent written and verbal communication skills ~ Bachelor’s degree in computer science, engineering, mathematics, GIS, or related field Recommended Qualifications Experience with IDEs, compilers, and development tools... 
    Senior
    Full time
    Relocation
    Relocation package

    Esri

    Portland, OR
    1 day ago
  • $205k - $235k

     ...management. We have become a multibillion-dollar asset manager, and we have ambitious goals for the future.  As a Senior Cluster Site Reliability Engineer (SRE), you will help scale our research compute cluster to meet our growing needs, and you will leverage engineering... 
    Senior
    Remote job
    Local area

    The Voleon Group

    Oregon State
    more than 2 months ago
  • $195.2k - $361.2k

     ...agent loop itself.This is core product engineering on an agent framework comparable to well...  ...builds the harness. Shipping and running it reliably in production is owned by the SRE /...  ...to split their time between working on-site at their assigned Intel site and off-site... 
    Senior
    Full time
    Internship
    Local area
    Immediate start
    Shift work

    Intel

    Hillsboro, OR
    2 days ago
  •  ...collaborative, and obsessed with building the best product in the industry. Come disrupt an industry with us. About the Role: The engineering team is scaling to meet demand that is outpacing our ability to ship. You'll build across our core platform products that... 
    Senior
    Full time
    Work at office
    Local area
    Immediate start
    Remote work
    Visa sponsorship
    Work visa
    Day shift

    Prophetic Technologies

    Portland, OR
    1 day ago
  •  ...Senior DevOps Engineer We are looking for senior DevOps engineers who excel in AWS/cloud multi-region/dc environments that can scale...  ...Design and develop tools and frameworks to improve security, reliability, maintainability, availability and performance for the... 
    Senior

    BizTek People

    Hillsboro, OR
    1 day ago
  • WHO WE’RE LOOKING FORWe are looking for a Senior Principal Software Engineer with deep expertise in FP&A, a strong understanding of P&L and financial drivers, and a proven track record delivering enterprise‑scale finance planning solutions using Anaplan. You are a recognized... 
    Senior
    Full time

    Nike

    Beaverton, OR
    3 days ago
  • $88.86k - $118.48k

     ...deliver meaningful impact, and help shape the future of AI‑ready connectivity, join us today. The Role The Senior IT Systems Engineer provides advanced Tier II support by troubleshooting and repairing network devices, tools, and services for a nationwide fiber... 
    Senior
    Full time
    Temporary work
    Work at office
    Remote work
    Shift work
    Night shift

    Lumen

    Salem, OR
    3 days ago
  •  ...automation development. Including: Style guides, versioning practices, source control, branching and merging patterns and advising other engineers on development standards Develop and advocate for Operations best practices, standards, and processes Provide support to... 
    Senior

    3B Staffing LLC

    Happy Valley, OR
    5 days ago
  • $111.72k - $138.08k

     ...our successes. Learn more. Position Summary The Sr. Salesforce Developer is a senior technical contributor...  ...accountable for delivering production-ready solutions, improving engineering quality and platform reliability, and driving continuous modernization and optimization... 
    Senior
    Work experience placement
    Work at office
    Remote work

    ISC2

    Salem, OR
    1 day ago
  • $152k - $241.5k

     ...Make the choice, join our diverse team today!The Advanced Technology Group is looking for a highly motivated Senior Systems Software Engineer to join our group. Do you have a proven software development background in advanced computational methods for semiconductor... 
    Senior
    Full time

    Nvidia

    Hillsboro, OR
    4 days ago
  • $155k - $278.3k

     ...complex customer problems. You will partner closely with Product, Engineering, and executive leadership to shape strategy, influence outcomes...  ...? Please search for open jobs and apply internally (not on this external site).SummaryLocation: Portland, OR, USAType: Full time... 
    Senior
    Full time
    For contractors
    Work at office

    Autodesk

    Portland, OR
    8 hours ago
  •  ...Senior System Engineer Reporting to the Manager, IT System Operations, the (Senior) System Engineer is responsible for maintaining,...  ...tooling, and partners with teams across the business to deliver reliable and scalable technology solutions. At the senior level, the position... 
    Senior

    Digital Realty

    Portland, OR
    4 days ago
  •  ...industrial process data, serving sectors like pharmaceuticals, energy, and manufacturing. Our core product is a robust calculation engine capable of executing advanced math and machine learning algorithms on streaming time-series data. By leveraging generative AI, we enhance... 
    Senior
    Remote work

    Randstad

    Myrtle Point, OR
    1 day ago
  • $130.56k - $179.52k

    DescriptionPOSITION:Senior Software Engineer-- or -- Senior Staff Software Engineer -- or --Principal Software EngineerPosition Level: The best fit candidate selected for this position will be offered a job title/level (Senior Software Engineer vs. Senior Staff Software... 
    Senior
    Remote work

    Hitachi

    Hillsboro, OR
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Sr. Site Reliability Engineer. Be the first to apply!