Senior Manager, Site Reliability Engineering - Paylo Platform
PDI Technologies
Job Description
Job Description
At PDI Technologies, we empower some of the world's leading convenience retail and petroleum brands with cutting-edge technology solutions that drive growth and operational efficiency. By “Connecting Convenience” across the globe, we empower businesses to increase productivity, make more informed decisions, and engage faster with customers through loyalty programs, shopper insights, and unmatched real-time market intelligence via mobile applications, such as GasBuddy. We’re a global team committed to excellence, collaboration, and driving real impact. Explore our opportunities and become part of a company that values diversity, integrity, and growth.
Role Overview
PDI Technologies is looking for a Senior Manager, Site Reliability Engineering to lead the SRE organization supporting Paylo, PDI’s payments, loyalty, and fuel-pricing product suite. This role owns the reliability, infrastructure, and operational strategy for a portfolio of high-traffic, customer- and partner-facing platforms that power payment transactions, fuel pricing, loyalty and rewards, and offer/coupon redemption for convenience retail and fuel customers around the world.
This is a hands-on, leadership-first role. You will manage a team of three SRE Managers/Leads who together lead approximately 20 engineers, while staying technically engaged yourself — reviewing architecture, unblocking hard infrastructure problems, and setting the technical bar across the organization. You will bring strong, current, hands-on expertise across AWS, Azure, Kubernetes, Helm, Argo CD, Terraform/OpenTofu, Jenkins, and Datadog, and you will be a strong, visible people leader who can coach managers and represent SRE to senior engineering and business stakeholders.
Key ResponsibilitiesDirectly manage and develop 3 SRE Managers/Leads and own the overall health, growth, and performance of an ~20-person SRE organization supporting the Paylo product suite.
Set the vision, priorities, and operating cadence for the SRE function; translate business and product priorities into a reliability roadmap your managers can execute against.
Build a strong bench by hiring, coaching, and developing managers and senior engineers while creating clear career paths and succession plans.
Foster a blameless, learning-oriented culture around incidents, on-call, and operational excellence.
Partner closely with engineering directors, product managers, and business stakeholders across the Paylo organization to align reliability investments with business risk and customer impact.
Stay technically engaged day to day by participating in architecture and design reviews, troubleshooting complex production issues, and directly contributing to infrastructure-as-code, Kubernetes manifests/Helm charts, and CI/CD pipelines when needed.
Set and enforce engineering standards for multi-cloud infrastructure across AWS and Azure and for container orchestration on Kubernetes at scale.
Own adoption and standards for GitOps-based continuous delivery using Argo CD/Argo Workflows, including deployment strategy, rollout policy, and multi-cluster promotion.
Own the Infrastructure-as-Code strategy across teams (Terraform, OpenTofu), including module standards, state management, drift detection, and remediation.
Own CI/CD pipeline architecture and standards built on Jenkins, driving build/deploy automation, pipeline reliability, and progressive delivery practices such as blue-green/canary deployments and automated rollback.
Evaluate and guide adoption of new infrastructure tooling and patterns as the platform evolves across AWS and Azure.
Own the observability strategy across all supported products, with deep, hands-on expertise in Datadog (APM, infrastructure monitoring, log management, dashboards, and alerting) as the standard platform for metrics, tracing, and alerting.
Define and drive adoption of SLIs/SLOs, error budgets, and reliability KPIs across the organization, holding managers and teams accountable to them.
Own the incident management program end to end, including on-call structure, escalation paths, severity definitions, postmortems, and follow-through on remediation actions.
Drive root-cause analysis and long-term reliability investments that reduce Sev1/Sev2 frequency and recurrence.
Ensure appropriate resilience, disaster recovery, and capacity planning practices are in place given the sensitivity of payment- and transaction-related systems.
Partner with Security and Compliance to maintain awareness of PCI DSS and related compliance requirements and ensure the SRE organization supports audit and compliance readiness.
Track and report cost, capacity, and operational KPIs to senior leadership.
8+ years of experience in Site Reliability Engineering, DevOps, or Infrastructure/Platform Engineering, including 4+ years in a people-leadership role.
Proven experience managing managers — you have directly led team leads/managers, not just individual contributors, and are comfortable operating at the scale of ~20 total reports.
Strong, hands-on expertise across AWS and Azure — you can architect, troubleshoot, and operate multi-cloud infrastructure yourself, not just direct others to do so.
Strong, hands-on expertise with Kubernetes and Helm — cluster operations, troubleshooting at scale, and chart design/maintenance.
Strong, hands-on expertise with Argo CD/Argo Workflows for GitOps-based continuous delivery.
Strong, hands-on expertise with Infrastructure as Code (Terraform, OpenTofu), including module design and state management.
Strong, hands-on expertise with Jenkins for CI/CD pipeline design, administration, and automation.
Strong, hands-on expertise with Datadog (or equivalent enterprise observability platform), including designing monitoring/alerting strategy, dashboards, and APM/tracing at scale.
Demonstrated track record of driving incident management, on-call, and postmortem programs for high-traffic, customer-facing systems.
Excellent communication and stakeholder-management skills; able to represent SRE to engineering leadership and business partners with equal credibility.
A strong, visible leadership style — someone who sets clear direction, holds teams accountable, and builds trust across the organization.
- Applicants must be legally authorized to work in the United States without the need for employer sponsorship, now or in the future. PDI Technologies is unable to offer visa sponsorship for this role.
Experience supporting payments, fuel/retail, or loyalty platforms, or other systems with PCI DSS or similar compliance obligations.
Relevant certifications such as CKA/CKAD, AWS Certified Solutions Architect, Microsoft Certified: Azure Solutions Architect, or HashiCorp Terraform Associate.
Experience with messaging systems (Kafka/SQS/SNS), PagerDuty (or similar), and multi-region/multi-AZ resilience patterns.
Prior experience consolidating or standardizing SRE and DevOps practices across multiple product lines or recently- integrated/acquired teams.
Experience partnering with product and business stakeholders to translate reliability investments into business outcomes.
A stable, well-led SRE organization with clear ownership, career paths, and low regrettable attrition among your managers and their teams.
Consistent, Datadog-driven observability and SLOs in place across the organization, with measurable reduction in Sev1/Sev2 incidents and mean time to detect/resolve.
Modern, standardized infrastructure practices — GitOps delivery via Argo, IaC via Terraform/OpenTofu, and reliable CI/CD via Jenkins — adopted consistently across teams and clouds.
A mature, blameless incident-management culture with strong postmortem follow-through.
Strong cross-functional trust with engineering, product, and security/compliance stakeholders.
PDI is committed to offering a well-rounded benefits program, designed to support and care for you, and your family throughout your life and career. This includes a competitive salary, market-competitive benefits, and a quarterly perks program. We encourage a good work-life balance with ample time off [time away] and, where appropriate, hybrid working arrangements. Employees have access to continuous learning, professional certifications, and leadership development opportunities. Our global culture fosters diversity, inclusion, and values authenticity, trust, curiosity, and diversity of thought, ensuring a supportive environment for all.
We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.
- Core Responsibilities:Evaluate applications, platforms, and vendors to assess resiliency, reliability, and operational risk.Design and implement processes that... ...tooling.Actively participate in reliability engineering and resilience communities of practice, contributing...SeniorFull time
$104.9k - $174.7k
...Credit Risk mitigation and Customer Data Management. You can learn more about LexisNexis... ...the Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate,... ...environmentsExperience operating monitoring and uptime platforms such as Grafana, Pingdom, and...SeniorFull timeWork at officeLocal areaRemote workWork from home- ...Motion Recruitment Partners LLC is seeking a Senior Java Applications Administrator / SRE in... .... You will optimize performance, drive reliability, and mentor teammates while aligning... ...observability to improve operational efficiency and platform strategy. #J-18808-Ljbffr...Senior
- ...with software developers, platform engineers, and IT staff to improve system... ..., service quality, reliability, security, and compliance needs... ...8+ years of experience in Site Reliability Engineering, DevOps... ...and configuration management. Experience managing AI tooling...SeniorWork at officeRemote work
- ...payments and financial platform for global businesses.... ...solutions to manage everything from business... ...is responsible for the reliability, performance, security... ...databases invisible: product engineers should be able to... .... What you'll do As a Senior/Staff Software Engineer...SeniorWorldwide
$180.5k - $236.91k
...'re Oscar. We're hiring a Senior Software Engineer, Cloud Infrastructure / SRE... ...a full stack technology platform and a relentless focus on... ...technical domains such as DevOps, site reliability, and cloud best practices... ...with partners, product managers, and designers to solve...SeniorFull timeWork at officeRemote work$185k - $227k
...bestengineers, scientists, designers, product managers, operations experts, and customer... .... ROLE AND RESPONSIBILITIES: A Senior Site Reliability Engineer (SRE) is expected to own the... ...is scalable and efficient. Nutanix Platform Management Design, deploy, and maintain...SeniorRemote work$92.7k - $203.94k
...Summary: About the Team Our Site Reliability Engineering team is the execution... ...We operate across pharmacy platforms, Point of Sale (POS) systems... ...produce. About the Role As a Senior Software Engineer - SRE,... ...mindset to change management: for any configuration or...SeniorHourly payFull timeTemporary workLocal areaShift work- Reference: BBBH92606_1786569999** Senior Preconstruction Manager - Dallas, Texas (Award-Winning ENR Construction Firm) **My client, a nationally recognized, award-winning ENR General Contractor are seeking a Senior Preconstruction Manager in Dallas, Texas to join their...SeniorFor contractors
- ...Finance, the ideal candidate should possess exceptional analytical abilities and strong client engagement skills. This role involves managing client relationships and guiding teams through the complexities of tax regulations while fostering a motivating work environment....Senior
$124k - $280k
...PwC, our people in data and analytics engineering focus on leveraging advanced technologies... ...Sets You Apart- Certification in Cloud Platforms [e.g., AWS Certified Solutions... ...solutions using cloud services- Designing and managing data warehouses and data lakes- Implementing...SeniorFull timeH1b$156.18k
...Dynamics 365 Developer Senior... ...Technology Consulting Senior Manager to join our growing Microsoft... ...within the Dynamics 365 platform. Advanced proficiency... ...Computer Science, Software Engineering, or related field). 7+... ...offices and on client sites, which can include local...SeniorFull timeTemporary workWork at officeLocal areaRemote workFlexible hours$138.4k - $173k
...well as help improve the reliability, quality of services... ..., configuration management, DDoS protection, infrastructure... ...or embed with engineering teams, helping them to... ...AppFolio Real Estate Platform. You’ll help build the... ...locations by visiting our site.Compensation & BenefitsThe...SeniorFull timeFlexible hours- ...cyber defense, application security, and managed service solutions to rethink the entire... ...engagement. A Cybersecurity Forward Deployed Engineer is a production engineer who works... ...multi-system integration risk across cloud platforms (AWS, Azure, or GCP)Lead AI governance...SeniorFull timeWork experience placementLive inWork at officeLocal area
- DescriptionWe are looking for an experienced Senior Manager to oversee financial reporting processes within the dynamic oil and gas industry. Based in Dallas, Texas, this role involves managing a team responsible for external reporting obligations, ensuring compliance...SeniorWork at office
$115k - $150k
...Onshore Oracle Retail Cloud Functional Lead / Senior Consultant, you will have the ability to... ...for supporting Deloitte’s Application Management Services (AMS) engagement providing... ...Science, Information Technology, Computer Engineering, or related IT discipline; or equivalent...SeniorLocal areaRemote workVisa sponsorship$100k - $130k
Business Network Consulting is seeking a Senior Consultant based in Dallas, Texas. The role demands a minimum of 5 years of consulting... ...and advanced virtualization. Responsibilities include managing IT resources for multiple clients, providing IT support, and mentoring...SeniorFull time$125k - $150k
...Dallas, TXJob Code: JM073126Job Type: PermanentCategory: EngineeringPlacement Manager: James MeeksSalary Range: $125,000-$150,000An established Texas engineering firm is seeking a licensed Senior Structural Project Manager for its FortWorth office. This leadership role...SeniorFor contractorsWork at office$99k - $232k
...OpportunityAs a Delivery Excellence - Tech Enablement - Senior Developer - Manager, you will play a pivotal role in driving software and product... ...digital transformation- Managing and mentoring software engineering teams to enhance performance and deliver quality outcomes...SeniorFull timeH1b- About the RoleThe Enterprise Platforms team within our Data & Agentic Platform Solutions... ...and DevOps capabilities. As an IT Site Reliability Engineer within the Enterprise Platforms team,... ...— the backbone of TI's API management and integration strategy — while also...Local area
- Site Reliability Engineer - Vice PresidentSite Reliability Engineering (SRE) is... ...of the firm’s most critical platform services and ensures they meet... ...leadership, mentoring senior engineers, and collaborating... ...enterprise.Complex Incident Management & Post-Mortem Analysis: Lead...
- ...Job Responsibilities Program & Portfolio Management – Oversee multiple projects/programs, manage interdependencies... ...Management – Partner with Product Owners, Engineering Managers, COE teams, finance, and senior leadership to align priorities and drive successful...Senior
- ...Seeking a Senior Facilities Manager (SFM) to join our team in Texas market, to serve as the primary point of contact for our client while ensuring exceptional service delivery and tenant satisfaction across the portfolio. This strategic role supporti Facilities Manager...Senior
- ...role in any organization who has a strong understanding of Clients culture and operating modelThe individual should possess strong management coordination and stakeholder engagement skills and be capable of effectively managing priorities driving execution and handling...SeniorWork experience placement
- ...Technology Centers (ATCs) are the engine for reinvention in our... ...client challenges.You Are:A Senior Forward Deployed Engineer leading... ...teams on Snowflake: data management, analytics, AI/ML, and BI integration... ...experience working with AI platforms across OpenAI, Claude, Vertex...SeniorFull timeWork experience placementLive inWork at officeLocal area
- ...Job description Senior HR Business Partner, Dallas, TX We'... ..., onsite people resource for managers and employees Coach managers... ..., ideally for technical, engineering, or development roles ~ A collaborative... ...the steady HR presence on site ~ Broad generalist depth...SeniorFull timeWork at officeLocal areaVisa sponsorship
$86.8k - $165.2k
...than 100 years of experience and renowned engineering expertise to meet the needs of today’s... ...aerospace and defense.The Constellation Manager State Aggregation & Payload Management team... ...of whether the role is designated as on-site, hybrid or remote.The salary range for...SeniorTemporary workWork experience placementWork at officeRemote workRelocation packageFlexible hours- ...Trace3 is seeking a Senior Project Manager to oversee cross‑functional enterprise system implementations and enhancements. You will manage full lifecycle projects, ensuring on‑time delivery within scope and budget, while aligning with business goals and executive expectations...Senior
- ...Strategic Human Resources Leadership Develop and implement HR strategies aligned with business goals. Partner with senior management to identify workforce needs and talent gaps. Lead organizational development, succession planning, and workforce planning initiatives...SeniorInterim role
$65k - $140k
...You Will DO In This Role:As a Sr. Project Manager, you will have a significant impact on... ...Manager, Project Controller, Project Lead Engineer, and the project core team to translate the... ...the following:Cooperation with the Site Manager and the EHS&S Group on the safety...SeniorFull timeContract workLocal area
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Manager, Site Reliability Engineering - Paylo Platform. Be the first to apply!
- site reliability engineer sre Dallas, TX
- site reliability engineer Dallas, TX
- data platform engineer Dallas, TX
- client platform engineer Dallas, TX
- platform engineering manager Dallas, TX
- platform developer Dallas, TX
- senior platform engineer Dallas, TX
- platform engineer Dallas, TX
- senior maintenance supervisor Dallas, TX
- senior operations associate Dallas, TX



