Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Cloud Infrastructure & Reliability Director

$211k - $317k

RADEMACHER GERÄTE-ELEKTRONIK GmbH

THE POSITION
Our roster has an opening with your name on it

As the Infrastructure Engineering Director, you'll own FanDuel's foundational cloud and outpost-based infrastructure layer. You'll establish operational excellence, standardization, and reliability across compute, networking, Kubernetes, and cloud provisioning that everything else runs on. This is a hands-on technical leadership role where you set the standard for how infrastructure operates, not just keeping systems running but making them run predictably, with clear processes, documented procedures, and a culture that values doing things right.

This role lands at a critical juncture. FanDuel is scaling rapidly across new markets and cloud footprints. Your infrastructure foundation must handle that growth without creating operational chaos or cost surprises. You'll drive standardization (provisioning, observability, disaster recovery, cost attribution), establish discipline (operational runbooks, change management, incident response), and build an infrastructure team that operates with rigor and purpose. You'll manage vendor relationships, own cloud spend and capacity planning, and balance cost optimization with reliability and resilience.

By connecting infrastructure behavior to real business impact, you'll drive improvements in reliability, performance, cost, and operational maturity. You'll partner closely with Platform Engineering leadership and product teams to ensure our foundation actually supports what they need to build.

In addition to the specific responsibilities outlined above, employees may be required to perform other such duties as assigned by the Company. This ensures operational flexibility and allows the Company to meet evolving business needs.

THE GAME PLAN
Everyone on our team has a part to play

  • Establishing infrastructure strategy and standards across cloud provisioning, observability, disaster recovery, and cost governance. Standardization removes ambiguity and makes operations safer and faster.
  • Building and scaling the infrastructure engineering team with operational rigor and discipline. Hiring for people who care about documentation, reliability, and unglamorous work. Creating career paths and developing infrastructure engineers into senior and leadership roles.
  • Owning disaster recovery strategy, RTO/RPO definitions, testing, and organizational readiness. If DR hasn't been tested, it doesn't exist.
  • Establishing SLOs/SLIs for foundational infrastructure services. Making clear tradeoffs between cost optimization and reliability; these aren't opposing goals but they require transparency.
  • Designing and evolving cloud infrastructure for resilience, cost, and operational clarity. Managing AWS, outposts, Kubernetes orchestration, networking, mesh, and compute provisioning.
  • Setting standards for operational discipline: change management, incident response, post-mortems, and continuous improvement. Protecting infrastructure engineers from burnout while celebrating unglamorous work.
  • Managing vendor relationships, cloud costs, and capacity planning. Owning the business relationship with infrastructure: negotiating contracts, optimizing spend, planning cloud adoption roadmap.
  • Leading complex production incident response across infrastructure domains: Kubernetes, networking, DNS, load balancing, service mesh, cloud services.
  • Stewarding infrastructure platforms operating under strict regulatory and compliance requirements. Ensuring secure operation, auditability, and disciplined change management.
  • Partnering with engineering and product leadership to influence platform direction, roadmap priorities, and long-term infrastructure strategy.
  • Evaluating build-versus-buy decisions and vendor capabilities against infrastructure requirements and operational sustainability goals.
  • Mentoring engineers and raising infrastructure, platform networking, and operational maturity across the organization.

A Sneak Peek Into Our Tech Stack

AWS (including Outposts), Kubernetes (EKS, 100+ clusters) , Kong Gateway & Kong Mesh, Terraform, Helm and Datadog

THE STATS
What we're looking for in our next teammate

  • Significant experience leading platform engineering, SRE, or infrastructure organizations.
  • Deep expertise in Kubernetes architecture, operations, and scaling. Working knowledge of container orchestration, cluster design, workload isolation, and multi-cluster management.
  • Strong hands-on understanding of AWS cloud infrastructure: VPCs, networking, IAM, EC2, EKS, cost management, and operating hybrid AWS environments including Outposts.
  • Experience defining and driving platform, reliability, or infrastructure strategy across teams, with the ability to influence technical direction and set standards.
  • Proven ability to establish operational discipline: documented runbooks, change management processes, incident response frameworks, and blameless post-mortem culture.
  • Experience designing and implementing disaster recovery, RTO/RPO definitions, failover testing, and resilience improvements across production systems.
  • Experience defining and implementing SLOs, SLIs, and alerting strategies for infrastructure services using user-centric and business-aligned metrics.
  • Strong software engineering fundamentals with proficiency in at least one modern programming language. Ability to build scalable tooling and automation.
  • Track record of driving large-scale operational improvements through automation, reducing organizational toil, and preventing recurring classes of issues.
  • Understanding of cloud cost drivers and the ability to optimize spend without sacrificing reliability, resilience, or team velocity.
  • Strong analytical and communication skills. Ability to influence technical and non-technical stakeholders. Can translate infrastructure constraints into business and customer impact.
  • A mindset of ownership, continuous improvement, and long-term platform sustainability. You know when standards are working and when they need to change.

Bonus

  • Experience with service mesh platforms (Kong Mesh, Istio, Linkerd, or Envoy-based systems).
  • Hands-on experience managing AWS Outposts or hybrid cloud infrastructure.
  • Experience operating infrastructure in regulated industries where network segmentation, auditability, security controls and uptime are critical.
  • CNCF Kubernetes certifications (CKA, CKS, CKAD).
  • Familiarity with observability platforms like Datadog, Prometheus, and tracing systems.

Don’t check all the boxes? That’s okay! We encourage you to still apply if you feel like you possess an adjacent skill set and are interested in learning more about this position.

ABOUT FANDUEL

FanDuel Group is the premier mobile gaming company in the United States and Canada. FanDuel Group consists of a portfolio of leading brands across mobile wagering including: America’s #1 Sportsbook, FanDuel Sportsbook; its leading iGaming platform, FanDuel Casino; the industry’s unquestioned leader in horse racing and advance-deposit wagering, FanDuel Racing; and its daily fantasy sports product.

In addition, FanDuel Group operates FanDuel TV, its broadly distributed linear cable television network and FanDuel TV+, its leading direct-to-consumer OTT platform. FanDuel Group has a presence across all 50 states, Canada, and Puerto Rico.

The company is based in New York with US offices in Los Angeles, Atlanta, and Jersey City, as well as global offices in Canada and Scotland. The company’s affiliates have offices worldwide, including in Ireland, Portugal, Romania, and Australia.

FanDuel Group is a subsidiary of Flutter Entertainment, the world's largest sports betting and gaming operator with a portfolio of globally recognized brands and traded on the New York Stock Exchange (NYSE: FLUT).

PLAYER BENEFITS
We treat our team right

We offer amazing benefits above and beyond the basics. We have an array of health plans to choose from (some as low as $0 per paycheck) that include programs for fertility and family planning, mental health support, and fitness benefits. We offer generous paid time off (PTO & sick leave), annual bonus and long-term incentive opportunities (based on performance), 401k with up to a 5% match, commuter benefits, pet insurance, and more - check out all our benefits here: FanDuel Total Rewards . *Benefits differ across location, role, and level.

FanDuel is an equal opportunities employer and we believe, as one of our principles states, “We are One Team!”. As such, we are committed to equal employment opportunity regardless of race, color, ethnicity, ancestry, religion, creed, sex, national origin, sexual orientation, age, citizenship status, marital status, disability, gender identity, gender expression, veteran status, or any other characteristic protected by state, local or federal law. We believe FanDuel is strongest and best able to compete if all employees feel valued, respected, and included.

FanDuel is committed to providing reasonable accommodations for qualified individuals with disabilities. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please email View email address on click.appcast.io .

The applicable salary range for this position is $211,000 - $317,000 USD, which is dependent on a variety of factors including relevant experience, location, business needs and market demand. This role may offer the following benefits: medical, vision, and dental insurance; life insurance; disability insurance; a 401(k) matching program; among other employee benefits. This role may also be eligible for short-term or long-term incentive compensation, including, but not limited to, cash bonuses and stock program participation. This role includes paid personal time off and 14 paid company holidays. FanDuel offers paid sick time in accordance with all applicable state and federal laws.

It is unlawful in Massachusetts to require or administer a lie detector test as a condition of employment or continued employment. An employer who violates this law shall be subject to criminal penalties and civil liability.

#LI-Hybrid

#J-18808-Ljbffr

Vacancy posted 12 hours ago
Similar jobs that could be interesting for youBased on the Cloud Infrastructure & Reliability Director in Atlanta, GA vacancy
  • $168k - $200k

     ...What We're Looking For We're looking for a Senior Site Reliability Engineer to join our Data & ML Platform team. You'll be at...  ...for someone who combines a strong SRE mindset with deep cloud infrastructure and data platform experience . You're comfortable operating... 
    Cloud
    Remote work

    Datavant

    Atlanta, GA
    2 days ago
  •  ...Join to apply for the Site Reliability Engineer role at Motion Recruitment Join to...  ...have previous experience in automation, infrastructure orchestration, configuration management...  ...~ Experience working with AWS cloud environments and on-prem, Linux and Windows... 
    Cloud
    Contract work
    Worldwide

    Motion Recruitment

    Atlanta, GA
    2 days ago
  •  ...Site Reliability Engineering (SRE) Architect Location: Atlanta, GA Duration: 12Months+ Extension Hourly Rate: Depending...  ...expertise in software engineering, distributed systems, cloud infrastructure, and SRE principles to influence technology choices,... 
    Cloud
    Hourly pay
    Permanent employment
    Contract work
    Local area
    Early shift

    ETHEREUM TECHNOLOGIES LLC

    Atlanta, GA
    1 day ago
  •  ...Senior Infrastructure Manager Join Starr, a global leader in commercial insurance with over...  ...is responsible for the operation, reliability, and modernization of enterprise infrastructure...  ..., backup, disaster recovery, and cloud-adjacent services. This is a hands-on leadership... 
    Cloud
    Worldwide

    Starr Companies

    Atlanta, GA
    5 days ago
  •  ...enterprise data center and core network infrastructure.This is a true player-coach role, ideal...  ...environments supporting enterprise, cloud, and external connectivity.Support Nexus...  ...observability, incident response, and service reliability.Required QualificationsBachelor's... 
    Cloud
    Full time
    Worldwide
    Flexible hours

    NCR

    Atlanta, GA
    2 days ago
  •  ...Head Of It Infrastructure Engineering & Ops The Company is one of the fastest growing and best...  ...to bring the organization to the public cloud. Amidst this, The Company is looking to...  ...that the bank runs a highly available, reliable, resilient, and scalable infrastructure... 
    Cloud
    Full time
    Contract work
    Temporary work

    Professional Recruiters

    Atlanta, GA
    2 days ago
  •  ...part of that transformation. As our GenAI Cloud Engineer, you will have the opportunity...  ...for overseeing the end to end GenAI Infrastructure across exploration phase, build, deploy...  ...workloads ensuring security, performance and reliability and building repeatable infrastructure... 
    Cloud
    Full time
    Work at office

    American International Group (AIG)

    Atlanta, GA
    1 day ago
  •  ...Technical Support Specialist In Site Reliability Engineering (Sre) Mandatory skills: Scripting and programming languages like Python, Java, Ruby. Cloud and infrastructure management – AWS, Google cloud and Azure is a plus- CI/CD Automation, Database Management. The... 
    Cloud

    Omni Inclusive

    Atlanta, GA
    2 days ago
  • $60 - $68 per hour

     ...Site Reliability Engineer Immediate need for a talented Site Reliability Engineer. This is a 12+ months contract opportunity with...  ...engineering software within an Amazon Web Services (AWS) cloud infrastructure or other prominent enterprise cloud provider. Knowledge... 
    Cloud
    Contract work
    Local area
    Immediate start

    Pyramid Corporation

    Atlanta, GA
    2 days ago
  •  ...Site Reliability Engineer We're looking for an experienced Site Reliability Engineer (...  ...scientists to build, automate, and maintain the infrastructure that powers our core platform—including...  ..., and operational readiness across our cloud infrastructure Drive post-incident... 
    Cloud

    Alembic Technologies

    Atlanta, GA
    2 days ago
  •  ...Position Objective: In this role, the Reliability Engineer will ensure the reliability, scalability...  ...and operational health of EDAV’s Azure cloud environment. Terraform is central to...  ..., build, review, and troubleshoot Infrastructure-as-Code for mission critical... 
    Cloud
    Permanent employment

    GAP Solutions, Inc. (GAPSI)

    Atlanta, GA
    1 day ago
  • $130k - $150k

     ...We are seeking a highly skilled Site Reliability Engineer (SRE) to join our team and help...  ...capabilities, and experience working with cloud-native technologies and platforms. This...  ..., and maintain scalable and reliable infrastructure using cloud-native technologies... 
    Cloud
    Full time
    Remote work

    Prestige Staffing

    Atlanta, GA
    3 days ago
  •  ...expertise. We deliver faster, smarter, more reliable insights to insurance carriers and...  ...and operational health of our AWS-hosted infrastructure. You'll lead incident response, build the...  ...Looking For   ~4+ years in SRE, DevOps, or cloud operations roles supporting production... 
    Cloud
    Full time
    Flexible hours

    Seek Now

    Atlanta, GA
    8 days ago
  •  ...Senior Site Reliability Engineer Inspire Brands is hiring two Senior Site Reliability...  ...reflect true service health, not just infrastructure metrics Incident Management...  ...knowledge and expertise in at least one major cloud platform (Azure, AWS, or GCP) Expertise... 
    Cloud
    Worldwide

    Inspire Brands Inc

    Atlanta, GA
    4 days ago
  •  ...Lead Engineer, Site Reliability Engineering Team As a lead engineer with Retail, Site...  ...team, you will be at the forefront of Cloud and Big Data technology. In this role you...  ...security and latency considerations for new infrastructure and service provisioning as appropriate... 
    Cloud

    Next Level Business Services, Inc.

    Atlanta, GA
    5 days ago
  •  ...possible. Job Summary Acrisure is seeking a Sr. Cloud Engineer based in our Infrastructure Operations group that is responsible for supporting enterprise...  ...systems, and business teams to deliver scalable, reliable, well-governed cloud services that align with... 
    Cloud
    Immediate start
    Flexible hours

    Acrisure LLC

    Atlanta, GA
    2 days ago
  •  ...artifacts in the back log from root cause analysis. • Evolve the cloud infrastructure ecosystem for our application suite by experimenting with...  ...critical application components. • 1+ Years in Site Reliability Engineering organization preferred • Overall 4-6years of... 
    Cloud
    Work experience placement

    3B Staffing LLC

    Atlanta, GA
    2 days ago
  • Cloud Engineer (AWS) Employment Type: Full-Time, Experienced We are seeking a Cloud Engineer (AWS) who will be responsible for supporting...  ...strong architectural skills to ensure availability, reliability, etc. Hands-on experience with AWS (Required) or other cloud services... 
    Cloud
    Full time
    Flexible hours
    Shift work

    Contact Government Services LLC

    Atlanta, GA
    6 days ago
  •  ...Senior Cloud Application ArchitectClient: State of Georgia – Department of Human ServicesDuration...  ...for analyzing and designing the infrastructure architecture and management of...  ...Elasticsearch to ensure optimal performance, reliability, and scalability.Infrastructure as Code... 
    Cloud
    Work at office
    Local area

    nLeague

    Atlanta, GA
    4 days ago
  • $98.18k - $115.5k

     ...analytics and reporting solutions from on-premises platforms to cloud-based technologies. The position serves as a trusted advisor...  ..., and operational support processes to improve platform reliability and user experience. Track production defects, enhancement requests... 
    Cloud
    Temporary work
    Work experience placement
    Local area
    3 days per week

    U.S. Bank

    Atlanta, GA
    18 hours ago
  •  ...design and operate FirstKey Homes’s enterprise network infrastructure across corporate locations, cloud environments and remote users. This is a hands-on...  ...completing work tasks. Dependability— Job requires being reliable, responsible, and dependable, and fulfilling... 
    Cloud
    Work at office
    Remote work

    FirstKey Homes LLC

    Atlanta, GA
    6 days ago
  •  ...initiatives. The ideal candidate brings strong technical depth in cloud-based data engineering along with a practical, solution-...  ...someone who enjoys turning complex data challenges into scalable and reliable platforms.Responsibilities:• Build and maintain robust data... 
    Cloud

    Robert Half

    Atlanta, GA
    1 day ago
  •  ...architecting and delivering end-to-end cloud analytics solutions. In a fully cloud-native...  ...monitoring to ensure our ML systems are reliable, secure, and scalable. You will be...  ...Databricks workspaces and its associated Azure infrastructure. You will manage and optimize data and... 
    Cloud
    Contract work
    Live in
    Remote work

    3B Staffing LLC

    Atlanta, GA
    1 day ago
  •  ...professional services company with leading capabilities in digital, cloud and security. Combining unmatched experience and specialized...  .... Visit us at . .We are seeking a Manager to join our Infrastructure Advisory & Transformation practice. In this role, you will help... 
    Cloud
    Full time
    Work experience placement
    Live in
    Work at office
    Local area

    Accenture

    Atlanta, GA
    1 day ago
  • We are seeking a highly experienced Senior Manager to join our Infrastructure Advisory & Transformation practice. In this role, you will...  ...development of infrastructure strategies, transformation roadmaps, cloud adoption initiatives, operating model redesigns, and large-... 
    Cloud
    Full time
    Work experience placement
    Live in
    Work at office
    Local area

    Accenture

    Atlanta, GA
    1 day ago
  •  ...pipelines and platforms using Python, Snowflake or Databricks, and distributed processing technologies. Develops reliable data models and integrations across cloud, database, streaming, and analytics environments. For applications and inquiries, contact:hirings@... 
    Cloud
    Contract work

    Openkyber

    Atlanta, GA
    2 days ago
  •  ...systems by building and maintaining the infrastructure, deployment workflows, and platform...  ...capabilities required to run Applied AI solutions reliably in production. This role focuses on...  ..., reliability, and cost efficiency in cloud environments. Collaborate with... 
    Cloud
    Casual work
    Remote work
    Flexible hours

    Speria

    Atlanta, GA
    2 days ago
  • $95k - $130k

     ...technical solutions that integrate relevant data sources to Oden’s cloud endpoints, including system configurations, data architecture,...  ...and debugging issues with performance, correctness, and reliability of these integrations. Comfortable with Google Cloud. ~ Technically... 
    Cloud
    Remote work
    Shift work

    Oden Technologies

    Atlanta, GA
    2 days ago
  • As the Senior Site Reliability Engineer, you will serve as a trusted technical resource responsible...  ...AI, HPC, Kubernetes, and enterprise infrastructure environments. This role transforms...  ...education to maintain expertise in cloud, infrastructure, AI, and platform technologies... 
    Cloud
    Work at office
    Immediate start
    Worldwide
    Shift work

    Wesco International

    Atlanta, GA
    8 hours ago
  •  ...maintain a balance of capacity, performance, and reliability Proactively drive lifecycle management initiatives to keep infrastructure platforms current and vendor supported...  ...switching, wireless, SD-WAN, security, and cloud-managed network operations Strong understanding... 
    Cloud

    Insight Global

    Atlanta, GA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Cloud Infrastructure & Reliability Director. Be the first to apply!