Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer

sporttrade inc

Site Reliability Engineer

Sporttrade operates a regulated sports betting exchange that runs the way a financial exchange does. A successful Site Reliability Engineer at Sporttrade will be an operations-minded self-starter who treats the exchange like the production trading system it is: sessions open and close on schedule, every trade is captured and reported, and incidents are resolved quickly and documented thoroughly. In this role, you will own the day-to-day health of the exchange, spanning cloud infrastructure and on-premises datacenters, and you will automate away the toil that comes with operating a market that never wants to miss an open. This role works closely with the DevOps, Technical Operations, and Engineering teams to keep a live, regulated marketplace fast, observable, and compliant.

Duties
  1. Own the daily operation of the exchange trading lifecycle - market startup and shutdown, enabling and disabling trading, pre- and post-session sanity checks, and capture of settlement, clearing, and trade-reporting artifacts.
  2. Participate in an on-call rotation for a live regulated marketplace; lead incident response, drive incidents to resolution, author postmortems, and turn one-off fixes into runbooks and automation
  3. Operate and improve our observability stack (Datadog, Wazuh, Prometheus, Grafana) - dashboards, alert quality, SLOs, and reducing time-to-detection for market-impacting issues
  4. Run and maintain hybrid infrastructure: Kubernetes clusters (GKE and EKS) with an Istio service mesh, AWS and GCP accounts, and exchange servers in geographically distributed on-premises datacenters
  5. Automate infrastructure and operational procedures with Ansible, Terraform, and Jenkins pipelines, with secrets managed in HashiCorp Vault
  6. Support the data platform behind the exchange: PostgreSQL (Cloud SQL), Kafka (Confluent Cloud) change-data-capture and streaming pipelines, Redis, and backup/restore and disaster-recovery procedures — including proving that backups actually restore
  7. Support market maker and partner connectivity (site-to-site VPNs and datacenter cross-connects), as well as conformance testing and onboarding support for partners joining the exchange.
  8. Contribute to ongoing process improvement and the establishment of new policy and procedure for monitoring, incident management, change control, and exchange operations
Your Portfolio
  • 5+ years of experience in a Site Reliability Engineering, DevOps, production engineering, or technical operations role supporting a 24/7 production system
  • Strong Linux fundamentals and scripting ability
  • Experience supporting and debugging Java applications in production - reading stack traces and thread dumps, working with JVM memory and garbage-collection behavior, and diagnosing service issues from logs and metrics
  • Solid working knowledge of TCP/IP networking — comfortable reasoning about connections, ports, routing, and firewalls to debug connectivity between exchange components, partners, and datacenters
  • Hands-on experience operating Kubernetes in production and managing infrastructure as code (Terraform, Ansible) with CI/CD pipelines (Jenkins or similar)
  • Experience with modern observability tooling (Datadog, Prometheus, Grafana, or equivalent) and a track record of being on-call for systems that matter
  • Working knowledge of SQL and relational databases (PostgreSQL preferred); experience with Kafka or other streaming platforms a plus
  • Self-starter who can deliver results with minimal guidance
  • Comfortable working independently and with a team
  • Excellent communication and organizational skills — especially written incident communication and documentation
  • Background and interest in trading, capital markets, exchange operations, or sports betting a plus; familiarity with exchange protocols a plus
  • Previous experience in a regulated industry (gaming, finance) a plus
  • Startup experience preferred but not required
Perks
  • Medical, Dental, and Vision Benefits: Company pays 100% Employee premium and 50% Spouse & Dependent premiums
  • Short- & Long-Term Disability; Group Term Life and AD&D Voluntary Life and AD&D
  • 401(k) Plan
  • Equity Options
  • Flexible time off
  • MacBooks issued to all employees

Diverse workforces create the best companies, and we at Sporttrade are committed to an inclusive culture that celebrates the uniqueness and contributions of each individual. Sporttrade is an equal opportunity employer, and does not discriminate on the basis of sex, race, religion, national origin, disability status, protected veteran status, or any other characteristic protected by law.

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in United States vacancy
  • $165k - $225.6k

     ...core infrastructure to enterprise platforms, we partner across functions to drive scale, reliability, and innovation through technology. THE SENIOR SITE RELIABILITY ENGINEER OPPORTUNITY Reporting to the Manager, Site Reliability Engineering, this role will help... 
    Suggested
    Permanent employment
    Full time
    Local area
    Worldwide
    Flexible hours

    Okta

    San Francisco, CA
    1 day ago
  • $133.6k - $183.7k

     ...transforming our infrastructure to support a more modern, containerized, and highly scalable architecture. We’re seeking a Sr. Site Reliability Engineer (SRE) to help lead that transformation. This role will play a critical part in evolving our platform from legacy Azure-... 
    Suggested
    Full time
    Local area
    Flexible hours

    Synapse Health

    United States
    5 hours ago
  • $130k - $145k

     ...and your desire to team up with some of the best and brightest in technology and entertainment. The Role The Site Reliability Engineer (SRE) II is responsible for designing, implementing, and maintaining scalable and reliable systems and applications. Focus... 
    Suggested
    Full time
    Local area
    Worldwide
    Flexible hours

    AXS

    Los Angeles, CA
    12 hours ago
  • $130k - $180k

     ...of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference...  ...and AI R&D. THE ROLE Nebius is looking for a Site Reliability Engineer in Hardware Infrastructure team. You’re welcome to... 
    Suggested
    Full time
    Temporary work
    Work at office
    Immediate start
    Remote work

    Nebius

    United States
    4 days ago
  • $105.6k - $145.2k

    ARCHITECT THE FUTURE AS OUR SITE RELIABILITY ENGINEER! Are you ready to take your skills to the next level as a self-motivated and enthusiastic Site Reliability Engineer with hands-on experience supporting multiple connected Cloud-based products? Trimble is a... 
    Suggested
    Full time
    Work at office
    Local area
    Worldwide

    Trimble Inc.

    Westminster, CO
    4 days ago
  • $120k - $175k

     ...of sports fandom. Ready to reimagine the DFS industry together? We are seeking a highly skilled and experienced Senior Site Reliability Engineer to join our team. We are passionate about delivering cutting-edge solutions and pushing the boundaries of what's possible.... 
    Full time
    Remote work
    Work visa
    Flexible hours

    PrizePicks

    United States
    1 day ago
  • $100k - $180k

     ...Site Reliability Engineer (SRE) - Remote Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. This is a fantastic opportunity to join an established and... 
    Full time
    H1b
    Local area
    Immediate start
    Remote work
    Visa sponsorship

    Bright Vision Technologies

    United States
    5 hours ago
  • $62k - $141k

    Site Reliability Engineer The Opportunity: Engineering to make a system more resilient and efficient frees up time and money to build more capabilities. Whether you come from a background in network engineering, systems administration, or software development, if you... 
    Full time
    Contract work
    Part time
    Work at office
    Local area
    Remote work

    Booz Allen Hamilton

    Aurora, CO
    1 day ago
  • $118.6k - $195.68k

    About the Job The Red Hat IT OpenShift team is looking for a Senior Site Reliability Engineer (SRE) to design, develop, scale, and operate our Red Hat Hybrid OpenShift Platforms (on-prem & cloud). As a Senior Engineer, you will contribute to running Red Hat OpenShift at... 
    Permanent employment
    Full time
    Contract work
    Work experience placement
    Work at office
    Remote work
    Flexible hours

    Red River

    North Carolina
    1 day ago
  • $121.4k

     ...will be responsible for ensuring best-in-class uptime and reliability of our AI hardware infrastructure offerings. Partner...  ...indicators and defend them when they are breached. As a Senior Site Reliability Engineer, you will be responsible for: * Developing and scaling... 
    Full time
    Work experience placement
    Work at office

    Akamai

    United States
    2 days ago
  • About the Role We are seeking a Senior Site Reliability Engineer to join our cloud engineering team. You will own the reliability, scalability, and observability of our critical financial SaaS applications and infrastructure, working across cloud platforms to ensure our... 
    Full time

    MeridianLink

    United States
    5 hours ago
  • $189k - $283.6k

     ...the SRE team, you will proactively and reactively improve the reliability of Block's platform and critical infrastructure. You are metrics...  ...of accountability * A strong desire to perform and grow as an engineer * 5+ years of software development experience Technologies... 
    Full time
    Local area
    Remote work
    Relocation package
    Flexible hours
    Shift work

    Block

    California
    12 hours ago
  • $151.5k - $252.5k

     ...artifacts rather than getting direct access to environments from day one. This is a ground-up role — you'll help define how reliability engineering works here by mapping systems, writing runbooks, setting baselines, and building the practices this team will run on going... 
    Base plus commission
    Full time
    Local area
    Remote work
    Worldwide

    Veeam Software

    United States
    2 days ago
  • $76k - $127k

     ...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Site Reliability Engineer II The BizOps team at Mastercard is looking for a Site Reliability Engineer (SRE) who thrives on solving complex... 
    Full time
    Part time
    Worldwide
    Flexible hours
    Early shift

    Mastercard

    O Fallon, MO
    2 days ago
  • Job title: Site Reliability Engineer (SRE) Bill rate: $52/hr W2 Client address: 2900 W Plano Pkwy Plano, TX 75075 - Role is hybrid (3 days/wk) Years of experience required: 11+ Mandatory skills: Azure DevOps (ADO), GitHub & GitHub Actions, JFrog Artifactory Site... 
    Full time

    IPolarity LLC

    Hanover, PA
    5 hours ago
  • $140k - $165k

     ...deployed services and infrastructure components, empowering cloud engineering teams to move fast without sacrificing stability is essential...  ...CI/CD pipelines to ensure cloud software changes are deployed reliably and efficiently. Own and manage developer self-service... 
    Full time
    Remote work

    Flock

    United States
    12 hours ago
  • The Senior Site Reliability Engineer is responsible for improving the reliability, availability, scalability, and operational excellence of our critical infrastructure platforms and services. This role partners closely with Engineering, Security, and Infrastructure teams... 
    Full time
    Work at office
    Local area

    Castleton Commodities International

    Houston, TX
    2 days ago
  • $75k - $150k

     ...culture where all of our employees feel respected, valued and have an opportunity to contribute to the company’s success. As a Site Reliability Engineer within PNC's Technology organization, you can be based in Pittsburgh PA, Strongsville OH, Birmingham AL, Denver CO,... 
    Full time
    Temporary work
    Part time
    Work experience placement
    Work at office

    The PNC Financial Services Group

    Strongsville, OH
    4 days ago
  • $133k - $190k

    Site Reliability Engineer needed for a full time opportunity with SOC's direct client based in Herndon, VA. Direct Hire Role **Due to federal requirements, candidates must hold and possess an Active DOW TS/SCI security clearance to be considered for this role.** SOC is... 
    Full time

    SOC Support Services

    McLean, VA
    4 days ago
  • Site Reliability Engineers are responsible for ensuring the availability, reliability, scalability, and performance of the firm’s most critical customer-facing microservices that power all eCommerce channels. This role applies Google-inspired SRE principles to balance... 
    Local area
    Remote work
    Flexible hours
    Shift work

    O'Reilly Auto Parts

    Springfield, MO
    3 days ago
  • $152.6k - $191.5k

     ...is responsible for partnering with leaders across engineering and technology to define objective reliability goals for services. Key responsibilities include composing...  ...improvement.Position Summary:The Senior GCP Site Reliability Engineer acts as an advanced senior individual... 
    Full time
    Work at office
    Day shift

    Bank of America

    Plano, TX
    8 hours ago
  • $185k - $230k

    As a Sr. Site Reliability Engineer (SRE) III, you’ll work as part of a collaborative and high-performing team providing your expertise to deliver technical solutions within the highest levels of the federal government.We know that you can’t have great technology services... 
    Full time
    Local area
    Immediate start

    MetroStar Systems

    Washington DC
    1 day ago
  • $158.5k - $172k

     ...exceptional value they deserve.About The OpportunityAs a Senior Engineer on the Runtime Automation team, you will design, automate, and...  .... This is a high-impact position driving continuous reliability, deep system optimization, and automation across our entire technology... 
    Full time
    Work at office
    3 days per week

    GrubHub

    Chicago, IL
    1 day ago
  • $165k - $241.4k

     ...very effective.We’re looking for talented engineers with a software or operations background...  ...development teams to ensure the reliability, performance and security of our infrastructure...  ...insurance. Please see the Cisco careers site to discover more benefits and perks.... 
    Full time
    Temporary work
    Work at office
    Local area
    Flexible hours
    1 day per week

    CISCO Systems

    Austin, TX
    1 day ago
  • $165k - $270k

     ...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARLINK)At SpaceX we’re leveraging our experience in building rockets and spacecraft to deploy Starlink, the world’s most... 
    Permanent employment
    Temporary work
    Worldwide
    Weekend work

    SpaceX

    Redmond, WA
    2 days ago
  •  ...home!Where you’ll be:This position will be based at our Corporate Headquarters located in Charlotte, NC.About the Role:The Site Reliability Engineer plays a critical role in designing, building, and maintaining scalable, secure, and highly available cloud infrastructure... 
    Full time
    Flexible hours

    Electrolux

    Charlotte, NC
    3 days ago
  • $152k - $241.5k

     ...infrastructure platforms for automated host lifecycle management, fleet reliability/auto-healing, E2E observability or data-driven operations (...  ...languages such as Python, Go, Perl, or Ruby.Mentored other engineers and influenced technical direction through design reviews,... 
    Full time

    Nvidia

    Durham, NC
    2 days ago
  • $109.4k - $146.7k

     ...of business as well as other initiatives including MyDisneyExperience and Hey, Disney!This role sits within the Commerce Site Reliability Engineering organization in Technology & Digital for Disney Experiences. It works closely with leaders across Commerce and Consumer... 
    Worldwide

    Disney Interactive

    Orlando, FL
    3 days ago
  • $80k - $133k

     ...degree, Four (4) years additional experience will be needed.Minimum Four (4) years of experience in IT administration, software engineering, or platform engineering, with a focus on AWS cloud infrastructure and enterprise systems.One(1)+ years of experience deploying and... 
    Permanent employment
    Full time
    Contract work
    Remote work
    Flexible hours

    Guidehouse

    McLean, VA
    4 days ago
  •  ...functional teamsRequired Skill and ExperienceReliability Engineering· Support SLIs, SLOs, error budgets, and reliability KPIs.· Drive service availability, resiliency,...  ...Technical/Domain Skill 2Technology|DevOps|Site Reliability Engineering(SRE) Technical/Domain Skill... 
    Full time
    Temporary work
    Relocation

    Infosys Technologies

    San Antonio, TX
    8 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!