Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer

sporttrade inc

Site Reliability Engineer

Sporttrade operates a regulated sports betting exchange that runs the way a financial exchange does. A successful Site Reliability Engineer at Sporttrade will be an operations-minded self-starter who treats the exchange like the production trading system it is: sessions open and close on schedule, every trade is captured and reported, and incidents are resolved quickly and documented thoroughly. In this role, you will own the day-to-day health of the exchange, spanning cloud infrastructure and on-premises datacenters, and you will automate away the toil that comes with operating a market that never wants to miss an open. This role works closely with the DevOps, Technical Operations, and Engineering teams to keep a live, regulated marketplace fast, observable, and compliant.

Duties
  • Own the daily operation of the exchange trading lifecycle - market startup and shutdown, enabling and disabling trading, pre- and post-session sanity checks, and capture of settlement, clearing, and trade-reporting artifacts.
  • Participate in an on-call rotation for a live regulated marketplace; lead incident response, drive incidents to resolution, author postmortems, and turn one-off fixes into runbooks and automation
  • Operate and improve our observability stack (Datadog, Wazuh, Prometheus, Grafana) - dashboards, alert quality, SLOs, and reducing time-to-detection for market-impacting issues
  • Run and maintain hybrid infrastructure: Kubernetes clusters (GKE and EKS) with an Istio service mesh, AWS and GCP accounts, and exchange servers in geographically distributed on-premises datacenters
  • Automate infrastructure and operational procedures with Ansible, Terraform, and Jenkins pipelines, with secrets managed in HashiCorp Vault
  • Support the data platform behind the exchange: PostgreSQL (Cloud SQL), Kafka (Confluent Cloud) change-data-capture and streaming pipelines, Redis, and backup/restore and disaster-recovery procedures — including proving that backups actually restore
  • Support market maker and partner connectivity (site-to-site VPNs and datacenter cross-connects), as well as conformance testing and onboarding support for partners joining the exchange.
  • Contribute to ongoing process improvement and the establishment of new policy and procedure for monitoring, incident management, change control, and exchange operations
Your Portfolio
  • 5+ years of experience in a Site Reliability Engineering, DevOps, production engineering, or technical operations role supporting a 24/7 production system
  • Strong Linux fundamentals and scripting ability
  • Experience supporting and debugging Java applications in production - reading stack traces and thread dumps, working with JVM memory and garbage-collection behavior, and diagnosing service issues from logs and metrics
  • Solid working knowledge of TCP/IP networking — comfortable reasoning about connections, ports, routing, and firewalls to debug connectivity between exchange components, partners, and datacenters
  • Hands-on experience operating Kubernetes in production and managing infrastructure as code (Terraform, Ansible) with CI/CD pipelines (Jenkins or similar)
  • Experience with modern observability tooling (Datadog, Prometheus, Grafana, or equivalent) and a track record of being on-call for systems that matter
  • Working knowledge of SQL and relational databases (PostgreSQL preferred); experience with Kafka or other streaming platforms a plus
  • Self-starter who can deliver results with minimal guidance
  • Comfortable working independently and with a team
  • Excellent communication and organizational skills — especially written incident communication and documentation
  • Background and interest in trading, capital markets, exchange operations, or sports betting a plus; familiarity with exchange protocols a plus
  • Previous experience in a regulated industry (gaming, finance) a plus
  • Startup experience preferred but not required
Perks
  • Medical, Dental, and Vision Benefits: Company pays 100% Employee premium and 50% Spouse & Dependent premiums
  • Short- & Long-Term Disability; Group Term Life and AD&D Voluntary Life and AD&D
  • 401(k) Plan
  • Equity Options
  • Flexible time off
  • MacBooks issued to all employees

Diverse workforces create the best companies, and we at Sporttrade are committed to an inclusive culture that celebrates the uniqueness and contributions of each individual. Sporttrade is an equal opportunity employer, and does not discriminate on the basis of sex, race, religion, national origin, disability status, protected veteran status, or any other characteristic protected by law.

Vacancy posted 5 days ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in United States vacancy
  • About the RoleWe’re looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability, observability, and operational excellence at the core. You’ll partner with engineers and data scientists to build, automate, and maintain the infrastructure... 
    Suggested

    Alembic

    San Francisco, CA
    5 days ago
  • $148.5k - $223.9k

     ...right place! Agentforce is the future of AI, and you are the future of Salesforce.Salesforce is seeking a senior engineering candidate to join the Site Reliability organization in San Francisco. Working closely with counterparts in the Infrastructure and R&D organizations,... 
    Suggested
    Full time
    Worldwide
    Weekend work

    Salesforce

    San Francisco, CA
    4 days ago
  • $104.9k - $174.7k

     ...Data Management. You can learn more about LexisNexis Risk at the link below, About the Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the reliability of our production systems. This is not a purely advisory... 
    Suggested
    Full time
    Work at office
    Local area
    Remote work
    Work from home

    RELX Group

    Atlanta, GA
    1 day ago
  • $167.7k - $245.2k

     ...very effective.We’re looking for talented engineers with a software or operations background...  ...development teams to ensure the reliability, performance and security of our infrastructure...  ...insurance. Please see the Cisco careers site to discover more benefits and perks.... 
    Suggested
    Full time
    Temporary work
    Work at office
    Local area
    Flexible hours
    1 day per week

    CISCO Systems

    New York, NY
    4 days ago
  • $182.8k - $247.3k

     ...mission to develop education for our half a billion (and growing!) learners around the world.About the role...As a Senior Site Reliability Engineer, you will work closely with both product and platform engineering teams to ensure Duolingo’s sophisticated distributed systems... 
    Suggested
    Work experience placement

    Duolingo

    Pittsburgh, PA
    2 days ago
  • $160k - $180k

     ...expertise, and world-class customer satisfaction. The Platform Engineering group at CentralReach builds the underlying technologies that...  ...in Software Engineering to drive adoption of modern reliability practices like SLOs, error budget policies, actionable alerts... 
    Full time
    Worldwide

    CentralReach

    Holmdel, NJ
    22 hours ago
  •  ...and responsible for ensuring the availability, scalability, and reliability of systems and applications.What will be your responsibilities...  ...using tools like Terraform or CloudFormation.Mentor junior engineers and provide technical guidance.Stay up-to-date with industry trends... 
    Work at office
    Remote work

    Interactive Brokers

    Greenwich, CT
    4 days ago
  • Qualifications: 8+ years of Software Engineering experience, or equivalent demonstrated through...  ...implement and maintain scalable and reliable infrastructure on Google Cloud Platform...  ...vendor resources Willingness to work on-site at stated location in the job openingDepartment... 
    Contract work
    For contractors
    Work experience placement

    Cedent Consulting

    Dallas, TX
    6 hours ago
  • $230k - $250k

     ...network. It's the foundation for autonomous networking, giving engineers and AI agents the ability to know the impact of every change...  ...how things have always been done.Forward is looking for a Site Reliability EngineerAbout the Role This is not a "keep the lights on"... 
    Night shift

    Forward Networks

    Santa Clara, CA
    6 hours ago
  • $158.5k - $172k

     ...exceptional value they deserve.About The OpportunityAs a Senior Engineer on the Runtime Automation team, you will design, automate, and...  .... This is a high-impact position driving continuous reliability, deep system optimization, and automation across our entire technology... 
    Full time
    Temporary work
    Work at office
    Flexible hours
    3 days per week

    GrubHub

    New York, NY
    6 hours ago
  • Senior Site Reliability EngineerLocation: Exton or Philadelphia, PA (Hybrid - 3 times a week in-office)Position SummaryAre you ready to start...  ...looking for you!We are looking for a Senior Site Reliability Engineer to take on the responsibility of automating cloud-based... 
    Casual work
    Work at office
    Worldwide

    Bentley Systems

    Philadelphia, PA
    3 days ago
  • We are looking for an experienced Site Reliability Engineer (SRE) to strengthen observability and operational resilience across a Microsoft Azure environment. This long-term Contract role will work closely with DevOps and engineering teams to establish monitoring standards... 
    Long term contract

    Robert Half

    Maumee, OH
    1 day ago
  • The Senior Site Reliability Engineer is responsible for improving the reliability, availability, scalability, and operational excellence of our critical infrastructure platforms and services. This role partners closely with Engineering, Security, and Infrastructure teams... 
    Full time
    Work at office
    Local area

    Castleton Commodities International

    Stamford, CT
    5 days ago
  •  ...and foster a dynamic work environment where new ideas thrive. Are you ready to join our team and make an impact?As a Senior Site Reliability Engineer at TeamViewer, you’ll be a key player in ensuring the reliability, scalability, and performance of our Azure-based SaaS... 
    Temporary work
    Casual work
    Worldwide

    TeamViewer

    Austin, TX
    1 day ago
  • $143k - $191k

     ...mission critical capabilities to our customers. System Deployment Engineers work in complex environments with shared environmental...  ...customersParticipate in customer demonstrations and exercisesWork with site reliability engineers to provide and refine requirements for tooling and... 
    Full time
    Temporary work
    Work experience placement
    Immediate start

    Anduril Industries

    Seattle, WA
    6 hours ago
  •  ...importance of in-office collaboration and fully intend for the selected candidate for this role to work on site in the specified location(s). As a Senior Site Reliability Engineer within the CET SAvE organization, you will play a critical leadership role advancing the... 
    Full time
    Work at office

    The Charles Schwab Corporation

    Southlake, TX
    3 hours ago
  •  ...selected candidate for this role to work on site in the specified location(s).The Client...  ...team is responsible for ensuring the reliability, scalability, and operational excellence...  ...around the clock. As a Site Reliability Engineer, you will partner across application engineering... 
    Full time
    Work at office

    The Charles Schwab Corporation

    Austin, TX
    22 hours ago
  •  ...and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the IP, you will solve complex and broad business problems with simple and straightforward solutions... 

    JP Morgan Chase

    Plano, TX
    4 days ago
  • $141k - $216.6k

     ...—it means helping shape the future of emergency response and building a safer, more connected world.Position OverviewAs a Site Reliability Engineer, you'll own the reliability, observability, and operational excellence of our Unified Call (UC) platform—the mission-critical... 
    Work experience placement
    Work at office

    Axon

    New York, NY
    3 days ago
  •  ...Fluency: English (Required)Work Shift:1st shift (United States of America)Please review the following job description:The Site Reliability Engineer role focuses on enhancing the reliability and operational excellence of enterprise platforms across hybrid cloud and on-premises... 
    Permanent employment
    Full time
    Part time
    H1b
    Work at office
    Local area
    Immediate start
    Work visa
    Monday to Friday
    Shift work
    Day shift

    Truist

    Raleigh, NC
    5 days ago
  • $45 - $85 per hour

    DescriptionThe Resy Site Reliability Engineering groups goal is to ensure Resy Customers can always use the service reliably. We're looking for engineers to be part of an empowered, self-organizing group, with the opportunity to use modern languages and tools and to operate... 
    Contract work
    Temporary work

    TEKsystems

    Phoenix, AZ
    5 days ago
  • Job ID: 28091697Reference Number: 26-00413Title: Site Reliability EngineerLocation: Iselin, NJ, 08830Posted Date: 2026-04-27Contact: Deepak...  ...Phone: (***) ***-****Company: HAN Staffing As a Site Reliability Engineer at JPMorgan Chase within the Commercial & Investment Banking,... 

    HAN Staffing

    Iselin, NJ
    1 day ago
  •  ...importance of in-office collaboration and fully intend for the selected candidate for this role to work on site in the specified location(s).As a Site Reliability Engineer supporting the Cashiering organization, you will play a critical role in ensuring the stability,... 
    Full time
    Work at office

    The Charles Schwab Corporation

    Southlake, TX
    22 hours ago
  •  ...and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Corporate and Investment Banking team, you will solve complex and broad business problems... 

    JP Morgan Chase

    Orem, UT
    5 days ago
  •  ...and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Commercial and Investment Bank, Healthcare Payments team, you will solve complex and broad... 

    JP Morgan Chase

    Irvine, CA
    3 days ago
  •  ...Evaluate applications, platforms, and vendors to assess resiliency, reliability, and operational risk.Design and implement processes that...  ...and reliability tooling.Actively participate in reliability engineering and resilience communities of practice, contributing to... 
    Full time

    Vanguard

    Dallas, TX
    3 days ago
  • $117k - $209.33k

    Job Requisition ID #26WD99276Position OverviewWant to help make a better world? As a Senior Site Reliability Engineer at Autodesk, you can help us build and operate reliable, secure, and scalable cloud services for Autodesk GovCloud products.As part of a new SRE team supporting... 
    Full time
    For contractors
    Remote work

    Autodesk

    Plano, TX
    1 day ago
  • $145.7k - $218.5k

     ...synonymous with entertainment excellence and creativity.Service Reliability EngineerDo you want to use transformative technologies to...  ...scalability and efficiency? Do you want a career that combines your engineering skills and your passion for video gaming? Are you fascinated... 
    Work experience placement
    Shift work

    Sony Interactive Entertainment America

    Aliso Viejo, CA
    1 day ago
  • $100k - $120k

    OverviewThe Site Reliability Engineer is a key force behind improving Origami’s time to resolution and advancing overall site reliability and scalability. This person participates in efforts to identify root causes during post-incident investigations, while also identifying... 
    Full time
    Temporary work
    Work experience placement
    Flexible hours

    Origami Risk

    Chicago, IL
    4 days ago
  •  ...English (Required)Work Shift:1st Shift (United States of America)Please review the following job description:Lead Site Reliability & Environment Monitoring Engineer (Azure / Dynatrace / ServiceNow)We are seeking a Lead Site Reliability & Environment Monitoring Engineer to... 
    Full time
    Temporary work
    Shift work
    Day shift

    TIH

    Charlotte, NC
    5 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!