Site Reliability Engineer
sporttrade inc
Site Reliability Engineer
Sporttrade operates a regulated sports betting exchange that runs the way a financial exchange does. A successful Site Reliability Engineer at Sporttrade will be an operations-minded self-starter who treats the exchange like the production trading system it is: sessions open and close on schedule, every trade is captured and reported, and incidents are resolved quickly and documented thoroughly. In this role, you will own the day-to-day health of the exchange, spanning cloud infrastructure and on-premises datacenters, and you will automate away the toil that comes with operating a market that never wants to miss an open. This role works closely with the DevOps, Technical Operations, and Engineering teams to keep a live, regulated marketplace fast, observable, and compliant.
Duties
- Own the daily operation of the exchange trading lifecycle - market startup and shutdown, enabling and disabling trading, pre- and post-session sanity checks, and capture of settlement, clearing, and trade-reporting artifacts.
- Participate in an on-call rotation for a live regulated marketplace; lead incident response, drive incidents to resolution, author postmortems, and turn one-off fixes into runbooks and automation
- Operate and improve our observability stack (Datadog, Wazuh, Prometheus, Grafana) - dashboards, alert quality, SLOs, and reducing time-to-detection for market-impacting issues
- Run and maintain hybrid infrastructure: Kubernetes clusters (GKE and EKS) with an Istio service mesh, AWS and GCP accounts, and exchange servers in geographically distributed on-premises datacenters
- Automate infrastructure and operational procedures with Ansible, Terraform, and Jenkins pipelines, with secrets managed in HashiCorp Vault
- Support the data platform behind the exchange: PostgreSQL (Cloud SQL), Kafka (Confluent Cloud) change-data-capture and streaming pipelines, Redis, and backup/restore and disaster-recovery procedures — including proving that backups actually restore
- Support market maker and partner connectivity (site-to-site VPNs and datacenter cross-connects), as well as conformance testing and onboarding support for partners joining the exchange.
- Contribute to ongoing process improvement and the establishment of new policy and procedure for monitoring, incident management, change control, and exchange operations
Your Portfolio
- 5+ years of experience in a Site Reliability Engineering, DevOps, production engineering, or technical operations role supporting a 24/7 production system
- Strong Linux fundamentals and scripting ability
- Experience supporting and debugging Java applications in production - reading stack traces and thread dumps, working with JVM memory and garbage-collection behavior, and diagnosing service issues from logs and metrics
- Solid working knowledge of TCP/IP networking — comfortable reasoning about connections, ports, routing, and firewalls to debug connectivity between exchange components, partners, and datacenters
- Hands-on experience operating Kubernetes in production and managing infrastructure as code (Terraform, Ansible) with CI/CD pipelines (Jenkins or similar)
- Experience with modern observability tooling (Datadog, Prometheus, Grafana, or equivalent) and a track record of being on-call for systems that matter
- Working knowledge of SQL and relational databases (PostgreSQL preferred); experience with Kafka or other streaming platforms a plus
- Self-starter who can deliver results with minimal guidance
- Comfortable working independently and with a team
- Excellent communication and organizational skills — especially written incident communication and documentation
- Background and interest in trading, capital markets, exchange operations, or sports betting a plus; familiarity with exchange protocols a plus
- Previous experience in a regulated industry (gaming, finance) a plus
- Startup experience preferred but not required
Perks
- Medical, Dental, and Vision Benefits: Company pays 100% Employee premium and 50% Spouse & Dependent premiums
- Short- & Long-Term Disability; Group Term Life and AD&D Voluntary Life and AD&D
- 401(k) Plan
- Equity Options
- Flexible time off
- MacBooks issued to all employees
Diverse workforces create the best companies, and we at Sporttrade are committed to an inclusive culture that celebrates the uniqueness and contributions of each individual. Sporttrade is an equal opportunity employer, and does not discriminate on the basis of sex, race, religion, national origin, disability status, protected veteran status, or any other characteristic protected by law.
$165k - $225.6k
...core infrastructure to enterprise platforms, we partner across functions to drive scale, reliability, and innovation through technology. THE SENIOR SITE RELIABILITY ENGINEER OPPORTUNITY Reporting to the Manager, Site Reliability Engineering, this role will help...SuggestedPermanent employmentFull timeLocal areaWorldwideFlexible hours$133.6k - $183.7k
...transforming our infrastructure to support a more modern, containerized, and highly scalable architecture. We’re seeking a Sr. Site Reliability Engineer (SRE) to help lead that transformation. This role will play a critical part in evolving our platform from legacy Azure-...SuggestedFull timeLocal areaFlexible hours$130k - $145k
...and your desire to team up with some of the best and brightest in technology and entertainment. The Role The Site Reliability Engineer (SRE) II is responsible for designing, implementing, and maintaining scalable and reliable systems and applications. Focus...SuggestedFull timeLocal areaWorldwideFlexible hours$130k - $180k
...of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference... ...and AI R&D. THE ROLE Nebius is looking for a Site Reliability Engineer in Hardware Infrastructure team. You’re welcome to...SuggestedFull timeTemporary workWork at officeImmediate startRemote work$105.6k - $145.2k
ARCHITECT THE FUTURE AS OUR SITE RELIABILITY ENGINEER! Are you ready to take your skills to the next level as a self-motivated and enthusiastic Site Reliability Engineer with hands-on experience supporting multiple connected Cloud-based products? Trimble is a...SuggestedFull timeWork at officeLocal areaWorldwide$120k - $175k
...of sports fandom. Ready to reimagine the DFS industry together? We are seeking a highly skilled and experienced Senior Site Reliability Engineer to join our team. We are passionate about delivering cutting-edge solutions and pushing the boundaries of what's possible....Full timeRemote workWork visaFlexible hours$100k - $180k
...Site Reliability Engineer (SRE) - Remote Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. This is a fantastic opportunity to join an established and...Full timeH1bLocal areaImmediate startRemote workVisa sponsorship$62k - $141k
Site Reliability Engineer The Opportunity: Engineering to make a system more resilient and efficient frees up time and money to build more capabilities. Whether you come from a background in network engineering, systems administration, or software development, if you...Full timeContract workPart timeWork at officeLocal areaRemote work$118.6k - $195.68k
About the Job The Red Hat IT OpenShift team is looking for a Senior Site Reliability Engineer (SRE) to design, develop, scale, and operate our Red Hat Hybrid OpenShift Platforms (on-prem & cloud). As a Senior Engineer, you will contribute to running Red Hat OpenShift at...Permanent employmentFull timeContract workWork experience placementWork at officeRemote workFlexible hours$121.4k
...will be responsible for ensuring best-in-class uptime and reliability of our AI hardware infrastructure offerings. Partner... ...indicators and defend them when they are breached. As a Senior Site Reliability Engineer, you will be responsible for: * Developing and scaling...Full timeWork experience placementWork at office- About the Role We are seeking a Senior Site Reliability Engineer to join our cloud engineering team. You will own the reliability, scalability, and observability of our critical financial SaaS applications and infrastructure, working across cloud platforms to ensure our...Full time
$189k - $283.6k
...the SRE team, you will proactively and reactively improve the reliability of Block's platform and critical infrastructure. You are metrics... ...of accountability * A strong desire to perform and grow as an engineer * 5+ years of software development experience Technologies...Full timeLocal areaRemote workRelocation packageFlexible hoursShift work$151.5k - $252.5k
...artifacts rather than getting direct access to environments from day one. This is a ground-up role — you'll help define how reliability engineering works here by mapping systems, writing runbooks, setting baselines, and building the practices this team will run on going...Base plus commissionFull timeLocal areaRemote workWorldwide$76k - $127k
...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Site Reliability Engineer II The BizOps team at Mastercard is looking for a Site Reliability Engineer (SRE) who thrives on solving complex...Full timePart timeWorldwideFlexible hoursEarly shift- Job title: Site Reliability Engineer (SRE) Bill rate: $52/hr W2 Client address: 2900 W Plano Pkwy Plano, TX 75075 - Role is hybrid (3 days/wk) Years of experience required: 11+ Mandatory skills: Azure DevOps (ADO), GitHub & GitHub Actions, JFrog Artifactory Site...Full time
$140k - $165k
...deployed services and infrastructure components, empowering cloud engineering teams to move fast without sacrificing stability is essential... ...CI/CD pipelines to ensure cloud software changes are deployed reliably and efficiently. Own and manage developer self-service...Full timeRemote work- The Senior Site Reliability Engineer is responsible for improving the reliability, availability, scalability, and operational excellence of our critical infrastructure platforms and services. This role partners closely with Engineering, Security, and Infrastructure teams...Full timeWork at officeLocal area
$75k - $150k
...culture where all of our employees feel respected, valued and have an opportunity to contribute to the company’s success. As a Site Reliability Engineer within PNC's Technology organization, you can be based in Pittsburgh PA, Strongsville OH, Birmingham AL, Denver CO,...Full timeTemporary workPart timeWork experience placementWork at office$133k - $190k
Site Reliability Engineer needed for a full time opportunity with SOC's direct client based in Herndon, VA. Direct Hire Role **Due to federal requirements, candidates must hold and possess an Active DOW TS/SCI security clearance to be considered for this role.** SOC is...Full time- Site Reliability Engineers are responsible for ensuring the availability, reliability, scalability, and performance of the firm’s most critical customer-facing microservices that power all eCommerce channels. This role applies Google-inspired SRE principles to balance...Local areaRemote workFlexible hoursShift work
$152.6k - $191.5k
...is responsible for partnering with leaders across engineering and technology to define objective reliability goals for services. Key responsibilities include composing... ...improvement.Position Summary:The Senior GCP Site Reliability Engineer acts as an advanced senior individual...Full timeWork at officeDay shift$185k - $230k
As a Sr. Site Reliability Engineer (SRE) III, you’ll work as part of a collaborative and high-performing team providing your expertise to deliver technical solutions within the highest levels of the federal government.We know that you can’t have great technology services...Full timeLocal areaImmediate start$158.5k - $172k
...exceptional value they deserve.About The OpportunityAs a Senior Engineer on the Runtime Automation team, you will design, automate, and... .... This is a high-impact position driving continuous reliability, deep system optimization, and automation across our entire technology...Full timeWork at office3 days per week$165k - $241.4k
...very effective.We’re looking for talented engineers with a software or operations background... ...development teams to ensure the reliability, performance and security of our infrastructure... ...insurance. Please see the Cisco careers site to discover more benefits and perks....Full timeTemporary workWork at officeLocal areaFlexible hours1 day per week$165k - $270k
...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARLINK)At SpaceX we’re leveraging our experience in building rockets and spacecraft to deploy Starlink, the world’s most...Permanent employmentTemporary workWorldwideWeekend work- ...home!Where you’ll be:This position will be based at our Corporate Headquarters located in Charlotte, NC.About the Role:The Site Reliability Engineer plays a critical role in designing, building, and maintaining scalable, secure, and highly available cloud infrastructure...Full timeFlexible hours
$152k - $241.5k
...infrastructure platforms for automated host lifecycle management, fleet reliability/auto-healing, E2E observability or data-driven operations (... ...languages such as Python, Go, Perl, or Ruby.Mentored other engineers and influenced technical direction through design reviews,...Full time$109.4k - $146.7k
...of business as well as other initiatives including MyDisneyExperience and Hey, Disney!This role sits within the Commerce Site Reliability Engineering organization in Technology & Digital for Disney Experiences. It works closely with leaders across Commerce and Consumer...Worldwide$80k - $133k
...degree, Four (4) years additional experience will be needed.Minimum Four (4) years of experience in IT administration, software engineering, or platform engineering, with a focus on AWS cloud infrastructure and enterprise systems.One(1)+ years of experience deploying and...Permanent employmentFull timeContract workRemote workFlexible hours- ...functional teamsRequired Skill and ExperienceReliability Engineering· Support SLIs, SLOs, error budgets, and reliability KPIs.· Drive service availability, resiliency,... ...Technical/Domain Skill 2Technology|DevOps|Site Reliability Engineering(SRE) Technical/Domain Skill...Full timeTemporary workRelocation
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- site reliability engineering manager United States
- site reliability engineer United States
- site reliability engineer sre United States
- site reliability engineer remote United States
- after school site coordinator United States
- site services specialist United States
- construction site safety United States
- site merchandiser United States
- site leader United States
- official site United States
