Senior Site Reliability Engineer
Gearbox Software
Who We AreAt 2K, we create some of the most iconic and culture-shaping video games in entertainment, including NBA 2K, one of the top-selling franchises in the world, and legendary titles like BioShock, Borderlands, Mafia, Sid Meier’s Civilization, and XCOM, as well as fan favorites WWE 2K, TopSpin, and PGA TOUR 2K. We build unforgettable experiences by pushing the boundaries of creativity, authenticity and innovation across every genre. Our portfolio is brought to life by some of the most influential game development studios in the world. Visual Concepts, Firaxis Games, Hangar 13, Cat Daddy Games, 31st Union, Cloud Chamber, Gearbox, HB Studios, and 2K SportsLab create world-class experiences across platforms. But what truly powers 2K is our people. We believe the best ideas come from teams that feel empowered, supported, and inspired. As an equal opportunity employer, we are committed to fostering a diverse, inclusive workplace where people are encouraged to come as they are and do their best work. The TeamThe 2K SRE team owns the infrastructure behind every player connection—All 2K game services, account platforms, CI/CD pipelines, and developer tooling spanning AWS, GCP, and on-premises data centers across multiple global regions. Global launch windows and live-service events push systems to their limits, and this team is expected to hold the line.Post-mortems here focus on systems, not people. Automation is the default answer to repetitive work. The infrastructure keeps millions of players connected, and the team takes that seriously!The RoleThe Senior SRE at 2K is a hands-on technical leader—shaping production infrastructure across multiple clouds and regions while partnering with network engineers, systems architects, and game studio developers. This is an ownership role: driving technical direction, influencing reliability from architecture review through production operation, and closing the gap between what engineering ships and what players experience.What You'll DoPlatform & InfrastructureDesign, build, and operate scalable multi-cloud and hybrid infrastructure using Terraform, Pulumi, and GitOps workflows (ArgoCD, Flux).Own Kubernetes platforms (EKS, GKE) end-to-end cluster lifecycle, multi-tenancy, networking (Istio, Cilium), and autoscaling.Push progressive delivery patterns (blue/green, canary) across game service deployments.Observability & ReliabilityBuild and run the full observability stack: Prometheus + Grafana + Datadog.Define SLI/SLO/error budget policies and build alerting that cuts through the noise.Lead chaos engineering exercises to surface failure modes before players encounter them.Drive incident response and post-mortems with a focus on systemic fixes and real follow-through.Automation, Security & Developer ExperienceEliminate toil through self-service provisioning, automated remediation, and intelligent scaling.Harden CI/CD pipelines (GitHub Actions, Jenkins, ArgoCD).Embed security at the platform layer through secrets management (PasswordState, 1Password, and AWS Secrets Manager) and policy-as-code (OPA/Gatekeeper).LeadershipPromote SRE practices across 2K studios through reliability reviews, runbooks, and embedded collaboration.Shape architectural decisions and author engineering RFCs that move the platform forward.Required QualificationsExperience: 5+ years in SRE, Platform Engineering, or equivalent infrastructure work at production scale.Kubernetes: Deep experience in cloud environments (EKS or GKE preferred), including networking, storage, and multi-cluster patterns.Infrastructure as Code (IaC): Strong proficiency with Terraform and/or Pulumi; hands-on with Helm, Terragrunt, and GitOps tooling (ArgoCD or GitHub Actions).Environments: Experience with modern and legacy tech, including AWS, GCP, VMware, and Bare metal servers.Configuration Management: Server configuration using Ansible, Puppet, and AWS Systems Manager.Observability: Experience with Datadog, Prometheus + Grafana, and OpenTelemetry; fluency in operationalizing SLI/SLO/error budgets inside engineering teams.Software Engineering: Production-quality code in Go, Python, or TypeScript for tools, automation, and internal libraries.Systems & Networking: Solid understanding of Linux internals, TCP/IP networking, DNS, and TLS proven enough to debug at the system level.Incident Management: Incident response and post-mortem leadership with a track record of systemic follow-through.Preferred QualificationsLive-service game or large-scale consumer internet experience dealing with millions of concurrent users.Deep knowledge of Service mesh (Istio, Cilium) and advanced Kubernetes networking.Experience with FinOps and managing resources efficiently at cloud scale.Experience with AI and Agentic Development.Cloud certifications (AWS Solutions Architect, GCP Professional Cloud Architect, CKA/CKS, or equivalent).Experience mentoring SREs or leading reliability working groups.As an equal opportunity employer, we are committed to ensuring that qualified individuals with disabilities are provided reasonable accommodation to participate in the job application or interview process, to perform their essential job functions, and to receive other benefits and privileges of employment. Please contact us if you need reasonable accommodation. Please note that 2K Games and its studios never uses instant messaging apps or personal email accounts to contact prospective employees or conduct interviews and when emailing, only use 2K.com accounts. #LI-Hybrid
- Job Description:About the Role: We are looking for a Senior SRE to join our Platform Engineering team as the operations owner of our observability platforms. You’ll be responsible for the reliability, scalability, and continued evolution of the tools that give our engineering...SeniorFull time
$127k - $249k
The TeamPlatform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions... ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper)....SeniorWork at officeLocal areaRemote workWorldwideFlexible hours- ...the rapidly evolving state of the art to engineer scalable, innovative, and research... ...business. About the Role:We are looking for a Senior SRE to serve as the operations owner for... ...tooling ecosystemsOwn the operational reliability of developer tooling ecosystems, including...SeniorFull timeLocal area
$152k - $241.5k
...artificial intelligence.We’re looking for a Senior SRE to join our Compute Farm team and... ...host lifecycle management, fleet reliability/auto-healing, E2E observability or data-... ...Python, Go, Perl, or Ruby.Mentored other engineers and influenced technical direction through...SeniorFull time$81.1k - $187k
.... You’ll partner with customer support, service owners, and engineering teams around the globe to ensure high-quality service for customers... ...posted.Career Level - IC3Escalation points for junior site reliability engineers during complex or high-impact incidents.Manage and...SeniorTemporary workMonday to FridayFlexible hoursShift workNight shift$165k - $241.4k
...very effective.We’re looking for talented engineers with a software or operations background... ...development teams to ensure the reliability, performance and security of our infrastructure... ...insurance. Please see the Cisco careers site to discover more benefits and perks....SeniorFull timeTemporary workWork at officeLocal areaFlexible hours1 day per week$127k - $249k
...Eastern or Central time zones. We are looking for an experienced Senior Engineer for our SRE, Atlas team to support, maintain and grow the... ...crucial workloads. Role OverviewWe are seeking a talented Site Reliability Engineer (SRE) with a strong infrastructure background....SeniorLocal areaRemote workWorldwideFlexible hours- ...Schwab. We are an integrated product, engineering, strategy and risk team, all based in San... ...how we serve our clients. As a Senior Engineer on AI.x, you will play a key role... ...areas of technology today.As a Senior AI Site Reliability Engineer you will support reliability efforts...SeniorFull time
$127k - $249k
We are looking for an experienced Senior or Staff Engineer for our SRE, InfraSec team, to guide the security of our cloud-based infrastructure. As a Staff SRE, you will be very hands-on technically while also mentoring a small team of SREs.The InfraSec team collaborates...SeniorLocal areaRemote workWorldwideFlexible hours$110.7k - $171.8k
...components Participation in oncall rotation as a platform reliability escalation point Incident response, postincident reviews... ..., and internal control requirements. Collaborate with engineering teams across the organization to influence platform adoption,...SeniorWork experience placementWork at officeLocal area$185k - $227k
...professionals. If the opportunity to build your career is compelling, read on for more details. ROLE AND RESPONSIBILITIES: A Senior Site Reliability Engineer (SRE) is expected to own the operational stability and performance ofJuul’s hybrid cloud infrastructure (Nutanix, AWS/...SeniorRemote work- Recognized as the No. 1 site trusted by real estate professionals, Realtor.com has been at the forefront of online real estate... ...confidence through expert guidance.We are seeking a Senior Site Reliability Engineer to join our newly formed Operations Excellence organization...SeniorWork at officeLocal area
- ...in-office collaboration and fully intend for the selected candidate for this role to work on site in the specified location(s).We are seeking a Kafka Site Reliability Engineer to help build, operate, and continuously improve Schwab's enterprise streaming platform ecosystem...SeniorFull timeWork at office
$100k - $125k
...a dynamic work environment where new ideas thrive. Are you ready to join our team and make an impact?ResponsibilitiesAs a Site Reliability Engineer at TeamViewer, you’ll be a key player in ensuring the reliability, scalability, and performance of our Azure-based SaaS platform...Temporary workCasual workWorldwide$98.58k - $138.02k
...Northern California / Silicon Valley Region / Denver, COProduct Engineering - DevOps /Full Time /HybridRestaurant365 is a SaaS company... ...office locations: Austin, TX; Irvine, CA; or Akron, OH. The Site Reliability Engineer II will be responsible for supporting, enhancing,...Full timeWork at office$196k - $269.5k
Senior Principal AI Agent EngineerThe Software Engineering team delivers next-generation software application enhancements and new products for a changing world. Working at the cutting edge, we design and develop software for platforms, peripherals, applications and diagnostics...Senior$167.18k - $203.61k
...remotely part of the weekTravel %NoWork ShiftJob DescriptionCox Automotive Corporate Services, LLCLEAD SITE RELIABILITY ENGINEERJob Description: Lead Site Reliability Engineer positions offered by Cox Automotive Corporate Services, LLC (Austin, Texas). Lead the...Full timeWork at officeRemote workFlexible hours$180k - $230k
...CAD, building the services and partner integrations that power reliable money movement across the region. This role will lead... ...Qualifications Key responsibilities Strong product mindset, connecting engineering decisions to customer outcomes, payout volume, speed, quality,...SeniorFull time$127k - $249k
...MongoDB, Inc. is seeking an experienced Senior or Staff Engineer for their SRE, InfraSec team, responsible for guiding the security of cloud-based infrastructure. The role involves hands-on technical work and mentorship of a small team while collaborating with engineering...Remote workFlexible hours$127k - $249k
...A leading technology company is seeking an experienced Senior or Staff Engineer for their SRE, InfraSec team in Austin. This role focuses on leading the design and implementation of security solutions for cloud platforms while mentoring a team. Candidates should have...- ...operational performance and availability of critical business platforms and cloud services. With a strong technical background in site reliability engineering, the ideal applicant will have excellent communication skills and a focus on continuous improvement through automation....
- ...METRIX IT SOLUTIONS INC is seeking a senior database engineer to design, deploy, and manage multi-region CockroachDB clusters in production. The role focuses on high availability, data consistency, and scalable capacity planning for global deployments. You will monitor...
$85 - $90 per hour
...00/hr - $90.00/hr Date Posted: 04/01/2025 Hiring Organization: Rose International Position Number: 480571 Job Title: Senior Principal Software Engineer Work Model: Onsite Employment Type: Temporary Min Hourly Rate($): 85.00 Max Hourly Rate($): 90.00 Job Description ***...SeniorHourly payTemporary workFlexible hours$184k - $287.5k
...to do their best work. Come join the team and see how you can make a lasting impact on the world.NVIDIA Nsight Compute helps CUDA engineers around the world to innovate in Artificial Intelligence (AI) and High Performance Computing. Join our team and help develop groundbreaking...SeniorFull time- Senior Principal Software Engineer (ServiceNow Information Architect)Be a part of a team that’s ensuring Dell Technologies' product integrity and customer satisfaction. Our IT Software Engineer team turns business requirements into technology solutions by designing, coding...Senior
$198k - $247.5k
...of the open source community, Cloudera advances digital transformation for the world’s largest enterprises.About Forward Deployed Engineering (FDE) at ClouderaCloudera’s Forward Deployed Engineering (FDE) function sits within the Applied AI organization, and is a...SeniorFull timeWork from homeRelocation$160k - $200k
...Founded by data scientists and engineers, Striveworks set out to make... ...building systems that remain reliable, adaptable, and ready to... ...environments. The Role As a Senior Software Engineer at... ...environment, or you can work hybrid/on site at our office in northwest...SeniorFull timeWork at officeRemote work$184k - $287.5k
...learning, supercomputing, gaming, and visualization. As a Senior System Software Engineer on the NvSci team, you will play an integral role in... ...tools, and generative AI technologies to improve software reliability, maintainability, and scalability.What we need to see:BS...SeniorFull timeRemote work$184k - $287.5k
.... Come join the team and see how you can make a lasting impact on the world!NVIDIA is searching for a highly motivated, technical engineer to join the Tegra system-on-chip (SoC) software organization. You will work on key aspects of our ARM SW ecosystem and system software...SeniorFull timeRemote work$152k - $241.5k
Join the NVIDIA Developer Tools team and empower engineers throughout the world developing innovative products in Automotive, VR, Gaming... ...compute profiler stack.Write fast, effective, maintainable, reliable, and well documented code.Work closely with internal and external...SeniorFull timeWorldwide
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- site reliability engineer Austin, TX
- site reliability engineer sre Austin, TX
- senior manufacturing manager Austin, TX
- senior business analyst Austin, TX
- senior cost estimator Austin, TX
- senior manager tax Austin, TX
- senior devops Austin, TX
- senior recruiter Austin, TX
- senior property manager Austin, TX
- senior paralegal Austin, TX

