Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer, Infrastructure Platforms — AMER (Intermediate to Senior Staff)

$126.4k - $314.4k
Full-time

GitLab

GitLab is the intelligent orchestration platform for DevSecOps. GitLab enables organizations to increase developer productivity, improve operational efficiency, reduce security and compliance risk, and accelerate digital transformation. More than 50 million registered users and more than 50% of the Fortune 100* trust GitLab to ship better, more secure software faster.

The same principles built into our products are reflected in how our team works: we embrace AI as a core productivity multiplier, with all team members expected to incorporate AI into their daily workflows to drive efficiency, innovation, and impact. GitLab is where careers accelerate, innovation flourishes, and every voice is valued. Our high-performance culture is driven by our values and continuous knowledge exchange, enabling our team members to reach their full potential while collaborating with industry leaders to solve complex problems. Co-create the future with us as we build technology that transforms how the world develops software.

* Fortune 500® is a registered trademark of Fortune Media IP Limited, used under license. Claim based on GitLab data. Fortune 100 refers to the top 20% ranked companies in the 2025 Fortune 500 list, published in June 2025. Fortune and Fortune Media IP Limited are not affiliated with, and do not endorse products or services of GitLab.

An overview of this role

Site Reliability Engineers keep GitLab's user-facing services and production systems running reliably at scale. They combine software engineering with operational excellence, applying sound engineering principles, automation, and continuous improvement to build, operate, and evolve our production infrastructure.

This is a single application for Site Reliability Engineering opportunities across Infrastructure Platforms. Rather than asking you to choose the right team or level upfront, we evaluate your skills holistically and match you to the opportunity that best aligns with your experience and our hiring needs. We hire Site Reliability Engineers from Intermediate through Senior Staff across multiple Infrastructure Platforms teams.

We don't expect every candidate to have experience with every technology in our environment. We're looking for engineers with strong technical fundamentals, a growth mindset, and the ability to learn quickly. We'll support you in becoming successful with GitLab's tools, systems, and ways of working.

Please note: This position is open to candidates based in the United States and Canada only. Candidates based in the United Kingdom can apply to this posting: Site Reliability Engineer, Infrastructure Platforms — UK (Intermediate to Senior Staff)

How our SRE hiring works

Because this is a single application for SRE roles across Infrastructure Platforms, our process is built to evaluate you once and match you well, rather than interviewing separately for every team.

  • Recruiter Screen: A conversation about your background, what you're looking for, and the level and teams that fit, so we can point your process in the right direction.
  • Core Technical: The shared assessment every SRE candidate takes, regardless of eventual team. A low-stress, collaborative discussion covering system architecture and incident review.
  • Hiring Manager Interview: A conversation about ownership, judgment, execution, collaboration, and growth, the non-technical signals that make an SRE effective at GitLab.
  • Peer Technical: Team-specific depth, run by SREs from the team you're most likely to join, focused on the problems that team actually works on.
  • Skip-Level Interview: A conversation with a senior leader on values alignment, and how you'll work across teams.

After your interviews, we consider your performance alongside our current hiring needs to confirm the level and team where you'll do your best work. Interview results are a major factor, and final placement also reflects our active hiring priorities at the time.

We’ll calibrate your level throughout the interview process based on the scope and impact of your experience.

  • Intermediate: You independently deliver meaningful reliability improvements within a defined area.
  • Senior: You own complex reliability work end to end and raise the effectiveness of your team.
  • Staff: You shape reliability across multiple teams, solving systemic problems and creating approaches others can reuse.
  • Senior Staff: You set technical direction across a broader Infrastructure area and influence reliability strategy at organizational scale.

What you'll do

  • Keep user-facing services and production systems reliable, scalable, and efficient
  • Build automation and tooling that reduces toil and replaces manual work with repeatable, infrastructure-as-code-driven workflows
  • Operate and troubleshoot production systems on Kubernetes, including deployments, rollouts, and scaling
  • Write and maintain infrastructure as code, and ship changes safely through CI/CD and GitOps
  • Participate in on-call, triage alerts, follow and improve runbooks, and escalate appropriately
  • Contribute to the observability stack, using metrics, logs, and SLOs to detect symptoms early rather than just outages
  • Take part in incident response and post-incident reviews, turning learnings into changes in automation and process
  • Document runbooks, architecture decisions, and reviews so your findings become repeatable practices

What you'll bring

  • Experience keeping production systems reliable, combining an operations mindset with real software engineering practice
  • Experience building net-new infrastructure tooling and automation, not just configuring existing tools. For example, Terraform modules, Kubernetes operators or controllers, or production automation and services written from scratch
  • The ability to read, debug, and reason about code. Most of our teams work in Go; some work in Ruby. You can discuss a piece of code's behavior, performance, and failure modes
  • Experience with infrastructure as code, and with Kubernetes and its ecosystem, at a depth appropriate to your level
  • Hands-on experience with at least one major cloud provider (GCP or AWS)
  • Familiarity with observability practices, including metrics, logging, alerting, and SLOs or SLIs, and using data to inform operational decisions
  • Comfort participating in on-call and incident response, with a structured approach to troubleshooting under pressure
  • Strong written communication and the ability to operate as a manager-of-one in an async, distributed environment
  • A track record of using automation, and increasingly AI, to reduce toil and improve how you and your team work
  • Alignment with GitLab's values and a commitment to working in accordance with them

About the team

Infrastructure Platforms is responsible for the availability, reliability, performance, and scalability of GitLab’s user-facing services, most notably GitLab.com. The organization spans teams across Production Engineering, GitLab Dedicated, GitLab Delivery, and Developer Experience, covering everything from the production fleet and networking platform to observability, incident response, deployment infrastructure, tenant scale, and our single-tenant Dedicated offering.

We are a globally distributed, remote-first organization that works asynchronously, favors automation over toil, and uses monitoring, metrics, and clear ownership to continuously improve the reliability of GitLab at scale. For more on how we work, see the Infrastructure Handbook Page.

The base salary range for this role’s listed level is currently for residents of the United States only. This range is intended to reflect the role's base salary rate in locations throughout the US. Grade level and salary ranges are determined through interviews and a review of education, experience, knowledge, skills, abilities of the applicant, equity with other team members, alignment with market data, and geographic location. The base salary range does not include any bonuses, equity, or benefits. See more information on our benefits and equity. Sales roles are also eligible for incentive pay targeted at up to 100% of the offered base salary.

United States Salary Range

$126,400—$314,400 USD

How GitLab Supports Full-Time Employees

  • Benefits to support your health, finances, and well-being
  • Flexible Paid Time Off
  • Team Member Resource Groups
  • Equity Compensation & Employee Stock Purchase Plan
  • Growth and Development Fund
  • Parental Leave

Please note that we welcome interest from candidates with varying levels of experience; many successful candidates do not meet every single requirement. Additionally, studies have shown that people from underrepresented groups are less likely to apply to a job unless they meet every single qualification. If you're excited about this role, please apply and allow our recruiters to assess your application.

Country Hiring Guidelines: GitLab hires new team members in countries around the world. All of our roles are remote, however some roles may carry specific location-based eligibility requirements. Our Talent Acquisition team can help answer any questions about location after starting the recruiting process.

Privacy Policy: Please review our Recruitment Privacy Policy. Your privacy is important to us.

GitLab is proud to be an equal opportunity workplace and is an affirmative action employer. GitLab’s policies and practices relating to recruitment, employment, career development and advancement, promotion, and retirement are based solely on merit, regardless of race, color, religion, ancestry, sex (including pregnancy, lactation, sexual orientation, gender identity, or gender expression), national origin, age, citizenship, marital status, mental or physical disability, genetic information (including family medical history), discharge status from the military, protected veteran status (which includes disabled veterans, recently separated veterans, active duty wartime or campaign badge veterans, and Armed Forces service medal veterans), or any other basis protected by law. GitLab will not tolerate discrimination or harassment based on any of these characteristics. See also GitLab’s EEO Policy and EEO is the Law. If you have a disability or special need that requires accommodation, please let us know during the recruiting process.

Vacancy posted 10 days ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer, Infrastructure Platforms — AMER (Intermediate to Senior Staff) in Remote vacancy
  • $127k - $249k

    We are looking for an experienced Senior or Staff Engineer for our SRE, InfraSec team, to guide the security of our cloud-based infrastructure. As a Staff SRE, you will be very hands-on...  ...implement controls that reinforce the platform’s security posture.This is an SRE team... 
    Senior
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    Austin, TX
    1 day ago
  •  ...The Team Platform Engineering is the department within SRE that is responsible...  ...for a range of critical infrastructure and operational functions that...  ...and maintaining the reliable and globally connected multi...  ...We are seeking a talented Site Reliability Engineer (SRE)... 
    Senior
    Full time
    Work at office
    Remote work
    Worldwide

    Mongodb

    Austin, TX
    14 days ago
  • $117k - $209.33k

     ...to help make a better world? As a Senior Site Reliability Engineer at Autodesk, you can help us...  ...engineering, security, compliance, platform, and infrastructure teams to ensure services are reliable...  ...: Idaho, USA - Remote; AMER - United States - Texas - PlanoType... 
    Senior
    Full time
    For contractors
    Remote work

    Autodesk

    Plano, TX
    2 days ago
  • $55 - $60 per hour

     ...advanced), Python (intermediate), AWS GovCloud (...  ...: This senior-level role focuses...  ...maintaining secure cloud infrastructure within AWS...  ...strong cloud engineering judgment and the...  ...infrastructure, DevOps, platform engineering, or...  ...of reliability, security, cost... 
    Senior
    Hourly pay
    Contract work
    Remote work

    Akraya

    United States
    3 days ago
  • Senior SRE - Platform - Kubernetes Engineer
    Senior
    Full time

    SFE

    Remote
    2 days ago
  •  .... Overview: The Site Reliability Engineer (SRE) is responsible for...  ...scalability of Precisely's infrastructure platforms across CEDAR (CCX) — an...  ...to meet them. As a Senior SRE, this role is...  ...workloads in AWS at an intermediate level: EC2, ECS, S3, VPC... 
    Senior
    Work experience placement
    Remote work

    Precisely US Jobs

    United States
    2 days ago
  • $169k - $215k

     ...live-streaming platform used by millions...  ..., scalability, reliability, data, machine...  ...seeking a remote Site Reliability Engineer who will elevate our infrastructure resilience and...  ...levels, from intermediate engineers with...  ...experience through senior and staff engineers who... 
    Senior
    Temporary work
    Work at office
    Remote work

    Multi Media, LLC

    United States
    4 days ago
  • $112.3k - $160.6k

     ...Western National is seeking a Site Reliability Engineer III to join our team!...  ...cloud-based application platforms. This individual monitors...  ...application deployment, infrastructure stability, and reliability...  ...teams. Demonstrated intermediate knowledge and application... 
    Senior
    Full time
    Work at office
    Local area
    Remote work
    Work visa
    Flexible hours

    Western National Insurance Group / Umialik Insurance Company

    Minneapolis, MN
    23 hours ago
  •  ...Role: Senior/Lead SRE - Platform - Platform Services Engineer Location: Remote Term: Full Time Job Description...  ...in SRE, platform engineering, infrastructure, distributed systems, or cloud...  ...-cloud systems with strong reliability, security, and observability... 
    Senior
    Full time
    Remote work

    SFE

    Remote
    2 days ago
  • $170k - $265k

     ...performance, cost-efficient infrastructure for AI-native...  ...metal up through the platform services teams actually...  ...The Role This is a senior SRE role for someone who sets the reliability bar and then pulls the...  ...the automation other engineers build on, the services... 
    Senior
    Full time
    Flexible hours

    Nscale

    Remote
    3 days ago
  • $129.1k - $191.03k

     ...semiconductor solutions are the essential building blocks of the data infrastructure that connects our world. Across enterprise, cloud and AI, and...  ...and technically strong ASIC CAD Flow & Infrastructure Engineer to drive the development, deployment, and support of next-... 
    Senior
    Permanent employment
    Full time
    Internship
    Work from home

    Marvell

    Austin, TX
    3 days ago
  • $190.9k - $334.1k

     ...DescriptionIt all started when engineer Fred Luddy wrote code that...  .... Our ServiceNow AI platform brings together any AI, any...  ...atmosphere.About the roleAs a Sr Staff Cloud Architect, you will directly...  ...that shape ServiceNow's infrastructure across private, sovereign, and... 
    Senior
    Work at office
    Immediate start
    Remote work
    Flexible hours

    ServiceNow

    Kirkland, WA
    4 days ago
  •  ...Bridgeway is seeking a Senior Platform Engineer to architect, design, develop...  ...ensure Bridgeway's cloud infrastructure platform meets the needs...  ...in DevOps, DevSecOps, Site Reliability Engineering (SRE), and modern...  ...workflows ~ Intermediate proficiency with Terraform... 
    Senior
    Remote work

    Bridgeway Benefit Technologies

    Tampa, FL
    21 days ago
  •  ...industry, is looking for a Senior Engineer to join our Platform team with a singular...  ...organization faster, more reliable, and more joyful. In...  ...Kafka, Redis, Redshift Infrastructure & Orchestration: Kubernetes...  ...careers of associate and intermediate engineers. Prior experience... 
    Senior
    Work at office
    Local area
    Remote work
    Work from home
    Flexible hours
    Shift work

    Worth AI

    New York, NY
    2 days ago
  • $160k - $220k

     ...-world experiences. This is an opportunity to work on a platform that blends marketplace dynamics, rich user interaction,...  ...sophisticated integrations—at meaningful scale. As a Senior or Staff Full Stack Engineer, you will play a central role in shaping the core user-... 
    Senior
    Full time
    Work at office
    Remote work
    Flexible hours

    Calliere

    New York, NY
    1 day ago
  • $163k

     ...realize their greatest potential. Title and Summary Senior Site Reliability Engineer Overview-The ProCOM team is looking for a Site...  ...operation and refinement. • Analyze ITSM activities of the platform and provide feedback loop to development teams on... 
    Senior
    Full time
    Part time
    Immediate start
    Worldwide
    Flexible hours

    Mastercard

    O Fallon, MO
    11 hours ago
  • $96k - $163k

     ...potential. Title and Summary Senior Site Reliability Engineer Who is Mastercard? At...  ...responsible for ensuring that our platform is stable and healthy. We break down...  ...reliability. • Cloud Computing and Infrastructure - Ability to design, deploy, and manage... 
    Senior
    Full time
    Part time
    Worldwide
    Flexible hours

    Mastercard

    O Fallon, MO
    11 hours ago
  • $96k - $163k

     ...realize their greatest potential. Title and Summary Senior Site Reliability Engineer Overview The BizOps team is looking for a Senior...  ...capabilities by supporting applications, services, and platforms. This position is to support the Data Delivery and... 
    Senior
    Full time
    Part time
    Worldwide
    Flexible hours
    Shift work

    Mastercard

    O Fallon, MO
    11 hours ago
  • $165.5k - $289.6k

    It all started when engineer Fred Luddy wrote code...  ...reinvention. Our ServiceNow AI platform brings together any...  ...a highly experienced Senior Staff Cloud FinOps Analyst...  ...Platform, private infrastructure, and emerging AI...  ...you are leaving this site and going to a third-... 
    Senior
    Work at office
    Immediate start
    Remote work
    Flexible hours

    Servicenow

    Santa Clara, CA
    3 days ago
  •  ...Senior Site Reliability Engineer Company: CyberArk Work Type: Remote Employment: Full Time Location: US Seniority: Mid Level Technologies: AWS, Kubernetes, Terraform, CloudFormation, Ansible, CloudWatch, Grafana, Datadog, OpenSearch, PagerDuty Requirements: Senior SRE... 
    Senior
    Full time
    Remote work

    CyberArk

    United States
    3 days ago
  •  ...Tech SolutionsJob Title: Senior Site Reliability EngineerLocation: Plano ,...  ...skilled Senior Site Reliability Engineer (SRE) to join our dynamic...  ...teams to enhance our infrastructure and deployment processes.ResponsibilitiesMaintain...  ...in AWS or GCP cloud platforms.Hands on experience with... 
    Senior
    Remote work

    SRI Tech

    Plano, TX
    1 day ago
  • Senior Platform Engineer (Hybrid)Position OverviewWe are seeking a Senior Platform Engineer to own and evolve our core infrastructure and internal developer platform in a hybrid work model. You will design, build, and operate reliable, scalable, and secure platform services... 
    Senior
    Work at office
    Remote work

    CyberCoders

    New York, NY
    1 day ago
  •  ...Senior Site Reliability Engineer Company: Sphera Work Type: Remote Employment: Full Time Location: US Seniority: Senior Level Technologies: Terraform, ARM templates, Kubernetes, Azure, SonarCloud, CheckPoint, Hadoop, Kafka, Presto, NewRelic, CI/CD, Linux, Windows, Redis... 
    Senior
    Full time
    Remote work

    Sphera

    United States
    3 days ago
  • $184k - $287.5k

     ...consolidating a dual-scheduler estate onto a single LSF platform, and we are looking for an engineer who knows LSF at the level of its internals — not just...  ...Hands-on MultiCluster experience in a production, multi-site environment.Strong Linux systems fundamentals, system... 
    Senior
    Full time
    Remote work

    Nvidia

    Austin, TX
    4 days ago
  •  ...a leader in AI cloud infrastructure serving tens of thousands...  ...is currently Tuesday.Engineering at Lambda is...  ...tenant cloud networking platform and SDN infrastructureOperate...  ...to improve service reliability and deployment...  ...years of experience in Site Reliability Engineering... 
    Senior
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    3 days ago
  • $210k - $230k

    GovCIO is currently hiring for a Senior Site Reliability Engineer (SRE) to design, implement, and maintain...  ...available, scalable, and resilient infrastructure systems. The ideal candidate will...  ...• Build self-service tools and platforms to enable development teamsReliability... 
    Senior
    Currently hiring
    Remote work

    Govcio

    Arlington, VA
    3 days ago
  • $174k - $252k

     ...developing software platforms and frameworks, capacity...  ...changes that improve reliability and velocity.Practice...  ...in Computer Science, Engineering, a related field, or...  ...or Engineering.Site Reliability Engineering...  ...built by the Technical Infrastructure team to keep it running... 
    Senior

    Google

    Sunnyvale, TX
    1 day ago
  •  ...Cloud, is a leader in AI cloud infrastructure serving tens of thousands...  ...day is currently Tuesday.Engineering at Lambda is responsible for...  ...automate the validation of platform quality.Design, build, and...  ...services, workloads, and platform reliability.You6+ years of experience... 
    Senior
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    12 hours ago
  • $104.9k - $174.7k

     ...the Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and...  ...be directly involved in designing infrastructure, writing Terraform, improving observability...  ...operating monitoring and uptime platforms such as Grafana, Pingdom, and... 
    Senior
    Full time
    Work at office
    Local area
    Remote work
    Work from home

    LexisNexis Risk Solutions Group

    Alpharetta, GA
    12 hours ago
  • $149.4k - $202k

     ...Noctua Technology is seeking a Senior Software Engineer specializing in Site Reliability Engineering to join their team. This role focuses on the reliability...  ...of cloud-native applications, emphasizing Infrastructure as Code and automation. The ideal candidate will... 
    Senior
    Remote work

    Noctua Technology

    Virginia, MN
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer, Infrastructure Platforms — AMER (Intermediate to Senior Staff). Be the first to apply!