Senior Site Reliability Engineer
Alembic Limited
About the RoleWe’re looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability, observability, and operational excellence at the core. You’ll partner with engineers and data scientists to build, automate, and maintain the infrastructure that powers our core platform—including data pipelines, ML workloads, and real-time analytics systems.This is a hands-on, high-impact role with visibility across the stack and the opportunity to shape the future of our infrastructure and operations.Key ResponsibilitiesDesign, build, and maintain scalable infrastructure to support real-time analytics and machine learning workloadsImprove system reliability and performance through automation, observability, and proactive capacity planningOwn and evolve CI/CD pipelines, deployment automation, rollback mechanisms, and config managementImplement and maintain monitoring, alerting, and incident response processes (SLOs, runbooks, on-call rotations)Collaborate across engineering and data science teams to drive a culture of performance and reliabilityEnsure security, compliance, and operational readiness across our cloud infrastructureDrive post-incident analysis and continuous improvement initiativesWhat Will Help You Succeed 8+ years of experience in SRE, DevOps, or infrastructure engineering roles5+ years of experience with datacenter operations and/or system and network administrationExperience with containerization (Docker), and orchestration (Kubernetes)Strong knowledge of Linux systems, networking, and systems performance tuning, including strong-to-expert facility with SSH, terminal, shell and the Linux command line.Solid understanding of infrastructure-as-code (e.g., Terraform, Ansible)Good programming skills and ability to apply sound coding principles to IaC and scripting code with languages such as Terraform, Ansible, Bash (shell scripting), and/or Python.Experience with monitoring and observability stacks (e.g., Prometheus, Grafana, Datadog, ELK, OpenTelemetry)Proficiency with CI/CD tools and pipelines (e.g., GitHub Actions, ArgoCD, etc.)Ability to debug complex systems and automate solutions in scripting languagesExcellent communication skills and the ability to work cross-functionallyNice-to-HaveExperience with cloud and managed services (e.g. AWS)Experience supporting data-intensive platforms (Spark, Airflow, Kafka, etc.)Familiarity with security practices for cloud-native applications and infrastructureExperience in high-compliance or SOC-2 environmentsWhat You’ll GetOwnership of mission-critical infrastructure in a company solving real-world enterprise problemsA front-row seat to a high-performance engineering cultureThe ability to influence how our platform scales—from deployment to incident managementAn environment that values curiosity, accountability, and impactCompensation Range: $210K - $240KLocationSan Francisco HQAddressSan Francisco, CaliforniaEmployment TypeFull timeLocation TypeOn-siteDepartmentEngineeringTechnical OperationsCompensationOur compensation philosophy ensures new hires earn 50%+ current benchmarks. We'll discuss further with you. Most recent San Francisco benchmark data: $210K – $240KOur CompensationPhilosophy:Market-based: Our formula ensures new hires earn at or above real-time benchmarks. Ownership: Our generous equity program ensures new hires are owners, not just employees.Transparent:We openly discusssalary expectations to avoid surprises later in the process.Data-driven:We use objective data to removebias and ensure consistency in compensation decisions.
- About the RoleWe’re looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability, observability, and operational excellence at the core. You’ll partner with engineers and data scientists to build, automate, and maintain the infrastructure...Senior
$190.8k - $267.1k
...helping Reddit grow its business. The reliability of our Ads systems directly impacts advertiser... ...team partners closely with Ads Engineering to improve reliability, scalability, operational... ...advertiser trust. We’re looking for a Senior Site Reliability Engineer to build, operate,...SeniorFor contractorsWork experience placement$127k - $249k
The TeamPlatform Engineering sits within SRE and builds the core infrastructure powering MongoDB... ...a pivotal role in engineering the reliable, globally connected, multi-cloud... ...Role OverviewWe are seeking a talented Senior Site Reliability Engineer (SRE) with a strong...SeniorLocal areaRemote workWorldwideFlexible hours$152.5k - $205k
...flexible work environment where new ideas are encouraged and everyone is a stakeholder.What you’ll be responsible forThe Site Reliability Engineer builds and maintains shared platform capabilities, common libraries, and infrastructure that help Circle teams ship secure...SeniorFlexible hours$117k - $209.33k
Job Requisition ID #26WD99273Position OverviewWant to help make a better world? As a Senior Site Reliability Engineer at Autodesk, you can help us build and operate reliable, secure, and scalable cloud services for Autodesk GovCloud products.As part of a new SRE team supporting...SeniorFull timeFor contractors$127k - $249k
The TeamPlatform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions... ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper)....SeniorWork at officeLocal areaRemote workWorldwideFlexible hours$165k - $227k
...opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk.The Engineering OpportunityWe are looking for an experienced Senior Site Reliability Engineer to join Okta's Emerging Products Group (EPG). Our mission is to build highly reliable...SeniorLocal areaWorldwideFlexible hours- ...’s build what’s next.About the teamThe Engineering team at Airwallex is a diverse group of... ...ownership, working together to build scalable, reliable, and secure products that empower... ...our Global services.What you’ll doAs a Senior Site Reliability Engineer, you’ll work...SeniorTemporary workLocal areaWorldwide
$148.5k - $223.9k
...right place! Agentforce is the future of AI, and you are the future of Salesforce.Salesforce is seeking a senior engineering candidate to join the Site Reliability organization in San Francisco. Working closely with counterparts in the Infrastructure and R&D organizations...SeniorFull timeWorldwideWeekend work$55k - $151.47k
...ApplicableSpecialismIFS - Internal Firm Services - OtherManagement LevelSenior AssociateJob Description & SummaryThe OpportunityAs a Site Reliability Engineer - Senior Associate, you will play a pivotal role in enhancing the reliability, scalability, and performance of our...SeniorFull timeH1b- ...getting here.)About the RoleWe're building infrastructure that has to perform under real-world scale, reliability, and security demands — and we're looking for an engineer who wants to own the foundation it runs on. This isn't a traditional "keep the lights on" role.You'...Senior
$167.7k - $245.2k
...within Cisco’s Networking, Security, Collaboration, and Observability portfolios.Your ImpactWe are seeking a skilled Senior Site Reliability Engineer (SRE) in Production Engineering with a strong background in SaaS and operations. You will design and manage large-scale...SeniorFull timeTemporary workWork at officeLocal areaFlexible hours1 day per week$232k - $319k
...to help us continue to scale the service with great people and reliable, cost-effective, and efficient infrastructure, processes, and... ...enabled with self-serviceAccelerate the velocity of SRE and product engineering by developing robust platforms, powerful tooling, and...SeniorPermanent employmentLocal areaWorldwideFlexible hours$175k - $250k
...00.00/yr - $250,000.00/yr Job Title: Senior Cloud Infrastructure Engineer Location: San Francisco, CA. Remote unavailable. Modality: On-Site only. Must live within commuting distance... ...scalability, performance, and reliability across environments. What You’ll Do...SeniorFull timeRemote workRelocationRelocation package$250k
...across Europe, while now significantly expanding its footprint in the United States. The company is looking for a Senior / Staff Site Reliability Engineer to support and scale large-scale HPC and cloud environments powering GPU-intensive workloads. The role involves...SeniorFull timeRemote work- ...The Team Platform Engineering is the department within SRE that is responsible for a range... ...role in developing and maintaining the reliable and globally connected multi-cloud network... ...Role Overview We are seeking a talented Site Reliability Engineer (SRE) with a strong...SeniorFull timeWork at officeRemote workWorldwide
$15k
...benefits packages, technology talks by our experts, a beautiful modern office, daily catered lunches, and more.As a Senior Cluster Site Reliability Engineer (SRE), you will help scale our research compute cluster to meet our growing needs, and you will leverage...SeniorWork at officeLocal areaRemote work$139.76k - $287.75k
...to grow their business.We are seeking a Senior Site ReliabilityEngineer to help operate,... ...will be instrumental in advancing the reliability, scalability, automation, observability... ...The ideal candidate is a highly hands-on engineer with strong production experience and a...SeniorWork at officeLocal areaRelocationRelocation package$106k - $130k
..., for any employer, at the date of hire. This position is ineligible for employment Visa sponsorship.Role Summary The Senior Site Reliability Engineer applies software engineering and systems engineering practices to improve the reliability, resilience, scalability, and...SeniorHourly payFull timeImmediate startVisa sponsorshipWork visaFlexible hours$174.92k - $209.91k
...same: to make access to data as simple and reliable as electricity. With Fivetran, customer... ..., canonical and ready to query, with no engineering or maintenance required. We’re proud... ...integrate our teams, systems, and career sites. About the Role Fivetran is building...SeniorFull timeWork at officeRemote work$262k - $364k
Lead a team of software/systems engineers on projects for users and be directly responsible for uptime.Own end-to-end availability... ...Engineering.Experience with machine learning infrastructure.Site Reliability Engineering (SRE) combines software and systems engineering...Senior$300k
...thousands of H100s, H200s, and B200s, ready for experimentation, full-scale model training, or inference. As a Platform Engineer/Senior Site Reliability Engineer, you’ll own the reliability, performance, and automation of this GPU-powered infrastructure, ensuring...SeniorPermanent employment- ...Infrastructure team builds the platforms and tooling that help engineering teams develop, deploy, and operate production systems safely... ...safe shipping the default for every product team.As a Staff Site Reliability Engineer on Release Engineering, you'll define and scale...Permanent employmentWork experience placementWork at officeLocal area
- ...and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Enterprise Technology, Infrastructure Platforms team, you will solve complex and broad business...
$113.4k - $162k
...break down barriers to communication and free the flow of conversation for people everywhere.TextNow is looking for motivated Site Reliability Engineer to own infrastructure, monitoring, logging, ci/cd, reliability and everything in between!This role is about impact at...Temporary work- ...complex, distributed, cloud-native systems. As a Staff Platform Engineer, you will play a critical role in ensuring these systems... ...hands-on engineering and technical leadership role. You will own reliability for major platform domains, design scalable solutions on Kubernetes...Senior
$195k - $257.5k
...is a stakeholder.What you’ll be responsible for:As a Staff Site Reliability Engineer on Circle’s Platform team, you’ll design, build, and operate... ...contributors:Staff Site Reliability Engineer (IV) Senior Site Reliability Engineer (III)What you’ll bring to Circle...Flexible hours$204k - $306k
...all in on this mission. If you are too, let's talk.Manager, Site Reliability EngineeringSan Francisco, CaliforniaSecure Every Identity, from... ...in our San Francisco Office. The IDaaS Site Reliability Engineering GroupOkta authenticates, authorizes and provisions millions...Permanent employmentWork at officeLocal areaWorldwideFlexible hours2 days per week$217k - $303.9k
...information, visit .As Reddit continues to scale globally, reliability and performance are more critical than ever. The Site Experience SRE team sits at the intersection of infrastructure, product engineering, and user experience - ensuring that every interaction across...For contractorsWork experience placement$150k - $220k
...teams, and innovators in this way. The Role: As an engineering organization, we pride ourselves on engineering as a creative... ...can achieve autonomy, mastery, and purpose. The Manager, Site Reliability Engineering will lead Forge’s SRE team responsible for keeping...Local area
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- site reliability engineer sre San Francisco, CA
- site reliability engineer San Francisco, CA
- site reliability engineer remote San Francisco, CA
- senior operations technician San Francisco, CA
- senior operations associate San Francisco, CA
- senior cloud service delivery manager San Francisco, CA
- senior it service manager San Francisco, CA
- senior project engineer San Francisco, CA
- senior chief engineer San Francisco, CA
- sr operations manager San Francisco, CA


