Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer

$150k

VantageScore

Job Description

Job Description

About The Role  

We are seeking an experienced Site Reliability Engineer (SRE) with a strong focus on DevSecOps to join our growing engineering team. In this role, you will oversee and maintain the reliability, security posture, and operational hygiene of our cloud infrastructure, APIs, and software supply chain. You will drive patch management programs, harden our Cloud infrastructure, and maintain our code repositories to ensure all systems remain compliant, secure, and scalable. 

This role is ideal for an engineer who thrives at the intersection of operations and security, is passionate about automation, and takes pride in keeping complex environments clean, auditable, and resilient. 

Key Responsibilities  

  • Own and execute end-to-end patch management across AWS compute resources (EC2, ECS, Lambda runtimes, EKS nodes), third-party dependencies, and OS-level packages. 
  • Monitor, triage, and remediate vulnerabilities identified by security scanning tools (e.g., AWS Inspector, Dependabot, Security Hub, or equivalent), prioritizing by CVSS severity and business impact. 
  • Maintain and enforce branch protection rules, secret scanning policies, and dependency update workflows across all code repositories. 
  • Design and implement automated pipelines for continuous compliance checking, security testing (SAST/DAST/SCA), and infrastructure drift detection. 
  • Collaborate with IT & Info-Sec SMEs on AWS IAM roles and policies, VPC configurations, Security Groups, CloudTrail, Config, and GuardDuty to ensure least-privilege access and auditability. 
  • Collaborate with development teams to embed security controls into CI/CD pipelines (GitHub Actions, CodePipeline, or equivalent) without impeding developer velocity. 
  • Support the reliability and availability of production APIs — including uptime monitoring, incident response, runbook creation, and post-incident reviews. 
  • Partner with Legal and Data Governance SMEs on API access procedures and monitoring. 
  • Define and track SLOs/SLAs for internal and external APIs; implement alerting and dashboards using observability tooling (e.g., CloudWatch, Datadog, Grafana). 
  • Lead periodic infrastructure and dependency audits; produce clear reports on patch compliance status and open risk items for engineering and security leadership. 
  • Maintain thorough documentation of patching schedules, runbooks, access policies, and environment configurations. 
  • Participate in on-call rotation and contribute to a culture of continuous improvement. 

Required Qualifications  

  • Bachelor's Degree in Computer Science, Information Systems, or a related field (or equivalent practical experience). 
  • 5+ years of professional experience in a Site Reliability Engineering, Software Engineering, DevOps, or DevSecOps role. 
  • Demonstrated expertise managing AWS environments — including EC2, Lambda, ECS/EKS, S3, RDS, IAM, VPC, CloudTrail, Config, and GuardDuty. 
  • Experience with various cloud environments: AWS, Azure, GPC 
  • Strong experience with GitHub administration: branch protection, Actions workflows, secret scanning, Dependabot, and code owners. 
  • Hands-on experience with patch management and vulnerability remediation at scale, including OS-level patching (Amazon Linux, Ubuntu) and dependency lifecycle management. 
  • Proficiency with infrastructure-as-code tools (Terraform, CloudFormation, or AWS CDK). 
  • Experience integrating security tooling (SAST, DAST, SCA, container scanning) into CI/CD pipelines. 
  • Solid understanding of API reliability patterns: health checks, rate limiting, circuit breakers, and observability. 
  • Familiarity with compliance frameworks relevant to cloud environments (SOC 2, CIS Benchmarks, NIST CSF). 
  • Strong scripting skills in Python, Bash, or similar for automation and tooling. 
  • Excellent communication skills and ability to translate technical risk for non-technical stakeholders. 
  • Build observation (logging, metrics, alerting) systems to make sure system works well, and develop response plans. 

Preferred Qualifications  

  • AWS certifications (e.g., AWS Certified Security – Specialty, AWS Certified DevOps Engineer – Professional). 
  • Experience with container security and Kubernetes (EKS) hardening. 
  • Familiarity with CSPM tools (e.g., Wiz, Prisma Cloud, AWS Security Hub) for continuous cloud posture management. 
  • Experience managing API gateways (AWS API Gateway, Kong, or similar) including security policy enforcement. 
  • Exposure to secrets management solutions (AWS Secrets Manager, HashiCorp Vault). 
  • Knowledge of SBOM (Software Bill of Materials) generation and management. 
  • Experience with incident response playbooks and tabletop exercises. 
  • Familiarity with Agile/Scrum methodologies and cross-functional engineering teams. 

Compensation

The anticipated base salary range for this position is $150,000 annually, plus eligibility for a 15% annual performance bonus. Actual compensation will be determined based on several factors, including skills, experience, education, certifications, and geographic location.

In addition to base salary and bonus eligibility, we offer a competitive benefits package, including medical, dental, vision, 401(k), paid time off, and other employee benefits.

Vacancy posted 17 days ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in San Francisco, CA vacancy
  • $180.1k - $278.7k

     ...Staff Infrastructure Reliability EngineerThe Staff Infrastructure Reliability Engineer is responsible for the technical leadership of Redfin's production database and storage systems. They will work with the database team manager and other database and storage engineers... 
    Suggested
    Minimum wage
    Immediate start

    Rocket Companies

    San Francisco, CA
    3 days ago
  •  ...and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the Enterprise Technology, Infrastructure Platforms team, you will solve complex and broad business... 
    Suggested

    J.P. Morgan

    San Francisco, CA
    1 day ago
  •  ...Lambda Inc. in San Francisco is seeking a Storage Engineer to own the reliability, performance, and capacity of our production storage fleet across multiple data centers, using a software-defined data plane. You will build monitoring, dashboards, and alerting for storage... 
    Suggested

    Lambda

    San Francisco, CA
    3 days ago
  • $113.4k - $162k

     ...break down barriers to communication and free the flow of conversation for people everywhere.TextNow is looking for motivated Site Reliability Engineer to own infrastructure, monitoring, logging, ci/cd, reliability and everything in between!This role is aboutimpactatscale.... 
    Suggested

    TextNow

    San Francisco, CA
    23 hours ago
  •  ...Apple Service Engineering (ASE) seeks a senior SRE software engineer to own the architectural direction of Kubernetes internals powering...  ...You will define controllers and namespace management, raise reliability, and contribute to upstream Kubernetes. The role includes... 
    Suggested

    Socket

    San Francisco, CA
    4 days ago
  • Job TitleAt U.S. Bank, we're on a journey to do our best. Helping the customers and businesses we serve to make better and smarter financial decisions and enabling the communities we support to grow and succeed. We believe it takes all of us to bring our shared ambition...

    Phenom People

    San Francisco, CA
    3 days ago
  • $200.7k - $250.9k

     ...washed away in a flood in 1942, the Royal Engineers rebuilt it. Then it washed away again in...  ...opportunities for improvements in reliability/observability/performance/preparedness and...  ...candidate for the role: Has past Site Reliability Engineering or DevOps experience... 

    Embedded Shishya

    San Francisco, CA
    3 days ago
  • $200k - $260k

     ...the future of professional services is being written today — and we're just getting started.Role OverviewAs a Software Engineer on the Site Reliability team at Harvey, you will ensure the reliability, scalability, and performance of our legal AI platform. You'll join a... 
    Relocation package

    Harvey

    San Francisco, CA
    3 days ago
  • $120k - $168.49k

     ...Site Reliability Engineer, Cloud Infrastructure About Quizlet At Quizlet, our mission is to help every learner achieve their outcomes in the most effective and delightful way. Our $1B+ learning platform serves tens of millions of students every month, including two-thirds... 
    Internship
    Work at office
    3 days per week

    Quizlet

    San Francisco, CA
    2 days ago
  • $194k - $267k

     ...all in on this mission. If you are too, let's talk.Position Overview:We are seeking a highly technical StaffObservabilitySite Reliability Engineer with a specialty in Splunk to own and evolve our Splunk ecosystem. In this role, you will move beyond simple monitoring to... 
    Permanent employment
    Work at office
    Local area
    Worldwide
    Flexible hours

    Okta, Inc.

    San Francisco, CA
    23 hours ago
  • $127k - $249k

     ...The TeamPlatform Engineering sits within SRE and builds the core infrastructure powering MongoDB...  ...plays a pivotal role in engineering the reliable, globally connected, multi-cloud network...  ...are seeking a talented Senior Site Reliability Engineer (SRE) with a strong... 
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    San Francisco, CA
    23 hours ago
  •  ...A tech startup in San Francisco is looking for Site Reliability Engineers to enhance system reliability and performance. Ideal candidates have over 5 years of relevant experience and strong expertise in cloud infrastructure, including AWS and Kubernetes. The role involves... 

    Breakout Tools

    San Francisco, CA
    2 days ago
  • $200k - $240k

     ...systems across all product teams. You will collaborate closely with engineering leadership, product managers, and cross-functional teams to...  ...and Helm ~ Understand the importance of performant and reliable systems ~ Education - Ideally looking for a B.A. / B.S. degree... 
    Work at office
    Immediate start
    3 days per week

    Altruist

    San Francisco, CA
    2 days ago
  • $350k

     ...with leading AI companies and infrastructure providers to build reliable, high-performance platforms supporting next-generation AI workloads. This opportunity is for a Staff Site Reliability Engineer to lead the reliability of large-scale GPU infrastructure, covering... 

    Hamilton Barnes Associates Limited

    San Francisco, CA
    3 days ago
  •  ...troubleshooting, providing rubric-based written feedback. This role requires hands-on Kubernetes expertise in EKS/GKE/AKS or self-managed clusters, with strong scripting in Go, Python, or TypeScript, and ability to document findings clearly for engineering #J-18808-Ljbffr

    Obsidian

    San Francisco, CA
    16 hours ago
  •  ...guarantees and certifications. We're hiring staff-level SREs to help run and evolve that infrastructure, working alongside the senior engineers already on the team. You'll contribute to architecture decisions for how we deploy, observe, and secure the platform, and help... 
    Remote work
    Flexible hours

    Akka

    San Francisco, CA
    19 days ago
  • $114.3k - $235.32k

     ...verification who have now purpose-built a CTV performance platform advertisers can trust to grow their business.We are seeking a Site Reliability Engineer to help operate, scale, and continuously improve a cloud-native platform built on AWS, Kubernetes/EKS, and ArgoCD-driven... 
    Work at office
    Local area
    Relocation
    Relocation package

    Pinterest

    San Francisco, CA
    a month ago
  • $250k

     ...Europe, while now significantly expanding its footprint in the United States. The company is looking for a Senior / Staff Site Reliability Engineer to support and scale large-scale HPC and cloud environments powering GPU-intensive workloads. The role involves working... 
    Full time
    Remote work
    San Francisco, CA
    more than 2 months ago
  • $117k - $209.33k

     ...Job Requisition ID #26WD99273Position OverviewWant to help make a better world? As a Senior Site Reliability Engineer at Autodesk, you can help us build and operate reliable, secure, and scalable cloud services for Autodesk GovCloud products.As part of a new SRE team supporting... 
    Full time
    For contractors

    Autodesk

    San Francisco, CA
    a month ago
  •  ...design of information and operational support systems.  Required Skills/Qualifications: BS/MS degree in Computer Science, Engineering, or a related subject. Equivalent experience accepted.   Proven working experience in installing, configuring, and troubleshooting... 
    Full time
    Work experience placement
    Remote work
    Flexible hours
    San Francisco, CA
    more than 2 months ago
  •  ...Job Description Job Description Site Reliability Engineer II Bay Area, offices in San Jose · Hybrid · 24/7 FedRAMP Operations · Rotational Shift · Initial Contract till March 27. KEY REQUIREMENT This role requires US citizenship and residence on US soil.... 
    Hourly pay
    Contract work
    For contractors
    Shift work
    Night shift
    Weekend work

    C-Serv

    San Francisco, CA
    17 days ago
  • $152.5k - $205k

     ...flexible work environment where new ideas are encouraged and everyone is a stakeholder.What you’ll be responsible forThe Site Reliability Engineer builds and maintains shared platform capabilities, common libraries, and infrastructure that help Circle teams ship secure... 
    Flexible hours

    Circle

    San Francisco, CA
    22 days ago
  • $165k - $227k

     ...opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Engineering Opportunity We are looking for an experienced Senior Site Reliability Engineer to join Okta's Emerging Products Group (EPG). Our mission is to build highly... 
    Local area
    Worldwide
    Flexible hours

    Okta

    San Francisco, CA
    12 days ago
  • $190.8k - $267.1k

     ...while helping Reddit grow its business. The reliability of our Ads systems directly impacts...  ...Reliability team partners closely with Ads Engineering to improve reliability, scalability,...  ...advertiser trust. We’re looking for a Senior Site Reliability Engineer to build, operate,... 
    For contractors
    Work experience placement

    Reddit

    San Francisco, CA
    6 days ago
  • $204k - $306k

     ...all in on this mission. If you are too, let's talk.Manager, Site Reliability EngineeringSan Francisco, CaliforniaSecure Every Identity, from...  ...week in our San Francisco Office.The IDaaS Site Reliability Engineering GroupOkta authenticates, authorizes and provisions millions... 
    Permanent employment
    Work at office
    Local area
    Worldwide
    Flexible hours
    2 days per week

    Okta, Inc.

    San Francisco, CA
    23 hours ago
  •  ...SRE to define the future of our cloud platform and champion engineering excellence across Ironclad. In this role, you will pair deep...  ...Provide technical leadership and strategic direction for the Site Reliability Engineering team and our broader Cloud Platform Define... 
    Full time
    Contract work
    Work at office

    Ironclad Inc

    San Francisco, CA
    3 days ago
  • $175k - $250k

     ...0/yr Job Title: Senior Cloud Infrastructure Engineer Location: San Francisco, CA. Remote unavailable. Modality: On-Site only. Must live within commuting distance of...  ...while ensuring scalability, performance, and reliability across environments. What You’ll Do Design,... 
    Full time
    Remote work
    Relocation
    Relocation package

    The Recruiting Guy

    San Francisco, CA
    3 days ago
  • $210k - $240k

    Join to apply for the Senior Site Reliability Engineer role at Alembic Technologies This range is provided by Alembic Technologies. Your actual pay will be based on your skills and experience — talk with your recruiter to learn more. Base pay range $210,000.00/yr - $2... 
    Full time

    Alembic Technologies

    San Francisco, CA
    3 days ago
  •  ...Infrastructure team builds the platforms and tooling that help engineering teams develop, deploy, and operate production systems safely...  ...safe shipping the default for every product team.As a Staff Site Reliability Engineer on Release Engineering, you'll define and scale... 
    Permanent employment
    Work experience placement
    Work at office
    Local area

    Plaid Financial

    San Francisco, CA
    a month ago
  • $195k - $257.5k

     ...flexible work environment where new ideas are encouraged and everyone is a stakeholder.What you’ll be responsible for:As a Staff Site Reliability Engineer on Circle’s Platform team, you’ll design, build, and operate the infrastructure that powers our blockchain platform at... 
    Flexible hours

    Circle

    San Francisco, CA
    29 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!