Site Reliability Engineer
$140k - $205kCooley
Senior Technology Site Reliability Engineer
Cooley is seeking a Senior Site Reliability Engineer to join the Infrastructure & Development Operationsteam.
Position summary: The Senior Technology Site Reliability Engineer("SRE") is responsible for ensuring the reliability, scalability, and performance of the firm's critical infrastructure and applications. The SREblends software engineering with systems engineering to build and maintain automated, resilient, and observable systems that support high availability and operational excellence. In addition to being technically advanced, the SRE will have a high degree of emotional intelligence and the ability to work as a team towards complex and layered objectives. Specific duties and responsibilities include, but are not limited to, the following:
Position responsibilities:
- Monitor and maintain production systems to ensure high availability and performance
- Implement and manage service-level indicators (SLIs), objectives (SLO's), agreements (SLA's), and error budgets
- Participate in on-call rotations and incident response, including root cause analysis and postmortems
- Develop and maintain infrastructure as code (IaC) using Terraform
- Automate deployment, scaling, and recovery processes to reduce manual intervention
- Partner with DevOps to build and maintain CI/CD pipelines to support safe and efficient software delivery
- Implement observability solutions using metrics, logs, traces, and alerting systems (Prometheus, Grafana, DataDog, etc.)
- Proactively identify and resolve system bottlenecks and reliability risks
- Work closely with Infrastructure, DevOps, Development, and security teams to embed reliability into the development lifecycle
- Contribute to a culture of blameless post-mortems and continuous improvement
- Document operational procedures and share knowledge across teams
- All other duties as assigned or required
Skills and experience :
Required:
- After orientation at Cooley LLP, exhibit proficiency in the Microsoft Office suite, iManage and other firm applications
- Ability to work extended and/or weekend hours, as required
- Ability to travel, as required
- 6+ years direct applicable experience (e.g. site reliability engineering or related field)
- Proficiency in Terraform and programming languages such as Python, Go, or Java
- Deep expertise in cloud platforms, particularly AWS, and container orchestration
- Strong background in distributed systems, performance tuning, and automation
- Hands-on experience with configuration management tools such as Puppet, Chef, or Salt
Preferred:
- Bachelor's Degree in Computer Science, Information Technology, Engineering, or associated discipline
- Experience working with advanced ETL data workflows including technologies such as AWS EMR, Azure Synapse, Azure Data Factory, or Apache Hive/Spark/Airflow
- Experience with IaC deployment of AKS/EKS/GKE architecture
- Experience with enterprise Data Lake environments using technologies such as DataBricks or Snowflake
Competencies :
- Expert analytical/quantitative, problem-solving, and deductive reasoning skills, experience performing advanced troubleshooting and root cause analysis of complex technical issues
- Excellent organizational, planning, and time management skills and ability to work independently and in a team environment to manage competing priorities and meet deadlines
- Advanced verbal and written communication skills with the ability to present findings, conclusions, alternatives, and information clearly and concisely
- Experience working with all levels of business professionals, management, stakeholders, and vendors with the ability to build effective relationships through trust and diplomacy
Cooley offers a competitive compensation and excellent benefits package and is committed to fair and equitable employment practices.
EOE.
The expected annual pay range for this position with a full-time schedule is $140,000 - $205,000. Please note that final offer amount will be dependent on geographic location, applicable experience and skillset of the candidate.
We offer a full range of elective benefits including medical, health savings account (with applicable medical plan), dental, vision, health and/or dependent care flexible spending accounts, pre-tax commuter benefits, life insurance, AD&D, long-term care coverage, backup care for children and/or adults and other parental support benefits. In addition to elective benefit options, benefited employees receive firm-paid life insurance, AD&D, LTD, short term medical benefits as well as 21 days of Paid Time Off ("PTO") and 10 paid holidays each year. We provide generous parental leave and fertility benefits. New employees will attend a detailed benefit orientation to learn more about our many benefits and resources.
- ...our Series B and have grown 800% over the last 12 months. Engineering at Ivo Engineers at Ivo are inventors. Ivo was first-to-... ...still expect us to hit our SLAs. What? We’re looking for a Site level Reliability Engineer as part of Infrastructure team to: Own uptime,...SuggestedFull timeContract workWork at officeRemote workVisa sponsorshipRelocation packageFlexible hours
- ...and actionable to everyone, everywhere. That everyone now includes AI agents. The Role: You'll be the infrastructure and reliability engineer on the Data Replication team - a full-stack product team running over 3 million sync jobs a week powering thousands of data...SuggestedFull timeWork at officeLocal areaFlexible hours
$204k - $281k
...This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. MANAGER, SITE RELIABILITY ENGINEERING San Francisco, California Secure Every Identity, from AI to Human Identity is the key to unlocking the potential...SuggestedPermanent employmentFull timeWork at officeLocal areaWorldwideFlexible hours2 days per week$210k - $240k
...Join to apply for the Senior Site Reliability Engineer role at Alembic Technologies This range is provided by Alembic Technologies. Your actual pay will be based on your skills and experience — talk with your recruiter to learn more. Base pay range $210,000.00/yr - $24...SuggestedFull time$175k - $250k
...00/yr Job Title: Senior Cloud Infrastructure Engineer Location: San Francisco, CA. Remote unavailable. Modality: On-Site only. Must live within commuting distance of... ...while ensuring scalability, performance, and reliability across environments. What You’ll Do Design, build...SuggestedFull timeRemote workRelocationRelocation package- ...company valued at $10 billion. We work in‑person five days a week in our new SanFrancisco headquarters. About the Role As a Site Reliability Engineer (SRE) at Mercor, you’ll own production reliability across our most critical systems, partnering directly with...
$155k - $222.6k
...global cloud platform. As a team of six engineers distributed across the US, Canada, and the... ...with a strong focus on automation, reliability, and operational excellence. We are one... ...Qualifications ~2+ years of experience in Site Reliability Engineering, DevOps,...Permanent employmentFull timeTemporary workLocal areaWorldwideFlexible hours$120k - $168.49k
...Site Reliability Engineer, Cloud Infrastructure About Quizlet At Quizlet, our mission is to help every learner achieve their outcomes in the most effective and delightful way. Our $1B+ learning platform serves tens of millions of students every month, including two-thirds...InternshipWork at office3 days per week- ...The role We're looking for a world-class Site Reliability Engineer to ensure the reliability, performance, and scalability of our AI infrastructure platform. You'll be building and operating the core systems that power agentic AI at scale. Your mission: keep...
$260k - $300k
...Software Agents We are the makers of Devin, the first AI software engineer. Our team is extremely talent-dense. Among our founding... ...than anyone expects. You will own both the production reliability of our user-facing products and the platform engineering that...- ...JOB DESCRIPTION Project Outline: We are looking for a Site Reliability Engineer with experience in incident response. In this role, you will help Shipt understand where we can improve stability and reliability. There will be a focus on the intersection of systems...
$148.5k - $223.9k
...you are not duplicating efforts. Job Category Software Engineering Job Details About Salesforce Salesforce is the #1... ...is seeking a senior engineering candidate to join the Site Reliability organization in San Francisco. Working closely with counterparts...WorldwideWeekend work- ...Site Reliability Engineer Job Location: San Francisco, CA or Charlotte, NC. Job Type: Contract Work with local API development squads, platform teams, product owners, scrum masters, and architects. The SRE ensures that both our internally critical and our externally...Contract workLocal area
$230k - $310k
...millions of daily users while enabling our engineering teams to ship fast. You'll own the... ...building automation and tooling that improves reliability and partnering with engineering to... ...What You'll Bring ~5+ years in site reliability engineering, DevOps, or systems...Full timeWork at officeWork from home- ...Open Source LLM Gateway Engineer LiteLLM is an open-source LLM Gateway with 34K+ stars on GitHub and trusted by companies like NASA... ...expanding and seeking our 6th Engineer focused on owning reliability, performance, and infrastructure stability for the LiteLLM proxy...
- ...enterprise that runs the real economy. Learn more about our vision in our manifesto. About the Role We're looking for a Site Reliability Engineer to take the lead on scaling our operational resilience as we grow. You'll own the stability, observability, and debugging...WorldwideShift work
- ...globe. Join us on this journey to redefine resource management-and change lives along the way. The Role As a Site Reliability Engineer (SRE) at Air Apps, you will be responsible for ensuring the reliability, availability, and scalability of our systems. You...Temporary workWorldwide
- ...A tech startup in San Francisco is looking for Site Reliability Engineers to enhance system reliability and performance. Ideal candidates have over 5 years of relevant experience and strong expertise in cloud infrastructure, including AWS and Kubernetes. The role involves...
- A leading technology firm is looking for a Manager to expand their Cloud Site Reliability team. The ideal candidate will have extensive Linux administration experience, a passion for automation, and be comfortable in a remote, diverse workplace. This position emphasizes...Remote work
- ...Arena Intelligence Engineer Arena Intelligence is looking for an engineer to build the core infrastructure that sits beneath our online... ...foundational infrastructure for our users that scales, is reliable, and makes the complexities of operating this infrastructure at...Permanent employmentShift work
- ...an SRE to join our infrastructure team. This role will be responsible for building software to ensure the reliability of our back-end systems, working with engineers who develop them, and planning for our future growth. You will work with our existing production...WorldwideHome officeFlexible hours
- ...Senior Site Reliability Engineer Location: Global Remote / San Francisco • Full-Time About Andromeda Andromeda Cluster was founded by Nat Friedman and Daniel Gross to give early-stage startups access to the kind of scaled AI infrastructure once reserved only...Full timeRemote work
$81.1k - $187k
...Site Reliability Engineer 3 We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The role focuses on improving service reliability, reducing operational risk, automating repetitive tasks, and driving...Temporary workImmediate startFlexible hoursShift work- ...Site Reliability Engineer Specter's mission is to help automate the physical world. Today, we build video sensors with state-of-the-art AI agents that answer any question, anywhere in their environments. Our systems can automatically detect and reason about any physical...Remote work
$166.9k - $225.9k
...Summary: Drata's SRE team operates as both a central engineering function and an embedded reliability practice. You'll be part of a close-knit SRE team... ...What you'll bring: ~6+ years of experience in Site Reliability Engineering, Cloud Engineering, or building...Work at officeImmediate startWorldwideMonday to FridayFlexible hours$220k - $235k
...Staff/Senior Staff Site Reliability Engineer Ironclad is the leading AI contracting platform that transforms agreements into assets. Contracts move faster, insights surface instantly, and agents push work forward, all with you in control. Whether you're buying or selling...Full timeContract workWork at office$200k - $260k
...Join the Cloud Infrastructure Team as a technical leader driving reliability, automation, and scalability across the systems running Sight... ...and reliability practices across teams, mentor senior engineers, and be a primary escalation point for the org's hardest systems...Casual workWork at officeRemote workFlexible hours$181k - $263k
...supporting deployments of global products, and providing first line operational support. We are looking for a Senior Staff Site Reliability Engineer who will set the technical direction for reliability engineering across LiveRamp's global infrastructure. This is a senior...Work from homeFlexible hoursNight shift- ...Lead Site Reliability Engineer Stuut is transforming accounts receivable for B2B companies—making collections smarter and faster for companies that have historically relied on manual processes that are labor intensive and costly. Our platform is gaining traction with...Full timeFlexible hours
$210.8k - $272.8k
About Thumbtack Thumbtack helps millions of people confidently care for their homes. About the Site Reliability Engineering Team The Site Reliability Engineering team focuses on creating and maintaining a reliable, secure, and scalable platform vital for a seamless user...Local area
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- construction site safety Daly City, CA
- on-site clinical research associate (traveling/remote) Daly City, CA
- site reliability engineering manager
- junior site reliability engineer
- site reliability engineer
- site reliability engineer remote
- site reliability engineer sre
- lead site reliability engineer
- manufacturing reliability site resource
- construction site safety



