Senior Lead Site Reliability Engineer
$146.7kSocket
What you can expect
As a Senior Lead Site Reliability Engineer, you can anticipate opportunities to work on our hybrid systems across the globe. You will be responsible for installing, configuring, and monitoring new systems within a network of global data centers. Additionally, you will patch and maintain thousands of physical and cloud systems worldwide. To streamline operations, you will develop automation to reduce repetitive tasks and analyze and address performance bottlenecks. Furthermore, you will update and troubleshoot user access permissions, resolve network connectivity issues, and maintain system firewalls.
About the Team
Zoom's SRE team is committed to delivering customer happiness, improving business efficiency, and promoting agility through innovation, data-driven insights, and automation. Our impact is reflected in smooth user experiences, optimized processes, and support for Zoom's expansion in the realm of communication and collaboration.
Responsibilities
Providing technical direction for cross-team initiatives and major incidents. Mentor SRE's and developers; define best practices and design patterns. Partner with Security, Networking, and Platform teams on architecture roadmaps. Influence vendor and hardware strategy for on-prem and cloud workloads. Design self-healing platforms using automation, chaos engineering, and fault-tolerant patterns. Optimize Linux systems at scale: performance tuning, kernel parameters, networking, storage, and security hardening. Define best practices and advocate for them across the company. Excellent communication skills and experience driving cross team projects as a technical lead. Able to participate in on-call shifts and incident management and work after hours/weekends for application releases/deployments.
What we’re looking for
- 10+ years in SRE, production engineering, or large-scale systems administration
- Have experience of Linux system administration (systemd, cgroups, networking, filesystems, performance analysis)
- Demonstrate coding ability with at least one programming language e.g. Python
- Have experience with configuration management (Ansible), IaC (Terraform, Packer), CI/CD pipelines (Jenkins, GitLab), container orchestration (k8s, Docker) and observability platforms.
- Have experience with incident response for mission-critical environments.
- Possess a security -first mindset (TPM, secure boot, identity, secrets management).
- Demonstrate networking expertise: BGP, load balancing, DNS, TLS, traffic engineering.
- Have experience with chaos engineering and resilience testing. Have experience with distributed storage systems such as Ceph
- Occasional weekend work may be required
- Ability to work across the globe or multiple time zones
Salary Range or On Target Earnings
Minimum: $146,700.00
Maximum: $339,300.00
In addition to the base salary and/or OTE listed Zoom has a Total Direct Compensation philosophy that takes into consideration; base salary, bonus and equity value.
Note: Starting pay will be based on a number of factors and commensurate with qualifications & experience.
We also have a location based compensation structure; there may be a different range for candidates in this and other locations
Anticipated Position Close Date
09/11/26
Ways of Working
Our structured hybrid approach is centered around our offices and remote work environments. The work style of each role, Hybrid, Remote, or In-Person is indicated in the job description/posting.
Benefits
As part of our award-winning workplace culture and commitment to delivering happiness, our benefits program offers a variety of perks, benefits, and options to help employees maintain their physical, mental, emotional, and financial health; support work-life balance; and contribute to their community in meaningful ways. Learn for more information.
About Us
Zoomies help people stay connected so they can get more done together. We set out to build the best collaboration platform for the enterprise, and today help people communicate better with products like Zoom Contact Center, Zoom Phone, Zoom Events, Zoom Apps, Zoom Rooms, and Zoom Webinars. We’re problem-solvers, working at a fast pace to design solutions with our customers and users in mind. Find room to grow with opportunities to stretch your skills and advance your career in a collaborative, growth-focused environment.
Our Commitment
At Zoom, we believe great work happens when people feel supported and empowered. We’re committed to fair hiring practices that ensure every candidate is evaluated based on skills, experience, and potential. If you require an accommodation during the hiring process, let us know—we’re here to support you at every step.
Our interviews are supported by BrightHire, a tool that helps us create a consistent and thoughtful interview experience and may include recordings. Please refer to candidate privacy statement for more information of how we use your data.
#J-18808-Ljbffr- ...work from home day is currently Tuesday.Engineering at Lambda is responsible for building and... ...and networking teams to improve service reliability and deployment workflowsDeploy and... ...rotationYouHave 5+ years of experience in Site Reliability Engineering, Production Engineering...SeniorWork at officeLocal areaWork from homeFlexible hours
- ...Lambda’s designated work from home day is currently Tuesday.Engineering at Lambda is responsible for building and scaling our cloud offering... ...and SLIs for Kubernetes services, workloads, and platform reliability.You6+ years of experience in a SRE, operations engineer, or...SeniorWork at officeLocal areaWork from homeFlexible hours
- Senior Site Reliability Engineer ILocationSan Jose, Costa Rica - RemoteSummary of roleOwn availability, the most important product feature, by continually striving for sustained operational excellence of Sumo’s planet-scale observability and security products. Work with...SeniorFlexible hours
$267k - $356k
...day is currently Tuesday.Lambda's Storage Engineering team is the backbone behind our world-... ...workloads in the industry, which means reliability and performance aren't just goals—they're... ...defined storage across new and existing sites using tools such as Ansible, Jenkins etc...SeniorWork experience placementWork at officeLocal areaWork from homeFlexible hours- LeanData helps the world’s fastest-growing companies automate, simplify, and accelerate revenue.We are looking for a Senior Site Reliability Engineer to lead the strategic evolution of our cloud infrastructure. Reporting directly to the SVP of Engineering, this role is...SeniorFull timeWork at office2 days per week
$168k - $270.25k
NVIDIA is leading the way in groundbreaking developments in Artificial Intelligence, High-Performance Computing, and Visualization... ...of artificial intelligence.Join our team at NVIDIA as a Senior Site reliability engineer focused on HPC storage and play a crucial role in...SeniorFull time$148k - $235.75k
...on the world.Join our team of innovative engineers who are building an AI Data Center AIOps... ...that turns raw, high-volume telemetry into reliable, job-centric insights and automation for... ...canary checks, post-deploy validation), and lead rollbacks/remediations when needed.Lead...SeniorFull time$152k - $241.5k
...and amazing people. NVIDIA is leading the way in groundbreaking... ...intelligence.We’re looking for a Senior SRE to join our Compute Farm... ...lifecycle management, fleet reliability/auto-healing, E2E observability... ...Perl, or Ruby.Mentored other engineers and influenced technical...SeniorFull time$101k - $161k
...prestigious awards, such as Best Engineering Team, Best Company for... ...Work WithWe’re looking for Site Reliability Engineers to join our growing... ...chance to be drive, develop, and lead projects in any of the... ...EngineeringExperience level: Mid-Senior LevelIndustry: Computer...Senior- ...Oracle Cloud Infrastructure (OCI) seeks a Senior Principal Engineer to lead the design and implementation of reliability validation for OCI control plane services, focusing on a high-performance, low-level systems approach. You will mentor engineers, define validation...Senior
$192.4k - $275.8k
...demanding enterprise customers, blending Site Reliability Engineering, Systems Engineering, and Service... ...you Your ImpactYou will be the most senior technical individual contributor on the... ...audiences 4+ years experience leading post-mortems and root cause analysis...SeniorFull timeTemporary workLocal areaFlexible hours- ...Bitdeer is a world-leading technology company for AI and Bitcoin mining infrastructure. Bitdeer is committed to providing comprehensive... ...the AIOps substrate The remediation-actuator and workflow engine land here — you make the control plane safe for automated...SeniorLocal area
$200k - $322k
...are seeking a highly skilled Senior Staff SRE to join our dynamic... ...endeavor!What you will be doing:Lead initiatives to transform IT... ...building for performance and reliability at global scale, covering... ...with NVIDIA leadership, senior engineers, program managers, and product...SeniorFull timeRemote work- ...We are seeking a Senior Database Reliability Engineer (DBRE) to design, operate, and improve reliable, scalable, secure, and highly available database... ...data platforms. The role combines database engineering, site reliability engineering, Linux systems administration, and...Senior
$132.6k - $214.5k
...you will collaborate closely with our engineering teams to develop innovative solutions that... ...’ performance and health. As a Senior Staff SRE with the Cortex Observability... ...operability of the product and ensure the reliability and availability of our services. Qualifications...SeniorFull timeWork at officeVisa sponsorshipWork visa$187.04k - $359.72k
...pushing for changes that improve reliability and velocity. Qualifications... ...Computer Science, Electrical Engineering, Computer Engineering or... ...drive. About USDS TikTok is the leading destination for short-form... ...Corporate Functions and more. On-site presence across teams allows...SeniorTemporary workLocal areaOverseasShift work$168k - $333.5k
...NVIDIA infrastructure. Work with NVIDIA's DGX Cloud team as a Senior Site Reliability Engineer to maintain high-performance DGX Cloud clusters for AI... ...pushing for changes that improve reliability and velocity. Lead triage and root-cause analysis of high-severity incidents....SeniorFull timeWorldwide$196k - $310.5k
...Silicon Codesign Group is seeking a versatile engineer to join the HW bring-up methodology team.... ...ambitious NPI and production schedules.Lead post-action reviews and convert learnings... ...of experience leading complex, multi-site technical programs in semiconductor, SoC,...SeniorFull time$160k - $253k
...software, operations, and partner technologies to build optimized and efficient AI factories. We are looking for a Senior Technical Marketing Engineer to lead and develop technical content for DSX. Want to create and tell an end-to-end story across the full stack? This...SeniorFull time$160k - $240k
...times a day - quickly, reliably, and securely. Any time... ...Fiserv.Job TitleSenior Site Reliability EngineerWhat... ...Site Reliability Engineer do at Fiserv?You will join... ...on-call rotations and lead incident response activities... ...or DevOps at a mid-to-senior level.Strong shell...SeniorFull time- ...robust product portfolio includes world leading MCUs, SoCs, Analog and power products, plus... ..., validation, packaging, and product engineering. Define chip-level architecture, integration... ...first-pass silicon execution. Mentor senior engineers and provide technical...SeniorWorldwideNight shift
$262k - $364k
...within the AViD ecosystem have reliability and uptime appropriate to... ...and performance.Build creative engineering solutions to operations and infrastructure... ...for production operations.Lead and contribute to the cross-... ...in a strategic way.Site Reliability Engineering (SRE)...Senior$222k - $300.5k
...TeamIntuit's Infrastructure and Site Reliability organization owns the... .... The Fintech Platform Systems Engineering team builds and operates the AWS... ....The OpportunityWe're hiring a Senior Manager, Site Reliability Engineering to lead a hands-on team of 10-15 systems...SeniorWorldwideShift work$145k - $165k
...A technology solutions firm in Sunnyvale, CA is looking for a highly experienced Site Reliability Engineer (SRE). This role involves maintaining uptime and performance across systems. Exceptional Linux expertise and automation skills in Bash and Python are crucial. Key...Senior$207.4k - $259.2k
...built specifically for aviation. We’re seeking exceptional engineers, operators and builders to join us on our mission to build the... ...are seeking a highly experienced and passionate Sr. Staff Site Reliability Engineer (SRE) to join our growing team. In this critical role...SeniorPermanent employmentLocal areaVisa sponsorshipNight shift- WD, a leader in data infrastructure, seeks a detail-orientedSupply Chain Analyst to drive cross-functional initiatives across logistics, planning, ERP, revenue and external audits. You will document processes, build analytics, and support system enhancements while identifying...Senior
$180.2k - $297.2k
...us!Role and ResponsibilitiesAs a Senior Staff Physical Implementation CAD Engineer, you will shape the development and... ...contributor role, you will lead a team of engineers focusing on physical... ..., and area (PPA), silicon reliability, and design turnaround time for GPU...SeniorHourly payFull timeWorldwideRelocation- Waters Corporation is seeking a Sr. Principal, Lifecycle Program Lead in Milpitas, California. This role involves driving transformative improvements in instrument replacement processes across business units. You'll collaborate with sales and operational teams, lead cross...Senior
$191.5k - $239.4k
...is hiring a seasoned Product Manager to lead the international expansion of our core Accounts... ...execution as a connective force across engineering, design, data, finance, compliance,... ...culture, benefits, and teams on our career site, LinkedIn Life, or YouTube pages.BILL is...SeniorTemporary workWork at officeLocal areaRemote workFlexible hours$102.5k - $187.9k
EY is seeking a ServiceNow Senior Consultant to lead transformation teams and interact directly with clients. This role involves overseeing project phases from design to implementation. Candidates should possess robust experience in ServiceNow IT/OT Asset Management and...SeniorFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Lead Site Reliability Engineer. Be the first to apply!
- lead engineer San Jose, CA
- lead operating engineer San Jose, CA
- site reliability engineer sre San Jose, CA
- site reliability engineer San Jose, CA
- senior technical service engineer San Jose, CA
- senior director product management San Jose, CA
- senior vice president human resources San Jose, CA
- senior automation controls engineer San Jose, CA
- senior grant accountant San Jose, CA
- senior tax San Jose, CA


