Senior SRE: Live Streaming CDN Reliability & Automation
$100kNetflix, Inc.
We are seeking a seasoned Reliability Engineer with extensive experience in *nix, networking, data analysis, and large-scale platform operations experience to design, scale, operate, automate, and analyze our globally distributed CDN. Come join us and play a meaningful role in our journey to entertain the world! Responsibilities Drive continual improvement in resiliency, observability, monitoring, instrumentation, and automation with the primary goal of maintaining a highly scalable and reliable CDN platform worldwide. Aggregate, analyze and correlate large amounts of server and application performance data. Use the innovative Netflix Big Data platform as a highly flexible, specialized, and efficient toolset to identify opportunities for platform optimization and system reliability improvements, as well as identifying patterns/anomalies for further investigation. Provide technical design and engineering assistance to ISP partners to integrate our Open Connect Appliances. Handle Tier 3 escalation and participate in an on-call rotation for the CDN platform production issues. Qualifications 5+ years of Service Reliability/Operational experience running large-scale, high-performance systems & internet services with a focus on performance and reliability. Preferred - B.S. in Computer Science, Electrical or Computer Engineering (or equivalent professional experience) Strong working knowledge of networking concepts and application protocols, especially TCP/IP, BGP, DNS, TLS, and with focused experience on CDNs and cache/proxy technologies Skilled in designing, creating, and maintaining automation written in a programming language such as Python Expert-level knowledge of managing and debugging Unix/Linux systems (engineering fundamentals, networking, storage, operating systems) at scale. Experience with distributed analytic processing technologies (Hive, Presto/Trino, Spark SQL, etc) Strong understanding of applied statistics and the ability to code systems that identify outlier behavior in large systems Some experience with container and container orchestration technologies (Docker, Kubernetes) Ability to work in a highly collaborative environment and to communicate cross-functionally with internal and external partners Things that show how we think FreeBSD optimization used by Netflix to serve 800 Gb/s from a single server Resiliency Practices in Managing CDN Measuring Real-Life Latency of the Internet: A Netflix Story Mastering Near-Real Time Telemetry and Big Data Does this sound interesting? Or does this sound interesting but intimidating? Please don’t self-select out, let’s figure it out together. We’d love to talk to you! Netflix is a global company with a diverse member base, which is why the content we produce reflects that: global perspectives and global stories. As we grow globally, we must have the most talented employees with diverse backgrounds, cultures, perspectives, and experiences to support our innovation and creativity. We are an equal-opportunity employer and strive to build balanced teams from all walks of life. Our culture is unique, and we tend to live by our values, so it’s worth learning more about Netflix here. At Netflix, we carefully consider a wide range of compensation factors to determine your personal top of market. We rely on market indicators to determine compensation and consider your specific job, skills, and experience to get it right. These considerations can cause your compensation to vary and will also be dependent on your location. The overall market range for roles in this area of Netflix is typically $100,000 - $720,000. This market range is based on total compensation (vs. only base salary), which is in line with our compensation philosophy. #J-18808-Ljbffr Netflix, Inc.
- NVIDIA is looking for a Senior Systems Software Engineer (SRE) in Santa Clara, California to design and maintain... ...foundation in infrastructure automation and distributed systems, ensuring... ...GPU cloud services deliver maximum reliability and uptime. You will also enhance...Senior
$120k - $145k
Fortinet, Inc. is seeking a Staff SRE to scale FortiSASE’s cloud infrastructure. The ideal candidate will have over 7 years of SRE... ...initiatives across teams, optimizing performance, and improving reliability. The position offers a salary range of $120,000-$145,000, along...Senior$118k - $170k
RUCKUS Networks is seeking a Senior Site Reliability Engineer to enhance platform reliability and efficiency. This hybrid role requires being on-site in Sunnyvale, CA, 3 days a week, engaging with teams to solve production problems and improve operational readiness. The...Senior3 days per week- Google in Sunnyvale, California is seeking a Senior Staff Software Engineer for Site Reliability Engineering. This role involves optimizing existing systems, building infrastructure, and managing large-scale challenges unique to Google. The ideal candidate will have significant...Senior
$118k - $170k
Vistance Networks in Sunnyvale, California, is seeking a Senior Site Reliability Engineer to enhance platform reliability and customer experience across cloud services. You will troubleshoot and improve scalable services while working closely with cross-functional teams...Senior3 days per week- A technology services company is seeking a Senior Site Reliability Engineer / DevOps Engineer in Sunnyvale, CA. The ideal candidate will have over 8 years of experience in DevOps, expertise in Docker and Kubernetes, and proficiency with Terraform or Ansible. Responsibilities...Senior
$166k - $244k
Overview Site Reliability Engineering (SRE) combines software and systems engineering to build and run... ...infrastructure and eliminating work through automation. On the SRE team, you\'ll have the... .... Monitor services once they are live by measuring availability, latency, and...SeniorFull time$145k - $165k
...firm in Sunnyvale, CA is looking for a highly experienced Site Reliability Engineer (SRE). This role involves maintaining uptime and performance across systems. Exceptional Linux expertise and automation skills in Bash and Python are crucial. Key responsibilities include...Senior- ...for its iOS/iPadOS/tvOS applications, located in Los Gatos, California. This role emphasizes building tools for automation and ensuring a high-quality streaming experience. The ideal candidate has strong skills in Swift and Objective-C, with experience in media playback...Senior
- ...Role: Site Reliability Engineer (SRE) Location: Santa Clara Valley (Cupertino), California, Hybrid. Duration: 6+ Months... ...and run systems, infrastructure and applications through automation. Participate in periodic on-call duties Must have...
$118k - $170k
Site Reliability Engineer - Senior Staff Req ID: 81736 Location: Sunnyvale, California, United States, 9... .... You will help improve reliability, automation, observability, and customer experience... ...while working with modern cloud and SRE technologies in a collaborative...SeniorWork at officeRelocation3 days per week$150k - $195k
..., California, seeks engineers passionate about automation. You will enhance the Lacework Cloud Security Platform... ...design, and automated tooling for reliable deployments. Candidates should have 3+ years in DevOps/SRE, strong skills in Kubernetes, Terraform, and cloud...$101k - $161k
Senior Site Reliability Engineer (SRE) - CloudVision as a Service (CVaaS) Full-time Arista Networks is an industry... ...deeply believe in building highly automated and self-sustaining environments,... ...enterprise network management and streaming telemetry SaaS offering....SeniorFull time- Arista Networks seeks a Senior Site Reliability Engineer to join CloudVision-as-a-Service (CVaaS) in Santa Clara, CA. You will design and operate scalable, reliable cloud services with emphasis on automation, observability, and resilient deployments in a Kubernetes-native...Senior
$201.6k - $302k
...Description The Role: As the Senior Engineering Manager for Hybrid Services & Reliability (HSR) within AV Core... ...such as DHCP, PXE, and CDN, to ensure seamless... ...Reliability Engineering (SRE) and defining SLO/SLI... ...Required: Opinionated view on automated observability, incident...SeniorLocal areaRemote workWork from homeRelocationRelocation packageFlexible hours- NVIDIA is seeking a Senior Site Reliability Engineer (SRE) to join the Compute Farm team in Santa Clara, California. This role involves building and... ...solutions and cloud environments. Key responsibilities include automating service provisioning, conducting capacity management,...Senior
$134.4k - $280k
Omnissa, LLC is seeking a DevOps Lead, Cloud Automation & Reliability to guide a team of engineers in Mountain View, California. This leadership role requires 10+ years of experience including technical leadership and SaaS service management. The ideal candidate will drive...Senior- ...is looking for multiple Kubernetes Site Reliability Engineers to join their remote team. This... ..., improving reliability through automation, and developing Infrastructure as Code with... ...practices. The position welcomes mid-level, senior, and technical leadership candidates, aiming...Remote job
- EITACIES Inc. is hiring a Kubernetes Site Reliability Engineer to work fully remote. The role... ...improving platform reliability through automation and Infrastructure as Code practices. Candidates... ...This position is ideal for mid-level or senior professionals looking to make an impact...Remote job
- ...AI/ML Engineer to lead the development of advanced AI systems that ensure reliability and operational excellence across its tech ecosystem. You will design and implement intelligent automation solutions impacting millions of users. This role involves architecting next...Senior
$248k - $396.75k
Site Reliability Engineering (SRE) at NVIDIA is an engineering discipline to design, build and maintain... ...focuses on eliminating manual work through automation, performance tuning and growing... ...refinement Support services before they go live through activities such as system...- Roku is seeking a Senior Software Engineer specializing in Python Automation to enhance its products. In this role, you will collaborate with various teams to develop automated tests and improve product quality. The ideal candidate has 5+ years in Software Engineering,...Senior
- Fortinet is seeking a talented Site Reliability Engineer to join our engineering team in the United States. This hands‑on role focuses on building, maintaining, and troubleshooting cloud service clusters, infrastructure, and monitoring systems to ensure high availability...Senior
- Senior Software Engineer, Python Automation Roku 15 March 2025 Teamwork makes the stream work. Roku is changing how the world watches TV. Roku is the #1 TV streaming platform... ...and features for the industry's most reliable streaming media platform. Our goal is to help...SeniorLocal areaRemote work
$207k - $300k
...Sunnyvale, CA is seeking a Software Engineering Manager II for Site Reliability Engineering. You'll lead a team to ensure uptime and optimize... ..., and performance of key services. With a focus on automation and system reliability, the ideal candidate will possess extensive...- Kody is seeking a Senior Site Reliability Engineer (Payments Infrastructure) in Sunnyvale, California, to ensure the reliability and operational... ...SLOs, and driving reliability improvements through automation and optimizations. The ideal candidate has strong hands-on...Senior
- ...client applications and platforms on Smart TVs, Set-top boxes, and streaming players. You will partner with the Client & Partner... ...with engineering from platform teams to ensure rapid product innovation and reliable deployment across devices. #J-18808-Ljbffr NetflixSenior
- Apple Inc. is seeking a Senior Site Reliability Engineer in Cupertino to drive reliability, scalability, and observability of our cloud platform... ...performance for millions of users. The role emphasizes automation, CI/CD, incident response, and capacity planning, with on-...Senior
- A leading technology firm is in search of a Senior Wireless Network Site Reliability Engineer to manage and enhance their wireless network infrastructure... ...high availability of the network, implementing automation, and monitoring performance to ensure client satisfaction...Senior
$165.5k - $289.6k
ServiceNow is seeking a Sr Staff Site Reliability Engineer to lead critical infrastructure initiatives in Santa Clara, California. The ideal candidate will have over 7 years of experience in Site Reliability Engineering and a strong background in AWS and Kubernetes. You...SeniorFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior SRE: Live Streaming CDN Reliability & Automation. Be the first to apply!
- senior magento developer Los Gatos, CA
- senior manager pmo Los Gatos, CA
- senior manager legal Los Gatos, CA
- senior human factors engineer Los Gatos, CA
- senior aws cloud engineer Los Gatos, CA
- senior tableau developer Los Gatos, CA
- senior project manager contract Los Gatos, CA
- senior cloud data engineer Los Gatos, CA
- senior manager m&a tax Los Gatos, CA
- senior accountant remote Los Gatos, CA

