Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior SRE: Live Streaming CDN Reliability & Automation

$100k

Netflix, Inc.

We are seeking a seasoned Reliability Engineer with extensive experience in *nix, networking, data analysis, and large-scale platform operations experience to design, scale, operate, automate, and analyze our globally distributed CDN. Come join us and play a meaningful role in our journey to entertain the world! Responsibilities Drive continual improvement in resiliency, observability, monitoring, instrumentation, and automation with the primary goal of maintaining a highly scalable and reliable CDN platform worldwide. Aggregate, analyze and correlate large amounts of server and application performance data. Use the innovative Netflix Big Data platform as a highly flexible, specialized, and efficient toolset to identify opportunities for platform optimization and system reliability improvements, as well as identifying patterns/anomalies for further investigation. Provide technical design and engineering assistance to ISP partners to integrate our Open Connect Appliances. Handle Tier 3 escalation and participate in an on-call rotation for the CDN platform production issues. Qualifications 5+ years of Service Reliability/Operational experience running large-scale, high-performance systems & internet services with a focus on performance and reliability. Preferred - B.S. in Computer Science, Electrical or Computer Engineering (or equivalent professional experience) Strong working knowledge of networking concepts and application protocols, especially TCP/IP, BGP, DNS, TLS, and with focused experience on CDNs and cache/proxy technologies Skilled in designing, creating, and maintaining automation written in a programming language such as Python Expert-level knowledge of managing and debugging Unix/Linux systems (engineering fundamentals, networking, storage, operating systems) at scale. Experience with distributed analytic processing technologies (Hive, Presto/Trino, Spark SQL, etc) Strong understanding of applied statistics and the ability to code systems that identify outlier behavior in large systems Some experience with container and container orchestration technologies (Docker, Kubernetes) Ability to work in a highly collaborative environment and to communicate cross-functionally with internal and external partners Things that show how we think FreeBSD optimization used by Netflix to serve 800 Gb/s from a single server Resiliency Practices in Managing CDN Measuring Real-Life Latency of the Internet: A Netflix Story Mastering Near-Real Time Telemetry and Big Data Does this sound interesting? Or does this sound interesting but intimidating? Please don’t self-select out, let’s figure it out together. We’d love to talk to you! Netflix is a global company with a diverse member base, which is why the content we produce reflects that: global perspectives and global stories. As we grow globally, we must have the most talented employees with diverse backgrounds, cultures, perspectives, and experiences to support our innovation and creativity. We are an equal-opportunity employer and strive to build balanced teams from all walks of life. Our culture is unique, and we tend to live by our values, so it’s worth learning more about Netflix here. At Netflix, we carefully consider a wide range of compensation factors to determine your personal top of market. We rely on market indicators to determine compensation and consider your specific job, skills, and experience to get it right. These considerations can cause your compensation to vary and will also be dependent on your location. The overall market range for roles in this area of Netflix is typically $100,000 - $720,000. This market range is based on total compensation (vs. only base salary), which is in line with our compensation philosophy. #J-18808-Ljbffr Netflix, Inc.

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Senior SRE: Live Streaming CDN Reliability & Automation in Los Gatos, CA vacancy
  • NVIDIA is looking for a Senior Systems Software Engineer (SRE) in Santa Clara, California to design and maintain...  ...foundation in infrastructure automation and distributed systems, ensuring...  ...GPU cloud services deliver maximum reliability and uptime. You will also enhance... 
    Senior

    NVIDIA

    Santa Clara, CA
    1 day ago
  • $120k - $145k

    Fortinet, Inc. is seeking a Staff SRE to scale FortiSASE’s cloud infrastructure. The ideal candidate will have over 7 years of SRE...  ...initiatives across teams, optimizing performance, and improving reliability. The position offers a salary range of $120,000-$145,000, along... 
    Senior

    Fortinet, Inc.

    Sunnyvale, CA
    4 days ago
  • $118k - $170k

    RUCKUS Networks is seeking a Senior Site Reliability Engineer to enhance platform reliability and efficiency. This hybrid role requires being on-site in Sunnyvale, CA, 3 days a week, engaging with teams to solve production problems and improve operational readiness. The... 
    Senior
    3 days per week

    RUCKUS Networks

    Sunnyvale, CA
    4 days ago
  • Google in Sunnyvale, California is seeking a Senior Staff Software Engineer for Site Reliability Engineering. This role involves optimizing existing systems, building infrastructure, and managing large-scale challenges unique to Google. The ideal candidate will have significant... 
    Senior

    Google

    Sunnyvale, CA
    4 days ago
  • $118k - $170k

    Vistance Networks in Sunnyvale, California, is seeking a Senior Site Reliability Engineer to enhance platform reliability and customer experience across cloud services. You will troubleshoot and improve scalable services while working closely with cross-functional teams... 
    Senior
    3 days per week

    Vistance Networks

    Sunnyvale, CA
    4 days ago
  • A technology services company is seeking a Senior Site Reliability Engineer / DevOps Engineer in Sunnyvale, CA. The ideal candidate will have over 8 years of experience in DevOps, expertise in Docker and Kubernetes, and proficiency with Terraform or Ansible. Responsibilities... 
    Senior

    Donato Technologies Inc

    Sunnyvale, CA
    22 hours ago
  • $166k - $244k

    Overview Site Reliability Engineering (SRE) combines software and systems engineering to build and run...  ...infrastructure and eliminating work through automation. On the SRE team, you\'ll have the...  .... Monitor services once they are live by measuring availability, latency, and... 
    Senior
    Full time

    Google

    Sunnyvale, CA
    4 days ago
  • $145k - $165k

     ...firm in Sunnyvale, CA is looking for a highly experienced Site Reliability Engineer (SRE). This role involves maintaining uptime and performance across systems. Exceptional Linux expertise and automation skills in Bash and Python are crucial. Key responsibilities include... 
    Senior

    Bolt Graphics, Inc.

    Sunnyvale, CA
    3 days ago
  •  ...for its iOS/iPadOS/tvOS applications, located in Los Gatos, California. This role emphasizes building tools for automation and ensuring a high-quality streaming experience. The ideal candidate has strong skills in Swift and Objective-C, with experience in media playback... 
    Senior

    Netflix

    Los Gatos, CA
    4 days ago
  •  ...Role: Site Reliability Engineer (SRE) Location: Santa Clara Valley (Cupertino), California, Hybrid. Duration: 6+ Months...  ...and run systems, infrastructure and applications through automation. Participate in periodic on-call duties Must have... 

    Zortech Solutions

    Santa Clara, CA
    2 days ago
  • $118k - $170k

    Site Reliability Engineer - Senior Staff Req ID: 81736 Location: Sunnyvale, California, United States, 9...  .... You will help improve reliability, automation, observability, and customer experience...  ...while working with modern cloud and SRE technologies in a collaborative... 
    Senior
    Work at office
    Relocation
    3 days per week

    Vistance Networks

    Sunnyvale, CA
    4 days ago
  • $150k - $195k

     ..., California, seeks engineers passionate about automation. You will enhance the Lacework Cloud Security Platform...  ...design, and automated tooling for reliable deployments. Candidates should have 3+ years in DevOps/SRE, strong skills in Kubernetes, Terraform, and cloud... 

    Fortinet, Inc.

    Sunnyvale, CA
    4 days ago
  • $101k - $161k

    Senior Site Reliability Engineer (SRE) - CloudVision as a Service (CVaaS) Full-time Arista Networks is an industry...  ...deeply believe in building highly automated and self-sustaining environments,...  ...enterprise network management and streaming telemetry SaaS offering.... 
    Senior
    Full time

    Arista Networks

    Santa Clara, CA
    4 days ago
  • Arista Networks seeks a Senior Site Reliability Engineer to join CloudVision-as-a-Service (CVaaS) in Santa Clara, CA. You will design and operate scalable, reliable cloud services with emphasis on automation, observability, and resilient deployments in a Kubernetes-native... 
    Senior

    Arista Networks

    Santa Clara, CA
    4 days ago
  • $201.6k - $302k

     ...Description The Role: As the Senior Engineering Manager for Hybrid Services & Reliability (HSR) within AV Core...  ...such as DHCP, PXE, and CDN, to ensure seamless...  ...Reliability Engineering (SRE) and defining SLO/SLI...  ...Required: Opinionated view on automated observability, incident... 
    Senior
    Local area
    Remote work
    Work from home
    Relocation
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    4 days ago
  • NVIDIA is seeking a Senior Site Reliability Engineer (SRE) to join the Compute Farm team in Santa Clara, California. This role involves building and...  ...solutions and cloud environments. Key responsibilities include automating service provisioning, conducting capacity management,... 
    Senior

    NVIDIA

    Santa Clara, CA
    4 days ago
  • $134.4k - $280k

    Omnissa, LLC is seeking a DevOps Lead, Cloud Automation & Reliability to guide a team of engineers in Mountain View, California. This leadership role requires 10+ years of experience including technical leadership and SaaS service management. The ideal candidate will drive... 
    Senior

    Omnissa, LLC

    Mountain View, CA
    22 hours ago
  •  ...is looking for multiple Kubernetes Site Reliability Engineers to join their remote team. This...  ..., improving reliability through automation, and developing Infrastructure as Code with...  ...practices. The position welcomes mid-level, senior, and technical leadership candidates, aiming... 
    Remote job

    Eitacies Inc

    Santa Clara, CA
    4 days ago
  • EITACIES Inc. is hiring a Kubernetes Site Reliability Engineer to work fully remote. The role...  ...improving platform reliability through automation and Infrastructure as Code practices. Candidates...  ...This position is ideal for mid-level or senior professionals looking to make an impact... 
    Remote job

    EITACIES Inc.

    Santa Clara, CA
    4 days ago
  •  ...AI/ML Engineer to lead the development of advanced AI systems that ensure reliability and operational excellence across its tech ecosystem. You will design and implement intelligent automation solutions impacting millions of users. This role involves architecting next... 
    Senior

    Walmart

    Sunnyvale, CA
    4 days ago
  • $248k - $396.75k

    Site Reliability Engineering (SRE) at NVIDIA is an engineering discipline to design, build and maintain...  ...focuses on eliminating manual work through automation, performance tuning and growing...  ...refinement Support services before they go live through activities such as system... 

    Nvidia Corporation

    Santa Clara, CA
    4 days ago
  • Roku is seeking a Senior Software Engineer specializing in Python Automation to enhance its products. In this role, you will collaborate with various teams to develop automated tests and improve product quality. The ideal candidate has 5+ years in Software Engineering,... 
    Senior

    Roku

    Los Gatos, CA
    4 days ago
  • Fortinet is seeking a talented Site Reliability Engineer to join our engineering team in the United States. This hands‑on role focuses on building, maintaining, and troubleshooting cloud service clusters, infrastructure, and monitoring systems to ensure high availability... 
    Senior

    Fortinet

    Sunnyvale, CA
    4 days ago
  • Senior Software Engineer, Python Automation Roku 15 March 2025 Teamwork makes the stream work. Roku is changing how the world watches TV. Roku is the #1 TV streaming platform...  ...and features for the industry's most reliable streaming media platform. Our goal is to help... 
    Senior
    Local area
    Remote work

    Roku

    Los Gatos, CA
    4 days ago
  • $207k - $300k

     ...Sunnyvale, CA is seeking a Software Engineering Manager II for Site Reliability Engineering. You'll lead a team to ensure uptime and optimize...  ..., and performance of key services. With a focus on automation and system reliability, the ideal candidate will possess extensive... 

    Google Inc.

    Sunnyvale, CA
    22 hours ago
  • Kody is seeking a Senior Site Reliability Engineer (Payments Infrastructure) in Sunnyvale, California, to ensure the reliability and operational...  ...SLOs, and driving reliability improvements through automation and optimizations. The ideal candidate has strong hands-on... 
    Senior

    Kody

    Sunnyvale, CA
    4 days ago
  •  ...client applications and platforms on Smart TVs, Set-top boxes, and streaming players. You will partner with the Client & Partner...  ...with engineering from platform teams to ensure rapid product innovation and reliable deployment across devices. #J-18808-Ljbffr Netflix
    Senior

    Netflix

    Los Gatos, CA
    3 days ago
  • Apple Inc. is seeking a Senior Site Reliability Engineer in Cupertino to drive reliability, scalability, and observability of our cloud platform...  ...performance for millions of users. The role emphasizes automation, CI/CD, incident response, and capacity planning, with on-... 
    Senior

    Apple Inc.

    Cupertino, CA
    1 day ago
  • A leading technology firm is in search of a Senior Wireless Network Site Reliability Engineer to manage and enhance their wireless network infrastructure...  ...high availability of the network, implementing automation, and monitoring performance to ensure client satisfaction... 
    Senior

    TechDigital Group

    Santa Clara, CA
    2 days ago
  • $165.5k - $289.6k

    ServiceNow is seeking a Sr Staff Site Reliability Engineer to lead critical infrastructure initiatives in Santa Clara, California. The ideal candidate will have over 7 years of experience in Site Reliability Engineering and a strong background in AWS and Kubernetes. You... 
    Senior
    Flexible hours

    ServiceNow

    Santa Clara, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior SRE: Live Streaming CDN Reliability & Automation. Be the first to apply!