Staff Site Reliability Engineer
$210.5k - $263.1kCrunchyroll
About Crunchyroll
Founded by fans, Crunchyroll delivers the art and culture of anime to a passionate community. We super-serve over 100 million anime and manga fans across 200+ countries and territories, and help them connect with the stories and characters they crave. Whether that experience is online or in-person, streaming video, theatrical, games, merchandise, events and more, it’s powered by the anime content we all love.
Join our team, and help us shape the future of anime!
About the role
We are hiring a Staff Site Reliability Engineer (SRE) to join the Center for Data & Insights (CDI) in the US and play a critical role in advancing the reliability, scalability, performance, and security of Crunchyroll’s consumer-facing data platforms. As a senior technical leader, you will partner closely with Engineering, Data, Infrastructure, Product, and Security teams to design and operate resilient cloud-native systems that power critical business and customer experiences. You will drive initiatives across observability, incident management, automation, capacity planning, disaster recovery, and operational excellence while helping teams adopt modern SRE practices such as SLIs, SLOs, and error budgets.
The ideal candidate combines deep expertise in large-scale distributed systems with a strong sense of ownership, collaboration, and service leadership. You are passionate about building highly reliable platforms, eliminating operational toil through automation, and enabling engineering teams to move quickly and safely. In addition, you will champion SecOps best practices by driving vulnerability management, supporting penetration testing initiatives, improving security observability, strengthening cloud and Kubernetes security controls, and ensuring operational readiness for emerging threats. This is a unique opportunity to shape reliability and security engineering practices across CDI while helping build a world-class data and insights ecosystem that enables informed decision-making throughout Crunchyroll.
Core Areas of Responsibility
- Reliability Engineering : Define, measure, and continuously improve the reliability, availability, and performance of CDI platforms through SLIs, SLOs, and error budgets.
- Operational Excellence : Establish and drive best practices for incident management, root cause analysis, postmortems, and service ownership across engineering teams.
- Observability & Monitoring : Build and evolve comprehensive monitoring, logging, tracing, and alerting capabilities to enable proactive issue detection and rapid resolution.
- Automation : Identify operational inefficiencies and develop automation, self-service capabilities, and self-healing mechanisms to improve engineering productivity.
- Platform Scalability : Design and optimize cloud-native infrastructure and services to support growing business demands while maintaining performance and cost efficiency.
- Infrastructure Engineering : Drive Infrastructure as Code (IaC), platform standardization, and deployment automation to improve consistency, reliability, and operational agility.
- Capacity Planning & Performance : Lead capacity planning and performance optimization initiatives to ensure platforms can scale predictably and efficiently.
- Disaster Recovery & Resilience : Develop and regularly validate disaster recovery, backup, and business continuity strategies to ensure platform resiliency.
- Security Operations (SecOps) : Partner with Crunchyroll’s security team to integrate security controls, operational risk management, and security best practices into platform operations and engineering workflows.
- Vulnerability Management : Own the triage and remediation of identified vulnerabilities across infrastructure, platform, container, and application security vulnerabilities through established Crunchyroll vulnerability management processes.
- Penetration Testing & Security Remediation : Support penetration test scoping activities by providing technical context on CDI platforms. Own the triage, prioritization, and remediation of resulting findings to drive timely resolution and strengthen platform security posture.
- Cloud & Kubernetes Security : Implement and maintain secure cloud, container, and Kubernetes environments following least-privilege, defense-in-depth, and Zero Trust principles.
- Cross-Functional Leadership : Collaborate with Engineering, Data, Product, Infrastructure, and Security teams to drive reliability, scalability, and security initiatives across CDI.
- Mentorship & Engineering Excellence : Mentor engineers and champion a culture of operational excellence, reliability, ownership, continuous improvement, and security awareness.
About You
We get excited about candidates like you, because…
- 12+ years of experience in Site Reliability Engineering (SRE), Platform Engineering, Infrastructure Engineering, or related disciplines, with a proven track record of operating and scaling production-critical systems.
- Deep expertise in Kubernetes and GCP , including the design, deployment, and operation of highly available, cloud-native platforms at scale.
- Strong Infrastructure as Code (IaC) experience , preferably with Terraform, and a commitment to automation, standardization, and operational efficiency.
- Solid foundation in Linux systems administration, networking, and distributed systems , with the ability to troubleshoot complex production issues across multiple layers of the technology stack.
- Proficiency in one or more programming and scripting languages , such as Go, Python, Java, or Shell, with a focus on automation and platform engineering.
- Hands-on experience with modern observability platforms and practices , including Prometheus, Grafana, OpenTelemetry, Datadog, or equivalent monitoring and telemetry solutions.
- Demonstrated expertise in incident management, service reliability, capacity planning, performance optimization, and operational excellence , including the implementation of SLIs, SLOs, and error budgets.
- Strong understanding of cloud and platform security , including container security, Kubernetes security, CI/CD security, vulnerability management, and secure infrastructure operations.
- Good to have knowledge of security frameworks and best practices , including OWASP Top 10, Identity and Access Management (IAM), secrets management, Secure Software Development Lifecycle (SSDLC), and security-by-design principles.
- Excellent collaboration, communication, and technical leadership skills , with experience influencing architectural decisions, driving cross-functional initiatives, and mentoring engineers in reliability and operational best practices.
About the Team
The Center for Data and Insights (CDI) is a service-oriented, horizontal organization uniquely positioned within the company to serve as the trusted, unbiased source of timely, data-driven insights for Crunchyroll. Our vision is to inspire, support, and guide our stakeholders to be data-aware and build the systems of intelligence to discover insights and act on them. We have built a highly functional organization that truly believes in being a responsive partner, with the utmost curiosity, unwavering accountability and the courage to lead with actions.
Why you will love working at Crunchyroll
In addition to getting to work with fun, passionate and inspired colleagues, you will also enjoy the following benefits and perks:
- Receive a great compensation package including salary plus performance bonus earning potential, paid annually.
- Flexible time off policies allowing you to take the time you need to be your whole self.
- Generous medical, dental, vision, STD, LTD, and life insurance
- Health Saving Account HSA program
- Health care and dependent care FSA
- 401(k) plan, with employer match
- Employer paid commuter benefit
- Support program for new parents
- Pet insurance and some of our offices are pet friendly!
#LifeAtCrunchyroll #LI-Hybrid
The Pay Range for this position is listed. Actual pay will vary based on factors including, but not limited to location, experience, and performance. The range listed is just one component of Crunchyroll’s Total Rewards offerings for employees. Other rewards may include performance bonuses, employer matched retirement savings, time-off programs, and progressive health benefits and perks.
Pay Transparency - Los Angeles, CA
$210,500—$263,100 USD
About our Values
We want to be everything for someone rather than something for everyone and we do this by living and modeling our values in all that we do. We value
-
Courage. We believe that when we overcome fear, we enable our best selves.
-
Curiosity. We are curious, which is the gateway to empathy, inclusion, and understanding.
-
Kaizen. We have a growth mindset committed to constant forward progress.
-
Service. We serve our community with humility, enabling joy and belonging for others.
Our commitment to diversity and inclusion
Our mission of helping people belong reflects our commitment to diversity & inclusion. It’s just the way we do business.
We are an equal opportunity employer and value diversity at Crunchyroll. Pursuant to applicable law, we do not discriminate on the basis of race, religion, color, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status.
Crunchyroll, LLC is an independently operated joint venture between US-based Sony Pictures Entertainment, and Japan’s Aniplex, a subsidiary of Sony Music Entertainment (Japan) Inc., both subsidiaries of Tokyo-based Sony Group Corporation.
Questions about Crunchyroll’s hiring process? Please check out our Hiring FAQs:
Please refer to ourCandidate Privacy Policy for more information about how we process your personal information, and your data protection rights:
#J-18808-Ljbffr$125k - $145k
...SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SITE RELIABILITY ENGINEER, GNCSpaceX’s mission is to make humanity multiplanetary by developing fully and rapidly reusable launch systems capable of...SuggestedPermanent employmentTemporary workFlexible hoursWeekend work$165k - $265k
...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARLINK)At SpaceX we’re leveraging our experience in building rockets and spacecraft to deploy Starlink, the world’s most...SuggestedPermanent employmentTemporary workWorldwideWeekend work- ...SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SITE RELIABILITY ENGINEER (HIGH PERFORMANCE COMPUTING)SpaceX HPC is a shared compute platform used across the company — vehicle and structures...SuggestedPermanent employmentTemporary workWeekend work
- ...your big ideas, and your desire to team up with some of the best and brightest in technology and entertainment. The RoleThe Site Reliability Engineer (SRE) II is responsible for designing, implementing, and maintaining scalable and reliable systems and applications. Focus...SuggestedFull timeLocal areaWorldwideFlexible hours
$165k - $265k
...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER - TOP SECRET CLEARANCE (STARLINK)At SpaceX we’re leveraging our experience in building rockets and spacecraft to deploy Starlink...SuggestedPermanent employmentTemporary workWorldwideWeekend work$155k - $195k
...you to join us on our mission of providing humankind access to the galaxy beyond our planet. About the RoleWe are seeking a Site Reliability Engineer to join our Ground Software team. As a Site Reliability Engineer, you will design, build, and operate the ground and site...Permanent employmentFull timeWork at office$125k - $150k
...possible, with the ultimate goal of enabling human life on Mars.SITE RELIABILITY ENGINEER (RAPTOR)SpaceX is looking for a Site Reliability Engineer... ...infrastructure systems.Work with propulsion engineering staff to solve critical bottlenecks.Coordinate and communicate with...Permanent employmentTemporary work$145k - $195k
...SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SITE RELIABILITY ENGINEER (TOP SECRET CLEARANCE)As a member of the Classified IT Systems Engineering team, the Site Reliability Engineer is involved...Permanent employmentTemporary workWeekend work$165k - $270k
...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARSHIELD)Starshield leverages SpaceX’s Starlink technology and launch capability to support national security efforts....Permanent employmentTemporary workImmediate startWeekend work$164k - $270k
...exactly who we’re looking for.The Role What You’ll DoOwn the reliability of our robotics systems, from PLCs through ROS2/middleware to... ...automated remediation.Partner with controls, robotics, and platform engineering teams to bake reliability in early. Review designs, develop...Permanent employmentFull timeRelocation packageFlexible hours$175k - $285k
Hadrian - Manufacturing the FutureHadrian is building autonomous factories to reindustrialize America. By combining AI, advanced software, robotics, and full-stack manufacturing, we help aerospace and defense companies build rockets, satellites, aircraft, ships, and other...Permanent employmentFull timeRemote workRelocation packageFlexible hours$160k - $200k
...what it means for businesses and their employees to truly feel safe. POSITION OVERVIEW: HiveWatch is seeking a Senior Site Reliability Engineer to join our Platform Team, where you'll build and operate mission-critical edge infrastructure that connects our SaaS...Flexible hours- ...join us on our mission of providing humankind access to the galaxy beyond our planet. About the Role We are seeking a Site Reliability Engineer to join our Ground Software team. As a Site Reliability Engineer, you will design, build, and operate the ground and site...Full timeWork at office
- ...teammates, and make a direct, measurable impact on our mission and the future of the company - regardless of your function. Site Reliability Engineer We are seeking a Site Reliability Engineer to own the reliability, performance, observability, and operational health...Contract workFor contractorsWork at officeLocal area
$81.5k - $141.3k
...branded generic medicines. Our 122,000 colleagues serve people in more than 160 countries. JOB DESCRIPTION: Position Title: Site Reliability Engineer II Team: CRM DevOps Employment Type: Full-Time About the Role This Site Reliability Engineer II position works on-site...Full timeRemote workShift work$81.5k - $141.3k
...generic medicines. Our 122,000 colleagues serve people in more than 160 countries. JOB DESCRIPTION: Position Title: Site Reliability Engineer II Team: CRM DevOps Employment Type: Full-Time About the Role This Site Reliability Engineer II position works...Full timeRemote workShift work$107.8k - $162k
...an expectation of a minimum of three days per week working in the office and flexibility to work remotely on the remaining days. On-site expectations may evolve over time to support business needs, with clear communication provided in advance. Job Description Operates...Work at officeLocal areaRemote work3 days per week$140k - $180k
...fundamentally different class of spacecraft. Engineered to survive the harshest radiation... ...create highly available, deployable, and reliable products Reduce operational toil through... ...experience in Software Engineering, Site Reliability Engineering or DevOps ~ Deep...Permanent employmentShift work$180k - $200k
...to you through an Ateme solution created by our award-winning engineering teams. Ateme (PARIS: ATEME) is the global leader in video... ...Culture: Collaborate with talented international teams that value reliability, innovation, knowledge sharing, and continuous improvement....- ...on one unified cloud. One cloud for compute, inference, and agents. Role Overview We are seeking a skilled Site Reliability Engineer to join the GMI Global Infrastructure team. This role is hands-on and critical to ensuring the stability, efficiency, and...
$181k - $225k
...Senior Site Reliability Engineer Los Angeles, CA Altruist is transforming the multi-trillion dollar wealth management industry by building an AI platform for wealth professionals. We partner with financial advisors nationwide, empowering them to grow, optimize time...Work at officeImmediate start3 days per week- ...Senior Site Reliability Engineer (SRE) Our client is a global technology consulting and digital solutions company that enables enterprises across industries to reimagine business models, accelerate innovation, and maximize growth by harnessing digital technologies....Local area
- ...Role: Site Reliability Engineering (SRE) Location: Los Angeles, CA Remote position Fulltime position JD Site Reliability Engineer Experience in Cloud platforms (AWS, Azure, Google Cloud) and hybrid environments. Proficiency...Full timeRemote work
$230k - $260k
...to operate with clarity, control, and confidence across the reimbursement journey. About The Role We’re hiring a Staff Site Reliability Engineer to define and strengthen how reliability, scalability, and operational excellence are built into Pivotal’s platform....Remote workFlexible hours$100k - $200k
...backed by top tier investors. Our lean, world-class team of engineers and operators is applying a first-principles approach to... ...culture of urgency, accountability and transparency. DevOps / Site Reliability Engineer We are seeking a highly capable DevOps / Site...Full timeWeekend work$139.9k - $199.3k
...Lead Site Reliability Engineer We're looking for talented professionals to join us in bringing smart money management and payment solutions to everyone's fingertips. This position is classified as structured hybrid, with an expectation of a minimum of three (3) days...Work experience placementWork at officeRemote work3 days per week$145k - $160k
...We are seeking a specialized Observability & Infrastructure Engineer to lead the monitoring, telemetry, and platform-as-code initiatives critical to our multi-region disaster recovery roadmap. You will architect and implement robust observability pipelines, ensure deep...Temporary workRemote workFlexible hours$138k - $234k
...compassionate world. About the Role As a Principal Site Reliability Engineer (SRE), you will own the end-to-end reliability, scale, and... ..."regular employees" refers to those who are not temporary staff, such as interns, and some benefits may not apply to employees...Work at officeLocal areaFlexible hours- ...Pivotal Health, a healthcare technology platform, is seeking a Staff Site Reliability Engineer to embed reliability and operational excellence into our production systems. You will lead hands-on engineering across cloud, observability, incident response, and automation...
- ..., and thrive! KēSTA I.T. is actively seeking a Principal Engineer for an immediate full-time opportunity with our industry creating... ...An innovative technology company is seeking experienced Site Reliability Engineers to take ownership of building reliable, scalable platforms...Full timeContract workImmediate startWork from homeFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff Site Reliability Engineer. Be the first to apply!
- assistant engineer Los Angeles, CA
- senior staff systems engineer Los Angeles, CA
- technology administrator Los Angeles, CA
- engineering aide Los Angeles, CA
- software engineer staff Los Angeles, CA
- staff engineer Los Angeles, CA
- site reliability engineer sre Los Angeles, CA
- site reliability engineer Los Angeles, CA
- official site Los Angeles, CA
- site merchandiser Los Angeles, CA


