Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Machine Learning Engineer, DevOps/SRE

$148.75k - $361k

Roku, Building C

Teamwork makes the stream work.Roku is changing how the world watches TVRoku is the #1 TV streaming platform in the U.S., Canada, and Mexico, and we've set our sights on powering every television in the world. Roku pioneered streaming to the TV. Our mission is to be the TV streaming platform that connects the entire TV ecosystem. We connect consumers to the content they love, enable content publishers to build and monetize large audiences, and provide advertisers unique capabilities to engage consumers.From your first day at Roku, you'll make a valuable - and valued - contribution. We're a fast-growing public company where no one is a bystander. We offer you the opportunity to delight millions of TV streamers around the world while gaining meaningful experience across a variety of disciplines.About the team The Advertising Performance group focuses on performance for all participants in the Advertising ecosystem - Advertisers, Publishers, and Roku. The systems and solutions span multiple disciplines and technologies to perform real-time multi-objective optimization across distributed systems at large scale and with low latency. We use Machine Learning, Reinforcement Learning, AI, Control and Optimization Systems, and Auction Dynamics to solve a large set of complex problems. At the core of this is our Machine Learning, Experimentation, and Inference Platform that powers the entire landscape, which we continuously evolve over time. About the role We are seeking a talented and experienced Senior Software Engineer, MLOps/DevOps, to join the Advertising Performance team and play a critical role in supporting and scaling our Machine Learning infrastructure. The ideal candidate has a strong background in DevOps/SRE practices, cloud infrastructure management, and MLOps tooling — with a passion for building platforms that accelerate ML experimentation and deployment at internet scale. You will partner closely with ML Scientists and Engineers to streamline the end-to-end ML lifecycle across training, evaluation, deployment, and monitoring — on top of a modern, cloud-native stack running on GCP and AWS using Kubernetes, Apache Airflow, Spark, Ray, MLflow, Chronon, etc.For California Only - The estimated annual salary for this position is between $148,750 - $361,000 annually. Compensation packages are based on factors unique to each candidate, including but not limited to skill set, certifications, and specific geographical location. This role is eligible for health insurance, equity awards, life insurance, disability benefits, parental leave, wellness benefits, and paid time off. What you’ll be doing Lead the design and operation of scalable, production-grade cloud infrastructure for ML workloads across AWS and GCP, including GPU/TPU-based training and inference environmentsArchitect and improve CI/CD systems for ML models and platform services to enable fast, reliable, and safe production releasesOwn and evolve low-latency infrastructure for real-time model inference, including KV store and vector databasesDefine and enforce observability standards for ML systems, including model performance monitoring, drift detection, capacity planning, and pipeline health metricsParticipate in on-call rotation, leading incident response and root-cause analysis for critical ML training and serving infrastructurePartner with data scientists and ML engineers to improve platform usability, accelerate model iteration, and implement strong MLOps and SRE best practicesChampion operational excellence across ML infrastructure through automation, resilience engineering, disaster recovery planning, and continuous improvementWe’re excited if you have BS or MS in Computer Science, Engineering, or a related quantitative fieldAbility to demonstrate AI tool (Claude, Cursor, ChatGPT etc.) coding efficiency at an advanced skill level8+ years of experience in DevOps, SRE, or ML infrastructure, including 4+ years supporting large-scale ML or AI systemsStrong programming skills in Python, and/or Scala, or Java for platform automation and toolingDeep experience with Kubernetes and container orchestration on GCP (GKE) and/or AWS (EKS)Expertise with NoSQL or low-latency data stores such as Aerospike or similar technologiesHands-on experience with data and orchestration technologies such as Apache Spark, Apache Flink, Apache Airflow, and KafkaExperience building and maintaining CI/CD systems using tools such as Jenkins or GitLab RunnerFamiliarity with feature engineering platforms such as Chronon and model lifecycle tools such as MLflowStrong infrastructure-as-code experience with Terraform or similar toolingExperience with observability platforms such as Prometheus, Grafana, and DatadogExcellent communication and cross-functional collaboration skillsExperience in the Advertising domain is a plus #LI-DH2What's Roku's approach to hybrid working?Roku fosters an inclusive and collaborative environment where teams generally work in the office Monday through Thursday. Fridays are generally flexible for remote work, except for employees whose specific roles or assigned office location require five days' a week attendance.What are some of the benefits?Roku is committed to offering a diverse range of benefits as part of our compensation package to support our employees and their families. Our comprehensive benefits include global access to mental health and financial wellness support and resources. Local benefits include statutory and voluntary benefits which may include healthcare (medical, dental, and vision), life, accident, disability, commuter, and retirement options (401(k)/pension). Employees are supported in taking time off, in accordance with local leave policies and other personal needs to support their evolving work and life needs. It's important to note that not every benefit is available in all locations or for every role. For details specific to your location, please consult with your recruiter.AccommodationsRoku welcomes applicants of all backgrounds and provides reasonable accommodations and adjustments in accordance with applicable law. If you require reasonable accommodation at any point in the hiring process, please direct your inquiries to View email address on click.appcast.io should I know about Roku's culture?Roku is a great place for people who want to work in a fast-paced environment where everyone is focused on the company's success rather than their own. We try to surround ourselves with people who are great at their jobs, who are easy to work with, and who keep their egos in check. We appreciate a sense of humor. We believe a fewer number of very talented folks can do more for less cost than a larger number of less talented teams. We're independent thinkers with big ideas who act boldly, move fast and accomplish extraordinary things through collaboration and trust. In short, at Roku you'll be part of a company that's changing how the world watches TV. We have a unique culture that we are proud of. We think of ourselves primarily as problem-solvers, which itself is a two-part idea. We come up with the solution, but the solution isn't real until it is built and delivered to the customer. That penchant for action gives us a pragmatic approach to innovation, one that has served us well since 2002. To learn more about Roku, our global footprint, and how we've grown, visit .By providing your information, you acknowledge that you want Roku to contact you about job roles, that you have read Roku's Applicant Privacy Notice, and understand that Roku will use your information as described in that notice. If you do not wish to receive any communications from Roku regarding this role or similar roles in the future, you may unsubscribe at any time by emailing View email address on click.appcast.io.

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Senior Machine Learning Engineer, DevOps/SRE in San Jose, CA vacancy
  • $98.9k - $228.7k

    What you can expectWe are hiring a Senior DevOps Engineer to ensure reliability, scalability, and operational excellence for our real-time communications...  ...rollout strategies across teams.Acting as the primary SRE partner for multiple engineering teams building real-time... 
    Senior
    Devops
    Full time
    Work at office
    Remote work
    Flexible hours

    Zoom

    San Jose, CA
    3 days ago
  • $203k - $258.6k

     ...techniques including reinforcement learning, and more. This is your...  ...testing, intelligence on edge, and DevOps automation. This role...  ...with product management and engineering teams to deliver impactful, scalable...  ..., validate, and deploy machine learning models that drive measurable... 
    Senior
    Devops
    Full time
    Temporary work
    Local area
    Flexible hours

    CISCO Systems

    San Jose, CA
    4 days ago
  • $224k - $356.5k

    NVIDIA is hiring engineers to scale up the introduction of next generation architecture...  ...with at scale infrastructure, DevOps and/or SRE practices and/or Platform Engineering....  ...communication skillsPassion for continuous learning and knowledge transfer. Ability to work... 
    Senior
    Devops
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $184k - $287.5k

    We are seeking a Senior DevOps / Cloud Simulation Infrastructure Engineer to own the complete end-to-end cloud execution pipeline for SimReady assets! This role...  ...simulation.Extensive experience in production-grade DevOps, SRE, or Infrastructure Engineering, with a focus on GPU-... 
    Senior
    Devops
    Full time
    Local area

    Nvidia

    Santa Clara, CA
    4 days ago
  •  ...Senior Backend Engineer (Infrastructure & AI Platform) FlexAI is looking for a Senior Backend Engineer (Infrastructure & AI Platform) with...  ...cloud-native, Kubernetes-native services Work with DevOps/SRE on CI/CD, deployment automation, and scalability Contribute... 
    Senior
    Devops
    Work at office

    FlexAI

    Santa Clara, CA
    3 days ago
  • $184k - $287.5k

    We are looking for a Senior System Software Engineer, Software Defined Networking to design, build, and operate highly performant and scalable SDN...  ...monitoring, and performance tuningCollaborate with SRE, DevOps, and network engineering teams on production readiness and... 
    Senior
    Devops
    Full time

    Nvidia

    Santa Clara, CA
    16 hours ago
  • $147k - $237.5k

     ...outcomes.Job SummaryWe are seeking a highly experienced ML Engineer with a good understanding of machine learning principles and algorithms to join our team and drive...  ...closely with cross-functional teams (product, QA, DevOps, and customer support) to align development efforts... 
    Devops
    Full time
    Work at office

    Palo Alto Networks

    Santa Clara, CA
    1 day ago
  •  ...Job Title: Mid-Senior Site Reliability Engineer Kubernetes Platform Location: San Jose, CA Full-Time Job Description Must...  ...Technical/Functional Skills: 8+ years of experience in SRE, DevOps, or platform engineering Hands-on experience with... 
    Senior
    Devops
    Full time

    SFE

    San Jose, CA
    2 days ago
  • $185.9k - $300.68k

     ...drives great outcomes.Job SummaryAs the Senior Manager of Product Management for...  ...services. You will work across Product, Engineering, SRE, and Support to enable seamless global...  ...systems, reliability engineering, and modern DevOps practices.Experience partnering with... 
    Senior
    Devops
    Full time
    Work at office

    Palo Alto Networks

    Santa Clara, CA
    16 hours ago
  • $148k - $235.75k

     ...on the world.Join our team of innovative engineers who are building an AI Data Center AIOps...  ...automation for GPU fleets. We’re hiring a DevOps Engineer to operate the platform itself (...  ...operating production distributed systems as SRE/DevOps/Platform Ops.Proven ownership of... 
    Senior
    Devops
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  •  ..., simplify, and accelerate revenue.We are looking for a Senior Site Reliability Engineer to lead the strategic evolution of our cloud infrastructure...  ...You AreExperienced Architect: 5+ years of experience in SRE, DevOps, or Systems Engineering, with a proven track record of... 
    Senior
    Devops
    Full time
    Work at office
    2 days per week

    LeanData

    Santa Clara, CA
    17 hours ago
  • $272k - $431.25k

     .... You will work closely with engineering and AI teams, help shape how...  ...collaborative, inclusive, and learning focused environment.Designing...  ...Partnering with engineering, DevOps, and AI/ML teams to improve data...  ...managing and scaling SRE/Production Engineering teams... 
    Senior
    Devops
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $174.72k - $295.68k

     ...the future of transportation through cutting-edge R&D in AI, machine learning, and smart connectivity.Our mission is to build strong...  ...or compilation stack.Strong Python programming and software engineering skills.Ability to work effectively across research, systems,... 
    Senior
    Full time

    XPENG Motors

    Santa Clara, CA
    2 days ago
  • $174.72k - $295.68k

     ...is dedicated to reshaping the future of transportation through cutting-edge R&D in AI, machine learning, and smart connectivity.We are looking for a full-time Machine Learning Engineer - AI Foundation, with deep knowledge and strong enthusiasm towards establishing a... 
    Senior
    Full time

    XPENG Motors

    Santa Clara, CA
    2 days ago
  • $224k - $356.5k

    NVIDIA is looking for a Machine Learning Engineer to join the GPU accelerated Apache Spark team.Apache Spark is the most popular data processing engine in data centers for running massive scale workloads for ETL, SQL, and ML/DL model training and inference pipelines, spanning... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $248k - $391k

     ...leader to build and lead a high-performance engineering organization that architects, delivers,...  ...for success.What You'll Be Doing:As a Senior Engineering Manager, you will own the...  ...in at least one of: Infrastructure, SRE, DevOps, or Production Engineering. 5+ years leading... 
    Senior
    Devops
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $184k - $287.5k

    Intelligent machines powered by Artificial Intelligence computers that can learn, reason and interact with people are no longer science fiction. GPU Deep Learning...  ....We are now looking for an extraordinary Senior Perception Engineer to develop and productize NVIDIA’s... 
    Senior
    Odd job
    Full time
    Work experience placement
    Remote work

    Nvidia

    Santa Clara, CA
    1 day ago
  • $172.5k - $306.63k

     ...of Adobe’s creative ecosystem. Our mission is to employ machine learning to enhance our comprehension of the creative content...  ...artificial intelligence to new heights.What You'll DoAs a Senior Machine Learning Engineer on the Content Intelligence team, you will lead the... 
    Senior
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Jose, CA
    3 days ago
  • $153.75k - $225k

     ...that values ownership, collaboration, and high standards. Our engineers, product leaders, and go-to-market teams work closely...  ...command execution.Qualifications:Knowledge and passion in machine learning algorithms, GenAI, LLMs, and Agentic AIUnderstanding of agent... 
    Senior
    Work experience placement
    Work at office
    3 days per week

    Eightfold

    Santa Clara, CA
    2 days ago
  • $150k - $250k

     ...future of autonomy, Plus is looking for talented individuals to join its fast-growing teams.We are seeking a highly skilled Machine Learning Engineer with deep expertise in developing Bird’s Eye View (BEV) fusion models using multimodal sensor inputs, particularly LiDAR.... 
    Senior
    Full time

    Plus.ai

    Santa Clara, CA
    3 days ago
  • $130k - $200k

    Santa Clara, CAUS Engineering - Simulation /Full-time /HybridPlusAI is a Physical AI company pioneering AI-based virtual driver...  ...individuals to join its fast-growing teams.We are seeking a Senior Machine Learning Engineer with expertise in deep learning and data analysis... 
    Senior
    Full time

    Plus.ai

    Santa Clara, CA
    2 days ago
  • $229.5k - $360k

     ...experiences across our platform by leveraging state-of-the-art machine learning. Our mission is to deliver meaningful, context-aware...  ...at the forefront of technology. Our work blends innovation, engineering excellence, and a deep commitment to understanding our users... 
    Senior
    Work at office
    Local area
    Remote work
    Monday to Thursday
    Flexible hours

    Roku

    San Jose, CA
    1 day ago
  •  ...Job Title: Senior Site Reliability Engineer Kubernetes Platform Location: San Jose, CA Full-Time Job Description Must Have...  ...Technical/Functional Skills: 10+ years of experience in SRE, DevOps, or infrastructure engineering Strong experience... 
    Senior
    Devops
    Full time

    SFE

    San Jose, CA
    2 days ago
  • $151.8k - $265.35k

     ...Media & Entertainment, marketing, and consumer retail, and is expanding rapidly into adjacent verticals. We are hiring a Senior Machine Learning Engineer to build the pipelines and services that turn Firefly Foundry’s models into reliable, enterprise-grade products. You... 
    Senior
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Jose, CA
    1 day ago
  • $203k - $258.6k

     ...TeamThe Cisco AI Research team brings together AI researchers, machine learning engineers, data engineers, and networking domain experts to build the...  ...scalable systems and real-world impact.Your ImpactAs a Senior Machine Learning Engineer, you will build and improve the data... 
    Senior
    Full time
    Temporary work
    Work at office
    Local area
    Flexible hours
    Shift work

    CISCO Systems

    San Jose, CA
    4 days ago
  • $152k - $230k

     ...release models and AI systems that integrate with existing machine learning, design automation, and visualization tools within the organization...  ...QOR.What we need to see:MS/PhD in Electrical/Computer Engineering, Computer Science, Applied Mathematics, or equivalent experience... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $130k - $220k

    Santa Clara, CASoftware Engineering - Motion Planning /Full-time /HybridPlusAI is a Physical...  ...Responsibilities Design, develop, and deploy learned behavior planning models using...  ...Computer/Software Engineering, Robotics, Machine Learning or related field.Experience developing... 
    Senior
    Full time

    Plus.ai

    Santa Clara, CA
    3 days ago
  • $244.14k - $413.16k

     ...future of transportation through cutting-edge R&D in AI, machine learning, and smart connectivity.The Mission: We are building...  ...infinite long-tail scenarios of global driving. As a Senior Staff Machine Learning Engineer, you will architect the transition from behavior... 
    Senior
    Full time
    Overseas

    XPENG Motors

    Santa Clara, CA
    2 days ago
  • $220k - $300k

     ...co-design, so decisions about a model shape decisions about the silicon it runs on. About the role As a Senior Principal Machine Learning Engineer, you will be responsible for designing, developing, and optimizing machine learning models-with a focus on cutting... 
    Senior
    Full time
    Temporary work
    Local area
    Worldwide
    Flexible hours

    SambaNova

    San Jose, CA
    2 days ago
  • $130k - $220k

    Santa Clara, CASoftware Engineering - Motion Planning /Full-time /HybridPlusAI is a Physical...  ...real roads, today.About the RoleAs a Senior Software Engineer on the Planning team,...  ...driving conditions.Bring novel robotics and machine learning techniques from research into a... 
    Senior
    Full time

    Plus.ai

    Santa Clara, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Machine Learning Engineer, DevOps/SRE. Be the first to apply!