Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Machine Learning Engineer, DevOps/SRE

$148.75k - $361k

Roku, Building C

Teamwork makes the stream work.Roku is changing how the world watches TVRoku is the #1 TV streaming platform in the U.S., Canada, and Mexico, and we've set our sights on powering every television in the world. Roku pioneered streaming to the TV. Our mission is to be the TV streaming platform that connects the entire TV ecosystem. We connect consumers to the content they love, enable content publishers to build and monetize large audiences, and provide advertisers unique capabilities to engage consumers.From your first day at Roku, you'll make a valuable - and valued - contribution. We're a fast-growing public company where no one is a bystander. We offer you the opportunity to delight millions of TV streamers around the world while gaining meaningful experience across a variety of disciplines.About the team The Advertising Performance group focuses on performance for all participants in the Advertising ecosystem - Advertisers, Publishers, and Roku. The systems and solutions span multiple disciplines and technologies to perform real-time multi-objective optimization across distributed systems at large scale and with low latency. We use Machine Learning, Reinforcement Learning, AI, Control and Optimization Systems, and Auction Dynamics to solve a large set of complex problems. At the core of this is our Machine Learning, Experimentation, and Inference Platform that powers the entire landscape, which we continuously evolve over time. About the role We are seeking a talented and experienced Senior Software Engineer, MLOps/DevOps, to join the Advertising Performance team and play a critical role in supporting and scaling our Machine Learning infrastructure. The ideal candidate has a strong background in DevOps/SRE practices, cloud infrastructure management, and MLOps tooling — with a passion for building platforms that accelerate ML experimentation and deployment at internet scale. You will partner closely with ML Scientists and Engineers to streamline the end-to-end ML lifecycle across training, evaluation, deployment, and monitoring — on top of a modern, cloud-native stack running on GCP and AWS using Kubernetes, Apache Airflow, Spark, Ray, MLflow, Chronon, etc.For California Only - The estimated annual salary for this position is between $148,750 - $361,000 annually. Compensation packages are based on factors unique to each candidate, including but not limited to skill set, certifications, and specific geographical location. This role is eligible for health insurance, equity awards, life insurance, disability benefits, parental leave, wellness benefits, and paid time off. What you’ll be doing Lead the design and operation of scalable, production-grade cloud infrastructure for ML workloads across AWS and GCP, including GPU/TPU-based training and inference environmentsArchitect and improve CI/CD systems for ML models and platform services to enable fast, reliable, and safe production releasesOwn and evolve low-latency infrastructure for real-time model inference, including KV store and vector databasesDefine and enforce observability standards for ML systems, including model performance monitoring, drift detection, capacity planning, and pipeline health metricsParticipate in on-call rotation, leading incident response and root-cause analysis for critical ML training and serving infrastructurePartner with data scientists and ML engineers to improve platform usability, accelerate model iteration, and implement strong MLOps and SRE best practicesChampion operational excellence across ML infrastructure through automation, resilience engineering, disaster recovery planning, and continuous improvementWe’re excited if you have BS or MS in Computer Science, Engineering, or a related quantitative fieldAbility to demonstrate AI tool (Claude, Cursor, ChatGPT etc.) coding efficiency at an advanced skill level8+ years of experience in DevOps, SRE, or ML infrastructure, including 4+ years supporting large-scale ML or AI systemsStrong programming skills in Python, and/or Scala, or Java for platform automation and toolingDeep experience with Kubernetes and container orchestration on GCP (GKE) and/or AWS (EKS)Expertise with NoSQL or low-latency data stores such as Aerospike or similar technologiesHands-on experience with data and orchestration technologies such as Apache Spark, Apache Flink, Apache Airflow, and KafkaExperience building and maintaining CI/CD systems using tools such as Jenkins or GitLab RunnerFamiliarity with feature engineering platforms such as Chronon and model lifecycle tools such as MLflowStrong infrastructure-as-code experience with Terraform or similar toolingExperience with observability platforms such as Prometheus, Grafana, and DatadogExcellent communication and cross-functional collaboration skillsExperience in the Advertising domain is a plus #LI-DH2What's Roku's approach to hybrid working?Roku fosters an inclusive and collaborative environment where teams generally work in the office Monday through Thursday. Fridays are generally flexible for remote work, except for employees whose specific roles or assigned office location require five days' a week attendance.What are some of the benefits?Roku is committed to offering a diverse range of benefits as part of our compensation package to support our employees and their families. Our comprehensive benefits include global access to mental health and financial wellness support and resources. Local benefits include statutory and voluntary benefits which may include healthcare (medical, dental, and vision), life, accident, disability, commuter, and retirement options (401(k)/pension). Employees are supported in taking time off, in accordance with local leave policies and other personal needs to support their evolving work and life needs. It's important to note that not every benefit is available in all locations or for every role. For details specific to your location, please consult with your recruiter.AccommodationsRoku welcomes applicants of all backgrounds and provides reasonable accommodations and adjustments in accordance with applicable law. If you require reasonable accommodation at any point in the hiring process, please direct your inquiries to View email address on click.appcast.io should I know about Roku's culture?Roku is a great place for people who want to work in a fast-paced environment where everyone is focused on the company's success rather than their own. We try to surround ourselves with people who are great at their jobs, who are easy to work with, and who keep their egos in check. We appreciate a sense of humor. We believe a fewer number of very talented folks can do more for less cost than a larger number of less talented teams. We're independent thinkers with big ideas who act boldly, move fast and accomplish extraordinary things through collaboration and trust. In short, at Roku you'll be part of a company that's changing how the world watches TV. We have a unique culture that we are proud of. We think of ourselves primarily as problem-solvers, which itself is a two-part idea. We come up with the solution, but the solution isn't real until it is built and delivered to the customer. That penchant for action gives us a pragmatic approach to innovation, one that has served us well since 2002. To learn more about Roku, our global footprint, and how we've grown, visit .By providing your information, you acknowledge that you want Roku to contact you about job roles, that you have read Roku's Applicant Privacy Notice, and understand that Roku will use your information as described in that notice. If you do not wish to receive any communications from Roku regarding this role or similar roles in the future, you may unsubscribe at any time by emailing View email address on click.appcast.io.

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Senior Machine Learning Engineer, DevOps/SRE in San Jose, CA vacancy
  •  ...cybersecurity firm based in Santa Clara is seeking a Sr Site Reliability Engineer. The candidate will be responsible for maintaining highly...  ...candidates should have at least 5 years of experience in SRE or DevOps and a solid background in cloud services, automation, and... 
    Senior
    Devops

    Palo Alto Networks

    Santa Clara, CA
    3 days ago
  • $224k - $356.5k

     ...three unnecessary? We're looking for a senior engineer to invent, set, and construct the...  ...2+ years of infrastructure, platform, DevOps, or SRE engineeringDeep expertise in modern CI...  ...system(s) and all the lessons you've learned from supporting themOutstanding communication... 
    Senior
    Devops
    Full time
    Immediate start

    Nvidia

    Santa Clara, CA
    3 days ago
  • $184k - $287.5k

    We are seeking a Senior DevOps / Cloud Simulation Infrastructure Engineer to own the complete end-to-end cloud execution pipeline for SimReady assets! This role...  ...simulation.Extensive experience in production-grade DevOps, SRE, or Infrastructure Engineering, with a focus on GPU-... 
    Senior
    Devops
    Full time
    Local area

    Nvidia

    Santa Clara, CA
    1 day ago
  •  ..., simplify, and accelerate revenue.We are looking for a Senior Site Reliability Engineer to lead the strategic evolution of our cloud infrastructure...  ...You AreExperienced Architect: 5+ years of experience in SRE, DevOps, or Systems Engineering, with a proven track record of... 
    Senior
    Devops
    Full time
    Work at office
    2 days per week

    LeanData

    Santa Clara, CA
    4 days ago
  •  ...ensure we take care of ourselves, each other, and our communities. Job Summary: Job Description: PayPal, Inc. seeks Senior Staff Machine Learning Engineer in San Jose, CA Job Duties: Define and drive the strategic vision for implementing machine learning (ML)... 
    Senior
    Full time
    Work at office
    Local area
    Immediate start
    Remote work
    Flexible hours

    PayPal

    San Jose, CA
    2 days ago
  • $175.8k - $312.2k

    SummaryWe are looking for a Machine Learning Engineer who will be converting abstract, high-level goals into concrete, measurable requirements. They will be proposing, implementing, evaluating, and shipping different AI/ML technologies and resulting data to achieve a given... 
    Senior
    Relocation

    Apple

    Cupertino, CA
    4 days ago
  • $185.9k - $300.68k

     ...drives great outcomes.Job SummaryAs the Senior Manager of Product Management for...  ...services. You will work across Product, Engineering, SRE, and Support to enable seamless global...  ...systems, reliability engineering, and modern DevOps practices.Experience partnering with... 
    Senior
    Devops
    Full time
    Work at office

    Palo Alto Networks

    Santa Clara, CA
    2 days ago
  • $174.72k - $295.68k

     ...is dedicated to reshaping the future of transportation through cutting-edge R&D in AI, machine learning, and smart connectivity.We are looking for a full-time Machine Learning Engineer - AI Foundation, with deep knowledge and strong enthusiasm towards establishing a... 
    Senior
    Full time

    XPENG Motors

    Santa Clara, CA
    4 days ago
  • $148k - $235.75k

     ...on the world.Join our team of innovative engineers who are building an AI Data Center AIOps...  ...automation for GPU fleets. We’re hiring a DevOps Engineer to operate the platform itself (...  ...operating production distributed systems as SRE/DevOps/Platform Ops.Proven ownership of... 
    Senior
    Devops
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $184k - $287.5k

    We are seeking a Senior Machine Learning Engineer to join our end‑to‑end autonomous driving team! You will help build, train, and deploy large‑scale E2E driving models that leverage VLM/VLA architectures, and build a data flywheel that continuously improves our systems... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $184k - $287.5k

    Intelligent machines powered by Artificial Intelligence computers that can learn, reason, and interact with people are no longer science fiction. Today, a self-driving...  ...AV. We are seeking the best Machine Learning Engineers with a background in computer vision, LiDAR &... 
    Senior
    Full time
    Worldwide
    Night shift

    Nvidia

    Santa Clara, CA
    1 day ago
  • $130k - $200k

    Santa Clara, CAUS Engineering - Simulation /Full-time /HybridPlusAI is a Physical AI company pioneering AI-based virtual driver...  ...individuals to join its fast-growing teams.We are seeking a Senior Machine Learning Engineer with expertise in deep learning and data analysis... 
    Senior
    Full time

    Plus.ai

    Santa Clara, CA
    4 days ago
  • $172.5k - $306.63k

     ...of Adobe’s creative ecosystem. Our mission is to employ machine learning to enhance our comprehension of the creative content...  ...artificial intelligence to new heights.What You'll DoAs a Senior Machine Learning Engineer on the Content Intelligence team, you will lead the... 
    Senior
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Jose, CA
    4 days ago
  • $153.75k - $225k

     ...that values ownership, collaboration, and high standards. Our engineers, product leaders, and go-to-market teams work closely...  ...command execution.Qualifications:Knowledge and passion in machine learning algorithms, GenAI, LLMs, and Agentic AIUnderstanding of agent... 
    Senior
    Work experience placement
    Work at office
    3 days per week

    Eightfold

    Santa Clara, CA
    4 days ago
  • $224k - $356.5k

    We are looking for outstanding Machine Learning Engineers to join our Physical AI teams. As the pioneers of the GPU—the visual cortex of modern computing—we are building the foundation for the next wave of AI that interacts with the physical world.This role is at the forefront... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $361.3k

     ...techniques, including LLMs, reinforcement learning, multi-objective optimization, and...  ...love.About the RoleWe are looking for a Senior Machine Learning Tech Lead to drive ranking and...  ...have built fluency across the agentic engineering toolchain — coding harnesses like Claude... 
    Senior
    Work at office
    Local area
    Remote work
    Monday to Thursday
    Flexible hours

    Roku

    San Jose, CA
    1 day ago
  • $272k - $431.25k

     .... You will work closely with engineering and AI teams, help shape how...  ...collaborative, inclusive, and learning focused environment.Designing...  ...Partnering with engineering, DevOps, and AI/ML teams to improve data...  ...managing and scaling SRE/Production Engineering teams... 
    Senior
    Devops
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $151.8k - $265.35k

     ...Media & Entertainment, marketing, and consumer retail, and is expanding rapidly into adjacent verticals. We are hiring a Senior Machine Learning Engineer to build the pipelines and services that turn Firefly Foundry’s models into reliable, enterprise-grade products. You... 
    Senior
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Jose, CA
    3 days ago
  •  ...Collaborate with product managers, UX designers, and other engineers to define requirements and deliver impactful solutions. Diagnose...  ...chat-like command execution. Knowledge and passion in machine learning algorithms, GenAI, LLMs, and Agentic AI Understanding of agent... 
    Senior
    Full time
    Work experience placement

    Eightfold

    Santa Clara, CA
    13 hours ago
  • $152k - $230k

     ...release models and AI systems that integrate with existing machine learning, design automation, and visualization tools within the organization...  ...QOR.What we need to see:MS/PhD in Electrical/Computer Engineering, Computer Science, Applied Mathematics, or equivalent experience... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $244.14k - $413.16k

     ...future of transportation through cutting-edge R&D in AI, machine learning, and smart connectivity.The Mission: We are building...  ...infinite long-tail scenarios of global driving. As a Senior Staff Machine Learning Engineer, you will architect the transition from behavior... 
    Senior
    Full time
    Overseas

    XPENG Motors

    Santa Clara, CA
    4 days ago
  • $174.72k - $295.68k

     ...is dedicated to reshaping the future of transportation through cutting-edge R&D in AI, machine learning, and smart connectivity.We are looking for a full-time Machine Learning Engineer / Research Scientist to drive the modeling and algorithmic development of XPENG’s next... 
    Senior
    Full time

    XPENG Motors

    Santa Clara, CA
    4 days ago
  • $130k - $220k

     ...drive the future of autonomy, Plus is looking for talented individuals to join its fast-growing teams.We’re looking for a machine learning engineer to train and deploy the latest generation of ML-based planning algorithms on the extensive data we collect every day across... 
    Senior
    Full time

    Plus.ai

    Santa Clara, CA
    4 days ago
  • $160k - $200k

    Santa Clara, CAData Engineering - Machine Learning and Data Engineer /Full-time /HybridPlusAI is a Physical AI company pioneering AI-based virtual...  ...talented individuals to join its fast-growing teams.As a Senior ML Infrastructure Engineer at Plus, you will design scalable... 
    Senior
    Full time

    Plus.ai

    Santa Clara, CA
    4 days ago
  • $174.72k - $295.68k

     ...is dedicated to reshaping the future of transportation through cutting-edge R&D in AI, machine learning, and smart connectivity.We are seeking a Machine Learning Data Curation Engineer to spearhead the data pipeline development and dataset management for our core AI... 
    Senior
    Full time

    XPENG Motors

    Santa Clara, CA
    1 day ago
  • $280k - $380k

     ...disciplines.About the TeamOur DevOps/SRE team runs an active-active,...  ...reliability and automation, engineering systems that perform under...  ...Site Reliability Engineering) Senior Software Engineer to join...  ...great balance of skills in learning, organizing, building, and enjoy... 
    Senior
    Devops
    Work at office
    Local area
    Remote work
    Monday to Thursday
    Flexible hours

    Roku

    San Jose, CA
    4 days ago
  • $110k - $145k

     ...Position Overview We are looking for a talented and experienced Deep Learning Engineer specializing in Large Language Models (LLMs) to join our dynamic team. In this role, you will play a pivotal part in enhancing the reliability, safety, and performance of AI models and... 
    Senior

    A10 Networks

    San Jose, CA
    3 days ago
  • $224k - $356.5k

    We are seeking exceptional Senior Machine Learning and Simulation Engineers to join NVIDIA's Autonomous Vehicles (AV) Simulation team! This role requires strong technical leadership and outstanding software engineering skills, coupled with deep expertise in both simulation... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $160.2k - $322.3k

     ...business results. This rich data environment is ripe for innovative machine learning applications, offering an unparalleled opportunity to positively impact our customer’s business. As the Sr Engineering Manager of Applied Machine Learning, you will be in charge of... 
    Senior
    Temporary work

    Adobe

    San Jose, CA
    3 days ago
  • $190k

     ...community is the kind of work that matters to you, you are in the right place. About the Role :We're seeking an experienced Staff Machine Learning Engineer to join our Engineering team. In this role, you'll drive the design and implementation of production machine learning... 
    Senior
    Local area
    Immediate start
    Remote work

    GrabJobs

    San Jose, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Machine Learning Engineer, DevOps/SRE. Be the first to apply!