Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

ML Scheduler Engineer: Optimizing GPU/CPU Orchestration

ByteDance

ByteDance's Volcano Ark team is seeking a software engineer to design and develop resource scheduling systems for machine learning workloads across data centers and clusters. The role involves optimizing orchestration of GPUs, CPUs, storage, and networking to support offline training and online inference. The candidate should have a CS degree, strong programming skills (Go/Java/Python), experience with ML frameworks (TensorFlow/PyTorch), and solid knowledge of Kubernetes, Docker, and distributed #J-18808-Ljbffr ByteDance

Vacancy posted 13 hours ago
Similar jobs that could be interesting for youBased on the ML Scheduler Engineer: Optimizing GPU/CPU Orchestration in San Jose, CA vacancy
  • $159.05k - $199.3k

     ...About the role We are looking for a software engineer with deep experience in optimizing ML models and deploying them on production‑grade embedded runtime...  ...field 3+ years of experience with ML accelerators, GPU, CPU, SoC architecture and micro‑architecture Strong software... 
    Suggested
    Full time
    For contractors
    For subcontractor

    Decisive Point

    Sunnyvale, CA
    2 days ago
  • $150.4k - $277.6k

     ...Services Apple’s Compute Frameworks team in GPU, Graphics and Displays org provides a...  ...extraordinary machine learning and GPU programming engineers who are passionate about providing robust...  ...architectures. Responsibilities Adding optimized GPU compute kernels across Machine... 
    Suggested
    Relocation

    Apple Inc.

    Cupertino, CA
    23 hours ago
  • Apple Inc. in Cupertino, CA, is seeking extraordinary machine learning and GPU programming engineers to deliver robust compute solutions on Apple Silicon. You will contributes to optimized kernels, ML workflows, and high-performance GPU architectures across iOS, macOS and... 
    Suggested

    Apple Inc.

    Cupertino, CA
    23 hours ago
  • $162k - $316.8k

    Responsibilities The Data-AML-Engine Orchestration team builds large-...  ...the orchestration, scheduling, and resource...  ...infrastructure with production ML workloads. The Data-...  ...that directly affect GPU utilization, serving...  ..., performance optimization, or highly available... 
    Suggested
    Temporary work
    Internship
    Local area

    ByteDance

    San Jose, CA
    3 days ago
  • TikTok Search Ads is seeking talented engineers to push the boundaries of monetization across...  ..., ranking, and IR related challenges to optimize ad delivery and performance. You will collaborate...  ...product strategy, implement advanced ML models, and enhance search relevance,... 
    Suggested

    Tik Tok

    San Jose, CA
    1 day ago
  • $165.2k - $223.6k

     ...a Software Development Engineer to own the design and implementation...  ...who can write and optimize low-level code for...  ...kernels for a custom ML accelerator...  ...vLLM, PyTorch), including scheduler extensions, memory management...  ...correctness validation across CPU, GPU, simulator, and... 
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    23 hours ago
  • ByteDance is building a Data-AML-Engine Orchestration platform to power online...  ...to orchestration, scheduling, and resource management systems...  ...heterogeneous compute with production ML workloads. Join a team...  ...availability, and efficient GPU usage while tackling challenging... 

    ByteDance

    San Jose, CA
    3 days ago
  • Apple Inc. is seeking a senior engineer to design and optimize the core auction system powering our advertising platform. You will develop production-grade code to generate high quality ad recommendations, collaborating with product and business teams to drive new products... 

    Apple Inc.

    Cupertino, CA
    1 day ago
  •  ...in California is seeking a Research Software Engineer to work at the intersection of computer vision and machine learning. The role involves optimizing network performance and developing technologies for next-generation GPU platforms. Candidates should have a background... 

    Google

    Mountain View, CA
    1 day ago
  • $174.72k - $295.68k

     ...performance trade-offs.Develop PTQ and QAT orchestration workflows.Serve as the primary...  ...Strong Python programming and software engineering skills.Ability to work effectively across...  ...hardware.Contributions to model optimization, inference, compiler, or serving projects... 
    Full time

    XPENG Motors

    Santa Clara, CA
    1 day ago
  • $174.72k - $295.68k

     ...a full-time Machine Learning Engineer - AI Foundation, with deep knowledge...  ...establishing a state-of-art ML infrastructure for training...  ...driving.Job Responsibilities:Optimize transformer-based LLMs for low...  ...Python and C++Familiarity with GPU CPU, NPU, DSP architecture.Deep understanding... 
    Full time

    XPENG Motors

    Santa Clara, CA
    1 day ago
  • $184k - $287.5k

     ...next era of computing. An era where our GPU serves as the intelligence behind computers...  ...systems at scale. We seek a Senior ML Engineer to compose and deliver next-generation Metropolis...  ..., quantization, and real-time inference optimization for production deployments.Previous... 
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $165.2k - $223.6k

     ...possible. AWS Neuron is the SDK that optimizes the performance of complex ML models executed on AWS Inferentia...  ...workloads.This role is for a software engineer in the Compiler team for AWS Neuron...  ...Experience in compiler design for CPU/GPU/Vector engines/ML-accelerators.-... 
    Internship
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    1 day ago
  • Apple seeks a senior engineer to design the core auction system powering its ads platform at global scale. You will drive optimal outcomes for advertisers, users, and the platform, applying advanced auction theory and machine learning in production environments. You will... 

    Socket

    Cupertino, CA
    3 days ago
  • $138k - $206k

     ...and communities.Senior Engineer, Architecture & Performance...  ...next-generation CPU (RISC-V) microarchitecture...  ...performance bottlenecks and optimization opportunities across...  ...x86 CPU cores, or with GPU/NPU vector unit...  ...support. All candidates scheduled for an interview will receive... 
    Work at office
    Flexible hours
    Shift work

    Samsung Semiconductor

    San Jose, CA
    1 day ago
  • $153.2k - $234.1k

     ...world scenarios. As a Senior ML Infra Engineer, you will work on the core...  ...ML training across large GPU/CPU clusters or specialized accelerators...  ...profiling and training optimization techniques and their impact...  ...with containerization and orchestration technologies (e.g., Docker,... 
    Full time
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    4 days ago
  • $213k - $263k

     ...across 15+ U.S. states. The ML Platform team at Waymo...  ...management, model development, optimization and monitoring. These efforts...  ...Simulation. We are looking for engineers with ML software or ML systems...  ...hardware accelerators (e.g., GPU/TPU). Deep knowledge of model... 
    Full time
    Remote work

    Waymo

    Mountain View, CA
    4 days ago
  •  ...are hiring AI / ML Platform Engineers to build the platform...  ...management, and GPU cluster...  ...framework across kernel optimization, RTL/PPA...  ..., reduce manual orchestration, and move successful...  ...job submission, scheduling, orchestration,...  ..., firmware, and CPU/GPU performance... 

    AMD

    Santa Clara, CA
    2 days ago
  • $100k

     ...developed a high performance RISC-V CPU from scratch, and share a...  ...is seeking an Physical Design Engineer to lead cross-functional...  ...architect, integrate, and deploy AI/ML-driven solutions into...  ...efficiency, turnaround time, and QoR.Optimize EDA tools and custom CAD flows... 
    Permanent employment

    Tenstorrent

    Santa Clara, CA
    3 days ago
  • $184k - $287.5k

     ...modeling and constrained optimization, for real-time routing and scheduling at scale.This platform encompasses...  .... Our data include GPU availability/lifecycle,...  ...and deploy scalable ML/AI and optimization models...  ...develop user-specific feature engineering for real-time cloud... 
    Full time

    Nvidia

    Santa Clara, CA
    18 hours ago
  •  ...Clara is seeking a skilled software engineer with over 8 years of experience to optimize machine learning models for real-...  ...training strategies, collaborating with ML researchers, and developing tools...  ...with PyTorch and NVIDIA GPU ecosystems. #J-18808-Ljbffr Odyssey

    Odyssey

    Santa Clara, CA
    4 days ago
  • Responsibilities The Data-AML-Engine Orchestration team builds large-scale...  ...develop the orchestration, scheduling, and resource management systems...  ...infrastructure with production ML workloads. Responsibilities...  ...that directly affect GPU utilization, serving latency... 
    Internship

    ByteDance

    San Jose, CA
    3 days ago
  • Applied Intuition, Inc. is a Silicon Valley leader powering the future of physical AI. We seek a Performance Engineer to optimize large-scale ML workloads, focusing on distributed training, batch inference, and cost-effective data processing. You will own profiling across... 

    Decisive Point

    Sunnyvale, CA
    3 days ago
  • $147.4k - $272.1k

     ...front and center? The Video Engineering group at Apple is responsible...  ...across different platforms while optimizing models for memory efficiency,...  ...workloads across CPU, GPU, Apple Neural Engine, and other...  ...programming skills for common ML frameworks like PyTorch or TensorFlow... 
    Relocation

    Apple Inc.

    Cupertino, CA
    23 hours ago
  • ByteDance seeks a PhD-level engineer to design and build scalable orchestration for ML platforms, including Kubernetes Operators and lifecycle management for jobs and services. You will improve GPU utilization, implement multi-cluster serving, and advance autoscaling and... 

    ByteDance

    San Jose, CA
    4 days ago
  • Rhoda AI in Mountain View is seeking a Research Engineer to build and maintain a training platform that powers model...  ...will develop tooling for experiment management and optimize training processes across large-scale GPU clusters. This role offers high visibility, direct... 

    Rhoda AI

    Mountain View, CA
    2 days ago
  • $184k - $287.5k

     ...computing. An era in which our GPU acts as the brains of...  ...CPUs, and a fully optimized NVIDIA AI and HPC...  ...for a highly motivated engineer to lead performance benchmarking...  ...:Experience with AI/ML frameworks (PyTorch,...  ...cloud provisioning and scheduling tools (Docker, Kubernetes... 
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  •  ...THE ROLEAMD is seeking an experienced AI/ML Engineer to join our Applied Research and...  ...technologies. In this role, you will develop and optimize state-of-the-art AI models for image...  ...AI solutions that leverage AMD’s latest GPU technologies and accelerate the future of... 

    AMD

    San Jose, CA
    1 day ago
  • $136.5k - $276.5k

    AI/ML Engineer - AgenticThis role has been designed as ‘Hybrid’ with...  ...a production-grade agentic orchestration platform, including multi-agent...  ...search, and relevance optimization.Develop and operate high-performance...  ...and resource optimization (CPU/memory, autoscaling).... 
    Full time
    Work experience placement
    Work at office
    Local area
    Immediate start
    2 days per week

    Hewlett Packard Enterprise

    San Jose, CA
    4 days ago
  • $150k - $350k

     ...cutting‑edge generative AI to assist engineers in RTL design, simulation, and verification...  ...Position Overview We are seeking an ML Systems Engineer to optimize the performance and efficiency of...  ...at the systems level—from GPU kernel execution to memory bandwidth... 

    ChipAgents

    San Jose, CA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to ML Scheduler Engineer: Optimizing GPU/CPU Orchestration. Be the first to apply!