Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

ML Systems Engineer, Physical AI

Orbifold AI

Scale our Ray + PyTorch infrastructure for the multimodal video, training, and RL pipelines that power frontier robotics and world model teams. 3+ yrs distributed systems / ML infra. About Orbifold AI Orbifold AI is building the foundational infrastructure that the next generation of physical AI runs on . We work directly with leading robotics and world model research teams. Our work spans evaluation, model training, reinforcement learning, and the multimodal data systems that fuel them — one integrated research loop. The bottleneck for physical AI is no longer model scale or computation. It is whether evaluation, training, and data can close the loop tightly enough to drive real progress. That loop is itself the infrastructure the next generation of physical AI will stand on, and it is what we are building. Role Overview We are hiring a Machine Learning Engineer to scale and optimize the ML infrastructure behind our pipelines. We process massive volumes of multimodal data — video, image, sensor, action — for some of the most demanding physical AI and world model teams in the field. Our foundation is built on PyTorch and Ray . You will own the systems that turn raw multimodal data into the training, evaluation, and RL signals our partners depend on. Your work is the bridge between our research and our distributed compute infrastructure: making the pipelines performant, fault-tolerant, and ready to scale to the next order of magnitude. This is highly applied infrastructure work with direct impact on what our partner models can do in the real world. What You Will Work On Architect, build, and optimize distributed ML pipelines on Ray (Ray Core, Ray Train, Ray Serve) and PyTorch , designed for the demands of multimodal video, image, and sensor data at scale Profile and tune distributed training jobs and inference deployments to maximize GPU/CPU utilization and reduce latency Build robust abstractions and internal tools that let our researchers and product engineers deploy PyTorch models onto our Ray clusters seamlessly Design and maintain high-throughput video processing pipelines (e.g. FFmpeg, NVDEC/NVENC, frame-level indexing) that feed our curation, training, and evaluation workloads Ensure the high availability, fault tolerance, and observability of our distributed compute systems Build the serving infrastructure for our evaluation harnesses, verification models, and RL environments Collaborate with research, data, and product engineering teams to translate modeling constraints into scalable infrastructure solutions What We Are Looking For 3+ years of software engineering experience with a strong focus on backend, distributed systems, or ML infrastructure Strong proficiency in Python and production-grade code Deep practical knowledge of PyTorch — including model serving, data loading bottlenecks, and memory management Hands-on experience with Ray for scaling Python and machine learning applications Solid understanding of distributed systems concepts: networking, concurrency, fault tolerance, parallel processing Comfortable owning systems end to end in fast-paced applied research or startup environments Nice to Have Experience with large-scale video or multimodal data pipelines (e.g. FFmpeg, NVDEC/NVENC, 3D / point cloud handling) Cloud-native infrastructure experience (Kubernetes, Docker) and major cloud providers (AWS, GCP, Azure) Hardware accelerator experience (GPUs, TPUs) and low-level optimization (CUDA, C++) Background in MLOps and automated CI/CD pipelines for machine learning Familiarity with VLA models, world models, or robotics middleware (e.g. ROS/ROS2) Experience with reinforcement learning environments or simulation infrastructure Why This Role Build the infrastructure that the next generation of physical AI will stand on Work directly with the labs and companies shipping frontier robotics and world model systems Own a critical layer of the stack end to end — from raw video and sensor ingest to distributed training and real-time evaluation serving High ownership, fast iteration, and direct impact on deployed systems #J-18808-Ljbffr Orbifold AI

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the ML Systems Engineer, Physical AI in Palo Alto, CA vacancy
  • $250k - $350k

     ...About Periodic Labs We're an AI and physical sciences company building state-of-the-art models to accelerate breakthroughs across...  ...You'll work alongside some of the world's leading ML systems engineers, including leaders behind Megatron-LM, SGLang, Liger Kernel... 
    Suggested
    Visa sponsorship

    Periodic Labs

    Menlo Park, CA
    2 days ago
  • $300k - $400k

     ...a traditional lab. We're an AI and physical sciences company building state...  ...the Role You will own the systems layer that makes our frontier...  ...and benchmarking distributed ML systems to identify and eliminate...  ...’s best — the scientists, engineers, and problem-solvers who don’... 
    Suggested
    Visa sponsorship
    Flexible hours
    Shift work

    Periodic Labs

    Menlo Park, CA
    4 days ago
  • $90.1k - $191.8k

     ...Understand The World!The Data Labeling Engineering team designs, builds, and...  ..., data engineering, and ML, defining labeling strategies,...  ...leadership, and work directly on systems that unblock the next...  ...cost, and latency goals.Champion AI-assisted engineering Use and advocate... 
    Suggested
    Work experience placement
    Flexible hours

    General Motors

    Mountain View, CA
    1 day ago
  • $100k

     ...the industry on cutting-edge AI technology, revolutionizing...  ...seniorities.Tenstorrent is seeking an Physical Design Engineer to lead cross-functional...  ..., integrate, and deploy AI/ML-driven solutions into...  ...indirect access to information, systems, or technologies subject to... 
    Suggested
    Permanent employment

    Tenstorrent

    Santa Clara, CA
    4 days ago
  • Nebius B.V. is building an AI training and model post-training capability focused...  ...at the intersection of distributed systems, GPU performance, and ML framework integration. The role requires strong Python and PyTorch engineering skills, hands-on experience with distributed... 
    Suggested

    Nebius B.V.

    Palo Alto, CA
    4 days ago
  • Sanas is building real-time speech AI deployed on-premise at scale inside sovereign data...  ...of hardware and software performance engineering. We seek a senior engineer who will shape...  ...they implement high-throughput inference systems. You will own the inference engine, optimize... 

    Sanas

    Palo Alto, CA
    5 days ago
  • General Motors’ Data Labeling Engineering team is building cutting‑edge...  ...that power autonomous vehicle ML models. The role sits at the intersection...  ...focusing on scalable labeling systems and foundations for foundation...  ...labeling features and enable AI‑assisted engineering across... 

    General Motors

    Mountain View, CA
    3 days ago
  • Rhoda AI is hiring a Senior/Staff-level Research Engineer to ensure our robot-learning pipeline is reliable from data collection through model training, inference...  ...real-robot evaluation. You will build validation systems, observability, and robust operating practices to... 

    Socket.dev

    Mountain View, CA
    3 days ago
  •  ...Token Factory is building an AI training and model post-training...  ...intersection of distributed systems, GPU performance, model training...  ...RL pipelines, and production engineering. Your responsibilities Build and...  ...model training, large-scale ML systems, or GPU cluster workloads... 

    Nebius B.V.

    Palo Alto, CA
    4 days ago
  • Rhoda AI in Mountain View is seeking a Staff / Principal ML Training Systems Engineer to lead the performance of large-scale multimodal training systems. This role involves improving training efficiency and collaborating closely with research teams to accelerate model iteration... 

    Rhoda AI

    Mountain View, CA
    5 days ago
  • $169k - $338k

     ...OfficePosition Summary...As a Distinguished AI/ML Engineer within Walmart Global Tech's Site...  ...development of next-generation agentic AI systems and intelligent automation solutions...  ...interaction across Walmart's digital and physical ecosystem.We are the guardians of... 
    Full time
    Temporary work
    Part time

    Walmart

    Sunnyvale, CA
    4 days ago
  • $224k - $356.5k

     ...the boundaries of multimodal AI, simulation, and world models....  ...phase, we are building agentic systems that can reason about, build,...  ...creating the meta-layer of modern ML: the agents, tooling,...  ...We are looking for exceptional engineers who are passionate about the idea... 
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $224k - $356.5k

    We are looking for outstanding Machine Learning Engineers to join our Physical AI teams. As the pioneers of the GPU—the visual cortex of modern computing...  ...mining techniques.Contribute to the full lifecycle of ML software, including performance optimization, testing, and... 
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $174.9k - $261.3k

     ...understand the world!The Data Labeling Engineering team designs, builds, and operates hybrid...  ...software engineering, data engineering, and AI/ML, defining the strategies, tooling, and...  ...leadership, and direct impact on systems that unblock the next generation of AV capabilities... 
    Full time
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    2 days ago
  • Rhoda AI is seeking a Staff / Principal ML Training Systems Engineer in Mountain View to enhance the training systems performance. This role focuses on large-scale multimodal training, driving efficiency, and scalability in compute use across thousands of GPUs. The ideal... 

    Rhoda AI

    Mountain View, CA
    2 days ago
  • $230k - $260k

     ...intersection of creativity and AI with real impact. Join us to...  ...a Principal Machine Learning Engineer, you will operate at the...  ...lead the design of large-scale ML systems and shared platforms that power...  ...resources to support your mental and physical healthDaily lunch &... 
    Work at office
    Immediate start
    3 days per week

    Typeface

    Palo Alto, CA
    2 days ago
  • $190k - $234k

     ...intersection of creativity and AI with real impact. Join...  ...StaffMachine Learning Engineer/Applied Scientist, you...  ...learning models/systems and innovative web applications...  ...building and evolving ML Training and...  ...support your mental and physical healthDaily lunch & snacksMentorship... 
    Work at office
    Local area
    3 days per week

    Typeface

    Palo Alto, CA
    2 days ago
  • $200k - $300k

     ...Machine Learning Systems EngineerLocation - Palo Alto, CA (On-site)...  ...Machine Learning, Generative AI, AI Infrastructure, Developer...  ...technical, and research-driven, with engineers working directly alongside...  ...infrastructure supporting large-scale ML training and inference... 
    H1b
    Work at office

    Recruiting from Scratch

    Palo Alto, CA
    1 day ago
  •  ...ML Systems Engineer Role Overview: We're seeking an experienced engineer to build our ML data infrastructure platform. You'll create the...  ...data processing ~ Cloud Storage for data lakes ~ Vertex AI Feature Store ~ Cloud Composer (managed Airflow) ~ Dataproc... 
    Local area

    My3Tech Inc

    Sunnyvale, CA
    5 days ago
  •  ...View, CA and seeks an experienced software engineer with machine learning expertise to broaden Moveworks NLU and agentic AI capabilities, delivering refined user experiences...  ...products. Responsibilities include applying ML/system engineering to create value, tackling... 

    ServiceNow

    Mountain View, CA
    4 days ago
  • HP, Inc. in Palo Alto is seeking a Senior Software Engineer to build software platforms spanning client apps, backend services, cloud integrations, device interfaces, and enterprise-ready workflows. This role participates in a start-up-like incubation within HP's PC Design... 

    HP Inc.

    Palo Alto, CA
    1 day ago
  • Waymo is seeking a seasoned ML/Computer Vision Engineer to advance the Waymo Driver stack. You will own ML tasks, optimize performance, and...  ...in a production setting. You will analyze and monitor ML systems, build AI-aided debugging tools, and develop metrics for safety-critical... 

    Neura Market

    Mountain View, CA
    2 days ago
  •  ...electricity grid with cutting-edge AI and machine learning. Our...  ...About The Role As an AI/ML Engineer at Powerline, you will be instrumental...  ...MLOps and building AI/ML systems end to end. ~ Strong...  ...systems. Experience with physically informed neural networks, physical... 
    Full time

    Powerline

    Palo Alto, CA
    1 day ago
  •  ...experiment, and interact with the physical world. We are starting out with understanding...  ...and building hardware; electronics systems and semiconductors where AI can design and create beyond human...  ...we're Looking For Strong AI/ML engineering skills from top tier CS, EECS, Math... 
    Full time

    Voltai

    Palo Alto, CA
    1 day ago
  •  ...About the Role We’re looking for an Applied ML Engineer to design, evaluate, and scale recommendation and ranking systems that power how content, ads, and interactive experiences...  ...Horowitz SR04 and top angels across the video AI space. We’re still early, and this is your... 
    Full time

    Darwin

    Palo Alto, CA
    1 day ago
  • $176k - $253k

     ...the state of the art in AI, robotics, driving, and...  ...automated driving systems, tightly coupled with AD...  ...senior researchers and engineers to develop methods that...  ...implementing and evaluating ML models using Python (...  ...national origin, age, physical or mental disability, medical... 
    Local area
    Shift work

    Tri

    Los Altos, CA
    1 day ago
  • A leading AI solutions provider in Palo Alto is looking for a Senior ML Engineer to take ownership of the entire machine learning lifecycle. The ideal candidate will...  ...Python, and significant expertise with production systems using PyTorch. You will work closely with... 

    MetAntz

    Palo Alto, CA
    3 days ago
  • $152k - $241.5k

     ...weight models are foundational to American AI leadership and cybersecurity, and that...  ...scrutiny. Our AI Safety & Security Engineering team builds and evaluates AI-powered tooling...  ...first. We are looking for an Evaluation/ML-Systems Engineer to own how we measure the program... 
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    4 days ago
  • $132k - $189k

     ...tools, utilizing machine learning (ML) automation to streamline image...  ...qualifications:Bachelor’s degree in Electrical Engineering, Computer Science, Imaging Science, Physics, or a related field, or equivalent...  ...that combine the best of Google AI, software, and hardware. Teams... 

    Google

    Mountain View, CA
    4 days ago
  • $209k - $313k

     ...digital services.Snap Engineering teams build fun and technically...  ...modern ML techniques to solve large...  ...driven featuresUtilize AI tools to design and ship...  ...and applying evolving AI systems and tools to remain at...  ...national origin, ancestry, physical disability, mental disability... 
    Full time
    Live in
    Work at office
    Local area

    Snap

    Palo Alto, CA
    10 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to ML Systems Engineer, Physical AI. Be the first to apply!