Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

ML Systems Engineer - Large-Scale, Low-Latency Infra

Jobleads-US

Meta is hiring a Software Engineer for the Systems ML team in Menlo Park, CA. You will design and optimize large-scale ML training and inference systems, spanning the full stack from pipelines to hardware-aware optimizations.

You will collaborate with researchers, platform engineers, and product teams to accelerate workloads and improve AI infrastructure for billions of users, ensuring reliability and low-latency performance.

#J-18808-Ljbffr Jobleads-US
Vacancy posted 10 hours ago
Similar jobs that could be interesting for youBased on the ML Systems Engineer - Large-Scale, Low-Latency Infra in Menlo Park, CA vacancy
  •  ...Institute of Foundation Models is seeking a distributed ML infrastructure engineer to extend and scale our training systems. You’ll work with world‑class researchers to...  ...aim to improve reliability and performance of large‑scale pre‑training pipelines. #J-18808-Ljbffr... 
    Suggested

    Jobleads-US

    Sunnyvale, CA
    10 hours ago
  • $300k - $400k

     ...You will own the systems layer that makes our...  ..., minimizing latency and maximizing utilization...  .../O bottlenecks in large-scale training runs...  ...RL loop tight and low-latency Build...  ...benchmarking distributed ML systems to...  ...— the scientists, engineers, and problem-solvers... 
    Suggested
    Visa sponsorship
    Flexible hours
    Shift work

    Periodic Labs

    Menlo Park, CA
    2 days ago
  • $124k - $250k

     ...LifeAs a member of our software engineering infra team, you'll solve...  ...rapid development of novel new systems that integrate into that ecosystem...  ...with high throughput and low latency to our bidding ecosystem....  ..., develop, and maintain large-scale distributed systemsCollaborate... 
    Suggested

    AppLovin

    Palo Alto, CA
    2 days ago
  • $165.2k - $223.6k

     ...self-driven and talented engineers to revel in designing,...  ...extremely high volume (internet-scale), low latency systems to drive revenue to...  ...will: - You’ll build ML based solutions & infra that (1) handles supremely...  ...and developing large-scale, multi-tiered, multi... 
    Suggested
    Internship
    Local area
    Immediate start
    Remote work
    Flexible hours

    Amazon

    Palo Alto, CA
    1 day ago
  •  ...Conversational AI system that integrates seamlessly...  ....About the AI & ML Platform TeamOur...  ...Senior ML System Engineer on the AI & ML...  ...design and optimize large-scale model serving...  ...-scaling) to deep low-level optimizations...  ...KV cache).Optimize latency and throughput of... 
    Suggested
    Work at office
    Local area

    Atlassian

    Mountain View, CA
    1 day ago
  • $171.7k - $303.9k

     ...The Data Labeling Engineering team designs, builds...  ..., and AI/ML, defining the strategies...  ...training data at scale. Our tools and platform...  ...across teams and systems that unblock the...  ...quality, cost, and latency goals.Level up how...  ...or tools used by large labeling workforces... 
    Full time
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    3 days ago
  •  ...Machine Learning Systems Engineer (P60) to lead technical...  ...on building and scaling the systems that...  ...Design and Build ML SystemsArchitect and...  ..., and serving large language models and...  ...Develop tools and infra to support rapid experimentation...  ...services.Optimize latency, throughput, and... 
    Work at office
    Local area

    Atlassian

    Mountain View, CA
    19 hours ago
  •  ...General Motors is seeking a Senior Performance Engineer to join the AV Capacity and Performance Engineering team. You will help develop and optimize large-scale ML infrastructure and GPU platforms for autonomous-vehicle research and deployment. Responsibilities include... 

    Jobleads-US

    Sunnyvale, CA
    3 days ago
  •  ...developing end-to-end ML models for robot...  ...foundational systems that directly accelerate...  ...multimodal data, scaling distributed...  ...throughput Build low-latency inference pipeline...  ...dataset management at large scale Optimize...  ...Strong software engineering and systems fundamentals... 

    Sunday

    Redwood City, CA
    3 days ago
  •  ...hiring a Member of Technical Staff, ML Systems to accelerate model training and inference...  ...kernels, runtimes, and distributed engines that power production-scale ML stacks. You’ll optimize GPU...  ...with Nsight, and implement low-level CUDA and Triton improvements.... 

    Jobleads-US

    Menlo Park, CA
    19 hours ago
  •  ...logging, processing, and ML model-training pipelines. Develop and scale infrastructure for...  ...scalable ML-based ranking systems and online ML inference...  ...managers, indexing platform engineers, ranking engineers, and...  ...high-throughput and low-latency services, and cloud... 
    Full time

    Coupang

    Mountain View, CA
    17 days ago
  • $220k - $350k

     ...Description Job Description ML Infrastructure Engineer Company: Dyna...  ...Do Architect and scale distributed training across large GPU clusters,...  ...never starve. Build low-latency inference pipelines for...  ...infrastructure hire Multimodal systems (video, audio,... 
    Full time
    H1b
    Work at office
    Visa sponsorship

    Transparent Search Group

    Redwood City, CA
    17 days ago
  • $169k - $338k

     ...Segment: Home OfficeAI/ML & Agentic Systems Technical Leadership:...  ...complex reliability engineering workflows, predictive...  ...improve reliability, latency, availability, and...  ...all domains, 2) Enable scaling by providing technical...  ...and analysis of large-scale distributed systems... 
    Full time
    Temporary work
    Part time

    Walmart

    Sunnyvale, CA
    4 days ago
  •  ...search quality, tail latency, reliability, scalability...  ...determine where new ML capabilities can meaningfully...  ...agentic search systems, addressing tail latency...  ..., mentor senior engineers, and align technical and...  ...designing and operating large-scale, low-latency search, recommendation... 
    Work at office
    Local area

    Atlassian

    Mountain View, CA
    19 hours ago
  • $159.3k - $230.7k

     ...on real roads, at real scale. Making self-driving safe...  .... The team delivers ML models that move the product...  ...We work with the very large datasets GM already has...  ....As a Senior AI/ML Engineer in the Embodied AI Data...  ...models into onboard driving systems, and document learnings... 
    Full time
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    4 days ago
  • $170.6k - $261.3k

     ...hardware and battery systems to intuitive...  ...on a global scale.As a Senior Machine Learning Engineer on the State Estimation...  ...and improve the ML perception model...  .../recall, latency, robustness under...  ...metrics. Analyze large‑scale datasets,...  ...with software and infra engineers to integrate... 
    Full time
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    1 day ago
  •  ...Google LLC in Mountain View, CA seeks software engineers to develop next-generation technologies that impact billions of users. You...  ...spanning information retrieval, distributed computing, and large-scale systems, with opportunities to switch teams as our fast-paced business... 

    Jobleads-US

    Mountain View, CA
    4 days ago
  •  ...their business systems through natural...  ...Moveworks’ Reasoning Engine and natural...  ...backed by the global scale of ServiceNow...  ...build cutting edge ML infrastructure...  ...systems. The ML infra team covers a...  ...inference pipeline for large language models(...  ...framework, LLM latency optimization,... 
    Permanent employment
    Work at office
    Remote work
    Flexible hours

    Moveworks

    Mountain View, CA
    4 days ago
  • $189k - $300k

     ...breakthrough hardware and battery systems to intuitive design,...  ...on a global scale. The Data Scaling team...  ...works on and delivers ML models to the product that...  ...team uses existing very large datasets that GM has access...  ...-impact team of AI/ML engineers, data scientists and... 
    Full time
    Local area
    Remote work
    Work from home
    Relocation
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    1 day ago
  • $207k - $300k

     ...through advanced context engineering and agentic feedback...  ...stack.Resolve key system-level bottlenecks in recommending...  ...efficiency for low-latency surfaces.Navigate high...  ...frameworks for Large Language Models (e.g.,...  ...information at massive scale, and extend well beyond... 

    Google

    Mountain View, CA
    2 days ago
  •  ...Machine Learning Engineer Chicago, IL;...  ...each month. Our systems process and move...  ...design, build, and scale the systems that...  ...transparent, low-fee financial services...  ...our production ML systems and...  ...Experience building low-latency online model...  ...working with large, messy, real-... 
    Work at office
    Immediate start
    Remote work

    Attain

    Redwood City, CA
    4 days ago
  • $221.38k - $263.67k

     ...technical challenges at scale, and helping to...  ...analytical data engineering, product analytics...  ....WHY ENGINE INFRA?Engine Infra team...  ...allocation, and edge latency.You Will:Define Metrics...  ...and decision systems that help Engine Infra...  ...multiplexed and large-server processes healthy... 
    Full time
    Work experience placement
    H1b
    Work at office
    Local area
    Visa sponsorship
    Monday to Friday

    Roblox

    San Mateo, CA
    12 hours ago
  •  ...for a Senior MLOps engineer to work closely with...  ...build and deploy ML models on a modern...  ...learning algorithms on large compute clusters,...  ...batch model serving systems, hyper-parameter tuning at scale, model monitoring,...  ...supports high throughput, low latency applications which... 

    JP Morgan Chase

    Palo Alto, CA
    1 day ago
  • $150k - $300k

     ...an accomplished Senior Staff ML Engineer who will serve as a technical...  ...design, develop, and deploy systems that ensure scalability, reliability...  ...by example to solve complex, large‑scale, cross‑functional problems....  ...workflows via both no code/low code and traditional high‑... 
    Hourly pay
    Work experience placement
    Local area

    Government Employees Insurance Company

    Palo Alto, CA
    1 day ago
  • $210k - $275k

     ...responsible for ML and work...  ...scientists and engineers. As a Staff Machine...  ...in order to scale and optimize our ML systems—creating and transforming...  ..., and evolve large-scale SID /...  ...quality, latency, cost, and...  ...with product, infra, research, and...  ...represents the low and high end of... 
    Permanent employment
    Immediate start

    Cacheflow

    Mountain View, CA
    2 days ago
  • $180k

     ...mission is to create AI systems that can accurately...  ..., and focused on engineering excellence. This...  ...THE ROLE: As an ML Infrastructure...  ...Designing, building, and scaling GPU compute...  ...pipelines and integrating large-scale data,...  ...JAX or PyTorch Low-level understanding... 
    Temporary work
    Work experience placement

    SpaceXAI

    Palo Alto, CA
    3 days ago
  • $275.8k - $340.5k

     ...About the team: The AV ML Infra team at GM builds ML infrastructure...  ...the productivity of ML engineers, and drive the adoption of...  ...model performance by running large-scale simulation workloads and managing...  ...Together, these tools and systems empower GM to tackle the... 
    Full time
    Local area
    Remote work
    Work from home
    Relocation
    Relocation package
    Flexible hours

    General Motors

    Mountain View, CA
    4 days ago
  •  ...Snap Inc in Palo Alto is seeking a Machine Learning Engineer to build and deploy ML models powering core products used by millions of Snapchatters. You will apply modern ML methods to solve large-scale problems and own the full lifecycle from data analysis to production... 

    Jobleads-US

    Palo Alto, CA
    1 day ago
  • $270k

     ...Inworld in Mountain View is seeking engineers to advance real-time, agentic AI systems at scale. You will own end-to-end deployment from research to production, optimizing latency and throughput on multi-GPU infrastructure. Relocation assistance is offered; base salary... 
    Relocation package

    Jobleads-US

    Mountain View, CA
    19 hours ago
  • $115k - $230k

     ...Implementation of ML Models: Lead the architecture...  ...Units, and Engineering teams. Build...  ...pipelines, ensuring that systems are reliable and performant at scale. Write...  ...systems that handle large‑scale data and high...  ...inference pipelines and low‑latency model serving. Familiar... 
    Hourly pay
    Work experience placement
    Local area

    Jobleads-US

    Palo Alto, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to ML Systems Engineer - Large-Scale, Low-Latency Infra. Be the first to apply!