Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff Engineer, Inference & RL Systems — Scale Production ML

$225k

Dormont Manufacturing Co

Dormont Manufacturing Co is looking for a Software Engineer on the Inference & RL Systems team in San Francisco. The role involves designing distributed systems, optimizing performance, and ensuring high reliability for RL and post-training workflows. The ideal candidate will possess strong software engineering fundamentals and experience with large-scale systems. Compensation includes a competitive salary range from $225K to $550K, along with equity, health benefits, and unlimited paid time off. #J-18808-Ljbffr Dormont Manufacturing Co

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Staff Engineer, Inference & RL Systems — Scale Production ML in San Francisco, CA vacancy
  •  ...in San Francisco is seeking a Staff ML Systems Engineer to design and prototype algorithms...  ...low-latency, high-throughput inference. You will implement changes in production-grade inference engines, including...  ...cost. You will also co-design RL and post-training pipelines,... 
    Suggested

    Together

    San Francisco, CA
    2 days ago
  • Additive is seeking an experienced backend engineer in San Francisco to own and evolve core ML-enabled systems. You will design, build, and scale services that handle ML inference, monitoring, and tax-related logic, while interfacing with external APIs and data stores.... 
    Suggested

    additiveai

    San Francisco, CA
    3 days ago
  • $220k

    We build and run the inference engine behind every Perplexity...  ...model architectures at scale with tight latency and...  ...to and learn from production incidents. Who we're...  ...similar). Any other deep systems programming experience...  ...if you touched any of ML compilers and framework... 
    Suggested

    Perplexity

    San Francisco, CA
    3 days ago
  •  ...Member of Technical Staff to advance state-of-...  ...the-art models, ship production-ready code, and...  ...research with production systems. You’ll work across research and engineering to push performance and scale using cutting-edge...  ...engineers, Python and ML framework expertise,... 
    Suggested
    Remote work

    Cohere

    San Francisco, CA
    4 days ago
  • $207k - $290k

     ...enterprises don't scale expertise—...  ...bringing AI systems to market...  ...actually run in production, handle real...  ...experienced AI Engineer with deep...  ...Reinforcement Learning (RL) to join our...  ...as a Senior Staff Architect. In...  ...in AI/ML engineering,...  ..., including inference-time search,... 
    Suggested
    Worldwide
    Flexible hours

    JazzX AI

    San Francisco, CA
    more than 2 months ago
  •  ...area of unprecedented scale and complexity. In no...  ...You Will Make: As a staff software engineer, you will lead two areas...  ...with different AI & ML engineering teams, cross...  ...of many AI driven products for our community....  ...teams to develop backend systems and enhance AI prompt... 
    Work experience placement
    Flexible hours

    airbnb, Inc.

    San Francisco, CA
    1 day ago
  •  ...consequences on. As a Staff Machine Learning Engineer, you’ll own AI-driven products end to end, from...  ...to the production system the mission...  ...something reliable at scale, drawing on real...  ...high-concurrency inference (Triton, vLLM, GPU...  ...shipping and operating ML-driven... 
    Full time
    Contract work
    Remote work
    Flexible hours

    Primer.ai

    San Francisco, CA
    12 hours ago
  • $180k - $336k

     ...seeking talented Senior/Staff Machine Learning Engineers with expertise in LLM training, evaluation, and production-oriented ML systems. You’ll work on improving...  .... Experience with RL post-training, such as RLHF...  ...at unprecedented speed, scale, and impact across... 
    Full time
    Work at office
    Local area
    Flexible hours

    Lila Sciences

    San Francisco, CA
    12 hours ago
  • Black Forest Labs in San Francisco, with Freiburg ties, is hiring a senior training-systems engineer to optimize production training of large multimodal models. You will work closely with researchers to improve attention performance, kernels, data movement, and training... 

    United States Digital Space LLC

    San Francisco, CA
    4 days ago
  •  ...Job Description Staff Machine Learning Engineer, Artificial Intelligence...  ...reliable, scalable, production-grade Machine Learning (ML) systems.  This role sits at...  ...evaluation systems, inference architecture, and deployment...  ...reduction, and scaling policies. - Collaborate... 
    Remote work
    Work from home

    Ginas Tech Jobs

    San Francisco, CA
    26 days ago
  • Kindredventures is recruiting infrastructure engineers to scale large-scale inference and evaluation around a physics-based LPM initiative. You will work on high-throughput systems, latency-optimized serving, and distributed orchestration across Kubernetes, Ray, and Slurm... 

    Kindredventures

    San Francisco, CA
    3 days ago
  • Sail Research in San Francisco is seeking a talented engineer to design and implement robust systems that ensure fast and cost-efficient AI inference at global scale. You will be responsible for building high-performance schedulers and optimizing global routing while focusing... 

    Sail Research

    San Francisco, CA
    3 days ago
  • $250k - $300k

     ...who believe in the scale of our ambition and...  ...and more reliably in production. That means owning the inference stack end to end:...  ...enough. This is core systems and performance...  ...directly with customer engineering teams to tailor...  ...across many kinds of ML models, with an emphasis... 
    Temporary work

    Crusoe

    San Francisco, CA
    4 days ago
  • $252k - $315k

    About Scale AIScale AI is the data foundation...  ...and deploy reliable production AI applications. We...  ...through frontier AI systems that solve real business...  ...one of the hardest engineering challenges.As a Staff Frontier Agent...  ....Unlike traditional ML roles that focus on... 
    Full time

    Scale AI

    San Francisco, CA
    2 days ago
  • $220k - $280k

     ...Staff MLOps Engineer — Machine Learning Platform Location...  ...to run them in production—at sub-10ms latency...  ...and enterprise scale. The world's largest...  ...production-grade systems that operate...  ...be our internal ML research scientists...  ...deployment, distributed inference pipelines, and... 
    Remote work

    GrabJobs

    San Francisco, CA
    5 days ago
  •  ...learning, optimization, and systems engineering to the core decisions...  ...parcel, and catering.ML models and...  ...modeling patterns that scale across DoorDash’s logistics...  ...RoleWe’re looking for a Staff Machine Learning Engineer...  ...of large-scale production ML systems that drive... 
    Hourly pay
    Work at office
    Local area
    Remote work
    Flexible hours

    Doordash

    San Francisco, CA
    4 days ago
  • $240k - $360k

     ...looking for a Senior Staff Data Engineer to be the...  ...backbone of our Data & ML Platform team — the...  ...powering analytics, product experiences, and...  ...any single team or system. You'll own the...  ...company's AI ambitions scale. You'll work in a...  ...training and inference workflows reliably... 
    Work at office
    Local area
    Immediate start
    Remote work
    Worldwide
    3 days per week

    Hinge Health

    San Francisco, CA
    12 hours ago
  • B Capital is seeking a skilled engineer for GPU infrastructure in San Francisco. This role involves designing and operating high-performance systems for model inference, synthetic data generation, and reinforcement learning. The ideal candidate has strong GPU systems experience... 

    B Capital

    San Francisco, CA
    4 days ago
  • $150k - $300k

    Prime Intellect is looking for a skilled ML Systems Engineer to build and optimize LLM serving infrastructure and inference systems. This hybrid role involves contributing to the scalability of their reinforcement learning training. Successful candidates will have over... 
    Relocation package

    Prime Intellect

    San Francisco, CA
    2 days ago
  •  ...to create their own products. Plaid powers the tools...  ...to build the shared ML and AI infrastructure...  ...develop the foundational systems, models, and data...  ...lifecycle — from large-scale data curation and...  ...across Plaid. As a Staff Machine Learning Engineer, you will lead the... 
    Full time
    Work experience placement
    Local area
    Immediate start

    Plaid Inc.

    San Francisco, CA
    12 hours ago
  •  ...business. Founded by engineers — and customer...  ...interfacing with data to scaling our services and infrastructure...  ...and Reliability systems.As a Sr. Staff Production Engineer, you will...  ...for leveraging AI/ML to revolutionize...  ...infrastructure, training/inference pipelines, or... 
    Worldwide

    DataBricks

    San Francisco, CA
    3 days ago
  •  ...Member of Technical Staff in Research...  ...build and own the ML platform researchers...  ...from data schemas to production services serving millions...  ...training and inference pipelines, lead...  ...architecture for scale, and own the GPU cluster...  ...have deep ML/system experience, Python... 

    Simile

    San Francisco, CA
    3 days ago
  •  ...Staff Machine Learning Engineer About Sprinter Health At...  ...the healthcare system, driving over $3...  ...first dedicated ML engineering hire and build the production systems that train...  ...training and inference pipelines, serving...  ...or meaningfully scaled ML infrastructure... 
    Full time
    Temporary work
    Work at office
    Relocation package
    Monday to Friday
    Monday to Thursday
    Flexible hours

    Sprinter Health

    San Francisco, CA
    12 hours ago
  • $264.8k - $331k

    Scale AI is the data foundation for AI, helping...  ...and deploy reliable production AI applications. We partner...  ..., production-grade systems that drive real...  ...the RoleAs a Senior/Staff Machine Learning Engineer (MLE) on the General...  ...Experience deploying ML systems in cloud environments... 
    Full time

    Scale AI

    San Francisco, CA
    1 day ago
  •  ...Inc. is seeking an AI-focused software engineer to design and deploy AI-powered solutions...  ...fast-paced environment. You will work with product, operations, and engineering to ship...  ...turning data and workflows into scalable systems with attention to latency, cost, and reliability... 

    Enam, Inc.

    San Francisco, CA
    2 days ago
  •  ...inspection company on the search for a Staff Systems Engineer. Our client helps manufacturers make...  ...safer, and more reliable micro- and nano-scale materials, the building blocks of...  ...and power systems. As demand surges and production grows more complex and faster, small undetected... 
    Full time

    VoltForce

    Oakland, CA
    2 days ago
  • General Motors is seeking a Staff Machine Learning Engineer to lead end-to-end ML and CV pipelines for automated map reconstruction...  .... You will architect scalable systems and collaborate across Mapping,...  ..., robust design, and national-scale deployments, with opportunities... 
    Remote job

    General Motors

    San Francisco, CA
    4 days ago
  • Sail is hiring for an engineering role in San Francisco to design and implement high-performance...  ...fleet. You will also build LV routing systems to dispatch workloads with latency...  ...caching for memory/compute trade-offs in LLM inference stacks. You will contribute to deep... 

    Sail

    San Francisco, CA
    2 days ago
  • $179k - $218k

     ...who believe in the scale of our ambition and...  ...are seeking a Senior Staff Data Center Operations Engineer, GPU Hardware Architecture...  ...: Leverage AI/ML methodologies to analyze...  ...failures in the production environment. Lead Root...  ...Analysis (RCA) on systemic issues that span the... 
    Temporary work

    Crusoe

    San Francisco, CA
    3 days ago
  • Jaide Health is seeking an engineer for their Model Efficiency team in San Francisco. The role focuses on building reliable ML systems while enhancing core performance metrics across model...  ...or Python and insights into the LLM inference ecosystem. A commitment to diversity... 
    Remote job

    Jaide Health

    San Francisco, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff Engineer, Inference & RL Systems — Scale Production ML. Be the first to apply!