Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff Engineer, Inference & RL Systems — Scale Production ML

$225k

Dormont Manufacturing Co

Dormont Manufacturing Co is looking for a Software Engineer on the Inference & RL Systems team in San Francisco. The role involves designing distributed systems, optimizing performance, and ensuring high reliability for RL and post-training workflows. The ideal candidate will possess strong software engineering fundamentals and experience with large-scale systems. Compensation includes a competitive salary range from $225K to $550K, along with equity, health benefits, and unlimited paid time off. #J-18808-Ljbffr Dormont Manufacturing Co

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Staff Engineer, Inference & RL Systems — Scale Production ML in San Francisco, CA vacancy
  •  ...in San Francisco is seeking a Staff ML Systems Engineer to design and prototype algorithms...  ...low-latency, high-throughput inference. You will implement changes in production-grade inference engines, including...  ...cost. You will also co-design RL and post-training pipelines,... 
    Suggested

    Together

    San Francisco, CA
    2 days ago
  •  ...infrastructure for high-throughput model inference and mid-training workloads. Join a team that powers synthetic data generation, RL pipelines, and distributed model evaluation...  ...performance tuning, kernel optimization, and scalable distributed systems. #J-18808-Ljbffr Visa Hunt
    Suggested

    Visa Hunt

    San Francisco, CA
    4 days ago
  • $220k

    We build and run the inference engine behind every Perplexity...  ...model architectures at scale with tight latency and...  ...to and learn from production incidents. Who we're...  ...similar). Any other deep systems programming experience...  ...if you touched any of ML compilers and framework... 
    Suggested

    Perplexity

    San Francisco, CA
    3 days ago
  •  ...Member of Technical Staff to advance state-of-...  ...the-art models, ship production-ready code, and...  ...research with production systems. You’ll work across research and engineering to push performance and scale using cutting-edge...  ...engineers, Python and ML framework expertise,... 
    Suggested
    Remote work

    Cohere

    San Francisco, CA
    4 days ago
  •  ...consequences on. As a Staff Machine Learning Engineer, you’ll own AI-driven products end to end, from...  ...to the production system the mission...  ...something reliable at scale, drawing on real...  ...high-concurrency inference (Triton, vLLM, GPU...  ...shipping and operating ML-driven... 
    Suggested
    Full time
    Contract work
    Remote work
    Flexible hours

    Primer.ai

    San Francisco, CA
    18 hours ago
  • $180k - $336k

     ...seeking talented Senior/Staff Machine Learning Engineers with expertise in LLM training, evaluation, and production-oriented ML systems. You’ll work on improving...  .... Experience with RL post-training, such as RLHF...  ...at unprecedented speed, scale, and impact across... 
    Full time
    Work at office
    Local area
    Flexible hours

    Lila Sciences

    San Francisco, CA
    18 hours ago
  • $203.5k - $299.3k

     ...causal decisioning systems for New Verticals:...  ...Causal Machine Learning Engineer to help build the causal ML foundation behind...  ...deeply worked on production causal systems: uplift...  ...spine for a large-scale consumer...  ...experience with causal inference, econometrics, experimentation... 
    Hourly pay
    Work at office
    Local area
    Remote work
    Flexible hours

    Doordash

    San Francisco, CA
    1 day ago
  • $190.2k - $345.65k

     ...paired with creative production workflows and a...  ...hiring a Senior Staff Machine Learning Engineer to architect and...  ...— the systems that turn massive...  ...index media at scale, the hybrid and...  ...that improve them.ML Engineering leadership...  ...models and the inference paths that produce... 
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Francisco, CA
    1 day ago
  • Jack & Jill in San Francisco builds RL environments and robust backend services to stress-test frontier AI models...  ...with Applied Researchers to translate designs into production-grade software and evaluation systems. The role demands hands-on work with React/TypeScript... 

    Jack & Jill

    San Francisco, CA
    4 days ago
  • $200k - $250k

     ...for frontier AI systems in financial...  ...and full-stack engineers from our network...  ...stage team, and is scaling quickly as...  ...for high-quality RL environments grows...  ...engineering, ML infrastructure,...  ...systems, and product development, with...  ...training and inference infrastructure... 
    Full time
    H1b
    Work at office
    Relocation package

    CoffeeSpace

    San Francisco, CA
    3 days ago
  •  ...causal decisioning systems for New Verticals:...  ...Causal Machine Learning Engineer to help build the causal ML foundation behind...  ...deeply worked on production causal systems: uplift...  ...spine for a large-scale consumer...  ...experience with causal inference, econometrics, experimentation... 
    Hourly pay
    Work at office
    Local area
    Flexible hours

    DoorDash, Inc.

    San Francisco, CA
    4 days ago
  • $203.5k - $299.3k

     ...causal decisioning systems for New Verticals:...  ...Causal Machine Learning Engineer to help build the causal ML foundation behind...  ...deeply worked on production causal systems: uplift...  ...spine for a large-scale consumer...  ...experience with causal inference, econometrics, experimentation... 
    Hourly pay
    Work at office
    Local area
    Remote work
    Flexible hours

    Visa Hunt

    San Francisco, CA
    4 days ago
  • Kindredventures is recruiting infrastructure engineers to scale large-scale inference and evaluation around a physics-based LPM initiative. You will work on high-throughput systems, latency-optimized serving, and distributed orchestration across Kubernetes, Ray, and Slurm... 

    Kindredventures

    San Francisco, CA
    3 days ago
  • Sail Research in San Francisco is seeking a talented engineer to design and implement robust systems that ensure fast and cost-efficient AI inference at global scale. You will be responsible for building high-performance schedulers and optimizing global routing while focusing... 

    Sail Research

    San Francisco, CA
    3 days ago
  •  ...building the world's most efficient software for inference and agent hosting. In this role, you'll be one of the first engineers on Sailboxes, contributing across the stack—...  ...custom networking stack to building large-scale systems that maximize efficiency. You'll focus on... 
    Work at office

    Sail Research Inc.

    San Francisco, CA
    18 hours ago
  • HausHaus is seeking a senior ML Engineer to drive high-impact projects in advanced marketing planning, analysis, and...  ...deliver trustworthy results, building scalable ML systems for cMMM and turning ideas into production-ready solutions. You will mentor engineers,... 

    Haus.com

    San Francisco, CA
    9 hours ago
  • $215k - $260k

     ...who believe in the scale of our ambition and...  ...and more reliably in production. That means owning the inference stack end to end:...  ...enough. This is core systems and performance...  ...directly with customer engineering teams to tailor...  ...across many kinds of ML models, with an emphasis... 
    Temporary work

    Crusoe

    San Francisco, CA
    4 days ago
  • $252k - $315k

    About Scale AIScale AI is the data foundation...  ...and deploy reliable production AI applications. We...  ...through frontier AI systems that solve real business...  ...one of the hardest engineering challenges.As a Staff Frontier Agent...  ....Unlike traditional ML roles that focus on... 
    Full time

    Scale AI

    San Francisco, CA
    3 days ago
  • Inception is seeking engineers and scientists to design, optimize, and scale the diffusion LLM serving systems powering production inference. Your work will help make inference faster, more cost-efficient, and more reliable. You will extend orchestration frameworks (Kubernetes... 

    Inception LLC

    San Francisco, CA
    3 days ago
  •  ...marketing decisions at scale. The Role...  ..., and causal inference. We are looking...  ...in writing production code, turning...  ...ideas to scalable systems. This role...  ...scientists, data engineers and other MLEs...  ...machine learning (ML) solutions for...  ...reinforcement learning (RL), Bayesian... 
    Full time
    Work at office
    Work from home
    Worldwide
    Flexible hours

    Haus Analytics

    San Francisco, CA
    18 hours ago
  •  ...Member of Technical Staff for the Frontier...  ...team to design RL environments, evaluations...  ...of tasks at scale and collaborate...  ...researchers and engineers across AI labs. This...  ...role spanning systems engineering, ML tooling, and research...  ..., with production hardening and open... 

    Roboflow, Inc.

    San Francisco, CA
    1 day ago
  •  ...seeking a Member of Technical Staff for their infrastructure team...  ...role, you will own the cloud systems that serve our compression...  ...latency, high-throughput GPU ML inference infrastructure. The ideal candidate...  ...track record in building production environments. Additional benefits... 
    Visa sponsorship

    The Token Company

    San Francisco, CA
    2 days ago
  • B Capital is seeking a skilled engineer for GPU infrastructure in San Francisco. This role involves designing and operating high-performance systems for model inference, synthetic data generation, and reinforcement learning. The ideal candidate has strong GPU systems experience... 

    B Capital

    San Francisco, CA
    4 days ago
  • $150k - $300k

    Prime Intellect is looking for a skilled ML Systems Engineer to build and optimize LLM serving infrastructure and inference systems. This hybrid role involves contributing to the scalability of their reinforcement learning training. Successful candidates will have over... 
    Relocation package

    Prime Intellect

    San Francisco, CA
    2 days ago
  • $207k - $290k

     ...enterprises don't scale expertise—...  ...bringing AI systems to market...  ...actually run in production, handle real...  ...experienced AI Engineer with deep...  ...Reinforcement Learning (RL) to join our...  ...as a Senior Staff Architect. In...  ...in AI/ML engineering,...  ..., including inference-time search,... 
    Worldwide
    Flexible hours

    JazzX AI

    San Francisco, CA
    15 days ago
  • $240k - $360k

     ...looking for a Senior Staff Data Engineer to be the...  ...backbone of our Data & ML Platform team — the...  ...powering analytics, product experiences, and...  ...any single team or system. You'll own the...  ...company's AI ambitions scale. You'll work in a...  ...training and inference workflows reliably... 
    Work at office
    Local area
    Immediate start
    Remote work
    Worldwide
    3 days per week

    Hinge Health

    San Francisco, CA
    5 days ago
  •  ...to create their own products. Plaid powers the tools...  ...to build the shared ML and AI infrastructure...  ...develop the foundational systems, models, and data...  ...lifecycle — from large-scale data curation and...  ...across Plaid. As a Staff Machine Learning Engineer, you will lead the technical... 
    Full time
    Work experience placement
    Local area
    Immediate start

    Plaid Inc.

    San Francisco, CA
    18 hours ago
  •  ...Member of Technical Staff in Research...  ...build and own the ML platform researchers...  ...from data schemas to production services serving millions...  ...training and inference pipelines, lead...  ...architecture for scale, and own the GPU cluster...  ...have deep ML/system experience, Python... 

    Simile

    San Francisco, CA
    3 days ago
  • Inception is seeking engineers and scientists to design, optimize, and maintain core systems enabling scalable reinforcement learning...  ...orchestration, and ensuring production-readiness. Responsibilities include...  ...building infrastructure for RL workloads, boosting training... 

    Inception

    San Francisco, CA
    3 days ago
  • A leading AI research firm in San Francisco is seeking a Member of Technical Staff specialized in Model Efficiency. In this role, you will enhance LLM inference systems by tackling performance issues and collaborating with cross-functional teams. Ideal candidates have over... 
    Remote work

    Cohere

    San Francisco, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff Engineer, Inference & RL Systems — Scale Production ML. Be the first to apply!