Staff Engineer, Inference & RL Systems — Scale Production ML
$225kDormont Manufacturing Co
Dormont Manufacturing Co is looking for a Software Engineer on the Inference & RL Systems team in San Francisco. The role involves designing distributed systems, optimizing performance, and ensuring high reliability for RL and post-training workflows. The ideal candidate will possess strong software engineering fundamentals and experience with large-scale systems. Compensation includes a competitive salary range from $225K to $550K, along with equity, health benefits, and unlimited paid time off. #J-18808-Ljbffr Dormont Manufacturing Co
- ...in San Francisco is seeking a Staff ML Systems Engineer to design and prototype algorithms... ...low-latency, high-throughput inference. You will implement changes in production-grade inference engines, including... ...cost. You will also co-design RL and post-training pipelines,...Suggested
- ...infrastructure for high-throughput model inference and mid-training workloads. Join a team that powers synthetic data generation, RL pipelines, and distributed model evaluation... ...performance tuning, kernel optimization, and scalable distributed systems. #J-18808-Ljbffr Visa HuntSuggested
$220k
We build and run the inference engine behind every Perplexity... ...model architectures at scale with tight latency and... ...to and learn from production incidents. Who we're... ...similar). Any other deep systems programming experience... ...if you touched any of ML compilers and framework...Suggested- ...Member of Technical Staff to advance state-of-... ...the-art models, ship production-ready code, and... ...research with production systems. You’ll work across research and engineering to push performance and scale using cutting-edge... ...engineers, Python and ML framework expertise,...SuggestedRemote work
- ...consequences on. As a Staff Machine Learning Engineer, you’ll own AI-driven products end to end, from... ...to the production system the mission... ...something reliable at scale, drawing on real... ...high-concurrency inference (Triton, vLLM, GPU... ...shipping and operating ML-driven...SuggestedFull timeContract workRemote workFlexible hours
$180k - $336k
...seeking talented Senior/Staff Machine Learning Engineers with expertise in LLM training, evaluation, and production-oriented ML systems. You’ll work on improving... .... Experience with RL post-training, such as RLHF... ...at unprecedented speed, scale, and impact across...Full timeWork at officeLocal areaFlexible hours$203.5k - $299.3k
...causal decisioning systems for New Verticals:... ...Causal Machine Learning Engineer to help build the causal ML foundation behind... ...deeply worked on production causal systems: uplift... ...spine for a large-scale consumer... ...experience with causal inference, econometrics, experimentation...Hourly payWork at officeLocal areaRemote workFlexible hours$190.2k - $345.65k
...paired with creative production workflows and a... ...hiring a Senior Staff Machine Learning Engineer to architect and... ...— the systems that turn massive... ...index media at scale, the hybrid and... ...that improve them.ML Engineering leadership... ...models and the inference paths that produce...Full timeTemporary workLocal areaWorldwide- Jack & Jill in San Francisco builds RL environments and robust backend services to stress-test frontier AI models... ...with Applied Researchers to translate designs into production-grade software and evaluation systems. The role demands hands-on work with React/TypeScript...
$200k - $250k
...for frontier AI systems in financial... ...and full-stack engineers from our network... ...stage team, and is scaling quickly as... ...for high-quality RL environments grows... ...engineering, ML infrastructure,... ...systems, and product development, with... ...training and inference infrastructure...Full timeH1bWork at officeRelocation package- ...causal decisioning systems for New Verticals:... ...Causal Machine Learning Engineer to help build the causal ML foundation behind... ...deeply worked on production causal systems: uplift... ...spine for a large-scale consumer... ...experience with causal inference, econometrics, experimentation...Hourly payWork at officeLocal areaFlexible hours
$203.5k - $299.3k
...causal decisioning systems for New Verticals:... ...Causal Machine Learning Engineer to help build the causal ML foundation behind... ...deeply worked on production causal systems: uplift... ...spine for a large-scale consumer... ...experience with causal inference, econometrics, experimentation...Hourly payWork at officeLocal areaRemote workFlexible hours- Kindredventures is recruiting infrastructure engineers to scale large-scale inference and evaluation around a physics-based LPM initiative. You will work on high-throughput systems, latency-optimized serving, and distributed orchestration across Kubernetes, Ray, and Slurm...
- Sail Research in San Francisco is seeking a talented engineer to design and implement robust systems that ensure fast and cost-efficient AI inference at global scale. You will be responsible for building high-performance schedulers and optimizing global routing while focusing...
- ...building the world's most efficient software for inference and agent hosting. In this role, you'll be one of the first engineers on Sailboxes, contributing across the stack—... ...custom networking stack to building large-scale systems that maximize efficiency. You'll focus on...Work at office
- HausHaus is seeking a senior ML Engineer to drive high-impact projects in advanced marketing planning, analysis, and... ...deliver trustworthy results, building scalable ML systems for cMMM and turning ideas into production-ready solutions. You will mentor engineers,...
$215k - $260k
...who believe in the scale of our ambition and... ...and more reliably in production. That means owning the inference stack end to end:... ...enough. This is core systems and performance... ...directly with customer engineering teams to tailor... ...across many kinds of ML models, with an emphasis...Temporary work$252k - $315k
About Scale AIScale AI is the data foundation... ...and deploy reliable production AI applications. We... ...through frontier AI systems that solve real business... ...one of the hardest engineering challenges.As a Staff Frontier Agent... ....Unlike traditional ML roles that focus on...Full time- Inception is seeking engineers and scientists to design, optimize, and scale the diffusion LLM serving systems powering production inference. Your work will help make inference faster, more cost-efficient, and more reliable. You will extend orchestration frameworks (Kubernetes...
- ...marketing decisions at scale. The Role... ..., and causal inference. We are looking... ...in writing production code, turning... ...ideas to scalable systems. This role... ...scientists, data engineers and other MLEs... ...machine learning (ML) solutions for... ...reinforcement learning (RL), Bayesian...Full timeWork at officeWork from homeWorldwideFlexible hours
- ...Member of Technical Staff for the Frontier... ...team to design RL environments, evaluations... ...of tasks at scale and collaborate... ...researchers and engineers across AI labs. This... ...role spanning systems engineering, ML tooling, and research... ..., with production hardening and open...
- ...seeking a Member of Technical Staff for their infrastructure team... ...role, you will own the cloud systems that serve our compression... ...latency, high-throughput GPU ML inference infrastructure. The ideal candidate... ...track record in building production environments. Additional benefits...Visa sponsorship
- B Capital is seeking a skilled engineer for GPU infrastructure in San Francisco. This role involves designing and operating high-performance systems for model inference, synthetic data generation, and reinforcement learning. The ideal candidate has strong GPU systems experience...
$150k - $300k
Prime Intellect is looking for a skilled ML Systems Engineer to build and optimize LLM serving infrastructure and inference systems. This hybrid role involves contributing to the scalability of their reinforcement learning training. Successful candidates will have over...Relocation package$207k - $290k
...enterprises don't scale expertise—... ...bringing AI systems to market... ...actually run in production, handle real... ...experienced AI Engineer with deep... ...Reinforcement Learning (RL) to join our... ...as a Senior Staff Architect. In... ...in AI/ML engineering,... ..., including inference-time search,...WorldwideFlexible hours$240k - $360k
...looking for a Senior Staff Data Engineer to be the... ...backbone of our Data & ML Platform team — the... ...powering analytics, product experiences, and... ...any single team or system. You'll own the... ...company's AI ambitions scale. You'll work in a... ...training and inference workflows reliably...Work at officeLocal areaImmediate startRemote workWorldwide3 days per week- ...to create their own products. Plaid powers the tools... ...to build the shared ML and AI infrastructure... ...develop the foundational systems, models, and data... ...lifecycle — from large-scale data curation and... ...across Plaid. As a Staff Machine Learning Engineer, you will lead the technical...Full timeWork experience placementLocal areaImmediate start
- ...Member of Technical Staff in Research... ...build and own the ML platform researchers... ...from data schemas to production services serving millions... ...training and inference pipelines, lead... ...architecture for scale, and own the GPU cluster... ...have deep ML/system experience, Python...
- Inception is seeking engineers and scientists to design, optimize, and maintain core systems enabling scalable reinforcement learning... ...orchestration, and ensuring production-readiness. Responsibilities include... ...building infrastructure for RL workloads, boosting training...
- A leading AI research firm in San Francisco is seeking a Member of Technical Staff specialized in Model Efficiency. In this role, you will enhance LLM inference systems by tackling performance issues and collaborating with cross-functional teams. Ideal candidates have over...Remote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff Engineer, Inference & RL Systems — Scale Production ML. Be the first to apply!
- project engineer assistant project manager San Francisco, CA
- senior staff systems engineer San Francisco, CA
- staff data engineer San Francisco, CA
- assistant chief engineer San Francisco, CA
- assistant engineer San Francisco, CA
- assistant electrical engineer San Francisco, CA
- engineering aide San Francisco, CA
- software engineer staff San Francisco, CA
- staff design engineer San Francisco, CA
- research assistant engineering San Francisco, CA

