Staff Engineer, Inference & RL Systems — Scale Production ML
$225kDormont Manufacturing Co
Dormont Manufacturing Co is looking for a Software Engineer on the Inference & RL Systems team in San Francisco. The role involves designing distributed systems, optimizing performance, and ensuring high reliability for RL and post-training workflows. The ideal candidate will possess strong software engineering fundamentals and experience with large-scale systems. Compensation includes a competitive salary range from $225K to $550K, along with equity, health benefits, and unlimited paid time off. #J-18808-Ljbffr Dormont Manufacturing Co
- ...in San Francisco is seeking a Staff ML Systems Engineer to design and prototype algorithms... ...low-latency, high-throughput inference. You will implement changes in production-grade inference engines, including... ...cost. You will also co-design RL and post-training pipelines,...Suggested
- Additive is seeking an experienced backend engineer in San Francisco to own and evolve core ML-enabled systems. You will design, build, and scale services that handle ML inference, monitoring, and tax-related logic, while interfacing with external APIs and data stores....Suggested
$220k
We build and run the inference engine behind every Perplexity... ...model architectures at scale with tight latency and... ...to and learn from production incidents. Who we're... ...similar). Any other deep systems programming experience... ...if you touched any of ML compilers and framework...Suggested- ...Member of Technical Staff to advance state-of-... ...the-art models, ship production-ready code, and... ...research with production systems. You’ll work across research and engineering to push performance and scale using cutting-edge... ...engineers, Python and ML framework expertise,...SuggestedRemote work
$207k - $290k
...enterprises don't scale expertise—... ...bringing AI systems to market... ...actually run in production, handle real... ...experienced AI Engineer with deep... ...Reinforcement Learning (RL) to join our... ...as a Senior Staff Architect. In... ...in AI/ML engineering,... ..., including inference-time search,...SuggestedWorldwideFlexible hours- ...area of unprecedented scale and complexity. In no... ...You Will Make: As a staff software engineer, you will lead two areas... ...with different AI & ML engineering teams, cross... ...of many AI driven products for our community.... ...teams to develop backend systems and enhance AI prompt...Work experience placementFlexible hours
- ...consequences on. As a Staff Machine Learning Engineer, you’ll own AI-driven products end to end, from... ...to the production system the mission... ...something reliable at scale, drawing on real... ...high-concurrency inference (Triton, vLLM, GPU... ...shipping and operating ML-driven...Full timeContract workRemote workFlexible hours
$180k - $336k
...seeking talented Senior/Staff Machine Learning Engineers with expertise in LLM training, evaluation, and production-oriented ML systems. You’ll work on improving... .... Experience with RL post-training, such as RLHF... ...at unprecedented speed, scale, and impact across...Full timeWork at officeLocal areaFlexible hours- Black Forest Labs in San Francisco, with Freiburg ties, is hiring a senior training-systems engineer to optimize production training of large multimodal models. You will work closely with researchers to improve attention performance, kernels, data movement, and training...
- ...Job Description Staff Machine Learning Engineer, Artificial Intelligence... ...reliable, scalable, production-grade Machine Learning (ML) systems. This role sits at... ...evaluation systems, inference architecture, and deployment... ...reduction, and scaling policies. - Collaborate...Remote workWork from home
- Kindredventures is recruiting infrastructure engineers to scale large-scale inference and evaluation around a physics-based LPM initiative. You will work on high-throughput systems, latency-optimized serving, and distributed orchestration across Kubernetes, Ray, and Slurm...
- Sail Research in San Francisco is seeking a talented engineer to design and implement robust systems that ensure fast and cost-efficient AI inference at global scale. You will be responsible for building high-performance schedulers and optimizing global routing while focusing...
$250k - $300k
...who believe in the scale of our ambition and... ...and more reliably in production. That means owning the inference stack end to end:... ...enough. This is core systems and performance... ...directly with customer engineering teams to tailor... ...across many kinds of ML models, with an emphasis...Temporary work$252k - $315k
About Scale AIScale AI is the data foundation... ...and deploy reliable production AI applications. We... ...through frontier AI systems that solve real business... ...one of the hardest engineering challenges.As a Staff Frontier Agent... ....Unlike traditional ML roles that focus on...Full time$220k - $280k
...Staff MLOps Engineer — Machine Learning Platform Location... ...to run them in production—at sub-10ms latency... ...and enterprise scale. The world's largest... ...production-grade systems that operate... ...be our internal ML research scientists... ...deployment, distributed inference pipelines, and...Remote work- ...learning, optimization, and systems engineering to the core decisions... ...parcel, and catering.ML models and... ...modeling patterns that scale across DoorDash’s logistics... ...RoleWe’re looking for a Staff Machine Learning Engineer... ...of large-scale production ML systems that drive...Hourly payWork at officeLocal areaRemote workFlexible hours
$240k - $360k
...looking for a Senior Staff Data Engineer to be the... ...backbone of our Data & ML Platform team — the... ...powering analytics, product experiences, and... ...any single team or system. You'll own the... ...company's AI ambitions scale. You'll work in a... ...training and inference workflows reliably...Work at officeLocal areaImmediate startRemote workWorldwide3 days per week- B Capital is seeking a skilled engineer for GPU infrastructure in San Francisco. This role involves designing and operating high-performance systems for model inference, synthetic data generation, and reinforcement learning. The ideal candidate has strong GPU systems experience...
$150k - $300k
Prime Intellect is looking for a skilled ML Systems Engineer to build and optimize LLM serving infrastructure and inference systems. This hybrid role involves contributing to the scalability of their reinforcement learning training. Successful candidates will have over...Relocation package- ...to create their own products. Plaid powers the tools... ...to build the shared ML and AI infrastructure... ...develop the foundational systems, models, and data... ...lifecycle — from large-scale data curation and... ...across Plaid. As a Staff Machine Learning Engineer, you will lead the...Full timeWork experience placementLocal areaImmediate start
- ...business. Founded by engineers — and customer... ...interfacing with data to scaling our services and infrastructure... ...and Reliability systems.As a Sr. Staff Production Engineer, you will... ...for leveraging AI/ML to revolutionize... ...infrastructure, training/inference pipelines, or...Worldwide
- ...Member of Technical Staff in Research... ...build and own the ML platform researchers... ...from data schemas to production services serving millions... ...training and inference pipelines, lead... ...architecture for scale, and own the GPU cluster... ...have deep ML/system experience, Python...
- ...Staff Machine Learning Engineer About Sprinter Health At... ...the healthcare system, driving over $3... ...first dedicated ML engineering hire and build the production systems that train... ...training and inference pipelines, serving... ...or meaningfully scaled ML infrastructure...Full timeTemporary workWork at officeRelocation packageMonday to FridayMonday to ThursdayFlexible hours
$264.8k - $331k
Scale AI is the data foundation for AI, helping... ...and deploy reliable production AI applications. We partner... ..., production-grade systems that drive real... ...the RoleAs a Senior/Staff Machine Learning Engineer (MLE) on the General... ...Experience deploying ML systems in cloud environments...Full time- ...Inc. is seeking an AI-focused software engineer to design and deploy AI-powered solutions... ...fast-paced environment. You will work with product, operations, and engineering to ship... ...turning data and workflows into scalable systems with attention to latency, cost, and reliability...
- ...inspection company on the search for a Staff Systems Engineer. Our client helps manufacturers make... ...safer, and more reliable micro- and nano-scale materials, the building blocks of... ...and power systems. As demand surges and production grows more complex and faster, small undetected...Full time
- General Motors is seeking a Staff Machine Learning Engineer to lead end-to-end ML and CV pipelines for automated map reconstruction... .... You will architect scalable systems and collaborate across Mapping,... ..., robust design, and national-scale deployments, with opportunities...Remote job
- Sail is hiring for an engineering role in San Francisco to design and implement high-performance... ...fleet. You will also build LV routing systems to dispatch workloads with latency... ...caching for memory/compute trade-offs in LLM inference stacks. You will contribute to deep...
$179k - $218k
...who believe in the scale of our ambition and... ...are seeking a Senior Staff Data Center Operations Engineer, GPU Hardware Architecture... ...: Leverage AI/ML methodologies to analyze... ...failures in the production environment. Lead Root... ...Analysis (RCA) on systemic issues that span the...Temporary work- Jaide Health is seeking an engineer for their Model Efficiency team in San Francisco. The role focuses on building reliable ML systems while enhancing core performance metrics across model... ...or Python and insights into the LLM inference ecosystem. A commitment to diversity...Remote job
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff Engineer, Inference & RL Systems — Scale Production ML. Be the first to apply!
- software engineer staff San Francisco, CA
- assistant engineer San Francisco, CA
- engineering aide San Francisco, CA
- staff engineer San Francisco, CA
- staff security engineer San Francisco, CA
- assistant mechanical engineer San Francisco, CA
- assistant engineering manager San Francisco, CA
- senior staff systems engineer San Francisco, CA
- technology administrator San Francisco, CA
- project engineer assistant project manager San Francisco, CA




