Senior Inference & RL Systems Engineer (Scalable ML Infra)
Magic AI, Inc
Magic AI, Inc. is seeking a Member of Technical Staff to design and operate distributed systems for serving models in production and driving large-scale post-training workflows. You will work where model execution meets distributed infrastructure, influencing latency, throughput, and reliability of RL and training loops. You will own the infrastructure enabling fast inference and scalable RL iteration, balancing KV-cache strategies, batching, and long-context workloads while collaborating with #J-18808-Ljbffr Magic AI, Inc
$250k
A Series A Funded start-up in California is seeking a Systems Engineer to design and optimize systems handling complex ML pipelines. The role involves building scalable infrastructure, developing CI/CD pipelines, and ensuring system performance. Key qualifications include...Senior- ...company in San Francisco is looking for a Senior Software Engineer to build scalable infrastructure for large‑scale... ...You will design distributed training systems and optimize GPU utilization while... ...have over 5 years of experience in ML infrastructure and a strong background...Senior
- ...a Member of Technical Staff to design and optimize inference systems. The role involves managing KV cache allocation and... ...components. Ideal candidates should have strong software engineering skills and experience with ML inference systems, particularly in Python and C++....Senior
- ...seeking experienced backend engineers to own the systems that serve our diffusion... ...infrastructure that handles billions of inference requests, optimizing for... ...at the intersection of ML systems and backend... ...responsibilities spanning scalable services, model serving, load...Senior
- ...AI in San Francisco is seeking a Staff ML Systems Engineer to design and prototype algorithms, architectures... ...for low-latency, high-throughput inference. You will implement changes in... ...latency and cost. You will also co-design RL and post-training pipelines, drive performance...Suggested
$225k
Dormont Manufacturing Co is looking for a Software Engineer on the Inference & RL Systems team in San Francisco. The role involves designing distributed systems, optimizing performance, and ensuring high reliability for RL and post-training workflows. The ideal candidate...- ...da Vinci surgical system and Ion—have transformed... ....We’re a team of engineers, clinicians, and... ...of PositionAs a Senior Systems GPU... ...research, SW/ HW/ ML engineering, regulatory... ...robust, validated and scalable medical device... ...real-time onboard inference—while serving as a...SeniorLocal areaWorldwideFlexible hours
$227.2k - $417k
...Role:As a Software Engineer on the ML Infrastructure... ...machine learning inference platforms. These... ...ML model serving systems that support Deep... ...and a mentor to senior engineers, fostering... ...Design and build scalable, high throughput,... ...efficiency of our infra. Lead large scale...Full timeTemporary workLocal areaFlexible hours- ...seeking a specialist to design and operate large-scale GPU infrastructure. This role requires expertise in deploying GPU systems for high-throughput inference and model performance optimization. The ideal candidate will have hands-on experience with modern inference...Senior
- MakerMaker in San Francisco is seeking a Senior ML systems engineer to build and operate production inference systems for large models. You will own performance, profiling, and optimizations to ensure high throughput and low latency in production. You will collaborate...Senior
- MakerMaker.AI is looking for a Senior Machine Learning Systems Engineer in San Francisco. In this role, you will build and operate production inference systems, optimizing for performance and reliability. The ideal candidate will have 3+ years of experience in production...Senior
- A leading AI research firm located in San Francisco is seeking a Senior ML Systems Engineer to build and maintain the training framework for large-scale language models. The role involves designing distributed training solutions and improving training throughput across...SeniorFlexible hours
- Senior ML Systems Engineer, Frameworks & Tooling at Cohere Our mission is to scale intelligence to serve... ...that enable fast, reliable, and scalable model training and build the tooling that... ...ergonomics. Collaborate closely with infra teams to ensure Slurm setups, container...SeniorFull timeWork at officeRemote workFlexible hours
- OpenAI in San Francisco is seeking an experienced systems generalist to build an automated inference optimization platform across hardware, compiler, and runtime... ...with research, infrastructure, security, product, and partnerships to deliver scalable, #J-18808-Ljbffr SlopeSenior
- ...entity. Responsibilities As a senior Machine Learning Systems Engineer on the Search Platform team, you... ...Platform EngineeringDesign and implement scalable search serving infrastructure,... ...search. Own end-to-end delivery of ML components from experimentation through...SeniorWork at officeLocal area
- Autodesk, Inc. in San Francisco seeks a Senior Principal AI/ML Developer to shape data-driven... ...customer lifecycle. You will design scalable ML pipelines, drive experimentation,... ...scientists and collaborate with product, engineering, and marketing teams to deploy robust...Senior
- Causal Labs in San Francisco is seeking an experienced Infrastructure Engineer to build high-throughput inference systems for large-scale evaluation and backtesting against historical physical observations. You will design techniques to improve latency and throughput, optimize...Senior
$124k - $280k
...people in data and analytics engineering focus on leveraging advanced... ...implementing advanced AI and ML solutions to drive innovation... ...optimising algorithms, models, and systems to enable intelligent... ...for LLM outputs- Developing scalable data storage solutions using...SeniorFull timeH1b- ...Parafin, Inc. is seeking a Software Engineer to lead the evolution of its ML Platform within the Infrastructure team. You will build scalable, reliable systems for model experimentation, training, evaluation, inference, and retraining powering underwriting and ML-driven...Senior
- Genesis AI in San Francisco is seeking a senior ML infrastructure engineer to design and optimize distributed training systems and performance-critical components. You will... ...‑node GPU clusters. Join a team focused on scalable AI foundations, monitoring tools, and robust...Senior
$194k - $266k
...San Francisco, CA About The Role AI inference is becoming core infrastructure. Every... ...one API change. We are looking for a Senior Systems Engineer to help build that layer. This is a deeply... ...Bonus Points Experience building AI/ML infrastructure, inference platforms,...SeniorTemporary workLocal areaFlexible hours- ...leading financial technology firm is seeking a Distinguished AI Engineer in San Francisco. In this role, you will collaborate with cross... ...teams to develop AI-powered products and contribute to scalable AI infrastructure. Ideal candidates have a strong background in...Senior
- ...Design, deploy, and maintain large distributed ML training and inference clusters Develop efficient, scalable end-to-end pipelines to manage petabyte-scale datasets... ...working on distributed task management systems and scalable model serving & deployment architectures...Senior
$250k
...building a serverless inference platform, beginning... ...chance to join as a Senior Inference Platform Engineer at an early stage... ...define the architecture, scalability, and technical... ...distributed inference systems to maximise GPU utilisation... ...systems (ML inference, HPC, or similar...SeniorFull time- Crusoe Energy Systems LLC in San Francisco seeks a Senior Systems Engineer to design and build agentic AI systems that advance... ...integration platforms to create scalable agent workflows and a citizen... ...engineering with 3+ years in AI/ML, strong Python/API skills, and experience...Senior
$215k - $275k
...ecosystem of libraries for scalable machine learning.... ...can scale an ML application from their... ...to be a distributed systems expert.Proud to be backed... ...Anyscale is looking for a Senior Site Reliability Engineer to join the... ...laptop. As part of the Infra team, we build the scalable...SeniorWork at office$250k - $280k
...Staff / Principal Founding Engineer (Backend-Leaning) – AI Systems Platform San Francisco... ...TypeScript, APIs, AWS, cloud infra) Background in 0→1 or... ...environments Experience building scalable, production-grade systems... ...agent frameworks, or data/ML pipelines, come from top 5...SeniorWork at officeImmediate startFlexible hours$200k - $240k
...secure world for all. The AI Engineering Team is chartered with... ...Models (LLMs) and agentic systems. Our mission is to build robust... ...faster than the market. As a Senior or Staff ML Systems Engineer - LLM ,... ...environments. Build out a modular and scalable AI infrastructure stack —...SeniorRemote workWorldwide$250k
...experimentation, and inference at scale. The... ...company is looking for a Senior / Staff Site Reliability Engineer to support and... ...closely with platform, ML, and infrastructure... ...growth and scalability. Don’t miss out... ...available infrastructure systems Improve CI/CD pipelines...SeniorFull timeRemote work$220k
...Perplexity is looking for an engineer to join their team in San Francisco. You will work on building and operating the inference engine, supporting new models, migrating GPU kernels... ...in software engineering with a focus on ML inference, familiarity with deep learning...Senior
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Inference & RL Systems Engineer (Scalable ML Infra). Be the first to apply!
- senior windows systems engineer San Francisco, CA
- software system engineer San Francisco, CA
- system test engineer San Francisco, CA
- mission system engineer San Francisco, CA
- healthcare systems engineer San Francisco, CA
- electronic systems engineer San Francisco, CA
- operating system engineer San Francisco, CA
- system engineer remote San Francisco, CA
- application system engineer San Francisco, CA
- advanced systems engineer San Francisco, CA



