Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Inference & RL Systems Engineer (Scalable ML Infra)

Magic AI, Inc

Magic AI, Inc. is seeking a Member of Technical Staff to design and operate distributed systems for serving models in production and driving large-scale post-training workflows. You will work where model execution meets distributed infrastructure, influencing latency, throughput, and reliability of RL and training loops. You will own the infrastructure enabling fast inference and scalable RL iteration, balancing KV-cache strategies, batching, and long-context workloads while collaborating with #J-18808-Ljbffr Magic AI, Inc

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Senior Inference & RL Systems Engineer (Scalable ML Infra) in San Francisco, CA vacancy
  • $250k

    A Series A Funded start-up in California is seeking a Systems Engineer to design and optimize systems handling complex ML pipelines. The role involves building scalable infrastructure, developing CI/CD pipelines, and ensuring system performance. Key qualifications include... 
    Senior

    Acceler8 Talent

    San Francisco, CA
    2 hours ago
  •  ...company in San Francisco is looking for a Senior Software Engineer to build scalable infrastructure for large‑scale...  ...You will design distributed training systems and optimize GPU utilization while...  ...have over 5 years of experience in ML infrastructure and a strong background... 
    Senior

    Baseten

    San Francisco, CA
    2 hours ago
  •  ...a Member of Technical Staff to design and optimize inference systems. The role involves managing KV cache allocation and...  ...components. Ideal candidates should have strong software engineering skills and experience with ML inference systems, particularly in Python and C++.... 
    Senior

    Gimlet Labs

    San Francisco, CA
    1 day ago
  •  ...seeking experienced backend engineers to own the systems that serve our diffusion...  ...infrastructure that handles billions of inference requests, optimizing for...  ...at the intersection of ML systems and backend...  ...responsibilities spanning scalable services, model serving, load... 
    Senior

    Inception

    San Francisco, CA
    3 days ago
  •  ...AI in San Francisco is seeking a Staff ML Systems Engineer to design and prototype algorithms, architectures...  ...for low-latency, high-throughput inference. You will implement changes in...  ...latency and cost. You will also co-design RL and post-training pipelines, drive performance... 
    Suggested

    Together

    San Francisco, CA
    2 days ago
  • $225k

    Dormont Manufacturing Co is looking for a Software Engineer on the Inference & RL Systems team in San Francisco. The role involves designing distributed systems, optimizing performance, and ensuring high reliability for RL and post-training workflows. The ideal candidate... 

    Dormont Manufacturing Co

    San Francisco, CA
    2 days ago
  •  ...da Vinci surgical system and Ion—have transformed...  ....We’re a team of engineers, clinicians, and...  ...of PositionAs a Senior Systems GPU...  ...research, SW/ HW/ ML engineering, regulatory...  ...robust, validated and scalable medical device...  ...real-time onboard inference—while serving as a... 
    Senior
    Local area
    Worldwide
    Flexible hours

    Intuitive Surgical

    San Francisco, CA
    1 day ago
  • $227.2k - $417k

     ...Role:As a Software Engineer on the ML Infrastructure...  ...machine learning inference platforms. These...  ...ML model serving systems that support Deep...  ...and a mentor to senior engineers, fostering...  ...Design and build scalable, high throughput,...  ...efficiency of our infra. Lead large scale... 
    Full time
    Temporary work
    Local area
    Flexible hours

    Tubi TV

    San Francisco, CA
    1 day ago
  •  ...seeking a specialist to design and operate large-scale GPU infrastructure. This role requires expertise in deploying GPU systems for high-throughput inference and model performance optimization. The ideal candidate will have hands-on experience with modern inference... 
    Senior

    Reflection AI

    San Francisco, CA
    3 days ago
  • MakerMaker in San Francisco is seeking a Senior ML systems engineer to build and operate production inference systems for large models. You will own performance, profiling, and optimizations to ensure high throughput and low latency in production. You will collaborate... 
    Senior

    MakerMaker

    San Francisco, CA
    1 day ago
  • MakerMaker.AI is looking for a Senior Machine Learning Systems Engineer in San Francisco. In this role, you will build and operate production inference systems, optimizing for performance and reliability. The ideal candidate will have 3+ years of experience in production... 
    Senior

    MakerMaker.AI

    San Francisco, CA
    5 days ago
  • A leading AI research firm located in San Francisco is seeking a Senior ML Systems Engineer to build and maintain the training framework for large-scale language models. The role involves designing distributed training solutions and improving training throughput across... 
    Senior
    Flexible hours

    Cohere

    San Francisco, CA
    1 day ago
  • Senior ML Systems Engineer, Frameworks & Tooling at Cohere Our mission is to scale intelligence to serve...  ...that enable fast, reliable, and scalable model training and build the tooling that...  ...ergonomics. Collaborate closely with infra teams to ensure Slurm setups, container... 
    Senior
    Full time
    Work at office
    Remote work
    Flexible hours

    Cohere

    San Francisco, CA
    1 day ago
  • OpenAI in San Francisco is seeking an experienced systems generalist to build an automated inference optimization platform across hardware, compiler, and runtime...  ...with research, infrastructure, security, product, and partnerships to deliver scalable, #J-18808-Ljbffr Slope
    Senior

    Slope

    San Francisco, CA
    5 days ago
  •  ...entity. Responsibilities As a senior Machine Learning Systems Engineer on the Search Platform team, you...  ...Platform EngineeringDesign and implement scalable search serving infrastructure,...  ...search. Own end-to-end delivery of ML components from experimentation through... 
    Senior
    Work at office
    Local area

    Atlassian

    San Francisco, CA
    2 days ago
  • Autodesk, Inc. in San Francisco seeks a Senior Principal AI/ML Developer to shape data-driven...  ...customer lifecycle. You will design scalable ML pipelines, drive experimentation,...  ...scientists and collaborate with product, engineering, and marketing teams to deploy robust... 
    Senior

    Autodesk

    San Francisco, CA
    1 day ago
  • Causal Labs in San Francisco is seeking an experienced Infrastructure Engineer to build high-throughput inference systems for large-scale evaluation and backtesting against historical physical observations. You will design techniques to improve latency and throughput, optimize... 
    Senior

    Causal Labs

    San Francisco, CA
    1 day ago
  • $124k - $280k

     ...people in data and analytics engineering focus on leveraging advanced...  ...implementing advanced AI and ML solutions to drive innovation...  ...optimising algorithms, models, and systems to enable intelligent...  ...for LLM outputs- Developing scalable data storage solutions using... 
    Senior
    Full time
    H1b

    PwC

    San Francisco, CA
    21 hours ago
  •  ...Parafin, Inc. is seeking a Software Engineer to lead the evolution of its ML Platform within the Infrastructure team. You will build scalable, reliable systems for model experimentation, training, evaluation, inference, and retraining powering underwriting and ML-driven... 
    Senior

    Parafin, Inc.

    San Francisco, CA
    2 hours ago
  • Genesis AI in San Francisco is seeking a senior ML infrastructure engineer to design and optimize distributed training systems and performance-critical components. You will...  ...‑node GPU clusters. Join a team focused on scalable AI foundations, monitoring tools, and robust... 
    Senior

    Genesis AI

    San Francisco, CA
    4 days ago
  • $194k - $266k

     ...San Francisco, CA About The Role AI inference is becoming core infrastructure. Every...  ...one API change. We are looking for a Senior Systems Engineer to help build that layer. This is a deeply...  ...Bonus Points Experience building AI/ML infrastructure, inference platforms,... 
    Senior
    Temporary work
    Local area
    Flexible hours

    Cloudflare

    San Francisco, CA
    4 days ago
  •  ...leading financial technology firm is seeking a Distinguished AI Engineer in San Francisco. In this role, you will collaborate with cross...  ...teams to develop AI-powered products and contribute to scalable AI infrastructure. Ideal candidates have a strong background in... 
    Senior

    Capital One

    San Francisco, CA
    2 hours ago
  •  ...Design, deploy, and maintain large distributed ML training and inference clusters Develop efficient, scalable end-to-end pipelines to manage petabyte-scale datasets...  ...working on distributed task management systems and scalable model serving & deployment architectures... 
    Senior

    Kindredventures

    San Francisco, CA
    5 days ago
  • $250k

     ...building a serverless inference platform, beginning...  ...chance to join as a Senior Inference Platform Engineer at an early stage...  ...define the architecture, scalability, and technical...  ...distributed inference systems to maximise GPU utilisation...  ...systems (ML inference, HPC, or similar... 
    Senior
    Full time
    San Francisco, CA
    more than 2 months ago
  • Crusoe Energy Systems LLC in San Francisco seeks a Senior Systems Engineer to design and build agentic AI systems that advance...  ...integration platforms to create scalable agent workflows and a citizen...  ...engineering with 3+ years in AI/ML, strong Python/API skills, and experience... 
    Senior

    Crusoe Energy Systems LLC

    San Francisco, CA
    1 day ago
  • $215k - $275k

     ...ecosystem of libraries for scalable machine learning....  ...can scale an ML application from their...  ...to be a distributed systems expert.Proud to be backed...  ...Anyscale is looking for a Senior Site Reliability Engineer to join the...  ...laptop. As part of the Infra team, we build the scalable... 
    Senior
    Work at office

    Anyscale

    San Francisco, CA
    3 days ago
  • $250k - $280k

     ...Staff / Principal Founding Engineer (Backend-Leaning) – AI Systems Platform San Francisco...  ...TypeScript, APIs, AWS, cloud infra) Background in 0→1 or...  ...environments Experience building scalable, production-grade systems...  ...agent frameworks, or data/ML pipelines, come from top 5... 
    Senior
    Work at office
    Immediate start
    Flexible hours

    Xpertalent

    San Francisco, CA
    more than 2 months ago
  • $200k - $240k

     ...secure world for all. The AI Engineering Team is chartered with...  ...Models (LLMs) and agentic systems. Our mission is to build robust...  ...faster than the market. As a Senior or Staff ML Systems Engineer - LLM ,...  ...environments. Build out a modular and scalable AI infrastructure stack —... 
    Senior
    Remote work
    Worldwide

    TRM Labs

    San Francisco, CA
    5 days ago
  • $250k

     ...experimentation, and inference at scale. The...  ...company is looking for a Senior / Staff Site Reliability Engineer to support and...  ...closely with platform, ML, and infrastructure...  ...growth and scalability. Don’t miss out...  ...available infrastructure systems Improve CI/CD pipelines... 
    Senior
    Full time
    Remote work
    San Francisco, CA
    more than 2 months ago
  • $220k

     ...Perplexity is looking for an engineer to join their team in San Francisco. You will work on building and operating the inference engine, supporting new models, migrating GPU kernels...  ...in software engineering with a focus on ML inference, familiarity with deep learning... 
    Senior

    Perplexity

    San Francisco, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Inference & RL Systems Engineer (Scalable ML Infra). Be the first to apply!