Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff Engineer, GPU AI Inference & RL Infrastructure

B Capital

B Capital is seeking a skilled engineer for GPU infrastructure in San Francisco. This role involves designing and operating high-performance systems for model inference, synthetic data generation, and reinforcement learning. The ideal candidate has strong GPU systems experience, optimization skills, and a passion for working in cutting-edge AI. Benefits include top-tier compensation, comprehensive health insurance, and paid parental leave. #J-18808-Ljbffr B Capital

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Staff Engineer, GPU AI Inference & RL Infrastructure in San Francisco, CA vacancy
  •  ...intelligence open and accessible. We design, build, and operate GPU-heavy infrastructure for high-throughput model inference and mid-training workloads. Join a team that powers synthetic data generation, RL pipelines, and distributed model evaluation across thousands of... 
    Suggested

    Visa Hunt

    San Francisco, CA
    2 days ago
  •  ...Barnes Associates Limited is seeking a Staff-level engineer to architect and evolve the AI infrastructure software stack for large-scale GPU workloads. You’ll drive orchestration,...  ...production environment. You’ll contribute to inference platforms, model serving, and high-... 
    Suggested

    Hamilton Barnes Associates Limited

    San Francisco, CA
    4 days ago
  • $200k - $400k

    A leading AI technology company located in San Francisco is seeking an infrastructure engineer to build distributed systems for their AI inference engine. The role involves designing systems that ensure minimal latency and maximum reliability. Candidates should have a... 
    Suggested
    Visa sponsorship

    Inferact

    San Francisco, CA
    2 days ago
  • Magic AI, Inc. is seeking a engineer for the Supercomputing Platform & Infrastructure to design, build, and operate large-scale GPU infrastructure powering model training and inference. You will implement Terraform-driven IaC across cloud and hybrid environments, manage... 
    Suggested
    Visa sponsorship
    Relocation package

    Magic AI Corp.

    San Francisco, CA
    4 days ago
  •  ...Company in San Francisco is seeking a Member of Technical Staff for their infrastructure team. In this role, you will own the cloud systems that...  ...API and build global low-latency, high-throughput GPU ML inference infrastructure. The ideal candidate will have solid experience... 
    Suggested
    Visa sponsorship

    The Token Company

    San Francisco, CA
    13 hours ago
  • Kindredventures is recruiting infrastructure engineers to scale large-scale inference and evaluation around a physics-based LPM initiative. You will work on high...  ...record in building efficient inference stacks, GPU-aware optimization, and deep learning frameworks like... 

    Kindredventures

    San Francisco, CA
    1 day ago
  •  ...hiring Members of Technical Staff to build systems that accelerate LLM inference and own customer workloads...  ...performance kernels, inference engine internals, and production infrastructure for named accounts. You...  ...collaborating with a fast-growing AI inference company. #J-1880... 

    Simplify

    San Francisco, CA
    4 days ago
  • $200k - $225k

     ...Recruiting Researchers & Engineers with Foundational AI Experience A fast-growing...  ...AI company is seeking a Staff Cloud Infrastructure Engineer to build and scale high-performance GPU compute infrastructure. You...  ...performance tools for AI inference Collaborating with... 
    Full time

    Lumicity

    San Francisco, CA
    3 days ago
  • $220k

    We build and run the inference engine behind every Perplexity query and deploy dozens of model...  ...and multimodal models in our inference infrastructure, from weight loading, request scheduling...  ...management to support in API Gateway. GPU kernels migration to CuTe DSL. Port our... 

    Perplexity

    San Francisco, CA
    1 day ago
  • San Francisco Tensor Company is hiring a Member of Technical Staff for Sandbox Infrastructure to build a serverless GPU container service across NVIDIA, AMD, TPU and Trainium, scalable and isolated for untrusted code. You will work with compiler, post‑training and kernel... 
    Relocation

    San Francisco Tensor Company

    San Francisco, CA
    1 day ago
  • Inception is seeking engineers and scientists to design, optimize, and maintain core systems enabling scalable reinforcement learning...  ...production-readiness. Responsibilities include building infrastructure for RL workloads, boosting training throughput, and creating... 

    Inception

    San Francisco, CA
    1 day ago
  • Causal is building a Large Physics foundation Model and seeks an infrastructure engineer to design, deploy, and operate large distributed GPU clusters. You will extend schedulers, build self-serve interfaces, and own storage and lineage for checkpoints and logs. You will... 

    causal

    San Francisco, CA
    1 day ago
  • Vmax, an applied research lab, seeks an infrastructure engineer to build the systems layer for large-scale RL. You will enable researchers and ML engineers to run, debug, and reproduce experiments across thousands of GPUs, with a focus on reliability and scalability. The... 

    Vmax

    San Francisco, CA
    1 day ago
  • $350k

     ...opportunities? Join a rapidly growing AI infrastructure provider delivering large-...  ...solutions for AI training and inference across global cloud and GPU environments. The organization...  ...This opportunity is for a Staff Site Reliability Engineer to lead the reliability of... 
    Full time
    San Francisco, CA
    a month ago
  • $190.9k - $232.8k

    A leading data and AI company is seeking a Staff Software Engineer for GenAI inference to lead the architecture and optimization of the inference engine. The role requires expertise in CUDA, GPU programming, and distributed systems design. Ideal candidates will have a strong... 

    Jobleads-US

    San Francisco, CA
    4 days ago
  • $200k - $250k

     ...anonymous venture-backed AI startup building...  ...learning environments and infrastructure for frontier AI...  ...platform and full-stack engineers from our network. If there...  ...demand for high-quality RL environments grows....  ...scalable training and inference infrastructure for RL... 
    Full time
    H1b
    Work at office
    Relocation package

    CoffeeSpace

    San Francisco, CA
    1 day ago
  • $275k - $315k

    SF Tensor is building the fastest GPU compiler and model foundry to enable AI across clouds and chips. We’re hiring a Member of Technical Staff to own the modeling side from data curation...  ...distillation to deployment, working with SFT, RL, and DPO to ship post-trained models in... 
    Relocation package

    SF Tensor

    San Francisco, CA
    3 days ago
  • $225k

    Dormont Manufacturing Co is looking for a Software Engineer on the Inference & RL Systems team in San Francisco. The role involves designing distributed systems, optimizing performance, and ensuring high reliability for RL and post-training workflows. The ideal candidate... 

    Dormont Manufacturing Co

    San Francisco, CA
    13 hours ago
  •  ...building the world's most efficient software for inference and agent hosting. In this role, you'll be one of the first engineers on Sailboxes, contributing across the stack...  ...You'll focus on distributed systems, cloud infrastructure, and scalable, ergonomic tooling. The SF... 
    Work at office

    Sail Research Inc.

    San Francisco, CA
    3 days ago
  • $215k - $260k

     ...intelligence. As the only vertically integrated AI infrastructure company built from the ground up, we...  ...in production. That means owning the inference stack end to end: profiling where time...  ...will also work directly with customer engineering teams to tailor deployments to their... 
    Temporary work

    Crusoe

    San Francisco, CA
    2 days ago
  • $179k - $218k

     ...and intelligence. As the only vertically integrated AI infrastructure company built from the ground up, we own and operate...  ...Silicon Reality" must be bridged.We are seeking a Senior Staff Data Center Operations Engineer, GPU Hardware Architecture to be the definitive technical... 
    Temporary work

    Crusoe

    San Francisco, CA
    1 day ago
  •  ...San Francisco is seeking a Member of Technical Staff to lead distributed systems at the core of our AI infrastructure. You’ll design, build, and operate scheduling,...  ...GPUs, and accelerators. You’ll work closely with inference, runtimes, compilers, kernels, and hardware... 

    Acceler8 Talent

    San Francisco, CA
    13 hours ago
  • Together AI is building the best inference infrastructure for voice applications. We seek a Staff ML Engineer to own the model serving stack and optimize latency and throughput for real-time voice workloads. You'll work with state-of-the-art accelerators and collaborate... 

    Together

    San Francisco, CA
    1 day ago
  • $220k - $280k

     ...About the Role Together AI is building the best inference infrastructure for voice applications. Our Voice AI...  ...reliability. We're looking for a Staff ML Engineer to drive the model serving layer...  ...to the frontier. You'll profile GPU utilization, design batching strategies... 
    Full time

    Together Ai

    San Francisco, CA
    13 hours ago
  • $207k - $290k

     ...Description About JazzX AI:   Vision:...  ...experienced AI Engineer with deep...  ...Learning (RL) to join our team as a Senior Staff Architect. In this...  ...ensuring the RL infrastructure can scale to support...  ..., including inference-time search,...  ...infrastructure (Kubernetes, GPU/TPU clusters,... 
    Worldwide
    Flexible hours

    JazzX AI

    San Francisco, CA
    4 days ago
  • $190.2k - $345.65k

     ...multimedia generative AI — deep-tuned image,...  ...are hiring a Senior Staff Machine Learning Engineer to architect and...  ...indexing, and search infrastructure behind Firefly Foundry...  ..., ANN index tuning, GPU-accelerated enrichment...  ...models and the inference paths that produce them... 
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Francisco, CA
    4 days ago
  •  ...the next generation of AI-driven game...  ...Senior Machine Learning Engineer for On-Device & Mobile...  ...significant parts of the inference stack — from a trained...  ...tuning across NPU, mobile GPU, and desktop/laptop GPU...  ...on-device benchmarking infrastructure, performance-... 
    Full time
    Work at office
    Remote work
    Worldwide

    Unity Technologies

    San Francisco, CA
    13 hours ago
  • $286.2k - $326.7k

     ...Overview Senior Staff Engineer, AI Compute (Remote Eligible) At Capital...  ...investments in technology infrastructure and world-class talent —...  ...infrastructure on top of CPU and GPU substrates. Your...  .../ DL model training, model inference and feature generation pipelines... 
    Full time
    Part time
    Local area
    Remote work

    Capital One

    San Francisco, CA
    6 days ago
  • $197.3k - $313.7k

     ...SalesforceSalesforce is the #1 AI CRM, where humans...  ...is looking for a Staff Machine Learning Engineer with deep expertise...  ...curation, training infrastructure, hyperparameter...  ...training pipelines on GPU infrastructure.Brainstorm...  ...optimization for inference (quantization,... 
    Full time

    Salesforce

    San Francisco, CA
    3 days ago
  • Sail builds the world’s most efficient software for inference and agent hosting. In this role, you’ll own token processing down to the lowest...  ...design and implement exotic parallelism schemes, write custom GPU kernels for regimes like cascade attention, and understand every... 

    Sail

    San Francisco, CA
    13 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff Engineer, GPU AI Inference & RL Infrastructure. Be the first to apply!