Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff Engineer, GPU AI Inference & RL Infrastructure

B Capital

B Capital is seeking a skilled engineer for GPU infrastructure in San Francisco. This role involves designing and operating high-performance systems for model inference, synthetic data generation, and reinforcement learning. The ideal candidate has strong GPU systems experience, optimization skills, and a passion for working in cutting-edge AI. Benefits include top-tier compensation, comprehensive health insurance, and paid parental leave. #J-18808-Ljbffr B Capital

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Staff Engineer, GPU AI Inference & RL Infrastructure in San Francisco, CA vacancy
  •  ...Barnes Associates Limited is seeking a Staff-level engineer to architect and evolve the AI infrastructure software stack for large-scale GPU workloads. You’ll drive orchestration,...  ...production environment. You’ll contribute to inference platforms, model serving, and high-... 
    Suggested

    Hamilton Barnes Associates Limited

    San Francisco, CA
    2 days ago
  • Together AI in San Francisco is seeking a Staff ML Systems Engineer to design and prototype algorithms, architectures...  ..., high-throughput inference. You will implement...  ...while profiling across GPU, networking, and memory...  ...You will also co-design RL and post-training pipelines... 
    Suggested

    Together

    San Francisco, CA
    4 days ago
  • United States Digital Space LLC is seeking an infrastructure leader to own a self-serve GPU compute platform for training and inference workloads. You will design and operate the...  ...and a coherent platform roadmap for scalable AI workloads. #J-18808-Ljbffr United States... 
    Suggested

    United States Digital Space LLC

    San Francisco, CA
    5 days ago
  • $200k - $400k

    A leading AI technology company located in San Francisco is seeking an infrastructure engineer to build distributed systems for their AI inference engine. The role involves designing systems that ensure minimal latency and maximum reliability. Candidates should have a... 
    Suggested
    Visa sponsorship

    Inferact

    San Francisco, CA
    1 day ago
  •  ...interactive world models for robotics and embodied AI. We are hiring a Member of Technical Staff to lead reinforcement learning infrastructure and model post-training. You will work with...  ...-generating agents through fine-tuning, RL, and scalable evaluation in a on-site San... 
    Suggested

    Doist

    San Francisco, CA
    5 days ago
  •  ...Company in San Francisco is seeking a Member of Technical Staff for their infrastructure team. In this role, you will own the cloud systems that...  ...API and build global low-latency, high-throughput GPU ML inference infrastructure. The ideal candidate will have solid experience... 
    Visa sponsorship

    The Token Company

    San Francisco, CA
    4 days ago
  • Wafer is building AI-powered GPU optimization systems and is seeking engineers to join a small, highly collaborative team. You will work directly with the founders to implement the agent framework, profiling, and compiler tooling that power our GPU optimization platform... 

    Wafer

    San Francisco, CA
    2 days ago
  • Kindredventures is recruiting infrastructure engineers to scale large-scale inference and evaluation around a physics-based LPM initiative. You will work on high...  ...record in building efficient inference stacks, GPU-aware optimization, and deep learning frameworks like... 

    Kindredventures

    San Francisco, CA
    5 days ago
  •  ...hiring Members of Technical Staff to build systems that accelerate LLM inference and own customer workloads...  ...performance kernels, inference engine internals, and production infrastructure for named accounts. You...  ...collaborating with a fast-growing AI inference company. #J-1880... 

    Simplify

    San Francisco, CA
    3 days ago
  • $350k

     ...opportunities? Join a rapidly growing AI infrastructure provider delivering large-...  ...solutions for AI training and inference across global cloud and GPU environments. The organization...  ...This opportunity is for a Staff Site Reliability Engineer to lead the reliability of... 
    Full time
    San Francisco, CA
    a month ago
  • $200k - $225k

     ...Recruiting Researchers & Engineers with Foundational AI Experience A fast-growing...  ...AI company is seeking a Staff Cloud Infrastructure Engineer to build and scale high-performance GPU compute infrastructure. You...  ...performance tools for AI inference Collaborating with... 
    Full time

    Lumicity

    San Francisco, CA
    1 day ago
  • $220k

    We build and run the inference engine behind every Perplexity query and deploy dozens of model...  ...and multimodal models in our inference infrastructure, from weight loading, request scheduling...  ...management to support in API Gateway. GPU kernels migration to CuTe DSL. Port our... 

    Perplexity

    San Francisco, CA
    5 days ago
  • Causal is building a Large Physics foundation Model and seeks an infrastructure engineer to design, deploy, and operate large distributed GPU clusters. You will extend schedulers, build self-serve interfaces, and own storage and lineage for checkpoints and logs. You will... 

    causal

    San Francisco, CA
    5 days ago
  • $190.9k - $232.8k

    A leading data and AI company is seeking a Staff Software Engineer for GenAI inference to lead the architecture and optimization of the inference engine. The role requires expertise in CUDA, GPU programming, and distributed systems design. Ideal candidates will have a strong... 

    Jobleads-US

    San Francisco, CA
    2 days ago
  • $225k

    Dormont Manufacturing Co is looking for a Software Engineer on the Inference & RL Systems team in San Francisco. The role involves designing distributed systems, optimizing performance, and ensuring high reliability for RL and post-training workflows. The ideal candidate... 

    Dormont Manufacturing Co

    San Francisco, CA
    4 days ago
  • $179k - $218k

     ...and intelligence. As the only vertically integrated AI infrastructure company built from the ground up, we own and operate...  ...Silicon Reality" must be bridged.We are seeking a Senior Staff Data Center Operations Engineer, GPU Hardware Architecture to be the definitive technical... 
    Temporary work

    Crusoe

    San Francisco, CA
    5 days ago
  • $250k - $300k

     ...intelligence. As the only vertically integrated AI infrastructure company built from the ground up, we...  ...in production. That means owning the inference stack end to end: profiling where time...  ...will also work directly with customer engineering teams to tailor deployments to their... 
    Temporary work

    Crusoe

    San Francisco, CA
    1 day ago
  • $207k - $290k

     ...Description About JazzX AI: Vision:...  ...experienced AI Engineer with deep...  ...Learning (RL) to join our team as a Senior Staff Architect. In this...  ...ensuring the RL infrastructure can scale to...  ...techniques , including inference-time search,...  ...(Kubernetes, GPU/TPU clusters,... 
    Worldwide
    Flexible hours

    JazzX AI

    San Francisco, CA
    more than 2 months ago
  • $160k - $300k

     ...Apiphany is a pioneering foundational AI company for physical product development...  ...Character, our mission is to revolutionize how engineering decisions are made, turning complexity...  .... About the Role As a Senior / Staff Infrastructure Engineer at Apiphany, you’ll design,... 
    Work at office
    Visa sponsorship
    Flexible hours

    Apiphany

    San Francisco, CA
    3 days ago
  • $220k - $280k

     ...About the Role Together AI is building the best inference infrastructure for voice applications. Our Voice AI...  ...reliability. We're looking for a Staff ML Engineer to drive the model serving layer...  ...to the frontier. You'll profile GPU utilization, design batching strategies... 
    Full time

    Together Ai

    San Francisco, CA
    1 day ago
  • $197.3k - $313.7k

     ...SalesforceSalesforce is the #1 AI CRM, where humans...  ...is looking for a Staff Machine Learning Engineer with deep expertise...  ...curation, training infrastructure, hyperparameter...  ...training pipelines on GPU infrastructure.Brainstorm...  ...optimization for inference (quantization,... 
    Full time

    Salesforce

    San Francisco, CA
    2 days ago
  •  ...the next generation of AI-driven game...  ...Senior Machine Learning Engineer for On-Device & Mobile...  ...significant parts of the inference stack — from a trained...  ...tuning across NPU, mobile GPU, and desktop/laptop GPU...  ...on-device benchmarking infrastructure, performance-... 
    Full time
    Work at office
    Remote work
    Worldwide

    Unity Technologies

    San Francisco, CA
    4 days ago
  • $307k - $352k

     ...About the Role Kikoff's infrastructure team builds the systems that enable engineering teams to move quickly without sacrificing...  ...Data Infrastructure. As a Staff Infrastructure Engineer, you will...  ...infrastructure boundaries. AI is increasing the amount of code... 
    Permanent employment
    Full time
    Local area

    Kikoff

    San Francisco, CA
    3 days ago
  • Sail builds the world’s most efficient software for inference and agent hosting. In this role, you’ll own token processing down to the lowest...  ...design and implement exotic parallelism schemes, write custom GPU kernels for regimes like cascade attention, and understand every... 

    SAIL

    San Francisco, CA
    4 days ago
  •  ...revolutionary AI‑powered enterprise...  ...experienced AI Engineer with deep...  ...Reinforcement Learning (RL) to join our team as a Senior Staff Architect. In...  ...Scalability & Infrastructure: Design...  ...techniques , including inference‑time search,...  ...infrastructure (Kubernetes, GPU/TPU clusters,... 
    Flexible hours

    JazzX AI

    San Francisco, CA
    3 days ago
  • A cutting-edge AI research firm in San Francisco is seeking talent to build and optimize GPU infrastructure for large-scale model inference and training workloads. The ideal candidate will have hands-on experience with GPU systems and optimization techniques, actively... 

    Reflection

    San Francisco, CA
    3 days ago
  •  ...building a Large Physics foundation Model and seeks an infrastructure engineer to design, deploy, and operate its GPU-driven compute environment. You will enable...  ...optimizing distributed clusters that power training and inference workloads. You will extend orchestration,... 

    Kindredventures

    San Francisco, CA
    5 days ago
  • $350k

    Mirendil in San Francisco is searching for an engineer to develop and optimize inference systems for cutting-edge AI models. You will handle the complete inference stack, enhancing performance and reliability. The role involves partnering with teams to deploy new architectures... 

    Mirendil

    San Francisco, CA
    5 days ago
  •  ...Francisco is seeking a Member of Technical Staff focused on kernels and GPU performance. This role involves...  ...optimizing GPU and accelerator kernels for AI workloads by analyzing performance...  ...candidates have strong software engineering foundations and experience with performance... 

    Gimlet Labs

    San Francisco, CA
    3 days ago
  •  ...Job Description Staff Machine Learning Engineer, Artificial Intelligence (AI) Required, Work From Home...  ...of research, infrastructure, and product. You are...  ..., evaluation systems, inference architecture, and deployment...  ...deployment, including GPU optimization, memory efficiency... 
    Remote work
    Work from home

    Ginas Tech Jobs

    San Francisco, CA
    28 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff Engineer, GPU AI Inference & RL Infrastructure. Be the first to apply!