Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff Engineer, GPU AI Inference & RL Infrastructure

B Capital

B Capital is seeking a skilled engineer for GPU infrastructure in San Francisco. This role involves designing and operating high-performance systems for model inference, synthetic data generation, and reinforcement learning. The ideal candidate has strong GPU systems experience, optimization skills, and a passion for working in cutting-edge AI. Benefits include top-tier compensation, comprehensive health insurance, and paid parental leave. #J-18808-Ljbffr B Capital

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Staff Engineer, GPU AI Inference & RL Infrastructure in San Francisco, CA vacancy
  •  ...intelligence open and accessible. We design, build, and operate GPU-heavy infrastructure for high-throughput model inference and mid-training workloads. Join a team that powers synthetic data generation, RL pipelines, and distributed model evaluation across thousands of... 
    Suggested

    Visa Hunt

    San Francisco, CA
    2 days ago
  •  ...Barnes Associates Limited is seeking a Staff-level engineer to architect and evolve the AI infrastructure software stack for large-scale GPU workloads. You’ll drive orchestration,...  ...production environment. You’ll contribute to inference platforms, model serving, and high-... 
    Suggested

    Hamilton Barnes Associates Limited

    San Francisco, CA
    3 days ago
  • United States Digital Space LLC is seeking an infrastructure leader to own a self-serve GPU compute platform for training and inference workloads. You will design and operate the...  ...and a coherent platform roadmap for scalable AI workloads. #J-18808-Ljbffr United States... 
    Suggested

    United States Digital Space LLC

    San Francisco, CA
    1 day ago
  •  ...interactive world models for robotics and embodied AI. We are hiring a Member of Technical Staff to lead reinforcement learning infrastructure and model post-training. You will work with...  ...-generating agents through fine-tuning, RL, and scalable evaluation in a on-site San... 
    Suggested

    Doist

    San Francisco, CA
    1 day ago
  • Kindredventures is recruiting infrastructure engineers to scale large-scale inference and evaluation around a physics-based LPM initiative. You will work on high...  ...record in building efficient inference stacks, GPU-aware optimization, and deep learning frameworks like... 
    Suggested

    Kindredventures

    San Francisco, CA
    1 day ago
  • $350k

     ...opportunities? Join a rapidly growing AI infrastructure provider delivering large-...  ...solutions for AI training and inference across global cloud and GPU environments. The organization...  ...This opportunity is for a Staff Site Reliability Engineer to lead the reliability of... 
    Full time
    San Francisco, CA
    a month ago
  • $190.9k - $232.8k

    A leading data and AI company is seeking a Staff Software Engineer for GenAI inference to lead the architecture and optimization of the inference engine. The role requires expertise in CUDA, GPU programming, and distributed systems design. Ideal candidates will have a strong... 

    Jobleads-US

    San Francisco, CA
    3 days ago
  • $179k - $218k

     ...and intelligence. As the only vertically integrated AI infrastructure company built from the ground up, we own and operate...  ...Silicon Reality" must be bridged.We are seeking a Senior Staff Data Center Operations Engineer, GPU Hardware Architecture to be the definitive technical... 
    Temporary work

    Crusoe

    San Francisco, CA
    1 day ago
  • $250k - $300k

     ...intelligence. As the only vertically integrated AI infrastructure company built from the ground up, we...  ...in production. That means owning the inference stack end to end: profiling where time...  ...will also work directly with customer engineering teams to tailor deployments to their... 
    Temporary work

    Crusoe

    San Francisco, CA
    2 days ago
  • $175k - $240k

     ...Scientific builds and deploys AI scientist agents to...  ...team run by scientists and engineers from leading institutions across...  ...As a Member of Technical Staff, Infrastructure Engineer, you'll play a key...  ...heterogeneous workloads (CPU, GPU, memory-intensive). Establish... 
    Full time
    Work at office

    Edison Scientific

    San Francisco, CA
    2 days ago
  • $207k - $290k

     ...Description About JazzX AI: Vision:...  ...experienced AI Engineer with deep...  ...Learning (RL) to join our team as a Senior Staff Architect. In this...  ...ensuring the RL infrastructure can scale to...  ...techniques , including inference-time search,...  ...(Kubernetes, GPU/TPU clusters,... 
    Worldwide
    Flexible hours

    JazzX AI

    San Francisco, CA
    more than 2 months ago
  • $160k - $300k

     ...Apiphany is a pioneering foundational AI company for physical product development...  ...Character, our mission is to revolutionize how engineering decisions are made, turning complexity...  .... About the Role As a Senior / Staff Infrastructure Engineer at Apiphany, you’ll design,... 
    Work at office
    Visa sponsorship
    Flexible hours

    Apiphany

    San Francisco, CA
    4 days ago
  • $220k - $280k

     ...About the Role Together AI is building the best inference infrastructure for voice applications. Our Voice AI...  ...reliability. We're looking for a Staff ML Engineer to drive the model serving layer...  ...to the frontier. You'll profile GPU utilization, design batching strategies... 
    Full time

    Together Ai

    San Francisco, CA
    10 hours ago
  •  ...the next generation of AI-driven game...  ...Senior Machine Learning Engineer for On-Device & Mobile...  ...significant parts of the inference stack — from a trained...  ...tuning across NPU, mobile GPU, and desktop/laptop GPU...  ...on-device benchmarking infrastructure, performance-... 
    Full time
    Work at office
    Remote work
    Worldwide

    Unity Technologies

    San Francisco, CA
    10 hours ago
  • $197.3k - $313.7k

     ...SalesforceSalesforce is the #1 AI CRM, where humans...  ...is looking for a Staff Machine Learning Engineer with deep expertise...  ...curation, training infrastructure, hyperparameter...  ...training pipelines on GPU infrastructure.Brainstorm...  ...optimization for inference (quantization,... 
    Full time

    Salesforce

    San Francisco, CA
    3 days ago
  • $307k - $352k

     ...About the Role Kikoff's infrastructure team builds the systems that enable engineering teams to move quickly without sacrificing...  ...Data Infrastructure. As a Staff Infrastructure Engineer, you will...  ...infrastructure boundaries. AI is increasing the amount of code... 
    Permanent employment
    Full time
    Local area

    Kikoff

    San Francisco, CA
    4 days ago
  • $276.5k - $300k

     ...accelerate people in the age of AI. As bots and autonomous...  ...software, AI, cryptography, mobile engineering, and global operations. Our...  .... About the Team Our Infrastructure team is a collaborative group...  ...Opportunity We are looking for a Staff Infrastructure Engineer to... 
    Full time
    Flexible hours

    Tools for Humanity

    San Francisco, CA
    11 days ago
  •  ...building a Large Physics foundation Model and seeks an infrastructure engineer to design, deploy, and operate its GPU-driven compute environment. You will enable...  ...optimizing distributed clusters that power training and inference workloads. You will extend orchestration,... 

    Kindredventures

    San Francisco, CA
    1 day ago
  •  ...Job Description Staff Machine Learning Engineer, Artificial Intelligence (AI) Required, Work From Home...  ...of research, infrastructure, and product. You are...  ..., evaluation systems, inference architecture, and deployment...  ...deployment, including GPU optimization, memory efficiency... 
    Remote work
    Work from home

    Ginas Tech Jobs

    San Francisco, CA
    29 days ago
  •  ...The role: SoFi’s Staff AI Engineer is a hands-on AI engineering role...  ...Develop robust, persistent infrastructure for agentic state...  ...high-throughput, low-latency inference across diverse hardware footprints...  ...the underlying Kubernetes/GPU orchestration for custom model... 
    Full time

    Sofi

    San Francisco, CA
    10 hours ago
  • $200k - $350k

     ...company building critical infrastructure that powers high-volume, real...  .... We are seeking a Staff AI Engineer to lead the design and deployment...  .... Design model serving, inference, evaluation, and...  ...PyTorch. Kubernetes and GPU infrastructure. RAG and... 
    Remote job
    Full time
    Immediate start

    Pragmatike

    San Francisco, CA
    10 hours ago
  • $192k - $260k

    A leading data and AI company is seeking a Staff Engineer to design and implement core systems for Foundation Model Serving. The ideal candidate will have...  ...across teams to ensure operational excellence in GPU serving workloads. Competitive salary range of $192,000... 

    Databricks Inc.

    San Francisco, CA
    1 day ago
  •  ...future and identify actions to alter it while scaling novel architectures across multimodal physical data. We are seeking infrastructure engineers who can design, implement, and optimize distributed training systems, profile performance, and collaborate with researchers... 

    causal

    San Francisco, CA
    1 day ago
  • $157k - $234k

     ...Description Job Description Waabi, founded by AI visionary Raquel Urtasun, is the leader...  ...- Collaborate with data scientists and engineers to understand model requirements and...  ...Bonus/nice to have:  - Experience with infrastructure-as-code (IaC) tools such as Terraform or... 
    Full time
    Work at office
    Work from home
    Flexible hours

    Waabi

    San Francisco, CA
    a month ago
  • $227.33k - $312.58k

    We’re looking for a Staff ML Data Engineer to join Procore’s AI & Frontier Models organization...  ..., evaluation, or inference workflows.Solid understanding...  ...and lineage toolsCloud & Infrastructure: AWS or GCP,...  ...Optimizing data pipelines for GPU‑backed training and large... 
    Full time
    Work at office
    Local area
    Immediate start
    3 days per week

    Procore Technologies

    San Francisco, CA
    10 hours ago
  • $173.5k - $331.05k

     ...on a mission to build the modern, AI-powered video rendering pipeline behind...  ...looking for a senior, hands-on engineer to own and evolve the cross-platform GPU rendering platform at the heart of...  ...GPU driver / hardware vendors ML inference integration (e.g., TensorRT, ONNX... 
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Francisco, CA
    10 hours ago
  • $168k - $247k

     ...the RoleAs a Senior/Staff Deep RL Engineer, you will design, train...  ...to on-vehicle inference. You'll help define how...  ...deep RL agents using GPU-accelerated simulation...  ...distributed training infrastructure in JAX across large compute...  ...proficiency in using AI coding tools (e.g.,... 
    Hourly pay
    Work at office
    Local area
    Remote work
    Flexible hours

    Doordash

    San Francisco, CA
    2 days ago
  • $156k - $190k

     ...only vertically integrated AI infrastructure company built from the...  ...Crusoe.About the Role:As a Staff Cloud Support Engineer, you are a technical authority...  ...NCCL, IB, GPU driver/firmware issues, distributed...  ...AI workloads (training + inference) with performance tuning and... 
    Temporary work

    Crusoe

    San Francisco, CA
    4 days ago
  •  ...and running the world’s best data and AI infrastructure platform so our customers can use deep...  ...to improve their business. Founded by engineers — and customer obsessed — we leap at every...  ...in production, across model serving, inference, retrieval, and agent frameworks, and... 
    Worldwide

    DataBricks

    San Francisco, CA
    1 day ago
  •  ...providing trusted decision-ready AI to the world's most critical organizations...  ...real consequences on. As a Staff Machine Learning Engineer, you’ll own AI-driven products end...  ...low-latency, high-concurrency inference (Triton, vLLM, GPU-backed serving) that stays fast and... 
    Full time
    Contract work
    Remote work
    Flexible hours

    Primer.ai

    San Francisco, CA
    10 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff Engineer, GPU AI Inference & RL Infrastructure. Be the first to apply!