Staff Engineer, GPU AI Inference & RL Infrastructure
B Capital
B Capital is seeking a skilled engineer for GPU infrastructure in San Francisco. This role involves designing and operating high-performance systems for model inference, synthetic data generation, and reinforcement learning. The ideal candidate has strong GPU systems experience, optimization skills, and a passion for working in cutting-edge AI. Benefits include top-tier compensation, comprehensive health insurance, and paid parental leave. #J-18808-Ljbffr B Capital
- ...intelligence open and accessible. We design, build, and operate GPU-heavy infrastructure for high-throughput model inference and mid-training workloads. Join a team that powers synthetic data generation, RL pipelines, and distributed model evaluation across thousands of...Suggested
- ...Barnes Associates Limited is seeking a Staff-level engineer to architect and evolve the AI infrastructure software stack for large-scale GPU workloads. You’ll drive orchestration,... ...production environment. You’ll contribute to inference platforms, model serving, and high-...Suggested
$200k - $400k
A leading AI technology company located in San Francisco is seeking an infrastructure engineer to build distributed systems for their AI inference engine. The role involves designing systems that ensure minimal latency and maximum reliability. Candidates should have a...SuggestedVisa sponsorship- Magic AI, Inc. is seeking a engineer for the Supercomputing Platform & Infrastructure to design, build, and operate large-scale GPU infrastructure powering model training and inference. You will implement Terraform-driven IaC across cloud and hybrid environments, manage...SuggestedVisa sponsorshipRelocation package
- ...Company in San Francisco is seeking a Member of Technical Staff for their infrastructure team. In this role, you will own the cloud systems that... ...API and build global low-latency, high-throughput GPU ML inference infrastructure. The ideal candidate will have solid experience...SuggestedVisa sponsorship
- Kindredventures is recruiting infrastructure engineers to scale large-scale inference and evaluation around a physics-based LPM initiative. You will work on high... ...record in building efficient inference stacks, GPU-aware optimization, and deep learning frameworks like...
- ...hiring Members of Technical Staff to build systems that accelerate LLM inference and own customer workloads... ...performance kernels, inference engine internals, and production infrastructure for named accounts. You... ...collaborating with a fast-growing AI inference company. #J-1880...
$200k - $225k
...Recruiting Researchers & Engineers with Foundational AI Experience A fast-growing... ...AI company is seeking a Staff Cloud Infrastructure Engineer to build and scale high-performance GPU compute infrastructure. You... ...performance tools for AI inference Collaborating with...Full time$220k
We build and run the inference engine behind every Perplexity query and deploy dozens of model... ...and multimodal models in our inference infrastructure, from weight loading, request scheduling... ...management to support in API Gateway. GPU kernels migration to CuTe DSL. Port our...- San Francisco Tensor Company is hiring a Member of Technical Staff for Sandbox Infrastructure to build a serverless GPU container service across NVIDIA, AMD, TPU and Trainium, scalable and isolated for untrusted code. You will work with compiler, post‑training and kernel...Relocation
- Inception is seeking engineers and scientists to design, optimize, and maintain core systems enabling scalable reinforcement learning... ...production-readiness. Responsibilities include building infrastructure for RL workloads, boosting training throughput, and creating...
- Causal is building a Large Physics foundation Model and seeks an infrastructure engineer to design, deploy, and operate large distributed GPU clusters. You will extend schedulers, build self-serve interfaces, and own storage and lineage for checkpoints and logs. You will...
- Vmax, an applied research lab, seeks an infrastructure engineer to build the systems layer for large-scale RL. You will enable researchers and ML engineers to run, debug, and reproduce experiments across thousands of GPUs, with a focus on reliability and scalability. The...
$350k
...opportunities? Join a rapidly growing AI infrastructure provider delivering large-... ...solutions for AI training and inference across global cloud and GPU environments. The organization... ...This opportunity is for a Staff Site Reliability Engineer to lead the reliability of...Full time$190.9k - $232.8k
A leading data and AI company is seeking a Staff Software Engineer for GenAI inference to lead the architecture and optimization of the inference engine. The role requires expertise in CUDA, GPU programming, and distributed systems design. Ideal candidates will have a strong...$200k - $250k
...anonymous venture-backed AI startup building... ...learning environments and infrastructure for frontier AI... ...platform and full-stack engineers from our network. If there... ...demand for high-quality RL environments grows.... ...scalable training and inference infrastructure for RL...Full timeH1bWork at officeRelocation package$275k - $315k
SF Tensor is building the fastest GPU compiler and model foundry to enable AI across clouds and chips. We’re hiring a Member of Technical Staff to own the modeling side from data curation... ...distillation to deployment, working with SFT, RL, and DPO to ship post-trained models in...Relocation package$225k
Dormont Manufacturing Co is looking for a Software Engineer on the Inference & RL Systems team in San Francisco. The role involves designing distributed systems, optimizing performance, and ensuring high reliability for RL and post-training workflows. The ideal candidate...- ...building the world's most efficient software for inference and agent hosting. In this role, you'll be one of the first engineers on Sailboxes, contributing across the stack... ...You'll focus on distributed systems, cloud infrastructure, and scalable, ergonomic tooling. The SF...Work at office
$215k - $260k
...intelligence. As the only vertically integrated AI infrastructure company built from the ground up, we... ...in production. That means owning the inference stack end to end: profiling where time... ...will also work directly with customer engineering teams to tailor deployments to their...Temporary work$179k - $218k
...and intelligence. As the only vertically integrated AI infrastructure company built from the ground up, we own and operate... ...Silicon Reality" must be bridged.We are seeking a Senior Staff Data Center Operations Engineer, GPU Hardware Architecture to be the definitive technical...Temporary work- ...San Francisco is seeking a Member of Technical Staff to lead distributed systems at the core of our AI infrastructure. You’ll design, build, and operate scheduling,... ...GPUs, and accelerators. You’ll work closely with inference, runtimes, compilers, kernels, and hardware...
- Together AI is building the best inference infrastructure for voice applications. We seek a Staff ML Engineer to own the model serving stack and optimize latency and throughput for real-time voice workloads. You'll work with state-of-the-art accelerators and collaborate...
$220k - $280k
...About the Role Together AI is building the best inference infrastructure for voice applications. Our Voice AI... ...reliability. We're looking for a Staff ML Engineer to drive the model serving layer... ...to the frontier. You'll profile GPU utilization, design batching strategies...Full time$207k - $290k
...Description About JazzX AI: Vision:... ...experienced AI Engineer with deep... ...Learning (RL) to join our team as a Senior Staff Architect. In this... ...ensuring the RL infrastructure can scale to support... ..., including inference-time search,... ...infrastructure (Kubernetes, GPU/TPU clusters,...WorldwideFlexible hours$190.2k - $345.65k
...multimedia generative AI — deep-tuned image,... ...are hiring a Senior Staff Machine Learning Engineer to architect and... ...indexing, and search infrastructure behind Firefly Foundry... ..., ANN index tuning, GPU-accelerated enrichment... ...models and the inference paths that produce them...Full timeTemporary workLocal areaWorldwide- ...the next generation of AI-driven game... ...Senior Machine Learning Engineer for On-Device & Mobile... ...significant parts of the inference stack — from a trained... ...tuning across NPU, mobile GPU, and desktop/laptop GPU... ...on-device benchmarking infrastructure, performance-...Full timeWork at officeRemote workWorldwide
$286.2k - $326.7k
...Overview Senior Staff Engineer, AI Compute (Remote Eligible) At Capital... ...investments in technology infrastructure and world-class talent —... ...infrastructure on top of CPU and GPU substrates. Your... .../ DL model training, model inference and feature generation pipelines...Full timePart timeLocal areaRemote work$197.3k - $313.7k
...SalesforceSalesforce is the #1 AI CRM, where humans... ...is looking for a Staff Machine Learning Engineer with deep expertise... ...curation, training infrastructure, hyperparameter... ...training pipelines on GPU infrastructure.Brainstorm... ...optimization for inference (quantization,...Full time- Sail builds the world’s most efficient software for inference and agent hosting. In this role, you’ll own token processing down to the lowest... ...design and implement exotic parallelism schemes, write custom GPU kernels for regimes like cascade attention, and understand every...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff Engineer, GPU AI Inference & RL Infrastructure. Be the first to apply!
- project engineer assistant project manager San Francisco, CA
- senior staff systems engineer San Francisco, CA
- staff data engineer San Francisco, CA
- assistant chief engineer San Francisco, CA
- assistant engineer San Francisco, CA
- assistant electrical engineer San Francisco, CA
- engineering aide San Francisco, CA
- software engineer staff San Francisco, CA
- staff design engineer San Francisco, CA
- research assistant engineering San Francisco, CA


