Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior RL Infrastructure Engineer - Scalable GPU Systems

Vmax AI Corp

Vmax AI Corp is seeking a strong infrastructure engineer to build the systems layer for RL at scale. You will enable thousands of GPUs to run, debug, and reproduce large-scale RL experiments, tackling training orchestration, data pipelines, and observability. You will own infra projects end to end—from architecture to deployment—while aligning with ML researchers to translate complex experiments into durable, scalable platforms within our San Francisco office (hybrid option possible). #J-18808-Ljbffr Vmax AI Corp

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Senior RL Infrastructure Engineer - Scalable GPU Systems in San Francisco, CA vacancy
  •  ...design and operate distributed systems for serving models in...  ...model execution meets distributed infrastructure, influencing latency, throughput, and reliability of RL and training loops. You will own...  ...infrastructure enabling fast inference and scalable RL iteration, balancing KV-... 
    Senior

    Magic AI, Inc

    San Francisco, CA
    1 day ago
  • $250k - $280k

     ...Description Job Description Staff / Principal Founding Engineer (Backend-Leaning) – AI Systems Platform San Francisco (in-office) $250–280K base +...  ...-stage startup environments Experience building scalable, production-grade systems end-to-end Comfortable in... 
    Senior
    Work at office
    Immediate start
    Flexible hours

    Xpertalent

    San Francisco, CA
    8 days ago
  •  ...technology company in San Francisco is looking for a Senior Software Engineer to build scalable infrastructure for large‑scale training and fine-tuning of...  ...models. You will design distributed training systems and optimize GPU utilization while collaborating with cross-... 
    Senior

    Baseten

    San Francisco, CA
    1 day ago
  • $117.2k - $313.7k

     ...immediate opportunities for Lead software engineers who want their lines of code to have...  ...and drive innovations that improve system scalability, robustness, and availability....  ...developing SAAS products over public cloud infrastructure - AWS/Azure/GCP.* Proven experience... 
    Senior
    Full time
    Immediate start
    Remote work

    Salesforce

    San Francisco, CA
    17 hours ago
  •  ...company in San Francisco is seeking a Senior Software Engineer for Backend (Systems / Infrastructure). You will architect and deliver backend systems to maintain scalability as demand grows. This role involves optimizing APIs, managing GPU workloads, and collaborating with... 
    Senior

    Vizcom

    San Francisco, CA
    4 days ago
  •  ...the da Vinci surgical system and Ion—have transformed...  ....We’re a team of engineers, clinicians, and innovators...  ...of PositionAs a Senior Systems GPU Engineer - AI & Robotics...  ...robust, validated and scalable medical device products and infrastructures, including edge and cloud... 
    Senior
    Local area
    Worldwide
    Flexible hours

    Intuitive Surgical

    San Francisco, CA
    17 hours ago
  • Valkai, Inc. is seeking an experienced infrastructure engineer to build and operate secure, scalable systems underpinning our agents and products. You will own deployment, observability, security, and performance across cloud and self-hosted environments while collaborating... 
    Senior

    Valkai, Inc.

    San Francisco, CA
    17 hours ago
  • $260k - $325k

    Harvey, Inc in San Francisco seeks an experienced Infrastructure Engineer to design and maintain the infrastructure supporting its AI platform. You will ensure system stability and scalability while collaborating with various teams on critical infrastructure projects.... 
    Senior

    Harvey

    San Francisco, CA
    4 days ago
  • Inception is seeking engineers and scientists to design, optimize, and maintain core systems enabling scalable reinforcement learning for large models. You will work at...  ...readiness. Responsibilities include building infrastructure for RL workloads, boosting training throughput... 

    Inception

    San Francisco, CA
    1 day ago
  • Google is seeking a Senior Software Engineer, Embedded for Platforms and Devices in the US. You will contribute to low-level systems development, focusing on C/C++ and embedded operating systems to deliver scalable software for Google’s platforms and devices. The role... 
    Senior

    Google Inc.

    San Francisco, CA
    2 days ago
  • Vmax, an applied research lab, seeks an infrastructure engineer to build the systems layer for large-scale RL. You will enable researchers and ML engineers to run,...  ...thousands of GPUs, with a focus on reliability and scalability. The role emphasizes end-to-end ownership, from... 

    Vmax

    San Francisco, CA
    1 day ago
  • Lightning AI in New York/USA is seeking a Senior Software Engineer to design, build, and scale core backend systems powering our Studio platform. You will create APIs...  ...AI engineering teams to improve reliability, scalability, and developer experience. You will mentor engineers... 
    Senior

    DeepCamp

    San Francisco, CA
    4 days ago
  • OpenAI in San Francisco is seeking an experienced systems generalist to build an automated inference...  ...model and hardware data. This cross-stack role collaborates with research, infrastructure, security, product, and partnerships to deliver scalable, #J-18808-Ljbffr Slope
    Senior

    Slope

    San Francisco, CA
    3 days ago
  • Coinbase is building the future of the financial system by creating a scalable, onchain platform. We seek a data platform engineer to design and operate data-heavy services, enabling fast analytics and ML across teams. You will build self-service data tooling, ensure data... 
    Senior

    Coinbase

    San Francisco, CA
    2 days ago
  • Chime is seeking a Senior Software Engineer in Lending to own end-to-end backend systems that connect member-facing lending products with shared platforms. You’ll lead...  ..., shape architecture, and drive safety, scalability, and reliability across critical fintech systems... 
    Senior

    Chime

    San Francisco, CA
    17 hours ago
  •  ...applied AI research, flexible infrastructure, and seamless developer...  ...and help build the platform engineers turn to to ship AI products....  ...building the global operating system for distributed, heterogeneous...  ...foundational engineers to lead our GPU Networking efforts, making... 
    Full time
    Flexible hours

    Baseten

    San Francisco, CA
    17 hours ago
  • $130k - $170k

     ...Architecture and AI Systems Compensation: USD 13...  ...services. This is a senior‑level, highly technical...  ...and implement scalable Generative AI and RAG...  ...processes. Automate infrastructure provisioning using AWS...  ...Apply advanced prompt engineering, function calling, and... 
    Senior
    Full time

    Mogi I/O : OTT/Podcast/Short Video Apps for you

    San Francisco, CA
    4 days ago
  •  ...Senior Infrastructure Engineer Vast.ai's cloud powers AI projects and businesses all over...  ...help design and scale the core systems that power Vast.ai's global GPU marketplace. You'll work closely...  ...APIs Develop scalable infrastructure for workload scheduling... 
    Senior
    Full time
    Work at office

    Vast

    San Francisco, CA
    3 days ago
  • Luma AI in SF Bay Area is hiring for a Research Scientist/Engineer in Training Infrastructure. You will design distributed training systems for thousands of GPUs and enable researchers to push multimodal foundation models. You should have deep experience with PyTorch, CUDA... 
    Senior
    Remote job

    Luma AI

    San Francisco, CA
    1 day ago
  • $190k - $280k

     ...Senior Network Engineer San Francisco About the Role Together...  ...the global network infrastructure supporting our...  ...available, reliable, scalable, and performant. The...  ...with infrastructure, systems, security, and application...  ...supporting GPU clusters, HPC environments... 
    Senior
    Full time

    Together AI

    San Francisco, CA
    4 days ago
  •  ...in San Francisco, CA, seeks a Sr. Mgr, AI Engineering to lead the design, deployment, and optimization of large-scale AI systems in production. You will guide cross-functional...  ...and ensure performance, cost, latency and scalability goals are met while advancing Capital One’... 
    Senior

    Capital One

    San Francisco, CA
    1 day ago
  •  ...of hands-on experience in infrastructure engineering with demonstrated depth and...  ...-as-code — with a focus on scalability versus resource use. · Track...  ...Engineering, Information Systems, or a related technical field...  ..., AMD EPYC, NVMe storage, GPU accelerators) sufficient to... 
    Senior
    Temporary work

    PB consulting

    San Francisco, CA
    26 days ago
  • $117.2k - $313.7k

     ..., and you are the future of Salesforce.Distributed Systems Software Engineer - Public Cloud (Senior/Lead/Principal) Note: By applying to the Public Cloud...  ..., and production-ready.Your Impact:Deliver cloud infrastructure automation tools, frameworks, workflows, and validation... 
    Senior
    Full time

    Salesforce

    San Francisco, CA
    4 days ago
  • A fast-growing SaaS company in San Francisco is looking for a Senior AI Engineer to design and build scalable, production-grade AI systems. You will develop LLM-powered solutions and architect optimized AI systems. The role demands extensive software engineering experience... 
    Senior

    Harnham

    San Francisco, CA
    1 day ago
  • An innovative AI company is seeking a Senior Backend Engineer to own their core infrastructure. You will design scalable backend systems and collaborate closely with cross-functional teams to implement the company's technical vision. Ideal candidates have 5+ years of experience... 
    Senior

    Trust In SODA

    San Francisco, CA
    2 days ago
  • $225k

     ...rapidly scaling AI cloud infrastructure provider building next-generation GPU platforms designed for...  ...is looking for a Senior Network Engineer to design, deploy, and...  ...Cumulus Linux to build scalable 400G network fabrics optimised...  ...ECMP) Strong Linux systems knowledge and network... 
    Senior
    Full time
    Remote work
    San Francisco, CA
    more than 2 months ago
  • $216k - $270k

    Scale AI, Inc. is seeking a software engineer to design, build, and maintain scalable systems within its Generative AI Data Engine. As part of a dynamic hybrid team based in San Francisco or New York City, you will play a crucial role in producing high-quality AI data... 
    Senior

    Scale AI, Inc.

    San Francisco, CA
    3 days ago
  • Nscale seeks a Senior Infrastructure Support Engineer to own the health of GPU fleets and high‑performance fabrics. You will operate across GPU hardware, Linux, and data centre operations, bridging Support, DC Operations, and Engineering. You’ll diagnose complex issues,... 
    Senior
    Remote work

    Nscale

    San Francisco, CA
    1 day ago
  • $250k

    A Series A Funded start-up in California is seeking a Systems Engineer to design and optimize systems handling complex ML pipelines. The role involves building scalable infrastructure, developing CI/CD pipelines, and ensuring system performance. Key qualifications include... 
    Senior

    Acceler8 Talent

    San Francisco, CA
    17 hours ago
  • Mapbox, the leading real-time location platform, seeks an Engineer to design, develop and operate significant areas of our routing...  ...and OEMs. You'll work with Rust/Go, AWS, and distributed systems, shaping scalable backend services while delivering performant navigation... 
    Senior

    Mapbox

    San Francisco, CA
    17 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior RL Infrastructure Engineer - Scalable GPU Systems. Be the first to apply!