Senior Software Engineer, RL Post-Training Frameworks
$184k - $287.5kNVIDIA
Reinforcement learning post-training is driving some of the most significant capability gains in AI today. It is the process that teaches a model to reason through hard problems, follow complex instructions, and act as an autonomous agent. It is also one of the hardest infrastructure challenges in the field. RL requires inference, rollout generation, and training running in a continuous loop. The rollout step is what makes it hard: the model must interact with environments, tools, and other models to produce the signal that drives learning. Coordinating actor, critic, and reward models across heterogeneous hardware at scale pushes the limits of what distributed systems can do. NVIDIA is building an RL Frameworks engineering team to develop the open-source tools and infrastructure that AI researchers and post-training teams depend on. The team spans the full software stack, from collaborating closely with the researchers and labs pushing the frontier, to contributing to RL frameworks like VeRL, Miles, and TorchTitan, to improving the distributed runtimes they depend on, including Ray and Monarch. Whether your strength is working with researchers to understand and address their need optimizing deep learning frameworks, or building distributed infrastructure, we want to hear from you. Come join us to build the systems that enable the next generation of AI. What you will be doing: You will architect and build RL post-training infrastructure that scales efficiently from experimentation on a single GPU to production across thousands of nodes. This means tuning RL training‑inference‑rollout loops on GPUs, CPUs, and LPUs for performance where it matters, contributing to and improving the performance and usability of open‑source RL frameworks, and partnering with the teams who own them. The role also spans fault tolerance, elastic scaling, and fast restarts so long‑running distributed training jobs survive failures, stragglers, and resource contention. Beyond GPU‑accelerated training, this work includes partnering with teams building CPU‑driven rollout workloads, including tool‑use, code execution, and agentic environments, supplying the systems and framework engineering needed to run them efficiently alongside GPU‑or‑LPU‑accelerated generation and GPU‑accelerated training. It also means advocating for researcher and partner needs with NVIDIA's networking, math library, and compiler teams so the capabilities RL workloads require get prioritised and delivered, and working with hardware teams to take advantage of next‑generation hardware capabilities in post‑training workloads. What we need to see: MS or PhD in Computer Science, Computer Engineering, or a related field (or equivalent experience) 5+ years of professional experience in distributed systems, high‑performance computing, deep learning infrastructure, or ML systems engineering Strong proficiency in Python and C/C++ Demonstrated experience building or contributing to large‑scale distributed systems or runtime frameworks in production at a frontier AI lab, hyperscaler, or major technology company Strong verbal and written communication skills and the ability to collaborate across organisational and geographic boundaries Depth in one or more of the following technical areas: Reinforcement learning for LLM post‑training (RLHF, PPO, GRPO, DPO, reward modelling), including how algorithms map to distributed execution and the systems challenges they create (heterogeneous placement, rollouts, environment execution, resharding between training and generation) PyTorch internals, including distributed training primitives (FSDP, tensor parallelism, pipeline parallelism) and their composition Kubernetes runtime internals (container lifecycle, pod scheduling, resource quotas, GPU allocation) End‑to‑end distributed systems design (service boundaries, data flows, consistency models, failure modes, recovery approaches) Experience in any of the following areas is a plus: Deep expertise in networking (NCCL, NVLink, InfiniBand), advanced multi‑dimensional parallelisms (Megatron‑LM, FSDP2, TP/DP/PP, MoE), or memory optimisations (quantisation‑aware training, mixed precision) Experience integrating high‑performance inference engines (vLLM, SGLang, TensorRT‑LLM) into RL training loops for GPU‑accelerated rollout Strong background in actor‑ and task‑based distributed programming (Ray, Monarch, or comparable systems) Familiarity with multi‑turn training, multi‑agent co‑evolution, or VLM post‑training Ways to stand out from the crowd: Open‑source contributions to RL post‑training or distributed training projects (e.g., VeRL, Miles, TorchTitan, OpenRLHF, NeMo‑Aligner, DeepSpeed‑Chat), including significant work on framework internals where applicable Kubernetes work beyond routine operations (custom operators, GPU device plugins, or scheduling contributions) Direct experience operating frontier‑scale training (RL post‑training at thousands of GPUs and/or large‑scale LLM or multimodal pre‑training) Hands‑on experience with production distributed failures at scale (stragglers, resource contention, hardware faults) Widely considered to be one of the technology world’s most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. As you plan your future, see what we can offer to you and your family Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD for Level 4, and 224,000 USD - 356,500 USD for Level5. You will also be eligible for equity and benefits. NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, colour, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law. #J-18808-Ljbffr NVIDIA
$204k - $259k
...states. The Waymo ML Frameworks & Efficiency team partners... ...autonomous driving software. We help our partners... ..., including pre-training and post-training. They are geared... .... We are looking for engineers with ML system... ...reinforcement learning (RL), building systems that...SeniorTrainingFull time$184k - $287.5k
...motivated Deep Learning engineer to bring advanced CUDA... ...demands, ranging from training on scales up to 100K... ...Runtime abstractions in AI frameworks: from PoC to... ...principles (aka systems software fundamentals). Adaptability... ...July 1, 2026. This posting is for an existing...SeniorTraining$148k - $226.2k
...performance of the autonomous driving software stack before it reaches public roads. As a software engineer on the SimCore team, you... ...learning model training. We are looking for an engineer... ...software stack, simulation frameworks, and RL model training. Required Skills...SeniorTrainingRemote work$152k - $241.5k
We are now looking for a Senior Software Engineer for Quantized Inference! NVIDIA... ...data drawn from SFT/RL pipelines. Each new recipe... ...the team: CI, build systems, training infrastructure, pipeline friction... ..., export) or equivalent framework Experience reading, modifying...SeniorTraining- ...developing deep learning frameworks for AMD GPUs. Your... ...models, and enabling RL training and SOTA LLM and Multimodal... ...across internal GPU software teams and engage with... ...THE PERSON: Skilled engineer with strong technical... ...available here. This posting is for an existing vacancy...SeniorTraining
$182k - $242k
...of research. We are building the Applied Training team to fix this problem. In this role you... ...with the backend team to enable RL training runs that spawn thousands of isolated... ...documentation for running popular OSS training frameworks on CoreWeave to unblock customers and...SeniorTrainingPermanent employmentTemporary workCasual workWork at officeFlexible hours$161k - $239k
...5 years of experience with software development in one or more programming... ...experience in AI/ML related engineering, developing, deploying,... ...range displayed on each job posting reflects the minimum and maximum... ..., and relevant education or training. Your recruiter can share...SeniorTrainingFull time$152k - $241.5k
NVIDIA seeks a senior software engineer to join the AI Networking co... ...within LLM training and inference stacks,... ...applications, machine learning frameworks, and communication... ...reinforcement learning, offline RL, supervised learning)... ...10, 2026. This posting is for an existing...SeniorTraining$152k - $241.5k
Join NVIDIA's Solution Engineering team that is shaping... ...with architecture and software development teams. Scale... ...related middleware frameworks. Experience with robot... ...reinforcement learning for training and validation of... ...sim-to-real transfer of RL policies being a plus....SeniorTraining$184k - $287.5k
...highly motivated software professional to work... ...of high-level DL frameworks and low-level CUDA... ...bottlenecks in both training and inference... ...Science, Computer Engineering, Electrical Engineering... ...Learning (RL) or highly parallel... ...May 18, 2026. This posting is for an existing...SeniorTraining- NVIDIA Gruppe is looking for an RL Frameworks Engineer in Santa Clara, California, to architect and build scalable RL post-training infrastructure. You will ensure efficient scaling from single GPU experimentation to production across thousands of nodes, while collaborating...SeniorTraining
$240k - $320k
Job Description As the Senior Principal Engineer, E2E AI Training Framework for Autonomous Driving Systems, you will spearhead... .... 10+ years of experience in software development and system... ...leave. Pay ranges included in the postings, when included, generally reflect...SeniorTrainingFull timeWork experience placementLocal areaFlexible hours$182k - $242k
...job orchestration. They came to train models. Instead, they're doing... ...the backend team, enabling RL training runs to spawn thousands... ...running popular OSS training frameworks on CoreWeave to unblock customers... ...We Offer The range we’ve posted represents the typical compensation...SeniorTrainingPermanent employmentFull timeTemporary workCasual workWork at officeFlexible hours$184k - $287.5k
...Overview NVIDIA is hiring senior engineers to develop its AI platform and... ...optimizations in deep learning frameworks using JAX, a tool that can... ...platform to handle data, training and analysis for a wide range... ...numeric libraries, modular software design. Highly motivated with...SeniorTraining- ...Club/ Walmart Job Title: Senior Software Engineer (Python) Location: Sunnyvale,... ...integration patterns. ~ Exposure to AI/ML frameworks such as PyTorch, TensorFlow, or... ...system integration rather than model training. ~ Good understanding of system...SeniorTraining
$184k - $287.5k
...for outstanding AI systems engineers to develop groundbreaking technologies... ...in the inference systems software stack. As a member of the... ...across deep learning frameworks, libraries, kernels, and GPU... ...solutions for LLM inference and training (e.g., FlashInfer, Flash Attention...SeniorTraining$152k - $241.5k
...and industries. Within our software stack, CUTLASS stands out as... ...learning models’ inference and training passes to identify key GPU... ...including GPU architecture, DL frameworks, and QA as the performance... ...Computer Science, Computer Engineering, or related field (or equivalent...SeniorTraining$184k - $287.5k
...for outstanding AI systems engineers to develop groundbreaking technologies... ...in the inference systems software stack! We build innovative... ...NVIDIA across deep learning frameworks, libraries, kernels, and GPU... ...for LLM inference and training (e.g. FlashInfer, Flash Attention...SeniorTraining$184k - $287.5k
...We're looking for a senior engineer at the intersection of... ...ecosystem. As a Senior Software Engineer, you will... ...agentic orchestration frameworks to support modeling,... ...supporting infrastructure for training, validation, and... ...July 4, 2026. This posting is for an existing...SeniorTraining- ...testing, data mining, training, and metrics as first-... ...in a unified analytics framework. By joining this team,... ...insights to engineering and leadership, including... ...reviews, and by following software-engineering best practices... ...Apply Now on the job posting of interest. The...SeniorTrainingLocal areaWork from home
- ...Services team builds cloud-native systems, frameworks, and services for managing data across... ...-scale, high-performance GPU-based training and inference jobs. Our work gives... ...production environments. Use modern software engineering practices, including AI-assisted and agentic...SeniorTraining
$152k - $241.5k
...that enable researchers and engineers to develop the next generation... .... We are seeking a Software Engineer to join our MARS team... ...— infrastructure capable of training frontier models and executing... ...large-scale computing, and AI frameworks to help shape the future direction...SeniorTraining- .... We are now looking for a Senior Software Engineer to help accelerate the next... ...researchers, enable them to focus on training and development by reducing... ...Slurm or custom scheduling frameworks in production ML... ...until July10, 2026. This posting is for an existing vacancy....SeniorTraining
$152k - $241.5k
...supercomputers to cloud‑scale training and simulation. Our... ..., AI‑powered engineers for the DRIVE Mapping... ...with a background in software design, embedded software... ...Implement evaluation frameworks to measure performance... ...June 14, 2026. This posting is for an existing vacancy...SeniorTraining$184k - $287.5k
...caliber Deep Learning Engineer to bridge the gap between... ...road. Architect the software interface to... ...similar machine learning frameworks. Sophisticated proficiency... ...proven track record of training, deploying, or... ...April 25, 2026. This posting is for an existing vacancy...SeniorTraining$184.7k - $277.6k
...Chain We’re looking for customer-facing Software Engineers who can rapidly identify operational... ...resolution Guide user adoption, training, and iterative improvement as workflows... ...Familiarity with modern AI tooling, agent frameworks, prompt engineering, and orchestration...SeniorTrainingRelocation$137.1k - $188.3k
.../Video algorithms and software starting with fresh proof... ...Science, Electrical Engineering or equivalent Passion... ...test automation frameworks Professional level experience... ...production and post‑production workflows.... ...relevant education or training. Your recruiter can share...SeniorTrainingFull timeLocal areaWorldwideFlexible hours$174k - $253k
...5 years of experience with software development in Java programming... ...(A/B testing, experiment frameworks, data analysis). Excellent... ...About the Job Google's software engineers develop the next-generation... ..., and relevant education or training. US: $174000 - $253000 (USD...SeniorTraining$152k - $241.5k
NVIDIA is seeking outstanding senior engineers to work on the CUDA driver,... .... You will join a versatile software engineering team that... ...with distributed system and training/inference patterns (data/model... ...parallelism) and deep learning frameworks Your base salary will be determined...SeniorTraining- What You’ll Do As a Senior Software Engineer II (IC4) on the AI Workload Orchestration team, you will help build... ...multiple orchestration and scheduling frameworks such as Kueue, Volcano, and Ray to support modern AI training and inference workflows. It complements SUNK...SeniorTrainingTemporary workCasual workWork at officeRemote workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Software Engineer, RL Post-Training Frameworks. Be the first to apply!
- ngo software engineer Santa Clara, CA
- software developer Santa Clara, CA
- software developer internship no experience Santa Clara, CA
- part time software developer remote Santa Clara, CA
- financial software developer Santa Clara, CA
- senior software engineer ruby on rails Santa Clara, CA
- software engineer amazon Santa Clara, CA
- senior software design engineer Santa Clara, CA
- software engineer remote Santa Clara, CA
- software engineer entry level Santa Clara, CA

