Member of Technical Staff - RL Training Framework
$180kSpaceXAI
Member of Technical Staff - RL Training Framework SpaceXAI’s mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge. Our team is small, highly motivated, and focused on engineering excellence. This organization is for individuals who appreciate challenging themselves and thrive on curiosity. We operate with a flat organizational structure. All employees are expected to be hands‑on and to contribute directly to the company’s mission. Leadership is given to those who show initiative and consistently deliver excellence. Work ethic and strong prioritization skills are important. All employees are expected to have strong communication skills. They should be able to concisely and accurately share knowledge with their teammates. ABOUT THE ROLE The RL infrastructure team is looking for an engineer to help develop our RL training framework. RESPONSIBILITIES Design and implement the systems backing all RL workloads at SpaceXAI, from small scale ablations to production training runs. Profile, debug, and optimize end-to-end training performance Improve scalability and observability of the RL stack BASIC QUALIFICATIONS Experience building, debugging, and optimizing efficiency of large-scale distributed systems Comfortable diving into unfamiliar areas and solving problems at all levels of the stack Proficiency in Python, Jax, Rust, and/or C++ PREFERRED SKILLS AND EXPERIENCE Experience with large scale LLM training infrastructure Strong knowledge of reinforcement learning techniques Experience with RL numerics
COMPENSATION AND BENEFITS $180,000 - $440,000 USD
Base salary is just one part of our total rewards package at SpaceXAI, which also includes equity, comprehensive medical, vision, and dental coverage, access to a 401(k) retirement plan, short & long-term disability insurance, life insurance, and various other discounts and perks. #J-18808-Ljbffr SpaceXAI- The RL infrastructure team is looking for an engineer to help develop our RL training framework Design and implement the systems backing all RL workloads at xAI, from small scale ablations to production training runs Profile, debug, and optimize end-to-end training performance...TrainingVisa sponsorshipFlexible hours
- Member of Technical Staff - RL Inference SpaceXAI’s mission is to create AI systems that can accurately... ...engineer to help with low precision RL training and inference. RESPONSIBILITIES:... ...languages such as Python, C++ and/or Rust; frameworks such as PyTorch, Jax, CUDA...Training
$180k
...teammates. ABOUT THE ROLE: You will work on the most critical post-training and reinforcement learning challenges at any given time — including reward modeling, preference optimization (RLHF/DPO), and RL for improving reasoning, truthfulness, and real-world capabilities...TrainingTemporary work- Member of Technical Staff — Kernel / Compiler / Communication About the Role RadixArk is seeking a Member... .... This role is critical to scaling training and inference across thousands of... ...and developed Miles, our large‑scale RL framework. We build world‑class systems for AI...TrainingFlexible hours
- RadixArk is seeking a Member of Technical Staff — Inference to push the limits of large-scale AI inference... ...and developed Miles (our large‑scale RL framework). We're on a mission to democratize... ...class open systems for inference and training. Our team has optimized kernels...TrainingWorldwideFlexible hours
- About The Role RadixArk is seeking a Member of Technical Staff — Training to build and scale the systems that... ...inference optimization for large-scale RL or other production workload.... ...etc.) Familiarity with post-training framework (e.g. Miles, Slime, veRL, Prime-RL,...TrainingFlexible hours
- RadixArk is seeking a Member of Technical Staff — Training to build and scale the systems that train frontier... ...running training jobs Develop training frameworks and infrastructure tooling... ...and developed Miles, our large-scale RL framework. We build world-class infrastructure...TrainingFlexible hours
- ...What You'll Do As a Founding Member of the Technical Staff at Architect, you'll be at the forefront of training AI models for chip design,... ...where you'll own the end‑to‑end RL workflow—from reward... ...computing, and distributed training frameworks (e.g., PyTorch, CUDA, QLoRA,...Training
- About The Role RadixArk is seeking a Member of Technical Staff: Accelerator Systems to push the limits of performance for frontier AI systems... ...with distributed inference systems (SGLang, vLLM) or training/RL frameworks (Miles, Megatron, veRL, TorchTitan) CPU inference...TrainingFlexible hours
- About the Role As a Member of Technical Staff [Research] at NeoCognition , you’ll... ...of LLM reasoning, post-training, and agentic system design.... ...training (instruction tuning, RL, reasoning) Data pipeline... ...familiarity with modern ML frameworks (e.g., PyTorch, JAX, or TensorFlow...Training
- About The Role RadixArk is seeking a Member of Technical Staff — Diffusion Model to advance the frontier... ...—from designing novel algorithms to training and deploying models at scale. Your... ...and developed Miles (our large‑scale RL framework). We’re on a mission to democratize...TrainingFlexible hours
$180k - $250k
Member of Technical Staff -- TPU Systems (JAX / XLA / PALLAS) About the Role RadixArk... ...-performance inference and training systems using JAX, XLA, and... ...JAX, XLA, or TPU-focused frameworks Bachelor's or Master's... ...developed Miles (our large-scale RL framework). We're on a...TrainingFull timeFlexible hours- Member of Technical Staff Physical AI (Robotics / World Models) Palo Alto, CA About Orbifold... ...of data, evaluation, and model training. We design evaluation harnesses that... ...-designing datasets, evaluation frameworks, and training and RL pipelines that shape how their...TrainingShift work
$180k
...About the Role The mid‑training team at xAI aims to provide... ...boost the ceiling for RL. Engineer long‑context... ...Spark, Ray, and other frameworks for large‑scale data... ...interview”) during which a member of our team will ask... ...which consists of four technical interviews: Coding assessment...TrainingTemporary workRelocation- RadixArk is hiring a Member of Technical Staff — CI Engineer to own the infrastructure that keeps SGLang... ...and developed Miles (our large-scale RL framework). We’re on a mission to democratize... ...class open systems for inference and training. Our team has optimized kernels...TrainingFlexible hoursNight shift
$180k
Member of Technical Staff - Multimodal Understanding About xAI xAI’s mission is... ...curation/acquisition, tokenizer training, large‑scale pre‑training,... ...models. Create evaluation frameworks, internal benchmarks,... ...stack (pre‑training > SFT/RL/post‑training) to enable reasoning...TrainingTemporary work- ...Role RadixArk is seeking a Member of Technical Staff, Developer Technology (DevTech... ...to make LLM inference and training dramatically faster,... ...reinforcement-learning post-training framework for large-scale LLM and MoE... ...Training systems: RL post-training with Miles, FP...TrainingFlexible hours
- About the Role As a Member of Technical Staff [Platform] at NeoCognition , you’ll design and build the... ...developer tools, automation frameworks, or internal platforms . Familiarity... ...with machine learning infrastructure , training pipelines, or model evaluation tooling...Training
- ...Reinforcement Learning experiments (GRPO/PPO/DPO), training data mixes and reward signal explorations... ...are also encouraged to apply. RL Knowledge: Strong academic understanding... ...proficiency in Python and deep learning frameworks (PyTorch). You should be able to write clean...TrainingInternship
- ...world models as simulators, feature extractors, and training grounds for real robot policies and model‑based RL approaches. Most new experiments here will fail;... ...ownership of the full ML stack, including core frameworks that Odyssey researchers and product engineers rely...TrainingRemote workFlexible hours
- ...who can push LLM inference and training systems to the limit across... ...workloads, cost-per-million tokens, RL rollout efficiency, and... ...and cloud partners on deep technical evaluations Contribute performance... ...Miles, our large-scale RL framework. We're on a mission to...TrainingFlexible hours
$350k
Member of Sciences San Francisco Bay Area | Bellevue | Hybrid Company... ...Description The Member of Technical Staff Scientist is responsible for... ...AI practices. Develop frameworks for model evaluation, bias mitigation... ...factors such as education, training, work experience, business...TrainingWork experience placementLocal areaFlexible hours$148.5k - $223.9k
Senior Member of Technical Staff - AI ResearchSkip to main content#Senior Member of Technical Staff... ...Hands-on experience with deep learning frameworks** *Strong understanding of ML... ...Experience implementing and debugging model training, evaluation, and inference pipelines*...TrainingWork at office$200k - $300k
...This team works on foundational model capabilities, post-training techniques, building RL infra and infrastructure that benefits the entire... ...model development Build robust and effective training frameworks (on top of Megatron/PyTorch) for post-training LLMs Implement...TrainingFull time$180k
...Responsibilities span data curation, modeling, training, inference serving, and product... ...long-horizon synthesis, agentic planning, RL training, and world simulation (including... ...visual and audio data. Design evaluation frameworks, metrics, benchmarks, evals, and reward models...TrainingTemporary work- ...real robot tasks. Post-training at Rhoda means taking a... ...levels — from senior to staff. What You'll Do Design and implement RL training pipelines to improve... ...Build evaluation frameworks for post-trained policies... ...are expected to define technical direction and drive research...TrainingShift work
- ...Developer Advocate to build and engage our technical community around SGLang, Miles, and our... ..., including LLM serving, distributed training, and GPU optimization. Proficiency in... ..., and developed Miles (our large-scale RL framework). We're on a mission to democratize frontier...TrainingFlexible hours
$200k - $420k
Member of Technical Staff, Software Engineering At River, our mission is to create personal AI owned... ...hardware for local inference, bespoke training infrastructure, next-generation UIs,... ...Hands-on experience with modern AI frameworks (e.g., PyTorch, JAX) and tooling for...TrainingLocal areaVisa sponsorshipRelocation package$180k
...dramatically enhance the user experience Write data pipelines and training jobs that continuously learn from product data. Iterate and... ...at industrial scale Skilled in one or more DL software frameworks such as JAX or PyTorch Exceptional candidates may be...TrainingTemporary work$180k
...design and scale safeguards, detection systems, and evaluation frameworks that prevent harm while preserving creativity and user freedom... ...feedback loops between user interactions, model outputs, and training data to continuously improve safety while maintaining high performance...TrainingTemporary workWorldwide
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Member of Technical Staff - RL Training Framework. Be the first to apply!
- life support technician Palo Alto, CA
- personal computer support technician Palo Alto, CA
- systems support technician Palo Alto, CA
- technical support analyst Palo Alto, CA
- user support analyst Palo Alto, CA
- help desk technical support Palo Alto, CA
- technical support specialist Palo Alto, CA
- IT assistant Palo Alto, CA
- help desk assistant Palo Alto, CA
- work from home technical support specialist Palo Alto, CA

