Member of Technical Staff - Post-Training and RL
$180kSpaceXAI
Member of Technical Staff - Post-Training and RL SpaceXAI’s mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge. Our team is small, highly motivated, and focused on engineering excellence. This organization is for individuals who appreciate challenging themselves and thrive on curiosity. We operate with a flat organizational structure. All employees are expected to be hands‑on and to contribute directly to the company’s mission. Leadership is given to those who show initiative and consistently deliver excellence. Work ethic and strong prioritization skills are important. All employees are expected to have strong communication skills. They should be able to concisely and accurately share knowledge with their teammates.
ABOUT THE ROLE
You will work on the most critical post‑training and reinforcement learning challenges at any given time — including reward modeling, preference optimization (RLHF/DPO), and RL for improving reasoning, truthfulness, and real‑world capabilities. You will get clarity on your first project before an offer.BASIC QUALIFICATIONS
You believe truth‑seeking AI is the most important and challenging problem. You are obsessed about building incredibly useful models through post‑training and RL techniques. You are a power user of AI models and eager to push the boundaries of what’s possible with reinforcement learning and alignment methods. If you previously worked on post‑training, RLHF, or trained models used by millions of people it’s a big plus, but relevant experience is not required. You take pride in your work and thrive in meritocratic environments.COMPENSATION AND BENEFITS
$180,000 - $600,000 USD
Base salary is just one part of our total rewards package at SpaceXAI, which also includes equity, comprehensive medical, vision, and dental coverage, access to a 401(k) retirement plan, short & long‑term disability insurance, life insurance, and various other discounts and perks. #J-18808-Ljbffr SpaceXAI$180k
...able to concisely and accurately share knowledge with their teammates. ABOUT THE ROLE: The RL infrastructure team is looking for an engineer to help develop our RL training framework. RESPONSIBILITIES: Design and implement the systems backing all RL workloads...TrainingTemporary work- ...able to concisely and accurately share knowledge with their teammates. ABOUT THE ROLE: The RL infrastructure team is looking for an engineer to help with low precision RL training and inference. RESPONSIBILITIES: Design and optimize our inference stack for all shapes...Training
- The RL infrastructure team is looking for an engineer to help develop our RL training framework Design and implement the systems backing all RL workloads at xAI, from small scale ablations to production training runs Profile, debug, and optimize end-to-end training performance...TrainingVisa sponsorshipFlexible hours
- About The Role RadixArk is seeking a Member of Technical Staff — Training to build and scale the systems that train... ...or operating large-scale agentic post-training systems. Experience working... ...inference optimization for large-scale RL or other production workload....TrainingFlexible hours
- About the Role As a Member of Technical Staff [Research] at NeoCognition , you’ll be part of the core... ...in the areas of LLM reasoning, post-training, and agentic system design. Develop... ...LLM post-training (instruction tuning, RL, reasoning) Data pipeline design and...Training
- RadixArk is seeking a Member of Technical Staff — Inference to push the limits of large-scale AI inference... ...and developed Miles (our large‑scale RL framework). We're on a mission to... ...‑class open systems for inference and training. Our team has optimized kernels serving...TrainingWorldwideFlexible hours
- Member of Technical Staff — Kernel / Compiler / Communication About the Role RadixArk is seeking a Member... .... This role is critical to scaling training and inference across thousands of GPUs... ...and developed Miles, our large‑scale RL framework. We build world‑class systems...TrainingFlexible hours
- About The Role RadixArk is seeking a Member of Technical Staff: Accelerator Systems to push the limits of performance for frontier AI systems.... ...Experience with distributed inference systems (SGLang, vLLM) or training/RL frameworks (Miles, Megatron, veRL, TorchTitan) CPU...TrainingFlexible hours
- ...intersection of data, evaluation, and model training. We design evaluation harnesses that... ...frameworks, and training and RL pipelines that shape how their models... ...training. About The Role We are hiring Members of Technical Staff to build the data and evaluation foundations...TrainingShift work
- RadixArk is hiring a Member of Technical Staff — CI Engineer to own the infrastructure that keeps SGLang... ...and developed Miles (our large-scale RL framework). We’re on a mission to... ...-class open systems for inference and training. Our team has optimized kernels serving...TrainingFlexible hoursNight shift
- ...What You’ll Do As a Founding Member of the Technical Staff at Architect, you'll be at the forefront of training AI models for chip design, verification... ..., scaling, and improving post-training techniques to... ...where you'll own the end-to-end RL workflow—from reward modeling...Training
- About The Role RadixArk is seeking a Member of Technical Staff — Diffusion Model to advance the frontier... ...—from designing novel algorithms to training and deploying models at scale. Your work... ...and developed Miles (our large‑scale RL framework). We’re on a mission to...TrainingFlexible hours
- RadixArk is seeking a Member of Technical Staff — Training to build and scale the systems that train frontier AI models. You will work on large-scale... ...LLM serving engine), and developed Miles, our large-scale RL framework. We build world-class infrastructure for AI training...TrainingFlexible hours
$180k
...teammates. About the Role The mid‑training team at xAI aims to provide an... ...to boost the ceiling for RL. Engineer long‑context data recipes... ...interview”) during which a member of our team will ask some... ...process, which consists of four technical interviews: Coding assessment...TrainingTemporary workRelocation$180k - $250k
Member of Technical Staff -- TPU Systems (JAX / XLA / PALLAS) About the Role RadixArk is looking for a... ...build high-performance inference and training systems using JAX, XLA, and Pallas. You... ...and developed Miles (our large-scale RL framework). We're on a mission to democratize...TrainingFull timeFlexible hours$180k
Member of Technical Staff - Multimodal Understanding About xAI xAI’s mission is to create AI... ...curation/acquisition, tokenizer training, large‑scale pre‑training, post‑training/alignment, infrastructure... ...across the stack (pre‑training > SFT/RL/post‑training) to enable...TrainingTemporary work- Member of Technical Staff — Developer Technology About the Role RadixArk is seeking... ...) to make LLM inference and training dramatically faster, cheaper... ...our reinforcement-learning post-training framework for large... ...silicon Training systems: RL post‑training with Miles, FP...TrainingFlexible hours
- Member of Technical Staff - Developer Experience About the Role RadixArk is seeking... ..., video guides, blog posts, and live streams. Host workshops... ...including LLM serving, distributed training, and GPU optimization.... ...developed Miles (our large-scale RL framework). We're on a...TrainingFlexible hours
$180k
...Distill the intelligence of flagship models into flash models through synthetic data generation. Optimize mid-training data mixtures to boost the ceiling for RL. Engineer long-context data recipes. Develop robust and diverse evaluation for mid-training checkpoints...TrainingTemporary work$180k
...Member of Technical Staff, Pre-training Data Infrastructure xAI’s mission is to create AI systems that can accurately understand the universe and aid... ...discoverability and data quality at scale for both pre‑training and post‑training across different modalities. Build, run, and...TrainingTemporary workRelocation- Kindredventures in Palo Alto is looking for a Founding Member of the Technical Staff to lead AI model training for chip design. You will work on cutting-edge Reinforcement Learning environments and build end-to-end ML pipelines. The successful candidate will hold a PhD...Training
- Member of Technical Staff — Kernel / Compiler / Communication RadixArk is seeking a deeply technical engineer who pushes the limits of performance... ...and interconnects. This role is critical to scaling training and inference across thousands of GPUs, where microseconds...TrainingFlexible hours
- About the Role As a Member of Technical Staff [Platform] at NeoCognition , you’ll design and build the internal systems that power everything... ...to have: Experience with machine learning infrastructure , training pipelines, or model evaluation tooling. Background in monitoring...Training
$180k
Member of Technical Staff - Pre-Training About xAI xAI’s mission is to create AI systems that can accurately understand the universe and aid humanity in... ...discoverability and data quality at scale for both pre‑training and post‑training across different modalities. Build, run, and...TrainingTemporary work- ...SuperIntelligence, xAI, Apple and Intel. What You’ll Do As a Founding Member of the Technical Staff (Applied AI) at Architect, you’ll sit at the... ...correct ones. Partner closely with the ML research, post‑training, and infra teams to turn hardware domain expertise into...Training
- ...About the Role We are looking for a Member of Technical Staff (MTS - Research) to help us build the... ...work will involve iterating on novel RL approaches and translating them into... ...Loop: Design and execute high-quality post-training runs (CPT, SFT, RL) to deliver frontier...Training
- ...will be working at the cusp of what’s possible, using world models as simulators, feature extractors, and training grounds for real robot policies and model‑based RL approaches. Most new experiments here will fail; your focus will be on maximally learning from failed...TrainingRemote workFlexible hours
- ...SuperIntelligence, xAI, Apple and Intel. What You’ll Do As a Founding Member of the Technical Staff on the RTL Design team at Architect, you’ll own the AI-... ...5/LPDDR5X command/address protocols, timing parameters, training sequences, and/or HBM2E/HBM3 pseudo‑channel architecture,...Training
- ...Reinforcement Learning experiments (GRPO/PPO/DPO), training data mixes and reward signal explorations. Contribute to research on post‑training techniques, running ablation studies... ...experience are also encouraged to apply. RL Knowledge: Strong academic understanding or...TrainingInternship
$324k - $396k
About the Role Member of Technical Staff (X.AI LLC; Palo Alto, CA): Introduce innovative techniques and analyses to the AI field to facilitate... ...and language understanding. Stabilize large language model training, pipeline parallelism training of large language models,...TrainingRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Member of Technical Staff - Post-Training and RL. Be the first to apply!
- life support technician Palo Alto, CA
- personal computer support technician Palo Alto, CA
- systems support technician Palo Alto, CA
- technical support analyst Palo Alto, CA
- user support analyst Palo Alto, CA
- help desk technical support Palo Alto, CA
- technical support specialist Palo Alto, CA
- IT assistant Palo Alto, CA
- help desk assistant Palo Alto, CA
- work from home technical support specialist Palo Alto, CA


