Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Member of Technical Staff - Post-Training and RL

$180k

SpaceXAI

Member of Technical Staff - Post-Training and RL SpaceXAI’s mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge. Our team is small, highly motivated, and focused on engineering excellence. This organization is for individuals who appreciate challenging themselves and thrive on curiosity. We operate with a flat organizational structure. All employees are expected to be hands‑on and to contribute directly to the company’s mission. Leadership is given to those who show initiative and consistently deliver excellence. Work ethic and strong prioritization skills are important. All employees are expected to have strong communication skills. They should be able to concisely and accurately share knowledge with their teammates.

ABOUT THE ROLE

You will work on the most critical post‑training and reinforcement learning challenges at any given time — including reward modeling, preference optimization (RLHF/DPO), and RL for improving reasoning, truthfulness, and real‑world capabilities. You will get clarity on your first project before an offer.

BASIC QUALIFICATIONS

You believe truth‑seeking AI is the most important and challenging problem. You are obsessed about building incredibly useful models through post‑training and RL techniques. You are a power user of AI models and eager to push the boundaries of what’s possible with reinforcement learning and alignment methods. If you previously worked on post‑training, RLHF, or trained models used by millions of people it’s a big plus, but relevant experience is not required. You take pride in your work and thrive in meritocratic environments.

COMPENSATION AND BENEFITS

$180,000 - $600,000 USD

Base salary is just one part of our total rewards package at SpaceXAI, which also includes equity, comprehensive medical, vision, and dental coverage, access to a 401(k) retirement plan, short & long‑term disability insurance, life insurance, and various other discounts and perks. #J-18808-Ljbffr SpaceXAI

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Member of Technical Staff - Post-Training and RL in Palo Alto, CA vacancy
  • $180k

     ...able to concisely and accurately share knowledge with their teammates. ABOUT THE ROLE: The RL infrastructure team is looking for an engineer to help develop our RL training framework. RESPONSIBILITIES: Design and implement the systems backing all RL workloads... 
    Training
    Temporary work

    SpaceXAI

    Palo Alto, CA
    a month ago
  •  ...able to concisely and accurately share knowledge with their teammates. ABOUT THE ROLE: The RL infrastructure team is looking for an engineer to help with low precision RL training and inference. RESPONSIBILITIES: Design and optimize our inference stack for all shapes... 
    Training

    Pantera Capital

    Palo Alto, CA
    1 day ago
  • The RL infrastructure team is looking for an engineer to help develop our RL training framework Design and implement the systems backing all RL workloads at xAI, from small scale ablations to production training runs Profile, debug, and optimize end-to-end training performance... 
    Training
    Visa sponsorship
    Flexible hours

    Xai

    Palo Alto, CA
    3 days ago
  • About The Role RadixArk is seeking a Member of Technical Staff — Training to build and scale the systems that train...  ...or operating large-scale agentic post-training systems. Experience working...  ...inference optimization for large-scale RL or other production workload.... 
    Training
    Flexible hours

    RadixArk

    Palo Alto, CA
    3 days ago
  • About the Role As a Member of Technical Staff [Research] at NeoCognition , you’ll be part of the core...  ...in the areas of LLM reasoning, post-training, and agentic system design. Develop...  ...LLM post-training (instruction tuning, RL, reasoning) Data pipeline design and... 
    Training

    NeoCognition Inc.

    Palo Alto, CA
    4 days ago
  • RadixArk is seeking a Member of Technical Staff — Inference to push the limits of large-scale AI inference...  ...and developed Miles (our large‑scale RL framework). We're on a mission to...  ...‑class open systems for inference and training. Our team has optimized kernels serving... 
    Training
    Worldwide
    Flexible hours

    Dormont Manufacturing Co

    Palo Alto, CA
    4 days ago
  • Member of Technical Staff — Kernel / Compiler / Communication About the Role RadixArk is seeking a Member...  .... This role is critical to scaling training and inference across thousands of GPUs...  ...and developed Miles, our large‑scale RL framework. We build world‑class systems... 
    Training
    Flexible hours

    RadixArk

    Palo Alto, CA
    2 days ago
  • About The Role RadixArk is seeking a Member of Technical Staff: Accelerator Systems to push the limits of performance for frontier AI systems....  ...Experience with distributed inference systems (SGLang, vLLM) or training/RL frameworks (Miles, Megatron, veRL, TorchTitan) CPU... 
    Training
    Flexible hours

    RadixArk

    Palo Alto, CA
    6 days ago
  •  ...intersection of data, evaluation, and model training. We design evaluation harnesses that...  ...frameworks, and training and RL pipelines that shape how their models...  ...training. About The Role We are hiring Members of Technical Staff to build the data and evaluation foundations... 
    Training
    Shift work

    Orbifold AI

    Palo Alto, CA
    1 day ago
  • RadixArk is hiring a Member of Technical Staff — CI Engineer to own the infrastructure that keeps SGLang...  ...and developed Miles (our large-scale RL framework). We’re on a mission to...  ...-class open systems for inference and training. Our team has optimized kernels serving... 
    Training
    Flexible hours
    Night shift

    RadixArk

    Palo Alto, CA
    3 days ago
  •  ...What You’ll Do As a Founding Member of the Technical Staff at Architect, you'll be at the forefront of training AI models for chip design, verification...  ..., scaling, and improving post-training techniques to...  ...where you'll own the end-to-end RL workflow—from reward modeling... 
    Training

    Architect Labs

    Palo Alto, CA
    1 day ago
  • About The Role RadixArk is seeking a Member of Technical Staff — Diffusion Model to advance the frontier...  ...—from designing novel algorithms to training and deploying models at scale. Your work...  ...and developed Miles (our large‑scale RL framework). We’re on a mission to... 
    Training
    Flexible hours

    RadixArk

    Palo Alto, CA
    1 day ago
  • RadixArk is seeking a Member of Technical Staff — Training to build and scale the systems that train frontier AI models. You will work on large-scale...  ...LLM serving engine), and developed Miles, our large-scale RL framework. We build world-class infrastructure for AI training... 
    Training
    Flexible hours

    RadixArk

    Palo Alto, CA
    1 day ago
  • $180k

     ...teammates. About the Role The mid‑training team at xAI aims to provide an...  ...to boost the ceiling for RL. Engineer long‑context data recipes...  ...interview”) during which a member of our team will ask some...  ...process, which consists of four technical interviews: Coding assessment... 
    Training
    Temporary work
    Relocation

    Pantera Capital

    Palo Alto, CA
    2 days ago
  • $180k - $250k

    Member of Technical Staff -- TPU Systems (JAX / XLA / PALLAS) About the Role RadixArk is looking for a...  ...build high-performance inference and training systems using JAX, XLA, and Pallas. You...  ...and developed Miles (our large-scale RL framework). We're on a mission to democratize... 
    Training
    Full time
    Flexible hours

    RadixArk

    Palo Alto, CA
    4 days ago
  • $180k

    Member of Technical Staff - Multimodal Understanding About xAI xAI’s mission is to create AI...  ...curation/acquisition, tokenizer training, large‑scale pre‑training, post‑training/alignment, infrastructure...  ...across the stack (pre‑training > SFT/RL/post‑training) to enable... 
    Training
    Temporary work

    xAI

    Palo Alto, CA
    5 days ago
  • Member of Technical Staff — Developer Technology About the Role RadixArk is seeking...  ...) to make LLM inference and training dramatically faster, cheaper...  ...our reinforcement-learning post-training framework for large...  ...silicon Training systems: RL post‑training with Miles, FP... 
    Training
    Flexible hours

    RadixArk

    Palo Alto, CA
    4 days ago
  • Member of Technical Staff - Developer Experience About the Role RadixArk is seeking...  ..., video guides, blog posts, and live streams. Host workshops...  ...including LLM serving, distributed training, and GPU optimization....  ...developed Miles (our large-scale RL framework). We're on a... 
    Training
    Flexible hours

    RadixArk

    Palo Alto, CA
    4 days ago
  • $180k

     ...Distill the intelligence of flagship models into flash models through synthetic data generation. Optimize mid-training data mixtures to boost the ceiling for RL. Engineer long-context data recipes. Develop robust and diverse evaluation for mid-training checkpoints... 
    Training
    Temporary work

    SpaceXAI

    Palo Alto, CA
    a month ago
  • $180k

     ...Member of Technical Staff, Pre-training Data Infrastructure xAI’s mission is to create AI systems that can accurately understand the universe and aid...  ...discoverability and data quality at scale for both pre‑training and post‑training across different modalities. Build, run, and... 
    Training
    Temporary work
    Relocation

    xAI

    Palo Alto, CA
    7 hours ago
  • Kindredventures in Palo Alto is looking for a Founding Member of the Technical Staff to lead AI model training for chip design. You will work on cutting-edge Reinforcement Learning environments and build end-to-end ML pipelines. The successful candidate will hold a PhD... 
    Training

    Kindredventures

    Palo Alto, CA
    3 days ago
  • Member of Technical Staff — Kernel / Compiler / Communication RadixArk is seeking a deeply technical engineer who pushes the limits of performance...  ...and interconnects. This role is critical to scaling training and inference across thousands of GPUs, where microseconds... 
    Training
    Flexible hours

    RadixArk

    Palo Alto, CA
    1 day ago
  • About the Role As a Member of Technical Staff [Platform] at NeoCognition , you’ll design and build the internal systems that power everything...  ...to have: Experience with machine learning infrastructure , training pipelines, or model evaluation tooling. Background in monitoring... 
    Training

    NeoCognition

    Palo Alto, CA
    4 days ago
  • $180k

    Member of Technical Staff - Pre-Training About xAI xAI’s mission is to create AI systems that can accurately understand the universe and aid humanity in...  ...discoverability and data quality at scale for both pre‑training and post‑training across different modalities. Build, run, and... 
    Training
    Temporary work

    Xai

    Palo Alto, CA
    5 days ago
  •  ...SuperIntelligence, xAI, Apple and Intel. What You’ll Do As a Founding Member of the Technical Staff (Applied AI) at Architect, you’ll sit at the...  ...correct ones. Partner closely with the ML research, post‑training, and infra teams to turn hardware domain expertise into... 
    Training

    Architect Labs

    Palo Alto, CA
    2 days ago
  •  ...About the Role We are looking for a Member of Technical Staff (MTS - Research) to help us build the...  ...work will involve iterating on novel RL approaches and translating them into...  ...Loop: Design and execute high-quality post-training runs (CPT, SFT, RL) to deliver frontier... 
    Training

    Collinear AI

    Sunnyvale, CA
    6 days ago
  •  ...will be working at the cusp of what’s possible, using world models as simulators, feature extractors, and training grounds for real robot policies and model‑based RL approaches. Most new experiments here will fail; your focus will be on maximally learning from failed... 
    Training
    Remote work
    Flexible hours

    Odyssey

    Palo Alto, CA
    3 days ago
  •  ...SuperIntelligence, xAI, Apple and Intel. What You’ll Do As a Founding Member of the Technical Staff on the RTL Design team at Architect, you’ll own the AI-...  ...5/LPDDR5X command/address protocols, timing parameters, training sequences, and/or HBM2E/HBM3 pseudo‑channel architecture,... 
    Training

    Kindredventures

    Palo Alto, CA
    4 days ago
  •  ...Reinforcement Learning experiments (GRPO/PPO/DPO), training data mixes and reward signal explorations. Contribute to research on post‑training techniques, running ablation studies...  ...experience are also encouraged to apply. RL Knowledge: Strong academic understanding or... 
    Training
    Internship

    Architect Labs

    Palo Alto, CA
    1 day ago
  • $324k - $396k

    About the Role Member of Technical Staff (X.AI LLC; Palo Alto, CA): Introduce innovative techniques and analyses to the AI field to facilitate...  ...and language understanding. Stabilize large language model training, pipeline parallelism training of large language models,... 
    Training
    Remote work

    Xai

    Palo Alto, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Member of Technical Staff - Post-Training and RL. Be the first to apply!