Post-Training AI Research Engineer - RL & Agentic Infra
Storm3
Fundamental AI Research Institute in San Francisco is seeking researchers experienced in LLM post-training, agentic infrastructure, RL environments, and evaluation harnesses to accelerate flagship AI research. Responsibilities include building and scaling post-training platforms and RL infra, developing benchmarks, contributing to deployment of large-scale models, and publishing results at top conferences. Join a fast-growing team with ample GPU resources and a collaborative culture, offering a #J-18808-Ljbffr Storm3
- ...About the Team The Post-Training Frontiers team... ...training for the agentic models we ship in... ...go into the final RL run and deciding... ...(3) building the research and infra for horizontal integrations... ...will work across engineering and... ...OpenAI is an AI research and deployment...TrainingFull time
- ...web by building AI agents that can... ...-first, from training our own models... ...experience in AI research and product spanning... ...Scale infra for post-training of multimodal... ...(CPT, SFT, RL, search, reward... ...Scale infra for agentic inference (throughput... ...with product engineers to translate...TrainingWork at officeRelocationVisa sponsorship
- HeyMilo AI is hiring a Research Engineer to join our Applied AI team in San Francisco. You’ll design and build reinforcement learning environments... ...take a use case from problem definition to reproducible environments for training and evaluation. #J-18808-Ljbffr HeyMilo AITraining
- ...building next‑generation AI systems that help... ...applications. As an Applied AI Research Engineer, you’ll focus on human‑machine teaming and agentic AI to build systems... ..., experiment with post‑training techniques, and integrate... ...ideas in GenAI, RL, agentic workflows, evaluation...TrainingRemote workRelocation packageFlexible hours
$269.1k - $307.2k
Distinguished AI Engineer (Agentic AI Platform) At Capital... ...model minutiae or infra plumbing. You will... ...team of engineers, research scientists, technical... ...techniques for optimizing training and inference... ...at the time of this posting. Salaries for part-time...TrainingFull timePart timeWork at officeLocal area- ...to apply. Mission Design, train, ship, iterate on, and innovate on the AI brains behind The Path's AI Therapist. Combine research, data science, and engineering to create models, orchestration,... ...training data curation, building RL environments, new model architectures...Training
$110.7k - $372.9k
...need. Deloitte has a new AI-first effort, backed... ...reasoning models and agentic systems to rebuild how... ...lab. As an Agentic AI Engineer, you will design, build... ...with our modeling and post-training engineers to improve... ...systems and translate research into practical engineering...TrainingLocal areaVisa sponsorship- Preference Model is seeking Research Engineers or Research Scientists to advance self-directed learning in AI. The role involves training and evaluating models within proprietary RL environments and optimizing ML infrastructure. Candidates will benefit from competitive...Training
- A leading AI research organization is seeking a Research Engineer to develop the next generation of training environments for safe agentic AI. This role requires strong software engineering skills and the ability to balance research and implementation while working collaboratively...Training
- Fundamental AI Research Institute Come join one of the only research... ...Learning and Agentic AI. Hiring for those experienced in LLM Post-Training / Agentic infrastructure, building RL environments, evaluation... ...Reasoning, Agents or AI infra Comfortable navigating ambiguity...Training
- Frontier Arc in San Francisco seeks a Research Engineer to push the core RL research agenda. You’ll own the end-to-end loop—from designing policies to training, evaluation, and interpretation—working on a proprietary simulator that aims to replace costly real-world data...Training
- ...AI Research Engineer (Robot Learning) San Francisco AI & Software In office... ...hands and cutting edge AI trained on human observations, bring... ...specialized fine-tuning and post-training. You will also be responsible... ...Experience with RL fine-tuning of generative models...TrainingFull timeWork at officeImmediate start
$180k - $215k
...The flagship product—an AI-driven, non-invasive... ...founding member of our new Agentic AI initiatives team,... ...and greenfield product engineering. You won't just be consuming... ...recruitment, hiring, training, relocation, promotion,... ...termination.Positions posted for Heartflow are not...TrainingLocal areaWorldwideRelocation- OpenAI is seeking an exceptional health AI researcher to build frontier capabilities that translate into real... ...The candidate will contribute to pretraining, RL/post-training, and evaluation, working with researchers, engineers, clinicians, and product teams to deliver...Training
- OpenAI is seeking an exceptional researcher to build frontier health capabilities and translate... .... The candidate should have strong ML/AI research depth, be a hands‑on builder,... ...track record in advancing pretraining, RL/post‑training, or health‑focused AI problems. #J-188...Training
- ...Origin Origin is building Physical AI for the built world - starting... ...from data collection and model training through edge deployment on Jetson AGX Orin. Every research project will have a deployment... ...least two of: imitation learning, RL, vision-language models, robot...Training
$197.3k - $225.1k
...Overview Lead AI Engineer (MLX, Agentic AI, Gen AI platform Services) At... ...functional team of engineers, research scientists, technical... ...including foundation model training, large language model inference... ...to pay at the time of this posting. Salaries for part-time roles...TrainingFull timePart timeLocal area$315k
As a Research Engineer or Research Scientist in Applied Finetuning, you will directly train the models we launch to the public via Claude.AI and our API. In this role, you will design and iterate on state... ...Complex shared codebases and RL infrastructure Authoring...TrainingWork at officeHome officeVisa sponsorshipRelocation package$180k - $220k
...Nimble Nimble is an AI robotics company building... ...commerce. We're training robot AGI to power a proprietary... ...of the world's best engineers and operators. If you... ...looking for an AI Robotics Research Engineer to help us... ...VLMs, multi-agent deep RL policies and more Develop...TrainingLocal areaImmediate startFlexible hoursWeekend work$110.7k - $379.2k
Position Summary Research Engineer — Post-Training & Small Language Models (SLMs), Healthcare AI Three hundred fifty million Americans... ...the reasoning models and agentic systems to rebuild how that system... ...using verifiable-reward RL — designing reward signals and...TrainingLocal areaVisa sponsorship- Together AI in San Francisco is seeking a Staff ML Systems Engineer to design and prototype algorithms, architectures, and scheduling for low-latency, high... ...latency and cost. You will also co-design RL and post-training pipelines, drive performance improvements, and...Training
- Scale is hiring a Staff Agent Post-Training MLRE in New York to build and scale an Agent RL training platform for enterprise use-cases. You will train state-of-the-art models on both internal and community research and deploy them to customers. Ideal candidates have 5+...Training
- ...experiment with AI systems for code... ...detection, and agentic security tools Implement research ideas end-to-end... ...publications, blog posts, or benchmarks... ...model evaluation, training, and... ...with product and engineering teams to integrate... ...evaluation pipelines, or RL-style...TrainingFull time
$200k - $350k
...Roam Job Posting Roam is an Applied AI lab building World Models... ...-Moonshot (post-training & agents,... ...technical founders, engineers that made 100+ games... .... Our current research spans:... ...difficulty adjustment Agentic codegen & diffusion... ...models, or RL systems. Strong...TrainingVisa sponsorshipRelocation package- Code Metal, based in San Francisco, is looking for an Applied AI Research Engineer to design and build advanced AI systems that assist military decision-making. This role combines AI research with practical applications, focusing on creating human-machine collaboration...
- A leading technology company is seeking a Machine Learning Engineer to drive innovation in AI technology. The role involves designing and implementing sophisticated AI systems in collaboration with a talented team. Ideal candidates should possess exceptional programming...
$227.2k - $284k
...develop reliable AI systems for the... ...of what agentic applications can... ...with applied ML research, design, and evaluation... ...Research Engineer, you will operate... ...could mean training and fine-tuning... ...online or offline RL — and validate... ...displayed on each job posting reflects the...TrainingFull time- ...interpretable, and steerable AI systems. We want AI to... ...group of committed researchers, engineers, policy experts, and... ...the role Anthropic's RL Data Platform team... ...turn raw feedback into training signal, and the tooling... ...expert review a long agentic transcript, flag the step...TrainingWork at officeVisa sponsorshipFlexible hours
$286.2k - $326.7k
...Overview Senior Director, AI Engineering -Agentic AI Platform(Remote Eligible... ...team of engineers, research scientists, technical program... ...including foundation model training, large language model inference... ...to pay at the time of this posting. Salaries for part-time roles...TrainingFull timePart timeLocal areaRemote work- Thinking Machines Lab Inc. in San Francisco, CA, seeks a post-training researcher to bridge raw model intelligence with practical, safe, and collaborative AI. You will write high-performance code and digest technical reports, blending theory with hands-on experimentation...Training
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Post-Training AI Research Engineer - RL & Agentic Infra. Be the first to apply!
- ai research engineer San Francisco, CA
- machine learning ai engineer San Francisco, CA
- ai developer San Francisco, CA
- senior ai engineer San Francisco, CA
- ai engineer San Francisco, CA
- ai ml engineer San Francisco, CA
- ai prompt engineer San Francisco, CA
- ai engineer remote San Francisco, CA
- research assistant engineering San Francisco, CA
- senior research engineer San Francisco, CA




