Post-Training Research Scientist
Baseten
Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-Edge models into production. We're growing quickly and recently raised our $1.5B Series F, led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE This role sits at the frontier of our research agenda. You will pursue open problems at the intersection of post-training methodology and performant inference, and then collaborate with research engineering to translate findings into production systems. A meaningful portion of your time will be dedicated to research that deepens our understanding of how models learn, alignment, and architectural efficiency — questions that may not have immediate product application. The remainder will be directed toward research that solves concrete problems for Baseten's platform and customers, who are the fastest growing AI companies in the world like Cursor, Lovable, and Notion. We are looking for someone with sharp research taste and genuine creative instinct for problem selection. Someone who can identify questions that matter, design clean experiments to answer them, and push the state of the art. The environment here is not theoretical, but rather research that can be validated with eager customers who are serving billions of tokens a second. RECENT RESEARCH Towards infinite context windows: neural KV cache compaction Dense, on-policy or both? Repeated kv cache for long-running agents Distillation without the dark – replicating black-box on-policy distillation on Baseten RESPONSIBILITIES Define and pursue a research agenda spanning both foundational and applied work, with the applied component connected to Baseten's platform and customer needs. Design and execute rigorous experiments, frequently at meaningful scale (multi-node, trillion parameter models). Work with customers to translate domain‑specific requirements into research problems, where relevant to your agenda. Publish at top venues (NeurIPS, ICML, ICLR) and establish Baseten's research presence. Collaborate with model performance and training infrastructure teams to bridge research findings and inference production systems. Mentor junior researchers and shape the technical direction of the research organization as it grows. PREFERRED QUALIFICATIONS Master’s or PhD research depth in machine learning, with first-author publications at top venues Demonstrated ability to move from theory through implementation to empirical results — not exclusively theoretical or exclusively engineering work Judgment about problem selection, the ability to distinguish research that advances a metric from research that changes how systems are built Willingness to operate in a startup environment where the majority of research informs product decisions, with timelines measured in months rather than years Background spanning multiple research areas (e.g., both interpretability and RL, or both systems and training methodology) Track record of open‑source contributions or community building in ML research OUR VIEW ON TALENT Many of the labs that exist today run a credentialist talent model. Concentrate the most already‑legible researchers, and assume the concentration compounds. The best researchers in this field are very often not yet legible. Research engineers who have spent years inside a production stack and developed insights no PhD program teaches; PhDs working on the wrong‑shaped problem at the right‑shaped lab; operators who have been close to real systems long enough to see things credentialed researchers have never had to see. If you don’t have the traditional qualifications and are doing exceptional work, we’d love to chat. BENEFITS Competitive compensation, including meaningful equity. 100% coverage of medical, dental, and vision insurance for employee and dependents Flexible PTO policy including company wide Winter Break (our offices are closed from Christmas Eve to New Year's Day!) Paid parental leave Fertility and family‑building stipend through Carrot Company‑facilitated 401(k) Exposure to a variety of ML startups, offering unparalleled learning and networking opportunities. At Baseten, we are committed to fostering a diverse and inclusive workplace. We provide equal employment opportunities to all employees and applicants without regard to race, color, religion, gender, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, or veteran status. We are an Equal Opportunity Employer and will consider qualified applicants with criminal histories in a manner consistent with applicable law (by example, the requirements of the San Francisco Fair Chance Ordinance, where applicable). #J-18808-Ljbffr Baseten
- We are Genmo, a research lab dedicated to building open, state-of-the-art models for video generation towards... ...overview:We are seeking an exceptional Research Scientist to join our team, focusing on alignment and post-training techniques for large-scale video generation...TrainingRelocation
$180.6k - $225.75k
...'s leading AI labs to provide high quality data and accelerate progress in GenAI research. We are looking for Research Scientists and Research Engineers with expertise in LLM post-training (SFT, RLHF, reward modeling) and evaluation. This role is on the evaluation pod within...TrainingFull time$216k - $270k
Scale Labs, Research Scientist — Frontier Risk EvaluationsAs the leading data and evaluation partner... .... The range displayed on each job posting reflects the minimum and maximum target... ...performance, and relevant education or training. Scale employees in eligible roles are...TrainingFull time$180.6k - $225.75k
...’s leading AI labs to provide high quality data and accelerate progress in GenAI research. We are looking for Research Scientists and Research Engineers with expertise in LLM post-training (SFT, RLHF, reward modeling). This role will focus on optimizing data curation and...TrainingFull time$120.7k - $238.6k
The OpportunityAdobe Research is looking for research scientists in Generative AI to join a world-class research... ...Experience on large-scale generative model training· Experience of working with large-... ...in Colorado (as listed on the job posting), the application window will...TrainingFull timeTemporary workLocal areaWorldwide$117.2k - $313.7k
....The ExperienceSalesforce AI Research is a global leader in Enterprise... ...in multimodal AI; trained state-of-the-art large language... ...for entrepreneurial Research Scientists who want to build, ship, and... ...processing.Core Modeling and Post-Training: Pre-training and post...TrainingFull timeWorldwide- ...Researcher Position As one of our researchers - you will answer the question: how to train and scale a model that can serve a web index? You: Have deep intuition on modern models and training. Like to argue how search, recommendations, and transformer models can...Training
$147k - $210k
Drive projects by defining key research questions.Design, implement, and evaluate experiments... ...:3 years of experience with training, evaluating, and interpreting large language... ...simulation.We are looking for a Research Scientist to develop cutting-edge social simulation...Training- ...workersResponsibilitiesWe are looking for an exceptional AI Research Scientist to join our growing team. In this role, you will be responsible... ...work; hands‑on experience with large‑scale model training, transformer architectures, reinforcement‑learning techniques...TrainingRemote workFlexible hours
- ...Research Scientist Engineering · Full-time · San Francisco; New York Our mission is to automate coding. The first step in our journey... ...Research Scientist Cursor is building the future of coding. We train frontier coding agents and scale RL on real user data to make...TrainingFull time
$160k - $220k
...enjoy multi-year runway.About the RoleWe’re looking for an AI Research Scientist to advance the methodological frontier of AI in healthcare.... ...strategy. Your work may include novel architectures, new training or evaluation techniques, long-horizon research bets, peer-reviewed...TrainingTemporary workWork at officeMonday to FridayMonday to Thursday- ...Pantograph is training general models that start by watching internet-scale video and end up on robots. We think the path to capable... ...fleet of affordable, durable robots. We're looking for research scientists who want to scale simple methods across the largest datasets...Training
- Overview Join our core R&D team building end-to-end automated research systems. Zochi publishes the first fully-AI generated A*... ...problems at the forefront of long-horizon agentic capabilities, post-training for open-ended goals, and environment development. Publish...Training
$225k - $300k
...Research Scientist About Latent Health Healthcare today is only truly personalized for two groups: those with wealth and access... ...on: Verifiable reinforcement learning at scale Mid-training and post-training of foundation models Novel objectives derived...TrainingWork at officeImmediate start$234.3k - $349k
...future of work with AI. About the roleAI research at WRITER isn't just about publishing... ...deployments in the world. As an AI research scientist, you'll be at the center of that work.... ...possible. The work you do here — on post-training, planning, multi-step reasoning, and...TrainingFull timeWork at officeLocal area- ...a world-class team of engineers, designers, marketers, sellers, researchers, and operational experts to achieve our mission. Job: As one of our Researchers - you will answer the question: how to train and scale a model that can serve a web index? You: Have deep...TrainingWork at officeVisa sponsorshipFlexible hours
- ...idler is a frontier data research lab. We build the evals and environments... ...labs use to measure and train their models. After raising... ...About the role As a Research Scientist at idler, you'll own measuring... ...labs - to design novel post-training recipes and data quality...TrainingWork at officeRelocation package
- ...Researcher Position at Hedra Hedra is building a world-class Physical AI research team to push the boundaries of action-conditioned... ...modeling for embodied systems Design novel architectures, training objectives, and evaluation frameworks for VLMs, VLAs, and world...TrainingWork at office
- ...Machine Learning Scientist Sesame believes in a future where computers are lifelike -... ...Learning Scientist at Sesame, you are a research-oriented person with experience in NLP,... ...architectures, data curation, model evaluation, training & inference infrastructure, research,...TrainingFull timeContract workFlexible hours
- ...AI David AI is the first audio data research company. We bring an R&D approach to data... ...and new use cases emerge, high-quality training data is the bottleneck. This is where David... .... About This Role As a Research Scientist at David AI you'll build cutting-edge speech...TrainingWork at office
$200k - $335k
...The role As a research scientist, you will design, implement, and optimize the large-scale training infrastructure that powers our frontier reinforcement learning stack.... ...Partner with researchers to bring frontier post-training capabilities into production deployments...TrainingFull timeWork at officeVisa sponsorshipRelocation package- ...Senior Research Scientist Sciforium is an AI infrastructure company developing next-generation multimodal AI models and a proprietary,... ...generative media, model architecture, optimization, and scalable training systems. You will work hands-on with modern ML frameworks,...TrainingFlexible hours
- ...physical goods as easily as they post online, and we're building... ...is Varun Jampani, a leading researcher who co-authored Dreambooth and... ..., was formerly a Principal Scientist at Amazon. Together, we're pioneering... ..., and author proprietary training paradigms when existing open-...Training
$204k - $259k
...Waymo AI Foundations Team Scientist Waymo is an autonomous driving technology company... ...initiate and foster collaborations with other research teams in Alphabet. AI Foundations areas... ...in Waymo's Foundation World Model post-training and evaluation Research and develop...TrainingTemporary work- ...About AfterQuery AfterQuery is an applied research lab curating data solutions for... ...our data works. You will design and run training experiments that isolate the impact of our... ...behavior. This includes SFT and RL-based post-training, where you'll measure how different...TrainingLocal areaShift work
$300k - $320k
...is a quickly growing group of committed researchers, engineers, policy experts, and... ...We're seeking an exceptional Research Scientist to join our Life Sciences team at Anthropic... ...capabilities on scientific tasks through post-training, evaluation design, and RL environment...TrainingWork at officeVisa sponsorshipFlexible hours$176k - $304k
...Research Scientist, Frontier Capabilities Cambridge, MA USA; San Francisco, CA USA Your impact at LILA We're building a talent... ...Our work sits at the intersection of large language models, post-training, and scientific reasoning, with the goal of enabling...TrainingFull timeWork at officeLocal areaFlexible hoursShift work$155k - $269k
...realistic, scalable, controllable, and efficient simulation. As a Research Scientist in World Models, you will develop algorithms and... ...rich generative priors for downstream planning, testing, and training. You will... - Conduct fundamental and applied research...TrainingFull timeWork at officeWork from homeFlexible hours$131.2k - $204.1k
Women’s Reproductive Health Research Scholar PositionPHYSICIAN SCIENTIST, Women’s Reproductive Health Research Scholar... ...year of postgraduate residency training in obstetrics-gynecologyHave... ...or PD/PI on other such awardsThe posted UC salary scales set the minimum pay...Training- OpenAI is seeking a Research Scientist to join our Reasoning and Planning team. You will push the boundaries of AI by developing novel architectures and training paradigms for multi-step reasoning, mathematical problem-solving, and complex planning in frontier models. You...Training
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Post-Training Research Scientist. Be the first to apply!
- materials scientist San Francisco, CA
- scientist assay development San Francisco, CA
- entry level research scientist San Francisco, CA
- health scientist San Francisco, CA
- quality control scientist San Francisco, CA
- deep learning scientist San Francisco, CA
- application scientist San Francisco, CA
- scientist antibody discovery San Francisco, CA
- senior analytical scientist San Francisco, CA
- decision scientist San Francisco, CA

