Principal AI Research Scientist Post-Training · Alignment · Reinforcement Learning Autodesk AI Lab: London · San Francisco · Toronto · Remote (US/CA/EU
Autodesk
Job Requisition ID #26WD98667Position OverviewAutodesk's domains — architecture, engineering, construction, manufacturing, media & entertainment — provide a distinctive research environment: rich structured data, long-horizon reasoning tasks, and real-world evaluation grounded in professional workflows. Uniquely, decades of investment in physics simulation engines, CAD kernels, and computational design tools give us something most labs don't have: high-fidelity, domain-grounded verifiers that can serve as reward signals for post-training. Rather than relying solely on human preference data, we can ground reinforcement learning in the laws of physics and the constraints of real engineering. These are exactly the kinds of challenges — and assets — that make post-training and alignment research here genuinely distinctive.We publish at NeurIPS, ICML, ICLR, CVPR, and SIGGRAPH. We collaborate with leading academic and industry labs. And we have a direct line from research advances to product impact at scale. This is not a role where research sits behind a wall from engineering — you will see your work matter.RespoinsibilitiesPost-training for model development — from RLHF and preference optimization to agentic systems and long-horizon reasoningDevelop novel algorithms that improve model reliability, controllability, and alignmentMake principled architectural decisions about when to address challenges at the pre-training, post-training, or system levelDesign and run experiments that shape model behavior, robustness, and reasoning qualityPartner with infrastructure teams to build scalable, reproducible post-training workflowsContribute to publications, patents, and Autodesk's external research visibilityDesign evaluation frameworks for long-horizon reasoning, tool use, agentic behavior, safety, and real-world workflow completionLead rigorous model analysis and interpretability effortsDrive human-in-the-loop evaluation with high annotation quality and sound scientific methodologyEstablish model readiness criteria and provide go/no-go recommendations for releasesCommunicate technical risks, limitations, and trade-offs clearly to leadershipMinimum RequirementsDeep hands-on expertise in reinforcement learning for foundation models, and fluency with post-training methods (RLHF, RLAIF, DPO, PPO, or adjacent approaches)Proven experience leading or mentoring technical research teams — whether in an academic lab, AI research organization, or industry settingStrong intuition for model behavior, alignment challenges, and post-training trade-offsExperience designing evaluation systems and thinking rigorously about what it means for a model to be readyAbility to communicate complex technical trade-offs clearly to both technical and non-technical audiencesA PhD or equivalent depth of industry research experience in ML, RL, AI, or a related fieldExperience at a frontier model lab or advanced applied AI organizationA strong publication record at leading ML or AI venuesBackground in alignment research, preference learning, or agentic AIExperience deploying or supporting production AI systemsFamiliarity with large-scale training infrastructure and compute trade-offsAt Autodesk, we're building a diverse workplace and an inclusive culture to give more people the chance to imagine, design, and make a better world. Autodesk is proud to be an equal opportunity employer and considers all qualified applicants for employment without regard to race, color, religion, age, sex, sexual orientation, gender, gender identity, national origin, disability, veteran status or any other legally protected characteristic. We also consider for employment all qualified applicants regardless of criminal histories, consistent with applicable law.Are you an existing contractor or consultant with Autodesk? Please search for open jobs and apply internally (not on this external site). If you have any questions or require support, contact Autodesk Careers.SummaryLocation: San Francisco, CA, USA; Massachusetts, USA - Remote; Boston, MA, USA; AMER - Canada - British Columbia - Vancouver - Bentall Centre; Vancouver, BC, CAN; Plano, TX, USA; Portland, OR, USA; Texas, USA - Remote; New York, USA - RemoteType:
$192.6k - $344.85k
...Research Lead / Principal Scientist & ManagerPost-Training · Alignment · Reinforcement LearningAutodesk AI Lab: London · San Francisco · Toronto · Remote (US/CA/EU)The... ...still an open research problem.Autodesk touches... ...evolving — and post-training... ...learning in the laws...Remote workAutodeskTrainingFull timeFor contractors$163k - $292.82k
...generation machine learning platform.You... ...of Agentic AI features across Autodesk products.A successful... ..., ensuring alignment with... ...based in the San Francisco Bay Area. Hybrid/Remote.ResponsibilitiesLead... ...future? Join us!BenefitsFrom health... ...Francisco, CA, USAType: Full...Remote workAutodeskFull timeFor contractors$55k - $151.47k
...machine learning engineering... ...and reinforce professional... ...People Tech & AI team you... ..., and alignment with PwC... ...specialized training and/or... ...with data scientists and ML engineers... ...20%Job Posting End... ...Ordinance, the San Francisco Fair... ...Rosemont; CA-... ...Birmingham; US-Remote; AR-Fayetteville...Remote workTrainingFull timeWork experience placementH1b$73.5k - $212.28k
...and machine learning... ...Uphold and reinforce professional... ...People Tech & AI team you will... ...training and/or progressively... ...platforms aligned to PwC... ...to 20%Job Posting End DateThe... ...Ordinance, the San Francisco Fair... ...Rosemont; CA-Sacramento... ...Birmingham; US-Remote; AR-Fayetteville...Remote workTrainingFull timeWork experience placementH1b$220k - $300k
Sales - USA - San Francisco / New York City... ...companies working with us daily, we’re... ...you! WalkMe AI represents the... ...You will help align customer AI strategies... ...concept, both remotely and onsite,... ...continuous learning and offer opportunities... ...as: location, training, transferable...Remote workTrainingWork experience placementImmediate start- ...Senior Machine Learning Test EngineerLocation... ...in the Research Enablement team... ...competent in using Autodesk CAD software.... ...team, located in London, San Francisco, Toronto, and remotely. Autodesk is a... ...gates for training and deployment... ...engineering or QA for ML/AI systemsStrong...Remote workAutodeskTrainingFull timeFor contractorsWork at office
$152k - $272.25k
...protection across Autodesk's... ...familiarity with AI-assisted... ...This role is remote-friendly... ...be based in San Francisco, CA; Portland,... ...Denver, CO; or Toronto, ON. Travel... ...aligns to data minimization... ...including training data exposure... ...to come.Learn MoreAbout AutodeskWelcome... ...? Join us!...Remote workAutodeskTrainingFull timeFor contractorsWork at office$200k - $400k
...Machine Learning Engineer San Francisco, CA & New York, NY... ...Goodfire is a research company... ...and design AI systems. Our... ...structures lets us steer what... ...for training, evaluating... ...that best aligns with your strengths... ...frontier lab experience... ...-wide remote week per month...Remote workTrainingWork at office$162k - $243.1k
...projects help us create... ...accomplished Principal Structural Engineer... ...life-long learning and self-... ...modeling software (Autodesk Revit)... ...Pursuant to the San Francisco Fair Chance... ...in NYC & CA (Bay Area) &... ...Full timeJob Posting: 27/07/2026... ...benefits, job training,...AutodeskTrainingFull timeTemporary workPart timeCasual workWork at officeLocal area- ...ll join Didit Labs , our in-house AI lab, and... ...Strong deep-learning fundamentals — PyTorch, training pipelines, and... ...can move from research to a real, fast... ...Barcelona or San Francisco — work shoulder... ...process. About us Didit is infrastructure... ...by an EU member-state...TrainingFull timeLive inWork at office
$164k - $220k
...II - Android (Kotlin) San Francisco, CA or Remote (U.S.) We are looking... ...Some experience with AI agentic development... ...bugs Compensation The US total compensation range... ...on each job posting reflects the approximate... ...experience, and education/training. Please note that the...Remote jobTrainingFull timeLocal area- ...Territory Manager - San Francisco, CA role at Currax... ...all required training programs.... ...through self-learning and active participation... ...competition; Aligns work with... ...0 2 weeks ago Principal Program... ...Finance Operations (Remote, US) Oakland, CA $... ...the help of AI. #J-18808-Ljbffr...Remote workTrainingFull timeHome office
$23 per hour
...Mission Bit is a San Francisco-based nonprofit... ...science and AI education that... ...courses, self-paced learning, workshops, and... ...San Francisco, CA, and reports... ...proud to provide training to all... ...PM - 7:30 PM Remote: August 27, August... ...Have values aligned with Mission Bit...Remote workTrainingContract workTemporary workPart timeInternshipWork at officeLocal area$190k - $300k
...meaningful AI doesn’t... ...as a research project... ...Stanford AI Lab, to the... ...empower scientists,... ...to help us redefine... ...machine learning techniques... ...transition to post-sales... ...expectation alignment.Lead... ...education or training.... ...LocationsNew York; San Francisco Bay Area... ...Boston; Remote - US#LI-...Remote workTrainingLocal area$220k - $253k
...applications, helping us answer... ...Komodo’s Labs team builds... ...— our AI-native product... ...have: Learned the Marmot... ..., drive alignment, and build... ...#LI-Remote The pay... ...for each job posting reflects a... ..., relevant training and certifications... ...policy. San Francisco Bay Area and...Remote workTrainingFull timeFor contractorsWork experience placementWork at officeLocal areaFlexible hours$193.3k - $261.5k
...AWS Machine Learning... ...Generative AI on AWS. The... ...in-class ML training performance... ...like Snap, Autodesk, Amazon Alexa... ...Annapurna Labs team is responsible... ...engineering, research, and... ...Austin, or Toronto.About the teamInclusive... ...is reinforced within our... ...benefits at .USA, CA, Cupertino...AutodeskTrainingInternshipLocal areaWork from homeRelocationFlexible hours$97.4k - $141.2k
...to projects help us create buildings that... ...performs other CA tasks.Actively participates... .... AutoCAD, Revit, Autodesk Construction Cloud... ...FE) / engineer in training (EIT) or other... ...to the San Francisco Fair Chance Ordinance... ...YesSchedule: Full timeJob Posting: 11/03/2026 05:03:...AutodeskTrainingFull timeTemporary workPart timeCasual workWork at office$140.4k - $195k
...Staff Machine Learning Engineer at... ...: The AI & Machine Learning... ...the vision, alignment, development... ...learning, reinforcement learning,... ...across the US, with preferred... ...in San Francisco, CA (hybrid), New... ...York City, NY (remote), and Seattle... ...education or training. Your recruiter...Remote jobTrainingFull timeWork at officeLocal area3 days per week$176k - $230k
...era, we seek AI-native... ...hiring an AI Research Scientist (New Grad)... ...code, and learn at scale inside... ...AI and reinforcement learning, where... ...for aligning and improving... ...and curate training data pipelines... ...experience in LLM post-training,... ...a research lab... ...can notify us if you believe...TrainingFull timeFlexible hours$70k - $140k
...environment.Join us in transforming... ...looking for an AI-fluent... ...pharmacovigilance and learn the regulatory... ...practices.This is a remote, full-time... ..., and customer training, acting as the... ...Jersey; Boston, MA; San Francisco, CA; San Diego, CA;... ..., Raleigh, and Toronto. We create...Remote workTrainingPermanent employmentFull timeFreelanceH1bWork at officeLocal areaWork from home3 days per week$60 per hour
Accounting Professionals - AI Training We are looking for... ...by AI systems align with professional standards... ...expertise and machine learning. Qualifications and Requirements... ...on ad‑hoc, remote assignments that fit around... ...and Working Hours Researchers pay up to $60 per hour...Remote workTrainingHourly payWork from homeFlexible hours$150k
...Analytics Lab (Machine Learning Engineer)... ...excellence in AI/Machine... ...Lisbon and London. If you... ...induction training that will... ...interests are aligned... ...across the US, EMEA and... ...Chesterbrook, PA, San Francisco, Boston,... ...fraudulent job postings. Emails... ....com, @ca.bnpparibas...TrainingSummer workInternshipSummer internship- ...experiences—from AI and data... ...perspectives. Join us as we shape... ...hiring an AI Research Scientist, Recursive Self... ...AI Safety and Reinforcement Learning focused on recursive... ...their own training signals,... ...appropriate; align internal narrative... ...available here.This posting is for an...TrainingShift work
$160k - $220k
...looking for an AI Research Scientist to advance... ...agenda aligned with Sprinter... ...architectures, new training or... ...usually find us playing a board... ...multimodal learning, clinical AI... ...industry research labs, or research... ...Francisco, CA; Menlo Park,... ...Pacific Avenue , San Francisco, California...TrainingTemporary workWork at officeMonday to FridayMonday to Thursday- ...enterprise AI company. We... ...problems.We’re training and... ...Each one of us is responsible... ...is a team of researchers, engineers,... ...headquartered in Toronto and San Francisco, with key offices in London, New York City... ...we are remote-first and accept... ...for Machine Learning Infrastructure...Remote workTrainingFull timeWork at officeLocal areaHome office
$125k
...While this is a remote position not... ...of life.Check us out on LinkedIn... ...adaptability to learn, and the ability... ...least 1 physician training - potentially... ...meaningful impact. In alignment with our... ...exclusively for Principal-level roles and... ...: San Francisco, California, United...Remote workTrainingFull timeH1bWork at officeLocal areaImmediate startFlexible hours- ...Machine Learning Engineer (Operations) Location: South San Francisco CA (Hybrid, 3 days/week) (Not remote) Duration: Long term JD... ...SageMaker (for model building, training, and deployment), EC2 (for... ...foundation models in generative AI applications. Knowledge...TrainingFull time3 days per week
$234.3k - $349k
...orchestrate AI-powered... ...hubs in San Francisco, New... ...Chicago, and London, our... ...to join us on our journey... ...roleAI research at... ...research scientist, you'll... ...here — on post-training,... ...tuning, reinforcement learning from human... ...emerging alignment techniques... ...Francisco, CA; New...TrainingFull timeWork at officeLocal area$136.5k - $276.5k
HPE Labs - AI/ML Research Scientist IIIThis role has been... ...Machine Learning research team... ...in the San Francisco Bay Area. This... ...spanning reinforcement learning for... ...research outcomes align with... ...3 years of post-PhD research... ...education/training, and/or skill... ...in the US can be found...TrainingFull timeWork experience placementWork at officeLocal areaImmediate start$120.5k - $276.5k
AI and Machine Learning Engineer This role has been designed as 'Hybrid... ...Computing, AI and Labs are a critical element... ...insight & innovation. Join us and redefine what’s... ...supervision in a semi-remote setting Additional... ...experience, education/training, and/or skill level....Remote workTrainingFull timeWork experience placementWork at officeLocal areaImmediate start2 days per week
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Principal AI Research Scientist Post-Training · Alignment · Reinforcement Learning Autodesk AI Lab: London · San Francisco · Toronto · Remote (US/CA/EU. Be the first to apply!
- ai scientist Plano, TX
- principal Plano, TX
- senior principal cloud computing engineer Plano, TX
- associate principal Plano, TX
- senior principal scientist Plano, TX
- principal cloud computing engineer Plano, TX
- autodesk Plano, TX
- remote clinical Plano, TX
- remote team lead Plano, TX
- remote auto claims adjuster Plano, TX



