Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Principal AI Research Scientist Post-Training · Alignment · Reinforcement Learning Autodesk AI Lab: London · San Francisco · Toronto · Remote (US/CA/EU

Autodesk

Job Requisition ID #26WD98667Position OverviewAutodesk's domains — architecture, engineering, construction, manufacturing, media & entertainment — provide a distinctive research environment: rich structured data, long-horizon reasoning tasks, and real-world evaluation grounded in professional workflows. Uniquely, decades of investment in physics simulation engines, CAD kernels, and computational design tools give us something most labs don't have: high-fidelity, domain-grounded verifiers that can serve as reward signals for post-training. Rather than relying solely on human preference data, we can ground reinforcement learning in the laws of physics and the constraints of real engineering. These are exactly the kinds of challenges — and assets — that make post-training and alignment research here genuinely distinctive.We publish at NeurIPS, ICML, ICLR, CVPR, and SIGGRAPH. We collaborate with leading academic and industry labs. And we have a direct line from research advances to product impact at scale. This is not a role where research sits behind a wall from engineering — you will see your work matter.RespoinsibilitiesPost-training for model development — from RLHF and preference optimization to agentic systems and long-horizon reasoningDevelop novel algorithms that improve model reliability, controllability, and alignmentMake principled architectural decisions about when to address challenges at the pre-training, post-training, or system levelDesign and run experiments that shape model behavior, robustness, and reasoning qualityPartner with infrastructure teams to build scalable, reproducible post-training workflowsContribute to publications, patents, and Autodesk's external research visibilityDesign evaluation frameworks for long-horizon reasoning, tool use, agentic behavior, safety, and real-world workflow completionLead rigorous model analysis and interpretability effortsDrive human-in-the-loop evaluation with high annotation quality and sound scientific methodologyEstablish model readiness criteria and provide go/no-go recommendations for releasesCommunicate technical risks, limitations, and trade-offs clearly to leadershipMinimum RequirementsDeep hands-on expertise in reinforcement learning for foundation models, and fluency with post-training methods (RLHF, RLAIF, DPO, PPO, or adjacent approaches)Proven experience leading or mentoring technical research teams — whether in an academic lab, AI research organization, or industry settingStrong intuition for model behavior, alignment challenges, and post-training trade-offsExperience designing evaluation systems and thinking rigorously about what it means for a model to be readyAbility to communicate complex technical trade-offs clearly to both technical and non-technical audiencesA PhD or equivalent depth of industry research experience in ML, RL, AI, or a related fieldExperience at a frontier model lab or advanced applied AI organizationA strong publication record at leading ML or AI venuesBackground in alignment research, preference learning, or agentic AIExperience deploying or supporting production AI systemsFamiliarity with large-scale training infrastructure and compute trade-offsAt Autodesk, we're building a diverse workplace and an inclusive culture to give more people the chance to imagine, design, and make a better world. Autodesk is proud to be an equal opportunity employer and considers all qualified applicants for employment without regard to race, color, religion, age, sex, sexual orientation, gender, gender identity, national origin, disability, veteran status or any other legally protected characteristic. We also consider for employment all qualified applicants regardless of criminal histories, consistent with applicable law.Are you an existing contractor or consultant with Autodesk? Please search for open jobs and apply internally (not on this external site). If you have any questions or require support, contact Autodesk Careers.SummaryLocation: San Francisco, CA, USA; Massachusetts, USA - Remote; Boston, MA, USA; AMER - Canada - British Columbia - Vancouver - Bentall Centre; Vancouver, BC, CAN; Plano, TX, USA; Portland, OR, USA; Texas, USA - Remote; New York, USA - RemoteType:

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Principal AI Research Scientist Post-Training · Alignment · Reinforcement Learning Autodesk AI Lab: London · San Francisco · Toronto · Remote (US/CA/EU in San Francisco, CA vacancy
  •  ...distinctive research...  ...tools give us something most labs don't have...  ...signals for post-training. Rather...  ...ground reinforcement learning in the...  ...training and alignment research...  ..., and Autodesk's...  ...academic lab, AI research...  ...: San Francisco, CA, USA; Massachusetts...  ..., USA - Remote; Boston,... 
    Remote work
    Autodesk
    Training
    For contractors

    Autodesk

    Boston, MA
    1 day ago
  • $192.6k - $344.85k

     ...Research Lead / Principal Scientist & ManagerPost-Training · Alignment · Reinforcement LearningAutodesk AI Lab: London · San Francisco · Toronto · Remote (US/CA/EU)The...  ...still an open research problem.Autodesk touches...  ...evolving — and post-training...  ...learning in the laws... 
    Remote work
    Autodesk
    Training
    Full time
    For contractors

    Autodesk

    Boston, MA
    2 days ago
  • $163k - $292.82k

     ...generation machine learning platform.You...  ...of Agentic AI features across Autodesk products.A successful...  ..., ensuring alignment with...  ...based in the San Francisco Bay Area. Hybrid/Remote.ResponsibilitiesLead...  ...future? Join us!BenefitsFrom health...  ...Francisco, CA, USAType: Full... 
    Remote work
    Autodesk
    Full time
    For contractors

    Autodesk

    San Francisco, CA
    2 days ago
  • $55k - $151.47k

     ...machine learning engineering...  ...and reinforce professional...  ...People Tech & AI team you...  ..., and alignment with PwC...  ...specialized training and/or...  ...with data scientists and ML engineers...  ...20%Job Posting End...  ...Ordinance, the San Francisco Fair...  ...Rosemont; CA-...  ...Birmingham; US-Remote; AR-Fayetteville... 
    Remote work
    Training
    Full time
    Work experience placement
    H1b

    PwC

    Nashville, TN
    3 days ago
  • $73.5k - $212.28k

     ...and machine learning...  ...Uphold and reinforce professional...  ...People Tech & AI team you will...  ...training and/or progressively...  ...platforms aligned to PwC...  ...to 20%Job Posting End DateThe...  ...Ordinance, the San Francisco Fair...  ...Rosemont; CA-Sacramento...  ...Birmingham; US-Remote; AR-Fayetteville... 
    Remote work
    Training
    Full time
    Work experience placement
    H1b

    PwC

    Des Moines, IA
    3 days ago
  • $220k - $300k

    Sales - USA - San Francisco / New York City...  ...companies working with us daily, we’re...  ...you! WalkMe AI represents the...  ...You will help align customer AI strategies...  ...concept, both remotely and onsite,...  ...continuous learning and offer opportunities...  ...as: location, training, transferable... 
    Remote work
    Training
    Work experience placement
    Immediate start

    WalkMe

    San Francisco, CA
    4 days ago
  •  ...Senior Machine Learning Test EngineerLocation...  ...in the Research Enablement team...  ...competent in using Autodesk CAD software....  ...team, located in London, San Francisco, Toronto, and remotely. Autodesk is a...  ...gates for training and deployment...  ...engineering or QA for ML/AI systemsStrong... 
    Remote work
    Autodesk
    Training
    Full time
    For contractors
    Work at office

    Autodesk

    New York, NY
    1 day ago
  • $152k - $272.25k

     ...protection across Autodesk's...  ...familiarity with AI-assisted...  ...This role is remote-friendly...  ...be based in San Francisco, CA; Portland,...  ...Denver, CO; or Toronto, ON. Travel...  ...aligns to data minimization...  ...including training data exposure...  ...to come.Learn MoreAbout AutodeskWelcome...  ...? Join us!... 
    Remote work
    Autodesk
    Training
    Full time
    For contractors
    Work at office

    Autodesk

    Boston, MA
    3 days ago
  • $200k - $400k

     ...Machine Learning Engineer San Francisco, CA & New York, NY...  ...Goodfire is a research company...  ...and design AI systems. Our...  ...structures lets us steer what...  ...for training, evaluating...  ...that best aligns with your strengths...  ...frontier lab experience...  ...-wide remote week per month... 
    Remote work
    Training
    Work at office

    Goodfire

    San Francisco, CA
    2 days ago
  • $162k - $243.1k

     ...projects help us create...  ...accomplished Principal Structural Engineer...  ...life-long learning and self-...  ...modeling software (Autodesk Revit)...  ...Pursuant to the San Francisco Fair Chance...  ...in NYC & CA (Bay Area) &...  ...Full timeJob Posting: 27/07/2026...  ...benefits, job training,... 
    Autodesk
    Training
    Full time
    Temporary work
    Part time
    Casual work
    Work at office
    Local area

    Stantec

    San Francisco, CA
    2 days ago
  • $164k - $220k

     ...II - Android (Kotlin) San Francisco, CA or Remote (U.S.) We are looking...  ...Some experience with AI agentic development...  ...bugs Compensation The US total compensation range...  ...on each job posting reflects the approximate...  ...experience, and education/training. Please note that the... 
    Remote job
    Training
    Full time
    Local area

    Doximity, Inc.

    San Francisco, CA
    4 days ago
  •  ...ll join Didit Labs , our in-house AI lab, and...  ...Strong deep-learning fundamentals — PyTorch, training pipelines, and...  ...can move from research to a real, fast...  ...Barcelona or San Francisco — work shoulder...  ...process. About us Didit is infrastructure...  ...by an EU member-state... 
    Training
    Full time
    Live in
    Work at office

    Didit

    San Francisco, CA
    more than 2 months ago
  •  ...Territory Manager - San Francisco, CA role at Currax...  ...all required training programs....  ...through self-learning and active participation...  ...competition; Aligns work with...  ...0 2 weeks ago Principal Program...  ...Finance Operations (Remote, US) Oakland, CA $...  ...the help of AI. #J-18808-Ljbffr... 
    Remote work
    Training
    Full time
    Home office

    Currax Pharmaceuticals LLC

    San Francisco, CA
    2 days ago
  • $23 per hour

     ...Mission Bit is a San Francisco-based nonprofit...  ...science and AI education that...  ...courses, self-paced learning, workshops, and...  ...San Francisco, CA, and reports...  ...proud to provide training to all...  ...PM - 7:30 PM Remote: August 27, August...  ...Have values aligned with Mission Bit... 
    Remote work
    Training
    Contract work
    Temporary work
    Part time
    Internship
    Work at office
    Local area

    Mission Bit

    San Francisco, CA
    15 days ago
  • $190k - $300k

     ...meaningful AI doesn’t...  ...as a research project...  ...Stanford AI Lab, to the...  ...empower scientists,...  ...to help us redefine...  ...machine learning techniques...  ...transition to post-sales...  ...expectation alignment.Lead...  ...education or training....  ...LocationsNew York; San Francisco Bay Area...  ...Boston; Remote - US#LI-... 
    Remote work
    Training
    Local area

    Snorkel AI

    Redwood City, CA
    1 day ago
  • $220k - $253k

     ...applications, helping us answer...  ...Komodo’s Labs team builds...  ...— our AI-native product...  ...have: Learned the Marmot...  ..., drive alignment, and build...  ...#LI-Remote The pay...  ...for each job posting reflects a...  ..., relevant training and certifications...  ...policy.  San Francisco Bay Area and... 
    Remote work
    Training
    Full time
    For contractors
    Work experience placement
    Work at office
    Local area
    Flexible hours

    Komodo Health

    Remote
    10 hours ago
  • $193.3k - $261.5k

     ...AWS Machine Learning...  ...Generative AI on AWS. The...  ...in-class ML training performance...  ...like Snap, Autodesk, Amazon Alexa...  ...Annapurna Labs team is responsible...  ...engineering, research, and...  ...Austin, or Toronto.About the teamInclusive...  ...is reinforced within our...  ...benefits at .USA, CA, Cupertino... 
    Autodesk
    Training
    Internship
    Local area
    Work from home
    Relocation
    Flexible hours

    Amazon

    Cupertino, CA
    6 hours ago
  • $97.4k - $141.2k

     ...to projects help us create buildings that...  ...performs other CA tasks.Actively participates...  .... AutoCAD, Revit, Autodesk Construction Cloud...  ...FE) / engineer in training (EIT) or other...  ...to the San Francisco Fair Chance Ordinance...  ...YesSchedule: Full timeJob Posting: 11/03/2026 05:03:... 
    Autodesk
    Training
    Full time
    Temporary work
    Part time
    Casual work
    Work at office

    Stantec

    San Francisco, CA
    3 days ago
  • $140.4k - $195k

     ...Staff Machine Learning Engineer at...  ...: The AI & Machine Learning...  ...the vision, alignment, development...  ...learning, reinforcement learning,...  ...across the US, with preferred...  ...in San Francisco, CA (hybrid), New...  ...York City, NY (remote), and Seattle...  ...education or training.  Your recruiter... 
    Remote job
    Training
    Full time
    Work at office
    Local area
    3 days per week

    The Headspace

    New York, NY
    10 hours ago
  • $176k - $230k

     ...era, we seek AI-native...  ...hiring an AI Research Scientist (New Grad)...  ...code, and learn at scale inside...  ...AI and reinforcement learning, where...  ...for aligning and improving...  ...and curate training data pipelines...  ...experience in LLM post-training,...  ...a research lab...  ...can notify us if you believe... 
    Training
    Full time
    Flexible hours

    Snowflake

    Bellevue, WA
    20 hours ago
  • $60 per hour

    Accounting Professionals - AI Training We are looking for...  ...by AI systems align with professional standards...  ...expertise and machine learning. Qualifications and Requirements...  ...on ad‑hoc, remote assignments that fit around...  ...and Working Hours Researchers pay up to $60 per hour... 
    Remote work
    Training
    Hourly pay
    Work from home
    Flexible hours

    Prolific

    San Francisco, CA
    2 days ago
  • $70k - $140k

     ...environment.Join us in transforming...  ...looking for an AI-fluent...  ...pharmacovigilance and learn the regulatory...  ...practices.This is a remote, full-time...  ..., and customer training, acting as the...  ...Jersey; Boston, MA; San Francisco, CA; San Diego, CA;...  ..., Raleigh, and Toronto. We create... 
    Remote work
    Training
    Permanent employment
    Full time
    Freelance
    H1b
    Work at office
    Local area
    Work from home
    3 days per week

    Veeva Systems

    Boston, MA
    1 day ago
  • $150k

     ...Analytics Lab (Machine Learning Engineer)...  ...excellence in AI/Machine...  ...Lisbon and London. If you...  ...induction training that will...  ...interests are aligned...  ...across the US, EMEA and...  ...Chesterbrook, PA, San Francisco, Boston,...  ...fraudulent job postings. Emails...  ....com, @ca.bnpparibas... 
    Training
    Summer work
    Internship
    Summer internship

    BNP Paribas

    Jersey City, NJ
    10 hours ago
  •  ...enterprise AI company. We...  ...problems.We’re training and...  ...Each one of us is responsible...  ...is a team of researchers, engineers,...  ...headquartered in Toronto and San Francisco, with key offices in London, New York City...  ...we are remote-first and accept...  ...for Machine Learning Infrastructure... 
    Remote work
    Training
    Full time
    Work at office
    Local area
    Home office

    Cohere

    New York, NY
    4 days ago
  •  ...experiences—from AI and data...  ...perspectives. Join us as we shape...  ...hiring an AI Research Scientist, Recursive Self...  ...AI Safety and Reinforcement Learning focused on recursive...  ...their own training signals,...  ...appropriate; align internal narrative...  ...available here.This posting is for an... 
    Training
    Shift work

    AMD

    Santa Clara, CA
    6 hours ago
  • $125k

     ...While this is a remote position not...  ...of life.Check us out on LinkedIn...  ...adaptability to learn, and the ability...  ...least 1 physician training - potentially...  ...meaningful impact. In alignment with our...  ...exclusively for Principal-level roles and...  ...: San Francisco, California, United... 
    Remote work
    Training
    Full time
    H1b
    Work at office
    Local area
    Immediate start
    Flexible hours

    Medtronic

    San Jose, CA
    4 days ago
  •  ...Machine Learning Engineer (Operations) Location: South San Francisco CA (Hybrid, 3 days/week) (Not remote) Duration: Long term JD...  ...SageMaker (for model building, training, and deployment), EC2 (for...  ...foundation models in generative AI applications. Knowledge... 
    Training
    Full time
    3 days per week

    Esrhealthcare

    San Bruno, CA
    10 hours ago
  • $160k - $220k

     ...looking for an AI Research Scientist to advance...  ...agenda aligned with Sprinter...  ...architectures, new training or...  ...usually find us playing a board...  ...multimodal learning, clinical AI...  ...industry research labs, or research...  ...Francisco, CA; Menlo Park,...  ...Pacific Avenue , San Francisco, California... 
    Training
    Temporary work
    Work at office
    Monday to Friday
    Monday to Thursday

    Sprinter Health

    San Francisco, CA
    4 days ago
  • $136.5k - $276.5k

    HPE Labs - AI/ML Research Scientist IIIThis role has been...  ...Machine Learning research team...  ...in the San Francisco Bay Area. This...  ...spanning reinforcement learning for...  ...research outcomes align with...  ...3 years of post-PhD research...  ...education/training, and/or skill...  ...in the US can be found... 
    Training
    Full time
    Work experience placement
    Work at office
    Local area
    Immediate start

    Hewlett Packard Enterprise

    Milpitas, CA
    2 days ago
  • $234.3k - $349k

     ...orchestrate AI-powered...  ...hubs in San Francisco, New...  ...Chicago, and London, our...  ...to join us on our journey...  ...roleAI research at...  ...research scientist, you'll...  ...here — on post-training,...  ...tuning, reinforcement learning from human...  ...emerging alignment techniques...  ...Francisco, CA; New... 
    Training
    Full time
    Work at office
    Local area

    Writer

    San Francisco, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Principal AI Research Scientist Post-Training · Alignment · Reinforcement Learning Autodesk AI Lab: London · San Francisco · Toronto · Remote (US/CA/EU. Be the first to apply!