Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Principal AI Research Scientist Post-Training · Alignment · Reinforcement Learning

Full-time

Autodesk

Role Description

Autodesk's domains — architecture, engineering, construction, manufacturing, media & entertainment — provide a distinctive research environment: rich structured data, long-horizon reasoning tasks, and real-world evaluation grounded in professional workflows. Uniquely, decades of investment in physics simulation engines, CAD kernels, and computational design tools give us something most labs don't have: high-fidelity, domain-grounded verifiers that can serve as reward signals for post-training. Rather than relying solely on human preference data, we can ground reinforcement learning in the laws of physics and the constraints of real engineering. These are exactly the kinds of challenges — and assets — that make post-training and alignment research here genuinely distinctive.

We publish at NeurIPS, ICML, ICLR, CVPR, and SIGGRAPH. We collaborate with leading academic and industry labs. And we have a direct line from research advances to product impact at scale. This is not a role where research sits behind a wall from engineering — you will see your work matter.

  • Post-training for model development — from RLHF and preference optimization to agentic systems and long-horizon reasoning
  • Develop novel algorithms that improve model reliability, controllability, and alignment
  • Make principled architectural decisions about when to address challenges at the pre-training, post-training, or system level
  • Design and run experiments that shape model behavior, robustness, and reasoning quality
  • Partner with infrastructure teams to build scalable, reproducible post-training workflows
  • Contribute to publications, patents, and Autodesk's external research visibility
  • Design evaluation frameworks for long-horizon reasoning, tool use, agentic behavior, safety, and real-world workflow completion
  • Lead rigorous model analysis and interpretability efforts
  • Drive human-in-the-loop evaluation with high annotation quality and sound scientific methodology
  • Establish model readiness criteria and provide go/no-go recommendations for releases
  • Communicate technical risks, limitations, and trade-offs clearly to leadership

Qualifications

  • Deep hands-on expertise in reinforcement learning for foundation models, and fluency with post-training methods (RLHF, RLAIF, DPO, PPO, or adjacent approaches)
  • Proven experience leading or mentoring technical research teams — whether in an academic lab, AI research organization, or industry setting
  • Strong intuition for model behavior, alignment challenges, and post-training trade-offs
  • Experience designing evaluation systems and thinking rigorously about what it means for a model to be ready
  • Ability to communicate complex technical trade-offs clearly to both technical and non-technical audiences
  • A PhD or equivalent depth of industry research experience in ML, RL, AI, or a related field
  • Experience at a frontier model lab or advanced applied AI organization
  • A strong publication record at leading ML or AI venues
  • Background in alignment research, preference learning, or agentic AI
  • Experience deploying or supporting production AI systems
  • Familiarity with large-scale training infrastructure and compute trade-offs

Company Description

At Autodesk, we're building a diverse workplace and an inclusive culture to give more people the chance to imagine, design, and make a better world. Autodesk is proud to be an equal opportunity employer and considers all qualified applicants for employment without regard to race, color, religion, age, sex, sexual orientation, gender, gender identity, national origin, disability, veteran status or any other legally protected characteristic. We also consider for employment all qualified applicants regardless of criminal histories, consistent with applicable law.

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Principal AI Research Scientist Post-Training · Alignment · Reinforcement Learning in Remote vacancy
  • Figure is an AI robotics company developing autonomous general-purpose humanoid robots...  ...are looking for a Helix AI Engineer, Reinforcement Learning to develop learning systems that...  ...real-world and simulated environments Train policies that learn from interaction, feedback... 
    Training
    Full time
    Work at office

    Figure

    Remote
    13 hours ago
  • $99k - $225k

     ...Agentic Ai Machine Learning Engineer The Opportunity:...  ...data science or data research in a professional or...  ...academic environment, and training or deploying models...  ...Learning (DL), and Reinforcement Learning (RL), and...  ...facility frequently, in alignment with leadership... 
    Training
    Full time
    Contract work
    Part time
    Work at office
    Local area
    Remote work

    Booz Allen Hamilton

    Washington DC
    1 day ago
  • $176.6k - $294.3k

     ...candidate for the Applied AI / ML scientist position leads the...  ...Internal Medicine Research Unit (IMRU), translating...  ...interface of machine learning, computational...  ...efforts remain tightly aligned to real scientific needs...  ...environments. Experience and/or training in cardiovascular,... 
    Training
    Permanent employment
    H1b
    Local area
    Visa sponsorship
    Work visa
    Relocation package
    2 days per week

    Pfizer

    Cambridge, MA
    2 days ago
  •  ...Helix AI Engineer, Robot Learning Figure is an AI robotics company developing autonomous general...  .... Responsibilities Design, train, evaluate, and deploy learning-based...  ...techniques including behavior cloning, reinforcement learning, and VLA reasoning Train... 
    Training
    Full time

    Figure

    Remote
    13 hours ago
  • $150k - $230k

     ...powered by advanced AI, recommendation...  ...hands-on Machine Learning Engineer to drive the post-training of our large language...  ...emphasis on reinforcement learning (RL) . You...  ...with post-training research and turn promising...  ...failure modes in alignment/agent training.... 
    Training
    Full time
    Local area
    Work from home

    News Break

    Remote
    13 hours ago
  •  ...at the intersection of AI, biology, chemistry,...  ...also carefully designed learning systems that can scale...  ...on building and training those systems. The...  ...owning outcomes from research through production....  ...Have Experience with reinforcement learning, fine-tuning,... 
    Training
    Full time
    Remote work
    Flexible hours

    Absentia Labs

    San Francisco, CA
    13 hours ago
  • $100k - $150k

     ...Description We are looking for an AI Learning Systems Engineer to design, train, and deploy RL-based systems for...  ...requires deep familiarity with modern reinforcement learning algorithms, simulation...  .... The ideal candidate has both research depth and engineering pragmatism,... 
    Training
    Full time
    Local area
    Immediate start
    Remote work

    Bright Vision Technologies

    Remote
    13 hours ago
  • Role Description AITP is looking for an experienced learning professional with deep AI fluency to support our growing managed service practice. You will design, develop, and deliver AI curriculum, training programs, and enablement content across enterprise AI platforms,... 
    Training
    Hourly pay
    Contract work
    Work at office

    AI Technology Partners

    Remote
    1 day ago
  • $125k - $150k

     ...a dynamic  Data Scientist/ML Engineer  to join...  ...your background aligns with our...  ...not limited to: Research, design, implement...  ...and deploy Machine Learning algorithms for enterprise...  ...vision, or reinforcement learning. Benefits...  ...Disability ~ Training & Development ~... 
    Training
    Full time
    Temporary work
    Work experience placement

    Cathexis

    Remote
    13 hours ago
  • $198.22k - $297.33k

     ...Description The AI Scientist will play a...  ...Intelligence and Machine Learning solutions to support Alignment Healthcare’s...  ...at the Principal level who are...  ...AI (25%): ~Research, prototype, and...  ...field. ~Training Required: ~...  ...learning, NLP, reinforcement learning, and... 
    Training
    Full time
    Remote work

    Alignment Health

    Remote
    7 days ago
  • $80k - $90k

    Role Description The AI Learning Specialist will play a critical role in building the IRC’s organizational capacity to use AI tools effectively...  ...Delivery & Facilitation ~Facilitate live virtual hands-on training sessions with generative AI tools and applications. ~... 
    Training
    Full time
    Work at office
    Immediate start

    International Rescue Committee

    Remote
    7 days ago
  • $149k - $350k

     ...iterating with AI. From idea to product...  ...for applied scientists with a Machine Learning and Artificial Intelligence...  ...and applied research in this area....  ...(SFT), Reinforcement Learning (RL), prompt...  ...~ Experience training LLMs with Reinforcement...  ...doesn’t align perfectly with the... 
    Training
    Full time
    Temporary work
    Remote work
    Work from home

    Figma

    New York, NY
    2 days ago
  •  ...AI/Machine Learning Engineer Founded in 2007, Initiate Government...  ...and health outcomes research. Fine-tune models...  ...lifecycle, from training to deployment, including...  ..., clinicians, data scientists, and software developers...  ...is actionable and aligned with federal healthcare... 
    Training
    Contract work
    Temporary work
    Work experience placement
    Remote work
    Flexible hours

    Initiate Government Solutions

    United States
    4 days ago
  •  ...AI Scientist Senior II Hybrid role (3 days/...  ...generative AI, machine learning, deep learning,...  ..., Operations Research, Bioinformatics,...  ...technical vision aligned with business strategy...  ...-supervised, and reinforcement learning...  ...with distributed training, data parallelism... 
    Training
    Work experience placement
    Work at office
    Relocation
    3 days per week

    Cambia Health Solutions

    Lewiston, ID
    4 days ago
  •  ...Member of Technical Staff, Reinforcement Learning Research Our client is a well-funded, early-stage AI lab building a real-time, multimodal...  .... About the Role Own RL and post-training for large-scale multimodal...  ...in RL, post-training, alignment, or ML systems #J-18808-Ljbffr... 
    Training
    Internship
    Relocation package
    Shift work

    RecruitSeq

    Seattle, WA
    2 days ago
  • $145.17k - $177.43k

     ...Description We are seeking a Machine Learning Research Scientist who will support the...  ...techniques ~Conducting post-processing and validations of...  ...performance computing resources to train large GeoAI models....  ...programming languages to develop AI algorithms in the PyTorch... 
    Training
    Full time
    Remote work
    Relocation package
    Flexible hours

    Oak Ridge National Laboratory

    Remote
    6 days ago
  • $296k - $370k

    Role Description We are seeking a seasoned Principal AI/ML Researcher and Engineer with deep expertise in Bayesian Learning, and Distributional Reinforcement Learning (RL) to lead the advanced research and development of cutting-edge intelligence AI models. These systems... 
    Full time

    Airbnb

    Remote
    1 day ago
  • $168k - $211k

     ...Cambia's Applied AI Team is living...  ...better. AI Scientists work with various...  ...AI, machine learning, deep learning...  ..., Operations Research, Bioinformatics...  ...vision aligned with business...  ...supervised, and reinforcement learning paradigms...  ...with distributed training, data... 
    Training
    Work experience placement
    Work at office
    Immediate start
    Relocation
    3 days per week

    Cambia Health Solutions

    Boise, ID
    13 hours ago
  •  ...Senior Learning and Development Specialist Duration: 1+ Month 12/31 Hard set end date...  ...-learning courses, workshops, and other trainings. Familiarity with e-learning...  ...deliver leadership development programs that align with company values, strategy, and priorities... 
    Training
    Remote work
    Flexible hours

    eTeam

    United States
    2 days ago
  • $100k - $150k

     ...is in the middle of an AI transformation. Our...  ...Engineering running discovery, training teams to use AI in...  .... Contribute to and reinforce the patterns the team...  ...Close the loop. Feed learnings from deployed agents back...  ...package.   In alignment with pay transparency... 
    Training
    Full time
    Worldwide

    Pcs Wireless Global

    Remote
    13 hours ago
  • $160k - $170k

     ...Role As a Senior AI Engineer focused on...  ...open-source deep learning models for production...  ...for training and inference Apply...  ...workloads Apply reinforcement learning techniques...  ...to improve model alignment and task-specific...  ...production and translate research ideas into scalable... 
    Training
    Full time

    Octus

    Remote
    13 hours ago
  • $115.54k - $128.26k

     ...runs on STACK.THE POSITION:The Learning & Development Manager with...  ...responsible for curriculum development, training delivery, and learning...  ...full-cycle training programs aligned with our wider business objectives...  ...project management, AI, and time management skillsHigh... 
    Training
    Work experience placement
    Local area
    Flexible hours
    Shift work
    Night shift

    STACK Infrastructure US

    Denver, CO
    4 days ago
  • $30 - $90 per hour

     ...Apply your domain expertise to help train and deploy next-generation AI systems. In this contractor role you will shape how models learn, reason, and perform by...  ...scalability. Collaborate with data scientists, engineers, and researchers to translate complex business... 
    Training
    Hourly pay
    Contract work
    For contractors
    Remote work

    SaidGig

    United States
    13 hours ago
  • $110k - $164k

     ...Leveraging cutting edge AI-driven generative...  ...-language models trained on real systems...  ...hardware. You’ll learn spacecraft...  ...dataset by collecting, aligning, and annotating text...  ...in ML/AI through research, personal projects...  ...-like systems) ~ Reinforcement learning for constrained... 
    Training
    Full time
    Internship
    Worldwide
    Weekend work

    Oligo Space

    Remote
    13 hours ago
  •  ...About JazzX AI :   Vision: Enterprises...  ...with deep expertise in Reinforcement Learning (RL) to join our team...  ...includes building scalable training architectures,...  ...platform engineering, and research—to define architectural...  ...and platform teams to align RL solutions with strategic... 
    Training
    Full time
    Worldwide

    Jazzx Ai

    Remote
    13 hours ago
  • $155k - $175k

     ...Machine Learning Engineer - LLMs & Generative AI Truveta is the world’s first...  ...is to enable researchers to find cures faster...  ...models (LLMs), and reinforcement learning...  ...foundational models trained on vast clinical data...  ...) and cross-modal alignment techniques. Hands... 
    Training
    Full time
    For contractors
    Visa sponsorship
    Work visa
    Flexible hours

    Truveta

    Remote
    13 hours ago
  • $150k - $210k

     ...physiological data with clinical research and expert knowledge...  ...As a Senior Machine Learning Engineer on our Core...  ...with data scientists and MLOps engineers...  ...and product teams to align model development with...  ...relevant education or training.    In addition to... 
    Training
    Full time
    Work at office
    Relocation

    Whoop

    Remote
    13 hours ago
  •  ...U.S. Intelligence Community.   We are seeking to hire a AI/Machine Learning Engineer to our team! Role Overview: As an AI/ML Engineer...  ...& Sick leave ~ Health insurance coverage ~ Career training ~ Performance bonus programs ~401K contribution & Employer... 
    Training
    Remote job
    Full time
    Work experience placement
    Work at office

    Cybermedia Technologies

    Remote
    13 hours ago
  • $95.5k - $181.7k

     ...deploy agentic AI systems...  ...cutting-edge research into practical...  ...research scientists and engineers...  ...What You Will Learn:...  ...reliability, and alignment. Collaboration...  ...generative AI, reinforcement learning,...  ..., education/training, and key skills...  ...notice was posted. However,... 
    Training
    Temporary work
    Work experience placement
    Work at office
    Remote work
    Flexible hours

    Raytheon

    Arlington, VA
    4 days ago
  •  ...Overview The Learning Specialist I’s primary focus is...  ...online, and blended) align with the department strategy...  ...Assists with or leads training delivery (in-person...  ...hires to skills that reinforce our culture. Teaches...  ...recommendations. Conducts research in response to... 
    Training
    Work at office
    Remote work

    Stamford Health

    Stamford, CT
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Principal AI Research Scientist Post-Training · Alignment · Reinforcement Learning. Be the first to apply!