Principal AI Research Scientist Post-Training · Alignment · Reinforcement Learning
Autodesk
Role Description
Autodesk's domains — architecture, engineering, construction, manufacturing, media & entertainment — provide a distinctive research environment: rich structured data, long-horizon reasoning tasks, and real-world evaluation grounded in professional workflows. Uniquely, decades of investment in physics simulation engines, CAD kernels, and computational design tools give us something most labs don't have: high-fidelity, domain-grounded verifiers that can serve as reward signals for post-training. Rather than relying solely on human preference data, we can ground reinforcement learning in the laws of physics and the constraints of real engineering. These are exactly the kinds of challenges — and assets — that make post-training and alignment research here genuinely distinctive.
We publish at NeurIPS, ICML, ICLR, CVPR, and SIGGRAPH. We collaborate with leading academic and industry labs. And we have a direct line from research advances to product impact at scale. This is not a role where research sits behind a wall from engineering — you will see your work matter.
- Post-training for model development — from RLHF and preference optimization to agentic systems and long-horizon reasoning
- Develop novel algorithms that improve model reliability, controllability, and alignment
- Make principled architectural decisions about when to address challenges at the pre-training, post-training, or system level
- Design and run experiments that shape model behavior, robustness, and reasoning quality
- Partner with infrastructure teams to build scalable, reproducible post-training workflows
- Contribute to publications, patents, and Autodesk's external research visibility
- Design evaluation frameworks for long-horizon reasoning, tool use, agentic behavior, safety, and real-world workflow completion
- Lead rigorous model analysis and interpretability efforts
- Drive human-in-the-loop evaluation with high annotation quality and sound scientific methodology
- Establish model readiness criteria and provide go/no-go recommendations for releases
- Communicate technical risks, limitations, and trade-offs clearly to leadership
Qualifications
- Deep hands-on expertise in reinforcement learning for foundation models, and fluency with post-training methods (RLHF, RLAIF, DPO, PPO, or adjacent approaches)
- Proven experience leading or mentoring technical research teams — whether in an academic lab, AI research organization, or industry setting
- Strong intuition for model behavior, alignment challenges, and post-training trade-offs
- Experience designing evaluation systems and thinking rigorously about what it means for a model to be ready
- Ability to communicate complex technical trade-offs clearly to both technical and non-technical audiences
- A PhD or equivalent depth of industry research experience in ML, RL, AI, or a related field
- Experience at a frontier model lab or advanced applied AI organization
- A strong publication record at leading ML or AI venues
- Background in alignment research, preference learning, or agentic AI
- Experience deploying or supporting production AI systems
- Familiarity with large-scale training infrastructure and compute trade-offs
Company Description
At Autodesk, we're building a diverse workplace and an inclusive culture to give more people the chance to imagine, design, and make a better world. Autodesk is proud to be an equal opportunity employer and considers all qualified applicants for employment without regard to race, color, religion, age, sex, sexual orientation, gender, gender identity, national origin, disability, veteran status or any other legally protected characteristic. We also consider for employment all qualified applicants regardless of criminal histories, consistent with applicable law.
- Figure is an AI robotics company developing autonomous general-purpose humanoid robots... ...are looking for a Helix AI Engineer, Reinforcement Learning to develop learning systems that... ...real-world and simulated environments Train policies that learn from interaction, feedback...TrainingFull timeWork at office
$99k - $225k
...Agentic Ai Machine Learning Engineer The Opportunity:... ...data science or data research in a professional or... ...academic environment, and training or deploying models... ...Learning (DL), and Reinforcement Learning (RL), and... ...facility frequently, in alignment with leadership...TrainingFull timeContract workPart timeWork at officeLocal areaRemote work$176.6k - $294.3k
...candidate for the Applied AI / ML scientist position leads the... ...Internal Medicine Research Unit (IMRU), translating... ...interface of machine learning, computational... ...efforts remain tightly aligned to real scientific needs... ...environments. Experience and/or training in cardiovascular,...TrainingPermanent employmentH1bLocal areaVisa sponsorshipWork visaRelocation package2 days per week- ...Helix AI Engineer, Robot Learning Figure is an AI robotics company developing autonomous general... .... Responsibilities Design, train, evaluate, and deploy learning-based... ...techniques including behavior cloning, reinforcement learning, and VLA reasoning Train...TrainingFull time
$150k - $230k
...powered by advanced AI, recommendation... ...hands-on Machine Learning Engineer to drive the post-training of our large language... ...emphasis on reinforcement learning (RL) . You... ...with post-training research and turn promising... ...failure modes in alignment/agent training....TrainingFull timeLocal areaWork from home- ...at the intersection of AI, biology, chemistry,... ...also carefully designed learning systems that can scale... ...on building and training those systems. The... ...owning outcomes from research through production.... ...Have Experience with reinforcement learning, fine-tuning,...TrainingFull timeRemote workFlexible hours
$100k - $150k
...Description We are looking for an AI Learning Systems Engineer to design, train, and deploy RL-based systems for... ...requires deep familiarity with modern reinforcement learning algorithms, simulation... .... The ideal candidate has both research depth and engineering pragmatism,...TrainingFull timeLocal areaImmediate startRemote work- Role Description AITP is looking for an experienced learning professional with deep AI fluency to support our growing managed service practice. You will design, develop, and deliver AI curriculum, training programs, and enablement content across enterprise AI platforms,...TrainingHourly payContract workWork at office
$125k - $150k
...a dynamic Data Scientist/ML Engineer to join... ...your background aligns with our... ...not limited to: Research, design, implement... ...and deploy Machine Learning algorithms for enterprise... ...vision, or reinforcement learning. Benefits... ...Disability ~ Training & Development ~...TrainingFull timeTemporary workWork experience placement$198.22k - $297.33k
...Description The AI Scientist will play a... ...Intelligence and Machine Learning solutions to support Alignment Healthcare’s... ...at the Principal level who are... ...AI (25%): ~Research, prototype, and... ...field. ~Training Required: ~... ...learning, NLP, reinforcement learning, and...TrainingFull timeRemote work$80k - $90k
Role Description The AI Learning Specialist will play a critical role in building the IRC’s organizational capacity to use AI tools effectively... ...Delivery & Facilitation ~Facilitate live virtual hands-on training sessions with generative AI tools and applications. ~...TrainingFull timeWork at officeImmediate start$149k - $350k
...iterating with AI. From idea to product... ...for applied scientists with a Machine Learning and Artificial Intelligence... ...and applied research in this area.... ...(SFT), Reinforcement Learning (RL), prompt... ...~ Experience training LLMs with Reinforcement... ...doesn’t align perfectly with the...TrainingFull timeTemporary workRemote workWork from home- ...AI/Machine Learning Engineer Founded in 2007, Initiate Government... ...and health outcomes research. Fine-tune models... ...lifecycle, from training to deployment, including... ..., clinicians, data scientists, and software developers... ...is actionable and aligned with federal healthcare...TrainingContract workTemporary workWork experience placementRemote workFlexible hours
- ...AI Scientist Senior II Hybrid role (3 days/... ...generative AI, machine learning, deep learning,... ..., Operations Research, Bioinformatics,... ...technical vision aligned with business strategy... ...-supervised, and reinforcement learning... ...with distributed training, data parallelism...TrainingWork experience placementWork at officeRelocation3 days per week
- ...Member of Technical Staff, Reinforcement Learning Research Our client is a well-funded, early-stage AI lab building a real-time, multimodal... .... About the Role Own RL and post-training for large-scale multimodal... ...in RL, post-training, alignment, or ML systems #J-18808-Ljbffr...TrainingInternshipRelocation packageShift work
$145.17k - $177.43k
...Description We are seeking a Machine Learning Research Scientist who will support the... ...techniques ~Conducting post-processing and validations of... ...performance computing resources to train large GeoAI models.... ...programming languages to develop AI algorithms in the PyTorch...TrainingFull timeRemote workRelocation packageFlexible hours$296k - $370k
Role Description We are seeking a seasoned Principal AI/ML Researcher and Engineer with deep expertise in Bayesian Learning, and Distributional Reinforcement Learning (RL) to lead the advanced research and development of cutting-edge intelligence AI models. These systems...Full time$168k - $211k
...Cambia's Applied AI Team is living... ...better. AI Scientists work with various... ...AI, machine learning, deep learning... ..., Operations Research, Bioinformatics... ...vision aligned with business... ...supervised, and reinforcement learning paradigms... ...with distributed training, data...TrainingWork experience placementWork at officeImmediate startRelocation3 days per week- ...Senior Learning and Development Specialist Duration: 1+ Month 12/31 Hard set end date... ...-learning courses, workshops, and other trainings. Familiarity with e-learning... ...deliver leadership development programs that align with company values, strategy, and priorities...TrainingRemote workFlexible hours
$100k - $150k
...is in the middle of an AI transformation. Our... ...Engineering running discovery, training teams to use AI in... .... Contribute to and reinforce the patterns the team... ...Close the loop. Feed learnings from deployed agents back... ...package. In alignment with pay transparency...TrainingFull timeWorldwide$160k - $170k
...Role As a Senior AI Engineer focused on... ...open-source deep learning models for production... ...for training and inference Apply... ...workloads Apply reinforcement learning techniques... ...to improve model alignment and task-specific... ...production and translate research ideas into scalable...TrainingFull time$115.54k - $128.26k
...runs on STACK.THE POSITION:The Learning & Development Manager with... ...responsible for curriculum development, training delivery, and learning... ...full-cycle training programs aligned with our wider business objectives... ...project management, AI, and time management skillsHigh...TrainingWork experience placementLocal areaFlexible hoursShift workNight shift$30 - $90 per hour
...Apply your domain expertise to help train and deploy next-generation AI systems. In this contractor role you will shape how models learn, reason, and perform by... ...scalability. Collaborate with data scientists, engineers, and researchers to translate complex business...TrainingHourly payContract workFor contractorsRemote work$110k - $164k
...Leveraging cutting edge AI-driven generative... ...-language models trained on real systems... ...hardware. You’ll learn spacecraft... ...dataset by collecting, aligning, and annotating text... ...in ML/AI through research, personal projects... ...-like systems) ~ Reinforcement learning for constrained...TrainingFull timeInternshipWorldwideWeekend work- ...About JazzX AI : Vision: Enterprises... ...with deep expertise in Reinforcement Learning (RL) to join our team... ...includes building scalable training architectures,... ...platform engineering, and research—to define architectural... ...and platform teams to align RL solutions with strategic...TrainingFull timeWorldwide
$155k - $175k
...Machine Learning Engineer - LLMs & Generative AI Truveta is the world’s first... ...is to enable researchers to find cures faster... ...models (LLMs), and reinforcement learning... ...foundational models trained on vast clinical data... ...) and cross-modal alignment techniques. Hands...TrainingFull timeFor contractorsVisa sponsorshipWork visaFlexible hours$150k - $210k
...physiological data with clinical research and expert knowledge... ...As a Senior Machine Learning Engineer on our Core... ...with data scientists and MLOps engineers... ...and product teams to align model development with... ...relevant education or training. In addition to...TrainingFull timeWork at officeRelocation- ...U.S. Intelligence Community. We are seeking to hire a AI/Machine Learning Engineer to our team! Role Overview: As an AI/ML Engineer... ...& Sick leave ~ Health insurance coverage ~ Career training ~ Performance bonus programs ~401K contribution & Employer...TrainingRemote jobFull timeWork experience placementWork at office
$95.5k - $181.7k
...deploy agentic AI systems... ...cutting-edge research into practical... ...research scientists and engineers... ...What You Will Learn:... ...reliability, and alignment. Collaboration... ...generative AI, reinforcement learning,... ..., education/training, and key skills... ...notice was posted. However,...TrainingTemporary workWork experience placementWork at officeRemote workFlexible hours- ...Overview The Learning Specialist I’s primary focus is... ...online, and blended) align with the department strategy... ...Assists with or leads training delivery (in-person... ...hires to skills that reinforce our culture. Teaches... ...recommendations. Conducts research in response to...TrainingWork at officeRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Principal AI Research Scientist Post-Training · Alignment · Reinforcement Learning. Be the first to apply!
- ai data scientist Remote
- ai scientist Remote
- principal architect Remote
- principal consultant Remote
- principal solutions consultant Remote
- senior principal cloud computing engineer Remote
- epic principal trainer Remote
- principal scientist Remote
- senior principal scientist Remote
- principal cloud computing engineer Remote










