Talent.com
Autodesk
Research Lead / Principal Scientist & Manager Post-Training · Alignment · Reinforcement LearningAutodesk • Toronto, ON, CAN
Research Lead / Principal Scientist & Manager Post-Training · Alignment · Reinforcement Learning

Research Lead / Principal Scientist & Manager Post-Training · Alignment · Reinforcement Learning

Autodesk • Toronto, ON, CAN
30+ days ago
Salary
CA$192,600.00–CA$344,850.00 yearly
Job type
  • Full-time
  • Remote
Job description

Job Requisition ID #

26WD94883

Research Lead / Principal Scientist & Manager

Post-Training · Alignment · Reinforcement Learning

Autodesk AI Lab: London · San Francisco · Toronto · Remote (US/CA/EU)

The Opportunity

Foundation models are reshaping how engineers, architects, and designers work-but training foundation models that are reliable, domain-capable systems is still an open research problem.

Autodesk touches more of the physical world than almost any other software company. The products we build are used to design skyscrapers, manufacture aircraft, and produce films. AI is now central to how those workflows are evolving — and post-training is the layer that makes the difference between a capable model and one that is dependable and robust in our customers’ high-precision domains.

As Research Lead for Post-Training & Alignment, you will own Autodesk's research strategy for transforming foundation models into systems that are reliable, aligned, and genuinely useful in complex, domain-specific workflows. This is a deeply technical leadership role — you will shape research direction, drive key architectural decisions, and remain close to the work.

You will lead a growing team of AI scientists while continuing to contribute directly to research: running experiments, developing novel algorithms, and publishing at top-tier venues.

This role reports to the Senior Director of AI Research within Autodesk AI Lab.

Why This Role

Unique research surface area

Autodesk's domains — architecture, engineering, construction, manufacturing, media & entertainment — provide a distinctive research environment: rich structured data, long-horizon reasoning tasks, and real-world evaluation grounded in professional workflows. Uniquely, decades of investment in physics simulation engines, CAD kernels, and computational design tools give us something most labs don't have: high-fidelity, domain-grounded verifiers that can serve as reward signals for post-training. Rather than relying solely on human preference data, we can ground reinforcement learning in the laws of physics and the constraints of real engineering. These are exactly the kinds of challenges — and assets — that make post-training and alignment research here genuinely distinctive.

Research-first, with real impact

We publish at NeurIPS, ICML, ICLR, CVPR, and SIGGRAPH. We collaborate with leading academic and industry labs. And we have a direct line from research advances to product impact at scale. This is not a role where research sits behind a wall from engineering — you will see your work matter.

What You Will Do

Research & Technical Leadership

  • Own post-training strategy for model development — from RLHF and preference optimization to agentic systems and long-horizon reasoning
  • Develop novel algorithms that improve model reliability, controllability, and alignment
  • Make principled architectural decisions about when to address challenges at the pre-training, post-training, or system level
  • Design and run experiments that shape model behavior, robustness, and reasoning quality
  • Partner with infrastructure teams to build scalable, reproducible post-training workflows
  • Contribute to publications, patents, and Autodesk's external research visibility

Evaluation & Model Quality

  • Design evaluation frameworks for long-horizon reasoning, tool use, agentic behavior, safety, and real-world workflow completion
  • Lead rigorous model analysis and interpretability efforts
  • Drive human-in-the-loop evaluation with high annotation quality and sound scientific methodology
  • Establish model readiness criteria and provide go/no-go recommendations for releases
  • Communicate technical risks, limitations, and trade-offs clearly to leadership

Team & Organizational Leadership

  • Manage, mentor, and grow a team of AI scientists
  • Set technical direction and research priorities across post-training and alignment initiatives
  • Foster a research culture grounded in scientific rigor, reproducibility, and fast iteration
  • Help recruit world-class talent across ML, RL, alignment, and foundation models
  • Partner closely with pre-training teams, infrastructure, product organizations, and other stakeholders
  • Translate research trade-offs into clear, decision-ready guidance for leadership

What We Are Looking For

We care about research judgment and outcomes, not credential checklists. Strong candidates will typically have:

  • Deep hands-on expertise in reinforcement learning for foundation models, and fluency with post-training methods (RLHF, RLAIF, DPO, PPO, or adjacent approaches)
  • Proven experience leading or mentoring technical research teams — whether in an academic lab, AI research organization, or industry setting
  • Strong intuition for model behavior, alignment challenges, and post-training trade-offs
  • Experience designing evaluation systems and thinking rigorously about what it means for a model to be ready
  • Ability to communicate complex technical trade-offs clearly to both technical and non-technical audiences
  • A PhD or equivalent depth of industry research experience in ML, RL, AI, or a related field

We also value, but do not require:

  • Experience at a frontier model lab or advanced applied AI organization
  • A strong publication record at leading ML or AI venues
  • Background in alignment research, preference learning, or agentic AI
  • Experience deploying or supporting production AI systems
  • Familiarity with large-scale training infrastructure and compute trade-offs

What Success Looks Like

In the first year, success means:

  • Post-trained models show measurable improvements in reliability, alignment, reasoning quality, and domain usefulness
  • Evaluation metrics and release criteria are trusted and adopted across teams
  • The team delivers high-quality research with practical impact — and team members are growing into stronger, more independent researchers
  • Leadership relies on your judgment for model readiness, technical direction, and risk assessment
  • Autodesk AI Lab advances its reputation as a serious contributor to frontier AI research

About Autodesk AI Lab

Autodesk AI Lab advances state-of-the-art research across generative AI, multimodal foundation models, reasoning systems, and human-AI collaboration. Our work has direct impact across the industries that shape the physical world. We are an active contributor to the global research community and collaborate closely with leading academic and industry labs.

At Autodesk, we are building a diverse workplace and an inclusive culture to give more people the chance to imagine, design, and make a better world. Autodesk is proud to be an equal opportunity employer and considers all qualified applicants for employment without regard to race, color, religion, age, sex, sexual orientation, gender identity, national origin, disability, veteran status, or any other legally protected characteristic.

Learn More

About Autodesk

Welcome to Autodesk! Amazing things are created every day with our software – from the greenest buildings and cleanest cars to the smartest factories and biggest hit movies. We help innovators turn their ideas into reality, transforming not only how things are made, but what can be made.

We take great pride in our culture here at Autodesk – it’s at the core of everything we do. Our culture guides the way we work and treat each other, informs how we connect with customers and partners, and defines how we show up in the world.

When you’re an Autodesker, you can do meaningful work that helps build a better world designed and made for all. Ready to shape the world and your future? Join us!

Benefits

From health and financial benefits to time away and everyday wellness, we give Autodeskers the best, so they can do their best work. Learn more about our benefits in the U.S. by visiting

Salary transparency

Salary is one part of Autodesk’s competitive compensation package. For U.S.-based roles, we expect a starting base salary between $192,600 and $344,850. Offers are based on the candidate’s experience and geographic location, and may exceed this range. In addition to base salaries, our compensation package may include annual cash bonuses, commissions for sales roles, stock grants, and a comprehensive benefits package.

Equal Employment Opportunity

Create a job alert for this search

Research Lead / Principal Scientist & Manager Post-Training · Alignment · Reinforcement Learning Autodesk AI Lab: London · San Francisco · Toronto · Remote (US/CA/EU) • Toronto, ON, CAN

Similar jobs

Senior Principal Researcher - AI for Science

Huawei CanadaMarkham, York Region, CA
Permanent

Senior Principal Researcher - AI for Science.Senior Principal Researcher - AI for Science.Be among the first 25 applicants.The Technology Planning and Cooperation Department promotes strategic inno... Show more

 • Promoted

Principal Scientist in AI and Hardware Co-Design

huaweicanadaMarkham, ON, CA
Full-time

Elevate data analytics capabilities as a Principal Scientist focused on AI and hardware co-design.Contribute to groundbreaking advancements in algorithm performance in a collaborative team.As part ... Show more

 • Promoted

Principal Ai/Ml Scientist - $128,000 - $238,000 A Year

TelusToronto, Canada
Full-time

Lead AI/ML initiatives, developing and deploying advanced models and Generative AI applications.Drive data-driven decision-making, mentor teams, and represent the company as a thought leader. Show more

 • Promoted

AI & Applied Research Leader

Thomson ReutersToronto, ON, CA
Full-time

A leading technology firm in Toronto is seeking a Manager of AI and Applied Research.This role involves managing a high-performing team focused on state-of-the-art research in information retrieval... Show more

 • Promoted

Ai Research Manager/Scientist, Model Alignment - $175,300 - $299,999 A Year - Remote

AutodeskToronto County, Canada
Remote
Full-time

Lead a team of AI scientists in post-training and model alignment research, contributing hands-on to transform foundation models into reliable, production-ready systems. Show more

 • Promoted

Ai & Applied Research Leader

Thomson ReutersToronto, Canada
Full-time

A leading technology firm in Toronto is seeking a Manager of AI and Applied Research.This role involves managing a high-performing team focused on state-of-the-art research in information retrieval... Show more

 • Promoted

Machine Learning Research Team Lead

ODAIAToronto, ON, CA
Full-time

RBC Borealis is looking for an enthusiastic Research Lead who is excited by the opportunity of being at the forefront of machine learning technology and working on extremely challenging problems in... Show more

 • Promoted

Senior Principal Researcher – AI Agent & Multimodal Interaction System

Huawei CanadaMarkham, Ontario, Canada
Permanent

Huawei Canada has an immediate permanent opening for a Senior Principal Researcher.The Huawei Human-Machine Interaction Lab unites global researchers, engineers, and designers to redefine human tec... Show more

 • Promoted

AI Multimodal Interaction Senior Principal Researcher

Huawei Technologies Canada Co., Ltd.Markham, Ontario, Canada
Full-time

Transform AI interactions as a Senior Principal Researcher at Huawei Canada.Focus on creating state-of-the-art multimodal systems that integrate voice, touch, and visual experiences in innovative w... Show more

 • Promoted

Senior Principal Researcher – Ai Agent & Multimodal Interaction System

Huawei Technologies Canada Co., Ltd.Markham, Canada
Permanent

Huawei Canada has an immediate permanent opening for a Senior Principal Researcher.About the team: The Huawei Human-Machine Interaction Lab unites global researchers, engineers, and designers to re... Show more

 • Promoted

Lead Researcher: Energy-Efficient AI Systems

Fujitsutoronto, on, Canada
Full-time

Take on the Senior Researcher role at the University of Toronto, exploring energy-efficient AI platforms.Collaborate globally to advance AI accelerator research and innovation.This position situate... Show more

 • Promoted

AI Lead

Prospect 33Toronto, ON, CA
Full-time

Gen AI Lead / Data Scientist (PhD) – Prospect 33 (P33.Work Model: In-House Leadership Role.Prospect 33 is expanding its Data, AI & Research Technology (DART) practice and is seeking an exceptional ... Show more

 • Promoted

Ai And Applied Research Leadership Role

PowerToFlyToronto, Canada
Full-time

Join Thomson Reuters Labs in Toronto as a Manager of AI and Applied Research.Lead a team of talented scientists and engineers to deliver top-tier AI solutions.As the Manager, you will foster a high... Show more

 • Promoted

Remote Lead Scientist & Analytical PM – Gene Therapy

Biolink360Toronto, ON, CA
Remote
Full-time

An innovative company dedicated to gene therapy is seeking an Analytical Project Manager to lead vital projects in the development of transformative therapies.This remote role offers the opportunit... Show more

 • Promoted

Principal Data Scientist-Gen AI, Machine Learning (10042)

Extreme Networkstoronto, on, Canada
Full-time

Principal Data Scientist – (Gen AI, Machine Learning).This is a greenfield opportunity to shape next‑gen networking experiences at the cutting edge of Generative AI, Machine Learning, Big Data, and... Show more

 • Promoted

ExaCare AI Systems Lead Role

ExaCare AItoronto, on, Canada
Full-time

Shape the future of AI at ExaCare AI as the Systems Lead.This hands-on position focuses on building AI systems to enhance company-wide efficiencies in a collaborative environment.As the AI Systems ... Show more

 • Promoted

AI and Applied Research Leadership Role

PowerToFlytoronto, on, Canada
Full-time

Join Thomson Reuters Labs in Toronto as a Manager of AI and Applied Research.Lead a team of talented scientists and engineers to deliver top-tier AI solutions.As the Manager, you will foster a high... Show more

 • Promoted

Principal Researcher - Systems & Networking - Microsoft Research - C$142,400 - C$257,500 A Year

MicrosoftEast York, Canada
Full-time

Seeking a Principal Researcher in Systems & Networking with AI expertise to innovate intelligent systems and drive research in collaboration with diverse teams. Show more

 • Promoted

Ai Multimodal Interaction Senior Principal Researcher

Huawei Technologies Canada Co., Ltd.Markham, Canada
Full-time

Transform AI interactions as a Senior Principal Researcher at Huawei Canada.Focus on creating state-of-the-art multimodal systems that integrate voice, touch, and visual experiences in innovative w... Show more

 • Promoted

Lead, Applied AI Research & Production

Refinitivtoronto, on, Canada
Full-time

A leading fintech organization based in Toronto is seeking an experienced Manager of Applied Research.The role involves managing a high-performing team focused on problem-solving using cutting-edge... Show more