Principal AI Research Scientist Post-Training · Alignment · Reinforcement Learning Autodesk AI [...]
Principal AI Research Scientist Post-Training · Alignment · Reinforcement Learning Autodesk AI [...]
Posted yesterday
SENIORITY
Manager
About the role
- Deep hands-on expertise in reinforcement learning for foundation models, and fluency with post-training methods (RLHF, RLAIF, DPO, PPO, or adjacent approaches)
- Proven experience leading or mentoring technical research teams — whether in an academic lab, AI research organization, or industry setting
- Strong intuition for model behavior, alignment challenges, and post-training trade-offs
- Experience designing evaluation systems and thinking rigorously about what it means for a model to be ready
- Ability to communicate complex technical trade-offs clearly to both technical and non-technical audiencesA PhD or equivalent depth of industry research experience in ML, RL, AI, or a related field
- Experience at a frontier model lab or advanced applied AI organization
- A strong publication record at leading ML or AI venues
- Background in alignment research, preference learning, or agentic AIExperience deploying or supporting production AI systems
- Familiarity with large-scale training infrastructure and compute trade-offs
Before you apply
Applying takes about a minute. These four things decide how fast it moves after that.
Your profile is current
It's what we read first. Occupations, seniority and locations matter more than a long history.
Two examples you can talk through
Not a portfolio — just two pieces of work where you can explain the decisions and what you'd change.
A number in mind
What you're on now and what would make you move. We negotiate better when we know both.
Your notice period
Employers plan around it, and it's the question that stalls offers most often.
Once you apply, someone reads it and calls you before anything reaches the employer — usually within two working days.
More like this
