Principal AI Research Scientist Post-Training Alignment Reinforcement Learning Autodesk AI Lab:...
Portland, OR · On-site
Rather than relying solely on human preference data, we can ground reinforcement learning in the ... Drive human-in-the-loop evaluation with high annotation quality and sound scientific methodology
