AI Research Scientist
San Francisco, CA · On-site
Designing and training systems using RLHF, RLAIF, and reward modeling approaches, applied to scientific hypothesis generation and evaluation. * Process reward models and verifiers: Developing fine ...
