LLM Research Engineer
$90 - $121.86/hr
Knowledge of reinforcement learning and RLHF (Reinforcement Learning with Human Feedback). Compensation: $90 - $121.86 per hour ID#: 36408719
Quick apply
$90 - $121.86/hr
Knowledge of reinforcement learning and RLHF (Reinforcement Learning with Human Feedback). Compensation: $90 - $121.86 per hour ID#: 36408719
Quick apply
$90 - $121.86/hr
Knowledge of reinforcement learning and RLHF (Reinforcement Learning with Human Feedback). Compensation: $90 - $121.86 per hour ID#: 36408719
... reinforcement learning from human feedback (RLHF) and fine-tuning. • Collaborate with engineering and product teams to translate safety research into practical, scalable solutions and best ...
... reinforcement learning from human feedback (RLHF) and fine-tuning. • Collaborate with engineering and product teams to translate safety research into practical, scalable solutions and best ...
Design and implement supervised fine-tuning and reinforcement learning from human feedback (RLHF ... Collaborate with cross-functional teams to integrate alignment improvements into our production ...
Design and implement supervised fine-tuning and reinforcement learning from human feedback (RLHF ... Collaborate with cross-functional teams to integrate alignment improvements into our production ...
Design and implement supervised fine-tuning and reinforcement learning from human feedback (RLHF ... Collaborate with cross-functional teams to integrate alignment improvements into our production ...
Quick apply
Design and implement supervised fine-tuning and reinforcement learning from human feedback (RLHF ... Collaborate with cross-functional teams to integrate alignment improvements into our production ...
Sunnyvale, CA · On-site
... Reinforcement Learning with Human Feedback (RLHF). • Excellent data analysis skills. Company : Official account of Mohamed bin Zayed University of Artificial Intelligence. Dedicated to research ...
Sunnyvale, CA · On-site
... Reinforcement Learning with Human Feedback (RLHF). • Excellent data analysis skills. Company : Official account of Mohamed bin Zayed University of Artificial Intelligence. Dedicated to research ...
Sunnyvale, CA · On-site
... Reinforcement Learning with Human Feedback (RLHF). • Excellent data analysis skills. Company : Official account of Mohamed bin Zayed University of Artificial Intelligence. Dedicated to research ...
Sunnyvale, CA · On-site
... Reinforcement Learning with Human Feedback (RLHF). • Excellent data analysis skills. Company : Official account of Mohamed bin Zayed University of Artificial Intelligence. Dedicated to research ...
Figure is an AI Robotics company focused on developing autonomous humanoid robots with human-level intelligence. They are seeking a Reinforcement Learning Engineer to develop, train, deploy, and ...
Figure is an AI Robotics company focused on developing autonomous humanoid robots with human-level intelligence. They are seeking a Reinforcement Learning Engineer to develop, train, deploy, and ...
Hands-on experience with LLM algorithms, such as Supervised Fine-Tuning (SFT) and Reinforcement Learning with Human Feedback (RLHF). * Excellent data analysis skills. Visa Sponsorship This position ...
Quick apply
Hands-on experience with LLM algorithms, such as Supervised Fine-Tuning (SFT) and Reinforcement Learning with Human Feedback (RLHF). * Excellent data analysis skills. Visa Sponsorship This position ...
Outerport is seeking a Member of Technical Staff with expertise in reinforcement learning. The role involves training reinforcement learning-based LLMs for applications in materials science and ...
Outerport is seeking a Member of Technical Staff with expertise in reinforcement learning. The role involves training reinforcement learning-based LLMs for applications in materials science and ...
Figure is an AI Robotics company focused on developing autonomous general-purpose humanoid robots with human-level intelligence. They are seeking a Staff Reinforcement Learning Engineer to develop ...
Figure is an AI Robotics company focused on developing autonomous general-purpose humanoid robots with human-level intelligence. They are seeking a Staff Reinforcement Learning Engineer to develop ...
You'll design and implement advanced methods to align human feedback with the training of cutting-edge AI models, including techniques like Reinforcement Learning from Human Feedback (RLHF) , Direct ...
Quick apply
You'll design and implement advanced methods to align human feedback with the training of cutting-edge AI models, including techniques like Reinforcement Learning from Human Feedback (RLHF) , Direct ...
San Jose, CA · On-site
$150K/yr
The goal of the company is to ship humanoid robots with human level intelligence. Its robots are ... We are looking for a Reinforcement Learning Engineer to develop, train, deploy, and evaluate ...
San Jose, CA · On-site
$150K/yr
The goal of the company is to ship humanoid robots with human level intelligence. Its robots are ... We are looking for a Reinforcement Learning Engineer to develop, train, deploy, and evaluate ...
San Jose, CA · Hybrid
$150K/yr
The goal of the company is to ship humanoid robots with human level intelligence. Its robots are ... We are looking for a Reinforcement Learning Engineer to develop, train, deploy, and evaluate ...
San Jose, CA · Hybrid
$150K/yr
The goal of the company is to ship humanoid robots with human level intelligence. Its robots are ... We are looking for a Reinforcement Learning Engineer to develop, train, deploy, and evaluate ...
... feedback, and experience, focusing on reinforcement learning across simulation and real-world ... Experience with reward modeling or human-in-the-loop learning • Experience at leading AI labs ...
... feedback, and experience, focusing on reinforcement learning across simulation and real-world ... Experience with reward modeling or human-in-the-loop learning • Experience at leading AI labs ...
Redwood City, CA · On-site
$225K - $400K/yr
... like reinforcement learning from human feedback (RLHF) and fine-tuning. * Collaborate with engineering and product teams to translate safety research into practical, scalable solutions and best ...
Redwood City, CA · On-site
$225K - $400K/yr
... like reinforcement learning from human feedback (RLHF) and fine-tuning. * Collaborate with engineering and product teams to translate safety research into practical, scalable solutions and best ...
... feedback, and experience, focusing on reinforcement learning across simulation and real-world ... Experience with reward modeling or human-in-the-loop learning • Experience at leading AI labs ...
... feedback, and experience, focusing on reinforcement learning across simulation and real-world ... Experience with reward modeling or human-in-the-loop learning • Experience at leading AI labs ...
... feedback, and experience. Responsibilities : • Design and implement reinforcement learning ... Experience with reward modeling or human-in-the-loop learning • Experience at leading AI labs ...
... feedback, and experience. Responsibilities : • Design and implement reinforcement learning ... Experience with reward modeling or human-in-the-loop learning • Experience at leading AI labs ...
San Jose, CA · On-site
$200K - $400K/yr
Train policies that learn from interaction, feedback, and large-scale experience across diverse ... Experience with reward modeling or human-in-the-loop learning * Experience at leading AI labs such ...
San Jose, CA · On-site
$200K - $400K/yr
Train policies that learn from interaction, feedback, and large-scale experience across diverse ... Experience with reward modeling or human-in-the-loop learning * Experience at leading AI labs such ...
Sunnyvale, CA · On-site
$150K - $450K/yr
Hands-on experience with LLM algorithms, such as Supervised Fine-Tuning (SFT) and Reinforcement Learning with Human Feedback (RLHF). * Excellent data analysis skills. $150,000 - $450,000 a year Visa ...
Sunnyvale, CA · On-site
$150K - $450K/yr
Hands-on experience with LLM algorithms, such as Supervised Fine-Tuning (SFT) and Reinforcement Learning with Human Feedback (RLHF). * Excellent data analysis skills. $150,000 - $450,000 a year Visa ...
San Jose, CA · On-site
$200K - $400K/yr
Train policies that learn from interaction, feedback, and large-scale experience across diverse ... Experience with reward modeling or human-in-the-loop learning * Experience at leading AI labs such ...
San Jose, CA · On-site
$200K - $400K/yr
Train policies that learn from interaction, feedback, and large-scale experience across diverse ... Experience with reward modeling or human-in-the-loop learning * Experience at leading AI labs such ...
| Aspect | Reinforcement Learning With Human Feedback | Reinforcement Learning Engineer |
|---|---|---|
| Credentials | Typically requires knowledge of machine learning, AI, and data analysis | Requires similar credentials in machine learning, programming, and AI |
| Work Environment | Research labs, AI development teams, tech companies | Development teams, research labs, tech firms |
| Industry Usage | Used in AI training, human-in-the-loop systems, and model refinement | Designing, implementing, and optimizing reinforcement learning algorithms |
Reinforcement Learning With Human Feedback focuses on improving AI models through human input, while Reinforcement Learning Engineers develop and deploy these algorithms. Both roles require strong machine learning skills and often work in similar environments, but their core responsibilities differ in application and focus.
For Reinforcement Learning With Human Feedback jobs in California, the most frequently searched job titles are:
The top searched job categories for Reinforcement Learning With Human Feedback jobs in California are:
Cities in California with the most Reinforcement Learning With Human Feedback job openings:

Sourced by ZipRecruiter
We deliver consistently superior recruiting by virtue of trusting, communicative relationships with companies and candidates alike. From Fortune 100s to startups, clients lean on us to fulfill their range of needs from contract to full-time positions. With an intimate knowledge of the industries we serve, a keen sense of what makes for high-performing talent in any role, and shared sense of urgency, our clients will tell you: your solution begins here.
Recruiting and staffing services
51 - 200 Employees
Walnut Creek, CA, US
2005