LLM Research Engineer
$90 - $121.86/hr
Knowledge of reinforcement learning and RLHF (Reinforcement Learning with Human Feedback). Compensation: $90 - $121.86 per hour ID#: 36408719
Quick apply
$90 - $121.86/hr
Knowledge of reinforcement learning and RLHF (Reinforcement Learning with Human Feedback). Compensation: $90 - $121.86 per hour ID#: 36408719
Quick apply
$90 - $121.86/hr
Knowledge of reinforcement learning and RLHF (Reinforcement Learning with Human Feedback). Compensation: $90 - $121.86 per hour ID#: 36408719
... and reinforcement learning from human feedback (RLHF) pipelines for video generation models • ... with cross-functional teams to integrate alignment improvements into our production pipeline • ...
... and reinforcement learning from human feedback (RLHF) pipelines for video generation models • ... with cross-functional teams to integrate alignment improvements into our production pipeline • ...
... reinforcement learning from human feedback (RLHF) and fine-tuning. • Collaborate with engineering and product teams to translate safety research into practical, scalable solutions and best ...
... reinforcement learning from human feedback (RLHF) and fine-tuning. • Collaborate with engineering and product teams to translate safety research into practical, scalable solutions and best ...
Design and implement supervised fine-tuning and reinforcement learning from human feedback (RLHF ... Collaborate with cross-functional teams to integrate alignment improvements into our production ...
Quick apply
Design and implement supervised fine-tuning and reinforcement learning from human feedback (RLHF ... Collaborate with cross-functional teams to integrate alignment improvements into our production ...
Design and implement supervised fine-tuning and reinforcement learning from human feedback (RLHF ... Collaborate with cross-functional teams to integrate alignment improvements into our production ...
Design and implement supervised fine-tuning and reinforcement learning from human feedback (RLHF ... Collaborate with cross-functional teams to integrate alignment improvements into our production ...
Sunnyvale, CA · On-site
... Reinforcement Learning with Human Feedback (RLHF). • Excellent data analysis skills. Company : Official account of Mohamed bin Zayed University of Artificial Intelligence. Dedicated to research ...
Sunnyvale, CA · On-site
... Reinforcement Learning with Human Feedback (RLHF). • Excellent data analysis skills. Company : Official account of Mohamed bin Zayed University of Artificial Intelligence. Dedicated to research ...
Palo Alto, CA · On-site
$267K - $280K/yr
Leverage a wide range of cutting-edge technologies such as Semantic Search, Personalized Search, Supervised Fine-Tuning (SFT), Reinforcement Learning with Human Feedback (RLHF), Retrieval-Augmented ...
Palo Alto, CA · On-site
$267K - $280K/yr
Leverage a wide range of cutting-edge technologies such as Semantic Search, Personalized Search, Supervised Fine-Tuning (SFT), Reinforcement Learning with Human Feedback (RLHF), Retrieval-Augmented ...
Sunnyvale, CA · On-site
... Reinforcement Learning with Human Feedback (RLHF). • Excellent data analysis skills. Company : Official account of Mohamed bin Zayed University of Artificial Intelligence. Dedicated to research ...
Sunnyvale, CA · On-site
... Reinforcement Learning with Human Feedback (RLHF). • Excellent data analysis skills. Company : Official account of Mohamed bin Zayed University of Artificial Intelligence. Dedicated to research ...
Figure is an AI Robotics company focused on developing autonomous humanoid robots with human-level intelligence. They are seeking a Reinforcement Learning Engineer to develop, train, deploy, and ...
Figure is an AI Robotics company focused on developing autonomous humanoid robots with human-level intelligence. They are seeking a Reinforcement Learning Engineer to develop, train, deploy, and ...
$150K - $450K/yr
Hands-on experience with LLM algorithms, such as Supervised Fine-Tuning (SFT) and Reinforcement Learning with Human Feedback (RLHF). * Excellent data analysis skills. Visa Sponsorship This position ...
Quick apply
$150K - $450K/yr
Hands-on experience with LLM algorithms, such as Supervised Fine-Tuning (SFT) and Reinforcement Learning with Human Feedback (RLHF). * Excellent data analysis skills. Visa Sponsorship This position ...
Outerport is seeking a Member of Technical Staff with expertise in reinforcement learning. The role involves training reinforcement learning-based LLMs for applications in materials science and ...
Outerport is seeking a Member of Technical Staff with expertise in reinforcement learning. The role involves training reinforcement learning-based LLMs for applications in materials science and ...
Figure is an AI Robotics company focused on developing autonomous general-purpose humanoid robots with human-level intelligence. They are seeking a Staff Reinforcement Learning Engineer to develop ...
Figure is an AI Robotics company focused on developing autonomous general-purpose humanoid robots with human-level intelligence. They are seeking a Staff Reinforcement Learning Engineer to develop ...
Mountain View, CA · On-site
Knowledge of reinforcement learning and RLHF (Reinforcement Learning with Human Feedback). **No C2C resumes are considered** Thank you! FocusKPI Hiring Team Founded in 2010, FocusKPI, Inc. (FocusKPI ...
Quick apply
Mountain View, CA · On-site
Knowledge of reinforcement learning and RLHF (Reinforcement Learning with Human Feedback). **No C2C resumes are considered** Thank you! FocusKPI Hiring Team Founded in 2010, FocusKPI, Inc. (FocusKPI ...
San Jose, CA · On-site
$150K/yr
The goal of the company is to ship humanoid robots with human level intelligence. Its robots are ... We are looking for a Reinforcement Learning Engineer to develop, train, deploy, and evaluate ...
San Jose, CA · On-site
$150K/yr
The goal of the company is to ship humanoid robots with human level intelligence. Its robots are ... We are looking for a Reinforcement Learning Engineer to develop, train, deploy, and evaluate ...
San Jose, CA · Hybrid
$150K/yr
The goal of the company is to ship humanoid robots with human level intelligence. Its robots are ... We are looking for a Reinforcement Learning Engineer to develop, train, deploy, and evaluate ...
San Jose, CA · Hybrid
$150K/yr
The goal of the company is to ship humanoid robots with human level intelligence. Its robots are ... We are looking for a Reinforcement Learning Engineer to develop, train, deploy, and evaluate ...
You'll design and implement advanced methods to align human feedback with the training of cutting-edge AI models, including techniques like Reinforcement Learning from Human Feedback (RLHF) , Direct ...
Quick apply
You'll design and implement advanced methods to align human feedback with the training of cutting-edge AI models, including techniques like Reinforcement Learning from Human Feedback (RLHF) , Direct ...
Mountain View, CA · On-site
$115 - $131/hr
Knowledge of reinforcement learning and RLHF (Reinforcement Learning with Human Feedback). **No C2C resumes are considered** Thank you! FocusKPI Hiring Team Founded in 2010, FocusKPI, Inc. (FocusKPI ...
Mountain View, CA · On-site
$115 - $131/hr
Knowledge of reinforcement learning and RLHF (Reinforcement Learning with Human Feedback). **No C2C resumes are considered** Thank you! FocusKPI Hiring Team Founded in 2010, FocusKPI, Inc. (FocusKPI ...
... feedback, and experience, focusing on reinforcement learning across simulation and real-world ... Experience with reward modeling or human-in-the-loop learning • Experience at leading AI labs ...
... feedback, and experience, focusing on reinforcement learning across simulation and real-world ... Experience with reward modeling or human-in-the-loop learning • Experience at leading AI labs ...
San Francisco, CA · On-site
Required : • 3+ years of practical experience applying Machine Learning with Deep Learning ... or reinforcement learning Preferred : • Experience with diffusion policies, Vision-Language ...
San Francisco, CA · On-site
Required : • 3+ years of practical experience applying Machine Learning with Deep Learning ... or reinforcement learning Preferred : • Experience with diffusion policies, Vision-Language ...
Redwood City, CA · On-site
$225K - $400K/yr
... like reinforcement learning from human feedback (RLHF) and fine-tuning. * Collaborate with engineering and product teams to translate safety research into practical, scalable solutions and best ...
Redwood City, CA · On-site
$225K - $400K/yr
... like reinforcement learning from human feedback (RLHF) and fine-tuning. * Collaborate with engineering and product teams to translate safety research into practical, scalable solutions and best ...
| Aspect | Reinforcement Learning With Human Feedback | Reinforcement Learning Engineer |
|---|---|---|
| Credentials | Typically requires knowledge of machine learning, AI, and data analysis | Requires similar credentials in machine learning, programming, and AI |
| Work Environment | Research labs, AI development teams, tech companies | Development teams, research labs, tech firms |
| Industry Usage | Used in AI training, human-in-the-loop systems, and model refinement | Designing, implementing, and optimizing reinforcement learning algorithms |
Reinforcement Learning With Human Feedback focuses on improving AI models through human input, while Reinforcement Learning Engineers develop and deploy these algorithms. Both roles require strong machine learning skills and often work in similar environments, but their core responsibilities differ in application and focus.
For Reinforcement Learning With Human Feedback jobs in California, the most frequently searched job titles are:
The top searched job categories for Reinforcement Learning With Human Feedback jobs in California are:
Cities in California with the most Reinforcement Learning With Human Feedback job openings:

Mountain View, CA
$90 - $121.86/hr
Full-time
Re-posted 13 days ago
Sourced by ZipRecruiter
We deliver consistently superior recruiting by virtue of trusting, communicative relationships with companies and candidates alike. From Fortune 100s to startups, clients lean on us to fulfill their range of needs from contract to full-time positions. With an intimate knowledge of the industries we serve, a keen sense of what makes for high-performing talent in any role, and shared sense of urgency, our clients will tell you: your solution begins here.
Recruiting and staffing services
51 - 200 Employees
Walnut Creek, CA, US
2005