Figure is an AI Robotics company focused on developing autonomous general-purpose humanoid robots with human-level intelligence. They are seeking a Staff Reinforcement Learning Engineer to develop ...
Figure is an AI Robotics company focused on developing autonomous general-purpose humanoid robots with human-level intelligence. They are seeking a Staff Reinforcement Learning Engineer to develop ...
LLM Research Engineer
Mountain View, CA · On-site
Knowledge of reinforcement learning and RLHF (Reinforcement Learning with Human Feedback). **No C2C resumes are considered** Thank you! FocusKPI Hiring Team Founded in 2010, FocusKPI, Inc. (FocusKPI ...
Quick apply
LLM Research Engineer
Mountain View, CA · On-site
Knowledge of reinforcement learning and RLHF (Reinforcement Learning with Human Feedback). **No C2C resumes are considered** Thank you! FocusKPI Hiring Team Founded in 2010, FocusKPI, Inc. (FocusKPI ...
They are seeking a Reinforcement Learning Engineer to join their Manipulation team, focusing on ... feedback into learned grasp policies. • Experience with contact-rich manipulation and force ...
They are seeking a Reinforcement Learning Engineer to join their Manipulation team, focusing on ... feedback into learned grasp policies. • Experience with contact-rich manipulation and force ...
LLM Research Engineer
Mountain View, CA · On-site
$115 - $131/hr
Knowledge of reinforcement learning and RLHF (Reinforcement Learning with Human Feedback). **No C2C resumes are considered** Thank you! FocusKPI Hiring Team Founded in 2010, FocusKPI, Inc. (FocusKPI ...
LLM Research Engineer
Mountain View, CA · On-site
$115 - $131/hr
Knowledge of reinforcement learning and RLHF (Reinforcement Learning with Human Feedback). **No C2C resumes are considered** Thank you! FocusKPI Hiring Team Founded in 2010, FocusKPI, Inc. (FocusKPI ...
You'll design and implement advanced methods to align human feedback with the training of cutting-edge AI models, including techniques like Reinforcement Learning from Human Feedback (RLHF) , Direct ...
Quick apply
You'll design and implement advanced methods to align human feedback with the training of cutting-edge AI models, including techniques like Reinforcement Learning from Human Feedback (RLHF) , Direct ...
Reinforcement Learning Engineer - Whole Body Control
San Jose, CA · On-site
$200K - $350K/yr
The goal of the company is to ship humanoid robots with human level intelligence. Its robots are ... We are looking for a Reinforcement Learning Engineer to develop, train, deploy, and evaluate ...
Reinforcement Learning Engineer - Whole Body Control
San Jose, CA · On-site
$200K - $350K/yr
The goal of the company is to ship humanoid robots with human level intelligence. Its robots are ... We are looking for a Reinforcement Learning Engineer to develop, train, deploy, and evaluate ...
Reinforcement Learning Engineer - Whole Body Control
San Jose, CA · Hybrid
$150K/yr
The goal of the company is to ship humanoid robots with human level intelligence. Its robots are ... We are looking for a Reinforcement Learning Engineer to develop, train, deploy, and evaluate ...
Reinforcement Learning Engineer - Whole Body Control
San Jose, CA · Hybrid
$150K/yr
The goal of the company is to ship humanoid robots with human level intelligence. Its robots are ... We are looking for a Reinforcement Learning Engineer to develop, train, deploy, and evaluate ...
... with Prompt Engineering, RLHF (Reinforcement Learning from Human Feedback), or RAG (Retrieval-Augmented Generation) workflows. • the ability to audit complex model logic, identify training data ...
... with Prompt Engineering, RLHF (Reinforcement Learning from Human Feedback), or RAG (Retrieval-Augmented Generation) workflows. • the ability to audit complex model logic, identify training data ...
... with Prompt Engineering, RLHF (Reinforcement Learning from Human Feedback), or RAG (Retrieval-Augmented Generation) workflows. • the ability to audit complex model logic, identify training data ...
... with Prompt Engineering, RLHF (Reinforcement Learning from Human Feedback), or RAG (Retrieval-Augmented Generation) workflows. • the ability to audit complex model logic, identify training data ...
Machine Learning Engineer: Imitation and Reinforcement Learning for Robotics
San Francisco, CA · On-site
Required : • 3+ years of practical experience applying Machine Learning with Deep Learning ... or reinforcement learning Preferred : • Experience with diffusion policies, Vision-Language ...
Machine Learning Engineer: Imitation and Reinforcement Learning for Robotics
San Francisco, CA · On-site
Required : • 3+ years of practical experience applying Machine Learning with Deep Learning ... or reinforcement learning Preferred : • Experience with diffusion policies, Vision-Language ...
Reward model training, preference learning, human feedback integration • Direct optimization: DPO ... Software engineering beyond research with scalable pipelines and training infrastructure • ...
Reward model training, preference learning, human feedback integration • Direct optimization: DPO ... Software engineering beyond research with scalable pipelines and training infrastructure • ...
Research Engineer, AI Safety & Alignment
Redwood City, CA · On-site
$225K - $400K/yr
... like reinforcement learning from human feedback (RLHF) and fine-tuning. * Collaborate with engineering and product teams to translate safety research into practical, scalable solutions and best ...
Research Engineer, AI Safety & Alignment
Redwood City, CA · On-site
$225K - $400K/yr
... like reinforcement learning from human feedback (RLHF) and fine-tuning. * Collaborate with engineering and product teams to translate safety research into practical, scalable solutions and best ...
... feedback, and experience. Responsibilities : • Design and implement reinforcement learning ... Experience with reward modeling or human-in-the-loop learning • Experience at leading AI labs ...
... feedback, and experience. Responsibilities : • Design and implement reinforcement learning ... Experience with reward modeling or human-in-the-loop learning • Experience at leading AI labs ...
... feedback, and experience, focusing on reinforcement learning across simulation and real-world ... Experience with reward modeling or human-in-the-loop learning • Experience at leading AI labs ...
... feedback, and experience, focusing on reinforcement learning across simulation and real-world ... Experience with reward modeling or human-in-the-loop learning • Experience at leading AI labs ...
Machine Learning Engineer
Sunnyvale, CA · On-site
$150K - $450K/yr
Hands-on experience with LLM algorithms, such as Supervised Fine-Tuning (SFT) and Reinforcement Learning with Human Feedback (RLHF). * Excellent data analysis skills. $150,000 - $450,000 a year Visa ...
Machine Learning Engineer
Sunnyvale, CA · On-site
$150K - $450K/yr
Hands-on experience with LLM algorithms, such as Supervised Fine-Tuning (SFT) and Reinforcement Learning with Human Feedback (RLHF). * Excellent data analysis skills. $150,000 - $450,000 a year Visa ...
... feedback, and experience, focusing on reinforcement learning across simulation and real-world ... Experience with reward modeling or human-in-the-loop learning • Experience at leading AI labs ...
... feedback, and experience, focusing on reinforcement learning across simulation and real-world ... Experience with reward modeling or human-in-the-loop learning • Experience at leading AI labs ...
Research Associate in Systems and Information Engineering
Charlottesville, VA · On-site
$62K - $67K/yr
... for reinforcement learning from human feedback (RLHF) and epistemic control of large language ... The initial appointment will be for one year , with the possibility of renewal for an additional ...
Research Associate in Systems and Information Engineering
Charlottesville, VA · On-site
$62K - $67K/yr
... for reinforcement learning from human feedback (RLHF) and epistemic control of large language ... The initial appointment will be for one year , with the possibility of renewal for an additional ...
Senior Reinforcement Learning Engineer
Austin, TX · On-site
$103K - $142K/yr
... human demonstration data (mocap, teleoperation) into robust reference trajectories for reinforcement learning. Qualifications : Required : • Deep, hands-on expertise (5+ years) with common RL ...
Senior Reinforcement Learning Engineer
Austin, TX · On-site
$103K - $142K/yr
... human demonstration data (mocap, teleoperation) into robust reference trajectories for reinforcement learning. Qualifications : Required : • Deep, hands-on expertise (5+ years) with common RL ...
Machine Learning Engineer, Reinforcement Learning
Pittsburgh, PA · On-site
$100K - $300K/yr
Our team consists of individuals with varying levels of experience and backgrounds, from new ... Develop and implement state-of-the-art reinforcement learning algorithms for robotic applications.
Machine Learning Engineer, Reinforcement Learning
Pittsburgh, PA · On-site
$100K - $300K/yr
Our team consists of individuals with varying levels of experience and backgrounds, from new ... Develop and implement state-of-the-art reinforcement learning algorithms for robotic applications.
... edge reinforcement learning algorithms, conducting experiments, and optimizing these models to ... This will require close collaboration with our robotics, research, and engineering team. Your work ...
... edge reinforcement learning algorithms, conducting experiments, and optimizing these models to ... This will require close collaboration with our robotics, research, and engineering team. Your work ...
Reinforcement Learning With Human Feedback information
See salary details
$29.79 is the 25th percentile. Wages below this are outliers.
$26.92 - $30.79
34% of jobs
The median wage is $34.32 / hr.
$30.79 - $34.66
18% of jobs
$34.66 - $38.53
12% of jobs
$38.53 - $42.40
3% of jobs
$42.40 - $46.26
5% of jobs
$46.26 - $50.13
2% of jobs
$50.45 is the 75th percentile. Wages above this are outliers.
$50.13 - $54
16% of jobs
$54 - $57.87
9% of jobs
$57.87 - $61.74
0% of jobs
$61.74 - $65.60
0% of jobs
$65.60 - $69.47
1% of jobs
$26
$40
$69
How much do reinforcement learning with human feedback jobs pay per hour?
What are the key skills and qualifications needed to thrive as a reinforcement learning with human feedback engineer?
What is the difference between Reinforcement Learning With Human Feedback vs Reinforcement Learning Engineer?
| Aspect | Reinforcement Learning With Human Feedback | Reinforcement Learning Engineer |
|---|---|---|
| Credentials | Typically requires knowledge of machine learning, AI, and data analysis | Requires similar credentials in machine learning, programming, and AI |
| Work Environment | Research labs, AI development teams, tech companies | Development teams, research labs, tech firms |
| Industry Usage | Used in AI training, human-in-the-loop systems, and model refinement | Designing, implementing, and optimizing reinforcement learning algorithms |
Reinforcement Learning With Human Feedback focuses on improving AI models through human input, while Reinforcement Learning Engineers develop and deploy these algorithms. Both roles require strong machine learning skills and often work in similar environments, but their core responsibilities differ in application and focus.
What is reinforcement learning with human feedback?
What collaborations are typical for a reinforcement learning with human feedback specialist within a machine learning team?

Full-time
Re-posted 28 days ago
Job description
Figure is an AI Robotics company focused on developing autonomous general-purpose humanoid robots with human-level intelligence. They are seeking a Staff Reinforcement Learning Engineer to develop, train, deploy, and evaluate advanced reinforcement learning algorithms for the whole body control of their humanoid robots.
Responsibilities:
• Develop, train, and deploy reinforcement learning algorithms for whole body control
• Determine the observations, actions, and model types that unlock maximum performance
• Identify and close the most important sim-to-real gaps
• Define, test, and evaluate performance metrics for learned policies
• Harden the control stack to ensure rock solid robustness
Qualifications:
Required:
• Strong background in dynamics and control, ideally of legged robots
• Experience with reinforcement learning algorithms for robotics: PPO, SAC, etc
• Experience tuning hyperparameters and cost functions for these RL algorithms
• Familiarity with common RL techniques such as: domain randomization, curriculum learning, reward shaping, etc.
• Capable of leading complex controls projects and mentoring junior engineers
Preferred:
• Experience with behavior cloning techniques (e.g. distillation)
Company:
Figure is an AI robotics company that develops autonomous general-purpose humanoid robots. Founded in 2022, the company is headquartered in San Jose, USA, with a team of 201-500 employees. The company is currently Growth Stage.