Apply and refine evaluation rubrics for RLHF , SFT , and AI safety benchmarking . * Identify unsafe outputs, hallucinations, reasoning failures, and policy violations. * Provide structured feedback ...
Quick apply
Apply and refine evaluation rubrics for RLHF , SFT , and AI safety benchmarking . * Identify unsafe outputs, hallucinations, reasoning failures, and policy violations. * Provide structured feedback ...
Quick apply
Apply and refine evaluation rubrics for RLHF , SFT , and AI safety benchmarking . * Identify unsafe outputs, hallucinations, reasoning failures, and policy violations. * Provide structured feedback ...
New York, NY · Remote
$60 - $70/hr
Apply and refine evaluation rubrics for RLHF , SFT , and AI safety benchmarking . * Identify unsafe outputs, hallucinations, reasoning failures, and policy violations. * Provide structured feedback ...
Quick apply
New York, NY · Remote
$60 - $70/hr
Apply and refine evaluation rubrics for RLHF , SFT , and AI safety benchmarking . * Identify unsafe outputs, hallucinations, reasoning failures, and policy violations. * Provide structured feedback ...
Apply and refine evaluation rubrics for RLHF , SFT , and AI safety benchmarking . * Identify unsafe outputs, hallucinations, reasoning failures, and policy violations. * Provide structured feedback ...
Quick apply
Apply and refine evaluation rubrics for RLHF , SFT , and AI safety benchmarking . * Identify unsafe outputs, hallucinations, reasoning failures, and policy violations. * Provide structured feedback ...
Santa Clara, CA · On-site
$122K - $168K/yr
Preferred : • Experience with RLHF or preference learning. • Experience with LLM agents or tool-using AI systems. • Multi-agent systems or long-horizon planning. • Simulation environments for ...
Santa Clara, CA · On-site
$122K - $168K/yr
Preferred : • Experience with RLHF or preference learning. • Experience with LLM agents or tool-using AI systems. • Multi-agent systems or long-horizon planning. • Simulation environments for ...
Required : • Hands-on experience with data generation and evaluation for LLM post-training • Experience training or fine-tuning models using SFT, instruction tuning, RLHF, DPO, or similar ...
Required : • Hands-on experience with data generation and evaluation for LLM post-training • Experience training or fine-tuning models using SFT, instruction tuning, RLHF, DPO, or similar ...
Santa Clara, CA · On-site
$122K - $168K/yr
Preferred : • Experience with RLHF or preference learning. • Experience with LLM agents or tool-using AI systems. • Multi-agent systems or long-horizon planning. • Simulation environments for ...
Santa Clara, CA · On-site
$122K - $168K/yr
Preferred : • Experience with RLHF or preference learning. • Experience with LLM agents or tool-using AI systems. • Multi-agent systems or long-horizon planning. • Simulation environments for ...
Redmond, WA · On-site
$220K - $331K/yr
You'll be embedded directly with applied researchers and ML engineers running RLHF, fine-tuning, and preference-tuning pipelines, reading model outputs, calling out what's wrong, and turning that ...
Redmond, WA · On-site
$220K - $331K/yr
You'll be embedded directly with applied researchers and ML engineers running RLHF, fine-tuning, and preference-tuning pipelines, reading model outputs, calling out what's wrong, and turning that ...
Boston, MA · On-site
$170K - $200K/yr
Build, fine‑tune, and productionize large language model (LLM) pipelines, including PEFT, RLHF, and DPO workflows. * Develop APIs, data pipelines, and orchestration systems for multi‑agent ...
Boston, MA · On-site
$170K - $200K/yr
Build, fine‑tune, and productionize large language model (LLM) pipelines, including PEFT, RLHF, and DPO workflows. * Develop APIs, data pipelines, and orchestration systems for multi‑agent ...
Responsibilities : • Conduct state-of-the-art research on AI safety topics such as RLHF, adversarial training, robustness, and more. • Implement new methods in OpenAI's core model training and ...
Responsibilities : • Conduct state-of-the-art research on AI safety topics such as RLHF, adversarial training, robustness, and more. • Implement new methods in OpenAI's core model training and ...
San Francisco, CA · On-site
$150 - $200/hr
Lead research spikes on emerging techniques (LLMs, multimodal, RLHF) to determine production viability. * Establish ML engineering best practices, including experimentation standards, model ...
New
San Francisco, CA · On-site
$150 - $200/hr
Lead research spikes on emerging techniques (LLMs, multimodal, RLHF) to determine production viability. * Establish ML engineering best practices, including experimentation standards, model ...
New
Charlottesville, VA · On-site
$62K - $67K/yr
Research Program Although RLHF and related methods, including direct preference optimization, are now widely used, their statistical properties remain only partially understood. Human feedback is ...
Charlottesville, VA · On-site
$62K - $67K/yr
Research Program Although RLHF and related methods, including direct preference optimization, are now widely used, their statistical properties remain only partially understood. Human feedback is ...
$180K - $600K/yr
If you previously worked on post-training, RLHF, or trained models used by millions of people it's a big plus, but relevant experience is not required. * You take pride in your work and thrive in ...
Quick apply
$180K - $600K/yr
If you previously worked on post-training, RLHF, or trained models used by millions of people it's a big plus, but relevant experience is not required. * You take pride in your work and thrive in ...
Preferred : • Demonstrated expertise in post-training methods and/or next generation use cases for large language models including instruction tuning, RLHF, tool use, reasoning, agents, and ...
Preferred : • Demonstrated expertise in post-training methods and/or next generation use cases for large language models including instruction tuning, RLHF, tool use, reasoning, agents, and ...
Required : • Significant hands-on experience with RL, RLHF, RLAIF, post-training, alignment, or large-scale fine-tuning for modern foundation models. • Deep understanding of RL/post-training ...
Required : • Significant hands-on experience with RL, RLHF, RLAIF, post-training, alignment, or large-scale fine-tuning for modern foundation models. • Deep understanding of RL/post-training ...
San Francisco, CA · On-site
RLHF, DPO, or equivalent • Experience designing evaluation frameworks for LLM or agentic systems • Strong proficiency in Python and PyTorch Preferred : • Background in Electrical or Computer ...
San Francisco, CA · On-site
RLHF, DPO, or equivalent • Experience designing evaluation frameworks for LLM or agentic systems • Strong proficiency in Python and PyTorch Preferred : • Background in Electrical or Computer ...
Santa Clara, CA · On-site
$122K - $168K/yr
Preferred : • Experience with RLHF or preference learning. • Experience with LLM agents or tool-using AI systems. • Multi-agent systems or long-horizon planning. • Simulation environments for ...
Santa Clara, CA · On-site
$122K - $168K/yr
Preferred : • Experience with RLHF or preference learning. • Experience with LLM agents or tool-using AI systems. • Multi-agent systems or long-horizon planning. • Simulation environments for ...
Palo Alto, CA · On-site
$300/hr
LLM providers**, enterprises, and internal teams to design solutions that span **data labeling, annotation, localization, RLHF, and multi-modal workflows**. This is a leadership position requiring ...
New
Palo Alto, CA · On-site
$300/hr
LLM providers**, enterprises, and internal teams to design solutions that span **data labeling, annotation, localization, RLHF, and multi-modal workflows**. This is a leadership position requiring ...
New
... RLHF Frameworks: provide the high-quality human feedback necessary to align models with human intent, safety, and helpfulness. • Analyze Model Reasoning: critically assess how an AI model navigates ...
... RLHF Frameworks: provide the high-quality human feedback necessary to align models with human intent, safety, and helpfulness. • Analyze Model Reasoning: critically assess how an AI model navigates ...
... RLHF Frameworks: provide the high-quality human feedback necessary to align models with human intent, safety, and helpfulness. • Analyze Model Reasoning: critically assess how an AI model navigates ...
... RLHF Frameworks: provide the high-quality human feedback necessary to align models with human intent, safety, and helpfulness. • Analyze Model Reasoning: critically assess how an AI model navigates ...
Cambridge, MA · On-site
$318K - $425K/yr
Retirement
PTO
Help shape the post-training recipe-SFT, preference tuning (RLHF/DPO-style), distillation, and RL-and own key stages of how each turns a strong base model into a deployable policy. * Reasoning, Tool ...
Cambridge, MA · On-site
$318K - $425K/yr
Retirement
PTO
Help shape the post-training recipe-SFT, preference tuning (RLHF/DPO-style), distillation, and RL-and own key stages of how each turns a strong base model into a deployable policy. * Reasoning, Tool ...
An RLHF (Reinforcement Learning with Human Feedback) job involves training AI models using human feedback to improve their responses. Professionals in this role analyze model outputs, provide evaluations, and refine AI behavior through reinforcement learning techniques. These roles are common in AI research, content moderation, and chatbot development.
| Aspect | Rlhf | Rn |
|---|---|---|
| Required Credentials | Licensed healthcare professional, often with specialized training in mental health or behavioral health | Licensed practical nurse or registered nurse, with nursing licensure and possibly additional certifications |
| Work Environment | Behavioral health facilities, clinics, hospitals, or community health settings | Hospitals, clinics, long-term care facilities, and community health settings |
| Employer & Industry Usage | Behavioral health and mental health services | General healthcare and nursing services |
| Common Search & Comparison | Rlhf vs Rn | Rlhf vs Rn |
While Rlhf (Registered Licensed Mental Health Facilitator) focuses on mental health support and behavioral health interventions, Rn (Registered Nurse) provides broader nursing care across various medical settings. Both roles require licensure, but Rlhf specializes in mental health, whereas Rn covers general patient care.
Cities with the most Rlhf job openings:
The most popular types of Rlhf jobs are:
States with the most job openings for Rlhf jobs include:
The top searched job categories for Rlhf jobs are:

$70/hr
Full-time
Posted 7 days ago
About the job
Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark, General Catalyst, Peter Thiel, Adam D'Angelo, Larry Summers, and Jack Dorsey.
Position: AI Safety Practitioner
Type: Contract
Compensation: $60–$70/hour
Location: Remote
Role Responsibilities
Qualifications
Must-Have
Preferred
Application Process (Takes 20–30 mins to complete)
Resources & Support
PS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this opportunity.