Ryz Labs is seeking a Head of Sales to own the GTM strategy, execution, and revenue growth for our RLHF vertical. Based in San Francisco, this leader will be responsible for building and managing the ...
Ryz Labs is seeking a Head of Sales to own the GTM strategy, execution, and revenue growth for our RLHF vertical. Based in San Francisco, this leader will be responsible for building and managing the ...
Head of Sales - RLHF Vertical
San Francisco, CA · On-site +1
Ryz Labs is seeking a Head of Sales to own the GTM strategy, execution, and revenue growth for our RLHF vertical. Based in San Francisco, this leader will be responsible for building and managing the ...
Head of Sales - RLHF Vertical
San Francisco, CA · On-site +1
Ryz Labs is seeking a Head of Sales to own the GTM strategy, execution, and revenue growth for our RLHF vertical. Based in San Francisco, this leader will be responsible for building and managing the ...
Build and lead the sales organization for the RLHF vertical, including hiring, training, and performance management. * Define and execute the GTM strategy for enterprise and mid-market clients in the ...
Build and lead the sales organization for the RLHF vertical, including hiring, training, and performance management. * Define and execute the GTM strategy for enterprise and mid-market clients in the ...
A tech startup in San Francisco is seeking a Head of Sales to own the GTM strategy and drive revenue growth for the RLHF vertical. This candidate will build and manage the sales function, establish ...
A tech startup in San Francisco is seeking a Head of Sales to own the GTM strategy and drive revenue growth for the RLHF vertical. This candidate will build and manage the sales function, establish ...
Incentive Scientist: ML, Economics & RLHF Research
San Francisco, CA · On-site
$80 - $100/hr
The team publishes regularly and ships findings into RLHF pipelines and evaluation suites. PhD or equivalent in ML / economics / decision theory expected. Posted on the board of The Incentives Lab ...
Incentive Scientist: ML, Economics & RLHF Research
San Francisco, CA · On-site
$80 - $100/hr
The team publishes regularly and ships findings into RLHF pipelines and evaluation suites. PhD or equivalent in ML / economics / decision theory expected. Posted on the board of The Incentives Lab ...
视频数据标注 & 评测专家 / 负责人
San Francisco, CA · On-site
... PT / SFT / RLHF / Eval 各阶段对数据的不同要求,推动数据标准持续迭代. 任职要求 • 3 年以上相关经验,做过视频 / 图像 / 多模态数据的标注或评测 ...
视频数据标注 & 评测专家 / 负责人
San Francisco, CA · On-site
... PT / SFT / RLHF / Eval 各阶段对数据的不同要求,推动数据标准持续迭代. 任职要求 • 3 年以上相关经验,做过视频 / 图像 / 多模态数据的标注或评测 ...
They combine platforms, tools and a large expert community to deliver training data, evaluation, RLHF and multilingual AI solutions for complex, high impact use cases. The work is global, fast moving ...
They combine platforms, tools and a large expert community to deliver training data, evaluation, RLHF and multilingual AI solutions for complex, high impact use cases. The work is global, fast moving ...
$80 - $100/hr
RLHF, Training, AI## To create truly intelligent and aligned AI, human expertise is indispensable. As an AI Training Specialist, you will be at the cutting edge of AI development, actively shaping ...
$80 - $100/hr
RLHF, Training, AI## To create truly intelligent and aligned AI, human expertise is indispensable. As an AI Training Specialist, you will be at the cutting edge of AI development, actively shaping ...
Member of Technical Staff - Post-Training and RL
$180K - $600K/yr
You will work on the most critical post-training and reinforcement learning challenges at any given time - including reward modeling, preference optimization (RLHF/DPO), and RL for improving ...
Member of Technical Staff - Post-Training and RL
$180K - $600K/yr
You will work on the most critical post-training and reinforcement learning challenges at any given time - including reward modeling, preference optimization (RLHF/DPO), and RL for improving ...
Mountain View, USA Founding Engineer - Reinforcement Learning
San Mateo, CA · On-site
$150 - $200/hr
Design human‑in‑the‑loop workflows for RLHF, preference data, expert grading, model behavior evaluation, and quality improvement. * Build tools to measure environment quality, task difficulty ...
Mountain View, USA Founding Engineer - Reinforcement Learning
San Mateo, CA · On-site
$150 - $200/hr
Design human‑in‑the‑loop workflows for RLHF, preference data, expert grading, model behavior evaluation, and quality improvement. * Build tools to measure environment quality, task difficulty ...
Senior AI Model Fine-Tuning Engineer
Austin, TX · On-site
$103K - $142K/yr
... and RLHF. Responsibilities : • Lead the fine-tuning process for large pre-trained models, focusing on making models behave appropriately in different contexts (e.g., following instructions ...
Senior AI Model Fine-Tuning Engineer
Austin, TX · On-site
$103K - $142K/yr
... and RLHF. Responsibilities : • Lead the fine-tuning process for large pre-trained models, focusing on making models behave appropriately in different contexts (e.g., following instructions ...
Proficiency in modern RLHF algorithms: PPO, DPO, GRPO, etc. * Hands-on experience training reward models and finetuning LLM/VLM/VLA * Knowledge of distributed RL training at scale * Proficiency with ...
Proficiency in modern RLHF algorithms: PPO, DPO, GRPO, etc. * Hands-on experience training reward models and finetuning LLM/VLM/VLA * Knowledge of distributed RL training at scale * Proficiency with ...
Proficiency in modern RLHF algorithms: PPO, DPO, GRPO, etc. * Hands-on experience training reward models and finetuning LLM/VLM/VLA * Knowledge of distributed RL training at scale * Proficiency with ...
Proficiency in modern RLHF algorithms: PPO, DPO, GRPO, etc. * Hands-on experience training reward models and finetuning LLM/VLM/VLA * Knowledge of distributed RL training at scale * Proficiency with ...
Research Scientist, Reinforcement Learning
Fremont, CA · On-site
$100 - $125/hr
Proficiency in modern RLHF algorithms: PPO, DPO, GRPO, etc. * Hands‑on experience training reward models and finetuning LLM/VLM/VLA * Knowledge of distributed RL training at scale * Proficiency ...
Research Scientist, Reinforcement Learning
Fremont, CA · On-site
$100 - $125/hr
Proficiency in modern RLHF algorithms: PPO, DPO, GRPO, etc. * Hands‑on experience training reward models and finetuning LLM/VLM/VLA * Knowledge of distributed RL training at scale * Proficiency ...
Preferred : • Demonstrated expertise in post-training methods and/or next generation use cases for large language models including instruction tuning, RLHF, tool use, reasoning, agents, and ...
Preferred : • Demonstrated expertise in post-training methods and/or next generation use cases for large language models including instruction tuning, RLHF, tool use, reasoning, agents, and ...
Preferred : • Demonstrated expertise in post-training methods and/or next generation use cases for large language models including instruction tuning, RLHF, tool use, reasoning, agents, and ...
Preferred : • Demonstrated expertise in post-training methods and/or next generation use cases for large language models including instruction tuning, RLHF, tool use, reasoning, agents, and ...
Preferred : • Demonstrated expertise in post-training methods and/or next generation use cases for large language models including instruction tuning, RLHF, tool use, reasoning, agents, and ...
Preferred : • Demonstrated expertise in post-training methods and/or next generation use cases for large language models including instruction tuning, RLHF, tool use, reasoning, agents, and ...
Proficiency in modern RLHF algorithms: PPO, DPO, GRPO, etc. * Hands-on experience training reward models and finetuning LLM/VLM/VLA * Knowledge of distributed RL training at scale * Proficiency with ...
Quick apply
Proficiency in modern RLHF algorithms: PPO, DPO, GRPO, etc. * Hands-on experience training reward models and finetuning LLM/VLM/VLA * Knowledge of distributed RL training at scale * Proficiency with ...
Principal Product Manager
Redmond, WA · On-site
$220K - $331K/yr
You'll be embedded directly with applied researchers and ML engineers running RLHF, fine-tuning, and preference-tuning pipelines, reading model outputs, calling out what's wrong, and turning that ...
Principal Product Manager
Redmond, WA · On-site
$220K - $331K/yr
You'll be embedded directly with applied researchers and ML engineers running RLHF, fine-tuning, and preference-tuning pipelines, reading model outputs, calling out what's wrong, and turning that ...
Senior Staff Research Engineer - Reinforcement Learning for AI Agents
Santa Clara, CA · On-site
$122K - $168K/yr
Preferred : • Experience with RLHF or preference learning. • Experience with LLM agents or tool-using AI systems. • Multi-agent systems or long-horizon planning. • Simulation environments for ...
Senior Staff Research Engineer - Reinforcement Learning for AI Agents
Santa Clara, CA · On-site
$122K - $168K/yr
Preferred : • Experience with RLHF or preference learning. • Experience with LLM agents or tool-using AI systems. • Multi-agent systems or long-horizon planning. • Simulation environments for ...
Rlhf information
What is an RLHF job?
An RLHF (Reinforcement Learning with Human Feedback) job involves training AI models using human feedback to improve their responses. Professionals in this role analyze model outputs, provide evaluations, and refine AI behavior through reinforcement learning techniques. These roles are common in AI research, content moderation, and chatbot development.
What are the key skills and qualifications needed to thrive as a Reinforcement Learning from Human Feedback (RLHF) engineer, and why are they important?
What are some common challenges faced by professionals working in Reinforcement Learning from Human Feedback (RLHF) roles?
What is the difference between Rlhf vs Rn?
| Aspect | Rlhf | Rn |
|---|---|---|
| Required Credentials | Licensed healthcare professional, often with specialized training in mental health or behavioral health | Licensed practical nurse or registered nurse, with nursing licensure and possibly additional certifications |
| Work Environment | Behavioral health facilities, clinics, hospitals, or community health settings | Hospitals, clinics, long-term care facilities, and community health settings |
| Employer & Industry Usage | Behavioral health and mental health services | General healthcare and nursing services |
| Common Search & Comparison | Rlhf vs Rn | Rlhf vs Rn |
While Rlhf (Registered Licensed Mental Health Facilitator) focuses on mental health support and behavioral health interventions, Rn (Registered Nurse) provides broader nursing care across various medical settings. Both roles require licensure, but Rlhf specializes in mental health, whereas Rn covers general patient care.
What cities are hiring for Rlhf jobs?
Cities with the most Rlhf job openings:
What are the most commonly searched types of Rlhf jobs?
The most popular types of Rlhf jobs are:
What states have the most Rlhf jobs?
States with the most job openings for Rlhf jobs include:
What job categories do people searching Rlhf jobs look for?
The top searched job categories for Rlhf jobs are:

Head of Sales - RLHF Vertical
San Francisco, CA • On-site
Full-time
Re-posted 28 days ago
Job description
About Ryz Labs
Sourced by ZipRecruiter
Industry
Software development
Company size
1 - 10 Employees
Headquarters location
Los Angeles, CA, US
Year founded
2021