Lead planning and execution of cross-functional programs spanning data collection, annotation pipelines, alignment workflows (RLHF, DPO, Constitutional AI), safety guardrails (adversarial testing ...
Lead planning and execution of cross-functional programs spanning data collection, annotation pipelines, alignment workflows (RLHF, DPO, Constitutional AI), safety guardrails (adversarial testing ...
Technical Program Manager, Model Alignment and Deployment
Redwood City, CA · On-site
$220K - $260K/yr
Lead planning and execution of cross-functional programs spanning data collection, annotation pipelines, alignment workflows (RLHF, DPO, Constitutional AI), safety guardrails (adversarial testing ...
Technical Program Manager, Model Alignment and Deployment
Redwood City, CA · On-site
$220K - $260K/yr
Lead planning and execution of cross-functional programs spanning data collection, annotation pipelines, alignment workflows (RLHF, DPO, Constitutional AI), safety guardrails (adversarial testing ...
... of AI/ML markets; familiarity with RLHF and its applications is strongly preferred. - Strong network in San Francisco and U.S. tech ecosystem. - Entrepreneurial mindset, with ability to thrive in a ...
Quick apply
... of AI/ML markets; familiarity with RLHF and its applications is strongly preferred. - Strong network in San Francisco and U.S. tech ecosystem. - Entrepreneurial mindset, with ability to thrive in a ...
Head of Sales - RLHF Vertical
San Francisco, CA · On-site +1
... of AI/ML markets; familiarity with RLHF and its applications is strongly preferred. - Strong network in San Francisco and U.S. tech ecosystem. - Entrepreneurial mindset, with ability to thrive in a ...
Head of Sales - RLHF Vertical
San Francisco, CA · On-site +1
... of AI/ML markets; familiarity with RLHF and its applications is strongly preferred. - Strong network in San Francisco and U.S. tech ecosystem. - Entrepreneurial mindset, with ability to thrive in a ...
... of AI/ML markets; familiarity with RLHF and its applications is strongly preferred. - Strong network in San Francisco and U.S. tech ecosystem. - Entrepreneurial mindset, with ability to thrive in a ...
... of AI/ML markets; familiarity with RLHF and its applications is strongly preferred. - Strong network in San Francisco and U.S. tech ecosystem. - Entrepreneurial mindset, with ability to thrive in a ...
Deep understanding of AI/ML markets; familiarity with RLHF and its applications is strongly preferred. * Strong network in San Francisco and U.S. tech ecosystem. * Entrepreneurial mindset, with ...
Deep understanding of AI/ML markets; familiarity with RLHF and its applications is strongly preferred. * Strong network in San Francisco and U.S. tech ecosystem. * Entrepreneurial mindset, with ...
Founding AI Engineer
San Francisco, CA · On-site
RLHF, DPO, or equivalent • Experience designing evaluation frameworks for LLM or agentic systems ... Bronco AI is an intelligent integration layer across your legacy IT so you can get things done ...
Founding AI Engineer
San Francisco, CA · On-site
RLHF, DPO, or equivalent • Experience designing evaluation frameworks for LLM or agentic systems ... Bronco AI is an intelligent integration layer across your legacy IT so you can get things done ...
Deep understanding of AI/ML markets; familiarity with RLHF and its applications is strongly preferred. * Strong network in San Francisco and U.S. tech ecosystem. * Entrepreneurial mindset, with ...
Deep understanding of AI/ML markets; familiarity with RLHF and its applications is strongly preferred. * Strong network in San Francisco and U.S. tech ecosystem. * Entrepreneurial mindset, with ...
AI Engagement Manager
$150K - $180K/yr
About This Role AI Engagement Managers own the relationships with Pareto's most important accounts ... Working knowledge of ML and data workflows: data labeling, evals, RLHF, red-teaming, or adjacent ...
AI Engagement Manager
$150K - $180K/yr
About This Role AI Engagement Managers own the relationships with Pareto's most important accounts ... Working knowledge of ML and data workflows: data labeling, evals, RLHF, red-teaming, or adjacent ...
They combine platforms, tools and a large expert community to deliver training data, evaluation, RLHF and multilingual AI solutions for complex, high impact use cases. The work is global, fast moving ...
They combine platforms, tools and a large expert community to deliver training data, evaluation, RLHF and multilingual AI solutions for complex, high impact use cases. The work is global, fast moving ...
Experience supporting AI model training, evaluation, RLHF, or large-scale annotation programs. * Background working with cross-functional teams including researchers, engineers, product managers, and ...
Experience supporting AI model training, evaluation, RLHF, or large-scale annotation programs. * Background working with cross-functional teams including researchers, engineers, product managers, and ...
Develop constitutional AI principles and assist with RLHF alignment pipelines. Qualifications: * Background in cybersecurity, prompt engineering, or adversarial ML. * Experience with jailbreak ...
Develop constitutional AI principles and assist with RLHF alignment pipelines. Qualifications: * Background in cybersecurity, prompt engineering, or adversarial ML. * Experience with jailbreak ...
Develop constitutional AI principles and assist with RLHF alignment pipelines. Qualifications: * Background in cybersecurity, prompt engineering, or adversarial ML. * Experience with jailbreak ...
Develop constitutional AI principles and assist with RLHF alignment pipelines. Qualifications: * Background in cybersecurity, prompt engineering, or adversarial ML. * Experience with jailbreak ...
Develop constitutional AI principles and assist with RLHF alignment pipelines. Qualifications: * Background in cybersecurity, prompt engineering, or adversarial ML. * Experience with jailbreak ...
Develop constitutional AI principles and assist with RLHF alignment pipelines. Qualifications: * Background in cybersecurity, prompt engineering, or adversarial ML. * Experience with jailbreak ...
Develop constitutional AI principles and assist with RLHF alignment pipelines. Qualifications: * Background in cybersecurity, prompt engineering, or adversarial ML. * Experience with jailbreak ...
Develop constitutional AI principles and assist with RLHF alignment pipelines. Qualifications: * Background in cybersecurity, prompt engineering, or adversarial ML. * Experience with jailbreak ...
Develop constitutional AI principles and assist with RLHF alignment pipelines. Qualifications: * Background in cybersecurity, prompt engineering, or adversarial ML. * Experience with jailbreak ...
Develop constitutional AI principles and assist with RLHF alignment pipelines. Qualifications: * Background in cybersecurity, prompt engineering, or adversarial ML. * Experience with jailbreak ...
... RLHF, tool use, reasoning, agents, and multimodal, etc. Company : Scale's mission is to develop reliable AI systems for the world's most important decisions. Founded in 2016, the company is ...
... RLHF, tool use, reasoning, agents, and multimodal, etc. Company : Scale's mission is to develop reliable AI systems for the world's most important decisions. Founded in 2016, the company is ...
Machine Learning Engineer (LLM) (Boston)
Boston, MA · On-site
$170K - $200K/yr
... RLHF, and DPO workflows. * Develop APIs, data pipelines, and orchestration systems for multi‑agent, multi‑turn AI conversations. * Integrate models with backend services, including voice ...
Machine Learning Engineer (LLM) (Boston)
Boston, MA · On-site
$170K - $200K/yr
... RLHF, and DPO workflows. * Develop APIs, data pipelines, and orchestration systems for multi‑agent, multi‑turn AI conversations. * Integrate models with backend services, including voice ...
Agentic AI Engineer Lead
Dallas, TX · On-site
$101K - $133K/yr
The ideal candidate will have deep expertise in LLM orchestration, knowledge graphs, reinforcement learning (RLHF/RLAIF), and real-world AI applications. As a leader in this space, they will be ...
Quick apply
Agentic AI Engineer Lead
Dallas, TX · On-site
$101K - $133K/yr
The ideal candidate will have deep expertise in LLM orchestration, knowledge graphs, reinforcement learning (RLHF/RLAIF), and real-world AI applications. As a leader in this space, they will be ...
... RLHF, tool use, reasoning, agents, and multimodal, etc. Company : Scale's mission is to develop reliable AI systems for the world's most important decisions. Founded in 2016, the company is ...
... RLHF, tool use, reasoning, agents, and multimodal, etc. Company : Scale's mission is to develop reliable AI systems for the world's most important decisions. Founded in 2016, the company is ...
Rlhf Ai information
See salary details
$35.58 - $42.92
13% of jobs
$45.44 is the 25th percentile. Wages below this are outliers.
$42.92 - $50.26
36% of jobs
The median wage is $50.58 / hr.
$50.26 - $57.60
24% of jobs
$64.95 is the 75th percentile. Wages above this are outliers.
$57.60 - $64.95
2% of jobs
$64.95 - $72.29
0% of jobs
$72.29 - $79.63
0% of jobs
$79.63 - $86.98
0% of jobs
$86.98 - $94.32
0% of jobs
$94.32 - $101.66
6% of jobs
$101.66 - $109
9% of jobs
$109 - $116.35
9% of jobs
$35
$65
$116
How much do rlhf ai jobs pay per hour?

Full-time
Re-posted 7 days ago
Job description
Character.AI empowers people to connect, learn and tell stories through interactive entertainment. They are seeking a Technical Program Manager to lead cross-functional programs that enhance model alignment and deployment, ensuring the integration of user experience, safety, and research into scalable AI products.
Responsibilities:
• Program ownership: Lead planning and execution of cross-functional programs spanning data collection, annotation pipelines, alignment workflows (RLHF, DPO, Constitutional AI), safety guardrails (adversarial testing, red-teaming), and model serving. Establish scopes, goals, timelines, risks, and success metrics.
• Cross-functional coordination: Serve as the connective tissue between Post-Training, Safety Engineering, Trust & Safety, ML Infra, UXR, and Product. Translate model development, safety, and user experience priorities into executable roadmaps, keeping tightly coupled workstreams aligned from post-training through to production deployment.
• Evaluation & quality: Develop and maintain custom evaluation frameworks to track model performance and user satisfaction. Drive comprehensive quality evaluation initiatives alongside rigorous safety and toxicity baselines. Partner with UXR, researchers, and engineers to identify quality signals, incorporate human feedback, and surface actionable insights on model behavior in production.
• Operational excellence: Drive visibility into data pipeline health, annotation quality, training run progress, and deployment readiness. Identify bottlenecks across teams and lead efforts to improve tooling, process, and developer velocity.
• Strategic partnership: Partner with research, safety, product, and UXR leadership on prioritization, sequencing, and tradeoffs—balancing aggressive capability scaling with strict safety requirements, user needs, and infrastructure constraints.
• Process development: Build and refine the operational patterns, ontologies, and frameworks used to scale new capability development—from prompt engineering and data generation to model behavior specification and safety guidelines.
• Vendor & partner management: Own external partner relationships supporting these workstreams, including general and safety-focused annotation vendors, evaluation tooling providers, and data partners.
Qualifications:
Required:
• 5+ years of experience in technical program management, research operations, or product execution in a fast-moving AI, ML, or research environment.
• Deep familiarity with post-training and alignment concepts (supervised fine-tuning, RLHF, AI safety frameworks, LLM evaluation) as well as model deployment/serving, sufficient to engage substantively with both research and infrastructure engineers.
• Proven ability to lead complex, multi-team programs in ambiguous, rapidly evolving environments; track record of shipping with quality and speed.
• Strong analytical mindset; comfortable working with data and user insights to measure program health, identify trends, and drive decisions.
• Proficiency in SQL and Python.
• Exceptional communication skills - able to translate deep technical work into clear narratives for leadership, and to hold detailed technical conversations with engineers across different disciplines.
• Obsessive about data integrity, operational rigor, and process quality without letting process slow teams down.
• BS in a quantitative, scientific, or technical field; MS or PhD a plus.
Preferred:
• Hands-on experience with data pipelines, annotation platforms, ML evaluation tooling, or human-in-the-loop workflows.
• Experience managing annotation vendors or external data partners.
• Familiarity with distributed training, experiment tracking, or ML infrastructure (Kubernetes, Docker, cloud) and model serving systems.
• Prior experience embedded in an AI research team, foundation model lab, or Trust & Safety engineering team.
• Direct experience managing AI safety, trust, quality eval, or red-teaming programs.
Company:
Character.ai provides open-ended conversational applications in which users create characters and converse with them. Founded in 2021, the company is headquartered in Menlo Park, USA, with a team of 51-200 employees. The company is currently Growth Stage.