This is a full-time, fully remote engagement requiring a commitment of approximately 35 hours per ... Experience supporting AI model training, evaluation, RLHF, or large-scale annotation programs.
This is a full-time, fully remote engagement requiring a commitment of approximately 35 hours per ... Experience supporting AI model training, evaluation, RLHF, or large-scale annotation programs.
AI Prompt Engineer
$60 - $70K/hr
Remote Salary: $60-$70K Must-Haves * Linguistic Precision: Exceptional command of English grammar ... Prior experience with RLHF, data labeling platforms (like Scale AI, Labelbox), or QA roles. * Basic ...
AI Prompt Engineer
$60 - $70K/hr
Remote Salary: $60-$70K Must-Haves * Linguistic Precision: Exceptional command of English grammar ... Prior experience with RLHF, data labeling platforms (like Scale AI, Labelbox), or QA roles. * Basic ...
... RLHF pipelines, and custom dataset delivery translate into long-term, multi-million-dollar ... Paid Time Off (vacation, sick leave, parental leave, holidays). * 100% remote work. * The ability ...
... RLHF pipelines, and custom dataset delivery translate into long-term, multi-million-dollar ... Paid Time Off (vacation, sick leave, parental leave, holidays). * 100% remote work. * The ability ...
Palo Alto, CA or Seattle, WA (Hybrid/Remote) About the Team Centific AI Research advances ... TD(0), TD(λ), eligibility traces, bootstrapping methods LLM Alignment & Post-Training • RLHF ...
Palo Alto, CA or Seattle, WA (Hybrid/Remote) About the Team Centific AI Research advances ... TD(0), TD(λ), eligibility traces, bootstrapping methods LLM Alignment & Post-Training • RLHF ...
Palo Alto, CA or Seattle, WA (Hybrid/Remote) About the Team Centific AI Research advances ... TD(0), TD(), eligibility traces, bootstrapping methods LLM Alignment & Post-Training RLHF pipelines:
Palo Alto, CA or Seattle, WA (Hybrid/Remote) About the Team Centific AI Research advances ... TD(0), TD(), eligibility traces, bootstrapping methods LLM Alignment & Post-Training RLHF pipelines:
Principal Applied Scientist, Agentic AI
$181K - $290K/yr
... RLHF, RLAIF, or DPO for multiobjective optimization. * Develop reward models and objective ... This role has been categorized as a Remote position. "Remote" employees do not have a permanent ...
Principal Applied Scientist, Agentic AI
$181K - $290K/yr
... RLHF, RLAIF, or DPO for multiobjective optimization. * Develop reward models and objective ... This role has been categorized as a Remote position. "Remote" employees do not have a permanent ...
Principal Machine Learning Engineer, Artificial Intelligence (AI) Required, Work From Home
San Francisco, CA · Remote
This position is 100% Remote. Principal Machine Learning Engineer Responsibilities: - Architect and ... with RLHF pipelines (PPO, DPO, ORPO). - Experience training or deploying multimodal or diffusion ...
Quick apply
Principal Machine Learning Engineer, Artificial Intelligence (AI) Required, Work From Home
San Francisco, CA · Remote
This position is 100% Remote. Principal Machine Learning Engineer Responsibilities: - Architect and ... with RLHF pipelines (PPO, DPO, ORPO). - Experience training or deploying multimodal or diffusion ...
Senior Machine Learning Engineer, Model Training and Reinforcement Learning
Palo Alto, CA · On-site +1
$122K - $168K/yr
... such as RLHF/RLAIF, PPO, and GRPO. * Build reward functions, judge models, verifiers, task ... Remote work reimbursement: Up to $85/month for mobile and internet. * Disability & life insurance
Senior Machine Learning Engineer, Model Training and Reinforcement Learning
Palo Alto, CA · On-site +1
$122K - $168K/yr
... such as RLHF/RLAIF, PPO, and GRPO. * Build reward functions, judge models, verifiers, task ... Remote work reimbursement: Up to $85/month for mobile and internet. * Disability & life insurance
Principal Machine Learning Engineer, Artificial Intelligence (AI) Required, Work From Home
San Francisco, CA · Remote
This position is 100% Remote. Principal Machine Learning Engineer Responsibilities: - Architect and ... with RLHF pipelines (PPO, DPO, ORPO). - Experience training or deploying multimodal or diffusion ...
Principal Machine Learning Engineer, Artificial Intelligence (AI) Required, Work From Home
San Francisco, CA · Remote
This position is 100% Remote. Principal Machine Learning Engineer Responsibilities: - Architect and ... with RLHF pipelines (PPO, DPO, ORPO). - Experience training or deploying multimodal or diffusion ...
Principal Machine Learning Engineer, Artificial Intelligence (AI) Required, Work From Home
San Francisco, CA · On-site +1
This position is 100% Remote. Principal Machine Learning Engineer Responsibilities: - Architect and ... with RLHF pipelines (PPO, DPO, ORPO). - Experience training or deploying multimodal or diffusion ...
Principal Machine Learning Engineer, Artificial Intelligence (AI) Required, Work From Home
San Francisco, CA · On-site +1
This position is 100% Remote. Principal Machine Learning Engineer Responsibilities: - Architect and ... with RLHF pipelines (PPO, DPO, ORPO). - Experience training or deploying multimodal or diffusion ...
High Volume (TOFU) Recruiter
Austin, TX · On-site +1
$55K - $100K/yr
San Francisco, CA preferred; open to other remote options About the Role HumanSignal Services runs ... Familiarity with AI data operations, annotation, or RLHF workforce programs * Experience with ATS ...
Quick apply
High Volume (TOFU) Recruiter
Austin, TX · On-site +1
$55K - $100K/yr
San Francisco, CA preferred; open to other remote options About the Role HumanSignal Services runs ... Familiarity with AI data operations, annotation, or RLHF workforce programs * Experience with ATS ...
AI Evaluation Experience Previous experience with AI evaluation, RLHF, prompt engineering, AI ... Maintain productivity and quality expectations while working independently in a remote environment.
AI Evaluation Experience Previous experience with AI evaluation, RLHF, prompt engineering, AI ... Maintain productivity and quality expectations while working independently in a remote environment.
High Volume (TOFU) Recruiter
Columbus, OH · On-site +1
$55K - $100K/yr
San Francisco, CA preferred; open to other remote options About the Role HumanSignal Services runs ... Familiarity with AI data operations, annotation, or RLHF workforce programs * Experience with ATS ...
Quick apply
High Volume (TOFU) Recruiter
Columbus, OH · On-site +1
$55K - $100K/yr
San Francisco, CA preferred; open to other remote options About the Role HumanSignal Services runs ... Familiarity with AI data operations, annotation, or RLHF workforce programs * Experience with ATS ...
High Volume (TOFU) Recruiter
San Francisco, CA · On-site +1
$55K - $100K/yr
San Francisco, CA preferred; open to other remote options About the Role HumanSignal Services runs ... Familiarity with AI data operations, annotation, or RLHF workforce programs * Experience with ATS ...
Quick apply
High Volume (TOFU) Recruiter
San Francisco, CA · On-site +1
$55K - $100K/yr
San Francisco, CA preferred; open to other remote options About the Role HumanSignal Services runs ... Familiarity with AI data operations, annotation, or RLHF workforce programs * Experience with ATS ...
AI Evaluation Experience Previous experience with AI evaluation, RLHF, prompt engineering, AI ... Maintain productivity and quality expectations while working independently in a remote environment.
AI Evaluation Experience Previous experience with AI evaluation, RLHF, prompt engineering, AI ... Maintain productivity and quality expectations while working independently in a remote environment.
Remote Scope of Work * Develop complex, adversarial multi-turn conversations and task-based ... Prior experience in AI human data environments (RLHF, SFT, evaluations, annotation, or prompt ...
Remote Scope of Work * Develop complex, adversarial multi-turn conversations and task-based ... Prior experience in AI human data environments (RLHF, SFT, evaluations, annotation, or prompt ...
Background in AI data vendors, GenAI platforms, or evaluation and RLHF providers. Experience with ... Hybrid or remote first, with occasional travel to customers and internal leadership meetings.
Background in AI data vendors, GenAI platforms, or evaluation and RLHF providers. Experience with ... Hybrid or remote first, with occasional travel to customers and internal leadership meetings.
AI Evaluation Experience Previous experience with AI evaluation, RLHF, prompt engineering, AI ... Maintain productivity and quality expectations while working independently in a remote environment.
AI Evaluation Experience Previous experience with AI evaluation, RLHF, prompt engineering, AI ... Maintain productivity and quality expectations while working independently in a remote environment.
AI Evaluation Experience Previous experience with AI evaluation, RLHF, prompt engineering, AI ... Maintain productivity and quality expectations while working independently in a remote environment.
AI Evaluation Experience Previous experience with AI evaluation, RLHF, prompt engineering, AI ... Maintain productivity and quality expectations while working independently in a remote environment.
High Volume (TOFU) Recruiter
Dallas, TX · On-site +1
$55K - $100K/yr
San Francisco, CA preferred; open to other remote options About the Role HumanSignal Services runs ... Familiarity with AI data operations, annotation, or RLHF workforce programs * Experience with ATS ...
Quick apply
High Volume (TOFU) Recruiter
Dallas, TX · On-site +1
$55K - $100K/yr
San Francisco, CA preferred; open to other remote options About the Role HumanSignal Services runs ... Familiarity with AI data operations, annotation, or RLHF workforce programs * Experience with ATS ...
Remote Rlhf information
What is a remote RLHF?
How does a remote RLHF specialist typically collaborate with other team members?
What are the key skills and qualifications needed to thrive as a remote RLHF engineer, and why are they important?
What is the difference between Remote Rlhf vs Remote Rlhf?
| Aspect | Remote Rlhf | Remote Rlhf |
|---|---|---|
| Credentials | Typically requires certification in mental health or counseling, such as LPC or LCSW | Similar credentials, often with additional training in specific therapy methods |
| Work Environment | Remote, client-facing sessions via telehealth platforms | Remote, providing therapy or support services online |
| Industry Usage | Common in mental health, therapy, and counseling sectors | Used in mental health and support services, often interchangeably with Rlhf |
Remote Rlhf and Remote Rlhf are similar roles in mental health support, primarily differing in specific certifications or training focus. Both roles involve providing remote therapy or support services via telehealth platforms, making them highly comparable in work environment and industry usage.
What cities are hiring for Remote Rlhf jobs?
Cities with the most Remote Rlhf job openings:
What are the most commonly searched types of Rlhf jobs?
The most popular types of Rlhf jobs are:
What states have the most Remote Rlhf jobs?
States with the most job openings for Remote Rlhf jobs include:
What job categories do people searching Remote Rlhf jobs look for?
The top searched job categories for Remote Rlhf jobs are:
What other helpful pages are available for Remote Rlhf?
Other pages related to Remote Rlhf:

AI Rater Guidelines Writer (Linguist / Instructional Designer)
Remote
$45 - $65/hr
Part-time
Re-posted 15 days ago
Job description
Compensation: $45-$65 per hour
Join a cutting-edge AI initiative focused on building the next generation of foundational AI models. We are seeking experienced Linguists, Instructional Designers, Technical Writers, and AI Content Specialists who excel at transforming complex, ambiguous program requirements into clear, structured, and actionable guidance for human evaluators.
In this role, you will play a critical part in improving AI training quality by creating precise rating guidelines, evaluation rubrics, and reviewer instructions across diverse domains including finance, retail, insurance, legal, and sports. Your expertise will help ensure consistency, accuracy, and reliability in human evaluations that directly shape advanced AI systems.
This is a full-time, fully remote engagement requiring a commitment of approximately 35 hours per week during standard weekdays.
Requirements
Key Responsibilities
- Translate complex and ambiguous program requirements into clear, concise, and easy-to-follow rater guidelines.
- Design comprehensive evaluation rubrics and scoring frameworks that enable consistent decision-making across a variety of AI tasks.
- Review existing documentation to identify ambiguity, inconsistencies, contradictions, and coverage gaps, then revise guidelines to improve clarity and usability.
- Convert subject-matter requirements from domains such as finance, retail, insurance, legal, and sports into structured evaluation instructions that non-domain raters can apply confidently.
- Collaborate with research teams, program managers, and subject matter experts to maintain consistency across multiple guideline sets.
- Develop documentation that addresses edge cases, exceptions, and complex evaluation scenarios while minimizing reviewer escalation.
- Minimum 3 years of professional experience in Linguistics, Instructional Design, Technical Writing, AI Content Development, or a closely related field.
- Direct experience creating, refining, or maintaining evaluation guidelines, rating rubrics, or reviewer instructions within Generative AI, RLHF, human evaluation, or AI data annotation environments.
- Demonstrated ability to work across multiple subject areas and convert domain-specific knowledge into clear, structured documentation.
- Proven experience resolving ambiguity and improving written specifications with measurable before-and-after improvements.
- Strong analytical thinking with exceptional attention to detail.
- Demonstrated career progression and professional growth.
- Ability to commit reliably to 35+ hours per week during weekdays.
- Outstanding written communication skills with the ability to explain nuanced concepts in a precise, consistent, and easy-to-understand manner.
- Experience supporting AI model training, evaluation, RLHF, or large-scale annotation programs.
- Background working with cross-functional teams including researchers, engineers, product managers, and subject matter experts.
- Familiarity with structured documentation standards, quality assurance methodologies, and guideline governance.
- Experience designing documentation that supports scalable, high-quality human evaluation processes.
- Help define how human evaluators assess next-generation AI systems.
- Influence the quality and consistency of AI training data across multiple industries.
- Collaborate with multidisciplinary experts working on advanced Generative AI initiatives.
- Solve complex language and reasoning challenges while improving AI evaluation standards at scale.
- Enjoy the flexibility of a fully remote engagement while contributing to high-impact AI research.
We are committed to creating an inclusive workplace and welcome applications from qualified professionals of all backgrounds. Reasonable accommodations are available throughout the application and engagement process.
Contract & Engagement Details
- Independent contractor engagement.
- Fully remote with flexible working arrangements.
- Expected commitment of approximately 35 hours per week during weekdays.
- Project duration may vary depending on business requirements and individual performance.
- Work does not require access to confidential or proprietary information from any current or former employer.
- Payments are issued weekly based on approved work completed.
- At this time, we are unable to support H1-B or STEM OPT candidates.