AI Data Science Expert - Remote
Miami, FL ยท Remote
$100 - $200/hr
Familiarity with prompt engineering, AI output evaluation, fact-checking, or RLHF is a plus. * Master's, MBA, PhD, or other advanced degree is preferred.
Quick apply
Miami, FL ยท Remote
$100 - $200/hr
Familiarity with prompt engineering, AI output evaluation, fact-checking, or RLHF is a plus. * Master's, MBA, PhD, or other advanced degree is preferred.
Quick apply
Miami, FL ยท Remote
$100 - $200/hr
Familiarity with prompt engineering, AI output evaluation, fact-checking, or RLHF is a plus. * Master's, MBA, PhD, or other advanced degree is preferred.
Miami, FL ยท Remote
$100 - $200/hr
Familiarity with prompt engineering, AI output evaluation, fact-checking, or RLHF is a plus. * Master's, MBA, PhD, or other advanced degree is preferred.
Quick apply
Miami, FL ยท Remote
$100 - $200/hr
Familiarity with prompt engineering, AI output evaluation, fact-checking, or RLHF is a plus. * Master's, MBA, PhD, or other advanced degree is preferred.
Prompt engineering, SFT, RLHF, red teaming and adversarial model training, model output ranking. * DATA COLLECTION & GENERATION: From institutional languages to remote field audio collection.
Prompt engineering, SFT, RLHF, red teaming and adversarial model training, model output ranking. * DATA COLLECTION & GENERATION: From institutional languages to remote field audio collection.
Tampa, FL ยท On-site
Working knowledge of model fine-tuning approaches (LoRA, RLHF, instruction tuning, or equivalent) * Experience designing and implementing custom LLM and agent evaluation frameworks, including ...
Tampa, FL ยท On-site
Working knowledge of model fine-tuning approaches (LoRA, RLHF, instruction tuning, or equivalent) * Experience designing and implementing custom LLM and agent evaluation frameworks, including ...
Miami, FL ยท On-site
... RLHF concepts. * Deep understanding of transformer architectures, token economics, context window management, and inference optimization. * Proven track record with RAG architectures, embedding ...
Miami, FL ยท On-site
... RLHF concepts. * Deep understanding of transformer architectures, token economics, context window management, and inference optimization. * Proven track record with RAG architectures, embedding ...
Tampa, FL ยท On-site
Working knowledge of model fine-tuning approaches (LoRA, RLHF, instruction tuning, or equivalent) * Experience designing and implementing custom LLM and agent evaluation frameworks, including ...
Tampa, FL ยท On-site
Working knowledge of model fine-tuning approaches (LoRA, RLHF, instruction tuning, or equivalent) * Experience designing and implementing custom LLM and agent evaluation frameworks, including ...
Working knowledge of model fine-tuning approaches (LoRA, RLHF, instruction tuning, or equivalent) * Experience designing and implementing custom LLM and agent evaluation frameworks, including ...
Working knowledge of model fine-tuning approaches (LoRA, RLHF, instruction tuning, or equivalent) * Experience designing and implementing custom LLM and agent evaluation frameworks, including ...
Working knowledge of model fine-tuning approaches (LoRA, RLHF, instruction tuning, or equivalent) * Experience designing and implementing custom LLM and agent evaluation frameworks, including ...
Working knowledge of model fine-tuning approaches (LoRA, RLHF, instruction tuning, or equivalent) * Experience designing and implementing custom LLM and agent evaluation frameworks, including ...
Build and optimize training using techniques such as SFT, RLHF, PPO, DPO, GRPO, RLAIF, and Constitutional AI, and understand how each affects reasoning quality, safety, latency, cost, and reliability.
Build and optimize training using techniques such as SFT, RLHF, PPO, DPO, GRPO, RLAIF, and Constitutional AI, and understand how each affects reasoning quality, safety, latency, cost, and reliability.
Tampa, FL ยท On-site
Working knowledge of model fine-tuning approaches (LoRA, RLHF, instruction tuning, or equivalent) * Experience designing and implementing custom LLM and agent evaluation frameworks, including ...
Tampa, FL ยท On-site
Working knowledge of model fine-tuning approaches (LoRA, RLHF, instruction tuning, or equivalent) * Experience designing and implementing custom LLM and agent evaluation frameworks, including ...
Build and optimize training using techniques such as SFT, RLHF, PPO, DPO, GRPO, RLAIF, and Constitutional AI, and understand how each affects reasoning quality, safety, latency, cost, and reliability.
Build and optimize training using techniques such as SFT, RLHF, PPO, DPO, GRPO, RLAIF, and Constitutional AI, and understand how each affects reasoning quality, safety, latency, cost, and reliability.
Working knowledge of model fine-tuning approaches (LoRA, RLHF, instruction tuning, or equivalent) * Experience designing and implementing custom LLM and agent evaluation frameworks, including ...
Working knowledge of model fine-tuning approaches (LoRA, RLHF, instruction tuning, or equivalent) * Experience designing and implementing custom LLM and agent evaluation frameworks, including ...
... RLHF concepts. * Deep understanding of transformer architectures, token economics, context window management, and inference optimization. * Proven track record with RAG architectures, embedding ...
... RLHF concepts. * Deep understanding of transformer architectures, token economics, context window management, and inference optimization. * Proven track record with RAG architectures, embedding ...
Build and optimize training using techniques such as SFT, RLHF, PPO, DPO, GRPO, RLAIF, and Constitutional AI, and understand how each affects reasoning quality, safety, latency, cost, and reliability.
Build and optimize training using techniques such as SFT, RLHF, PPO, DPO, GRPO, RLAIF, and Constitutional AI, and understand how each affects reasoning quality, safety, latency, cost, and reliability.
An RLHF (Reinforcement Learning with Human Feedback) job involves training AI models using human feedback to improve their responses. Professionals in this role analyze model outputs, provide evaluations, and refine AI behavior through reinforcement learning techniques. These roles are common in AI research, content moderation, and chatbot development.
| Aspect | Rlhf | Rn |
|---|---|---|
| Required Credentials | Licensed healthcare professional, often with specialized training in mental health or behavioral health | Licensed practical nurse or registered nurse, with nursing licensure and possibly additional certifications |
| Work Environment | Behavioral health facilities, clinics, hospitals, or community health settings | Hospitals, clinics, long-term care facilities, and community health settings |
| Employer & Industry Usage | Behavioral health and mental health services | General healthcare and nursing services |
| Common Search & Comparison | Rlhf vs Rn | Rlhf vs Rn |
While Rlhf (Registered Licensed Mental Health Facilitator) focuses on mental health support and behavioral health interventions, Rn (Registered Nurse) provides broader nursing care across various medical settings. Both roles require licensure, but Rlhf specializes in mental health, whereas Rn covers general patient care.
The most popular types of Rlhf jobs in Florida are:
For Rlhf jobs in Florida, the most frequently searched job titles are:
The top searched job categories for Rlhf jobs in Florida are:
Cities in Florida with the most Rlhf job openings:

$100 - $200/hr
Part-time
Posted 23 days ago
Job Type: Contractor (Part-Time)
Location: Remote
We are seeking experienced AI Data Science Domain Experts to contribute their expertise to an innovative project focused on advancing next-generation AI systems. In this role, you will review, evaluate, and refine AI-generated technical and analytical content to improve model accuracy, reasoning, and overall performance. No prior AI experience is required—your data science expertise, analytical thinking, and communication skills are what matter most.
Key Responsibilities