GenAI Tuning Analyst
Norcross, GA · On-site
RLHF & Quality Calibration:LeadReinforcement Learning from Human Feedback (RLHF)cycles. You will review AI-generated outputs and "grade" them based on accuracy, empathy, and adherence to compliance ...
Norcross, GA · On-site
RLHF & Quality Calibration:LeadReinforcement Learning from Human Feedback (RLHF)cycles. You will review AI-generated outputs and "grade" them based on accuracy, empathy, and adherence to compliance ...
Norcross, GA · On-site
RLHF & Quality Calibration:LeadReinforcement Learning from Human Feedback (RLHF)cycles. You will review AI-generated outputs and "grade" them based on accuracy, empathy, and adherence to compliance ...
Norcross, GA · On-site
RLHF & Quality Calibration:LeadReinforcement Learning from Human Feedback (RLHF)cycles. You will review AI-generated outputs and "grade" them based on accuracy, empathy, and adherence to compliance ...
Quick apply
Norcross, GA · On-site
RLHF & Quality Calibration:LeadReinforcement Learning from Human Feedback (RLHF)cycles. You will review AI-generated outputs and "grade" them based on accuracy, empathy, and adherence to compliance ...
... RLHF, PPO, DPO, or reward model training -- and understanding of how training data quality affects model behavior Familiarity with RL frameworks (Gymnasium, dm_env) and the ability to design or ...
... RLHF, PPO, DPO, or reward model training -- and understanding of how training data quality affects model behavior Familiarity with RL frameworks (Gymnasium, dm_env) and the ability to design or ...
Atlanta, GA · On-site
... RLHF, PPO, DPO, or reward model training -- and understanding of how training data quality affects model behavior Familiarity with RL frameworks (Gymnasium, dm_env) and the ability to design or ...
Atlanta, GA · On-site
... RLHF, PPO, DPO, or reward model training -- and understanding of how training data quality affects model behavior Familiarity with RL frameworks (Gymnasium, dm_env) and the ability to design or ...
... RLHF, PPO, DPO, or reward model training -- and understanding of how training data quality affects model behavior Familiarity with RL frameworks (Gymnasium, dm_env) and the ability to design or ...
... RLHF, PPO, DPO, or reward model training -- and understanding of how training data quality affects model behavior Familiarity with RL frameworks (Gymnasium, dm_env) and the ability to design or ...
Kennesaw, GA · On-site
... RLHF, PPO, DPO, or reward model training -- and understanding of how training data quality affects model behavior Familiarity with RL frameworks (Gymnasium, dm_env) and the ability to design or ...
Kennesaw, GA · On-site
... RLHF, PPO, DPO, or reward model training -- and understanding of how training data quality affects model behavior Familiarity with RL frameworks (Gymnasium, dm_env) and the ability to design or ...
Alpharetta, GA · On-site
Apply techniques such as prompt engineering, RAG (Retrieval-Augmented Generation), fine-tuning, and RLHF to enhance model performance. * Develop and deploy autonomous AI agents using frameworks like ...
Alpharetta, GA · On-site
Apply techniques such as prompt engineering, RAG (Retrieval-Augmented Generation), fine-tuning, and RLHF to enhance model performance. * Develop and deploy autonomous AI agents using frameworks like ...
Norcross, GA · On-site
RLHF & Quality Calibration: Lead Reinforcement Learning from Human Feedback (RLHF) cycles. You will review AI-generated outputs and "grade" them based on accuracy, empathy, and adherence to ...
Norcross, GA · On-site
RLHF & Quality Calibration: Lead Reinforcement Learning from Human Feedback (RLHF) cycles. You will review AI-generated outputs and "grade" them based on accuracy, empathy, and adherence to ...
Alpharetta, GA · On-site
Apply techniques such as prompt engineering, RAG (Retrieval-Augmented Generation), fine-tuning, and RLHF to enhance model performance. * Develop and deploy autonomous AI agents using frameworks like ...
Alpharetta, GA · On-site
Apply techniques such as prompt engineering, RAG (Retrieval-Augmented Generation), fine-tuning, and RLHF to enhance model performance. * Develop and deploy autonomous AI agents using frameworks like ...
Advance Workday's proprietary capabilities in pre-training, post-training (RLHF, DPO), and domain-specific alignment for HR and Finance workflows. * Publish & Open Source: Lead Workday's contribution ...
Advance Workday's proprietary capabilities in pre-training, post-training (RLHF, DPO), and domain-specific alignment for HR and Finance workflows. * Publish & Open Source: Lead Workday's contribution ...
Norcross, GA · On-site
RLHF & Quality Calibration: Lead Reinforcement Learning from Human Feedback (RLHF) cycles. You will review AI-generated outputs and "grade" them based on accuracy, empathy, and adherence to ...
Quick apply
Norcross, GA · On-site
RLHF & Quality Calibration: Lead Reinforcement Learning from Human Feedback (RLHF) cycles. You will review AI-generated outputs and "grade" them based on accuracy, empathy, and adherence to ...
... RLHF, RAG and Knowledge graph etc. • Experience in designing and implementing Model Context Protocol (MCP) servers to enable seamless integration between AI agents, enterprise systems, and external ...
... RLHF, RAG and Knowledge graph etc. • Experience in designing and implementing Model Context Protocol (MCP) servers to enable seamless integration between AI agents, enterprise systems, and external ...
Build and optimize training using techniques such as SFT, RLHF, PPO, DPO, GRPO, RLAIF, and Constitutional AI, and understand how each affects reasoning quality, safety, latency, cost, and reliability.
Build and optimize training using techniques such as SFT, RLHF, PPO, DPO, GRPO, RLAIF, and Constitutional AI, and understand how each affects reasoning quality, safety, latency, cost, and reliability.
Specialized expertise in other topics like fine-tuning, RLHF, RAG and Knowledge graph etc. * Experience in designing and implementing Model Context Protocol (MCP) servers to enable seamless ...
Specialized expertise in other topics like fine-tuning, RLHF, RAG and Knowledge graph etc. * Experience in designing and implementing Model Context Protocol (MCP) servers to enable seamless ...
Specialized expertise in other topics like fine-tuning, RLHF, RAG and Knowledge graph etc. * Experience in designing and implementing Model Context Protocol (MCP) servers to enable seamless ...
Specialized expertise in other topics like fine-tuning, RLHF, RAG and Knowledge graph etc. * Experience in designing and implementing Model Context Protocol (MCP) servers to enable seamless ...
Fine-tuning, instruction tuning, RLHF, and domain adaptation * Azure DevOps, App Insights, Log Analytics, Key Vault, and Managed Identity integration. * Tools for inference performance testing and ...
New
Fine-tuning, instruction tuning, RLHF, and domain adaptation * Azure DevOps, App Insights, Log Analytics, Key Vault, and Managed Identity integration. * Tools for inference performance testing and ...
New
... RLHF, RLAIF, fine-tuning, or distillation to smaller models) into production Published or open-source work in agent infrastructure or evaluation tooling Cost management (FinOps) experience for LLM ...
New
... RLHF, RLAIF, fine-tuning, or distillation to smaller models) into production Published or open-source work in agent infrastructure or evaluation tooling Cost management (FinOps) experience for LLM ...
New
... RLHF, RLAIF, fine-tuning, or distillation to smaller models) into production Published or open-source work in agent infrastructure or evaluation tooling Cost management (FinOps) experience for LLM ...
New
... RLHF, RLAIF, fine-tuning, or distillation to smaller models) into production Published or open-source work in agent infrastructure or evaluation tooling Cost management (FinOps) experience for LLM ...
New
| Aspect | Rlhf | Rn |
|---|---|---|
| Required Credentials | Licensed healthcare professional, often with specialized training in mental health or behavioral health | Licensed practical nurse or registered nurse, with nursing licensure and possibly additional certifications |
| Work Environment | Behavioral health facilities, clinics, hospitals, or community health settings | Hospitals, clinics, long-term care facilities, and community health settings |
| Employer & Industry Usage | Behavioral health and mental health services | General healthcare and nursing services |
| Common Search & Comparison | Rlhf vs Rn | Rlhf vs Rn |
While Rlhf (Registered Licensed Mental Health Facilitator) focuses on mental health support and behavioral health interventions, Rn (Registered Nurse) provides broader nursing care across various medical settings. Both roles require licensure, but Rlhf specializes in mental health, whereas Rn covers general patient care.
An RLHF (Reinforcement Learning with Human Feedback) job involves training AI models using human feedback to improve their responses. Professionals in this role analyze model outputs, provide evaluations, and refine AI behavior through reinforcement learning techniques. These roles are common in AI research, content moderation, and chatbot development.
****Applicants must be authorized to work for ANY employer in the U.S.
We are unable to sponsor or take over sponsorship of an employment Visa at this time.
No agencies please.
Role Overview
As aGenAI Tuning Analyst, you willbe responsible forthe continuous improvement, accuracy, and "brand voice" of our Generative AI deployments within the Contact Center ecosystem, including agentic bots. You will work at the intersection of data science, linguistics, and customer operations to ensure our AI agents and agent-assist tools provide precise, empathetic, and compliant resolutions.
This role requires aprofessional learner; ahigh-signal individual who pairs raw brainpower and adaptability with the hunger to master the evolving science of AI performance.
Your goal is to transform generic LLMoutputinto specialized,contactcenterdomain-aware intelligence.
Key Responsibilities
Ideal CandidateSkills
Analytical Pattern Recognition:Naturallyidentifiespatterns in data and customer language, using strong data and business analytics skills to interpret trends and organize information effectively.
Customer-facingskills for both the supervisory and executive sponsors of AI deployed in USAN's cloud contact center offerings.
Comfort with Ambiguity:Able to make thoughtful decisions in gray areas, applying sound judgment todeterminehow conversational intents shouldbe categorized, merged, or preserved.
Curiosity About Customer Communication:Interested in how customers naturally express their needsandareable totranslate that language into meaningful insights for AI optimization.
Process Improvement Mindset:Enjoys working with data repeatedly while finding smarter waysprocessandanalyzingit,such as building Excel formulas, spotting systematic issues, and improving workflows.
Trust-Based Collaboration:Builds credibility and trust with clients and internal teams while working collaboratively to improve AI performance and data quality.
Functional Requirements Writing:Translate customer feedback into actionable product improvements by gathering input, researching root causes, writing clear technical requirements for developers, and analyzing post-update results (e.g., for agent scorecards/GenAI tools) before returning changes to the customer.
Required Qualifications
Experience
Technical Skills
Communication
Preferred "Bonus" Skills
Why This Role Matters
In the modern contact center, the AI isthefirst impression. The Tuning Analyst ensures that impression is not just intelligent, but human-centric and helpful.
Job Benefits:
Healthcare benefits
401K plan
Paid company holidays
Paid vacation
Business casual work environment
Annual performance based bonus program
Company Description
United States Advanced Network, In. (USAN) is a privately held corporation based out of Norcross, GA (a suburb of Atlanta, GA). USAN is an AWS Advanced Tier Partner specializing in Amazon Connect, helping organizations design and deploy scalable, AI-driven customer interactions that accelerate time to value and maximize ROI. With over 35 years of deep contact center expertise, USAN delivers modern agentic CX solutions and a white-glove approach to optimizing and managing cloud contact center environments through its managed services.
For more information, please visit us at www.usan.com
Sourced by ZipRecruiter
Software development
51 - 200 Employees
Norcross, GA, US
1989