You'll identify everything from capability gaps and reasoning errors to robustness and alignment ... home and work life. * Experience with post-training techniques such as RLHF, preference modeling ...
You'll identify everything from capability gaps and reasoning errors to robustness and alignment ... home and work life. * Experience with post-training techniques such as RLHF, preference modeling ...
Post-Training Research Scientist
New York, NY · On-site
$165K - $300K/yr
... build RLHF, DPO, and reward modeling capabilities from the ground up. This is a greenfield role ... Flexible in-office days with budget for home office setup The base pay for this role will be ...
Post-Training Research Scientist
New York, NY · On-site
$165K - $300K/yr
... build RLHF, DPO, and reward modeling capabilities from the ground up. This is a greenfield role ... Flexible in-office days with budget for home office setup The base pay for this role will be ...
Hands‑on experience with LLM fine‑tuning, prompt engineering, RLHF, and deployment at scale ... Hybrid work environment with home and office schedule (2+ days in office per week) and work from ...
Hands‑on experience with LLM fine‑tuning, prompt engineering, RLHF, and deployment at scale ... Hybrid work environment with home and office schedule (2+ days in office per week) and work from ...
Quantitative Software Engineer: Techniques Engineering
New York, NY · On-site
$165K - $300K/yr
... from prototype-stage code to production systems. Preferred Skills * Experience with CUDA ... Experience in agent evaluation frameworks, RLHF, RLVR, policy optimization, or synthetic data ...
Quantitative Software Engineer: Techniques Engineering
New York, NY · On-site
$165K - $300K/yr
... from prototype-stage code to production systems. Preferred Skills * Experience with CUDA ... Experience in agent evaluation frameworks, RLHF, RLVR, policy optimization, or synthetic data ...
From Home Rlhf information
What is a from home RLHF?
What are the key skills and qualifications needed to thrive as a remote RLHF specialist, and why are they important?
What are some common challenges faced by remote RLHF professionals, and how can they be managed?
What is the difference between From Home Rlhf vs From Home Customer Service Representative?
| Aspect | From Home Rlhf | From Home Customer Service Representative |
|---|---|---|
| Required Credentials | High school diploma or equivalent, basic computer skills | High school diploma or equivalent, customer service experience |
| Work Environment | Remote, home-based | Remote, home-based |
| Industry Usage | Healthcare, insurance, or related fields | Retail, telecom, or service industries |
| Common Search Intent | Remote healthcare or insurance roles | Customer support jobs from home |
From Home Rlhf typically refers to remote roles in healthcare or insurance sectors, requiring specific industry knowledge. From Home Customer Service Representative positions are more general, focusing on customer support across various industries. Both roles are home-based, but they differ in industry focus and required experience.
What are the most commonly searched types of Rlhf jobs in New York?
The most popular types of Rlhf jobs in New York are:
What are popular job titles related to From Home Rlhf jobs in New York?
For From Home Rlhf jobs in New York, the most frequently searched job titles are:
- Machine Learning Ai Developer
- Machine Learning Engineer Intern
- Temporary Machine Learning Engineer
- Machine Learning Infrastructure Engineer
- Contract Machine Learning Software Engineer
- Senior Machine Learning Engineer
- Google Cloud Machine Learning Engineer
- Machine Learning Engineer Biotech
- Urgently Hiring Machine Learning Engineer New Grad
- Manager Edge Ai Machine Learning
What job categories do people searching From Home Rlhf jobs in New York look for?
The top searched job categories for From Home Rlhf jobs in New York are:
What cities in New York are hiring for From Home Rlhf jobs?
Cities in New York with the most From Home Rlhf job openings:
Machine Learning Research Scientist (Evaluations)
Manhattan, NY • On-site
Other
Medical, Dental, Vision, PTO
Posted 15 days ago
Scale AI rating
8.5
Based on 9 frontline employees who took The Breakroom Quiz
Job description
- Scale works with the industry’s leading AI labs to provide high quality data and accelerate progress in GenAI research
- We are looking for Research Scientists and Research Engineers with expertise in LLM post-training (SFT, RLHF, reward modeling) and evaluation
- This role is on the evaluation pod within the GenAI Research Organization and will focus on building benchmarks and diagnosing model failure modes in both text and multimodal modalities
- In this role, you will develop rigorous evaluations and diagnostic methods that reveal where frontier models fail and why
- You will collaborate with researchers and engineers to define best practices in evaluation-driven AI development
- You will also partner with top foundation model labs to translate failure analysis into technical and strategic input on the next generation of generative AI models
- Analyze model behavior to identify, characterize, and diagnose failure modes in frontier LLMs and Agents. You’ll identify everything from capability gaps and reasoning errors to robustness and alignment issues, all focusing on RCA
- Design and build benchmarks and evaluation methods that measure LLM capabilities in both text and multimodal modalities
- Apply post-training expertise (SFT, RLHF, reward modeling) to connect observed failures to the data and training interventions that address them
- Publish research findings in top-tier AI conferences
- Health & Wellbeing: Our holistic approach to supporting Scaliens includes comprehensive health coverage, dental and vision insurance, mental healthcare services, and more. PTO policies and accommodating schedules ensure you’ll get time off when you need it to relax and recharge. Note that our offerings may vary by region as we strive to respond to the unique needs of Scaliens around the globe.
- Personal & Career Growth: Continuously learn and grow through annual learning & development stipend, attending leadership breakfasts, manager training, speaker series, and joining an ERG.
- Building Scale Community: We welcome guests to our offices, and you can expect to see Scalien families and friends around. Join local happy hours, and accept invites to game nights, book clubs, and many other employee-led community events.
- Parental Support: Balancing work and family is essential, and Scale understands the importance of having adequate leave policies in place to promote a healthy home and work life.
- Experience with post-training techniques such as RLHF, preference modeling, or instruction tuning, and with LLM evaluation or benchmark development
- Excellent written and verbal communication skills
- Published research in areas of machine learning at major conferences (NeurIPS, ICML, ICLR, ACL, EMNLP, CVPR, etc.) and/or journals
- Ph.D. or Master’s degree in Computer Science, Machine Learning, AI, or a related field
- Previous experience in a customer facing role
- Deep understanding of deep learning, reinforcement learning, and large-scale model fine-tuning
About Scale AI
Sourced by ZipRecruiter
Industry
Software development
Company size
201 - 500 Employees
Headquarters location
San Francisco, CA, US
Year founded
2016