They are seeking a Tech Lead Manager for their ML Systems team to build and optimize a training and ... RLHF/RLVR and related algorithms like PPO/GRPO etc. • Strong software engineering skills ...
They are seeking a Tech Lead Manager for their ML Systems team to build and optimize a training and ... RLHF/RLVR and related algorithms like PPO/GRPO etc. • Strong software engineering skills ...
Post-Training Research Scientist
New York, NY · On-site
$165K - $300K/yr
RLHF, DPO, RLAIF, or related methods * Deep understanding of distributed training infrastructure: multi-node GPU clusters, training stability, checkpointing * Track record managing large-scale ...
Post-Training Research Scientist
New York, NY · On-site
$165K - $300K/yr
RLHF, DPO, RLAIF, or related methods * Deep understanding of distributed training infrastructure: multi-node GPU clusters, training stability, checkpointing * Track record managing large-scale ...
... RLHF) during pre-training, fine-tuning, and production evaluations. This is the main innovation ... Experience in program management, including planning, organizing, and managing resources.
... RLHF) during pre-training, fine-tuning, and production evaluations. This is the main innovation ... Experience in program management, including planning, organizing, and managing resources.
Developer
New York, NY · On-site
Deep understanding of Data preprocessing, Prompt management, Caching, Validation, Advanced RAG, RLHF, and success measurement. * Thorough understanding of LLMOps, data pipelines and other common ...
Developer
New York, NY · On-site
Deep understanding of Data preprocessing, Prompt management, Caching, Validation, Advanced RAG, RLHF, and success measurement. * Thorough understanding of LLMOps, data pipelines and other common ...
Experience with post-training methods like RLHF/RLVR and related algorithms like PPO/GRPO etc ... manage applicants' needs, provide our services, and comply with applicable laws. Any information we ...
Experience with post-training methods like RLHF/RLVR and related algorithms like PPO/GRPO etc ... manage applicants' needs, provide our services, and comply with applicable laws. Any information we ...
... RLHF) during pre-training, fine-tuning, and production evaluations. This is the main innovation ... access management, authorization, and authentication. You'll also get widespread exposure to the ...
... RLHF) during pre-training, fine-tuning, and production evaluations. This is the main innovation ... access management, authorization, and authentication. You'll also get widespread exposure to the ...
AI/LLM Product Director - Executive Director
New York, NY · On-site
$254K - $266K/yr
Oversees the product roadmap, vision, development, execution, risk management, and business growth ... Feedback (RLHF), Retrieval-Augmented Generation (RAG), and Agents to enhance user experiences.
AI/LLM Product Director - Executive Director
New York, NY · On-site
$254K - $266K/yr
Oversees the product roadmap, vision, development, execution, risk management, and business growth ... Feedback (RLHF), Retrieval-Augmented Generation (RAG), and Agents to enhance user experiences.
Delivery Lead
New York, NY · Remote
$110K - $140K/yr
... RLHF, annotation, model evaluation) * STEM background or strong technical fluency * Python & REACT working knowledge * Experience managing distributed contributor workforces at scale * Background in ...
Quick apply
Delivery Lead
New York, NY · Remote
$110K - $140K/yr
... RLHF, annotation, model evaluation) * STEM background or strong technical fluency * Python & REACT working knowledge * Experience managing distributed contributor workforces at scale * Background in ...
Software Engineer, Identity
Manhattan, NY · On-site
... RLHF) during pre-training, fine-tuning, and production evaluations. This is the main innovation ... access management, authorization, and authentication. You'll also get widespread exposure to the ...
Software Engineer, Identity
Manhattan, NY · On-site
... RLHF) during pre-training, fine-tuning, and production evaluations. This is the main innovation ... access management, authorization, and authentication. You'll also get widespread exposure to the ...
... RLHF systems. Compensation packages at Scale for eligible roles include base salary, equity, and ... manage applicants' needs, provide our services, and comply with applicable laws. Any information we ...
... RLHF systems. Compensation packages at Scale for eligible roles include base salary, equity, and ... manage applicants' needs, provide our services, and comply with applicable laws. Any information we ...
Direct experience with AI/ML platform products - data labeling, RLHF, fine-tuning workflows, or ... manage applicants' needs, provide our services, and comply with applicable laws. Any information we ...
Direct experience with AI/ML platform products - data labeling, RLHF, fine-tuning workflows, or ... manage applicants' needs, provide our services, and comply with applicable laws. Any information we ...
Manage the project lifecycle from ideation and scoping to deployment and post-launch support ... RLHF, multi-task learning). * Model Optimization: Expertise in model compression and quantization ...
Manage the project lifecycle from ideation and scoping to deployment and post-launch support ... RLHF, multi-task learning). * Model Optimization: Expertise in model compression and quantization ...
... to manage their own context in long-horizon tasks. This is applied research with direct product ... RLHF, GRPO, PPO, RLVR, reward modeling, RL scaling laws Code generation and coding agents ...
... to manage their own context in long-horizon tasks. This is applied research with direct product ... RLHF, GRPO, PPO, RLVR, reward modeling, RL scaling laws Code generation and coding agents ...
Familiarity with SFT, DPO, RLHF, or similar techniques. • Understanding of evaluation methodology ... GPUs, compute management, debugging common training failures. You don't need to be an infra ...
Familiarity with SFT, DPO, RLHF, or similar techniques. • Understanding of evaluation methodology ... GPUs, compute management, debugging common training failures. You don't need to be an infra ...
Manager, Data Science - AI Foundations Data is at the center of everything we do. As a startup, we ... RLHF. * You have an engineering mindset as shown by a track record of delivering models at scale ...
Manager, Data Science - AI Foundations Data is at the center of everything we do. As a startup, we ... RLHF. * You have an engineering mindset as shown by a track record of delivering models at scale ...
Manager, Data Science - AI Foundations Data is at the center of everything we do. As a startup, we ... RLHF. * You have an engineering mindset as shown by a track record of delivering models at scale ...
Manager, Data Science - AI Foundations Data is at the center of everything we do. As a startup, we ... RLHF. * You have an engineering mindset as shown by a track record of delivering models at scale ...
Manager, Data Science - AI Foundations Data is at the center of everything we do. As a startup, we ... RLHF. * You have an engineering mindset as shown by a track record of delivering models at scale ...
Manager, Data Science - AI Foundations Data is at the center of everything we do. As a startup, we ... RLHF. * You have an engineering mindset as shown by a track record of delivering models at scale ...
Senior Manager, Data Science - AI Foundations Data is at the center of everything we do. As a ... RLHF. * You have an engineering mindset as shown by a track record of delivering models at scale ...
Senior Manager, Data Science - AI Foundations Data is at the center of everything we do. As a ... RLHF. * You have an engineering mindset as shown by a track record of delivering models at scale ...
Senior Manager, Data Science - AI Foundations Data is at the center of everything we do. As a ... RLHF. * You have an engineering mindset as shown by a track record of delivering models at scale ...
Senior Manager, Data Science - AI Foundations Data is at the center of everything we do. As a ... RLHF. * You have an engineering mindset as shown by a track record of delivering models at scale ...
Senior Manager, Data Science - AI Foundations Data is at the center of everything we do. As a ... RLHF. * You have an engineering mindset as shown by a track record of delivering models at scale ...
Senior Manager, Data Science - AI Foundations Data is at the center of everything we do. As a ... RLHF. * You have an engineering mindset as shown by a track record of delivering models at scale ...
Manager Rlhf information
See Bronx, NY salary details
$25.5K - $34.2K
9% of jobs
$34.2K - $42.9K
15% of jobs
$43.5K is the 25th percentile. Wages below this are outliers.
$42.9K - $51.5K
17% of jobs
The median wage is $54.5K / yr.
$51.5K - $60.2K
27% of jobs
$65.5K is the 75th percentile. Wages above this are outliers.
$60.2K - $68.9K
12% of jobs
$68.9K - $77.5K
8% of jobs
$77.5K - $86.2K
4% of jobs
$86.2K - $94.8K
3% of jobs
$94.8K - $103.5K
2% of jobs
$103.5K - $112.2K
2% of jobs
$112.2K - $120.8K
1% of jobs
$25.5K
$62K
$120.8K
How much do manager rlhf jobs pay per year?
Scale AI rating
8.5
Based on 9 frontline employees who took The Breakroom Quiz
78th of 245 rated software companies
Job description
Scale AI is a company focused on developing reliable AI systems for critical decisions. They are seeking a Tech Lead Manager for their ML Systems team to build and optimize a training and inference framework that supports machine learning research and development.
Responsibilities:
• Build, profile and optimize our training and inference framework.
• Collaborate with ML and research teams to accelerate their research and development, and enable them to develop the next generation of models and data curation.
• Research and integrate state-of-the-art technologies to optimize our ML system.
Qualifications:
Required:
• Passionate about system optimization
• Experience with multi-node LLM training and inference
• Experience with developing large-scale distributed ML systems
• Experience with post-training methods like RLHF/RLVR and related algorithms like PPO/GRPO etc.
• Strong software engineering skills, proficient in frameworks and tools such as CUDA, Pytorch, transformers, flash attention, etc.
• Strong written and verbal communication skills to operate in a cross functional team environment.
Preferred:
• Demonstrated expertise in post-training methods and/or next generation use cases for large language models including instruction tuning, RLHF, tool use, reasoning, agents, and multimodal, etc.
Company:
Scale’s mission is to develop reliable AI systems for the world’s most important decisions. Founded in 2016, the company is headquartered in San Francisco, USA, with a team of 501-1000 employees. The company is currently Late Stage.
About Scale AI
Sourced by ZipRecruiter
Industry
Software development
Company size
201 - 500 Employees
Headquarters location
San Francisco, CA, US
Year founded
2016