They are seeking a Tech Lead Manager for their ML Systems team to build and optimize a training and ... RLHF/RLVR and related algorithms like PPO/GRPO etc. • Strong software engineering skills ...
They are seeking a Tech Lead Manager for their ML Systems team to build and optimize a training and ... RLHF/RLVR and related algorithms like PPO/GRPO etc. • Strong software engineering skills ...
Post-Training Research Scientist
New York, NY · On-site
$165K - $300K/yr
RLHF, DPO, RLAIF, or related methods * Deep understanding of distributed training infrastructure: multi-node GPU clusters, training stability, checkpointing * Track record managing large-scale ...
Post-Training Research Scientist
New York, NY · On-site
$165K - $300K/yr
RLHF, DPO, RLAIF, or related methods * Deep understanding of distributed training infrastructure: multi-node GPU clusters, training stability, checkpointing * Track record managing large-scale ...
Developer
New York, NY · On-site
Deep understanding of Data preprocessing, Prompt management, Caching, Validation, Advanced RAG, RLHF, and success measurement. * Thorough understanding of LLMOps, data pipelines and other common ...
Developer
New York, NY · On-site
Deep understanding of Data preprocessing, Prompt management, Caching, Validation, Advanced RAG, RLHF, and success measurement. * Thorough understanding of LLMOps, data pipelines and other common ...
AI/LLM Product Director - Executive Director
$254K - $266K/yr
Oversees the product roadmap, vision, development, execution, risk management, and business growth ... Feedback (RLHF), Retrieval-Augmented Generation (RAG), and Agents to enhance user experiences.
AI/LLM Product Director - Executive Director
$254K - $266K/yr
Oversees the product roadmap, vision, development, execution, risk management, and business growth ... Feedback (RLHF), Retrieval-Augmented Generation (RAG), and Agents to enhance user experiences.
Delivery Lead
New York, NY · Remote
$110K - $140K/yr
... RLHF, annotation, model evaluation) * STEM background or strong technical fluency * Python & REACT working knowledge * Experience managing distributed contributor workforces at scale * Background in ...
Quick apply
Delivery Lead
New York, NY · Remote
$110K - $140K/yr
... RLHF, annotation, model evaluation) * STEM background or strong technical fluency * Python & REACT working knowledge * Experience managing distributed contributor workforces at scale * Background in ...
Manage the project lifecycle from ideation and scoping to deployment and post-launch support ... RLHF, multi-task learning). * Model Optimization: Expertise in model compression and quantization ...
Manage the project lifecycle from ideation and scoping to deployment and post-launch support ... RLHF, multi-task learning). * Model Optimization: Expertise in model compression and quantization ...
... to manage their own context in long-horizon tasks. This is applied research with direct product ... RLHF, GRPO, PPO, RLVR, reward modeling, RL scaling laws Code generation and coding agents ...
... to manage their own context in long-horizon tasks. This is applied research with direct product ... RLHF, GRPO, PPO, RLVR, reward modeling, RL scaling laws Code generation and coding agents ...
Familiarity with SFT, DPO, RLHF, or similar techniques. • Understanding of evaluation methodology ... GPUs, compute management, debugging common training failures. You don't need to be an infra ...
Familiarity with SFT, DPO, RLHF, or similar techniques. • Understanding of evaluation methodology ... GPUs, compute management, debugging common training failures. You don't need to be an infra ...
Manager, Data Science - AI Foundations Data is at the center of everything we do. As a startup, we ... RLHF. * You have an engineering mindset as shown by a track record of delivering models at scale ...
Manager, Data Science - AI Foundations Data is at the center of everything we do. As a startup, we ... RLHF. * You have an engineering mindset as shown by a track record of delivering models at scale ...
Manager, Data Science - AI Foundations Data is at the center of everything we do. As a startup, we ... RLHF. * You have an engineering mindset as shown by a track record of delivering models at scale ...
Manager, Data Science - AI Foundations Data is at the center of everything we do. As a startup, we ... RLHF. * You have an engineering mindset as shown by a track record of delivering models at scale ...
Manager, Data Science - AI Foundations Data is at the center of everything we do. As a startup, we ... RLHF. * You have an engineering mindset as shown by a track record of delivering models at scale ...
Manager, Data Science - AI Foundations Data is at the center of everything we do. As a startup, we ... RLHF. * You have an engineering mindset as shown by a track record of delivering models at scale ...
Senior Manager, Data Science - AI Foundations Data is at the center of everything we do. As a ... RLHF. * You have an engineering mindset as shown by a track record of delivering models at scale ...
Senior Manager, Data Science - AI Foundations Data is at the center of everything we do. As a ... RLHF. * You have an engineering mindset as shown by a track record of delivering models at scale ...
Senior Manager, Data Science - AI Foundations Data is at the center of everything we do. As a ... RLHF. * You have an engineering mindset as shown by a track record of delivering models at scale ...
Senior Manager, Data Science - AI Foundations Data is at the center of everything we do. As a ... RLHF. * You have an engineering mindset as shown by a track record of delivering models at scale ...
Senior Manager, Data Science - AI Foundations Data is at the center of everything we do. As a ... RLHF. * You have an engineering mindset as shown by a track record of delivering models at scale ...
Senior Manager, Data Science - AI Foundations Data is at the center of everything we do. As a ... RLHF. * You have an engineering mindset as shown by a track record of delivering models at scale ...
GenAI Product Engineering Lead
New York, NY · Remote
$104K - $138K/yr
Key Responsibilities Technical Leadership & Team Management * Lead, mentor, and grow a high ... RLHF) or synthetic data augmentation. • Establish model-drift detection and retraining triggers ...
Quick apply
GenAI Product Engineering Lead
New York, NY · Remote
$104K - $138K/yr
Key Responsibilities Technical Leadership & Team Management * Lead, mentor, and grow a high ... RLHF) or synthetic data augmentation. • Establish model-drift detection and retraining triggers ...
ML Researcher, Apple Foundation Models
$184K - $324K/yr
... to manage their own context in long-horizon tasks. This is applied research with direct product ... RLHF, GRPO, PPO, RLVR, reward modeling, RL scaling laws Code generation and coding agents ...
ML Researcher, Apple Foundation Models
$184K - $324K/yr
... to manage their own context in long-horizon tasks. This is applied research with direct product ... RLHF, GRPO, PPO, RLVR, reward modeling, RL scaling laws Code generation and coding agents ...
AI Engineer
New York, NY · On-site
$200K - $300K/yr
Design strategies to manage latency, output variance, and graceful error handling at scale ... Research or applied experience with LLM agents, RL (offline/online, RLHF/RLAIF), constrained ...
AI Engineer
New York, NY · On-site
$200K - $300K/yr
Design strategies to manage latency, output variance, and graceful error handling at scale ... Research or applied experience with LLM agents, RL (offline/online, RLHF/RLAIF), constrained ...
Senior AI Engineer
New York, NY · On-site
$114K - $157K/yr
Apply reinforcement learning techniques (e.g., RLHF, RLAIF) to improve model alignment and task-specific performance * Architect and manage high-throughput, real-time data pipelines using Kafka
Senior AI Engineer
New York, NY · On-site
$114K - $157K/yr
Apply reinforcement learning techniques (e.g., RLHF, RLAIF) to improve model alignment and task-specific performance * Architect and manage high-throughput, real-time data pipelines using Kafka
Senior AI / Machine Learning Engineer
New York, NY · On-site
$115K - $200K/yr
... management. * Provide technical leadership through design reviews, mentorship, and cross-team ... Experience with reinforcement learning, fine-tuning, or preference-based optimization (e.g., RLHF)
Senior AI / Machine Learning Engineer
New York, NY · On-site
$115K - $200K/yr
... management. * Provide technical leadership through design reviews, mentorship, and cross-team ... Experience with reinforcement learning, fine-tuning, or preference-based optimization (e.g., RLHF)
Senior AI Engineer
$114K - $157K/yr
Apply reinforcement learning techniques (e.g., RLHF, RLAIF) to improve model alignment and task-specific performance * Architect and manage high-throughput, real-time data pipelines using Kafka
Senior AI Engineer
$114K - $157K/yr
Apply reinforcement learning techniques (e.g., RLHF, RLAIF) to improve model alignment and task-specific performance * Architect and manage high-throughput, real-time data pipelines using Kafka
Manager Rlhf information
Scale AI rating
8.5
Based on 9 frontline employees who took The Breakroom Quiz
69th of 217 rated software companies
Job description
Scale AI is a company focused on developing reliable AI systems for critical decisions. They are seeking a Tech Lead Manager for their ML Systems team to build and optimize a training and inference framework that supports machine learning research and development.
Responsibilities:
• Build, profile and optimize our training and inference framework.
• Collaborate with ML and research teams to accelerate their research and development, and enable them to develop the next generation of models and data curation.
• Research and integrate state-of-the-art technologies to optimize our ML system.
Qualifications:
Required:
• Passionate about system optimization
• Experience with multi-node LLM training and inference
• Experience with developing large-scale distributed ML systems
• Experience with post-training methods like RLHF/RLVR and related algorithms like PPO/GRPO etc.
• Strong software engineering skills, proficient in frameworks and tools such as CUDA, Pytorch, transformers, flash attention, etc.
• Strong written and verbal communication skills to operate in a cross functional team environment.
Preferred:
• Demonstrated expertise in post-training methods and/or next generation use cases for large language models including instruction tuning, RLHF, tool use, reasoning, agents, and multimodal, etc.
Company:
Scale’s mission is to develop reliable AI systems for the world’s most important decisions. Founded in 2016, the company is headquartered in San Francisco, USA, with a team of 501-1000 employees. The company is currently Late Stage.
About Scale AI
Sourced by ZipRecruiter
Industry
Software development
Company size
201 - 500 Employees
Headquarters location
San Francisco, CA, US
Year founded
2016