AWS Neuron is hiring a Principal Technical Product Manager to define and drive product strategy for ... This includes distributed training libraries, post-training workflows (RLHF, DPO, fine-tuning ...
AWS Neuron is hiring a Principal Technical Product Manager to define and drive product strategy for ... This includes distributed training libraries, post-training workflows (RLHF, DPO, fine-tuning ...
AWS Neuron is hiring a Principal Technical Product Manager to define and drive product strategy for ... This includes distributed training libraries, post-training workflows (RLHF, DPO, fine-tuning ...
AWS Neuron is hiring a Principal Technical Product Manager to define and drive product strategy for ... This includes distributed training libraries, post-training workflows (RLHF, DPO, fine-tuning ...
Senior Applied AI Engineer
San Francisco, CA · On-site
$200K - $300K/yr
... RLHF, and multi-step agentic reasoning into high-impact product workflows. You'll collaborate with engineers, product managers, and technical leads to iterate on intelligent systems that deliver real ...
Senior Applied AI Engineer
San Francisco, CA · On-site
$200K - $300K/yr
... RLHF, and multi-step agentic reasoning into high-impact product workflows. You'll collaborate with engineers, product managers, and technical leads to iterate on intelligent systems that deliver real ...
Senior AI Model Fine-Tuning Engineer
$128K - $176K/yr
Medical
Dental
Vision
Retirement
PTO
You will use advanced techniques like prompt engineering, RLHF, and instruction tuning to ensure ... you manage your health and achieve your goals across many areas of your life. This includes a ...
Senior AI Model Fine-Tuning Engineer
$128K - $176K/yr
Medical
Dental
Vision
Retirement
PTO
You will use advanced techniques like prompt engineering, RLHF, and instruction tuning to ensure ... you manage your health and achieve your goals across many areas of your life. This includes a ...
Technical Solutions Architect, Evals & Fine-Tuning
$140K - $160K/yr
SFT data curation, preference data collection for RLHF/DPO, golden datasets, custom benchmarks, LLM ... managers to keep solutions aligned with the original intent. * Feed customer signal back into ...
Technical Solutions Architect, Evals & Fine-Tuning
$140K - $160K/yr
SFT data curation, preference data collection for RLHF/DPO, golden datasets, custom benchmarks, LLM ... managers to keep solutions aligned with the original intent. * Feed customer signal back into ...
... managing large-scale dataset generation or annotation for LLMs, ideally with experience in RLHF or SFT pipelines. • Strong understanding of quality review mechanisms including prompt win rate ...
... managing large-scale dataset generation or annotation for LLMs, ideally with experience in RLHF or SFT pipelines. • Strong understanding of quality review mechanisms including prompt win rate ...
Senior AI Model Fine-Tuning Engineer
Phoenix, AZ · On-site
$128K - $176K/yr
Medical
Dental
Vision
Retirement
PTO
You will use advanced techniques like prompt engineering, RLHF, and instruction tuning to ensure ... you manage your health and achieve your goals across many areas of your life. This includes a ...
Senior AI Model Fine-Tuning Engineer
Phoenix, AZ · On-site
$128K - $176K/yr
Medical
Dental
Vision
Retirement
PTO
You will use advanced techniques like prompt engineering, RLHF, and instruction tuning to ensure ... you manage your health and achieve your goals across many areas of your life. This includes a ...
QA Specialist (GenAI/LLM)
Louisville, KY · On-site
Required Skills * 2-3+ years of Quality Assurance/Quality Management experience, preferably in AI ... Experience with GenAI, LLM, Data Annotation, RLHF, or AI training datasets * Experience providing ...
QA Specialist (GenAI/LLM)
Louisville, KY · On-site
Required Skills * 2-3+ years of Quality Assurance/Quality Management experience, preferably in AI ... Experience with GenAI, LLM, Data Annotation, RLHF, or AI training datasets * Experience providing ...
Senior AI Model Fine-Tuning Engineer
Phoenix, AZ · On-site
$128K - $176K/yr
Medical
Dental
Vision
Retirement
PTO
You will use advanced techniques like prompt engineering, RLHF, and instruction tuning to ensure ... you manage your health and achieve your goals across many areas of your life. This includes a ...
Senior AI Model Fine-Tuning Engineer
Phoenix, AZ · On-site
$128K - $176K/yr
Medical
Dental
Vision
Retirement
PTO
You will use advanced techniques like prompt engineering, RLHF, and instruction tuning to ensure ... you manage your health and achieve your goals across many areas of your life. This includes a ...
Strategic Projects Lead
$75K - $110K/yr
Experience managing distributed workforces or marketplace operations * Exposure to AI data, RLHF, or model evaluation workflows * Background in investment banking, private equity, or management ...
Strategic Projects Lead
$75K - $110K/yr
Experience managing distributed workforces or marketplace operations * Exposure to AI data, RLHF, or model evaluation workflows * Background in investment banking, private equity, or management ...
Research Lead / Principal Scientist & Manager Post-Training Alignment Reinforcement Learning Au...
California, MD · On-site
Own post-training strategy for model development - from RLHF and preference optimization to agentic ... Manage, mentor, and grow a team of AI scientists * Set technical direction and research priorities ...
Research Lead / Principal Scientist & Manager Post-Training Alignment Reinforcement Learning Au...
California, MD · On-site
Own post-training strategy for model development - from RLHF and preference optimization to agentic ... Manage, mentor, and grow a team of AI scientists * Set technical direction and research priorities ...
LLM Dataset Engineer
San Francisco, CA · On-site
Hands-on experience building datasets for RLHF, DPO, and multi-turn instruction following, including the management of human-labeling workflows and quality gold-sets. • Data Tooling: Mastery of ...
LLM Dataset Engineer
San Francisco, CA · On-site
Hands-on experience building datasets for RLHF, DPO, and multi-turn instruction following, including the management of human-labeling workflows and quality gold-sets. • Data Tooling: Mastery of ...
Research Lead / Principal Scientist & Manager Post-Training - Alignment - Reinforcement Learning Aut
San Francisco, CA · On-site +1
Own post-training strategy for model development - from RLHF and preference optimization to agentic ... Manage, mentor, and grow a team of AI scientists * Set technical direction and research priorities ...
Research Lead / Principal Scientist & Manager Post-Training - Alignment - Reinforcement Learning Aut
San Francisco, CA · On-site +1
Own post-training strategy for model development - from RLHF and preference optimization to agentic ... Manage, mentor, and grow a team of AI scientists * Set technical direction and research priorities ...
Own post-training strategy for model development - from RLHF and preference optimization to agentic ... Manage, mentor, and grow a team of AI scientists * Set technical direction and research priorities ...
Own post-training strategy for model development - from RLHF and preference optimization to agentic ... Manage, mentor, and grow a team of AI scientists * Set technical direction and research priorities ...
Research Lead / Principal Scientist & Manager Post-Training - Alignment - Reinforcement Learning Aut
San Francisco, CA · On-site
Own post-training strategy for model development - from RLHF and preference optimization to agentic ... Manage, mentor, and grow a team of AI scientists * Set technical direction and research priorities ...
Research Lead / Principal Scientist & Manager Post-Training - Alignment - Reinforcement Learning Aut
San Francisco, CA · On-site
Own post-training strategy for model development - from RLHF and preference optimization to agentic ... Manage, mentor, and grow a team of AI scientists * Set technical direction and research priorities ...
Own post-training strategy for model development - from RLHF and preference optimization to agentic ... Manage, mentor, and grow a team of AI scientists * Set technical direction and research priorities ...
Own post-training strategy for model development - from RLHF and preference optimization to agentic ... Manage, mentor, and grow a team of AI scientists * Set technical direction and research priorities ...
Own post-training strategy for model development - from RLHF and preference optimization to agentic ... Manage, mentor, and grow a team of AI scientists * Set technical direction and research priorities ...
Own post-training strategy for model development - from RLHF and preference optimization to agentic ... Manage, mentor, and grow a team of AI scientists * Set technical direction and research priorities ...
Software Engineer - RL Environments
San Francisco, CA · On-site
$180K - $220K/yr
You will build and refine reward signals for RLHF and RLVR pipelines. You will develop quantitative ... Create and manage both real world & synthetic data pipelines * Partner with lab research teams to ...
Software Engineer - RL Environments
San Francisco, CA · On-site
$180K - $220K/yr
You will build and refine reward signals for RLHF and RLVR pipelines. You will develop quantitative ... Create and manage both real world & synthetic data pipelines * Partner with lab research teams to ...
... g., RLHF, prompt evaluation). Scale & Operations: Experience scaling large data operations ... managing complex annotation workflows, and working directly with external data vendors. Technical ...
... g., RLHF, prompt evaluation). Scale & Operations: Experience scaling large data operations ... managing complex annotation workflows, and working directly with external data vendors. Technical ...
Experience supporting AI model training, evaluation, RLHF, or large-scale annotation programs. * Background working with cross-functional teams including researchers, engineers, product managers, and ...
Experience supporting AI model training, evaluation, RLHF, or large-scale annotation programs. * Background working with cross-functional teams including researchers, engineers, product managers, and ...
Manager Rlhf information
See salary details
$24.5K - $32.8K
9% of jobs
$32.8K - $41.1K
15% of jobs
$41.8K is the 25th percentile. Wages below this are outliers.
$41.1K - $49.5K
17% of jobs
The median wage is $52.3K / yr.
$49.5K - $57.8K
27% of jobs
$62.9K is the 75th percentile. Wages above this are outliers.
$57.8K - $66.1K
12% of jobs
$66.1K - $74.4K
8% of jobs
$74.4K - $82.7K
4% of jobs
$82.7K - $91K
3% of jobs
$91K - $99.4K
2% of jobs
$99.4K - $107.7K
2% of jobs
$107.7K - $116K
1% of jobs
$24.5K
$59.5K
$116K
How much do manager rlhf jobs pay per year?
What cities are hiring for Manager Rlhf jobs?
Cities with the most Manager Rlhf job openings:
What are the most commonly searched types of Rlhf jobs?
The most popular types of Rlhf jobs are:
What states have the most Manager Rlhf jobs?
States with the most job openings for Manager Rlhf jobs include:
What job categories do people searching Manager Rlhf jobs look for?
The top searched job categories for Manager Rlhf jobs are:

Full-time
Re-posted 16 hours ago
Amazon rating
7.4
Based on 7,086 frontline employees who took The Breakroom Quiz
6th of 39 rated national retailers
Job description
AWS Trainium is deployed at scale, with millions of chips in production, used for training and inference of frontier models. AWS Neuron is the software stack for Trainium, enabling customers to run deep learning and generative AI workloads with optimal performance and cost efficiency.
AWS Neuron is hiring a Principal Technical Product Manager to define and drive product strategy for training software on Trainium.
This includes distributed training libraries, post-training workflows (RLHF, DPO, fine-tuning), reinforcement learning frameworks, and training performance optimization. Your mission is to enable researchers and operators to train frontier models at scale on Trainium, from single-node experimentation to distributed training across thousands of nodes.
You will be the champion inside AWS for frontier model builders pushing the bounds of scale and resilience for current and emerging training paradigms.
You will work with customers inside and outside the company to identify key improvements and stay ahead of the training landscape. You will define how Neuron supports the training AI/ML ecosystem and what tools customers will use for their training workflows on Trainium.
To be successful, you will partner with engineering teams building training libraries and distributed training infrastructure, applied scientists developing optimization techniques, and PMs responsible for compiler, runtime, NKI, and infrastructure.
You will develop deep knowledge of AI/ML training architectures, distributed training systems, model parallelism strategies, and training performance optimization to effectively define product strategy and make informed technical decisions.
The Ideal Candidate
The ideal candidate will have solid understanding of large-scale model training, distributed training architectures, post-training workflows, and reinforcement learning. They should be able to assess technical implications of training software stack decisions, understand customer needs, and drive developer experience improvements.
The ideal candidate can navigate ambiguity in a fast-moving, early-stage initiative, balance competing priorities across multiple workstreams, and drive alignment across engineering and science stakeholders with excellent written and verbal communication abilities
Key job responsibilities
Training Product Strategy & Roadmap
Define and execute training product strategy and roadmap working backwards from customer requirements in collaboration with engineering leadership. Define the vision for how customers train frontier models at scale on Trainium, balancing performance, developer experience, and AI/ML ecosystem compatibility. Produce PRFAQs and PRDs for training capabilities.
Drive technical alignment across Neuron training libraries, distributed training infrastructure, and dependencies. Partner with PMs responsible for compiler, NKI, runtime, and infrastructure. Drive trade-offs between training performance, scalability, developer experience, and AI/ML ecosystem compatibility.
Define requirements for reusable training building blocks that compose into end-to-end workflows.
Post-Training, RL & Emerging Workflows
Drive strategy for post-training workflows including RLHF, DPO, reward modeling, and fine-tuning at scale. Define requirements for how Neuron supports emerging training paradigms, model architectures, and RL-based optimization loops.
Lead the product experience for RL research-to-production workflows on Trainium. Create and optimize RL libraries and frameworks to help researchers and production model builders.
Customer Engagement & Enablement
Work with BD, Solutions Architecture, and GTM teams to engage customers training frontier models on Trainium.
Understand their distributed training challenges, RL needs, performance optimization requirements, and framework preferences. Translate customer pain points into product requirements. Define success metrics for training adoption and performance.
Support customer enablement for training migration and optimization.
Training AI/ML Ecosystem & Delivery
Define how Neuron supports the training AI/ML ecosystem and what tools customers will use for their training workflows on Trainium. Own the technical depth on training-specific AI/ML ecosystem tools and define how Neuron's training libraries integrate with them.
Track training-specific AI/ML ecosystem trends and feed them into product planning. Drive open source community engagement and upstream contributions for training-related tools. Coordinate with BD on partnership discussions where training-specific technical input is needed.
Launch & Go-to-Market
Lead end-to-end launches for training capabilities, coordinating documentation, field enablement, and customer communications. Partner with Marketing and Solutions Architecture to drive awareness and adoption. Define launch success criteria and track adoption metrics.
About the team
Our team is dedicated to supporting new members. We have a broad mix of experience levels and tenures, and we're building an environment that celebrates knowledge sharing and mentorship. We operate with startup like velocity, prioritizing talent acquisition, hands on leadership, and flexible organization.
Our senior members enjoy one on one mentoring and thorough, but kind, code reviews. We care about your career growth and strive to assign projects that help our team members develop your engineering expertise so you feel empowered to take on more complex tasks in the future.
Diverse Experiences
AWS values diverse experiences.
Even if you do not meet all of the qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn't followed a traditional path, or includes alternative experiences, don't let it stop you from applying.
Inclusive Team Culture
Here at AWS, it's in our nature to learn and be curious.
Our employee led affinity groups foster a culture of inclusion that empower us to be proud of our differences. Ongoing events and learning experiences, including our Conversations on Race and Ethnicity (CORE) and AmazeCon conferences, inspire us to never stop embracing our uniqueness.
Work/Life Balance
We value work life harmony.
Achieving success at work should never come at the expense of sacrifices at home, which is why we strive for flexibility as part of our working culture. When we feel supported in the workplace and at home, there's nothing we can't achieve in the cloud.
Mentorship & Career Growth
We're continuously raising our performance bar as we strive to become Earth's Best Employer.
That's why you'll find endless knowledge sharing, mentorship and other career advancing resources here to help you develop into a better rounded professional.
About Amazon Annapurna Labs
Amazon Annapurna Labs team (our organization within AWS UC) is responsible for building innovation in silicon and software for our AWS customers. We are at the forefront of innovation by combining cloud scale with the world's most talented engineers.
Our team covers multiple disciplines including silicon engineering, hardware design, software and operations. Because of our teams breadth of talent, we have been able to improve AWS cloud infrastructure in high performance machine learning with AWS Neuron, Inferentia and Trainium ML chips, in networking and security with products such as AWS Nitro, Enhanced Network Adapter (ENA), and Elastic Fabric Adapter (EFA), and in computing with AWS Graviton and F1 EC2 instances.
About AWS Utility Computing (UC)
AWS Utility Computing (UC) provides product innovations that continue to set AWS's services and features apart in the industry.
As a member of the UC organization, you'll support the development and management of Compute, Database, Storage, Platform, and Productivity Apps services in AWS, including support for customers who require specialized security solutions for their cloud services. Additionally, this role may involve exposure to and experience with Amazon's growing suite of generative AI services and other cloud computing offerings across the AWS portfolio.
About AWS
Amazon Web Services (AWS) is the world's most comprehensive and broadly adopted cloud platform.
We pioneered cloud computing and never stopped innovating, that's why customers from the most successful startups to Global 500 companies trust our robust suite of products and services to power their businesses.
About Amazon
Sourced by ZipRecruiter
Amazon.com, Inc., commonly known as Amazon, is an American multinational technology company. It was founded by Jeff Bezos in 1994 and initially started as an online marketplace for books. Since then, Amazon has expanded its operations and become one of the largest e-commerce companies in the world. Amazon's primary business is its online retail platform, where customers can purchase a vast array of products, including electronics, clothing, books, home goods, and much more. The company offers a convenient and user-friendly shopping experience, with features such as fast shipping, customer reviews, and personalized recommendations. In addition to its e-commerce platform, Amazon has diversified its business into various other areas. One of its notable ventures is Amazon Web Services (AWS), a comprehensive cloud computing platform that provides services such as storage, compute power, and database management to individuals and businesses. AWS has become a leader in the cloud computing industry, powering many websites and applications worldwide. Amazon has also developed its own consumer electronics, including the popular Amazon Kindle e-reader, Fire tablets, Fire TV streaming devices, and the Alexa-powered Echo smart speakers. The Alexa voice assistant, integrated into these devices, allows users to interact with their devices using voice commands, perform tasks, and access information. Furthermore, Amazon has expanded into media and entertainment. It operates Prime Video, a streaming service that offers a wide range of movies, TV shows, and original content. Amazon Music provides a platform for streaming and purchasing digital music, while Audible offers audiobooks and other audio content. The company's commitment to customer satisfaction and convenience is demonstrated by its membership program, Amazon Prime. Prime members receive various benefits, including free two-day shipping, access to streaming services, exclusive deals, and more.
Industry
It services, book publishers, retail, real estate and computer and electronic product manufacturing
Company size
10,000+ Employees
Headquarters location
Seattle, WA, US