As a Research Scientist in the Reinforcement Learning team, you will develop novel approaches to reinforcement learning, contribute to large-scale training infrastructure, and maintain a productive ...
As a Research Scientist in the Reinforcement Learning team, you will develop novel approaches to reinforcement learning, contribute to large-scale training infrastructure, and maintain a productive ...
Research Scientist - Reinforcement Learning
Sunnyvale, CA · On-site
$150K - $450K/yr
Position Summary As a Research Scientist within our Reinforcement Learning team, you will play a fundamental role in establishing our scientific and technical directions toward the development of ...
Quick apply
Research Scientist - Reinforcement Learning
Sunnyvale, CA · On-site
$150K - $450K/yr
Position Summary As a Research Scientist within our Reinforcement Learning team, you will play a fundamental role in establishing our scientific and technical directions toward the development of ...
Senior Reinforcement Learning Engineer
Austin, TX · On-site
$103K - $142K/yr
JOB SUMMARY The Senior Reinforcement Learning Engineer is a key, hands-on role focused on achieving state-of-the-art performance on our humanoid robots. This engineer will leverage their deep ...
Senior Reinforcement Learning Engineer
Austin, TX · On-site
$103K - $142K/yr
JOB SUMMARY The Senior Reinforcement Learning Engineer is a key, hands-on role focused on achieving state-of-the-art performance on our humanoid robots. This engineer will leverage their deep ...
Reinforcement Learning Engineer (Cybersecurity)
$176K - $242K/yr
As a Reinforcement Learning Engineer you will advance the frontier of AI Reinforcement Learning development and delivery. You will build the infrastructure and tooling that transforms real-world ...
Reinforcement Learning Engineer (Cybersecurity)
$176K - $242K/yr
As a Reinforcement Learning Engineer you will advance the frontier of AI Reinforcement Learning development and delivery. You will build the infrastructure and tooling that transforms real-world ...
They are seeking a Helix AI Engineer, Reinforcement Learning to develop learning systems that enable robots to acquire skills through interaction, feedback, and experience, focusing on reinforcement ...
They are seeking a Helix AI Engineer, Reinforcement Learning to develop learning systems that enable robots to acquire skills through interaction, feedback, and experience, focusing on reinforcement ...
Research Scientist - Reinforcement Learning
Sunnyvale, CA · On-site
$150K - $450K/yr
Position Summary As a Research Scientist within our Reinforcement Learning team, you will play a fundamental role in establishing our scientific and technical directions toward the development of ...
Research Scientist - Reinforcement Learning
Sunnyvale, CA · On-site
$150K - $450K/yr
Position Summary As a Research Scientist within our Reinforcement Learning team, you will play a fundamental role in establishing our scientific and technical directions toward the development of ...
Senior Reinforcement Learning Engineer
Sunnyvale, CA · On-site
$230K - $260K/yr
JOB SUMMARY The Senior Reinforcement Learning Engineer is a key, hands-on role focused on achieving state-of-the-art performance on our humanoid robots. This engineer will leverage their deep ...
Senior Reinforcement Learning Engineer
Sunnyvale, CA · On-site
$230K - $260K/yr
JOB SUMMARY The Senior Reinforcement Learning Engineer is a key, hands-on role focused on achieving state-of-the-art performance on our humanoid robots. This engineer will leverage their deep ...
Research Engineer, Machine Learning (Reinforcement Learning)
San Francisco, CA · On-site
$500K - $850K/yr
About the teams Our Reinforcement Learning teams lead Anthropic's reinforcement learning research and development, playing a critical role in advancing our AI systems. We've contributed to all Claude ...
Research Engineer, Machine Learning (Reinforcement Learning)
San Francisco, CA · On-site
$500K - $850K/yr
About the teams Our Reinforcement Learning teams lead Anthropic's reinforcement learning research and development, playing a critical role in advancing our AI systems. We've contributed to all Claude ...
Machine Learning Scientist, Reinforcement Learning
Emeryville, CA · On-site +1
$200K - $330K/yr
We're looking for a motivated and creative Machine Learning (ML) Scientist to drive research into reinforcement learning for biomolecular design. This position offers an opportunity to work at the ...
Machine Learning Scientist, Reinforcement Learning
Emeryville, CA · On-site +1
$200K - $330K/yr
We're looking for a motivated and creative Machine Learning (ML) Scientist to drive research into reinforcement learning for biomolecular design. This position offers an opportunity to work at the ...
They are seeking a Helix AI Engineer, Reinforcement Learning to develop learning systems that enable robots to acquire skills through interaction, feedback, and experience, focusing on reinforcement ...
They are seeking a Helix AI Engineer, Reinforcement Learning to develop learning systems that enable robots to acquire skills through interaction, feedback, and experience, focusing on reinforcement ...
Helix AI Engineer, Reinforcement Learning
San Jose, CA · On-site
$200K - $400K/yr
We are looking for a Helix AI Engineer, Reinforcement Learning to develop learning systems that enable robots to acquire skills through interaction, feedback, and experience. This role focuses on ...
Helix AI Engineer, Reinforcement Learning
San Jose, CA · On-site
$200K - $400K/yr
We are looking for a Helix AI Engineer, Reinforcement Learning to develop learning systems that enable robots to acquire skills through interaction, feedback, and experience. This role focuses on ...
Senior Reinforcement Learning Engineer
Austin, TX · On-site
$103K - $142K/yr
JOB SUMMARY The Senior Reinforcement Learning Engineer is a key, hands-on role focused on achieving state-of-the-art performance on our humanoid robots. This engineer will leverage their deep ...
Senior Reinforcement Learning Engineer
Austin, TX · On-site
$103K - $142K/yr
JOB SUMMARY The Senior Reinforcement Learning Engineer is a key, hands-on role focused on achieving state-of-the-art performance on our humanoid robots. This engineer will leverage their deep ...
Senior Reinforcement Learning Engineer
Sunnyvale, CA · On-site
$170 - $210/hr
JOB SUMMARY The Senior Reinforcement Learning Engineer is a key, hands-on role focused on achieving state-of-the-art performance on our humanoid robots. This engineer will leverage their deep ...
Senior Reinforcement Learning Engineer
Sunnyvale, CA · On-site
$170 - $210/hr
JOB SUMMARY The Senior Reinforcement Learning Engineer is a key, hands-on role focused on achieving state-of-the-art performance on our humanoid robots. This engineer will leverage their deep ...
SWE (RL Environments) "Reinforcement Learning"
San Francisco, CA · On-site
$150 - $250/hr
Design and develop datasets that shape frontier model learning. * Collaborate with research teams at leading AI labs to build and refine reinforcement learning environments. * Implement ...
SWE (RL Environments) "Reinforcement Learning"
San Francisco, CA · On-site
$150 - $250/hr
Design and develop datasets that shape frontier model learning. * Collaborate with research teams at leading AI labs to build and refine reinforcement learning environments. * Implement ...
Applied Reinforcement Learning Engineer Location: Palo Alto, CA or Seattle, WA (Hybrid/Remote) About the Team Centific AI Research advances foundational AI models and applications through ...
Applied Reinforcement Learning Engineer Location: Palo Alto, CA or Seattle, WA (Hybrid/Remote) About the Team Centific AI Research advances foundational AI models and applications through ...
Reinforcement Learning Engineer, Grasping
Houston, TX · On-site
$90 - $130/hr
Role Overview We are looking for a Reinforcement Learning Engineer to join our Manipulation team, focused on dexterous grasping. Our goal is to ship capable, reliable grasping policies on real ...
Reinforcement Learning Engineer, Grasping
Houston, TX · On-site
$90 - $130/hr
Role Overview We are looking for a Reinforcement Learning Engineer to join our Manipulation team, focused on dexterous grasping. Our goal is to ship capable, reliable grasping policies on real ...
About the RL teams Our Reinforcement Learning teams lead Anthropic's reinforcement learning research and development, playing a critical role in advancing our AI systems. We've contributed to all ...
About the RL teams Our Reinforcement Learning teams lead Anthropic's reinforcement learning research and development, playing a critical role in advancing our AI systems. We've contributed to all ...
They are seeking a Helix AI Engineer, Reinforcement Learning to develop learning systems that enable robots to acquire skills through interaction, feedback, and experience. Responsibilities : • ...
They are seeking a Helix AI Engineer, Reinforcement Learning to develop learning systems that enable robots to acquire skills through interaction, feedback, and experience. Responsibilities : • ...
The role involves training reinforcement learning-based LLMs for applications in materials science and engineering, integrating simulation tools, and evaluating models. Responsibilities : • Train ...
The role involves training reinforcement learning-based LLMs for applications in materials science and engineering, integrating simulation tools, and evaluating models. Responsibilities : • Train ...
Helix AI Engineer, Reinforcement Learning
San Jose, CA · On-site
$200K - $400K/yr
We are looking for a Helix AI Engineer, Reinforcement Learning to develop learning systems that enable robots to acquire skills through interaction, feedback, and experience. This role focuses on ...
Helix AI Engineer, Reinforcement Learning
San Jose, CA · On-site
$200K - $400K/yr
We are looking for a Helix AI Engineer, Reinforcement Learning to develop learning systems that enable robots to acquire skills through interaction, feedback, and experience. This role focuses on ...
Reinforcement Learning information
See salary details
$28.5K - $33.2K
2% of jobs
$33.2K - $37.9K
3% of jobs
$37.9K - $42.5K
6% of jobs
$42.5K - $47.2K
3% of jobs
$51.4K is the 25th percentile. Wages below this are outliers.
$47.2K - $51.9K
12% of jobs
The median wage is $56.3K / yr.
$51.9K - $56.6K
25% of jobs
$56.6K - $61.3K
12% of jobs
$61.3K - $66K
9% of jobs
$66.9K is the 75th percentile. Wages above this are outliers.
$66K - $70.6K
12% of jobs
$70.6K - $75.3K
11% of jobs
$75.3K - $80K
5% of jobs
$28.5K
$58.3K
$80K
How much do reinforcement learning jobs pay per year?
What is a reinforcement learning?
A Reinforcement Learning (RL) job involves designing, developing, and optimizing algorithms that enable machines to learn from interactions with their environment. RL professionals work on applications in robotics, finance, gaming, and autonomous systems, leveraging techniques like deep reinforcement learning and policy optimization. Responsibilities often include researching new models, implementing RL algorithms, and improving AI performance. Strong programming skills, knowledge of machine learning frameworks, and an understanding of mathematical concepts like probability and optimization are essential.
What does a reinforcement learning professional do?
A typical day for a Reinforcement Learning professional involves designing and implementing learning algorithms, running experiments, analyzing data, and iterating on models to improve performance. You might collaborate closely with data scientists, software engineers, and product managers to integrate your solutions into broader systems or products. Regular activities also include reading recent research literature and participating in team meetings to discuss progress and obstacles. This dynamic role often balances deep technical work with teamwork to drive innovative applications in areas such as robotics, recommendation systems, or autonomous systems.
What are the key skills and qualifications needed to thrive in the reinforcement learning position?
To thrive in a Reinforcement Learning role, you need a solid background in mathematics, statistics, machine learning, and programming (commonly with Python), typically supported by a relevant degree such as in computer science or engineering. Experience with frameworks like TensorFlow, PyTorch, OpenAI Gym, and familiarity with large-scale computing systems are highly valued. Strong problem-solving abilities, curiosity, and effective collaboration and communication skills help you excel in multidisciplinary research and project teams. These capabilities are crucial for designing, implementing, and refining complex algorithms that learn from interaction to solve real-world problems.
What can you do with reinforcement learning?
What cities are hiring for Reinforcement Learning jobs?
Cities with the most Reinforcement Learning job openings:
What are the most commonly searched types of Reinforcement Learning jobs?
The most popular types of Reinforcement Learning jobs are:
What states have the most Reinforcement Learning jobs?
States with the most job openings for Reinforcement Learning jobs include:
What job categories do people searching Reinforcement Learning jobs look for?
The top searched job categories for Reinforcement Learning jobs are:

Research Scientist - Reinforcement Learning
Sunnyvale, CA • On-site
Full-time
Re-posted 11 days ago
Key responsibilities
Develop novel reinforcement learning approaches for large-scale self-play, agentic tasks, and proactive environment learning.
Contribute to large-scale reinforcement learning training and inference frameworks, including data curation, model architecture, and algorithm design.
Collaborate with internal and external partners and contribute to technical reports and research publications.
Job description
The Institute of Foundation Models is a dedicated research lab focused on advancing research and building capabilities in foundation models. As a Research Scientist in the Reinforcement Learning team, you will develop novel approaches to reinforcement learning, contribute to large-scale training infrastructure, and maintain a productive research portfolio while collaborating with internal and external partners.
Responsibilities:
• Develop novel research toward massive scale self-play for foundation model training, agentic tasks, and imbuing models with the capability to proactively learn from its environment.
• Initiate and pursue novel reinforcement learning algorithmic approaches to define and drive emergent capabilities in Foundation Models.
• Full-stack engineering from data curation, model architecture and algorithm design, to final production of models for end-users using high quality (documented, tested, maintainable) code.
• Contribute to technical reports and research publications.
• Represent MBZUAI at industry conferences and events, showcasing the institution’s technology and deep learning capabilities and establishing MBZUAI as a global leader in AI research and innovation.
• Proactively engage with the open-source community.
• Contribute to large-scale reinforcement learning training and inference frameworks.
• Facilitate internal and external collaboration
Qualifications:
Required:
• MSc/MEng or PhD Degree (or equivalent experience) in Machine Learning, Computer Science or related fields.
• 3+ years of hands-on experience with reinforcement learning.
• Demonstrated ability to independently identify limitations of current practice (internal and external), formulate and enact solution strategies for improvement.
• Proactive mindset with the ability to identify impactful research questions and execute on them with minimal supervision.
• Strong Python development skills with a focus on research-grade code and scalable data pipelines.
• Practical experience implementing complex mathematical concepts into reliable, well-documented code.
• Experience applying novel RL algorithms to practical applications.
• Strong experience contributing to academic and/or open-source research through publication, GitHub contributions, or professional presentations.
• Strong communication and collaboration skills for effective cross-functional work.
Preferred:
• Strong systems and engineering expertise in deep learning frameworks such as PyTorch, Jax, etc.
• Experience in large-scale model training (LLMs or Diffusion Models) on large clusters.
• Familiarity with current RL+LLM training libraries.
• Experience training policies in self-play, possibly demonstrated by publication, blog post, public code.
• Experience working with Diffusion Models in RL, possibly demonstrated by publication, blog post, public code.
• Strong publication record in leading AI and RL venues (e.g.ICLR, ICML, NeurIPS, RLC, JMLR, TMLR).
• Familiarity with performance constraints in production environments and the trade-offs in model design and execution.
• Prior contributions to open-source ML research or data tools.
• Demonstrated ability to solve complex system-level challenges and debug failures across training/inference stack (e.g. memory issues, deadlocks, I/O bottlenecks, multi-node communication failures).
Company:
Official account of Mohamed bin Zayed University of Artificial Intelligence. Dedicated to research, innovation, and empowering brilliant minds in AI. Founded in 2019, the company is headquartered in Abu Dhabi, ARE, with a team of 51-200 employees. The company is currently Growth Stage.