... help businesses automate and optimize their operations. We leverage cutting-edge technologies to ... Reinforcement Learning Engineer Job Title: Reinforcement Learning Engineer Location: 100% Remote ...
... help businesses automate and optimize their operations. We leverage cutting-edge technologies to ... Reinforcement Learning Engineer Job Title: Reinforcement Learning Engineer Location: 100% Remote ...
Reinforcement Learning Expert Dexmate is building the foundation for physical AI -- a unified ... If you want to help shape the next layer of human capability -- and believe the future of robotics ...
Reinforcement Learning Expert Dexmate is building the foundation for physical AI -- a unified ... If you want to help shape the next layer of human capability -- and believe the future of robotics ...
... help businesses automate and optimize their operations. We leverage cutting-edge technologies to ... Reinforcement Learning Engineer Job Title: Reinforcement Learning Engineer Location: 100% Remote ...
... help businesses automate and optimize their operations. We leverage cutting-edge technologies to ... Reinforcement Learning Engineer Job Title: Reinforcement Learning Engineer Location: 100% Remote ...
Reinforcement Learning Engineer
Fremont, CA · Remote
$100K - $150K/yr
... help businesses automate and optimize their operations. We leverage cutting-edge technologies to ... Job Title: Reinforcement Learning Engineer Location: 100% Remote (Continental United States ...
Reinforcement Learning Engineer
Fremont, CA · Remote
$100K - $150K/yr
... help businesses automate and optimize their operations. We leverage cutting-edge technologies to ... Job Title: Reinforcement Learning Engineer Location: 100% Remote (Continental United States ...
... help businesses automate and optimize their operations. We leverage cutting-edge technologies to ... Reinforcement Learning Engineer Job Title: Reinforcement Learning Engineer Location: 100% Remote ...
... help businesses automate and optimize their operations. We leverage cutting-edge technologies to ... Reinforcement Learning Engineer Job Title: Reinforcement Learning Engineer Location: 100% Remote ...
Help scale our systems to handle increasingly complex research workflows. * Design, implement, and test novel training environments, evaluations, and methodologies for reinforcement learning agents ...
Help scale our systems to handle increasingly complex research workflows. * Design, implement, and test novel training environments, evaluations, and methodologies for reinforcement learning agents ...
Research Engineer, Machine Learning (Reinforcement Learning)
Manhattan, NY · On-site +1
$500K - $850K/yr
Help scale our systems to handle increasingly complex research workflows. * Design, implement, and test novel training environments, evaluations, and methodologies for reinforcement learning agents ...
Research Engineer, Machine Learning (Reinforcement Learning)
Manhattan, NY · On-site +1
$500K - $850K/yr
Help scale our systems to handle increasingly complex research workflows. * Design, implement, and test novel training environments, evaluations, and methodologies for reinforcement learning agents ...
Research Engineer, Machine Learning (Reinforcement Learning)
San Francisco, CA · On-site
$500K - $850K/yr
Help scale our systems to handle increasingly complex research workflows. * Design, implement, and test novel training environments, evaluations, and methodologies for reinforcement learning agents ...
Research Engineer, Machine Learning (Reinforcement Learning)
San Francisco, CA · On-site
$500K - $850K/yr
Help scale our systems to handle increasingly complex research workflows. * Design, implement, and test novel training environments, evaluations, and methodologies for reinforcement learning agents ...
Our Team The OpenPipe team at CoreWeave is building tools to help agents learn from experience ... reinforcement learning or PhD + 2 years experience * Strong programming skills in Python and ...
Our Team The OpenPipe team at CoreWeave is building tools to help agents learn from experience ... reinforcement learning or PhD + 2 years experience * Strong programming skills in Python and ...
Senior Machine Learning Engineer, Reinforcement Learning - Egofold
Beverly Hills, CA · On-site +1
$150K - $185K/yr
Senior Machine Learning Engineer, Reinforcement Learning - Egofold About Snail Games USA Snail ... Work closely with engineers and other partners to help integrate successful ML work into usable ...
Senior Machine Learning Engineer, Reinforcement Learning - Egofold
Beverly Hills, CA · On-site +1
$150K - $185K/yr
Senior Machine Learning Engineer, Reinforcement Learning - Egofold About Snail Games USA Snail ... Work closely with engineers and other partners to help integrate successful ML work into usable ...
... help drive superior competitive differentiation, customer experiences, and business outcomes in a ... Senior ML Scientist (Pricing Reinforcement Learning) Work Location: Plano TX Work Mode: Remote Role ...
... help drive superior competitive differentiation, customer experiences, and business outcomes in a ... Senior ML Scientist (Pricing Reinforcement Learning) Work Location: Plano TX Work Mode: Remote Role ...
Research Intern - Applied Reinforcement Learning
$35 - $45/hr
We aim to help these organizations unlock significant business value by deploying GenAI at scale ... About Job PhD Research Intern - Applied Reinforcement Learning Centific AI Research Role Summary ...
Research Intern - Applied Reinforcement Learning
$35 - $45/hr
We aim to help these organizations unlock significant business value by deploying GenAI at scale ... About Job PhD Research Intern - Applied Reinforcement Learning Centific AI Research Role Summary ...
AI Research Scientist, Reinforcement Learning
New York, NY · On-site
$122K/yr
AI Research Scientist, Reinforcement Learning Responsibilities: * Explore and develop novel post ... Meta builds technologies that help people connect, find communities, and grow businesses. When ...
AI Research Scientist, Reinforcement Learning
New York, NY · On-site
$122K/yr
AI Research Scientist, Reinforcement Learning Responsibilities: * Explore and develop novel post ... Meta builds technologies that help people connect, find communities, and grow businesses. When ...
AI Research Scientist, Reinforcement Learning
New York, NY · On-site
$122K - $181K/yr
... reinforcement learning • Explore and develop novel LLM post-training recipes using 3D data • ... People who choose to build their careers by building with us at Meta help shape a future that will ...
AI Research Scientist, Reinforcement Learning
New York, NY · On-site
$122K - $181K/yr
... reinforcement learning • Explore and develop novel LLM post-training recipes using 3D data • ... People who choose to build their careers by building with us at Meta help shape a future that will ...
Applied Deep Learning PhD Research Intern, Reinforcement Learning for LLMs - Fall 2026
$17.50 - $23.50/hr
... reinforcement learning ... Our applied deep learning research team at NVIDIA has helped pioneer projects such as Megatron, MT ...
Applied Deep Learning PhD Research Intern, Reinforcement Learning for LLMs - Fall 2026
$17.50 - $23.50/hr
... reinforcement learning ... Our applied deep learning research team at NVIDIA has helped pioneer projects such as Megatron, MT ...
Applied Deep Learning PhD Research Intern, Reinforcement Learning for LLMs - Fall 2026
Santa Clara, CA · On-site
$17.50 - $23.50/hr
... reinforcement learning ... Our applied deep learning research team at NVIDIA has helped pioneer projects such as Megatron, MT ...
Applied Deep Learning PhD Research Intern, Reinforcement Learning for LLMs - Fall 2026
Santa Clara, CA · On-site
$17.50 - $23.50/hr
... reinforcement learning ... Our applied deep learning research team at NVIDIA has helped pioneer projects such as Megatron, MT ...
Our Team The OpenPipe team at CoreWeave is building tools to help agents learn from experience ... reinforcement learning or PhD + 2 years experience * Strong programming skills in Python and ...
Our Team The OpenPipe team at CoreWeave is building tools to help agents learn from experience ... reinforcement learning or PhD + 2 years experience * Strong programming skills in Python and ...
... help drive superior competitive differentiation, customer experiences, and business outcomes in a ... Senior ML Scientist (Pricing Reinforcement Learning) Work Location: Plano TX Work Mode: Remote Role ...
... help drive superior competitive differentiation, customer experiences, and business outcomes in a ... Senior ML Scientist (Pricing Reinforcement Learning) Work Location: Plano TX Work Mode: Remote Role ...
... helping developers understand our demos and encouraging them to build their own applications ... Ability to design and implement Reinforcement learning and post-training pipelines for LLM to ...
... helping developers understand our demos and encouraging them to build their own applications ... Ability to design and implement Reinforcement learning and post-training pipelines for LLM to ...
Senior Machine Learning Engineer, Reinforcement Learning - Egofold
Beverly Hills, CA · On-site +1
$150K - $185K/yr
Senior Machine Learning Engineer, Reinforcement Learning - Egofold About Snail Games USA Snail ... Work closely with engineers and other partners to help integrate successful ML work into usable ...
Senior Machine Learning Engineer, Reinforcement Learning - Egofold
Beverly Hills, CA · On-site +1
$150K - $185K/yr
Senior Machine Learning Engineer, Reinforcement Learning - Egofold About Snail Games USA Snail ... Work closely with engineers and other partners to help integrate successful ML work into usable ...
Helper Reinforcement Learning information
See salary details
$10.10 - $11.21
2% of jobs
$11.21 - $12.33
4% of jobs
$12.33 - $13.44
6% of jobs
$14.43 is the 25th percentile. Wages below this are outliers.
$13.44 - $14.55
14% of jobs
$14.55 - $15.67
20% of jobs
The median wage is $15.87 / hr.
$15.67 - $16.78
18% of jobs
$17.37 is the 75th percentile. Wages above this are outliers.
$16.78 - $17.90
19% of jobs
$17.90 - $19.01
10% of jobs
$19.01 - $20.13
4% of jobs
$20.13 - $21.24
1% of jobs
$21.24 - $22.36
1% of jobs
$10
$16
$22
How much do helper reinforcement learning jobs pay per hour?
What is the difference between Helper Reinforcement Learning vs Data Scientist?
| Aspect | Helper Reinforcement Learning | Data Scientist |
|---|---|---|
| Required Credentials | Degree in Computer Science, AI, or related fields; knowledge of reinforcement learning | Degree in Data Science, Statistics, Computer Science; proficiency in programming and analytics |
| Work Environment | Research labs, AI development teams, tech companies | Business analytics, research, consulting firms, tech companies |
| Industry Usage | AI development, machine learning projects | Data analysis, predictive modeling, business insights |
| Common Search/Comparison | Helper Reinforcement Learning vs Data Scientist |
Helper Reinforcement Learning focuses on developing algorithms that enable machines to learn through interactions, often requiring knowledge of reinforcement learning techniques. Data Scientists analyze data to extract insights, build models, and support decision-making. While both roles involve programming and data handling, Helper Reinforcement Learning is more specialized in AI algorithm development, whereas Data Scientists work broadly across data analysis and modeling in various industries.
- Senior Data Scientist Citizen
- Associate Machine Learning Chemistry
- Temporary Data Scientist Machine Learning
- Intern Data Scientist Machine Learning
- Nlp Master Practitioner
- Insitro
- Machine Learning Testing
- Amazon Machine Learning Scientist
- Senior Data Scientist Machine Learning
- Full Time No Experience Machine Learning

Full-time
This job post has expired today. Applications are no longer accepted.
Job description
Reinforcement Learning Engineer
Job Title: Reinforcement Learning Engineer Location: 100% Remote (Continental United States) Position Type: In-house Bright Vision Technologies SOW engagement (no third-party client or vendor) Experience: 6+ years Salary: 100k - 150k Sponsorship: No new H1B sponsorship available. H1B transfers welcomed for qualified candidates. Employment Type: Full-time, direct W2 with Bright Vision Technologies (no C2C, no 1099, no third-party) Engagement: Long-term, multi-year, aligned to the Bright Vision SOW delivery roadmap Compensation: Competitive base salary commensurate with experience, plus benefits. Employment Terms & Visa Policy This is a 100% remote, full-time, direct W2 position with Bright Vision Technologies. This role is part of Bright Vision Technologies' in-house Statement of Work (SOW) engagement. The client, end customer, and employer for this position is Bright Vision Technologies — there is no third-party client, vendor, or implementation partner involved. We do not engage in C2C, 1099, or third-party arrangements for this role. BUT STRICTLY NO C2C/1099/3RD PARTY COMPANIES. ALL OUR ROLES ARE W2 AND NO 3RD PARTY BROKERING PLEASE. Candidates must be willing to work directly as a full-time W2 employee of Bright Vision Technologies and contribute to our in-house SOW deliverables. No new H1B sponsorship is available for this role. However, candidates who are currently on a valid H1B visa and require a transfer are welcome to apply. We will support H1B transfers for qualified candidates. For every role, a technical coding assessment is mandatory. Please apply only if you are confident in your technical abilities and hands-on experience. Job Summary We are looking for a Reinforcement Learning Engineer to design, train, and deploy RL-based systems for high-impact decision-making problems where supervised learning alone is insufficient. The role requires deep familiarity with modern reinforcement learning algorithms, simulation environments, reward modeling, and the engineering complexity of training and evaluating policies at scale. The ideal candidate has both research depth and engineering pragmatism, with experience taking RL solutions out of the lab and into production where stability, safety, and ongoing improvement are critical. Key Responsibilities- Design and implement reinforcement learning solutions for sequential decision-making problems in real and simulated environments.
- Develop, calibrate, and maintain simulation environments suitable for large-scale agent training.
- Implement and evaluate modern RL algorithms including policy gradient, actor-critic, off-policy, and offline RL methods.
- Engineer reward functions and shaping strategies that align agent behavior with desired outcomes and safety constraints.
- Apply offline RL and imitation learning techniques where exploration is costly or unsafe.
- Use RLHF, DPO, and related techniques for fine-tuning large language models when relevant.
- Build scalable training infrastructure for distributed RL, including efficient experience collection and replay systems.
- Optimize training stability and sample efficiency through algorithmic and engineering improvements.
- Design rigorous evaluation protocols, including out-of-distribution and adversarial test cases.
- Implement safety mechanisms such as constraint enforcement, conservative policies, and human-in-the-loop oversight.
- Collaborate with applied scientists and product teams to identify high-value RL use cases.
- Monitor deployed policies and models in production for drift, regression, and unintended behaviors, building the alerting and dashboards that surface issues before they meaningfully affect users.
- Document methodology, design decisions, and operational characteristics for internal stakeholders.
- Stay current with RL research and translate promising techniques into production-ready solutions.
- Master's or PhD in Computer Science, Machine Learning, or a related field; or equivalent applied experience.
- Six or more years of combined RL research and engineering experience.
- Strong proficiency in Python and modern deep learning frameworks.
- Hands-on experience with at least one major RL library or in-house RL stack.
- Solid understanding of probability, optimization, and the theoretical foundations of RL.
- Experience designing and tuning reward functions in non-trivial environments.
- Familiarity with simulation environments and large-scale experience collection.
- Experience training neural network policies on GPU clusters.
- Strong written and verbal communication skills.
- Track record of shipping or publishing impactful RL work.
- Experience with RLHF for large language models.
- Familiarity with multi-agent RL or hierarchical RL.
- Exposure to robotics, control systems, or autonomous driving.
- Publications in RL or related research venues.
- Open-source contributions to RL libraries or environments.
Equal Employment Opportunity (EEO) Statement
Bright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall.
BV Teck expressly prohibits any form of workplace harassment or discrimination. Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.