They are seeking a Research Engineer specializing in Reinforcement Learning to develop and refine ... Company : Autonomous AI to help build and maintain data pipelines using your infrastructure.
They are seeking a Research Engineer specializing in Reinforcement Learning to develop and refine ... Company : Autonomous AI to help build and maintain data pipelines using your infrastructure.
Help scale our systems to handle increasingly complex research workflows. * Design, implement, and test novel training environments, evaluations, and methodologies for reinforcement learning agents ...
Help scale our systems to handle increasingly complex research workflows. * Design, implement, and test novel training environments, evaluations, and methodologies for reinforcement learning agents ...
Research Engineer, Machine Learning (Reinforcement Learning)
San Francisco, CA · On-site
$500K - $850K/yr
Help scale our systems to handle increasingly complex research workflows. * Design, implement, and test novel training environments, evaluations, and methodologies for reinforcement learning agents ...
Research Engineer, Machine Learning (Reinforcement Learning)
San Francisco, CA · On-site
$500K - $850K/yr
Help scale our systems to handle increasingly complex research workflows. * Design, implement, and test novel training environments, evaluations, and methodologies for reinforcement learning agents ...
Machine Learning Scientist, Reinforcement Learning
Emeryville, CA · On-site +1
$200K - $330K/yr
... to help shape the scientific and strategic vision of the company Qualifications * PhD (or ... reinforcement learning techniques * Publications at major machine learning conferences (NeurIPS ...
Machine Learning Scientist, Reinforcement Learning
Emeryville, CA · On-site +1
$200K - $330K/yr
... to help shape the scientific and strategic vision of the company Qualifications * PhD (or ... reinforcement learning techniques * Publications at major machine learning conferences (NeurIPS ...
Research Engineer, Chip Design RL (Reinforcement Learning)
San Francisco, CA · On-site
$500K - $850K/yr
About the RL teams Our Reinforcement Learning teams lead Anthropic's reinforcement learning ... help with this. We encourage you to apply even if you do not believe you meet every single ...
Research Engineer, Chip Design RL (Reinforcement Learning)
San Francisco, CA · On-site
$500K - $850K/yr
About the RL teams Our Reinforcement Learning teams lead Anthropic's reinforcement learning ... help with this. We encourage you to apply even if you do not believe you meet every single ...
... helping developers understand our demos and encouraging them to build their own applications Some ... Ability to design and implement Reinforcement learning and post-training pipelines for LLM to ...
... helping developers understand our demos and encouraging them to build their own applications Some ... Ability to design and implement Reinforcement learning and post-training pipelines for LLM to ...
Applied Deep Learning PhD Research Intern, Reinforcement Learning for LLMs - Fall 2026
$17.50 - $23.50/hr
... reinforcement learning ... Our applied deep learning research team at NVIDIA has helped pioneer projects such as Megatron, MT ...
Applied Deep Learning PhD Research Intern, Reinforcement Learning for LLMs - Fall 2026
$17.50 - $23.50/hr
... reinforcement learning ... Our applied deep learning research team at NVIDIA has helped pioneer projects such as Megatron, MT ...
Applied Deep Learning PhD Research Intern, Reinforcement Learning for LLMs - Fall 2026
$17.50 - $23.50/hr
... reinforcement learning ... Our applied deep learning research team at NVIDIA has helped pioneer projects such as Megatron, MT ...
Applied Deep Learning PhD Research Intern, Reinforcement Learning for LLMs - Fall 2026
$17.50 - $23.50/hr
... reinforcement learning ... Our applied deep learning research team at NVIDIA has helped pioneer projects such as Megatron, MT ...
Applied Deep Learning PhD Research Intern, Reinforcement Learning for LLMs - Fall 2026
Santa Clara, CA · On-site
$17.50 - $23.50/hr
... reinforcement learning ... Our applied deep learning research team at NVIDIA has helped pioneer projects such as Megatron, MT ...
Applied Deep Learning PhD Research Intern, Reinforcement Learning for LLMs - Fall 2026
Santa Clara, CA · On-site
$17.50 - $23.50/hr
... reinforcement learning ... Our applied deep learning research team at NVIDIA has helped pioneer projects such as Megatron, MT ...
... helping developers understand our demos and encouraging them to build their own applications ... Ability to design and implement Reinforcement learning and post-training pipelines for LLM to ...
... helping developers understand our demos and encouraging them to build their own applications ... Ability to design and implement Reinforcement learning and post-training pipelines for LLM to ...
Research Engineer, Code RL (Reinforcement Learning)
San Francisco, CA · On-site
$500K - $850K/yr
About the RL Teams Our Reinforcement Learning teams play a critical role in advancing our AI ... help with this. We encourage you to apply even if you do not believe you meet every single ...
Research Engineer, Code RL (Reinforcement Learning)
San Francisco, CA · On-site
$500K - $850K/yr
About the RL Teams Our Reinforcement Learning teams play a critical role in advancing our AI ... help with this. We encourage you to apply even if you do not believe you meet every single ...
... helping developers understand our demos and encouraging them to build their own applications ... Ability to design and implement Reinforcement learning and post-training pipelines for LLM to ...
... helping developers understand our demos and encouraging them to build their own applications ... Ability to design and implement Reinforcement learning and post-training pipelines for LLM to ...
Research Engineer, Performance RL (Reinforcement Learning)
San Francisco, CA · On-site
$350K - $850K/yr
About the RL Teams Our Reinforcement Learning teams lead Anthropic's reinforcement learning ... help with this. We encourage you to apply even if you do not believe you meet every single ...
Research Engineer, Performance RL (Reinforcement Learning)
San Francisco, CA · On-site
$350K - $850K/yr
About the RL Teams Our Reinforcement Learning teams lead Anthropic's reinforcement learning ... help with this. We encourage you to apply even if you do not believe you meet every single ...
Machine Learning Engineer: Imitation and Reinforcement Learning for Robotics
San Francisco, CA · On-site
Our team includes veterans who helped launch Waymo, scaled Segment to a $3.2B acquisition, and grew ... Design, train, validate, and launch models for behavior cloning and reinforcement learning * Build ...
Machine Learning Engineer: Imitation and Reinforcement Learning for Robotics
San Francisco, CA · On-site
Our team includes veterans who helped launch Waymo, scaled Segment to a $3.2B acquisition, and grew ... Design, train, validate, and launch models for behavior cloning and reinforcement learning * Build ...
Research Engineer, Chip Design RL (Reinforcement Learning)
San Francisco, CA · On-site
$500 - $850/hr
About the RL teams Our Reinforcement Learning teams lead Anthropic's reinforcement learning ... help with this. We encourage you to apply even if you do not believe you meet every single ...
Research Engineer, Chip Design RL (Reinforcement Learning)
San Francisco, CA · On-site
$500 - $850/hr
About the RL teams Our Reinforcement Learning teams lead Anthropic's reinforcement learning ... help with this. We encourage you to apply even if you do not believe you meet every single ...
Research Engineer (Cybersecurity Reinforcement Learning)
San Francisco, CA · On-site
$300 - $405/hr
Research Engineer, Cybersecurity Reinforcement Learning Anthropic | Anthropic | Posted Mar 9 ... As a Research Engineer, you'll help to safely advance the capabilities of our models in secure ...
Research Engineer (Cybersecurity Reinforcement Learning)
San Francisco, CA · On-site
$300 - $405/hr
Research Engineer, Cybersecurity Reinforcement Learning Anthropic | Anthropic | Posted Mar 9 ... As a Research Engineer, you'll help to safely advance the capabilities of our models in secure ...
Research Scientist - Reinforcement Learning, Robotics
Sunnyvale, CA · On-site
$126 - $423/hr
Improvements deployed to our system immediately help our customers with their programs and deliver ... Conduct research on reinforcement learning (RL) related topics including large‑scale ...
Research Scientist - Reinforcement Learning, Robotics
Sunnyvale, CA · On-site
$126 - $423/hr
Improvements deployed to our system immediately help our customers with their programs and deliver ... Conduct research on reinforcement learning (RL) related topics including large‑scale ...
Lead AI Infrastructure Engineer, Reinforcement Learning
Santa Clara, CA · On-site
$120 - $170/hr
THE ROLE: We are hiring a Lead AI Infrastructure Engineer, Reinforcement Learning to own ... AMD may use Artificial Intelligence to help screen, assess or select applicants for this position.
Lead AI Infrastructure Engineer, Reinforcement Learning
Santa Clara, CA · On-site
$120 - $170/hr
THE ROLE: We are hiring a Lead AI Infrastructure Engineer, Reinforcement Learning to own ... AMD may use Artificial Intelligence to help screen, assess or select applicants for this position.
Senior Staff Research Engineer - Reinforcement Learning for AI Agents
Santa Clara, CA · On-site
$122K - $168K/yr
You will help build the learning systems that power these agents . Key Responsibilities: * Reinforcement learning methods for LLM-driven agents and decision systems. * Policy optimization for long ...
Senior Staff Research Engineer - Reinforcement Learning for AI Agents
Santa Clara, CA · On-site
$122K - $168K/yr
You will help build the learning systems that power these agents . Key Responsibilities: * Reinforcement learning methods for LLM-driven agents and decision systems. * Policy optimization for long ...
Senior Staff Research Engineer - Reinforcement Learning for AI Agents
Santa Clara, CA · On-site
$122K - $168K/yr
You will help build the learning systems that power these agents . Key Responsibilities: * Reinforcement learning methods for LLM-driven agents and decision systems. * Policy optimization for long ...
Senior Staff Research Engineer - Reinforcement Learning for AI Agents
Santa Clara, CA · On-site
$122K - $168K/yr
You will help build the learning systems that power these agents . Key Responsibilities: * Reinforcement learning methods for LLM-driven agents and decision systems. * Policy optimization for long ...
Helper Reinforcement Learning information
What is the difference between Helper Reinforcement Learning vs Data Scientist?
| Aspect | Helper Reinforcement Learning | Data Scientist |
|---|---|---|
| Required Credentials | Degree in Computer Science, AI, or related fields; knowledge of reinforcement learning | Degree in Data Science, Statistics, Computer Science; proficiency in programming and analytics |
| Work Environment | Research labs, AI development teams, tech companies | Business analytics, research, consulting firms, tech companies |
| Industry Usage | AI development, machine learning projects | Data analysis, predictive modeling, business insights |
| Common Search/Comparison | Helper Reinforcement Learning vs Data Scientist |
Helper Reinforcement Learning focuses on developing algorithms that enable machines to learn through interactions, often requiring knowledge of reinforcement learning techniques. Data Scientists analyze data to extract insights, build models, and support decision-making. While both roles involve programming and data handling, Helper Reinforcement Learning is more specialized in AI algorithm development, whereas Data Scientists work broadly across data analysis and modeling in various industries.
What are the most commonly searched types of Reinforcement Learning jobs in California?
The most popular types of Reinforcement Learning jobs in California are:
What are popular job titles related to Helper Reinforcement Learning jobs in California?
For Helper Reinforcement Learning jobs in California, the most frequently searched job titles are:
What job categories do people searching Helper Reinforcement Learning jobs in California look for?
The top searched job categories for Helper Reinforcement Learning jobs in California are:
What cities in California are hiring for Helper Reinforcement Learning jobs?
Cities in California with the most Helper Reinforcement Learning job openings:
Research Engineer, Reinforcement Learning
San Francisco, CA • On-site
Full-time
Re-posted 21 days ago
Job description
TensorStax is building fully autonomous AI systems to manage and maintain mission-critical data infrastructure and pipelines. They are seeking a Research Engineer specializing in Reinforcement Learning to develop and refine reward functions, create RL gym environments, and fine-tune language models using advanced reinforcement learning techniques.
Responsibilities:
• Develop and refine reward functions to optimize agent behavior for complex data engineering tasks.
• Create RL gym environments for language model agents.
• Fine-tune language models using reinforcement learning techniques such as PPO, DPO, and KTO.
• Stay at the forefront of research on RL for language models, incorporating advancements like GRPO, SWE-Gym, and SWE-RL into practical applications.
• Curate and build high-quality datasets for supervised fine-tuning (SFT) and RLHF.
• Design experiments to evaluate and improve the agentic capabilities of language models in data environments.
Qualifications:
Required:
• Deep understanding of reinforcement learning, reward shaping, and optimization strategies.
• Strong familiarity with LLM fine-tuning techniques (PPO, DPO, KTO) and their applications in reinforcement learning.
• Knowledge of recent advancements in RL for language models (GRPO, SWE-Gym, SWE-RL).
• Experience curating and constructing high-quality datasets for fine-tuning.
• Strong problem-solving skills and a history of working on complex ML projects.
• High agency—ability to work independently, experiment proactively, and drive research initiatives forward.
Preferred:
• Experience with distributed training in PyTorch (DDP, FSDP).
• Hands-on experience designing RL environments for traditional RL problems.
• Contributions to open-source projects in RL, LLMs, or ML infrastructure.
• Familiarity with data lakes and warehouses (Snowflake, BigQuery, Redshift).
Company:
Autonomous AI to help build and maintain data pipelines using your infrastructure. Founded in 2024, the company is headquartered in San Francisco, USA, with a team of 2-10 employees. The company is currently Early Stage.