We are building next-generation end-to-end autonomous driving systems powered by reinforcement ... Proficiency in deep learning frameworks such as PyTorch * Experience with distributed training ...
We are building next-generation end-to-end autonomous driving systems powered by reinforcement ... Proficiency in deep learning frameworks such as PyTorch * Experience with distributed training ...
We are building next-generation end-to-end autonomous driving systems powered by reinforcement ... Proficiency in deep learning frameworks such as PyTorch * Experience with distributed training ...
We are building next-generation end-to-end autonomous driving systems powered by reinforcement ... Proficiency in deep learning frameworks such as PyTorch * Experience with distributed training ...
We are building next-generation end-to-end autonomous driving systems powered by reinforcement ... Proficiency in deep learning frameworks such as PyTorch * Experience with distributed training ...
Quick apply
We are building next-generation end-to-end autonomous driving systems powered by reinforcement ... Proficiency in deep learning frameworks such as PyTorch * Experience with distributed training ...
Research Engineer, Performance RL (Reinforcement Learning)
San Francisco, CA · On-site
$350K - $850K/yr
Our Reinforcement Learning teams sit at the intersection of cutting-edge research and engineering excellence, with a deep commitment to building high-quality, scalable systems that push the ...
Research Engineer, Performance RL (Reinforcement Learning)
San Francisco, CA · On-site
$350K - $850K/yr
Our Reinforcement Learning teams sit at the intersection of cutting-edge research and engineering excellence, with a deep commitment to building high-quality, scalable systems that push the ...
Deep understanding and practical experience with various reinforcement learning algorithms and techniques (model-free, model-based, multi-task, hierarchical, multi-agent, etc.). * Strong background ...
Deep understanding and practical experience with various reinforcement learning algorithms and techniques (model-free, model-based, multi-task, hierarchical, multi-agent, etc.). * Strong background ...
Machine Learning Engineer: Imitation and Reinforcement Learning for Robotics
San Francisco, CA · On-site
... with Deep Learning frameworks, such as PyTorch/Tensorflow/JAX to solve real-world problems * 3+ ... Practical experience in behavior cloning and/or reinforcement learning * Bonus: Experience with ...
Machine Learning Engineer: Imitation and Reinforcement Learning for Robotics
San Francisco, CA · On-site
... with Deep Learning frameworks, such as PyTorch/Tensorflow/JAX to solve real-world problems * 3+ ... Practical experience in behavior cloning and/or reinforcement learning * Bonus: Experience with ...
Proficiency with advanced deep learning techniques and architectures as well as reinforcement learning and or imitation learning. * Experience with distributed deep learning systems.
Proficiency with advanced deep learning techniques and architectures as well as reinforcement learning and or imitation learning. * Experience with distributed deep learning systems.
Optimization & Reinforcement Learning: Multi-armed bandits, Deep RL (PPO, DQN) for sequential ... Deep Learning: Modern architectures for demand sensing and price-response curves; Uncertainty ...
Optimization & Reinforcement Learning: Multi-armed bandits, Deep RL (PPO, DQN) for sequential ... Deep Learning: Modern architectures for demand sensing and price-response curves; Uncertainty ...
Machine Learning Engineer - Reinforcement Learning
Fremont, CA · On-site
$150 - $250/hr
Ship deep learning solutions (including LLM / VLM where appropriate) that improve human‑led triaging, automate high‑volume workflows, and support nuanced analysis of self‑driving behavior to ...
Machine Learning Engineer - Reinforcement Learning
Fremont, CA · On-site
$150 - $250/hr
Ship deep learning solutions (including LLM / VLM where appropriate) that improve human‑led triaging, automate high‑volume workflows, and support nuanced analysis of self‑driving behavior to ...
Machine Learning Engineer - Reinforcement Learning
Fremont, CA · On-site
$150 - $250/hr
Ship deep learning solutions (including LLM / VLM where appropriate) that improve human‑led triaging, automate high‑volume workflows, and support nuanced analysis of self‑driving behavior to ...
Machine Learning Engineer - Reinforcement Learning
Fremont, CA · On-site
$150 - $250/hr
Ship deep learning solutions (including LLM / VLM where appropriate) that improve human‑led triaging, automate high‑volume workflows, and support nuanced analysis of self‑driving behavior to ...
Lead AI Infrastructure Engineer, Reinforcement Learning
Santa Clara, CA · On-site
$120 - $170/hr
THE ROLE: We are hiring a Lead AI Infrastructure Engineer, Reinforcement Learning to own ... Deep experience with PyTorch (or JAX), NCCL/MPI‑style distributed training, and GPU cluster ...
Lead AI Infrastructure Engineer, Reinforcement Learning
Santa Clara, CA · On-site
$120 - $170/hr
THE ROLE: We are hiring a Lead AI Infrastructure Engineer, Reinforcement Learning to own ... Deep experience with PyTorch (or JAX), NCCL/MPI‑style distributed training, and GPU cluster ...
Machine Learning Research Scientist, Behavior Planning and Prediction (Mountain View)
Victorville, CA · On-site
$160K - $240K/yr
You have subject matter expertise and research in one or more of the following areas: sequential decision making, imitation learning, deep reinforcement learning, generative modeling, large models ...
Machine Learning Research Scientist, Behavior Planning and Prediction (Mountain View)
Victorville, CA · On-site
$160K - $240K/yr
You have subject matter expertise and research in one or more of the following areas: sequential decision making, imitation learning, deep reinforcement learning, generative modeling, large models ...
Machine Learning Engineer - Reinforcement Learning
$150K - $250K/yr
Ship deep learning solutions (including LLM / VLM where appropriate) that improve human-led triaging, automate high-volume workflows, and support nuanced analysis of self-driving behavior to surface ...
Machine Learning Engineer - Reinforcement Learning
$150K - $250K/yr
Ship deep learning solutions (including LLM / VLM where appropriate) that improve human-led triaging, automate high-volume workflows, and support nuanced analysis of self-driving behavior to surface ...
Machine Learning Engineer - Reinforcement Learning
Fremont, CA · On-site
$150K - $250K/yr
Ship deep learning solutions (including LLM / VLM where appropriate) that improve human-led triaging, automate high-volume workflows, and support nuanced analysis of self-driving behavior to surface ...
Quick apply
Machine Learning Engineer - Reinforcement Learning
Fremont, CA · On-site
$150K - $250K/yr
Ship deep learning solutions (including LLM / VLM where appropriate) that improve human-led triaging, automate high-volume workflows, and support nuanced analysis of self-driving behavior to surface ...
Machine Learning Engineer - Reinforcement Learning
Fremont, CA · On-site
$150K - $250K/yr
Ship deep learning solutions (including LLM / VLM where appropriate) that improve human-led triaging, automate high-volume workflows, and support nuanced analysis of self-driving behavior to surface ...
Machine Learning Engineer - Reinforcement Learning
Fremont, CA · On-site
$150K - $250K/yr
Ship deep learning solutions (including LLM / VLM where appropriate) that improve human-led triaging, automate high-volume workflows, and support nuanced analysis of self-driving behavior to surface ...
Research Engineer, Chip Design RL (Reinforcement Learning)
San Francisco, CA · On-site
$500 - $850/hr
Our Reinforcement Learning teams sit at the intersection of cutting‑edge research and engineering excellence, with a deep commitment to building high‑quality, scalable systems that push the ...
Research Engineer, Chip Design RL (Reinforcement Learning)
San Francisco, CA · On-site
$500 - $850/hr
Our Reinforcement Learning teams sit at the intersection of cutting‑edge research and engineering excellence, with a deep commitment to building high‑quality, scalable systems that push the ...
Machine Learning Research Scientist, Behavior Planning and Prediction
Mountain View, CA · On-site
$160K - $240K/yr
You have subject matter expertise and research in one or more of the following areas: sequential decision making, Imitation Learning, Deep Reinforcement Learning, generative modeling, large models ...
Machine Learning Research Scientist, Behavior Planning and Prediction
Mountain View, CA · On-site
$160K - $240K/yr
You have subject matter expertise and research in one or more of the following areas: sequential decision making, Imitation Learning, Deep Reinforcement Learning, generative modeling, large models ...
Applied Research Scientist, Proactive Intelligence, - Agentic Systems and Generative Modeling
Cupertino, CA · On-site
You will train Generative AI models and Agentic systems using deep reinforcement learning to solve hard problems. Where necessary, you will also work on integrating ML/RL frameworks into our products ...
Applied Research Scientist, Proactive Intelligence, - Agentic Systems and Generative Modeling
Cupertino, CA · On-site
You will train Generative AI models and Agentic systems using deep reinforcement learning to solve hard problems. Where necessary, you will also work on integrating ML/RL frameworks into our products ...
Machine Learning Research Scientist, Behavior Planning and Prediction
Mountain View, CA · On-site
$160K - $240K/yr
You have subject matter expertise and research in one or more of the following areas: sequential decision making, Imitation Learning, Deep Reinforcement Learning, generative modeling, large models ...
Quick apply
Machine Learning Research Scientist, Behavior Planning and Prediction
Mountain View, CA · On-site
$160K - $240K/yr
You have subject matter expertise and research in one or more of the following areas: sequential decision making, Imitation Learning, Deep Reinforcement Learning, generative modeling, large models ...
AI Research Scientist - Infrastructure Engineer, Reinforcement Learning
Santa Clara, CA · On-site
$178K/yr
We are hiring a AI Research Scientist - Infrastructure Engineer, Reinforcement Learning , to own ... Deep experience with PyTorch (or JAX), NCCL/MPI-style distributed training, and GPU cluster ...
AI Research Scientist - Infrastructure Engineer, Reinforcement Learning
Santa Clara, CA · On-site
$178K/yr
We are hiring a AI Research Scientist - Infrastructure Engineer, Reinforcement Learning , to own ... Deep experience with PyTorch (or JAX), NCCL/MPI-style distributed training, and GPU cluster ...
Deep Reinforcement Learning information
What is deep reinforcement learning?
A Deep Reinforcement Learning (DRL) job involves researching, developing, and applying AI models that use reinforcement learning techniques combined with deep learning. Professionals in this role design algorithms that enable agents to learn optimal decision-making policies through trial and error. Common applications include robotics, game AI, autonomous systems, and financial modeling. This job typically requires expertise in machine learning, neural networks, and programming languages like Python, along with frameworks such as TensorFlow or PyTorch.
What does a typical day look like for someone working in deep reinforcement learning?
A typical day for a Deep Reinforcement Learning professional involves designing algorithms, running experiments, analyzing results, and optimizing models to improve performance. You may collaborate regularly with data scientists, software engineers, and domain experts to integrate RL solutions into larger systems or products. Tasks often include reading the latest research, contributing to code reviews, and documenting findings while troubleshooting technical challenges. This dynamic environment encourages continuous learning and teamwork, ensuring you stay at the forefront of AI innovation.
What are the key skills and qualifications needed to thrive in deep reinforcement learning?
To thrive in Deep Reinforcement Learning, you need expertise in machine learning, programming (Python, TensorFlow, or PyTorch), and applied mathematics, often supported by an advanced degree in computer science or a related field. Familiarity with version control systems, cloud computing platforms, and relevant certifications in AI or data science are valuable assets. Strong problem-solving abilities, collaboration, and effective communication are important soft skills in this position. These skills are essential for developing, implementing, and iterating cutting-edge algorithms that solve complex real-world problems in dynamic environments.
What are the most commonly searched types of Deep Reinforcement Learning jobs in California?
The most popular types of Deep Reinforcement Learning jobs in California are:
What are popular job titles related to Deep Reinforcement Learning jobs in California?
For Deep Reinforcement Learning jobs in California, the most frequently searched job titles are:
- Remote Associates Degree In Artificial Intelligence
- Artificial Intelligence In Medicine
- Artificial Intelligence Programmer
- Director Artificial Intelligence Consultant
- Contract Artificial Intelligence Engineer
- Full Time Senior Artificial Intelligence Engineer
- Artificial Intelligence Cognitive Science
- Artificial Intelligence With Secret Clearance
- Associates Degree In Artificial Intelligence
- Artificial Intelligence Developer
What job categories do people searching Deep Reinforcement Learning jobs in California look for?
The top searched job categories for Deep Reinforcement Learning jobs in California are:

Full-time
Re-posted 10 days ago
Job description
You will work on applying RL in closed-loop, safety-critical environments, leveraging large-scale simulation and real-world driving data to improve safety, comfort, and robustness.
- Train and deploy RL policies in closed-loop driving environments
- Scale RL training using massively parallel simulation systems
- Design and optimize reward functions for complex driving behaviors
- Improve sim-to-real transfer for real-world robustness
- Collaborate with cross-functional teams to integrate models into production systems
Requirements
Core Technical Skills
- Proficiency in modern RL algorithms: DQN, PPO, SAC, TD3, etc.
- Proficiency in modern RLHF algorithms: PPO, DPO, GRPO, etc.
- Hands-on experience training reward models and finetuning LLM/VLM/VLA
- Knowledge of distributed RL training at scale
- Proficiency with massively parallel simulation environments
- Knowledge of sim-to-real transfer techniques and domain randomization
- Proficiency in Python, comfortable with C++
- Proficiency in deep learning frameworks such as PyTorch
- Experience with distributed training frameworks (Ray, Horovod, etc.)
- Knowledge of model optimization (quantization, pruning) and CUDA is a plus
- Knowledge of traffic rules, driving behavior modeling
Preferred Qualifications
- Publications in top-tier venues (ICML, NeurIPS, ICLR, CVPR, ICCV, ECCV, ICRA, IROS, etc.)
- Open-source contributions to RL libraries or autonomous driving projects
- Previous experience with LLM fine-tuning using RLHF
- Knowledge of safe RL, interpretable AI, or robustness techniques
- Familiarity with autonomous vehicle regulations and safety standards