1

Reinforcement Learning Game Jobs (NOW HIRING)

Senior Machine Learning Engineer

Mountain View, CA · On-site +1

$123K - $169K/yr

... game engine constraints, optimizing for latency, compute, and determinism. * Lead high-impact initiatives, including hierarchical reinforcement learning for scalable behaviors, AI planning and goal ...

Lead R&D efforts for game bots using machine learning, reinforcement learning, and neural networks to create realistic agent behaviors. * Design, implement, and maintain gameplay-focused AI ...

Develop and train reinforcement learning models for real-world applications, focusing on efficiency ... Experience in applying PPO to [specific domain, e.g., robotics, gaming, finance, etc.

Showing results 21-40

Reinforcement Learning Game information

See salary details

$7

$20

$37

How much do reinforcement learning game jobs pay per hour?

As of Sep 10, 2026, the average hourly pay for reinforcement learning game in the United States is $20.94, according to ZipRecruiter salary data. Most workers in this role earn between $14.18 and $23.80 per hour, depending on experience, location, and employer.

What is a reinforcement learning game?

A Reinforcement Learning (RL) Game is a simulation or environment designed for testing and training artificial intelligence (AI) agents using reinforcement learning techniques. In these games, an agent interacts with the environment by taking actions and receiving rewards based on its performance, allowing it to learn optimal strategies over time. RL games are widely used in research to benchmark algorithms and in industry to develop intelligent behaviors for robots, automated systems, or video game characters. Popular RL games include OpenAI Gym environments, Atari games, and custom simulations built for specific tasks.

What are the key skills and qualifications needed to thrive as a reinforcement learning engineer in the gaming industry?

To thrive as a Reinforcement Learning Engineer in game development, you need a strong background in machine learning, algorithms, and programming (typically Python), often supported by a degree in computer science or a related field. Familiarity with frameworks like TensorFlow, PyTorch, and RL-specific libraries (such as OpenAI Gym or Unity ML-Agents), as well as experience with simulation environments, is typically required. Critical thinking, creativity, and effective communication help you design innovative AI solutions and collaborate with interdisciplinary teams. These skills are crucial for developing intelligent game agents that enhance player experience and drive technological advancement in interactive entertainment.

How does a reinforcement learning game engineer typically collaborate with game designers and data scientists during development?

As a Reinforcement Learning Game Engineer, you will frequently collaborate with game designers to integrate RL agents in ways that enhance gameplay and balance. Close coordination with data scientists is also common, as they help analyze agent behaviors and performance data to refine training environments and reward structures. Regular cross-functional meetings and iterative testing sessions are standard, ensuring that RL-driven features both align with the game's vision and deliver measurable improvements. This collaborative environment fosters innovation and provides valuable insights into both AI and game design best practices.

What other helpful pages are available for Reinforcement Learning Game?

Other pages related to Reinforcement Learning Game:

Infographic showing various Reinforcement Learning Game job openings in the United States as of September 2026, with employment types broken down into 1% Internship, 1% As Needed, 74% Full Time, 23% Part Time, and 1% Contract. Highlights an 83% Physical, 2% Hybrid, and 15% Remote job distribution, with an average salary of $43,561 per year, or $20.9 per hour.

Machine Learning Engineer, Next-Generation Recommendation Systems

New York, NY • On-site, Remote

$127K - $191K/yr

Full-time

Medical, Life, Retirement, PTO

Re-posted 9 days ago


Key responsibilities

  • Design, build, and evaluate next-generation ranking and recommendation models that incorporate LLMs, RLHF, and preference learning to improve ad relevance and user experience.

  • Develop user understanding systems such as conversion prediction, behavioral modeling, and value estimation that operate across billions of impressions.

  • Apply reinforcement learning and optimization techniques to bidding strategy, auction dynamics, and real-time ad delivery.


Job description

The opportunity
Unity's Vector AI team builds the machine learning systems that decide which ads reach which players - across billions of monthly users on the world's leading game engine. Recommendation and ranking systems are the core of this work: predicting user value, optimizing bids, and delivering outcomes for advertisers at massive scale.

We are building the next generation of these systems. The frontier has shifted - large language models, reinforcement learning from human feedback, and agentic AI are reshaping what recommendation systems can do. We are looking for PhD graduates who have worked at that frontier and want to bring those ideas into production systems that matter.

What you'll be doing

  • Design, build, and evaluate next-generation ranking and recommendation models that incorporate LLMs, RLHF, and preference learning to improve ad relevance and user experience.
  • Develop user understanding systems - conversion prediction, behavioral modeling, and value estimation - that operate across billions of impressions.
  • Apply reinforcement learning and optimization techniques to bidding strategy, auction dynamics, and real-time ad delivery.
  • Design and run rigorous experiments using causal inference, A/B testing, and offline evaluation frameworks to measure and improve model quality.
  • Partner with engineering to bring research ideas into production, working across the full pipeline from training data to deployed model.
  • Communicate findings clearly to technical and non-technical stakeholders across engineering, product, and business teams.

What we're looking for

  • PhD in Computer Science, Machine Learning, Statistics, or a related field (graduating 2026 or recent graduate).
  • Strong research foundations in one or more of: recommendation systems, reinforcement learning, LLM post-training or alignment, human-AI collaboration, probabilistic modeling, or optimization.
  • Experience working with large-scale data and ML systems, whether through research or industry internships.
  • Fluency in Python; familiarity with ML frameworks such as PyTorch or TensorFlow.
  • A track record of rigorous, high-quality research - publications at top venues (NeurIPS, ICML, ICLR, KDD, RecSys, ACL, WWW, or similar) are a strong signal.
  • Strong written and verbal communication skills - able to make complex ideas accessible across technical and non-technical audiences.

You might also have

  • Industry experience in ads, recommendation, or user understanding systems (internship experience counts).
  • Hands-on experience with production ML pipelines - training at scale, feature engineering, or experimentation infrastructure.
  • Experience applying LLMs or generative models to ranking, retrieval, or structured prediction problems.
  • Familiarity with agentic AI approaches - multi-step reasoning, tool use, or human-AI collaboration frameworks.
  • Exposure to causal inference, uplift modeling, or A/B testing at scale.
  • Genuine curiosity about applied research and the drive to see ideas through to impact.

Additional information

  • Relocation support is not available for this position
$127,400.00 - $191,200.00

This range reflects the anticipated base salary for this position. Beyond base salary, this role may be eligible for equity awards and participation in our company incentive plans (such as annual discretionary bonuses or sales commissions). The final offer amount will depend on several factors, including geographic location and the candidate's relevant experience, professional background, and skill set.

Benefits


At Unity, we want our team members to thrive. We offer a wide range of benefits designed to support well-being and work-life balance.


Please note: Benefits eligibility, specific offerings, and coverage vary based on the country and employment status.


While specific benefits vary, here are some of the ways we strive to take care of our eligible team members globally: Comprehensive health, life, and disability insurance | Commute subsidy | Employee stock ownership | Competitive retirement/pension plans | Generous vacation and personal days | Support for new parents through leave and family-care programs | Office food snacks | Mental Health and Wellbeing programs and support | Employee Resource Groups | Global Employee Assistance Program | Training and development programs | Volunteering and donation matching program

Life at Unity


Unity [NYSE: U] is the world's leading game engine, powering play for more than 3 billion consumers each month. The top mobile games in the world, the most played PC indie titles, the most innovative console games, and virtually all of the top XR and Web Games are developed, deployed, and grown in Unity. Unity also enables teams across industries like automotive, manufacturing, and healthcare to design, simulate, and collaborate in 3D - closing the gap between ideas and reality. For more information, please visit www.unity.com.


Unity is a proud equal opportunity employer. We are committed to fostering an inclusive, innovative environment and celebrate our employees across age, race, color, ancestry, national origin, religion, disability, sex, gender identity or expression, sexual orientation, or any other protected status in accordance with applicable law. Our differences are strengths that enable us to support the growing and evolving needs of our customers, partners, and collaborators. If you have a disability that means there are preparations or accommodations we can make to help ensure you have a comfortable and positive interview experience, please fill out this form to let us know.


This position requires the incumbent to have a sufficient knowledge of English to have professional verbal and written exchanges in this language since the performance of the duties related to this position requires frequent and regular communication with colleagues and partners located worldwide and whose common language is English.
This posting is intended to fill an existing vacancy, and we are committed to providing applicants with updates throughout the hiring process in accordance with applicable law.


Headhunters and recruitment agencies may not submit resumes/CVs through this website or directly to managers. Unity does not accept unsolicited headhunter and agency resumes. Unity will not pay fees to any third-party agency or company that does not have a signed agreement with Unity.
Your privacy is important to us. Please take a moment to review our
Prospect and Applicant Privacy Policies. Should you have any concerns about your privacy, please contact us at DPO@unity.com.