1

Summer Reinforcement Learning Intern Jobs (NOW HIRING)

Intern, AI Engineering

San Francisco, CA

$19.75 - $25.50/hr

We are now filling intern positions for Winter 2026 and Spring 2027. Research Areas * LLM Agent ... Develop novel methods for parameter-efficient adaptation, alignment, and reinforcement learning for ...

Intern, AI Engineering

San Francisco, CA · On-site

$19.75 - $25.50/hr

We are now filling intern positions for Winter 2026 and Spring 2027. Research Areas * LLM Agent ... Develop novel methods for parameter-efficient adaptation, alignment, and reinforcement learning for ...

next page

Showing results 1-20

Summer Reinforcement Learning Intern information

See salary details

$8

$17

$24

How much do summer reinforcement learning intern jobs pay per hour?

As of Jul 23, 2026, the average hourly pay for summer reinforcement learning intern in the United States is $17.04, according to ZipRecruiter salary data. Most workers in this role earn between $14.42 and $19.23 per hour, depending on experience, location, and employer.

What types of projects or tasks can I expect as a Summer Reinforcement Learning Intern?

As a Summer Reinforcement Learning Intern, you can expect to work on projects ranging from implementing and testing RL algorithms to analyzing experiment results and optimizing model performance. Interns often collaborate with experienced researchers and engineers, contributing to both independent and team projects. You may also be involved in literature reviews, setting up simulation environments, and presenting findings to your team. The role provides hands-on experience with real-world RL applications, and you’ll have the opportunity to learn from feedback and mentorship throughout your internship.

What is the difference between Summer Reinforcement Learning Intern vs Summer Data Science Intern?

AspectSummer Reinforcement Learning InternSummer Data Science Intern
Required CredentialsUndergraduate or graduate in CS, AI, or related fields; some knowledge of machine learning and programmingUndergraduate or graduate in Data Science, Statistics, or related fields; strong analytical and programming skills
Work EnvironmentResearch-focused, experimental projects, often in AI and machine learning teamsData analysis, modeling, visualization, and reporting tasks across various departments
Employer & Industry UsageTech companies, AI startups, research labsTech firms, finance, healthcare, and consulting industries

The Summer Reinforcement Learning Intern role focuses on developing and testing reinforcement learning algorithms, often within AI research teams. In contrast, the Summer Data Science Intern role involves broader data analysis and modeling tasks. Both roles require programming skills and are common in tech industries, but they differ in their specific focus and project types.

What are Summer Reinforcement Learning Interns?

Summer Reinforcement Learning Interns are students or recent graduates who work temporarily, usually during the summer, to gain hands-on experience in reinforcement learning, a subfield of machine learning. Their responsibilities often include assisting with the development and testing of algorithms, analyzing data, and collaborating with research teams on projects related to artificial intelligence. This role provides an opportunity to apply theoretical knowledge from coursework to real-world problems, often resulting in valuable skills and networking opportunities for future careers in AI or data science.

What are the key skills and qualifications needed to thrive as a Summer Reinforcement Learning Intern, and why are they important?

To thrive as a Summer Reinforcement Learning Intern, you need a solid background in computer science, mathematics (particularly probability and linear algebra), and experience with machine learning frameworks. Familiarity with Python, TensorFlow or PyTorch, and a strong grasp of reinforcement learning algorithms are typically required, often supported by coursework or relevant certifications. Strong problem-solving skills, curiosity, and effective communication help you stand out in collaborative research and fast-paced project environments. These skills are crucial for contributing to innovative AI projects, rapidly learning new concepts, and effectively sharing findings with mentors and team members.
More about Summer Reinforcement Learning Intern jobs
What cities are hiring for Summer Reinforcement Learning Intern jobs? Cities with the most Summer Reinforcement Learning Intern job openings:
What states have the most Summer Reinforcement Learning Intern jobs? States with the most job openings for Summer Reinforcement Learning Intern jobs include:

Machine Learning Engineer, Next-Generation Recommendation Systems

Unitytech

Bellevue, WA • On-site, Remote

$127K - $191K/yr

Full-time

Medical, Life, Retirement, PTO

Posted 24 days ago


Job description

The opportunity
Unity's Vector AI team builds the machine learning systems that decide which ads reach which players - across billions of monthly users on the world's leading game engine. Recommendation and ranking systems are the core of this work: predicting user value, optimizing bids, and delivering outcomes for advertisers at massive scale.

We are building the next generation of these systems. The frontier has shifted - large language models, reinforcement learning from human feedback, and agentic AI are reshaping what recommendation systems can do. We are looking for PhD graduates who have worked at that frontier and want to bring those ideas into production systems that matter.

What you'll be doing

  • Design, build, and evaluate next-generation ranking and recommendation models that incorporate LLMs, RLHF, and preference learning to improve ad relevance and user experience.
  • Develop user understanding systems - conversion prediction, behavioral modeling, and value estimation - that operate across billions of impressions.
  • Apply reinforcement learning and optimization techniques to bidding strategy, auction dynamics, and real-time ad delivery.
  • Design and run rigorous experiments using causal inference, A/B testing, and offline evaluation frameworks to measure and improve model quality.
  • Partner with engineering to bring research ideas into production, working across the full pipeline from training data to deployed model.
  • Communicate findings clearly to technical and non-technical stakeholders across engineering, product, and business teams.

What we're looking for

  • PhD in Computer Science, Machine Learning, Statistics, or a related field (graduating 2026 or recent graduate).
  • Strong research foundations in one or more of: recommendation systems, reinforcement learning, LLM post-training or alignment, human-AI collaboration, probabilistic modeling, or optimization.
  • Experience working with large-scale data and ML systems, whether through research or industry internships.
  • Fluency in Python; familiarity with ML frameworks such as PyTorch or TensorFlow.
  • A track record of rigorous, high-quality research - publications at top venues (NeurIPS, ICML, ICLR, KDD, RecSys, ACL, WWW, or similar) are a strong signal.
  • Strong written and verbal communication skills - able to make complex ideas accessible across technical and non-technical audiences.

You might also have

  • Industry experience in ads, recommendation, or user understanding systems (internship experience counts).
  • Hands-on experience with production ML pipelines - training at scale, feature engineering, or experimentation infrastructure.
  • Experience applying LLMs or generative models to ranking, retrieval, or structured prediction problems.
  • Familiarity with agentic AI approaches - multi-step reasoning, tool use, or human-AI collaboration frameworks.
  • Exposure to causal inference, uplift modeling, or A/B testing at scale.
  • Genuine curiosity about applied research and the drive to see ideas through to impact.

Additional information

  • Relocation support is not available for this position
$127,400.00 - $191,200.00

This range reflects the anticipated base salary for this position. Beyond base salary, this role may be eligible for equity awards and participation in our company incentive plans (such as annual discretionary bonuses or sales commissions). The final offer amount will depend on several factors, including geographic location and the candidate's relevant experience, professional background, and skill set.

Benefits


At Unity, we want our team members to thrive. We offer a wide range of benefits designed to support well-being and work-life balance.


Please note: Benefits eligibility, specific offerings, and coverage vary based on the country and employment status.


While specific benefits vary, here are some of the ways we strive to take care of our eligible team members globally: Comprehensive health, life, and disability insurance | Commute subsidy | Employee stock ownership | Competitive retirement/pension plans | Generous vacation and personal days | Support for new parents through leave and family-care programs | Office food snacks | Mental Health and Wellbeing programs and support | Employee Resource Groups | Global Employee Assistance Program | Training and development programs | Volunteering and donation matching program

Life at Unity


Unity [NYSE: U] is the world's leading game engine, powering play for more than 3 billion consumers each month. The top mobile games in the world, the most played PC indie titles, the most innovative console games, and virtually all of the top XR and Web Games are developed, deployed, and grown in Unity. Unity also enables teams across industries like automotive, manufacturing, and healthcare to design, simulate, and collaborate in 3D - closing the gap between ideas and reality. For more information, please visit www.unity.com.


Unity is a proud equal opportunity employer. We are committed to fostering an inclusive, innovative environment and celebrate our employees across age, race, color, ancestry, national origin, religion, disability, sex, gender identity or expression, sexual orientation, or any other protected status in accordance with applicable law. Our differences are strengths that enable us to support the growing and evolving needs of our customers, partners, and collaborators. If you have a disability that means there are preparations or accommodations we can make to help ensure you have a comfortable and positive interview experience, please fill out this form to let us know.


This position requires the incumbent to have a sufficient knowledge of English to have professional verbal and written exchanges in this language since the performance of the duties related to this position requires frequent and regular communication with colleagues and partners located worldwide and whose common language is English.
This posting is intended to fill an existing vacancy, and we are committed to providing applicants with updates throughout the hiring process in accordance with applicable law.


Headhunters and recruitment agencies may not submit resumes/CVs through this website or directly to managers. Unity does not accept unsolicited headhunter and agency resumes. Unity will not pay fees to any third-party agency or company that does not have a signed agreement with Unity.
Your privacy is important to us. Please take a moment to review our
Prospect and Applicant Privacy Policies. Should you have any concerns about your privacy, please contact us at DPO@unity.com.