1

Reinforcement Learning With Human Feedback Jobs in Arizona

Senior AI Model Fine-Tuning Engineer

Phoenix, AZ · On-site

$128K - $176K/yr

Apply Reinforcement Learning from Human Feedback (RLHF) and other behavioral fine-tuning methods to improve the model's alignment with user needs and ethical standards. * Collaborate with data teams ...

Senior AI Model Fine-Tuning Engineer

Phoenix, AZ · On-site

$128K - $176K/yr

Apply Reinforcement Learning from Human Feedback (RLHF) and other behavioral fine-tuning methods to improve the model's alignment with user needs and ethical standards. * Collaborate with data teams ...

With a world-class team, we're unlocking the next era of autonomous transportation with technology ... reinforcement learning, generative models, foundation models, planning and search, and other ...

With a world-class team, we're unlocking the next era of autonomous transportation with technology ... reinforcement learning, generative models, foundation models, planning and search, and other ...

Provide targeted feedback and reinforcement based on observed learner performance and established ... Collaborate with instructional designers to share learner insights and contribute to continuous ...

Sr. Machine Learning Engineer

Phoenix, AZ

$103K - $142K/yr

How model choices interact with runtime constraints on edge hardware IMPORTANT: Please be informed ... reinforcement learning. * Ability to own and drive a research agenda independently, generating ...

The work we do directly aligns with the company's core areas of focus, and we ensure that ... partners, Learning & Development, Compensation & Benefits, and Shared Services. We enable ...

The work we do directly aligns with the company's core areas of focus, and we ensure that ... partners, Learning & Development, Compensation & Benefits, and Shared Services. We enable ...

next page

Showing results 1-20

Reinforcement Learning With Human Feedback information

What is reinforcement learning with human feedback?

Reinforcement Learning with Human Feedback (RLHF) is a machine learning technique where AI agents are trained not only through automated reward signals but also by incorporating feedback from humans. This approach helps align the agent’s behavior with human preferences, values, or safety requirements by allowing humans to guide or correct the learning process. RLHF is commonly used in developing advanced AI systems, such as language models, to ensure their outputs are helpful, safe, and aligned with user expectations. The process often involves human evaluators ranking or scoring the AI's responses, which are then used to fine-tune the model’s behavior.

What collaborations are typical for a reinforcement learning with human feedback specialist within a machine learning team?

As an RLHF specialist, you often work closely with data scientists, machine learning engineers, and domain experts to design effective feedback mechanisms and reward models. Collaboration with annotation teams or subject matter experts is common, as high-quality human feedback is crucial for training robust RLHF models. You may also partner with product managers and UX researchers to ensure that the models align with user needs and ethical considerations. Regular cross-functional meetings and code reviews help maintain alignment and foster innovation across teams.

What are the key skills and qualifications needed to thrive as a reinforcement learning with human feedback engineer?

To excel as a Reinforcement Learning with Human Feedback (RLHF) Engineer, you need a strong background in machine learning, reinforcement learning theory, statistics, and typically an advanced degree in computer science or a related field. Familiarity with deep learning frameworks (such as TensorFlow or PyTorch), RL libraries (like Ray RLlib), and experience with data collection and annotation systems are essential. Excellent problem-solving abilities, communication skills, and teamwork help you collaborate with researchers, data annotators, and other engineers. These skills enable you to design and implement RLHF systems that are robust, scalable, and aligned with human values.

What is the difference between Reinforcement Learning With Human Feedback vs Reinforcement Learning Engineer?

AspectReinforcement Learning With Human FeedbackReinforcement Learning Engineer
CredentialsTypically requires knowledge of machine learning, AI, and data analysisRequires similar credentials in machine learning, programming, and AI
Work EnvironmentResearch labs, AI development teams, tech companiesDevelopment teams, research labs, tech firms
Industry UsageUsed in AI training, human-in-the-loop systems, and model refinementDesigning, implementing, and optimizing reinforcement learning algorithms

Reinforcement Learning With Human Feedback focuses on improving AI models through human input, while Reinforcement Learning Engineers develop and deploy these algorithms. Both roles require strong machine learning skills and often work in similar environments, but their core responsibilities differ in application and focus.

What are popular job titles related to Reinforcement Learning With Human Feedback jobs in Arizona?

For Reinforcement Learning With Human Feedback jobs in Arizona, the most frequently searched job titles are:

What job categories do people searching Reinforcement Learning With Human Feedback jobs in Arizona look for?

The top searched job categories for Reinforcement Learning With Human Feedback jobs in Arizona are:

What cities in Arizona are hiring for Reinforcement Learning With Human Feedback jobs?

Cities in Arizona with the most Reinforcement Learning With Human Feedback job openings:

Infographic showing various Reinforcement Learning With Human Feedback job openings in Arizona as of August 2026, with employment types broken down into 100% Full Time. Highlights an 100% In-person job distribution.

Senior AI Model Fine-Tuning Engineer

Phoenix, AZ • On-site


TSMC
Computer and Electronic Product Manufacturing • 10K+ employees

8.0

Company rating: 8.0 out of 10

Based on 21 frontline employees who took The Breakroom Quiz

60th of 159 rated electronics manufacturers

Great coworkers

People enjoy working here

Good employer


$128K - $176K/yr

Full-time

Medical, Dental, Vision, Retirement, PTO

Re-posted 20 days ago


Job description

Senior AI Model Fine-Tuning Engineer

A job at TSMC Arizona offers an opportunity to work at the most advanced semiconductor fab in the United States. TSMC Arizona's first fab will operate it's leading-edge semiconductor process technology (N4 process), starting production in the first half of 2025. The second fab will utilize its leading edge N3 and N2 process technology and be operational in 2028. The recently announced third fab will manufacture chips using 2nm or even more advanced process technology, with production starting by the end of the decade. America's leading technology companies are ready to rely on TSMC Arizona for the next generations of chips that will power the digital future.

As a Senior AI Model Fine-Tuning Engineer, you will work on adjusting the behavior and functional abilities of our AI models to make them more adaptable, intelligent, and aligned with real-world use cases.

You will use advanced techniques like prompt engineering, RLHF, and instruction tuning to ensure our models produce high-quality, context-aware responses.

Responsibilities:

  • Lead the fine-tuning process for large pre-trained models, focusing on making models behave appropriately in different contexts (e.g., following instructions, answering questions, or performing tasks).
  • Design and implement prompt engineering strategies to help the model produce more accurate, relevant, and coherent outputs.
  • Apply Reinforcement Learning from Human Feedback (RLHF) and other behavioral fine-tuning methods to improve the model's alignment with user needs and ethical standards.
  • Collaborate with data teams to integrate relevant data and continuously improve model behavior.
  • Conduct model evaluations using various performance metrics (accuracy, bias detection, user feedback) to identify areas for improvement.
  • Iterate and experiment with different fine-tuning methods to achieve optimal performance for specific use cases.
  • Monitor model drift and ensure that models remain consistent, reliable, and safe over time.

Minimum Qualifications/Requirements:

Education: Minimum degree required: Bachelor's degree in Computer Science, Data Science, or a related field.

Technical Skills:

  • 5+ years of experience working on fine-tuning large-scale models such as GPT, T5, or BERT, with a strong focus on behavior and functionality.
  • Expertise in advanced tuning methods such as RLHF, prompt engineering, and zero-shot learning.
  • Experience with popular transformer architectures and frameworks like Hugging Face, TensorFlow, or PyTorch.
  • Deep understanding of LLM behaviors, including instruction-following, task completion, and ethical considerations in output.
  • Proficiency in Python and experience with libraries for model fine-tuning (e.g., Transformers, DeepSpeed).
  • Experience in evaluating model performance, including using metrics like BLEU, ROUGE, perplexity, and custom evaluation frameworks.
  • Bonus: Experience with ethical AI and safety considerations, such as minimizing bias and handling adversarial inputs.
  • Bonus: Experience with model deployment and real-time experimentation (A/B testing).

Interpersonal Skills:

  • Communication
  • Computer proficiency
  • Presentation skills
  • Listening
  • Teamwork

Candidates must be willing and able to work on-site at our Phoenix Arizona facility.

As a valued member of the TSMC family, we place a significant focus on your health and well-being. When you are at your best-physically, mentally, and financially-our company is at its best. We offer a comprehensive and competitive benefits program that provides the resources you need to help you manage your health and achieve your goals across many areas of your life. This includes a variety of medical, dental and vision plan offerings you can choose from that best fit your and your family's needs. Additionally, TSMC provides income-protection programs to financially assist you should you experience an injury or illness, and a 401(k)-retirement savings plan to help you secure your financial future. TSMC also offers competitive paid time-off programs and paid holidays allowing you to recharge and spend time with your family and loved ones.

Work Location: 5088 W. Innovation Circle, Phoenix, AZ 85083

TSMC is an equal opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees. All qualified applicants will receive consideration for employment without regard to race, color, religion, age, sex, sexual orientation, gender identity, national origin, disability, veteran status, or any other protected characteristic. We encourage all qualified individuals to apply, and we welcome applications from individuals with diverse backgrounds and experiences. Candidates must be able to perform the essential functions of the job with or without a reasonable accommodation. If you need a reasonable accommodation as part of this application process, please contact P_LOA@tsmc.com.

#LI-Onsite


TSMC logo

About TSMC

Sourced by ZipRecruiter

Industry

Computer and electronic product manufacturing

Company size

10,000+ Employees

Headquarters location

San Jose, CA, US


What TSMC employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom