1

Ai Alignment Jobs (NOW HIRING)

AI Safety Red Teamer Expert

New York, NY · On-site +1

$70 - $84/hr

Collaborate with AI researchers to improve model alignment, robustness, and safety. Qualifications Must-Have * Bachelor's degree or higher in Computer Science , Cybersecurity , Journalism ...

Collaborate with AI researchers to improve model alignment, robustness, and safety. Qualifications Must-Have * Bachelor's degree or higher in Computer Science , Cybersecurity , Journalism ...

Advance the field of AI alignment by developing cutting-edge methods, such as RLHF and novel approaches, that ensure AI systems reflect human preferences more accurately. * Improve the quality of ...

Developing a robust and holistic alignment strategy and platform that combines state-of-the-art evaluation and optimization techniques, Responsible AI best practices and the realities of running a ...

We're running a DARPA seedling on alignment and control, our policy team engages directly with the government on AI alignment, and our work has been featured many times in the Wall Street Journal. We ...

Job Summary The Director, AI Operations is responsible for defining and leading the AI alignment, roadmap, and execution portfolio for the Operations & Service Delivery (OSD) organization under ...

Director, AI Operations

PA · On-site

$105K/yr

Job Summary The Director, AI Operations is responsible for defining and leading the AI alignment, roadmap, and execution portfolio for the Operations & Service Delivery (OSD) organization under ...

Bridge product and engineering AI alignment. What We're Looking For (Required Qualifications) * 7+ years of technical program management in a software engineering organization, with at least 2 years ...

AVP, Head of AI Solutions

New York, NY · On-site

$171K - $215K/yr

Global-Local AI Alignment & Acceleration: Adapt and lead the regional AI strategy, ensuring global alignment while addressing zonal specificities. Drive the "Tech Accelerator" aspect by fostering ...

Showing results 21-40

Ai Alignment information

See salary details

$89.5K

$100K

$108.5K

How much do ai alignment jobs pay per year?

As of Aug 25, 2026, the average yearly pay for ai alignment in the United States is $99,999.00, according to ZipRecruiter salary data. Most workers in this role earn between $95,000.00 and $105,000.00 per year, depending on experience, location, and employer.

What is AI alignment?

AI alignment refers to the process of ensuring that artificial intelligence systems act in ways that are aligned with human values, intentions, and ethical standards. This field focuses on designing AI models that not only achieve their objectives but also do so safely and beneficially for humanity. As AI systems become more advanced, alignment becomes increasingly important to prevent unintended consequences or harmful behaviors. Researchers in AI alignment work on technical solutions, such as value learning and interpretability, as well as broader ethical and policy considerations.

What are some common challenges faced by professionals working in AI alignment roles?

Professionals in AI alignment roles often encounter the challenge of translating complex ethical principles and human values into machine-understandable objectives. Balancing technical constraints with theoretical considerations requires close collaboration with cross-functional teams, including ethicists, engineers, and product managers. Additionally, the rapidly evolving landscape of artificial intelligence demands continuous learning to stay current with new alignment techniques and research findings. Navigating these challenges can be intellectually stimulating and offers significant opportunities for interdisciplinary growth.

What are the key skills and qualifications needed to thrive as an AI alignment specialist, and why are they important?

To thrive as an AI Alignment Specialist, you need a strong background in computer science, mathematics, and machine learning, often evidenced by an advanced degree in a related field. Familiarity with technical tools such as Python, TensorFlow, PyTorch, and formal verification systems is typically required, along with understanding of AI safety principles. Analytical thinking, ethical reasoning, and effective communication are crucial soft skills for success in this role. These skills ensure that AI systems are developed safely, ethically, and in alignment with human values, which is essential for mitigating risks associated with advanced AI.

What is the difference between Ai Alignment vs Data Scientist?

AspectAi AlignmentData Scientist
Required CredentialsAdvanced degrees in AI, Machine Learning, or related fieldsDegree in Data Science, Statistics, Computer Science, or related fields
Work EnvironmentResearch labs, AI development companies, tech firmsTech companies, finance, healthcare, consulting firms
Industry UsageFocuses on ensuring AI systems behave as intendedAnalyzes data to extract insights and build predictive models

While both roles involve advanced technical skills, Ai Alignment specialists focus on aligning AI systems with human values and safety, whereas Data Scientists analyze data to inform business decisions. The roles often overlap in AI research environments but serve different primary objectives.

More about Ai Alignment jobs

What cities are hiring for Ai Alignment jobs?

Cities with the most Ai Alignment job openings:

What states have the most Ai Alignment jobs?

States with the most job openings for Ai Alignment jobs include:

Infographic showing various Ai Alignment job openings in the United States as of August 2026, with employment types broken down into 76% Full Time, 21% Part Time, and 3% Contract. Highlights an 64% Physical, 4% Hybrid, and 32% Remote job distribution, with an average salary of $99,999 per year, or $48.1 per hour.

AI Safety Red Teamer Expert

New York, NY • On-site, Remote

$70 - $84/hr

Full-time

Posted 9 days ago


Job description

About the job

Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark, General Catalyst, Peter Thiel, Adam D'Angelo, Larry Summers, and Jack Dorsey.

Position: AI Safety Red Teamer
Type: Contract
Compensation: $70–$84/hour
Location: Remote

Role Responsibilities

  • Design adversarial prompts to stress-test frontier AI models.
  • Identify jailbreaks, unsafe behaviors, hallucinations, and policy failures.
  • Evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content, and other sensitive domains.
  • Document vulnerabilities and contribute to safety benchmarking and red-teaming reports.
  • Collaborate with AI researchers to improve model alignment, robustness, and safety.

Qualifications

Must-Have

  • Bachelor's degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related discipline.
  • 5+ years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a related field.
  • Strong analytical reasoning, prompt design, and written communication skills.
  • Experience designing adversarial prompts or evaluating frontier AI systems.

Preferred

  • Experience with AI Red Teaming, RLHF, SFT, AI Alignment, or Trust & Safety.
  • Familiarity with jailbreak testing, prompt engineering, or adversarial evaluation methodologies.
  • Expertise in one or more grey-area domains, including cyber, biosecurity, political content, misinformation, or scientific safety.

Application Process (Takes 20–30 mins to complete)

  • Upload resume
  • AI interview based on your resume
  • Submit form

Resources & Support

  • For details about the interview process and platform information, please check: https://talent.docs.mercor.com/welcome
  • For any help or support, reach out to: support@mercor.com

PS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this opportunity.