1

Ai Safety Jobs (NOW HIRING)

AI Safety Experts -- English & Norwegian Type: Contract Compensation: $48-$62/hour Location: Remote Role Responsibilities * Red team conversational AI models and agents. Focus on jailbreaks, prompt ...

AI Safety Experts -- English & Swedish Type: Contract Compensation: $48-$62/hour Location: Remote Role Responsibilities * Red team conversational AI models and agents to identify jailbreaks, prompt ...

AI Safety Experts -- English & Portuguese (global) Type: Contract Compensation: $29-$45/hour Location: Remote Role Responsibilities * Red team conversational AI models and agents by performing ...

Position: AI Safety Experts -- English & Indonesian Type: Contract Compensation: $17-$25/hour Location: Remote Role Responsibilities * Red team conversational AI models and agents. Focus on ...

AI Safety Experts -- English & Bengali Type: Contract Compensation: $20-$22/hour Location: Remote Role Responsibilities * Red team conversational AI models and agents to identify jailbreaks, prompt ...

AI Safety Experts -- English & Bengali Type: Contract Compensation: $20-$22/hour Location: Remote Role Responsibilities * Red team conversational AI models and agents to identify jailbreaks, prompt ...

AI Safety Experts -- English & Vietnamese Type: Contract Compensation: $17-$25/hour Location: Remote Role Responsibilities * Red team conversational AI models and agents by conducting jailbreaks ...

Position: AI Safety Experts -- English & Indonesian Type: Contract Compensation: $17-$25/hour Location: Remote Role Responsibilities * Red team conversational AI models and agents. Focus on ...

AI Safety Experts -- English & Norwegian Type: Contract Compensation: $48-$62/hour Location: Remote Role Responsibilities * Red team conversational AI models and agents. Focus on jailbreaks, prompt ...

AI Safety Experts -- English & Danish Type: Contract Compensation: $48-$62/hour Location: Remote Role Responsibilities * Red team conversational AI models and agents to identify jailbreaks, prompt ...

Manager, AI Safety and Security Policy

Washington, DC · On-site

$122K - $131K/yr

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

Manager, AI Safety and Security Policy Full-time, 1-year termed FAS staff position with benefits Washington, DC | Hybrid To Sum It Up... What's the "elevator pitch" for the role? FAS is expanding its ...

AI Safety Experts -- English & Punjabi Type: Contract Compensation: $20-$22/hour Location: Remote Role Responsibilities * Red team conversational AI models and agents. Focus on jailbreaks, prompt ...

AI Safety Experts -- English & Portuguese (global) Type: Contract Compensation: $29-$45/hour Location: Remote Role Responsibilities * Red team conversational AI models and agents by performing ...

AI Safety Experts -- English & Punjabi Type: Contract Compensation: $20-$22/hour Location: Remote Role Responsibilities * Red team conversational AI models and agents. Focus on jailbreaks, prompt ...

Position: AI Safety Experts -- English & Assamese Type: Contract Compensation: $20-$22/hour Location: Remote Role Responsibilities * Red team conversational AI models and agents, focusing on ...

AI Safety Experts -- English & Dutch Type: Contract Compensation: $48-$62/hour Location: Remote Role Responsibilities * Red team conversational AI models and agents to identify jailbreaks, prompt ...

AI Safety Experts -- English & Odia Type: Contract Compensation: $20-$22/hour Location: Remote Role Responsibilities * Red team conversational AI models and agents to identify jailbreaks, prompt ...

AI Safety Experts -- English & Finnish Type: Contract Compensation: $48-$62/hour Location: Remote Role Responsibilities * Red team conversational AI models and agents to identify jailbreaks, prompt ...

AI Safety Experts -- English & Finnish Type: Contract Compensation: $48-$62/hour Location: Remote Role Responsibilities * Red team conversational AI models and agents to identify jailbreaks, prompt ...

Showing results 41-60

ai safety information

See salary details

$10

$32

$58

How much do ai safety jobs pay per hour?

As of Aug 18, 2026, the average hourly pay for ai safety in the United States is $32.38, according to ZipRecruiter salary data. Most workers in this role earn between $25.48 and $39.18 per hour, depending on experience, location, and employer.

What is an AI Safety?

An AI Safety job focuses on ensuring that artificial intelligence systems operate safely, reliably, and ethically. Professionals in this field work to prevent unintended consequences, mitigate risks, and align AI behavior with human values. Roles may involve technical research, policy development, or implementing safety frameworks in AI systems. AI Safety experts collaborate with engineers, ethicists, and policymakers to address concerns such as bias, robustness, and long-term risks of advanced AI.

What are the key skills and qualifications needed to thrive in AI Safety?

To thrive in AI Safety, a strong background in computer science, mathematics, and machine learning is typically required, often supported by an advanced degree in a related field. Familiarity with AI frameworks like TensorFlow or PyTorch, formal verification tools, and knowledge of ethical AI guidelines is highly beneficial. Strong analytical thinking, attention to detail, and collaborative communication skills are essential for success in multi-disciplinary teams. These competencies are crucial for identifying, evaluating, and mitigating risks in AI systems to ensure their safe and responsible deployment.

What are some common challenges faced by professionals working in AI Safety?

Professionals in AI Safety often encounter the challenge of anticipating and mitigating unforeseen consequences of complex machine learning systems. The field requires staying updated with rapid advancements in AI technologies while ensuring that safety protocols and ethical standards keep pace. Collaborating with diverse teams—including engineers, ethicists, and product managers—is typical, and clear communication is vital to align safety goals with broader organizational objectives. Addressing shifting regulatory landscapes and evolving public perceptions around AI also adds dynamic aspects to the role.

How to get into working on AI safety?

To work in AI safety, develop a strong foundation in machine learning, computer science, and ethics through relevant degrees or online courses. Gaining experience with AI research, programming skills in Python, and familiarity with safety frameworks or tools like formal verification can improve your prospects. Engaging with AI safety communities and staying updated on current research also helps build expertise in this specialized field.

What cities are hiring for Ai Safety jobs?

Cities with the most Ai Safety job openings:

What are the most commonly searched types of Ai Safety jobs?

The most popular types of Ai Safety jobs are:

What states have the most Ai Safety jobs?

States with the most job openings for Ai Safety jobs include:

Infographic showing various Ai Safety job openings in the United States as of August 2026, with employment types broken down into 65% Full Time, 9% Part Time, and 26% Contract. Highlights an 83% In-person, and 17% Remote job distribution, with an average salary of $67,344 per year, or $32.4 per hour.

AI Safety Expert - Red Team

Mercor

San Francisco, CA • Remote

$48 - $62/hr

Full-time

Re-posted 8 days ago


Job description

About the job

Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark, General Catalyst, Peter Thiel, Adam D'Angelo, Larry Summers, and Jack Dorsey.

Position: AI Safety Experts — English & Norwegian
Type: Contract
Compensation: $48–$62/hour
Location: Remote

Role Responsibilities

  • Red team conversational AI models and agents. Focus on jailbreaks, prompt injections, misuse cases, and bias exploitation.
  • Generate high-quality human data. Annotate failures, classify vulnerabilities, and flag systemic risks.
  • Apply structure by following taxonomies, benchmarks, and playbooks to maintain consistent testing.
  • Document reproducibly. Produce reports, datasets, and attack cases for customer action.

Qualifications

Must-Have

  • Fluent in English & Norwegian.
  • Prior red teaming experience in AI adversarial work, cybersecurity, or socio-technical probing.
  • Structured approach using frameworks or benchmarks.
  • Strong communication skills for technical and non-technical stakeholders.
  • Adaptability to move across projects and customers.

Preferred

  • Experience in Adversarial ML, Cybersecurity, or socio-technical risk.
  • Skills in creative probing such as psychology, acting, or writing.

Application Process (Takes 20–30 mins to complete)

  • Upload resume
  • AI interview based on your resume
  • Submit form

Resources & Support

  • For details about the interview process and platform information, please check: https://talent.docs.mercor.com/welcome
  • For any help or support, reach out to: support@mercor.com

PS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this opportunity.