1

Ai Safety Jobs (NOW HIRING)

AI Safety & Red Teaming Specialist Role Type: Contractor Location: Remote We are looking for an experienced AI Safety & Red Teaming Specialist to help evaluate and strengthen the safety, security ...

Moonshot is recruiting a Head of AI Safety to lead the delivery, development, and growth of our AI Safety portfolio. The role combines Moonshot's expertise in violence prevention, behavioral risk ...

Head of AI Safety

Washington, DC · On-site

$110 - $145/hr

Moonshot is recruiting a Head of AI Safety to lead the delivery, development, and growth of our AI Safety portfolio. The role combines Moonshot's expertise in violence prevention, behavioral risk ...

New

Moonshot is recruiting a Head of AI Safety to lead the delivery, development, and growth of our AI Safety portfolio. The role combines Moonshot's expertise in violence prevention, behavioral risk ...

Moonshot is recruiting a Head of AI Safety to lead the delivery, development, and growth of our AI Safety portfolio. The role combines Moonshot's expertise in violence prevention, behavioral risk ...

Moonshot is recruiting a Head of AI Safety to lead the delivery, development, and growth of our AI Safety portfolio. The role combines Moonshot's expertise in violence prevention, behavioral risk ...

Head of AI Safety

Atlanta, GA · On-site

$110 - $145/hr

Moonshot is recruiting a Head of AI Safety to lead the delivery, development, and growth of our AI Safety portfolio. The role combines Moonshot's expertise in violence prevention, behavioral risk ...

New

Moonshot is recruiting a Head of AI Safety to lead the delivery, development, and growth of our AI Safety portfolio. The role combines Moonshot's expertise in violence prevention, behavioral risk ...

Head of AI Safety

Denver, CO · On-site

$110 - $145/hr

Moonshot is recruiting a Head of AI Safety to lead the delivery, development, and growth of our AI Safety portfolio. The role combines Moonshot's expertise in violence prevention, behavioral risk ...

New

Moonshot is recruiting a Head of AI Safety to lead the delivery, development, and growth of our AI Safety portfolio. The role combines Moonshot's expertise in violence prevention, behavioral risk ...

Moonshot is recruiting a Head of AI Safety to lead the delivery, development, and growth of our AI Safety portfolio. The role combines Moonshot's expertise in violence prevention, behavioral risk ...

This role is for one of our clients Compensation: $65-$70 per hour Join an advanced AI safety initiative focused on improving how next-generation AI systems understand and reason about complex ...

AI Safety Red Teamer Type: Contract Compensation: $70-$84/hour Location: Remote Role Responsibilities * Design adversarial prompts to stress-test frontier AI models . * Identify jailbreaks, unsafe ...

New

next page

Showing results 1-20

Ai Safety information

See salary details

$10

$32

$58

How much do ai safety jobs pay per hour?

As of Aug 15, 2026, the average hourly pay for ai safety in the United States is $32.38, according to ZipRecruiter salary data. Most workers in this role earn between $25.48 and $39.18 per hour, depending on experience, location, and employer.

What is an AI Safety?

An AI Safety job focuses on ensuring that artificial intelligence systems operate safely, reliably, and ethically. Professionals in this field work to prevent unintended consequences, mitigate risks, and align AI behavior with human values. Roles may involve technical research, policy development, or implementing safety frameworks in AI systems. AI Safety experts collaborate with engineers, ethicists, and policymakers to address concerns such as bias, robustness, and long-term risks of advanced AI.

What are some common challenges faced by professionals working in AI Safety?

Professionals in AI Safety often encounter the challenge of anticipating and mitigating unforeseen consequences of complex machine learning systems. The field requires staying updated with rapid advancements in AI technologies while ensuring that safety protocols and ethical standards keep pace. Collaborating with diverse teams—including engineers, ethicists, and product managers—is typical, and clear communication is vital to align safety goals with broader organizational objectives. Addressing shifting regulatory landscapes and evolving public perceptions around AI also adds dynamic aspects to the role.

What are the key skills and qualifications needed to thrive in AI Safety?

To thrive in AI Safety, a strong background in computer science, mathematics, and machine learning is typically required, often supported by an advanced degree in a related field. Familiarity with AI frameworks like TensorFlow or PyTorch, formal verification tools, and knowledge of ethical AI guidelines is highly beneficial. Strong analytical thinking, attention to detail, and collaborative communication skills are essential for success in multi-disciplinary teams. These competencies are crucial for identifying, evaluating, and mitigating risks in AI systems to ensure their safe and responsible deployment.

What cities are hiring for Ai Safety jobs?

Cities with the most Ai Safety job openings:

What are the most commonly searched types of Ai Safety jobs?

The most popular types of Ai Safety jobs are:

What states have the most Ai Safety jobs?

States with the most job openings for Ai Safety jobs include:

Infographic showing various Ai Safety job openings in the United States as of August 2026, with employment types broken down into 1% As Needed, 79% Full Time, 17% Part Time, 2% Contract, and 1% Nights. Highlights an 99% Physical, and 1% Remote job distribution, with an average salary of $67,344 per year, or $32.4 per hour.

AI Safety & Red Teaming Specialist

Weekday AI

Remote

Contractor

Re-posted 14 days ago


Job description

This role is for one of our clients
$50 - $90/hourpay
Role Title: AI Safety & Red Teaming Specialist
Role Type: Contractor
Location: Remote
We are looking for an experienced AI Safety & Red Teaming Specialist to help evaluate and strengthen the safety, security, and robustness of next-generation AI systems. In this role, you will leverage your expertise in adversarial AI, LLM security, and AI safety to identify vulnerabilities, design evaluation methodologies, and improve the resilience of large language models against real-world threats.
This opportunity is ideal for professionals passionate about AI security, ethical hacking, and developing robust evaluation frameworks for advanced AI systems.
Requirements
Key Responsibilities
  • Design and implement advanced evaluation methodologies for AI system safety, including ethical jailbreak testing, prompt injection detection, LLM red teaming, and tool-use abuse scenarios.
  • Develop cross-domain adversarial testing strategies to uncover complex, multi-turn attack patterns and model vulnerabilities.
  • Build, maintain, and enhance regression test suites to continuously assess jailbreak susceptibility and prompt injection risks.
  • Create comprehensive evaluation frameworks that simulate real-world adversarial threats to improve AI robustness and reliability.
  • Collaborate with technical teams to translate security findings into actionable recommendations for AI safety improvements.
  • Document testing methodologies, findings, and best practices through clear technical reports and presentations for both technical and non-technical stakeholders.
Required Qualifications
  • 2+ years of experience in AI Safety, Adversarial Machine Learning, LLM Red Teaming, AI Security, or a related field.
  • Hands-on experience researching, testing, or identifying vulnerabilities involving prompt injection, ethical jailbreaks, adversarial attacks, or tool-use exploitation.
  • Strong understanding of modern LLM architectures, prompt engineering, and AI safety evaluation methodologies.
  • Experience developing structured security assessments, regression testing frameworks, and adversarial evaluation strategies.
  • Excellent analytical, documentation, and communication skills with the ability to explain complex technical findings clearly.
  • Ability to collaborate effectively within cross-functional technical teams.
Preferred Qualifications
  • Master's or PhD in Computer Science, Cybersecurity, Machine Learning, Artificial Intelligence, or a related discipline.
  • Contributions to AI security research, open-source AI safety tools, conference presentations, or published research.
  • Experience with AI model evaluation frameworks, prompt engineering techniques, and AI security assessment tools.
  • Background in multidisciplinary AI safety, cybersecurity, or adversarial machine learning projects.
Must-Have Skills
  • AI Safety
  • LLM Red Teaming
  • Prompt Injection
  • Adversarial Machine Learning
Good-to-Have Skills
  • Ethical Jailbreaking
  • AI Security
  • Prompt Engineering
  • AI Evaluation Frameworks