1

Ai Safety Jobs (NOW HIRING)

About the role mpathic is seeking AI Safety Experts for a temporary project to support confidential projects evaluating and improving the safety, reliability, and real-world behavior of frontier AI ...

Expression of Interest

San Francisco, CA · On-site

$19.25 - $25/hr

About The Center for AI Safety (CAIS) The Center for AI Safety (CAIS) is a leading research and advocacy organization focused on mitigating societal-scale risks from AI. We address AI's toughest ...

Founding Recruiter

San Francisco, CA · On-site

$145K - $185K/yr

About The Center for AI Safety (CAIS) The Center for AI Safety (CAIS) is a leading research and advocacy organization focused on mitigating societal-scale risks from AI. Some of our past achievements ...

About the Team The Safety Systems team is dedicated to ensuring the safety, robustness, and reliability of AI models and their deployment in the real world. Learn more about OpenAI's approach to ...

Senior Freelance Consultant, AI Safety

$108K - $123K/yr

Moonshot is seeking a Senior Freelance Consultant to support the delivery of our AI Safety portfolio. This is a hands on, technical engagement focused on red teaming, adversarial evaluation, and ...

Description We are seeking an AI/LLM Safety Engineer to join our AI team and take ownership of how safely our models and agents behave in production; with a focus on AI Safety, Trust & Safety, and ...

We are seeking an AI/LLM Safety Engineer to join our AI team and take ownership of how safely our models and agents behave in production; with a focus on AI Safety, Trust & Safety, and Responsible AI.

Showing results 41-60

ai safety information

See salary details

$10

$32

$58

How much do ai safety jobs pay per hour?

As of Sep 7, 2026, the average hourly pay for ai safety in the United States is $32.38, according to ZipRecruiter salary data. Most workers in this role earn between $25.48 and $39.18 per hour, depending on experience, location, and employer.

What is an AI Safety?

An AI Safety job focuses on ensuring that artificial intelligence systems operate safely, reliably, and ethically. Professionals in this field work to prevent unintended consequences, mitigate risks, and align AI behavior with human values. Roles may involve technical research, policy development, or implementing safety frameworks in AI systems. AI Safety experts collaborate with engineers, ethicists, and policymakers to address concerns such as bias, robustness, and long-term risks of advanced AI.

What are the key skills and qualifications needed to thrive in AI Safety?

To thrive in AI Safety, a strong background in computer science, mathematics, and machine learning is typically required, often supported by an advanced degree in a related field. Familiarity with AI frameworks like TensorFlow or PyTorch, formal verification tools, and knowledge of ethical AI guidelines is highly beneficial. Strong analytical thinking, attention to detail, and collaborative communication skills are essential for success in multi-disciplinary teams. These competencies are crucial for identifying, evaluating, and mitigating risks in AI systems to ensure their safe and responsible deployment.

What are some common challenges faced by professionals working in AI Safety?

Professionals in AI Safety often encounter the challenge of anticipating and mitigating unforeseen consequences of complex machine learning systems. The field requires staying updated with rapid advancements in AI technologies while ensuring that safety protocols and ethical standards keep pace. Collaborating with diverse teams—including engineers, ethicists, and product managers—is typical, and clear communication is vital to align safety goals with broader organizational objectives. Addressing shifting regulatory landscapes and evolving public perceptions around AI also adds dynamic aspects to the role.

How to get into working on AI safety?

To work in AI safety, develop a strong foundation in machine learning, computer science, and ethics through relevant degrees or online courses. Gaining experience with AI research, programming skills in Python, and familiarity with safety frameworks or tools like formal verification can improve your prospects. Engaging with AI safety communities and staying updated on current research also helps build expertise in this specialized field.

What cities are hiring for Ai Safety jobs?

Cities with the most Ai Safety job openings:

What are the most commonly searched types of Ai Safety jobs?

The most popular types of Ai Safety jobs are:

What states have the most Ai Safety jobs?

States with the most job openings for Ai Safety jobs include:

Infographic showing various Ai Safety job openings in the United States as of August 2026, with employment types broken down into 80% Full Time, and 20% Part Time. Highlights an 80% In-person, and 20% Remote job distribution, with an average salary of $67,344 per year, or $32.4 per hour.

Strategic Projects Lead- AI Safety

Raydar

San Francisco, CA • On-site

$260K - $450K/yr

Full-time

Posted 5 days ago


Job description

Join a rapidly scaling AI data and infrastructure company that partners with leading frontier AI labs to build complex datasets, adversarial evaluations, and experimentation systems for advanced models and agents. The company has grown quickly to more than $100M in revenue run rate, raised $34M, and offers early team members the chance to shape both client work and company-wide initiatives.

This is an on-site role in San Francisco for a highly driven operator who can own demanding AI safety projects from client relationship through delivery. The team typically works in the office six days per week and may work 60+ hours during intense periods.

What you'll do

• Partner directly with leading AI labs to design and deliver adversarial datasets, red-teaming programs, and safety benchmarks.

• Own client relationships and project execution end to end, delivering high-quality work against tight timelines.

• Lead teams of subject-matter experts across software engineering, finance, science, security, and other domains.

• Collaborate with internal teams on advanced AI research, evaluation, and data-infrastructure initiatives.

• Drive cross-functional company initiatives as an early team member and support operational priorities where needed.

Requirements

• 2+ years of experience in AI safety, red teaming, trust and safety, adversarial machine learning, security research, or a closely related field.

• Hands-on adversarial experience such as jailbreaking or stress-testing frontier models, prompt-injection research, offensive security, penetration testing, bug bounty programs, CTFs, or trust-and-safety investigations and enforcement.

• Experience at a frontier AI lab, major technology company, leading security firm, or an equivalent high-caliber environment.

• Degree from a top-tier university in the United States, Canada, or Europe, unless professional experience is exceptional.

• Strong communication and client-management skills, with high agency and a track record of owning projects end to end.

• Comfort operating in a high-intensity environment such as management consulting, investment banking, a fast-paced startup, or another demanding client-facing setting.

• Willingness to work on site in San Francisco, typically six days per week, and to work 60+ hours when needed.

Benefits

• Base salary: $130,000-$150,000.

• Public total compensation range: $260,000-$450,000, including bonus and meaningful equity.

• Full-time, on-site in San Francisco, California.

• Visa transfers such as OPT and H-1B transfers may be considered. Exceptional candidates may receive other sponsorship support, but new H-1B petitions are not available.

• In-person final-stage logistics are covered.