Research Engineer -- AI Alignment & Evaluation AI Safety / Research Engineering | San Francisco, CA | Hybrid / In-Person About the Company We are representing a high-growth AI research organization ...
Quick apply
Research Engineer -- AI Alignment & Evaluation AI Safety / Research Engineering | San Francisco, CA | Hybrid / In-Person About the Company We are representing a high-growth AI research organization ...
Quick apply
Research Engineer -- AI Alignment & Evaluation AI Safety / Research Engineering | San Francisco, CA | Hybrid / In-Person About the Company We are representing a high-growth AI research organization ...
Research Engineer -- AI Alignment & Evaluation AI Safety / Research Engineering | San Francisco, CA | Hybrid / In-Person About the Company We are representing a high-growth AI research organization ...
Research Engineer -- AI Alignment & Evaluation AI Safety / Research Engineering | San Francisco, CA | Hybrid / In-Person About the Company We are representing a high-growth AI research organization ...
Redwood City, CA ยท On-site
$225K - $400K/yr
About the role and team Joining us as a Research Engineer, you'll be at the forefront of tackling one of the most critical challenges in AI today: safety and alignment. Your work will be pivotal in ...
Redwood City, CA ยท On-site
$225K - $400K/yr
About the role and team Joining us as a Research Engineer, you'll be at the forefront of tackling one of the most critical challenges in AI today: safety and alignment. Your work will be pivotal in ...
Menlo Park, CA ยท On-site
$7.6K - $12K/mo
Research Scientist Intern, AI Alignment Responsibilities: * Develop novel state-of-the-art algorithms and corresponding systems, leveraging various deep learning techniques * Analyze and improve ...
Menlo Park, CA ยท On-site
$7.6K - $12K/mo
Research Scientist Intern, AI Alignment Responsibilities: * Develop novel state-of-the-art algorithms and corresponding systems, leveraging various deep learning techniques * Analyze and improve ...
$120K - $190K/yr
As a scientist, you will take ownership of and accelerate existing AI alignment research agendas. You can publish research findings broadly and engage with the AI alignment community. If you are an ...
$120K - $190K/yr
As a scientist, you will take ownership of and accelerate existing AI alignment research agendas. You can publish research findings broadly and engage with the AI alignment community. If you are an ...
Meta is committed to advancing the field of artificial intelligence by making fundamental advances in technologies to help interact with and understand
Meta is committed to advancing the field of artificial intelligence by making fundamental advances in technologies to help interact with and understand
San Francisco, CA ยท On-site
$256K - $276K/yr
Ensure AI safety and alignment principles are integrated throughout the agent lifecycle. * Mentor and grow technical staff, fostering an environment of collaboration and innovation. * Evaluate new ...
San Francisco, CA ยท On-site
$256K - $276K/yr
Ensure AI safety and alignment principles are integrated throughout the agent lifecycle. * Mentor and grow technical staff, fostering an environment of collaboration and innovation. * Evaluate new ...
Ensure AI safety, fairness, and alignment principles are integrated into model training processes and evaluated rigorously. * Mentor and support cross-functional teams on applied machine learning ...
Ensure AI safety, fairness, and alignment principles are integrated into model training processes and evaluated rigorously. * Mentor and support cross-functional teams on applied machine learning ...
Berkeley, CA ยท On-site
Since 2021, we have trained over 630 researchers. 75% of pre-2026 fellows continue to work in AI alignment. 10% have co-founded organizations. Our fellows have produced 215+ research papers with 17 ...
Berkeley, CA ยท On-site
Since 2021, we have trained over 630 researchers. 75% of pre-2026 fellows continue to work in AI alignment. 10% have co-founded organizations. Our fellows have produced 215+ research papers with 17 ...
San Francisco, CA ยท On-site +1
Advance the field of AI alignment by developing cutting-edge methods, such as RLHF and novel approaches, that ensure AI systems reflect human preferences more accurately. * Improve the quality of ...
San Francisco, CA ยท On-site +1
Advance the field of AI alignment by developing cutting-edge methods, such as RLHF and novel approaches, that ensure AI systems reflect human preferences more accurately. * Improve the quality of ...
San Francisco, CA ยท On-site
Advance the field of AI alignment by developing cutting-edge methods, such as RLHF and novel approaches, that ensure AI systems reflect human preferences more accurately. * Improve the quality of ...
San Francisco, CA ยท On-site
Advance the field of AI alignment by developing cutting-edge methods, such as RLHF and novel approaches, that ensure AI systems reflect human preferences more accurately. * Improve the quality of ...
San Francisco, CA ยท On-site
You'll design and implement advanced methods to align human feedback with the training of cutting-edge AI models, including techniques like Reinforcement Learning from Human Feedback (RLHF) , Direct ...
Quick apply
San Francisco, CA ยท On-site
You'll design and implement advanced methods to align human feedback with the training of cutting-edge AI models, including techniques like Reinforcement Learning from Human Feedback (RLHF) , Direct ...
Los Angeles, CA ยท On-site +1
We're running a DARPA seedling on alignment and control, our policy team engages directly with the government on AI alignment, and our work has been featured many times in the Wall Street Journal. We ...
Los Angeles, CA ยท On-site +1
We're running a DARPA seedling on alignment and control, our policy team engages directly with the government on AI alignment, and our work has been featured many times in the Wall Street Journal. We ...
San Francisco, CA ยท On-site
$256K - $276K/yr
Ensure AI safety and alignment principles are integrated throughout the agent lifecycle. * Mentor and grow technical staff, fostering an environment of collaboration and innovation. * Evaluate new ...
San Francisco, CA ยท On-site
$256K - $276K/yr
Ensure AI safety and alignment principles are integrated throughout the agent lifecycle. * Mentor and grow technical staff, fostering an environment of collaboration and innovation. * Evaluate new ...
San Francisco, CA ยท On-site
The Role As a research engineer intern here, you will work very closely with our researchers on projects in areas such as AI security, machine ethics, AI alignment, and benchmarking AI risks. We will ...
San Francisco, CA ยท On-site
The Role As a research engineer intern here, you will work very closely with our researchers on projects in areas such as AI security, machine ethics, AI alignment, and benchmarking AI risks. We will ...
San Francisco, CA ยท On-site +1
$114K - $157K/yr
These range from Generative AI alignment, evaluation, and mitigations - ensuring these models are safe, bias-free, and aligned with Pinterest's vision and policies - to championing ML Fairness and ...
San Francisco, CA ยท On-site +1
$114K - $157K/yr
These range from Generative AI alignment, evaluation, and mitigations - ensuring these models are safe, bias-free, and aligned with Pinterest's vision and policies - to championing ML Fairness and ...
As a Research Engineer, you will tackle critical challenges in AI safety and alignment, conducting foundational research and developing methodologies to ensure AI models behave in accordance with ...
As a Research Engineer, you will tackle critical challenges in AI safety and alignment, conducting foundational research and developing methodologies to ensure AI models behave in accordance with ...
San Francisco, CA ยท On-site
The Role As a research engineer intern here, you will work very closely with our researchers on projects in areas such as AI security, machine ethics, AI alignment, and benchmarking AI risks. We will ...
San Francisco, CA ยท On-site
The Role As a research engineer intern here, you will work very closely with our researchers on projects in areas such as AI security, machine ethics, AI alignment, and benchmarking AI risks. We will ...
San Francisco, CA ยท On-site
$218K - $288K/yr
Ensure AI safety, fairness, and alignment principles are integrated into model training processes and evaluated rigorously. * Mentor and support cross-functional teams on applied machine learning ...
San Francisco, CA ยท On-site
$218K - $288K/yr
Ensure AI safety, fairness, and alignment principles are integrated into model training processes and evaluated rigorously. * Mentor and support cross-functional teams on applied machine learning ...
San Francisco, CA ยท On-site
$256K - $276K/yr
Ensure AI safety and alignment principles are integrated throughout the agent lifecycle. * Mentor and grow technical staff, fostering an environment of collaboration and innovation. * Evaluate new ...
San Francisco, CA ยท On-site
$256K - $276K/yr
Ensure AI safety and alignment principles are integrated throughout the agent lifecycle. * Mentor and grow technical staff, fostering an environment of collaboration and innovation. * Evaluate new ...
| Aspect | Ai Alignment | Data Scientist |
|---|---|---|
| Required Credentials | Advanced degrees in AI, Machine Learning, or related fields | Degree in Data Science, Statistics, Computer Science, or related fields |
| Work Environment | Research labs, AI development companies, tech firms | Tech companies, finance, healthcare, consulting firms |
| Industry Usage | Focuses on ensuring AI systems behave as intended | Analyzes data to extract insights and build predictive models |
While both roles involve advanced technical skills, Ai Alignment specialists focus on aligning AI systems with human values and safety, whereas Data Scientists analyze data to inform business decisions. The roles often overlap in AI research environments but serve different primary objectives.
For Ai Alignment jobs in California, the most frequently searched job titles are:
The top searched job categories for Ai Alignment jobs in California are:
Cities in California with the most Ai Alignment job openings:

San Francisco, CA โข On-site
Full-time
Posted 18 days ago
Design and build evaluation environments for frontier AI models.
Own evaluation projects from concept to refinement, including testing and measurement.
Work with LLM-based agents to perform technical tasks, review outputs, and identify errors.
AI Safety / Research Engineering | San Francisco, CA | Hybrid / In-Person
About the CompanyWe are representing a high-growth AI research organization working at the intersection of frontier model evaluation, AI safety, and security.
The team develops sophisticated evaluation environments designed to surface undesirable or misaligned model behavior and help leading AI organizations better understand how advanced systems behave under complex, long-horizon conditions.
This is a technically rigorous environment for engineers who are interested in AI alignment, agent behavior, model evaluation, and building systems that help make increasingly capable AI more reliable and controllable.
The RoleThis is an opportunity to join a small, highly technical team as a Research Engineer with significant end-to-end ownership.
You will independently design and build evaluation environments that test frontier AI systems for subtle forms of undesirable behavior. You will own the full lifecycle of each environment, from initial concept and failure-mode identification through implementation, grader development, testing, measurement, and refinement.
A significant part of the role involves working directly with advanced LLM agents: prompting them to perform technical tasks, reviewing their output, identifying subtle errors, and making judgment calls where current models still fall short.
The role is ideal for a strong software engineer or technical researcher who enjoys ambiguous problems, learns new domains quickly, and is deeply interested in AI alignment and security.
What You'll DoConfidential details removed: salary, client name, founder names, exact address, company links, investor names, funding details, exact team size, founding year, and highly identifiable wording.