1

Ai Alignment Jobs in California (NOW HIRING)

$120K - $190K/yr

As a scientist, you will take ownership of and accelerate existing AI alignment research agendas. You can publish research findings broadly and engage with the AI alignment community. If you are an ...

Advance the field of AI alignment by developing cutting-edge methods, such as RLHF and novel approaches, that ensure AI systems reflect human preferences more accurately. * Improve the quality of ...

You'll design and implement advanced methods to align human feedback with the training of cutting-edge AI models, including techniques like Reinforcement Learning from Human Feedback (RLHF) , Direct ...

We're running a DARPA seedling on alignment and control, our policy team engages directly with the government on AI alignment, and our work has been featured many times in the Wall Street Journal. We ...

General Counsel

Los Angeles, CA ยท On-site

$225 - $300/hr

We're running a DARPA seedling on alignment and control, our policy team engages directly with the government on AI alignment, and our work has been featured many times in the Wall Street Journal. We ...

next page

Showing results 1-20

Ai Alignment information

What is AI alignment?

AI alignment refers to the process of ensuring that artificial intelligence systems act in ways that are aligned with human values, intentions, and ethical standards. This field focuses on designing AI models that not only achieve their objectives but also do so safely and beneficially for humanity. As AI systems become more advanced, alignment becomes increasingly important to prevent unintended consequences or harmful behaviors. Researchers in AI alignment work on technical solutions, such as value learning and interpretability, as well as broader ethical and policy considerations.

What is the difference between Ai Alignment vs Data Scientist?

AspectAi AlignmentData Scientist
Required CredentialsAdvanced degrees in AI, Machine Learning, or related fieldsDegree in Data Science, Statistics, Computer Science, or related fields
Work EnvironmentResearch labs, AI development companies, tech firmsTech companies, finance, healthcare, consulting firms
Industry UsageFocuses on ensuring AI systems behave as intendedAnalyzes data to extract insights and build predictive models

While both roles involve advanced technical skills, Ai Alignment specialists focus on aligning AI systems with human values and safety, whereas Data Scientists analyze data to inform business decisions. The roles often overlap in AI research environments but serve different primary objectives.

What are some common challenges faced by professionals working in AI alignment roles?

Professionals in AI alignment roles often encounter the challenge of translating complex ethical principles and human values into machine-understandable objectives. Balancing technical constraints with theoretical considerations requires close collaboration with cross-functional teams, including ethicists, engineers, and product managers. Additionally, the rapidly evolving landscape of artificial intelligence demands continuous learning to stay current with new alignment techniques and research findings. Navigating these challenges can be intellectually stimulating and offers significant opportunities for interdisciplinary growth.

What are the key skills and qualifications needed to thrive as an AI alignment specialist, and why are they important?

To thrive as an AI Alignment Specialist, you need a strong background in computer science, mathematics, and machine learning, often evidenced by an advanced degree in a related field. Familiarity with technical tools such as Python, TensorFlow, PyTorch, and formal verification systems is typically required, along with understanding of AI safety principles. Analytical thinking, ethical reasoning, and effective communication are crucial soft skills for success in this role. These skills ensure that AI systems are developed safely, ethically, and in alignment with human values, which is essential for mitigating risks associated with advanced AI.

What are popular job titles related to Ai Alignment jobs in California?

For Ai Alignment jobs in California, the most frequently searched job titles are:

What cities in California are hiring for Ai Alignment jobs?

Cities in California with the most Ai Alignment job openings:

Infographic showing various Ai Alignment job openings in California as of August 2026, with employment types broken down into 33% Full Time, and 67% Contract. Highlights an 100% In-person job distribution.

Research Engineer, AI Safety & Alignment

Character.ai

Redwood City, CA โ€ข On-site

$225K - $400K/yr

Full-time

Re-posted 13 days ago


Job description

About the role and team
Joining us as a Research Engineer, you'll be at the forefront of tackling one of the most critical challenges in AI today: safety and alignment. Your work will be pivotal in understanding and mitigating the risks of advanced AI, conducting foundational research to make our models safer, and solving the core technical problems of AI alignment-ensuring our models behave in accordance with human values and intentions.
The Safety team is dedicated to pioneering and implementing techniques that make our models more robust, honest, and harmless. As a Research Engineer, you will bridge the gap between theoretical research and practical application, writing high-quality code to test hypotheses and integrating successful safety solutions directly into our products. Your research will not only protect millions of users but also contribute to the broader scientific community's understanding of how to build safe, beneficial AI.
What you'll do
  • Develop and implement novel evaluation methodologies and metrics to assess the safety and alignment of large language models.
  • Research and develop cutting-edge techniques for model alignment, value learning, and interpretability.
  • Conduct adversarial testing to proactively uncover potential vulnerabilities and failure modes in our models.
  • Analyze and mitigate biases, toxicity, and other harmful behaviors in large language models through techniques like reinforcement learning from human feedback (RLHF) and fine-tuning.
  • Collaborate with engineering and product teams to translate safety research into practical, scalable solutions and best practices.
  • Stay abreast of the latest advancements in AI safety research and contribute to the academic community through publications and presentations.
Who you are
  • Hold a PhD (or equivalent experience) in a relevant field such as Computer Science, Machine Learning, or a related discipline.
  • Write clear and clean production-facing and training code
  • Experience working with GPUs (training, serving, debugging)
  • Experience with data pipelines and data infrastructure
  • Strong understanding of modern machine learning techniques, particularly transformers and reinforcement learning, with a focus on their safety implications.
  • Are passionate about the responsible development of AI and dedicated to solving complex safety challenges.
Nice to Have
  • Experience with product experimentation and A/B testing
  • Experience training large models in a distributed setting
  • Familiarity with ML deployment and orchestration (Kubernetes, Docker, cloud)
  • Experience with explainable AI (XAI) and interpretability techniques.
  • Have research in AI safety, alignment, ethics, or a related area.
  • Knowledge of the broader societal and ethical implications of AI, including policy and governance.
  • Publications in relevant academic journals or conferences in the field of machine learning

About Character.AI
Character.AI empowers people to connect, learn and tell stories through interactive entertainment. Over 20 million people visit Character.AI every month, using our technology to supercharge their creativity and imagination. Our platform lets users engage with tens of millions of characters, enjoy unlimited conversations, and embark on infinite adventures.
In just two years, we achieved unicorn status and were honored as Google Play's AI App of the Year-a testament to our innovative technology and visionary approach.
Join us and be a part of establishing this new entertainment paradigm while shaping the future of Consumer AI!
At Character, we value diversity and welcome applicants from all backgrounds. As an equal opportunity employer, we firmly uphold a non-discrimination policy based on race, religion, national origin, gender, sexual orientation, age, veteran status, or disability. Your unique perspectives are vital to our success.