1

Retrieval Augmented Generation Jobs in California

You will design, develop, and deploy intelligent systems that leverage LLMs, Retrieval-Augmented Generation (RAG), MLOps pipelines, and data visualization to optimize engineering productivity, test ...

AI Voice Engineer

Los Angeles, CA · On-site

$140 - $220/hr

Implement conversation memory, tool calling, function execution, and retrieval-augmented generation (RAG) workflows. * Collaborate with product managers, AI engineers, and software developers to ...

AI Engineer

Santa Clara, CA · On-site

$56 - $61/hr

Experience architecting and implementing Retrieval-Augmented Generation (RAG) pipelines to enhance model performance using external knowledge sources, including document chunking, embedding ...

AI and Recommendation Engineer

Oakland, CA · On-site

$119K - $164K/yr

Design and implement Retrieval-Augmented Generation (RAG) solutions, knowledge retrieval systems, and AI-powered assistants. * Evaluate emerging AI technologies, frameworks, and infrastructure to ...

Senior Agentic AI Builder

San Jose, CA · On-site

$107K - $136K/yr

... retrieval-augmented generation pipelines for context-aware AI applications • MCP (Model Context Protocol) - Practical experience with MCP client/server patterns for structured tool-to-data ...

AI/ML Engineer

Burbank, CA · On-site

$111K - $153K/yr

Build and deploy RAG (Retrieval-Augmented Generation) pipelines * Integrate LLMs via APIs (Azure OpenAI preferred) into enterprise applications * Develop and orchestrate agentic AI workflows with ...

This role focuses on building scalable systems leveraging Large Language Models (LLMs), Retrieval-Augmented Generation (RAG), and agentic AI workflows . The ideal candidate will bring deep expertise ...

Software Engineer (Java + GenAI)

San Jose, CA · On-site

$60.75 - $83.25/hr

... Retrieval-Augmented Generation (RAG) - Vector databases - Prompt engineering - Large Language Models (LLMs) - Application: Send suitable profiles and contact details to rams@vensoft.com

Advanced AI/ML: Strong expertise in Large Language Models (LLMs), including techniques like prompt engineering and Retrieval-Augmented Generation (RAG) * Coding Excellence: Proficiency in ...

Implement conversation memory, tool calling, function execution, and retrieval-augmented generation (RAG) workflows. * Collaborate with product managers, AI engineers, and software developers to ...

Senior AI Automation Engineer

San Francisco, CA · On-site

$122K - $160K/yr

Define reusable engineering patterns for prompt frameworks, retrieval-augmented generation, text-to-SQL, semantic data models, agent workflows, evaluation, guardrails, human review, and workflow ...

next page

Showing results 1-20

Retrieval Augmented Generation information

What is a retrieval augmented generation?

A Retrieval Augmented Generation (RAG) job typically involves developing and optimizing AI systems that enhance text generation by incorporating external knowledge retrieved from relevant sources. Professionals in this field work on integrating retrieval mechanisms with large language models to improve the relevance, accuracy, and factual grounding of generated content. Common responsibilities include designing retrieval systems, fine-tuning language models, optimizing performance, and ensuring the seamless integration of factual data into AI-generated text. This role is highly interdisciplinary, involving expertise in natural language processing (NLP), machine learning, and information retrieval.

What does a retrieval augmented generation engineer do?

A Retrieval Augmented Generation engineer typically spends their day designing and implementing systems that combine information retrieval with advanced generative models, such as large language models. This includes fine-tuning models, integrating external data sources, developing vector search pipelines, and evaluating output quality. Collaboration with data scientists, machine learning engineers, and product teams is common to ensure the solutions meet user requirements and scale effectively. Additionally, RAG engineers often troubleshoot issues, monitor model performance in production, and stay informed about the latest advancements in AI and information retrieval.

What skills and qualifications are needed for retrieval augmented generation?

To thrive in a Retrieval Augmented Generation (RAG) engineering role, you need a solid background in machine learning, natural language processing (NLP), and experience with scalable information retrieval systems, typically supported by a relevant degree in computer science or a related field. Familiarity with tools such as Python, PyTorch or TensorFlow, vector databases, and search platforms like Elasticsearch is essential, along with practical experience deploying and tuning RAG pipelines. Strong problem-solving skills, a collaborative mindset, and effective communication abilities set outstanding professionals apart in this field. These competencies are crucial for designing, implementing, and optimizing hybrid retrieval-generation AI systems that address complex, real-world information needs.

What are the most commonly searched types of Retrieval Augmented Generation jobs in California?

The most popular types of Retrieval Augmented Generation jobs in California are:

What are popular job titles related to Retrieval Augmented Generation jobs in California?

For Retrieval Augmented Generation jobs in California, the most frequently searched job titles are:

What job categories do people searching Retrieval Augmented Generation jobs in California look for?

The top searched job categories for Retrieval Augmented Generation jobs in California are:

What cities in California are hiring for Retrieval Augmented Generation jobs?

Cities in California with the most Retrieval Augmented Generation job openings:

Infographic showing various Retrieval Augmented Generation job openings in California as of August 2026, with employment types broken down into 65% Full Time, 32% Part Time, and 3% Contract. Highlights an 63% Physical, 3% Hybrid, and 34% Remote job distribution.

GenAI Full Stack Engineer

The Avian Consulting LLC

San Jose, CA • On-site

Other

Posted 2 days ago

New


Job description

Role Overview

We are seeking a highly skilled GenAI Full Stack Engineer with strong expertise in Angular frontend development and Python backend engineering to design, develop, and deploy enterprise-grade AI-powered applications. The ideal candidate will have hands-on experience building scalable full-stack solutions leveraging Large Language Models (LLMs), Retrieval-Augmented Generation (RAG), AI Agents, Vector Databases, and cloud-based AI services.

 

 

Required Skills

Frontend

  • Angular 14+ / TypeScript

  • HTML5, CSS3, JavaScript

  • Responsive UI development

 

Backend

  • Python

  • FastAPI / Flask / Django

  • REST API Development

  • Microservices Architecture

  • SQL & NoSQL Databases

Generative AI

  • Large Language Models (LLMs)

  • Prompt Engineering

  • Retrieval-Augmented Generation (RAG)

  • Vector Databases

  • LangChain / LangGraph / Semantic Kernel

  • AI Agents and Conversational AI Solutions

 

Skills

Mandatory Skills : AI/Generative AI, Angular, Python