Guided teams on LLM usage patterns, including prompt design, grounding strategies, and Retrieval-Augmented Generation (RAG) * Collaborated with AI architects and MLOps teams to ensure Responsible AI ...
Quick apply
Guided teams on LLM usage patterns, including prompt design, grounding strategies, and Retrieval-Augmented Generation (RAG) * Collaborated with AI architects and MLOps teams to ensure Responsible AI ...
Quick apply
Guided teams on LLM usage patterns, including prompt design, grounding strategies, and Retrieval-Augmented Generation (RAG) * Collaborated with AI architects and MLOps teams to ensure Responsible AI ...
... retrieval-augmented generation (RAG) * 7+ Years Exp - coding experience in Python, with proficiency in ML/NLP libraries * 7+ Years Exp - healthcare data standards like FHIR, HL7, ICD/CPT, X12 EDI ...
Quick apply
... retrieval-augmented generation (RAG) * 7+ Years Exp - coding experience in Python, with proficiency in ML/NLP libraries * 7+ Years Exp - healthcare data standards like FHIR, HL7, ICD/CPT, X12 EDI ...
Palo Alto, CA · On-site
$122K - $168K/yr
We are now hiring a Sr. AI Engineer - LLM, RAG to lead the development of Retrieval-Augmented Generation (RAG) systems that harness the power of large language models (LLMs) and real-world knowledge ...
Palo Alto, CA · On-site
$122K - $168K/yr
We are now hiring a Sr. AI Engineer - LLM, RAG to lead the development of Retrieval-Augmented Generation (RAG) systems that harness the power of large language models (LLMs) and real-world knowledge ...
$122K - $168K/yr
We are now hiring a Sr. AI Engineer - LLM, RAG to lead the development of Retrieval-Augmented Generation (RAG) systems that harness the power of large language models (LLMs) and real-world knowledge ...
$122K - $168K/yr
We are now hiring a Sr. AI Engineer - LLM, RAG to lead the development of Retrieval-Augmented Generation (RAG) systems that harness the power of large language models (LLMs) and real-world knowledge ...
Irvine, CA · On-site
$140 - $210/hr
Build Retrieval-Augmented Generation solutions using services such as BigQuery, Vertex AI Vector Search, Cloud Storage, and Document AI. * Develop APIs, microservices, agent tools, MCP integrations ...
Irvine, CA · On-site
$140 - $210/hr
Build Retrieval-Augmented Generation solutions using services such as BigQuery, Vertex AI Vector Search, Cloud Storage, and Document AI. * Develop APIs, microservices, agent tools, MCP integrations ...
Santa Clara, CA · On-site
$136 - $265/hr
Develop tools for training and fine-tuning large language models, advanced Retrieval-Augmented Generation (RAG) pipelines, vector databases and agentic frameworks. * Build efficient data pipelines to ...
Santa Clara, CA · On-site
$136 - $265/hr
Develop tools for training and fine-tuning large language models, advanced Retrieval-Augmented Generation (RAG) pipelines, vector databases and agentic frameworks. * Build efficient data pipelines to ...
Develop tools for training and fine-tuning large language models, advanced Retrieval-Augmented Generation (RAG) pipelines, vector databases and agentic frameworks. * Build efficient data pipelines to ...
Develop tools for training and fine-tuning large language models, advanced Retrieval-Augmented Generation (RAG) pipelines, vector databases and agentic frameworks. * Build efficient data pipelines to ...
San Jose, CA · On-site
$90/hr
Contribute to architecture, retrieval-augmented generation (RAG) design, prompt engineering, and CI/CD practices that support rapid, reliable delivery of AI features. Succeed in this role by shipping ...
San Jose, CA · On-site
$90/hr
Contribute to architecture, retrieval-augmented generation (RAG) design, prompt engineering, and CI/CD practices that support rapid, reliable delivery of AI features. Succeed in this role by shipping ...
Retrieval-augmented generation (RAG) systems for dynamic hotel and travel searches * Preference engines that learn and evolve from member interactions * Lightweight AI orchestration layers integrated ...
Retrieval-augmented generation (RAG) systems for dynamic hotel and travel searches * Preference engines that learn and evolve from member interactions * Lightweight AI orchestration layers integrated ...
Cupertino, CA · On-site
Expertise in building Retrieval-Augmented Generation (RAG) pipelines * Knowledge of semantic layers and knowledge graphs for structured reasoning * Proficient in Python and FastAPI for backend and ...
Quick apply
Cupertino, CA · On-site
Expertise in building Retrieval-Augmented Generation (RAG) pipelines * Knowledge of semantic layers and knowledge graphs for structured reasoning * Proficient in Python and FastAPI for backend and ...
Build and scale Retrieval-Augmented Generation (RAG) pipelines, multi-agent orchestrations, and API integrations connecting LLMs to internal databases and enterprise software. * Evaluations ...
Build and scale Retrieval-Augmented Generation (RAG) pipelines, multi-agent orchestrations, and API integrations connecting LLMs to internal databases and enterprise software. * Evaluations ...
Mountain View, CA · On-site
$213K - $263K/yr
Advanced AI/ML: Strong expertise in Large Language Models (LLMs), including techniques like prompt engineering and Retrieval-Augmented Generation (RAG) * Coding Excellence: Proficiency in ...
Mountain View, CA · On-site
$213K - $263K/yr
Advanced AI/ML: Strong expertise in Large Language Models (LLMs), including techniques like prompt engineering and Retrieval-Augmented Generation (RAG) * Coding Excellence: Proficiency in ...
San Jose, CA · On-site
$30 - $50/hr
Additionally, the candidate will implement Retrieval-Augmented Generation (RAG) models and fine-tune open-source models from leading providers like Meta, Google, Deepseek, and Nvidia. Collaboration ...
San Jose, CA · On-site
$30 - $50/hr
Additionally, the candidate will implement Retrieval-Augmented Generation (RAG) models and fine-tune open-source models from leading providers like Meta, Google, Deepseek, and Nvidia. Collaboration ...
Long Beach, CA · On-site
$106.20 - $182.98/hr
Implement retrieval-augmented generation, vector search, structured data access, and document processing where appropriate.* Partner with data and platform teams to ensure quality, lineage ...
Long Beach, CA · On-site
$106.20 - $182.98/hr
Implement retrieval-augmented generation, vector search, structured data access, and document processing where appropriate.* Partner with data and platform teams to ensure quality, lineage ...
Woodland Hills, CA · On-site
RAG (Retrieval-Augmented Generation) Pipelines * Document Extraction, Parsing & Chunking * Structured & Unstructured Data Processing * Embeddings & Vector Search * Vector Databases & MongoDB
Quick apply
Woodland Hills, CA · On-site
RAG (Retrieval-Augmented Generation) Pipelines * Document Extraction, Parsing & Chunking * Structured & Unstructured Data Processing * Embeddings & Vector Search * Vector Databases & MongoDB
Familiarity with Large Language Models (LLMs) and Generative AI (GenAI) technologies including Retrieval-Augmented Generation (RAG) and model tuning. * Familiarity with SLMs: model design and fine ...
Familiarity with Large Language Models (LLMs) and Generative AI (GenAI) technologies including Retrieval-Augmented Generation (RAG) and model tuning. * Familiarity with SLMs: model design and fine ...
San Francisco, CA · On-site
$90 - $120/hr
You will guide clients toward the best approaches--such as Retrieval-Augmented Generation (RAG), agents, or fine-tuning--for optimized and cost-effective impact. Beyond prompt creation, you will be ...
San Francisco, CA · On-site
$90 - $120/hr
You will guide clients toward the best approaches--such as Retrieval-Augmented Generation (RAG), agents, or fine-tuning--for optimized and cost-effective impact. Beyond prompt creation, you will be ...
... and retrieval-augmented generation. * Engineering integrations between data platforms, governance, risk, and compliance workflows, and enterprise systems using application programming interfaces ...
... and retrieval-augmented generation. * Engineering integrations between data platforms, governance, risk, and compliance workflows, and enterprise systems using application programming interfaces ...
Culver City, CA · On-site
$73/hr
Implement Retrieval-Augmented Generation (RAG), Model Context Protocol (MCP), and other advanced AI integration technologies. * Create AI-native workflow automation solutions using skills-based ...
Culver City, CA · On-site
$73/hr
Implement Retrieval-Augmented Generation (RAG), Model Context Protocol (MCP), and other advanced AI integration technologies. * Create AI-native workflow automation solutions using skills-based ...
Build and continuously improve the retrieval-augmented generation pipeline -- document ingestion from Slack, GitHub, Jira, and Confluence; chunking and embedding strategies; hybrid vector + keyword ...
Build and continuously improve the retrieval-augmented generation pipeline -- document ingestion from Slack, GitHub, Jira, and Confluence; chunking and embedding strategies; hybrid vector + keyword ...
A Retrieval Augmented Generation (RAG) job typically involves developing and optimizing AI systems that enhance text generation by incorporating external knowledge retrieved from relevant sources. Professionals in this field work on integrating retrieval mechanisms with large language models to improve the relevance, accuracy, and factual grounding of generated content. Common responsibilities include designing retrieval systems, fine-tuning language models, optimizing performance, and ensuring the seamless integration of factual data into AI-generated text. This role is highly interdisciplinary, involving expertise in natural language processing (NLP), machine learning, and information retrieval.
A Retrieval Augmented Generation engineer typically spends their day designing and implementing systems that combine information retrieval with advanced generative models, such as large language models. This includes fine-tuning models, integrating external data sources, developing vector search pipelines, and evaluating output quality. Collaboration with data scientists, machine learning engineers, and product teams is common to ensure the solutions meet user requirements and scale effectively. Additionally, RAG engineers often troubleshoot issues, monitor model performance in production, and stay informed about the latest advancements in AI and information retrieval.
To thrive in a Retrieval Augmented Generation (RAG) engineering role, you need a solid background in machine learning, natural language processing (NLP), and experience with scalable information retrieval systems, typically supported by a relevant degree in computer science or a related field. Familiarity with tools such as Python, PyTorch or TensorFlow, vector databases, and search platforms like Elasticsearch is essential, along with practical experience deploying and tuning RAG pipelines. Strong problem-solving skills, a collaborative mindset, and effective communication abilities set outstanding professionals apart in this field. These competencies are crucial for designing, implementing, and optimizing hybrid retrieval-generation AI systems that address complex, real-world information needs.
The most popular types of Retrieval Augmented Generation jobs in California are:
For Retrieval Augmented Generation jobs in California, the most frequently searched job titles are:
The top searched job categories for Retrieval Augmented Generation jobs in California are:
Cities in California with the most Retrieval Augmented Generation job openings:

Contractor
Re-posted 15 days ago
Sourced by ZipRecruiter
It services
51 - 200 Employees
Princeton, NJ, US