Python Full Stack AI
San Jose, CA ยท On-site
Develop Retrieval-Augmented Generation (RAG) solutions using vector databases. Integrate LLMs such as GPT, Claude, Gemini, or Llama into enterprise applications. Fine-tune, evaluate, and optimize ...
San Jose, CA ยท On-site
Develop Retrieval-Augmented Generation (RAG) solutions using vector databases. Integrate LLMs such as GPT, Claude, Gemini, or Llama into enterprise applications. Fine-tune, evaluate, and optimize ...
San Jose, CA ยท On-site
Develop Retrieval-Augmented Generation (RAG) solutions using vector databases. Integrate LLMs such as GPT, Claude, Gemini, or Llama into enterprise applications. Fine-tune, evaluate, and optimize ...
Irvine, CA ยท On-site
Build Retrieval-Augmented Generation (RAG) solutions using enterprise data. Develop prompt engineering strategies and context management pipelines. Fine-tune and optimize LLM-powered applications for ...
Irvine, CA ยท On-site
Build Retrieval-Augmented Generation (RAG) solutions using enterprise data. Develop prompt engineering strategies and context management pipelines. Fine-tune and optimize LLM-powered applications for ...
San Diego, CA ยท On-site
$95 - $100/hr
You'll evaluate and integrate large language models, build Retrieval-Augmented Generation systems, and partner with cross-functional teams to deliver scalable, reliable AI solutions. The ideal ...
New
Quick apply
San Diego, CA ยท On-site
$95 - $100/hr
You'll evaluate and integrate large language models, build Retrieval-Augmented Generation systems, and partner with cross-functional teams to deliver scalable, reliable AI solutions. The ideal ...
New
Woodland Hills, CA ยท On-site
AI/ML technologies (e.g., LLMs, NLP, retrieval-augmented generation) - 6 Yrs of Exp * Jira, Azure DevOps, Confluence, and Smartsheet - 6 Yrs of Exp Must have Certifications: PMP, PMI-ACP, or ...
Quick apply
Woodland Hills, CA ยท On-site
AI/ML technologies (e.g., LLMs, NLP, retrieval-augmented generation) - 6 Yrs of Exp * Jira, Azure DevOps, Confluence, and Smartsheet - 6 Yrs of Exp Must have Certifications: PMP, PMI-ACP, or ...
Guided teams on LLM usage patterns, including prompt design, grounding strategies, and Retrieval-Augmented Generation (RAG) * Collaborated with AI architects and MLOps teams to ensure Responsible AI ...
Quick apply
Guided teams on LLM usage patterns, including prompt design, grounding strategies, and Retrieval-Augmented Generation (RAG) * Collaborated with AI architects and MLOps teams to ensure Responsible AI ...
Sunnyvale, CA ยท On-site
$60 - $82.50/hr
... Retrieval-Augmented Generation) architecture * - Hands-on experience on designing solutions using agentic frameworks like LangChain, CrewAI, Semantic Kernel and AutoGen. * - Having experience in ...
Sunnyvale, CA ยท On-site
$60 - $82.50/hr
... Retrieval-Augmented Generation) architecture * - Hands-on experience on designing solutions using agentic frameworks like LangChain, CrewAI, Semantic Kernel and AutoGen. * - Having experience in ...
... retrieval-augmented generation (RAG) * 7+ Years Exp - coding experience in Python, with proficiency in ML/NLP libraries * 7+ Years Exp - healthcare data standards like FHIR, HL7, ICD/CPT, X12 EDI ...
Quick apply
... retrieval-augmented generation (RAG) * 7+ Years Exp - coding experience in Python, with proficiency in ML/NLP libraries * 7+ Years Exp - healthcare data standards like FHIR, HL7, ICD/CPT, X12 EDI ...
D. in Computer Science, Machine Learning, information retrieval, data mining, or a related field 8+ years of experience in leading engineering/applied research/ML experiences in natural language ...
D. in Computer Science, Machine Learning, information retrieval, data mining, or a related field 8+ years of experience in leading engineering/applied research/ML experiences in natural language ...
San Jose, CA ยท On-site
This role focuses on building robust Retrieval-Augmented Generation (RAG) pipelines to ensure AI agents and applications have access to the most relevant, timely, and high-quality information. You'll ...
San Jose, CA ยท On-site
This role focuses on building robust Retrieval-Augmented Generation (RAG) pipelines to ensure AI agents and applications have access to the most relevant, timely, and high-quality information. You'll ...
Develop tools for training and fine-tuning large language models, advanced Retrieval-Augmented Generation (RAG) pipelines, vector databases and agentic frameworks. * Build efficient data pipelines to ...
Develop tools for training and fine-tuning large language models, advanced Retrieval-Augmented Generation (RAG) pipelines, vector databases and agentic frameworks. * Build efficient data pipelines to ...
Retrieval-augmented generation (RAG) systems for dynamic hotel and travel searches * Preference engines that learn and evolve from member interactions * Lightweight AI orchestration layers integrated ...
Retrieval-augmented generation (RAG) systems for dynamic hotel and travel searches * Preference engines that learn and evolve from member interactions * Lightweight AI orchestration layers integrated ...
Develop tools for training and fine-tuning large language models, advanced Retrieval-Augmented Generation (RAG) pipelines, vector databases and agentic frameworks. * Build efficient data pipelines to ...
Develop tools for training and fine-tuning large language models, advanced Retrieval-Augmented Generation (RAG) pipelines, vector databases and agentic frameworks. * Build efficient data pipelines to ...
San Francisco, CA ยท On-site
The ideal candidate will have hands-on experience with Large Language Models (LLMs), Retrieval-Augmented Generation (RAG), AI Agents, prompt engineering, vector databases, and cloud-native AI ...
San Francisco, CA ยท On-site
The ideal candidate will have hands-on experience with Large Language Models (LLMs), Retrieval-Augmented Generation (RAG), AI Agents, prompt engineering, vector databases, and cloud-native AI ...
Build and scale Retrieval-Augmented Generation (RAG) pipelines, multi-agent orchestrations, and API integrations connecting LLMs to internal databases and enterprise software. * Evaluations ...
Build and scale Retrieval-Augmented Generation (RAG) pipelines, multi-agent orchestrations, and API integrations connecting LLMs to internal databases and enterprise software. * Evaluations ...
Cupertino, CA ยท On-site
Expertise in building Retrieval-Augmented Generation (RAG) pipelines * Knowledge of semantic layers and knowledge graphs for structured reasoning * Proficient in Python and FastAPI for backend and ...
Quick apply
Cupertino, CA ยท On-site
Expertise in building Retrieval-Augmented Generation (RAG) pipelines * Knowledge of semantic layers and knowledge graphs for structured reasoning * Proficient in Python and FastAPI for backend and ...
Create scalable RAG (Retrieval-Augmented Generation) pipelines * Design systems for managing context, grounding, and memory * Optimize retrieval using embeddings, vector search, and ranking ...
Create scalable RAG (Retrieval-Augmented Generation) pipelines * Design systems for managing context, grounding, and memory * Optimize retrieval using embeddings, vector search, and ranking ...
San Francisco, CA ยท On-site
$59.25 - $81.50/hr
Implement Retrieval-Augmented Generation (RAG) and orchestrate AI agents using tools like LangChain, LangGraph, or CrewAI. * Integrate Large Language Models (LLMs) with enterprise vector databases ...
San Francisco, CA ยท On-site
$59.25 - $81.50/hr
Implement Retrieval-Augmented Generation (RAG) and orchestrate AI agents using tools like LangChain, LangGraph, or CrewAI. * Integrate Large Language Models (LLMs) with enterprise vector databases ...
AI-Powered Knowledge Systems & Retrieval-Augmented Generation (RAG) * Architect and implement RAG platforms and AI-ready knowledge repositories. * Design intelligent retrieval systems using hybrid ...
AI-Powered Knowledge Systems & Retrieval-Augmented Generation (RAG) * Architect and implement RAG platforms and AI-ready knowledge repositories. * Design intelligent retrieval systems using hybrid ...
Mountain View, CA ยท On-site
$213K - $263K/yr
Advanced AI/ML: Strong expertise in Large Language Models (LLMs), including techniques like prompt engineering and Retrieval-Augmented Generation (RAG) * Coding Excellence: Proficiency in ...
Mountain View, CA ยท On-site
$213K - $263K/yr
Advanced AI/ML: Strong expertise in Large Language Models (LLMs), including techniques like prompt engineering and Retrieval-Augmented Generation (RAG) * Coding Excellence: Proficiency in ...
Mountain View, CA ยท On-site
$213K - $263K/yr
Advanced AI/ML: Strong expertise in Large Language Models (LLMs), including techniques like prompt engineering and Retrieval-Augmented Generation (RAG) * Coding Excellence: Proficiency in ...
Mountain View, CA ยท On-site
$213K - $263K/yr
Advanced AI/ML: Strong expertise in Large Language Models (LLMs), including techniques like prompt engineering and Retrieval-Augmented Generation (RAG) * Coding Excellence: Proficiency in ...
A Retrieval Augmented Generation engineer typically spends their day designing and implementing systems that combine information retrieval with advanced generative models, such as large language models. This includes fine-tuning models, integrating external data sources, developing vector search pipelines, and evaluating output quality. Collaboration with data scientists, machine learning engineers, and product teams is common to ensure the solutions meet user requirements and scale effectively. Additionally, RAG engineers often troubleshoot issues, monitor model performance in production, and stay informed about the latest advancements in AI and information retrieval.
A Retrieval Augmented Generation (RAG) job typically involves developing and optimizing AI systems that enhance text generation by incorporating external knowledge retrieved from relevant sources. Professionals in this field work on integrating retrieval mechanisms with large language models to improve the relevance, accuracy, and factual grounding of generated content. Common responsibilities include designing retrieval systems, fine-tuning language models, optimizing performance, and ensuring the seamless integration of factual data into AI-generated text. This role is highly interdisciplinary, involving expertise in natural language processing (NLP), machine learning, and information retrieval.
To thrive in a Retrieval Augmented Generation (RAG) engineering role, you need a solid background in machine learning, natural language processing (NLP), and experience with scalable information retrieval systems, typically supported by a relevant degree in computer science or a related field. Familiarity with tools such as Python, PyTorch or TensorFlow, vector databases, and search platforms like Elasticsearch is essential, along with practical experience deploying and tuning RAG pipelines. Strong problem-solving skills, a collaborative mindset, and effective communication abilities set outstanding professionals apart in this field. These competencies are crucial for designing, implementing, and optimizing hybrid retrieval-generation AI systems that address complex, real-world information needs.

Other
Posted 11 days ago
Key Responsibilities
Design, develop, and deploy AI/ML and Generative AI applications.
Build AI-powered assistants, copilots, chatbots, and autonomous AI agents.
Develop Retrieval-Augmented Generation (RAG) solutions using vector databases.
Integrate LLMs such as GPT, Claude, Gemini, or Llama into enterprise applications.
Fine-tune, evaluate, and optimize foundation models for business use cases.
Develop REST APIs and microservices for AI applications.
Implement prompt engineering, function calling, AI workflows, and agent orchestration.
Ensure AI solutions meet security, privacy, governance, and responsible AI standards.
Optimize AI model performance, latency, scalability, and cost.
Collaborate with product managers, architects, and data scientists throughout the software lifecycle.