This role focuses on building scalable, production-grade AI systems, leveraging advanced API development, retrieval-augmented generation (RAG), and agentic frameworks. You will work on cutting-edge ...
Quick apply
This role focuses on building scalable, production-grade AI systems, leveraging advanced API development, retrieval-augmented generation (RAG), and agentic frameworks. You will work on cutting-edge ...
Quick apply
This role focuses on building scalable, production-grade AI systems, leveraging advanced API development, retrieval-augmented generation (RAG), and agentic frameworks. You will work on cutting-edge ...
Mclean, VA · On-site
Job Summary : Closure Technologies is seeking an AI/ML Engineer who will implement and maintain Retrieval-Augmented Generation (RAG) pipelines and integrate Large Language Models (LLMs) into ...
Mclean, VA · On-site
Job Summary : Closure Technologies is seeking an AI/ML Engineer who will implement and maintain Retrieval-Augmented Generation (RAG) pipelines and integrate Large Language Models (LLMs) into ...
Santa Clara, CA · On-site
$56 - $61/hr
Experience architecting and implementing Retrieval-Augmented Generation (RAG) pipelines to enhance model performance using external knowledge sources, including document chunking, embedding ...
Quick apply
Santa Clara, CA · On-site
$56 - $61/hr
Experience architecting and implementing Retrieval-Augmented Generation (RAG) pipelines to enhance model performance using external knowledge sources, including document chunking, embedding ...
West Palm Beach, FL · On-site
$62.75 - $82.25/hr
The ideal candidate will combine strong software engineering skills with deep expertise in Generative AI, Agentic AI, Retrieval-Augmented Generation (RAG), and AWS cloud services to build scalable ...
West Palm Beach, FL · On-site
$62.75 - $82.25/hr
The ideal candidate will combine strong software engineering skills with deep expertise in Generative AI, Agentic AI, Retrieval-Augmented Generation (RAG), and AWS cloud services to build scalable ...
Washington, DC · On-site
Integrate with large language models (LLMs) and generative AI (GenAI) using prompt engineering, fine-tuning, and retrieval-augmented generation (RAG) techniques. * Implement MCP client and server ...
Washington, DC · On-site
Integrate with large language models (LLMs) and generative AI (GenAI) using prompt engineering, fine-tuning, and retrieval-augmented generation (RAG) techniques. * Implement MCP client and server ...
Boston, MA · On-site
You'll provide technical leadership while driving the development of enterprise-scale AI solutions leveraging Large Language Models (LLMs), Retrieval-Augmented Generation (RAG), Agentic AI, NLP ...
Boston, MA · On-site
You'll provide technical leadership while driving the development of enterprise-scale AI solutions leveraging Large Language Models (LLMs), Retrieval-Augmented Generation (RAG), Agentic AI, NLP ...
Warren, NJ · On-site
$56.50 - $74.75/hr
The ideal candidate will have strong skills in Python, React.js, Retrieval-Augmented Generation (RAG ), and hands-on experience with LangChain and LangGraph frameworks. You will be responsible for ...
Quick apply
Warren, NJ · On-site
$56.50 - $74.75/hr
The ideal candidate will have strong skills in Python, React.js, Retrieval-Augmented Generation (RAG ), and hands-on experience with LangChain and LangGraph frameworks. You will be responsible for ...
Chicago, IL · On-site
$144K - $177K/yr
The ideal candidate will bring deep expertise in Python, FastAPI, and Retrieval-Augmented Generation (RAG) solutions, with hands-on experience deploying scalable AI applications on Azure. This role ...
Quick apply
Chicago, IL · On-site
$144K - $177K/yr
The ideal candidate will bring deep expertise in Python, FastAPI, and Retrieval-Augmented Generation (RAG) solutions, with hands-on experience deploying scalable AI applications on Azure. This role ...
Boston, MA · On-site
$113K - $155K/yr
This role focuses on developing AI applications powered by large language models (LLMs), retrieval-augmented generation (RAG), Model Context Protocol (MCP) servers, and Agentic AI across the ...
Quick apply
Boston, MA · On-site
$113K - $155K/yr
This role focuses on developing AI applications powered by large language models (LLMs), retrieval-augmented generation (RAG), Model Context Protocol (MCP) servers, and Agentic AI across the ...
Retrieval-Augmented Generation (RAG) * Strong experience building production-grade services using Python and/or Java * Experience integrating LLMs or AI services into enterprise applications
Quick apply
Retrieval-Augmented Generation (RAG) * Strong experience building production-grade services using Python and/or Java * Experience integrating LLMs or AI services into enterprise applications
Retrieval-Augmented Generation (RAG) * Strong experience building production-grade services using Python and/or Java * Experience integrating LLMs or AI services into enterprise applications
Quick apply
Retrieval-Augmented Generation (RAG) * Strong experience building production-grade services using Python and/or Java * Experience integrating LLMs or AI services into enterprise applications
... retrieval augmented generation (RAG) and model context protocol (MCP)
... retrieval augmented generation (RAG) and model context protocol (MCP)
Retrieval-Augmented Generation (RAG) * Strong experience building production-grade services using Python and/or Java * Experience integrating LLMs or AI services into enterprise applications
Quick apply
Retrieval-Augmented Generation (RAG) * Strong experience building production-grade services using Python and/or Java * Experience integrating LLMs or AI services into enterprise applications
Dallas, TX · On-site
... Retrieval-Augmented Generation (RAG) pipelines, and Agent SDKs - Skilled in building and deploying AI/LLM systems in production environments - Familiarity with AI agents, including evaluation ...
Quick apply
Dallas, TX · On-site
... Retrieval-Augmented Generation (RAG) pipelines, and Agent SDKs - Skilled in building and deploying AI/LLM systems in production environments - Familiarity with AI agents, including evaluation ...
Retrieval-Augmented Generation (RAG) * Strong experience building production-grade services using Python and/or Java * Experience integrating LLMs or AI services into enterprise applications
Quick apply
Retrieval-Augmented Generation (RAG) * Strong experience building production-grade services using Python and/or Java * Experience integrating LLMs or AI services into enterprise applications
Phoenix, AZ · On-site
$100K - $110K/yr
... Retrieval-Augmented Generation (RAG), and Agentic AI technologies. Working within a highly regulated financial environment, you will contribute to building scalable data solutions and support AI ...
Phoenix, AZ · On-site
$100K - $110K/yr
... Retrieval-Augmented Generation (RAG), and Agentic AI technologies. Working within a highly regulated financial environment, you will contribute to building scalable data solutions and support AI ...
Retrieval-Augmented Generation (RAG) * Strong experience building production-grade services using Python and/or Java * Experience integrating LLMs or AI services into enterprise applications
Quick apply
Retrieval-Augmented Generation (RAG) * Strong experience building production-grade services using Python and/or Java * Experience integrating LLMs or AI services into enterprise applications
Dallas, TX · On-site
Develop and maintain Retrieval-Augmented Generation (RAG) architectures using vector databases and semantic search technologies * Create, test, and refine prompts, structured outputs, and evaluation ...
Dallas, TX · On-site
Develop and maintain Retrieval-Augmented Generation (RAG) architectures using vector databases and semantic search technologies * Create, test, and refine prompts, structured outputs, and evaluation ...
Retrieval-Augmented Generation (RAG) * Strong experience building production-grade services using Python and/or Java * Experience integrating LLMs or AI services into enterprise applications
Quick apply
Retrieval-Augmented Generation (RAG) * Strong experience building production-grade services using Python and/or Java * Experience integrating LLMs or AI services into enterprise applications
Manhattan, NY · On-site
New Jersey /New York (Onsite) We are looking for a GenAI Engineer with strong hands-on experience building RAG (Retrieval-Augmented Generation) solutions and document vectorization pipelines. Must ...
Manhattan, NY · On-site
New Jersey /New York (Onsite) We are looking for a GenAI Engineer with strong hands-on experience building RAG (Retrieval-Augmented Generation) solutions and document vectorization pipelines. Must ...
$15.14 - $16.19
6% of jobs
$17.16 is the 25th percentile. Wages below this are outliers.
$16.19 - $17.24
20% of jobs
$17.24 - $18.29
13% of jobs
The median wage is $19.04 / hr.
$18.29 - $19.34
15% of jobs
$19.34 - $20.39
13% of jobs
$20.95 is the 75th percentile. Wages above this are outliers.
$20.39 - $21.44
15% of jobs
$21.44 - $22.49
3% of jobs
$22.49 - $23.54
2% of jobs
$23.54 - $24.58
4% of jobs
$24.58 - $25.63
7% of jobs
$25.63 - $26.68
1% of jobs
$15
$20
$26
Cities with the most Retrieval Augmented Generation Rag job openings:
States with the most job openings for Retrieval Augmented Generation Rag jobs include:
The top searched job categories for Retrieval Augmented Generation Rag jobs are:

Contractor
Re-posted 22 days ago
Job Description :
We are seeking for highly skilled Software Engineer with strong expertise in modern Python development and Large Language Model (LLM) ecosystems.
This role focuses on building scalable, production-grade AI systems, leveraging advanced API development, retrieval-augmented generation (RAG), and agentic frameworks.
You will work on cutting-edge AI solutions, contributing to the design, development, and deployment of intelligent systems within a distributed enterprise environment
Core Programming & Backend Development
Develop robust, scalable applications using Python (intermediate to advanced level)
Implement asynchronous programming patterns for high-performance systems
Design and build RESTful APIs using FastAPI
Write clean, maintainable, production-grade code
Develop and execute unit and integration tests
Debug and resolve issues in complex distributed systems
LLM Fundamentals & Prompt Engineering
Design and optimize prompts for various LLM use cases
Understand tokenization, context windows, and model limitations
Select appropriate models based on performance and cost trade-offs
Mitigate hallucinations and ensure grounded, reliable responses
Retrieval-Augmented Generation (RAG)
Build and maintain document ingestion and preprocessing pipelines
Implement chunking strategies (semantic, recursive, sliding window)
Generate and manage embeddings
Work with vector databases (e.g., pgvector)
Design hybrid search systems combining keyword (BM25) and semantic search
Optimize re-ranking and relevance tuning mechanisms
Agentic Frameworks & Orchestration
Design and implement multi-agent systems
Manage conversational and long-term memory
Build workflow orchestration pipelines for AI agents
Required Qualifications
Strong proficiency in Python with experience in asynchronous programming
Hands-on experience with FastAPI or similar frameworks
Experience building scalable backend systems
Solid understanding of LLM concepts and prompt engineering
Experience with RAG pipelines and vector databases (pgvector preferred)
Familiarity with distributed systems and debugging techniques
Experience with Docker and CI/CD pipelines
Sourced by ZipRecruiter
It services
51 - 200 Employees
Sheridan, WY, US