This role focuses on building scalable, production-grade AI systems, leveraging advanced API development, retrieval-augmented generation (RAG), and agentic frameworks. You will work on cutting-edge ...
Quick apply
This role focuses on building scalable, production-grade AI systems, leveraging advanced API development, retrieval-augmented generation (RAG), and agentic frameworks. You will work on cutting-edge ...
Quick apply
This role focuses on building scalable, production-grade AI systems, leveraging advanced API development, retrieval-augmented generation (RAG), and agentic frameworks. You will work on cutting-edge ...
Boston, MA · On-site
$113K - $155K/yr
This role focuses on developing AI applications powered by large language models (LLMs), retrieval-augmented generation (RAG), Model Context Protocol (MCP) servers, and Agentic AI across the ...
Quick apply
Boston, MA · On-site
$113K - $155K/yr
This role focuses on developing AI applications powered by large language models (LLMs), retrieval-augmented generation (RAG), Model Context Protocol (MCP) servers, and Agentic AI across the ...
Palo Alto, CA · On-site
$123K - $168K/yr
They are seeking a Senior AI Engineer to lead the development of Retrieval-Augmented Generation ... Responsibilities : • Lead the architecture and development of RAG systems that combine LLMs (e.g ...
Palo Alto, CA · On-site
$123K - $168K/yr
They are seeking a Senior AI Engineer to lead the development of Retrieval-Augmented Generation ... Responsibilities : • Lead the architecture and development of RAG systems that combine LLMs (e.g ...
We are seeking a highly skilled AI Engineer with hands-on experience in agent development, Retrieval-Augmented Generation (RAG), agentic workflows , and platforms such as Cursor AI . The ideal ...
We are seeking a highly skilled AI Engineer with hands-on experience in agent development, Retrieval-Augmented Generation (RAG), agentic workflows , and platforms such as Cursor AI . The ideal ...
Charlotte, NC · On-site
$102K - $140K/yr
Evaluation frameworks Retrieval-Augmented Generation (RAG
Charlotte, NC · On-site
$102K - $140K/yr
Evaluation frameworks Retrieval-Augmented Generation (RAG
New York, NY · On-site
$170K/yr
This role designs and implements scalable artificial intelligence systems leveraging Large Language Models (LLMs), Retrieval-Augmented Generation (RAG) frameworks, AI agents, model orchestration ...
New York, NY · On-site
$170K/yr
This role designs and implements scalable artificial intelligence systems leveraging Large Language Models (LLMs), Retrieval-Augmented Generation (RAG) frameworks, AI agents, model orchestration ...
Austin, TX · On-site
$65 - $70/hr
We are seeking a highly skilled AI Engineer with hands-on experience in agent development, Retrieval-Augmented Generation (RAG), agentic workflows , and platforms such as Cursor AI . The ideal ...
Quick apply
Austin, TX · On-site
$65 - $70/hr
We are seeking a highly skilled AI Engineer with hands-on experience in agent development, Retrieval-Augmented Generation (RAG), agentic workflows , and platforms such as Cursor AI . The ideal ...
Warren, NJ · On-site
$56.50 - $74.75/hr
The ideal candidate will have strong skills in Python, React.js, Retrieval-Augmented Generation (RAG ), and hands-on experience with LangChain and LangGraph frameworks. You will be responsible for ...
Quick apply
Warren, NJ · On-site
$56.50 - $74.75/hr
The ideal candidate will have strong skills in Python, React.js, Retrieval-Augmented Generation (RAG ), and hands-on experience with LangChain and LangGraph frameworks. You will be responsible for ...
Chicago, IL · On-site
$144K - $177K/yr
The ideal candidate will bring deep expertise in Python, FastAPI, and Retrieval-Augmented Generation (RAG) solutions, with hands-on experience deploying scalable AI applications on Azure. This role ...
Quick apply
Chicago, IL · On-site
$144K - $177K/yr
The ideal candidate will bring deep expertise in Python, FastAPI, and Retrieval-Augmented Generation (RAG) solutions, with hands-on experience deploying scalable AI applications on Azure. This role ...
Annapolis, MD · Hybrid
The ideal candidate will possess strong expertise in Python development , LLM integration , retrieval-augmented generation (RAG) , chatbot development , workflow automation , and AI model deployment ...
Quick apply
Annapolis, MD · Hybrid
The ideal candidate will possess strong expertise in Python development , LLM integration , retrieval-augmented generation (RAG) , chatbot development , workflow automation , and AI model deployment ...
Develop and maintain Retrieval-Augmented Generation (RAG) architectures using vector databases and semantic search technologies * Create, test, and refine prompts, structured outputs, and evaluation ...
Develop and maintain Retrieval-Augmented Generation (RAG) architectures using vector databases and semantic search technologies * Create, test, and refine prompts, structured outputs, and evaluation ...
Washington, DC · On-site
Integrate with large language models (LLMs) and generative AI (GenAI) using prompt engineering, fine-tuning, and retrieval-augmented generation (RAG) techniques. * Implement MCP client and server ...
Washington, DC · On-site
Integrate with large language models (LLMs) and generative AI (GenAI) using prompt engineering, fine-tuning, and retrieval-augmented generation (RAG) techniques. * Implement MCP client and server ...
Dallas, TX · On-site
... Retrieval-Augmented Generation (RAG) pipelines, and Agent SDKs - Skilled in building and deploying AI/LLM systems in production environments - Familiarity with AI agents, including evaluation ...
Quick apply
Dallas, TX · On-site
... Retrieval-Augmented Generation (RAG) pipelines, and Agent SDKs - Skilled in building and deploying AI/LLM systems in production environments - Familiarity with AI agents, including evaluation ...
Medina, MN · Hybrid
$130K - $170K/yr
You will lead Retrieval-Augmented Generation (RAG) and knowledge-retrieval initiatives that support a variety of business functions. You will partner cross-functionally with business leaders, IT ...
Medina, MN · Hybrid
$130K - $170K/yr
You will lead Retrieval-Augmented Generation (RAG) and knowledge-retrieval initiatives that support a variety of business functions. You will partner cross-functionally with business leaders, IT ...
New York, NY · On-site
$134K - $176K/yr
As large language models and Retrieval-Augmented Generation (RAG) technologies reshape enterprise search and knowledge discovery, we are investing in next-generation AI systems that combine ...
New York, NY · On-site
$134K - $176K/yr
As large language models and Retrieval-Augmented Generation (RAG) technologies reshape enterprise search and knowledge discovery, we are investing in next-generation AI systems that combine ...
Dallas, TX · On-site
Develop and maintain Retrieval-Augmented Generation (RAG) architectures using vector databases and semantic search technologies * Create, test, and refine prompts, structured outputs, and evaluation ...
Dallas, TX · On-site
Develop and maintain Retrieval-Augmented Generation (RAG) architectures using vector databases and semantic search technologies * Create, test, and refine prompts, structured outputs, and evaluation ...
... retrieval augmented generation (RAG) and model context protocol (MCP)
... retrieval augmented generation (RAG) and model context protocol (MCP)
Dallas, TX · On-site
Develop and maintain Retrieval-Augmented Generation (RAG) architectures using vector databases and semantic search technologies * Create, test, and refine prompts, structured outputs, and evaluation ...
Quick apply
Dallas, TX · On-site
Develop and maintain Retrieval-Augmented Generation (RAG) architectures using vector databases and semantic search technologies * Create, test, and refine prompts, structured outputs, and evaluation ...
The ideal candidate has hands-on experience with machine learning, large language models (LLMs), Retrieval-Augmented Generation (RAG), and enterprise data systems. Collaborate with data engineers ...
The ideal candidate has hands-on experience with machine learning, large language models (LLMs), Retrieval-Augmented Generation (RAG), and enterprise data systems. Collaborate with data engineers ...
This role focuses on building an Agentic AI Platform for internal enterprise users using Python, Large Language Models (LLMs), and Retrieval-Augmented Generation (RAG) technologies. The ideal ...
Quick apply
This role focuses on building an Agentic AI Platform for internal enterprise users using Python, Large Language Models (LLMs), and Retrieval-Augmented Generation (RAG) technologies. The ideal ...
$15.14 - $16.19
6% of jobs
$17.16 is the 25th percentile. Wages below this are outliers.
$16.19 - $17.24
20% of jobs
$17.24 - $18.29
13% of jobs
The median wage is $19.04 / hr.
$18.29 - $19.34
15% of jobs
$19.34 - $20.39
13% of jobs
$20.95 is the 75th percentile. Wages above this are outliers.
$20.39 - $21.44
15% of jobs
$21.44 - $22.49
3% of jobs
$22.49 - $23.54
2% of jobs
$23.54 - $24.58
4% of jobs
$24.58 - $25.63
7% of jobs
$25.63 - $26.68
1% of jobs
$15
$20
$26

Contractor
Posted 11 days ago
Job Description :
We are seeking for highly skilled Software Engineer with strong expertise in modern Python development and Large Language Model (LLM) ecosystems.
This role focuses on building scalable, production-grade AI systems, leveraging advanced API development, retrieval-augmented generation (RAG), and agentic frameworks.
You will work on cutting-edge AI solutions, contributing to the design, development, and deployment of intelligent systems within a distributed enterprise environment
Core Programming & Backend Development
Develop robust, scalable applications using Python (intermediate to advanced level)
Implement asynchronous programming patterns for high-performance systems
Design and build RESTful APIs using FastAPI
Write clean, maintainable, production-grade code
Develop and execute unit and integration tests
Debug and resolve issues in complex distributed systems
LLM Fundamentals & Prompt Engineering
Design and optimize prompts for various LLM use cases
Understand tokenization, context windows, and model limitations
Select appropriate models based on performance and cost trade-offs
Mitigate hallucinations and ensure grounded, reliable responses
Retrieval-Augmented Generation (RAG)
Build and maintain document ingestion and preprocessing pipelines
Implement chunking strategies (semantic, recursive, sliding window)
Generate and manage embeddings
Work with vector databases (e.g., pgvector)
Design hybrid search systems combining keyword (BM25) and semantic search
Optimize re-ranking and relevance tuning mechanisms
Agentic Frameworks & Orchestration
Design and implement multi-agent systems
Manage conversational and long-term memory
Build workflow orchestration pipelines for AI agents
Required Qualifications
Strong proficiency in Python with experience in asynchronous programming
Hands-on experience with FastAPI or similar frameworks
Experience building scalable backend systems
Solid understanding of LLM concepts and prompt engineering
Experience with RAG pipelines and vector databases (pgvector preferred)
Familiarity with distributed systems and debugging techniques
Experience with Docker and CI/CD pipelines
Sourced by ZipRecruiter