1

Assistant Retrieval Augmented Generation Jobs in Riverside, NJ

Software Engineer

Newtown, PA · On-site

$100 - $125/hr

Help build and integrate Retrieval-Augmented Generation (RAG) pipelines with vector databases (e.g ... assistants, including writing effective prompts and reviewing AI-generated code for quality and ...

Software Engineer

Newtown, PA · On-site +1

$108K - $115K/yr

Help build and integrate Retrieval-Augmented Generation (RAG) pipelines with vector databases (e.g ... assistants, including writing effective prompts and reviewing AI-generated code for quality and ...

Software Engineer

Newtown, PA · On-site

$108K - $115K/yr

Help build and integrate Retrieval-Augmented Generation (RAG) pipelines with vector databases (e.g ... assistants, including writing effective prompts and reviewing AI-generated code for quality and ...

Hands-on experience designing and implementing retrieval-augmented generation (RAG) workflows * Understanding of: * LLM integration and orchestration * Embeddings and vector search * Retrieval ...

Hands-on experience designing and implementing retrieval-augmented generation (RAG) workflows * Understanding of: * LLM integration and orchestration * Embeddings and vector search * Retrieval ...

Hands-on experience designing and implementing retrieval-augmented generation (RAG) workflows * Understanding of: * LLM integration and orchestration * Embeddings and vector search * Retrieval ...

Agent Builder (Amazon Bedrock)

Malvern, PA · Remote

$14.50 - $19.25/hr

Develop Retrieval-Augmented Generation (RAG) solutions using enterprise data. * Integrate agents ... Experience building chatbots, virtual assistants, or AI agents. * Knowledge of RAG, vector ...

Evaluate emerging AI capabilities-including fine-tuning, prompt engineering, retrieval-augmented generation, and multi-agent architectures-and assess their relevance to Vanguard Workplace Solutions ...

Mistral). • Implement RAG (Retrieval-Augmented Generation); vector databases; embeddings; and prompt‐engineering strategies. • Collaborate with product; engineering; and domain teams to ...

Hands-on experience designing or implementing Retrieval-Augmented Generation (RAG) solutions. * Experience with AI standards and frameworks such as Model Context Protocol (MCP), Agent2Agent (A2A ...

next page

Showing results 1-20

Assistant Retrieval Augmented Generation information

See Riverside, NJ salary details

$10

$23

$41

How much do assistant retrieval augmented generation jobs pay per hour?

As of Sep 9, 2026, the average hourly pay for assistant retrieval augmented generation in Riverside, NJ is $23.40, according to ZipRecruiter salary data. Most workers in this role earn between $16.73 and $26.20 per hour, depending on experience, location, and employer.

What is the difference between Assistant Retrieval Augmented Generation vs Data Analyst?

AspectAssistant Retrieval Augmented GenerationData Analyst
Required CredentialsKnowledge of AI, NLP, and retrieval systemsBachelor's in Statistics, Data Science, or related fields
Work EnvironmentTech companies, AI development teamsBusiness, finance, healthcare sectors
Industry UsageAI, machine learning, natural language processingData analysis, reporting, decision support

Assistant Retrieval Augmented Generation focuses on developing AI models that combine retrieval techniques with language generation, often requiring expertise in AI and NLP. Data Analysts interpret data to generate insights, primarily using statistical tools. While both roles involve working with data, Assistant Retrieval Augmented Generation is centered on AI model development, whereas Data Analysts focus on data interpretation and reporting.

What cities near Riverside, NJ are hiring for Assistant Retrieval Augmented Generation jobs?

Cities near Riverside, NJ with the most Assistant Retrieval Augmented Generation job openings:

Infographic showing various Assistant Retrieval Augmented Generation job openings in Riverside, NJ as of August 2026, with employment types broken down into 1% As Needed, 71% Full Time, 23% Part Time, 2% Temporary, and 3% Contract. Highlights an 98% Physical, 1% Hybrid, and 1% Remote job distribution, with an average salary of $48,672 per year, or $23.4 per hour.

Software Developer / Engineer - Philadelphia, PA (Locals Only)

Philadelphia, PA • On-site

Apetan Consulting llc
IT Services • 1 - 10 employees

$80 - $150/hr

Contractor

Re-posted 17 days ago


Job description

Software Developer / Engineer
Location: Philadelphia, PA 
Work Schedule: Hybrid 3 days on site, 2 remote


Position Overview

We are seeking a Software Developer / Engineer to help design and implement an on-premises Large Language Model (LLM) platform with Retrieval-Augmented Generation (RAG) capabilities. This role will focus on deploying open-source AI models, integrating vector databases, and building secure, enterprise-grade AI solutions in a private environment.

This is an excellent opportunity for a developer with hands-on experience in modern AI technologies who enjoys building scalable, high-performance systems.

Responsibilities

  • Deploy and optimize open-source large language models (LLMs) such as Meta Llama 3 and Mistral/Mixtral in on-premises or private environments.
  • Develop Python-based applications for LLM inference, prompt engineering, and model integration.
  • Optimize CPU-based model inference through quantization and performance tuning.
  • Design and implement Retrieval-Augmented Generation (RAG) (RAG) pipelines.
  • Configure and manage open-source vector databases such as Qdrant, Chroma, Milvus, or pgvector.
  • Generate and manage embeddings while implementing metadata filtering strategies.
  • Support enterprise security requirements, including air-gapped deployments, access controls, data privacy, and audit logging.
  • Produce technical documentation, deployment guidance, and knowledge transfer materials for internal teams.
  • Build a working prototype integrating an LLM, vector database, and RAG architecture.

Required Qualifications

  • Professional experience deploying open-source LLMs (e.g., Meta Llama 3, Mistral/Mixtral) in on-premises or private environments.
  • Strong Python development experience.
  • Hands-on experience with LLM inference, prompt engineering, and AI application integration.
  • Experience optimizing CPU-based inference through model quantization and performance tuning.
  • Experience with vector databases such as Qdrant, Chroma, Milvus, or pgvector.
  • Proven experience implementing Retrieval-Augmented Generation (RAG) solutions.
  • Understanding of enterprise security, data privacy, air-gapped environments, access controls, and audit logging.

Preferred Qualifications

  • Experience with LangChain or LlamaIndex.
  • Familiarity with Docker and Kubernetes.
  • Experience with inference frameworks such as vLLM, llama.cpp, or Hugging Face Transformers.
  • Experience with Rust, Go, or C++.
  • Previous experience working in enterprise or regulated environments.

Deliverables

  • Reference architecture and deployment guidance.
  • Working prototype integrating an LLM, vector database, and RAG solution.
  • Technical documentation and knowledge transfer to internal teams.