1

Retrieval Augmented Generation Rag Jobs in Pleasanton, CA

The ideal candidate should have hands-on expertise with Retrieval-Augmented Generation (RAG), Agentic AI workflows, and LLM-based automation, with the ability to integrate AI systems into complex ...

... retrieval-augmented generation (RAG) systems. • Collaborate closely with product and engineering teams to turn early prototypes into production-grade systems. • Help shape the future of AI ...

... RAG (Retrieval-Augmented Generation) models for efficient AI-driven data retrieval. • Implement semantic search and high-speed querying mechanisms for AI-driven insights. • Develop AI-powered ...

Advanced AI/ML: Strong expertise in Large Language Models (LLMs), including techniques like prompt engineering and Retrieval-Augmented Generation (RAG) * Coding Excellence: Proficiency in ...

Familiarity with Large Language Models (LLMs) and Generative AI (GenAI) technologies including Retrieval-Augmented Generation (RAG) and model tuning. * Familiarity with SLMs: model design and fine ...

... retrieval-augmented generation (RAG) systems. • Collaborate closely with product and engineering teams to turn early prototypes into production-grade systems. • Help shape the future of AI ...

... retrieval-augmented generation (RAG) systems. • Collaborate closely with product and engineering teams to turn early prototypes into production-grade systems. • Help shape the future of AI ...

Architect and implement Retrieval-Augmented Generation (RAG) pipelines to enhance model performance using external knowledge sources, including document chunking, embedding generation, and retrieval ...

Lead AI Platform Engineer

Cupertino, CA · On-site

$126K - $166K/yr

Retrieval-Augmented Generation (RAG) architectures * Prompt engineering techniques * Agentic AI workflows and orchestration * Build intelligent systems using frameworks such as LangChain, LangGraph ...

Lead AI Engineer - Observability

Cupertino, CA · On-site

$126K - $166K/yr

Retrieval-Augmented Generation (RAG) architectures * Prompt engineering techniques * Agentic AI workflows and orchestration * Build intelligent systems using frameworks such as LangChain, LangGraph ...

New

Gen AI Architect

Santa Clara, CA · On-site

$74.50 - $98/hr

Architect enterprise-scale Generative AI solutions leveraging LLMs, embeddings, and retrieval-augmented generation (RAG) pipelines. * Design and implement scalable AI microservices and APIs ...

Senior Agentic AI Builder

San Jose, CA · On-site

$107K - $136K/yr

... RAG - Practical experience implementing retrieval-augmented generation pipelines for context-aware AI applications • MCP (Model Context Protocol) - Practical experience with MCP client/server ...

Showing results 21-40

Retrieval Augmented Generation Rag information

See Pleasanton, CA salary details

$16

$22

$29

How much do retrieval augmented generation rag jobs pay per hour?

As of Sep 5, 2026, the average hourly pay for retrieval augmented generation rag in Pleasanton, CA is $22.54, according to ZipRecruiter salary data. Most workers in this role earn between $19.28 and $23.56 per hour, depending on experience, location, and employer.

What are popular job titles related to Retrieval Augmented Generation Rag jobs in Pleasanton, CA?

For Retrieval Augmented Generation Rag jobs in Pleasanton, CA, the most frequently searched job titles are:

What job categories do people searching Retrieval Augmented Generation Rag jobs in Pleasanton, CA look for?

The top searched job categories for Retrieval Augmented Generation Rag jobs in Pleasanton, CA are:

What cities near Pleasanton, CA are hiring for Retrieval Augmented Generation Rag jobs?

Cities near Pleasanton, CA with the most Retrieval Augmented Generation Rag job openings:

Infographic showing various Retrieval Augmented Generation Rag job openings in Pleasanton, CA as of August 2026, with employment types broken down into 63% Full Time, 35% Part Time, and 2% Contract. Highlights an 67% Physical, 2% Hybrid, and 31% Remote job distribution, with an average salary of $46,875 per year, or $22.5 per hour.

Software Engineer(AI Agent Platform)_SanJose, CA Onsite_Face to Face interview Must(only local candi

Xoriant Corporation

San Jose, CA • On-site

$90/hr

Other

Posted 25 days ago


Key responsibilities

  • Build and enhance an internal AI agent platform that enables Cisco employees to interact with enterprise knowledge through conversational interfaces.

  • Design and develop multi-agent workflows powered by Large Language Models (LLMs) to automate knowledge retrieval, content generation, and competitive analysis.

  • Architect resilient data ingestion pipelines that synchronize and index unstructured content from sources like SharePoint, Confluence, and internal wikis.


Job description

//***NOTE- NEED ONLY LOCAL CANDIDATE TO SAN JOSE, CA****///////////

RoleSoftware Engineer (AI Agent Platform)

Location - San Jose, CA (Onsite)

Client - Confidential

Rate - $90

Interview Mode - Face to Face

Duration - 12+ Months (Possibility of extension)

Build and enhance an internal AI agent platform that empowers Cisco employees to interact with enterprise knowledge through conversational interfaces (Webex, web). Design and develop multi-agent workflows powered by Large Language Models (LLMs) to automate knowledge retrieval, content generation, and competitive analysis at scale. Architect resilient data ingestion pipelines that synchronize and index unstructured content from sources like SharePoint, Confluence, and internal wikis through Microsoft Graph and other enterprise APIs. Collaborate with product managers, AI/ML engineers, and InfoSec partners to deliver secure, observable, and high-quality solutions in an agile environment. Contribute to architecture, retrieval-augmented generation (RAG) design, prompt engineering, and CI/CD practices that support rapid, reliable delivery of AI features. Succeed in this role by shipping maintainable, production grade agentic systems that measurably improve employee productivity and advance Cisco's AI enabled platform strategy.

Minimum Qualifications

  • Bachelor's degree with 5+ years of related experience, or Master's degree with 2+ years of related experience, or PhD with 0+ years of related experience.
  • Strong proficiency in Python 3.10+ with experience building production services using FastAPI, Pytest, and asynchronous programming patterns (asyncio, httpx).
  • Hands-on experience with LLM orchestration frameworks such as LangGraph, LangChain, LlamaIndex, or equivalent agentic frameworks.
  • Experience designing and querying relational databases, preferably PostgreSQL (including JSONB, full-text search, and pgvector or similar vector extensions).
  • Working knowledge of JavaScript / TypeScript and at least one modern frontend framework (Angular, React, or Vue) for building chat or admin UI surfaces.
  • Experience with microservices, containerization (Docker, Kubernetes), and at least one major cloud platform (AWS, Azure, or Google Cloud Platform).
  • Experience integrating AI/LLM features into production applications, including prompt engineering, retrieval-augmented generation (RAG), embeddings, and tool/function calling.
  • Familiarity with RESTful API design, OAuth2 / OIDC authentication flows, and secure API integration patterns.

Preferred Qualifications

  • Experience with multi agent orchestration patterns (planner/worker/critic, supervisor graphs, state machines) using LangGraph or similar.
  • Experience integrating with enterprise SaaS APIs such as Microsoft Graph, SharePoint, Webex, Salesforce, or ServiceNow.
  • Experience with vector databases and embedding stores (pgvector, Pinecone, Weaviate, Milvus, FAISS, OpenSearch k-NN).
  • Experience with caching and event driven systems such as Redis, Apache Kafka, or RabbitMQ for real-time data pipelines.
  • Experience with chatbot / conversational interface development (Webex Bot SDK, Slack Bolt, Microsoft Bot Framework, or Adaptive Cards).
  • Experience with DevOps tools including Jenkins, GitHub Actions, ArgoCD, or Terraform.
  • Experience with secrets management (Hashi Corp Vault, AWS Secrets Manager) and object storage (MinIO, Amazon S3, Azure Blob).
  • Experience with document processing pipelines, PDF/DOCX/PPTX extraction (PyMuPDF, python-docx, python-pptx), OCR, or computer vision libraries (OpenCV, Pillow) for image filtering.
  • Familiarity with LLM observability and evaluation tooling (LangSmith, Langfuse, Arize, Helicone, or RAGAS).
  • Strong problem solving, written communication, and stakeholder facing delivery skills, especially in cross functional environments spanning Engineering, IT, and Security.