Experience building or integrating retrieval-augmented generation (RAG) systems * Experience working with enterprise security and compliance frameworks (i.e., SOC 2, GDPR, etc.) * Familiarity with ...
Experience building or integrating retrieval-augmented generation (RAG) systems * Experience working with enterprise security and compliance frameworks (i.e., SOC 2, GDPR, etc.) * Familiarity with ...
Full Stack Developer (Mobile) || Burlingame / Menlo Park CA & Seattle, WA (Onsite) || Fulltime
Burlingame, CA · On-site
LLM APIs, tool use, vector DBs, memory, retrieval-augmented generation (RAG), and evaluation loops * - Design secure, scalable backend services in JavaScript/TypeScript or Python to orchestrate ...
Quick apply
Full Stack Developer (Mobile) || Burlingame / Menlo Park CA & Seattle, WA (Onsite) || Fulltime
Burlingame, CA · On-site
LLM APIs, tool use, vector DBs, memory, retrieval-augmented generation (RAG), and evaluation loops * - Design secure, scalable backend services in JavaScript/TypeScript or Python to orchestrate ...
FullStack Developer (UX Design) || Burlingame / Menlo Park, CA / Seattle, WA (Onsite) || Fulltime
Burlingame, CA · On-site
LLM APIs, tool use, vector DBs, memory, retrieval-augmented generation (RAG), and evaluation loops * Design secure, scalable backend services in JavaScript/TypeScript or Python to orchestrate agentic ...
Quick apply
FullStack Developer (UX Design) || Burlingame / Menlo Park, CA / Seattle, WA (Onsite) || Fulltime
Burlingame, CA · On-site
LLM APIs, tool use, vector DBs, memory, retrieval-augmented generation (RAG), and evaluation loops * Design secure, scalable backend services in JavaScript/TypeScript or Python to orchestrate agentic ...
AI Engineer, Multimodal LLMs
San Francisco, CA · On-site
$120 - $150/hr
Experience with prompt engineering, parameter-efficient fine-tuning (PEFT), retrieval-augmented generation (RAG), reinforcement learning for LLMs. * Published AI research in top tier AI conferences ...
AI Engineer, Multimodal LLMs
San Francisco, CA · On-site
$120 - $150/hr
Experience with prompt engineering, parameter-efficient fine-tuning (PEFT), retrieval-augmented generation (RAG), reinforcement learning for LLMs. * Published AI research in top tier AI conferences ...
Experience building or integrating retrieval-augmented generation (RAG) systems * Experience working with enterprise security and compliance frameworks (i.e., SOC 2, GDPR, etc.) * Familiarity with ...
Quick apply
Experience building or integrating retrieval-augmented generation (RAG) systems * Experience working with enterprise security and compliance frameworks (i.e., SOC 2, GDPR, etc.) * Familiarity with ...
Backend Developer ___ Burlingame, CA / Menlo Park, CA / Seattle, WA (Onsite) ___ Fulltime FTE
Burlingame, CA · On-site
LLM APIs, tool use, vector DBs, memory, retrieval-augmented generation (RAG), and evaluation loops * Design secure, scalable backend services in JavaScript/TypeScript or Python to orchestrate agentic ...
Quick apply
Backend Developer ___ Burlingame, CA / Menlo Park, CA / Seattle, WA (Onsite) ___ Fulltime FTE
Burlingame, CA · On-site
LLM APIs, tool use, vector DBs, memory, retrieval-augmented generation (RAG), and evaluation loops * Design secure, scalable backend services in JavaScript/TypeScript or Python to orchestrate agentic ...
AI Engineer
Menlo Park, CA · On-site
$100 - $150/hr
Design and implement RAG (Retrieval-Augmented Generation) pipelines for contextual tour recommendations * Distill large models into efficient, deployable versions for production use * Quantize models ...
AI Engineer
Menlo Park, CA · On-site
$100 - $150/hr
Design and implement RAG (Retrieval-Augmented Generation) pipelines for contextual tour recommendations * Distill large models into efficient, deployable versions for production use * Quantize models ...
Senior AI Engineer
Oakland, CA · On-site
$150K - $170K/yr
... Retrieval Augmented Generation (RAG) • Build and optimize RAG pipelines across large unstructured datasets • Hands on experience with: o Embedding strategies o Vector databases and semantic ...
Senior AI Engineer
Oakland, CA · On-site
$150K - $170K/yr
... Retrieval Augmented Generation (RAG) • Build and optimize RAG pipelines across large unstructured datasets • Hands on experience with: o Embedding strategies o Vector databases and semantic ...
Partner with GenAI teams to explore multi-modal embeddings, avatar generation, and retrieval-augmented generation (RAG) in economic surfaces Lead efforts to optimize ML system performance: low ...
Partner with GenAI teams to explore multi-modal embeddings, avatar generation, and retrieval-augmented generation (RAG) in economic surfaces Lead efforts to optimize ML system performance: low ...
Senior Applied Researcher
San Francisco, CA · On-site
$107K - $147K/yr
Retrieval-augmented generation (RAG) pipelines and vector-based semantic search systems Representation learning and semantic embeddings for clustering, categorization, and content understanding Model ...
Senior Applied Researcher
San Francisco, CA · On-site
$107K - $147K/yr
Retrieval-augmented generation (RAG) pipelines and vector-based semantic search systems Representation learning and semantic embeddings for clustering, categorization, and content understanding Model ...
Build robust retrieval-augmented generation (RAG) pipelines with vector databases, embedding pipelines, and optimized chunking strategies * Design advanced prompting strategies including chain-of ...
Quick apply
Build robust retrieval-augmented generation (RAG) pipelines with vector databases, embedding pipelines, and optimized chunking strategies * Design advanced prompting strategies including chain-of ...
Gen AI Architect
Pleasanton, CA · Remote
Experience with vector databases (e.g., Pinecone, Weaviate, FAISS) and retrieval-augmented generation (RAG) systems. * Familiarity with responsible AI frameworks and privacy-preserving techniques.
Quick apply
Gen AI Architect
Pleasanton, CA · Remote
Experience with vector databases (e.g., Pinecone, Weaviate, FAISS) and retrieval-augmented generation (RAG) systems. * Familiarity with responsible AI frameworks and privacy-preserving techniques.
AI Engineer
San Francisco, CA · On-site
Experience building or integrating retrieval-augmented generation (RAG) systems * Experience working with enterprise security and compliance frameworks (e.g., SOC 2) * Familiarity with vector ...
AI Engineer
San Francisco, CA · On-site
Experience building or integrating retrieval-augmented generation (RAG) systems * Experience working with enterprise security and compliance frameworks (e.g., SOC 2) * Familiarity with vector ...
AI Engineer
San Francisco, CA · On-site
Experience building or integrating retrieval-augmented generation (RAG) systems * Experience working with enterprise security and compliance frameworks (e.g., SOC 2) * Familiarity with vector ...
Quick apply
AI Engineer
San Francisco, CA · On-site
Experience building or integrating retrieval-augmented generation (RAG) systems * Experience working with enterprise security and compliance frameworks (e.g., SOC 2) * Familiarity with vector ...
Chief AI Architect
San Francisco, CA · On-site
Prompt engineering Retrieval-Augmented Generation (RAG) Tool augmentation / API integration Ensure solutions are: Scalable Secure Production-ready 3. AI Industrialization & Platforms Drive AI ...
Chief AI Architect
San Francisco, CA · On-site
Prompt engineering Retrieval-Augmented Generation (RAG) Tool augmentation / API integration Ensure solutions are: Scalable Secure Production-ready 3. AI Industrialization & Platforms Drive AI ...
Chief AI Architect
Alameda, CA · On-site
Prompt engineering Retrieval-Augmented Generation (RAG) Tool augmentation / API integration Ensure solutions are: Scalable Secure Production-ready 3. AI Industrialization & Platforms Drive AI ...
Chief AI Architect
Alameda, CA · On-site
Prompt engineering Retrieval-Augmented Generation (RAG) Tool augmentation / API integration Ensure solutions are: Scalable Secure Production-ready 3. AI Industrialization & Platforms Drive AI ...
Chief AI Architect
Hayward, CA · On-site
Prompt engineering Retrieval-Augmented Generation (RAG) Tool augmentation / API integration Ensure solutions are: Scalable Secure Production-ready 3. AI Industrialization & Platforms Drive AI ...
Chief AI Architect
Hayward, CA · On-site
Prompt engineering Retrieval-Augmented Generation (RAG) Tool augmentation / API integration Ensure solutions are: Scalable Secure Production-ready 3. AI Industrialization & Platforms Drive AI ...
Qdrant is an open-source vector search engine powering the next generation of AI applications, from semantic search and retrieval-augmented generation (RAG) to AI agents and real-time recommendations.
Quick apply
Qdrant is an open-source vector search engine powering the next generation of AI applications, from semantic search and retrieval-augmented generation (RAG) to AI agents and real-time recommendations.
Chief AI Architect
San Mateo, CA · On-site
Prompt engineering Retrieval-Augmented Generation (RAG) Tool augmentation / API integration Ensure solutions are: Scalable Secure Production-ready 3. AI Industrialization & Platforms Drive AI ...
Chief AI Architect
San Mateo, CA · On-site
Prompt engineering Retrieval-Augmented Generation (RAG) Tool augmentation / API integration Ensure solutions are: Scalable Secure Production-ready 3. AI Industrialization & Platforms Drive AI ...
Qdrant is an open-source vector search engine powering the next generation of AI applications, from semantic search and retrieval-augmented generation (RAG) to AI agents and real-time recommendations.
Quick apply
Qdrant is an open-source vector search engine powering the next generation of AI applications, from semantic search and retrieval-augmented generation (RAG) to AI agents and real-time recommendations.
Retrieval Augmented Generation Rag information
See Berkeley, CA salary details
$18.54 - $19.83
6% of jobs
$21.01 is the 25th percentile. Wages below this are outliers.
$19.83 - $21.11
20% of jobs
$21.11 - $22.40
13% of jobs
The median wage is $23.31 / hr.
$22.40 - $23.68
15% of jobs
$23.68 - $24.97
13% of jobs
$25.65 is the 75th percentile. Wages above this are outliers.
$24.97 - $26.25
15% of jobs
$26.25 - $27.53
3% of jobs
$27.53 - $28.82
2% of jobs
$28.82 - $30.10
4% of jobs
$30.10 - $31.39
7% of jobs
$31.39 - $32.67
1% of jobs
$18
$24
$32
How much do retrieval augmented generation rag jobs pay per hour?
What are popular job titles related to Retrieval Augmented Generation Rag jobs in Berkeley, CA?
For Retrieval Augmented Generation Rag jobs in Berkeley, CA, the most frequently searched job titles are:
What job categories do people searching Retrieval Augmented Generation Rag jobs in Berkeley, CA look for?
The top searched job categories for Retrieval Augmented Generation Rag jobs in Berkeley, CA are:
What cities near Berkeley, CA are hiring for Retrieval Augmented Generation Rag jobs?
Cities near Berkeley, CA with the most Retrieval Augmented Generation Rag job openings:
Full-time
Medical, Dental, Vision, Retirement, PTO
Re-posted 21 hours ago
Key responsibilities
Design and implement agentic systems capable of interacting with browsers, operating systems, and enterprise filesystems.
Build search and retrieval pipelines for large-scale structured and unstructured data, and develop backend layers for RAG systems and information management.
Integrate language models, vision models, reinforcement learning, and scaffolding frameworks to enable autonomous, multi-step decision-making.
Job description
The Role:
As a Full-Stack Software Engineer, you will be a core contributor to Zyphra's Agentic Systems and Interaction projects. You will be at the forefront of building a next-generation desktop and browser-based agent (end-to-end) that can autonomously navigate the web, interact with filesystems, and complete complex user tasks. This role spans agentic orchestration, frontend interfaces, secure sandboxing environments, large-scale document search and retrieval, and language/vision model integration.
You'll Work Across:
- Design and implementation of different agentic system designs capable of interacting with browsers, operating systems, enterprise filesystems, collaboration tools, etc.
- Building search and retrieval pipelines across large-scale structured and unstructured data
- Build the backend layer for RAG systems and other information context management systems required for agents to operate
- Integrating LLMs, vision models, reinforcement learning, and scaffolding frameworks for autonomous, multi-step decision-making
- What matters most is your drive to build production-grade software
- We value velocity and curiosity, especially in fast-moving and ambiguous environments
What We're Looking For / Requirements:
- Proficiency in Python and a deep understanding of building and debugging complex end-to-end applications
- Experience working with SaaS environment and production workloads, web interfaces, databases, search engines, and queuing systems
- Experience developing browser extensions or automation tools with fine-grained control over the browser (mouse, tabs, DOM)
- Understanding of LLMs, prompting techniques, and orchestration frameworks for multi-step reasoning
- Ability to work across the whole stack from web interfaces to control plane and data infrastructure like queuing systems, databases, and search indexes
- Experience designing or working with secure and virtualized execution environments
- Excellent communication and collaboration skills across product, research, and engineering teams
Qualifications / Additional Skills:
- Experience building or integrating retrieval-augmented generation (RAG) systems
- Experience working with enterprise security and compliance frameworks (i.e., SOC 2, GDPR, etc.)
- Familiarity with embeddings, vector databases and large-scale document indexing
- Knowledge of web automation tools and headless browser environments (i.e., Puppeteer, Playwright)
- Understanding of sandboxed or containerized compute environments with strict access controls
- Comfort designing user-facing agentic workflows and reasoning systems that span multiple modalities (text, vision, actions)
- Experience using and fine-tuning models for screen reading, OCR, or UI understanding
- Background in HCI or interest in building intuitive agent interfaces that extend human capabilities
- Machine Learning experience as a double bonus
Why Work at Zyphra:
- Our research methodology is grounded in methodical, step-by-step approaches to ambitious goals. Both deep research and engineering excellence are equally valued
- We strongly value new and crazy ideas and are very willing to bet big on new ideas
- We move as quickly as we can; we aim to minimize the bar to impact as low as possible
- We all enjoy what we do and love discussing AI
Benefits and Perks:
- Comprehensive medical, dental, vision, and FSA plans
- Competitive compensation and 401(k) plan
- Relocation and immigration support on a case-by-case basis
- In-office snacks and meals provided
- Unlimited PTO and company holidays
- In-person team in San Francisco with a collaborative, high-energy environment
About Zyphra Technologies
Sourced by ZipRecruiter
Industry
Computer and peripheral equipment manufacturing
Company size
1 - 10 Employees
Headquarters location
San Francisco, CA, US
Year founded
2020