... retrieval-augmented generation (RAG) systems. • Collaborate closely with product and engineering teams to turn early prototypes into production-grade systems. • Help shape the future of AI ...
... retrieval-augmented generation (RAG) systems. • Collaborate closely with product and engineering teams to turn early prototypes into production-grade systems. • Help shape the future of AI ...
Staff Software Developer 2
San Diego, CA · On-site
Support implementation of retrieval-augmented generation workflows, including document chunking, metadata handling, vector embeddings, semantic search, and citation-enabled response generation.
Staff Software Developer 2
San Diego, CA · On-site
Support implementation of retrieval-augmented generation workflows, including document chunking, metadata handling, vector embeddings, semantic search, and citation-enabled response generation.
Expertise in LLM including model architecture, training/finetuning techniques, retrieval augmented generation (RAG), reasoning and action planning, etc. * Experience in planning, tool use, agent AI ...
Expertise in LLM including model architecture, training/finetuning techniques, retrieval augmented generation (RAG), reasoning and action planning, etc. * Experience in planning, tool use, agent AI ...
Founding Engineer - FlowGen Labs
San Francisco, CA · On-site
$240K - $400K/yr
Experience orchestrating agentic workflows and retrieval augmented generation. Familiarity with LangGraph is a plus. * Stand up inference paths with low latency serving and token-level observability
Founding Engineer - FlowGen Labs
San Francisco, CA · On-site
$240K - $400K/yr
Experience orchestrating agentic workflows and retrieval augmented generation. Familiarity with LangGraph is a plus. * Stand up inference paths with low latency serving and token-level observability
Build and optimize retrieval-augmented generation (RAG) systems. * Collaborate closely with product and engineering teams to turn early prototypes into production-grade systems. * Help shape the ...
Build and optimize retrieval-augmented generation (RAG) systems. * Collaborate closely with product and engineering teams to turn early prototypes into production-grade systems. * Help shape the ...
Machine Learning Engineer II at Glidewell Dental Irvine, CA
Irvine, CA · On-site
$117 - $146/hr
Designs, implements, and optimizes retrieval-augmented generation (RAG) pipelines that combine large language models (LLMs) with vector search/retrieval systems. * Builds data ingestion and embedding ...
New
Machine Learning Engineer II at Glidewell Dental Irvine, CA
Irvine, CA · On-site
$117 - $146/hr
Designs, implements, and optimizes retrieval-augmented generation (RAG) pipelines that combine large language models (LLMs) with vector search/retrieval systems. * Builds data ingestion and embedding ...
New
Research Scientist - AI Agent Memory Infrastructure - Global Frontier Tech Recruitment Program - 202
San Jose, CA · On-site
... retrieval-augmented generation (RAG), context engineering, retrieval systems, and long-term state management. • Familiarity with one or more key areas in memory systems: memory extraction and ...
Research Scientist - AI Agent Memory Infrastructure - Global Frontier Tech Recruitment Program - 202
San Jose, CA · On-site
... retrieval-augmented generation (RAG), context engineering, retrieval systems, and long-term state management. • Familiarity with one or more key areas in memory systems: memory extraction and ...
LLM Applications Engineer (San Francisco)
San Francisco, CA · On-site
$130K - $175K/yr
Implement and optimize RAG (Retrieval-Augmented Generation) pipelines, vector databases, and prompt management systems. * Feature Development: Build intuitive, AI-driven UI components that allow ...
LLM Applications Engineer (San Francisco)
San Francisco, CA · On-site
$130K - $175K/yr
Implement and optimize RAG (Retrieval-Augmented Generation) pipelines, vector databases, and prompt management systems. * Feature Development: Build intuitive, AI-driven UI components that allow ...
Software Engineer (Early Career)
Los Altos, CA · On-site
$100K - $150K/yr
Build and optimize retrieval-augmented generation (RAG) systems. * Collaborate closely with product and engineering teams to turn early prototypes into production-grade systems. * Help shape the ...
Software Engineer (Early Career)
Los Altos, CA · On-site
$100K - $150K/yr
Build and optimize retrieval-augmented generation (RAG) systems. * Collaborate closely with product and engineering teams to turn early prototypes into production-grade systems. * Help shape the ...
Software Engineer (Early Career)
$100K - $150K/yr
Build and optimize retrieval-augmented generation (RAG) systems. * Collaborate closely with product and engineering teams to turn early prototypes into production-grade systems. * Help shape the ...
Software Engineer (Early Career)
$100K - $150K/yr
Build and optimize retrieval-augmented generation (RAG) systems. * Collaborate closely with product and engineering teams to turn early prototypes into production-grade systems. * Help shape the ...
Build and optimize RAG (Retrieval-Augmented Generation) pipelines using Vertex AI Search and BigQuery, ensuring agents have "zero-hallucination" access to enterprise truths. • Governance & Security:
Build and optimize RAG (Retrieval-Augmented Generation) pipelines using Vertex AI Search and BigQuery, ensuring agents have "zero-hallucination" access to enterprise truths. • Governance & Security:
Experience building RAG (Retrieval-Augmented Generation) systems * Familiarity with semantic search and vector databases * Experience with computer vision models, such as: * SAM (Segment Anything ...
Experience building RAG (Retrieval-Augmented Generation) systems * Familiarity with semantic search and vector databases * Experience with computer vision models, such as: * SAM (Segment Anything ...
Backend Developer ___ Burlingame, CA / Menlo Park, CA / Seattle, WA (Onsite) ___ Fulltime FTE
Burlingame, CA · On-site
LLM APIs, tool use, vector DBs, memory, retrieval-augmented generation (RAG), and evaluation loops * Design secure, scalable backend services in JavaScript/TypeScript or Python to orchestrate agentic ...
Quick apply
Backend Developer ___ Burlingame, CA / Menlo Park, CA / Seattle, WA (Onsite) ___ Fulltime FTE
Burlingame, CA · On-site
LLM APIs, tool use, vector DBs, memory, retrieval-augmented generation (RAG), and evaluation loops * Design secure, scalable backend services in JavaScript/TypeScript or Python to orchestrate agentic ...
AI Engineer
San Francisco, CA · On-site
Experience building or integrating retrieval-augmented generation (RAG) systems * Experience working with enterprise security and compliance frameworks (e.g., SOC 2) * Familiarity with vector ...
AI Engineer
San Francisco, CA · On-site
Experience building or integrating retrieval-augmented generation (RAG) systems * Experience working with enterprise security and compliance frameworks (e.g., SOC 2) * Familiarity with vector ...
AI Engineer
San Francisco, CA · On-site
Experience building or integrating retrieval-augmented generation (RAG) systems * Experience working with enterprise security and compliance frameworks (e.g., SOC 2) * Familiarity with vector ...
Quick apply
AI Engineer
San Francisco, CA · On-site
Experience building or integrating retrieval-augmented generation (RAG) systems * Experience working with enterprise security and compliance frameworks (e.g., SOC 2) * Familiarity with vector ...
Full Stack Developer (Mobile) || Burlingame / Menlo Park CA & Seattle, WA (Onsite) || Fulltime
Burlingame, CA · On-site
LLM APIs, tool use, vector DBs, memory, retrieval-augmented generation (RAG), and evaluation loops * - Design secure, scalable backend services in JavaScript/TypeScript or Python to orchestrate ...
Quick apply
Full Stack Developer (Mobile) || Burlingame / Menlo Park CA & Seattle, WA (Onsite) || Fulltime
Burlingame, CA · On-site
LLM APIs, tool use, vector DBs, memory, retrieval-augmented generation (RAG), and evaluation loops * - Design secure, scalable backend services in JavaScript/TypeScript or Python to orchestrate ...
FullStack Developer (UX Design) || Burlingame / Menlo Park, CA / Seattle, WA (Onsite) || Fulltime
Burlingame, CA · On-site
LLM APIs, tool use, vector DBs, memory, retrieval-augmented generation (RAG), and evaluation loops * Design secure, scalable backend services in JavaScript/TypeScript or Python to orchestrate agentic ...
Quick apply
FullStack Developer (UX Design) || Burlingame / Menlo Park, CA / Seattle, WA (Onsite) || Fulltime
Burlingame, CA · On-site
LLM APIs, tool use, vector DBs, memory, retrieval-augmented generation (RAG), and evaluation loops * Design secure, scalable backend services in JavaScript/TypeScript or Python to orchestrate agentic ...
AI Engineer
Newark, CA · On-site
$80 - $90/hr
Experience with vector databases (e.g., Pinecone, Weaviate, FAISS) for retrieval-augmented generation. * Familiarity with MLOps tools (MLflow, Kubeflow, or similar). * Knowledge of prompt engineering ...
AI Engineer
Newark, CA · On-site
$80 - $90/hr
Experience with vector databases (e.g., Pinecone, Weaviate, FAISS) for retrieval-augmented generation. * Familiarity with MLOps tools (MLflow, Kubeflow, or similar). * Knowledge of prompt engineering ...
Founding Engineer
San Francisco, CA · On-site
$120K - $190K/yr
Hands-on experience with LLMs and Retrieval-Augmented Generation (RAG) architectures - this is a dealbreaker requirement. * Strong CS fundamentals, high attention to detail, fast learner, and clear ...
Founding Engineer
San Francisco, CA · On-site
$120K - $190K/yr
Hands-on experience with LLMs and Retrieval-Augmented Generation (RAG) architectures - this is a dealbreaker requirement. * Strong CS fundamentals, high attention to detail, fast learner, and clear ...
Founding Engineer
San Francisco, CA · On-site
$120K - $190K/yr
Hands-on experience with LLMs and Retrieval-Augmented Generation (RAG) architectures -- this is a dealbreaker requirement. * Strong CS fundamentals, high attention to detail, fast learner, and clear ...
Quick apply
Founding Engineer
San Francisco, CA · On-site
$120K - $190K/yr
Hands-on experience with LLMs and Retrieval-Augmented Generation (RAG) architectures -- this is a dealbreaker requirement. * Strong CS fundamentals, high attention to detail, fast learner, and clear ...
Entrylevel Retrieval Augmented Generation information
Full-time
Re-posted 2 days ago
Job description
Uare.ai is an AI startup focused on empowering people with their memories through innovative technology. The role involves designing and implementing scalable machine learning pipelines and integrating advanced models to enhance human interaction with AI-powered products.
Responsibilities:
• Design and implement machine learning workflows that combine structured and unstructured data.
• Work with state-of-the-art language models and voice technologies.
• Build and optimize retrieval-augmented generation (RAG) systems.
• Collaborate closely with product and engineering teams to turn early prototypes into production-grade systems.
• Help shape the future of AI-powered human interaction products.
Qualifications:
Required:
• Experience in machine learning, NLP, or AI.
• Strong experience with LLMs, RAG pipelines, vector databases, and prompt engineering.
• Hands-on experience with cloud-based AI platforms (AWS, Azure, or GCP).
• Excited about building 0-to-1 products in a fast-moving environment.
Preferred:
• Familiarity with voice cloning, speech synthesis, or related audio ML workflows is a plus.
Company:
At Uare.ai, we’re building the first platform for Individual AI - designed to protect, reflect, and amplify the uniqueness of every person. Founded in 2024, the company is headquartered in Palo Alto, California, US, , with a team of 11-50 employees. The company is currently Early Stage.