1

Retrieval Augmented Generation Jobs in Massachusetts

You'll work on distributed systems that combine traditional infrastructure automation with large language models (LLMs), retrieval-augmented generation (RAG), and intelligent agents. What You'll Do

Optimize LLM API interactions using prompt engineering, retrieval-augmented generation (RAG), contex management, and performance tuning. * Code Quality / Code Reviews: Write clean, maintainable, and ...

AI Engineer - Boston

Boston, MA · On-site

$140 - $175/hr

The role translates complex legal and operational needs into testableprototypes and proof-of-concept solutions--including generative AI applications,retrieval-augmented generation (RAG) solutions ...

New

AI Engineer

Boston, MA · On-site

$65 - $80/hr

Experience designing Retrieval-Augmented Generation (RAG) pipelines and working with vector databases. * Strong understanding of prompt engineering, embeddings, model evaluation, and AI application ...

Architect and deliver integrated AI solutions, including agentic workflows, retrieval-augmented generation pipelines, and enterprise platform integrations * Define and enforce governance, security ...

Showing results 21-40

Retrieval Augmented Generation information

What is a retrieval augmented generation?

A Retrieval Augmented Generation (RAG) job typically involves developing and optimizing AI systems that enhance text generation by incorporating external knowledge retrieved from relevant sources. Professionals in this field work on integrating retrieval mechanisms with large language models to improve the relevance, accuracy, and factual grounding of generated content. Common responsibilities include designing retrieval systems, fine-tuning language models, optimizing performance, and ensuring the seamless integration of factual data into AI-generated text. This role is highly interdisciplinary, involving expertise in natural language processing (NLP), machine learning, and information retrieval.

What does a retrieval augmented generation engineer do?

A Retrieval Augmented Generation engineer typically spends their day designing and implementing systems that combine information retrieval with advanced generative models, such as large language models. This includes fine-tuning models, integrating external data sources, developing vector search pipelines, and evaluating output quality. Collaboration with data scientists, machine learning engineers, and product teams is common to ensure the solutions meet user requirements and scale effectively. Additionally, RAG engineers often troubleshoot issues, monitor model performance in production, and stay informed about the latest advancements in AI and information retrieval.

What skills and qualifications are needed for retrieval augmented generation?

To thrive in a Retrieval Augmented Generation (RAG) engineering role, you need a solid background in machine learning, natural language processing (NLP), and experience with scalable information retrieval systems, typically supported by a relevant degree in computer science or a related field. Familiarity with tools such as Python, PyTorch or TensorFlow, vector databases, and search platforms like Elasticsearch is essential, along with practical experience deploying and tuning RAG pipelines. Strong problem-solving skills, a collaborative mindset, and effective communication abilities set outstanding professionals apart in this field. These competencies are crucial for designing, implementing, and optimizing hybrid retrieval-generation AI systems that address complex, real-world information needs.

What are the most commonly searched types of Retrieval Augmented Generation jobs in Massachusetts?

The most popular types of Retrieval Augmented Generation jobs in Massachusetts are:

What are popular job titles related to Retrieval Augmented Generation jobs in Massachusetts?

For Retrieval Augmented Generation jobs in Massachusetts, the most frequently searched job titles are:

What job categories do people searching Retrieval Augmented Generation jobs in Massachusetts look for?

The top searched job categories for Retrieval Augmented Generation jobs in Massachusetts are:

What cities in Massachusetts are hiring for Retrieval Augmented Generation jobs?

Cities in Massachusetts with the most Retrieval Augmented Generation job openings:

Infographic showing various Retrieval Augmented Generation job openings in Massachusetts as of August 2026, with employment types broken down into 100% Full Time. Highlights an 100% Remote job distribution.

Senior Software Engineer (AI / LLMs)

apiphani

Boston, MA • Remote

$120K - $160K/yr

Full-time

Re-posted 3 days ago


Job description

Apiphani is a technology-enabled managed services company dedicated to redefining what it means to support mission-critical enterprise workloads. We're a small but rapidly growing company, which means there's lots of room for growth and learning opportunities abound!

Apiphani is dedicated to creating a diverse and inclusive work environment for all as a fundamental component of our business. Diversity and inclusion are the bedrock of creativity and innovation. Without diversity of experience and thought, we would fail to progress as a company and as a team. Apiphani strives to foster an environment of belonging, where every employee feels respected, valued, and empowered. We embrace the unique experiences, perspective, and cultural background, which only you can bring to the table.

Senior Software Engineer — Agentic AI Platform

Location: Remote | Full-time | Competitive Compensation

Apiphani is building the future of intelligent infrastructure automation through agentic AI. We're looking for a Senior Backend Engineer to help design and build the systems that power Luumen — an AI-driven automation platform used by enterprise IT and managed service providers around the world.

This is a high-impact, zero-to-one engineering role focused on building the backend foundations for large-scale AI orchestration. You'll work on distributed systems that combine traditional infrastructure automation with large language models (LLMs), retrieval-augmented generation (RAG), and intelligent agents.

What You'll Do

  • Design and implement backend services that enable intelligent agent workflows and autonomous infrastructure actions
  • Develop APIs and orchestration layers in Python and TypeScript, integrating LLMs, vector databases, and observability pipelines
  • Build scalable systems to support LLM-based reasoning, retrieval, and decision-making across cloud infrastructure
  • Integrate with AWS Bedrock and other LLM platforms to support multi-model capabilities
  • Develop data access and semantic search layers using vector databases (e.g., pgvector, Pinecone, Qdrant)
  • Build robust monitoring, testing, and CI/CD systems to ensure reliability and reproducibility of AI workflows
  • Collaborate closely with the product and DevOps teams to design architecture diagrams, plan deployments, and monitor system health
  • Write clean, maintainable code with clear documentation and strong adherence to security and performance best practices
  • Participate in code reviews, design discussions, and iterative delivery cycles to improve product velocity and quality

What We're Looking For

  • 6+ years of backend engineering experience in production environments
  • Strong proficiency in Python and TypeScript for building distributed, event-driven systems
  • Deep understanding of AWS services (Lambda, ECS, Bedrock, S3, CloudWatch, etc.)
  • Experience designing APIs, microservices, and event pipelines that interface with LLMs or AI models
  • Familiarity with vector databases and concepts like embeddings, similarity search, and retrieval-augmented generation
  • Experience with infrastructure-as-code tools such as Terraform or AWS CDK
  • Understanding of SQL and schema migration workflows (PostgreSQL or similar)
  • Hands-on experience with Docker, GitHub Actions, and cloud-native CI/CD workflows
  • Ability to diagram systems, communicate architecture decisions clearly, and work asynchronously in a distributed team
  • Strong sense of ownership and ability to deliver in fast-moving, ambiguous environments

Bonus Points

  • Experience working with LangChain, OpenAI, or Anthropic APIs
  • Familiarity with agentic frameworks or AI orchestration systems
  • Background in observability or APM tooling (e.g., Datadog, Dynatrace)
  • Prior experience building automation or infrastructure management tools
  • Contributions to open-source LLM or MLOps projects
  • Interest in shaping how AI is applied to real-world IT operations

Why Join Apiphani

You'll be joining a globally distributed, high-performing team focused on redefining how enterprises manage infrastructure. Every feature you build will directly impact how engineers interact with intelligent systems in production environments.

This is an opportunity to help architect the foundations of a platform that blends infrastructure automation, AI, and agentic reasoning — where your technical decisions will shape the next generation of enterprise operations.

Base Salary
$120,000—$160,000 USD
Company Benefits*
  • Medical/dental/vision - 100% paid for employees, 50% paid for dependents
  • Life and disability - 100% paid for employees
  • 401K - 3% contribution, no employee contribution necessary
  • Education and tuition reimbursement
  • Accident, critical illness, hospital indemnity benefits offered through our providers
  • Employee Assistance Program
  • Legal assistance
  • Paid Time Off - up to 6 weeks per year
  • Sick Leave - up to 2 weeks per year
  • Parental Leave - up to 12 weeks

*Benefits listed in the job description apply to employees working in the United States. For international employees, Apiphani partners with an Employer of Record, Deel, and provides all statutory benefits required under local law; certain U.S.-specific programs (such as EAP, legal assistance, etc.) may not be available outside the United States. The specific benefits package will be outlined in the local employment agreement issued through Deel.