1

Rag Engineer Jobs in Edison, NJ (NOW HIRING)

Senior RAG Engineer

New York, NY · On-site

$134K - $176K/yr

... RAG systems real users depend on. Backend depth without production retrieval won't be enough for this role, but our Senior Backend Engineer role might be the better fit. * You've diagnosed a ...

Senior RAG Engineer

New York, NY · On-site

$134K - $176K/yr

... RAG systems real users depend on. Backend depth without production retrieval won't be enough for this role, but our Senior Backend Engineer role might be the better fit. * You've diagnosed a ...

Senior RAG Engineer

New York, NY · Remote

$134K - $176K/yr

... RAG systems real users depend on. Backend depth without production retrieval won't be enough for this role, but our Senior Backend Engineer role might be the better fit. * You've diagnosed a ...

Build scalable Agentic AI systems, autonomous AI workflows, and multi-modal RAG solutions ... Mentor engineering teams while establishing AI development best practices and monitoring model ...

Sr. GenAI Engineer

Jersey City, NJ · On-site

$108K - $149K/yr

Job Title: Sr. GenAI Engineer Location: Jersey City, NJ (5 Days Onsite) Job Type: 12+ Month ... Implement multi-modal RAG systems and Agentic AI architectures for enterprise solutions * Fine-tune ...

AI Engineer

New York, NY · On-site

$175K - $250K/yr

Strong understanding of LLMs, RAG, prompt engineering, conversational AI platforms, and modern AI ... powered user experiences * Ability to design, optimize, and evolve large-scale conversational AI ...

Responsibilities : • Proficiency in technologies like Agentic AI, Gen AI, RAG, Python, Lang Graph ... with GenAI engineers, application teams, MLOps, product, and business stakeholders to deliver ...

Senior Agentic AI Engineer

Iselin, NJ · On-site

$106K - $145K/yr

Develop reusable agent frameworks, SDKs, evaluation pipelines, RAG, and memory systems * Own ... engineering experience * Hands-on experience building and deploying LLM agents or tool-using AI ...

next page

Showing results 1-20

Rag Engineer information

See Edison, NJ salary details

$61.6K

$93.7K

$158.9K

How much do rag engineer jobs pay per year?

As of Aug 13, 2026, the average yearly pay for rag engineer in Edison, NJ is $93,702.00, according to ZipRecruiter salary data. Most workers in this role earn between $70,900.00 and $108,700.00 per year, depending on experience, location, and employer.

What is the difference between Rag Engineer vs Textile Technician?

AspectRag EngineerTextile Technician
Required CredentialsEngineering degree, technical certificationsDiploma or degree in textiles or related field
Work EnvironmentFactories, manufacturing plants, R&D labsTextile mills, production facilities, quality control labs
Industry UsageDesigning and improving rag production processesMonitoring textile quality, testing fabrics

While both roles involve working within the textile industry, a Rag Engineer primarily focuses on the engineering aspects of rag production, process optimization, and machinery, whereas a Textile Technician concentrates on fabric testing, quality control, and ensuring textile standards are met. The roles often overlap in industry settings but differ in technical focus and responsibilities.

What are popular job titles related to Rag Engineer jobs in Edison, NJ?

For Rag Engineer jobs in Edison, NJ, the most frequently searched job titles are:

What job categories do people searching Rag Engineer jobs in Edison, NJ look for?

The top searched job categories for Rag Engineer jobs in Edison, NJ are:

What cities near Edison, NJ are hiring for Rag Engineer jobs?

Cities near Edison, NJ with the most Rag Engineer job openings:

Infographic showing various Rag Engineer job openings in Edison, NJ as of August 2026, with employment types broken down into 91% Full Time, 4% Part Time, and 5% Contract. Highlights an 85% Physical, 6% Hybrid, and 9% Remote job distribution, with an average salary of $93,702 per year, or $45 per hour.

Senior RAG Engineer

Newcode.ai

New York, NY • On-site

$134K - $176K/yr

Full-time

Posted 14 days ago


Job description

About Newcode.ai
Newcode.ai is a fast-growing legal tech and agentic AI company transforming how legal work is done. With teams across Norway, Sweden, Ireland, the US and growing we work at the intersection of law, technology, and intelligence. We move fast, think big, and take pride in doing things the right way.
Note: We believe in being transparent about what it's like to work at Newcode. As a fast-growing startup, we're building and evolving every day. That means not every process, playbook, or framework is already in place, and priorities can shift quickly.
The people who thrive here are comfortable with ambiguity, take ownership, and don't wait for perfect direction. They are resourceful, proactive, and able to "figure it out"-solving problems, creating structure where needed, and helping build the company as they go. If you are good with this then, great! Keep reading to learn more.
The role
You'll own retrieval. When someone asks our product a question, something has to find the right passage across large sets of confidential documents, decide it really is the right passage, and hand it to a language model whose answer ends up in real professional advice. That whole path is yours: how documents get parsed, chunked and embedded, how search combines meaning with exact terms, how results get reranked, and how an agent decides what to read next and when it has read enough. This is where AI products fail without anyone noticing. The answer reads well, the citation is wrong, and the client is the one who finds out.
It isn't CRUD work, and it isn't a demo notebook either. Retrieval is where AI products quietly fail: the answer reads well, the citation is wrong, and nobody notices until a client does. Your job is to stop that happening, and to be able to show with numbers that it isn't happening. You'll make architecture calls early and live with them.
The role is fully remote. Work from anywhere in the EU.
Requirements
What You'll Do
  • Own the retrieval pipeline end to end: document parsing and chunking, embeddings, indexing, and keeping all of it current as clients add and change material. Exposed through FastAPI, with ingestion and reindexing running as background jobs.
  • Design how we search. Combining semantic and keyword retrieval in Qdrant, fusing ranked lists, filtering on metadata, and reranking so the handful of results we pass to the model are the right ones.
  • Build agentic retrieval: query decomposition, the tools a model uses to search and navigate documents, multi-step loops that know when to stop, and the cost and latency budgets that keep them honest.
  • Build the evaluation layer that tells us any of this is working: golden sets, retrieval metrics, regression tests on realistic client data, and tracing good enough that a bad answer leads you back to the chunk that caused it.
  • Ship it as fast, dependable services in Python and FastAPI, with the heavy work - ingestion, embedding, reindexing - running as background jobs that hold up under load.
  • Take data security and isolation seriously. Clients hand us privileged material, and tenants stay strictly separated.
  • Make retrieval hold up across our clients' languages as well as it does in English. Compounding and inflection break sparse retrieval in ways an English eval set never shows you.

Who You Are
  • At least five years building backend systems that run in production, and at least two of them shipping retrieval or RAG systems real users depend on. Backend depth without production retrieval won't be enough for this role, but our Senior Backend Engineer role might be the better fit.
  • You've diagnosed a retrieval regression in production and fixed it. You can tell us what broke, how you found it, and what the numbers were before and after. - You've built and maintained a golden set. How many queries, who labelled them, which metrics you trust, and a change you shipped or killed because of what they told you.
  • You've run a vector index in production: picked index parameters, dealt with the memory and latency trade-offs, and reindexed without taking search down.
  • Strong Python. Comfortable reasoning about FastAPI or a close equivalent, PostgreSQL and Redis under load.
  • You don't ship a retrieval change because the output looked better on the three queries you tried by hand.
  • You own things end to end - including the boring maintenance and the tech debt nobody assigned you - rather than building the interesting part and handing off the rest.
  • You track what's moving in vector databases and LLMs because you're curious, not because it's the job. Show us the side project, the benchmark you ran for fun, or the repo where you tried something before it showed up in everyone else's stack.
  • You've worked with AI coding assistants and have a view on where they help and where they don't. If your current employer forbids them, that's not a mark against you. Tell us how you'd review AI-written code instead.
  • You're fine in a startup that changes direction. Decisions get made, then revisited.

Experience we expect with the stack:
  • Python: 5+ years
  • FastAPI: 2+ years
  • Vector databases in production (we run Qdrant): 2+ years. You've picked index parameters, dealt with the memory and latency trade-offs, and reindexed without taking search down.
  • Embedding models, and the metrics you use to judge retrieval quality
  • PostgreSQL / SQL
  • OCR-heavy document ingestion at scale; an information-retrieval background (BM25, NDCG, recall@k); search abstractions spanning more than one backend.

The hiring process
Stage 1: Online coding challenge
Stage 2: Screening Call
Stage 3: Technical home assessment
Stage 4: Whiteboard challenge and presentation of take home assessment
Stage 5: CEO interview
Visa Sponsorship:
At this time, Newcode is unable to provide visa sponsorship. Candidates must be authorized to work in the applicable country without employer sponsorship.
Benefits
Why Join Us?
  • Fully remote across the EEA, Norway included. We employ through our own entity or a local employer of record depending on where you are, so ask about your country and we'll tell you straight away whether we can do it.
  • You'll need the existing right to work where you live. We don't sponsor visas for this role.
  • Be a driver of innovation in one of the most exciting AI ventures
  • Work with talented peers in a collaborative, high-energy team
  • Shape both product and culture as we grow
  • Flexible, English-speaking environment