1

Llm Prompt Review Jobs (NOW HIRING)

... review, testing and documentation • Continuous learning in Generative AI-related fields ... LLM, Prompt Engineering, RAG Architecture, Agentic AI. • Basic knowledge of Hybrid prompting ...

... review, testing and documentation • Continuous learning in Generative AI-related fields ... LLM, Prompt Engineering, RAG Architecture, Agentic AI. • Basic knowledge of Hybrid prompting ...

... review, testing and documentation • Continuous learning in Generative AI-related fields ... LLM, Prompt Engineering, RAG Architecture, Agentic AI. • Basic knowledge of Hybrid prompting ...

Prompt Engineer

Jersey City, NJ · On-site

$90 - $120/hr

Required Qualifications * 4+ Years of Strong understanding of LLM behavior, prompt design ... Familiarity with prompt registries, A/B testing, human review workflows, and evaluation tooling. #J ...

... review workflows, and evaluation tooling. Initial screening questions * Show how you improved a weak LLM output through prompt design and testing. * How do you evaluate prompt performance beyond ...

Prompt Engineer

Jersey City, NJ · On-site

$104K - $182K/yr

Prompt Engineer LLM Interaction Design / Prompt Optimization / GenAI Application Quality Role ... human review workflows, and evaluation tooling. NTT DATA provides a reasonable range of ...

Prompt Engineer LLM Interaction Design / Prompt Optimization / GenAI Application Quality Role ... human review workflows, and evaluation tooling. NTT DATA provides a reasonable range of ...

Prompt Engineer LLM Interaction Design / Prompt Optimization / GenAI Application Quality Role ... human review workflows, and evaluation tooling. NTT DATA provides a reasonable range of ...

Senior WEM AQM Prompt Engineer

San Ramon, CA · On-site

$116K - $160K/yr

... reviews and providing structured feedback. * Drive continuous improvement of the team's prompt design process based on performance data, customer outcomes, and LLM platform developments. Key ...

Senior WEM AQM Prompt Engineer

San Francisco, CA · On-site

$123K - $169K/yr

... reviews and providing structured feedback. * Drive continuous improvement of the team's prompt design process based on performance data, customer outcomes, and LLM platform developments. Key ...

Prompt patterns, system instructions, response templates, and conversation policies for AIRP LLM ... human review workflows, and evaluation tooling. #LI-NorthAmerica NTT DATA provides a reasonable ...

Senior WEM AQM Prompt Engineer

$107K - $146K/yr

... reviews and providing structured feedback. • Drive continuous improvement of the team's prompt design process based on performance data, customer outcomes, and LLM platform developments. Key ...

next page

Showing results 1-20

Llm Prompt Review information

See salary details

$37

$60

$77

How much do llm prompt review jobs pay per hour?

As of Aug 22, 2026, the average hourly pay for llm prompt review in the United States is $60.90, according to ZipRecruiter salary data. Most workers in this role earn between $54.09 and $69.71 per hour, depending on experience, location, and employer.

What is an LLM prompt reviewer?

An LLM Prompt Reviewer is a professional responsible for evaluating, refining, and optimizing prompts used with large language models (LLMs) like GPT-4. Their main goal is to ensure that prompts elicit accurate, useful, and safe responses from the AI. This role involves understanding both the technical and linguistic aspects of prompts, testing various phrasings, and documenting best practices. LLM Prompt Reviewers often collaborate with data scientists, AI trainers, and product teams to improve prompt quality and user experience.

What are the key skills and qualifications needed to thrive as an LLM prompt reviewer, and why are they important?

To thrive as an LLM Prompt Reviewer, you need a strong background in linguistics, critical thinking, and AI language model behavior, often supported by experience in content moderation or NLP. Familiarity with prompt engineering tools, annotation platforms, and basic understanding of large language model systems is typically required. Attention to detail, analytical skills, and clear written communication make someone stand out in this position. These skills ensure the creation and evaluation of high-quality prompts that drive accurate, safe, and useful AI model outputs.

What are some common challenges faced by professionals in LLM prompt review roles, and how can they be managed?

Professionals in LLM Prompt Review roles often encounter challenges such as ensuring prompt clarity, mitigating bias, and maintaining consistency across large volumes of prompts. Balancing creativity with precision is essential, as even small changes can significantly impact model outputs. To manage these challenges, reviewers typically rely on established guidelines, peer collaboration, regular calibration sessions, and continuous feedback from model performance metrics. Staying updated on best practices and working closely with data scientists and prompt engineers also helps maintain high-quality outputs.

What is the difference between Llm Prompt Review vs Data Annotator?

AspectLlm Prompt ReviewData Annotator
CredentialsBasic understanding of AI and NLP conceptsTypically high school diploma or equivalent, sometimes specialized training
Work EnvironmentRemote or office-based, focused on AI projectsRemote or on-site, working with datasets and labeling tools
Industry UsageUsed in AI development, NLP, and machine learning projectsUsed across various industries for data preparation and labeling
Search & Comparison IntentUnderstanding roles related to AI prompt evaluationComparing data labeling and annotation roles

While both roles involve working with data and AI, Llm Prompt Review focuses on evaluating and refining AI prompts, whereas Data Annotator involves labeling data for machine learning models. The roles differ mainly in their specific tasks and required skills, but both are essential in AI development workflows.

More about Llm Prompt Review jobs

What cities are hiring for Llm Prompt Review jobs?

Cities with the most Llm Prompt Review job openings:

What states have the most Llm Prompt Review jobs?

States with the most job openings for Llm Prompt Review jobs include:

Infographic showing various Llm Prompt Review job openings in the United States as of August 2026, with employment types broken down into 1% As Needed, 82% Full Time, 13% Part Time, 3% Contract, and 1% Nights. Highlights an 89% Physical, 3% Hybrid, and 8% Remote job distribution, with an average salary of $126,666 per year, or $60.9 per hour.

Principal Engineer - Context Engineering & LLM Optimization

Bank of America

Charlotte, NC • On-site

$140 - $180/hr

Other

Re-posted 23 days ago


Bank Of America rating

8.2

Company rating: 8.2 out of 10

Based on 530 frontline employees who took The Breakroom Quiz

50th of 171 rated banks


Job description

The Context Engineering & LLM Optimization Principal Engineer is responsible for defining and leading the engineering approach for solutions at the program or portfolio level, to deliver significant business outcomes. Key responsibilities include continuously improving the design, quality, and reuse of the solution and delivering technology enablers that improve development efficiencies for the solution. Job expectations include familiarity with at least one area of engineering, acting as a “go to” reference across the organization, and applying knowledge to improve technical competencies through recruitment and development activities.

Developer Experience (DevEx) provides enterprise technical standards and common technical services, platforms, and tools that are leveraged by delivery teams across all lines of business. Within the SDLC Software Delivery Lifecycle program, this role leads portfolio product delivery strategy and execution for enterprise software delivery capabilities, ensuring the right investments, operating model, governance, and prioritization are in place to improve how internal technical users build, test, and deliver software at scale.

The Context Engineering & LLM Optimization Principal Engineer is responsible for designing how information is selected, organized, compressed, prioritized, and presented to LLMs. This role focuses on context window management, prompt architecture, retrieval orchestration, grounding strategies, instruction design, tool‑use patterns, and evaluation of LLM behavior.

This engineer ensures that LLM applications receive the right context, in the right format, at the right time, with minimal token waste and maximum answer quality.

Responsibilities
  • Develops the engineering approach for the entire program/portfolio solution and works with Architecture, to develop/analyze/deliver the implementation of technical enablers
  • Leads the planning, definition, and design of the complex features which span multiple teams and explore solution alternatives
  • Creates ideas on designing complex technology and solution development approaches
  • Leads the technical oversight for teams in solution development including design reviews and code within own domain
  • Defines the technology tool stack for the solution within ranged of internally approved and supported technologies
  • Explores state-of-the-art technologies to improve development efficiencies, quality of test/QA coverage, and release management
  • Leads and is responsible for the end-to-end test strategy/creation/adherence, and the integration between teams for a program/portfolio solution
  • Design context engineering strategies for enterprise LLM and RAG applications.
  • Define prompt architectures for system prompts, developer instructions, user prompts, retrieved context, tool outputs, conversation history, and structured constraints.
  • Optimize context window usage through summarization, compression, ranking, filtering, deduplication, and context prioritization.
  • Design retrieval orchestration patterns that determine what data is retrieved, when it is retrieved, and how it is injected into the LLM prompt.
  • Partner with RAG database engineers to tune retrieval outputs for downstream reasoning quality.
  • Partner with data ingestion engineers to improve source formatting, metadata, and chunk structures for better contextual use.
  • Develop patterns for multi-turn conversation memory, session state, user intent preservation, and context refresh.
  • Define strategies for grounding, citation handling, source attribution, conflicting evidence resolution, and hallucination reduction.
  • Improve the experience for our developers, making it easier to deliver industry-leading solutions, while managing work efficiently and with the right controls
  • Advance our technology platforms through innovation
  • Reduce risk and improve quality across our technology portfolio by aligning to a single enterprise architecture strategy and delivering governance that enables consistency, integration and automation
  • Design LLM evaluation frameworks for answer quality, factuality, instruction adherence, relevance, safety, and token efficiency.
  • Establish prompt engineering and context engineering standards across product and platform teams.
  • Evaluate LLM model behavior across different context sizes, retrieval strategies, and prompt structures.
  • Define reusable patterns for agents, tool calling, function calling, dynamic prompt generation, and workflow-based reasoning.
  • Lead technical reviews for LLM application design, prompt safety, and context efficiency.
  • Serve as a senior technical authority for enterprise AI platform engineering.
  • Own architecture decisions that impact multiple teams, systems, or domains.
  • Create reusable patterns, reference architectures, standards, and engineering guardrails.
  • Mentor senior engineers and influence technical direction without requiring direct reporting authority.
  • Balance innovation with operational reliability, security, compliance, scalability, and cost management.
  • Communicate complex AI and data engineering concepts clearly to engineering, product, risk, security, and executive stakeholders.
Required Qualifications
  • 10+ years of software engineering, data engineering, platform engineering, or AI engineering experience.
  • 5+ years designing large-scale enterprise systems.
  • 2+ years working with LLM, RAG, vector search, semantic search, or AI platform capabilities.
  • Experience operating systems in regulated, security-conscious, or enterprise-scale environments.
  • Extensive experience building or architecting production LLM, RAG, or AI assistant systems.
  • Deep understanding of how LLMs use prompts, retrieved context, conversation history, system instructions, and tool outputs.
  • Strong knowledge of context window management, token budgeting, prompt construction, grounding, and response evaluation.
  • Experience with OpenAI, Azure OpenAI, Anthropic, Google Gemini, Meta Llama, or similar LLM ecosystems.
  • Experience designing prompt templates, retrieval‑augmented prompts, agent workflows, and tool‑use orchestration.
  • Familiarity with vector search, embeddings, reranking, semantic retrieval, and document chunking.
  • Experience with automated LLM evaluation, prompt regression testing, and quality measurement.
  • Ability to define enterprise standards for reliable, explainable, and secure LLM behavior.
  • Proven ability to lead architecture across multiple engineering teams.
  • Strong written and verbal communication skills.
  • Bachelor’s degree in Computer Science, Engineering, Information Systems, Applied Mathematics, or a related technical field
Desired Qualifications
  • Experience with agentic workflows, multi-agent orchestration, function calling, or tool‑augmented reasoning.
  • Experience with prompt injection mitigation, jailbreak resistance, and secure context handling.
  • Experience with token optimization, long‑context models, summarization pipelines, and contextual compression.
  • Experience with user personalization, enterprise memory patterns, or domain‑specific copilots.
  • Higher‑quality LLM responses with better grounding and reduced hallucination.
  • Lower token usage and improved response latency through efficient context construction.
  • Standardized prompt and context patterns reused across teams.
  • Improved evaluation coverage for LLM behavior, factuality, and instruction adherence.
  • Better alignment between retrieved enterprise knowledge and generated responses.
  • Enterprise architecture
  • Distributed systems design
  • AI platform engineering
  • Data governance and security
  • Cloud‑native engineering
  • Observability and operational excellence
  • Technical strategy and roadmap development
  • Cross‑functional influence
  • Vendor and platform evaluation
  • Production support and continuous improvement
Skills
  • Automation
  • Influence
  • Result Orientation
  • Stakeholder Management
  • Technical Strategy Development
  • Application Development
  • Architecture
  • Business Acumen
  • Risk Management
  • Solution Design
  • Agile Practices
  • Analytical Thinking
  • Collaboration
  • Data Management
  • Solution Delivery Process

Shift: 1st shift (United States of America)

Hours Per Week: 40

#J-18808-Ljbffr

What Bank Of America employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom


Bank Of America logo

About Bank Of America

Sourced by ZipRecruiter

At Bank of America, we are guided by a common purpose to help make financial lives better through the power of every connection. Responsible Growth is how we run our company and how we deliver for our clients, teammates, communities and shareholders every day. One of the keys to driving Responsible Growth is being a great place to work for our teammates around the world. We're devoted to being a diverse and inclusive workplace for everyone. We hire individuals with a broad range of backgrounds and experiences and invest heavily in our teammates and their families by offering competitive benefits to support their physical, emotional, and financial well-being.

Industry

Finance and insurance

Company size

10,000+ Employees

Headquarters location

Charlotte, NC, US

Social media