1

Llm Prompt Evaluation Jobs in Decatur, GA (NOW HIRING)

Senior AI/ML Engineer

Atlanta, GA · On-site

$100K - $138K/yr

Perform prompt engineering and evaluation for structured data extraction from unstructured inputs ... Experience with LLM APIs (OpenAI, Anthropic) and prompt pipelines * Strong SQL skills and ...

GenAI Systems Developer

Atlanta, GA · Remote

$185K - $215K/yr

  • Medical

  • Retirement

  • PTO

Go well beyond "prompt engineering" by building end-to-end solutions that solve real-world client ... Evaluation Frameworks: Develop custom methods to test and validate LLM performance. * Interfaces ...

... of LLM evaluation methodologies, testing frameworks, and quality assuranceUnderstanding of reasoning models, embedding models, and prompt optimization Experience with document processing and unit ...

Build robust evaluation frameworks to measure LLM efficacy, manage ground truth dataset quality ... prompt engineering, inference optimization, and fine-tuning. • GenAI Solutions: 3+ years of ...

Senior AI Engineer

Atlanta, GA

$100K - $138K/yr

Build and tune prompt-engineering patterns at production scale -- system prompts, structured output, tool and function calling. * Design and maintain LLM evaluation harnesses -- golden sets ...

Senior Product Manager

Atlanta, GA · On-site

$150K - $200K/yr

Hands-on experience with modern generative AI technologies, including prompt engineering, retrieval-augmented generation (RAG), LLM APIs, and model evaluation methodologies. * Experience working ...

Senior Product Manager

Atlanta, GA · Remote

$150K - $200K/yr

Hands-on experience with modern generative AI technologies, including prompt engineering, retrieval-augmented generation (RAG), LLM APIs, and model evaluation methodologies. Experience working within ...

Senior Product Manager

Atlanta, GA · On-site +1

$121K - $160K/yr

Hands-on experience with modern generative AI technologies, including prompt engineering, retrieval-augmented generation (RAG), LLM APIs, and model evaluation methodologies. * Experience working ...

Showing results 41-60

Llm Prompt Evaluation information

See Decatur, GA salary details

$7

$20

$40

How much do llm prompt evaluation jobs pay per hour?

As of Aug 17, 2026, the average hourly pay for llm prompt evaluation in Decatur, GA is $20.39, according to ZipRecruiter salary data. Most workers in this role earn between $14.09 and $26.97 per hour, depending on experience, location, and employer.

What is an LLM prompt evaluation?

An LLM Prompt Evaluation job involves assessing and optimizing prompts used in large language models (LLMs) to ensure they generate accurate, relevant, and high-quality responses. Evaluators test different prompts, analyze model outputs, and refine phrasing to improve performance. This role requires a strong understanding of AI behavior, critical thinking, and sometimes domain-specific expertise to create effective instructions for the model.

What does an LLM prompt evaluation do?

Professionals in LLM Prompt Evaluation spend their days reviewing, analyzing, and scoring the outputs of large language models based on specific prompts. This involves identifying inaccuracies, biases, or other issues in AI-generated responses, and providing detailed, constructive feedback that informs further model development. Collaboration with data scientists, machine learning engineers, and product teams is common to align evaluation efforts with organizational goals. The role also includes documenting findings, participating in regular team meetings, and staying updated on best practices in prompt design and AI ethics.

What are the key skills and qualifications needed to thrive in the LLM prompt evaluation position?

To thrive in LLM Prompt Evaluation, you need a solid understanding of natural language processing, critical thinking, and analytical skills, often supported by a background in linguistics, computer science, or related disciplines. Familiarity with annotation tools, large language model (LLM) platforms, and prompt engineering frameworks is important. Attention to detail, strong written communication, and the ability to provide objective, structured feedback are key soft skills. These qualifications ensure accurate assessments of AI-generated responses and support the continuous improvement of language models.

What are popular job titles related to Llm Prompt Evaluation jobs in Decatur, GA?

For Llm Prompt Evaluation jobs in Decatur, GA, the most frequently searched job titles are:

What job categories do people searching Llm Prompt Evaluation jobs in Decatur, GA look for?

The top searched job categories for Llm Prompt Evaluation jobs in Decatur, GA are:

What cities near Decatur, GA are hiring for Llm Prompt Evaluation jobs?

Cities near Decatur, GA with the most Llm Prompt Evaluation job openings:

Senior AI/ML Engineer

Sumeru

Atlanta, GA • On-site

$100K - $138K/yr

Other

Re-posted 3 days ago


Job description

Role: Senior AI/ML Engineer
Location: Bellevue/Seattle, WA ; Atlanta, GA, and Frisco, TX


Need Local Candidates


Job Overview

We are seeking an AI/ML Engineer to build the intelligent systems that power identity resolution and data accessibility within our Customer Data Platform (CDP) - the authoritative source of truth for customer data across the entire US adult population.

This role focuses on developing machine learning pipelines that deduplicate, link, and resolve customer identities across disparate data sources - the core capability that transforms raw data into trusted, unified customer profiles. You will also contribute to LLM-based solutions that enable natural language querying of CDP data, making the platform accessible to business users across the organization.

You will work on both classical ML techniques and modern LLM-based approaches to ensure that every customer identity in CDP is accurately resolved, every profile is trustworthy, and every user can access the data they need.

Job Responsibilities - Identity Resolution

  • Develop and deploy entity resolution models to match and deduplicate customer records across multiple systems - directly impacting the accuracy of CDP as the source of truth
  • Implement probabilistic matching techniques (e.g., Fellegi-Sunter) and ML models (gradient boosting, neural classifiers) for record linkage across the US adult population
  • Build candidate blocking pipelines using phonetic algorithms (Soundex, Double Metaphone), token similarity, and LSH to handle billions of potential match pairs efficiently
  • Apply fuzzy matching techniques (Levenshtein, Jaro-Winkler, Jaccard) for customer attributes such as name, address, phone, and identifiers
  • Develop clustering algorithms (DBSCAN, hierarchical clustering) to create unified "golden customer profiles" that serve as the authoritative representation of each individual
  • Build embedding-based similarity systems using Sentence-BERT or transformer-based models for semantic matching
  • Implement ANN/KNN retrieval systems (FAISS, Annoy) for large-scale entity matching across population-scale datasets

Job Responsibilities - AI/LLM

  • Use LLMs (e.g., GPT, Claude) for classification and disambiguation of entity matches, improving resolution accuracy where traditional methods fall short
  • Build and support RAG pipelines to enrich customer profiles with contextual data from unstructured sources
  • Perform prompt engineering and evaluation for structured data extraction from unstructured inputs feeding into CDP
  • Contribute to NLQ-to-SQL systems, enabling business users to query CDP data using natural language - making the authoritative source of truth accessible to non-technical stakeholders
  • Support integration with vector databases (e.g., Pinecone, pgvector, Qdrant) for semantic search across customer data

Education and Work Experience

  • Bachelor's or Master's degree in Computer Science, Data Science, or related field
  • 3+ years of experience in ML/AI engineering
  • At least 1 year of experience in entity resolution, record linkage, or deduplication - ideally at scale

Technical Skills

  • Programming: Python (required)
  • Libraries: scikit-learn, HuggingFace Transformers, RapidFuzz, jellyfish
  • Experience with LLM APIs (OpenAI, Anthropic) and prompt pipelines
  • Strong SQL skills and experience with Spark or Dask for distributed processing
  • Familiarity with vector databases and embedding-based retrieval
  • Experience with ML lifecycle tools (MLflow or similar)
  • Understanding of data quality metrics and how identity resolution impacts downstream trust

Knowledge, Skills, and Abilities

  • Strong understanding of ML fundamentals and similarity matching techniques applied to customer identity
  • Ability to work with large, messy, real-world datasets spanning hundreds of millions of records
  • Understanding of precision/recall tradeoffs in identity resolution and their impact on data trust
  • Good problem-solving and analytical skills
  • Ability to collaborate with data engineering, platform, and business teams to deliver accurate customer profiles

Sumeru logo

About Sumeru

Sourced by ZipRecruiter

Industry

It services

Company size

501 - 1,000 Employees

Headquarters location

Washington, DC, US

Year founded

2002