... Prompt Engineer AI Specialist Technical Analyst to support enterprise scale Generative AI and LLM ... Solid understanding of data science and ML fundamentals model evaluation feature engineering ...
Quick apply
... Prompt Engineer AI Specialist Technical Analyst to support enterprise scale Generative AI and LLM ... Solid understanding of data science and ML fundamentals model evaluation feature engineering ...
Quick apply
... Prompt Engineer AI Specialist Technical Analyst to support enterprise scale Generative AI and LLM ... Solid understanding of data science and ML fundamentals model evaluation feature engineering ...
Work with OpenAI SDK and other LLM providers (Anthropic, Azure OpenAI, Cohere, etc.). * Manage prompt engineering, prompt routing, safety guardrails, and evaluation metrics. Data & Vector Search ...
Quick apply
Work with OpenAI SDK and other LLM providers (Anthropic, Azure OpenAI, Cohere, etc.). * Manage prompt engineering, prompt routing, safety guardrails, and evaluation metrics. Data & Vector Search ...
Atlanta, GA · On-site +1
$100K - $138K/yr
Mentor engineers on agent patterns, prompt hygiene, eval discipline, and LLM failure modes ... Real eval experience golden sets, offline and online evaluations, used to make ship/no-ship calls.
Atlanta, GA · On-site +1
$100K - $138K/yr
Mentor engineers on agent patterns, prompt hygiene, eval discipline, and LLM failure modes ... Real eval experience golden sets, offline and online evaluations, used to make ship/no-ship calls.
Atlanta, GA · Remote
$107K - $146K/yr
Mentor engineers on agent patterns, prompt hygiene, eval discipline, and LLM failure modes ... Real eval experience golden sets, offline and online evaluations, used to make ship/no-ship calls.
Quick apply
Atlanta, GA · Remote
$107K - $146K/yr
Mentor engineers on agent patterns, prompt hygiene, eval discipline, and LLM failure modes ... Real eval experience golden sets, offline and online evaluations, used to make ship/no-ship calls.
Atlanta, GA · On-site
Develop prompt engineering, optimization, and safety techniques for agentic LLM interactions ... Participate in design reviews, architecture discussions, and model evaluations. * Document ...
Atlanta, GA · On-site
Develop prompt engineering, optimization, and safety techniques for agentic LLM interactions ... Participate in design reviews, architecture discussions, and model evaluations. * Document ...
Atlanta, GA · On-site
$100K - $138K/yr
Responsibilities : • Architect and deliver end-to-end LLM-powered applications and agentic ... Establish evaluation, reliability, and performance strategies (accuracy, latency, cost ...
Atlanta, GA · On-site
$100K - $138K/yr
Responsibilities : • Architect and deliver end-to-end LLM-powered applications and agentic ... Establish evaluation, reliability, and performance strategies (accuracy, latency, cost ...
Develop prompt engineering, optimization, and safety techniques for agentic LLM interactions ... Participate in design reviews, architecture discussions, and model evaluations. * Document ...
Quick apply
Develop prompt engineering, optimization, and safety techniques for agentic LLM interactions ... Participate in design reviews, architecture discussions, and model evaluations. * Document ...
Atlanta, GA · On-site
... LLM-powered applications or agentic product workflows • Strong understanding of prompt design, tool use, retrieval, orchestration, and evaluation techniques • Ability to move quickly from ...
Atlanta, GA · On-site
... LLM-powered applications or agentic product workflows • Strong understanding of prompt design, tool use, retrieval, orchestration, and evaluation techniques • Ability to move quickly from ...
Atlanta, GA · On-site
Develop prompt engineering, optimization, and safety techniques for agentic LLM interactions ... Participate in design reviews, architecture discussions, and model evaluations. * Document ...
Atlanta, GA · On-site
Develop prompt engineering, optimization, and safety techniques for agentic LLM interactions ... Participate in design reviews, architecture discussions, and model evaluations. * Document ...
Atlanta, GA · On-site
$120 - $160/hr
Own the full lifecycle: prototyping → evaluation → production deployment → iteration * Ensure ... Strong understanding of LLM capabilities and limitations * Experience with prompt engineering and ...
Atlanta, GA · On-site
$120 - $160/hr
Own the full lifecycle: prototyping → evaluation → production deployment → iteration * Ensure ... Strong understanding of LLM capabilities and limitations * Experience with prompt engineering and ...
Atlanta, GA · On-site
$100K - $138K/yr
Own the full lifecycle: prototyping → evaluation → production deployment → iteration * Ensure ... Strong understanding of LLM capabilities and limitations * Experience with prompt engineering and ...
Atlanta, GA · On-site
$100K - $138K/yr
Own the full lifecycle: prototyping → evaluation → production deployment → iteration * Ensure ... Strong understanding of LLM capabilities and limitations * Experience with prompt engineering and ...
... LLM-as-judge, regression suites) using AgentCore Evaluations, and wire them into CI * Implement guardrails around tool execution: auth scoping, input/output validation, PII and prompt-injection ...
... LLM-as-judge, regression suites) using AgentCore Evaluations, and wire them into CI * Implement guardrails around tool execution: auth scoping, input/output validation, PII and prompt-injection ...
Atlanta, GA · On-site
Palantir Foundry (data, ontology, pipelines, apps) Palantir AIP (LLM agents, Logic, RAG/OAG ... prompt design, orchestration, grounding). Solid understanding of AI/ML fundamentals (evaluation ...
New
Atlanta, GA · On-site
Palantir Foundry (data, ontology, pipelines, apps) Palantir AIP (LLM agents, Logic, RAG/OAG ... prompt design, orchestration, grounding). Solid understanding of AI/ML fundamentals (evaluation ...
New
LLM-driven content generation in consumer experiences, including prompt engineering, evaluation harnesses, and guardrails * Reduction-to-practice work with research scientists * Labs-to-flagship ...
LLM-driven content generation in consumer experiences, including prompt engineering, evaluation harnesses, and guardrails * Reduction-to-practice work with research scientists * Labs-to-flagship ...
Experience with LLM application development, embeddings, model evaluation, prompt optimization, and production AI/ML implementation patterns. * Strong understanding of Google Cloud AI and data ...
Experience with LLM application development, embeddings, model evaluation, prompt optimization, and production AI/ML implementation patterns. * Strong understanding of Google Cloud AI and data ...
Atlanta, GA · On-site
$61 - $83.75/hr
Experience with LLM application development, embeddings, model evaluation, prompt optimization, and production AI/ML implementation patterns. * Strong understanding of Google Cloud AI and data ...
Atlanta, GA · On-site
$61 - $83.75/hr
Experience with LLM application development, embeddings, model evaluation, prompt optimization, and production AI/ML implementation patterns. * Strong understanding of Google Cloud AI and data ...
LLM-driven content generation in consumer experiences, including prompt engineering, evaluation harnesses, and guardrails * Reduction-to-practice work with research scientists * Labs-to-flagship ...
LLM-driven content generation in consumer experiences, including prompt engineering, evaluation harnesses, and guardrails * Reduction-to-practice work with research scientists * Labs-to-flagship ...
Atlanta, GA · Remote
$65 - $89/hr
Experience with LLM application development, embeddings, model evaluation, prompt optimization, and production AI/ML implementation patterns. * Strong understanding of Google Cloud AI and data ...
Quick apply
Atlanta, GA · Remote
$65 - $89/hr
Experience with LLM application development, embeddings, model evaluation, prompt optimization, and production AI/ML implementation patterns. * Strong understanding of Google Cloud AI and data ...
Atlanta, GA · On-site +1
$116K - $195K/yr
Experience with LLM evaluation: building datasets, defining scorers/metrics, and using eval results to drive iteration * Strong understanding of prompt engineering, tool/function calling, and agent ...
Atlanta, GA · On-site +1
$116K - $195K/yr
Experience with LLM evaluation: building datasets, defining scorers/metrics, and using eval results to drive iteration * Strong understanding of prompt engineering, tool/function calling, and agent ...
Alpharetta, GA · On-site
$40.25 - $55/hr
Title: QA Test Engineer Location: Alpharetta, GA Duration: 7 months Position type: W2 contract. Face to Face interview is needed for this position. PS:
Quick apply
Alpharetta, GA · On-site
$40.25 - $55/hr
Title: QA Test Engineer Location: Alpharetta, GA Duration: 7 months Position type: W2 contract. Face to Face interview is needed for this position. PS:
$7.28 - $10.28
18% of jobs
$10.28 - $13.29
2% of jobs
$13.75 is the 25th percentile. Wages below this are outliers.
$13.29 - $16.30
30% of jobs
$16.30 - $19.31
15% of jobs
$19.31 - $22.32
9% of jobs
$22.77 is the 75th percentile. Wages above this are outliers.
$22.32 - $25.33
5% of jobs
$25.33 - $28.33
3% of jobs
$28.33 - $31.34
5% of jobs
$31.34 - $34.35
4% of jobs
$34.35 - $37.36
4% of jobs
$37.36 - $40.37
3% of jobs
$7
$20
$40
An LLM Prompt Evaluation job involves assessing and optimizing prompts used in large language models (LLMs) to ensure they generate accurate, relevant, and high-quality responses. Evaluators test different prompts, analyze model outputs, and refine phrasing to improve performance. This role requires a strong understanding of AI behavior, critical thinking, and sometimes domain-specific expertise to create effective instructions for the model.
Professionals in LLM Prompt Evaluation spend their days reviewing, analyzing, and scoring the outputs of large language models based on specific prompts. This involves identifying inaccuracies, biases, or other issues in AI-generated responses, and providing detailed, constructive feedback that informs further model development. Collaboration with data scientists, machine learning engineers, and product teams is common to align evaluation efforts with organizational goals. The role also includes documenting findings, participating in regular team meetings, and staying updated on best practices in prompt design and AI ethics.
To thrive in LLM Prompt Evaluation, you need a solid understanding of natural language processing, critical thinking, and analytical skills, often supported by a background in linguistics, computer science, or related disciplines. Familiarity with annotation tools, large language model (LLM) platforms, and prompt engineering frameworks is important. Attention to detail, strong written communication, and the ability to provide objective, structured feedback are key soft skills. These qualifications ensure accurate assessments of AI-generated responses and support the continuous improvement of language models.
For Llm Prompt Evaluation jobs in Decatur, GA, the most frequently searched job titles are:
The top searched job categories for Llm Prompt Evaluation jobs in Decatur, GA are:
Cities near Decatur, GA with the most Llm Prompt Evaluation job openings:
Sourced by ZipRecruiter
Cynet Systems Inc is a staffing and recruiting corporation nestled in Ashburn, VA, USA. Established in 2010, the company operates within the Information Technology and Services sector, specializing in providing effective workforce solutions to different business needs, including IT consulting, direct hire, and contract staffing services. Through the years, Cynet Systems has built an impressive portfolio, going beyond borders and expanding its operations internationally in Canada and India. Rooted in its core values of teamwork, leadership, and commitment, Cynet Systems helps businesses unlock their full potential by providing versatile and competent professionals that perfectly align with their needs. Fueled by their unwavering mission to deliver top-tier talent to businesses worldwide, Cynet Systems garnered various recognitions including SIA's fastest-growing staffing firms and Best Place to Work in Virginia for 2019.
It services
501 - 1,000 Employees
Sterling, VA, US
2010