1

Llm Prompt Evaluation Jobs in Bothell, WA (NOW HIRING)

Architect and maintain LLM-based systems, including prompt pipelines, agentic workflows, tool ... Contribute to internal standards for AI development: prompt management, evaluation frameworks ...

Senior AI Engineer

Seattle, WA · On-site

$170K - $200K/yr

Architect and maintain LLM-based systems, including prompt pipelines, agentic workflows, tool ... Contribute to internal standards for AI development: prompt management, evaluation frameworks ...

Senior AI Engineer

Seattle, WA · On-site

$170K - $200K/yr

Architect and maintain LLM-based systems, including prompt pipelines, agentic workflows, tool ... Contribute to internal standards for AI development: prompt management, evaluation frameworks ...

Technical Product Manager, LLM/ML Domain

Seattle, WA · On-site

$190K - $219K/yr

... prompt management, evaluation frameworks, etc.) • Comfortable discussing architecture, APIs, infrastructure trade-offs, and scalability concerns with senior engineers • Strong analytical mindset ...

... prompt injection techniques • Explore edge cases to provoke disallowed, harmful, or incorrect ... or evaluation tooling • Comfort with structured data annotation and rubric-based scoring • ...

Senior Software Engineer - Backend

Bellevue, WA · On-site

$138K - $182K/yr

... with LLM services. Designing, testing, and optimizing prompts and prompt flows for LLMs and RAG ... evaluation metrics. Engineering Practices: TDD/unit integration tests, API documentation ...

Lead AI Engineer

Bellevue, WA · On-site

$115K - $151K/yr

... prompt engineering, structured outputs (JSON schemas/function calling), and tool-augmented LLMs ... LLM evaluation using automated and human-in-the-loop techniques (offline + online). • Optimize ...

... prompt/context patterns. * Implement LLM application patterns including RAG, document ingestion/chunking, embeddings, vector/hybrid search, and retrieval/evaluation telemetry. * Deliver governed ...

Senior Software Engineer (AI)

Seattle, WA · On-site

$139K - $183K/yr

... LLM-powered features end-to-end. • Design and maintain evaluation frameworks for prompts ... prompt iteration. • Ability to move quickly and independently in a fast-paced environment. • ...

... LLM-based agents for reasoning, synthesis, or evaluation tasks. • Familiarity with prompt engineering and multi-step reasoning--designing structured flows that balance quality, latency, and cost ...

Senior AI Engineer

Seattle, WA · On-site

$139K - $183K/yr

... maintain LLM-based systems, including prompt pipelines, agentic workflows, tool-calling ... management, evaluation frameworks, model selection, and system architecture Qualifications

Lead AI Engineer

Bellevue, WA · On-site

$155K - $167K/yr

Conduct LLM evaluation using automated and human-in-the-loop techniques (offline + online ... Prompt engineering expertise * Embeddings and vector search * Experienced in backend API design ...

next page

Showing results 1-20

Llm Prompt Evaluation information

See Bothell, WA salary details

$8

$23

$46

How much do llm prompt evaluation jobs pay per hour?

As of Aug 27, 2026, the average hourly pay for llm prompt evaluation in Bothell, WA is $23.35, according to ZipRecruiter salary data. Most workers in this role earn between $16.11 and $30.91 per hour, depending on experience, location, and employer.

What is an LLM prompt evaluation?

An LLM Prompt Evaluation job involves assessing and optimizing prompts used in large language models (LLMs) to ensure they generate accurate, relevant, and high-quality responses. Evaluators test different prompts, analyze model outputs, and refine phrasing to improve performance. This role requires a strong understanding of AI behavior, critical thinking, and sometimes domain-specific expertise to create effective instructions for the model.

What does an LLM prompt evaluation do?

Professionals in LLM Prompt Evaluation spend their days reviewing, analyzing, and scoring the outputs of large language models based on specific prompts. This involves identifying inaccuracies, biases, or other issues in AI-generated responses, and providing detailed, constructive feedback that informs further model development. Collaboration with data scientists, machine learning engineers, and product teams is common to align evaluation efforts with organizational goals. The role also includes documenting findings, participating in regular team meetings, and staying updated on best practices in prompt design and AI ethics.

What are the key skills and qualifications needed to thrive in the LLM prompt evaluation position?

To thrive in LLM Prompt Evaluation, you need a solid understanding of natural language processing, critical thinking, and analytical skills, often supported by a background in linguistics, computer science, or related disciplines. Familiarity with annotation tools, large language model (LLM) platforms, and prompt engineering frameworks is important. Attention to detail, strong written communication, and the ability to provide objective, structured feedback are key soft skills. These qualifications ensure accurate assessments of AI-generated responses and support the continuous improvement of language models.

What job categories do people searching Llm Prompt Evaluation jobs in Bothell, WA look for?

The top searched job categories for Llm Prompt Evaluation jobs in Bothell, WA are:

What cities near Bothell, WA are hiring for Llm Prompt Evaluation jobs?

Cities near Bothell, WA with the most Llm Prompt Evaluation job openings:

Infographic showing various Llm Prompt Evaluation job openings in Bothell, WA as of August 2026, with employment types broken down into 3% As Needed, 77% Full Time, 16% Part Time, and 4% Contract. Highlights an 91% Physical, 2% Hybrid, and 7% Remote job distribution, with an average salary of $48,571 per year, or $23.4 per hour.

Senior AI Engineer

Seattle, WA

$170K - $200K/yr

Full-time

Medical, Life, Retirement

Re-posted 7 days ago


Job description

Who we are

The real world is the next frontier, and at Metropolis, we are creating the artificial intelligence to make it responsive. We are pioneering the Recognition Economy - a future where mundane repetition disappears and being known unlocks access, comfort, and belonging everywhere you go. From transforming parking into a seamless drive-in, drive-out experience for millions of Members to expanding our intelligence layer across retail and hospitality, we are building a world that feels instinctive and magical. The future isn't coming; it's here, and we need builders, innovators, and problem solvers to help us create it.

Who you are

Metropolis is seeking a Senior AI Engineer to join our Applied AI organization, a team purpose-built to rebuild how the company operates from an AI-first perspective. You will be the builder at the heart of this transformation, designing and shipping AI-powered tools and automation pipelines that replace manual, repetitive work across finance, operations, revenue, and beyond. This is a hands-on engineering role with enormous scope: you will own systems end-to-end, from the LLM prompt layer through to production deployment, working side-by-side with Process Analysts who understand the business deeply. If you want to see your work directly change how a company functions - not someday, but now - this is the role.

What you'll do
  • Design, build, and ship AI-powered automation tools and internal applications that transform manual business processes into AI-first workflows
  • Architect and maintain LLM-based systems, including prompt pipelines, agentic workflows, tool-calling integrations, and retrieval-augmented generation (RAG) setups
  • Build and maintain data pipelines that connect source systems (Snowflake, internal APIs, SaaS tools) to AI workflows, ensuring data quality and reliability
  • Partner with Process Analysts to understand how business functions work today, translate requirements into engineering specifications, and iterate based on real user feedback
  • Own production reliability for the systems you build - monitoring, alerting, and iterating on AI outputs to maintain quality over time
  • Evaluate and integrate emerging AI tooling (models, frameworks, APIs) to ensure Metropolis is always leveraging the best available capabilities
  • Contribute to internal standards for AI development: prompt management, evaluation frameworks, model selection, and system architecture
What we're looking for
  • 6+ years of software engineering experience, with 2+ years of hands-on experience designing, building, and deploying LLM-based or AI-powered applications in production
  • Strong Python skills - you write clean, maintainable code and are comfortable building APIs, data pipelines, and backend services
  • Practical experience with LLM APIs (OpenAI, Anthropic, or similar), prompt engineering, and agentic patterns (function calling, multi-step reasoning, tool use)
  • Familiarity with data engineering fundamentals: SQL, data transformation, working with cloud data warehouses (Snowflake preferred)
  • Experience with workflow orchestration tools (e.g., Airflow, Prefect, or similar) and cloud infrastructure (AWS preferred)
  • A strong instinct for product quality - you care about whether the tools you build are actually useful, not just technically functional
  • Ability to work autonomously in a fast-moving environment, manage competing priorities, and ship iteratively rather than waiting for perfection
While not required, these are a plus:
  • Experience with vector databases, embeddings, or semantic search is a plus
Our Stack
  • Languages + Frameworks: TypeScript, React, Scala (principally), Java (limited)
  • Datastores: MySQL, PostgreSQL, Snowflake
  • Cloud: AWS
  • Version control: Git & GitHub
  • AI Tooling: Copilot on GitHub and Claude Code
  • Observability: Datadog

4 Days in Office: Metropolis values in-person collaboration to drive innovation, strengthen culture, and enhance the Member experience. Our corporate team members hold to our office-first model, which requires employees to be on-site at least four days a week, fostering organic interactions that spark creativity and connection

When you join Metropolis, you'll join a team of world-class product leaders and engineers, building an ecosystem of technologies at the intersection of parking, mobility, and real estate. Our goal is to build an inclusive culture where everyone has a voice and the best idea wins. You will play a key role in building and maintaining this culture as our organization grows. The anticipated base salary for this position is $170,000.00 USD to $200,000.00 USD annually. The actual base salary offered is determined by a number of variables, including, as appropriate, the applicant's qualifications for the position, years of relevant experience, distinctive skills, level of education attained, certifications or other professional licenses held, and the location of residence and/or place of employment. Base salary is one component of Metropolis' total compensation package, which may also include access to or eligibility for healthcare benefits, a 401(k) plan, short-term and long-term disability coverage, basic life insurance, a lucrative stock option plan, bonus plans, and more. #LI-CM1 #LI-Onsite