1

Weekend Ai Evaluator Jobs (NOW HIRING)

Applied AI Data Analyst

New York, NY · On-site

$180K - $240K/yr

You will own the evaluation platform of our agentic workflows forecasting pipelines, as well as ... Unlimited PTO - plus regular 4-day holiday weekends we actually take * 401(k) through Vestwell

Posted today

Founding AI Engineer

San Francisco, CA · On-site

$160K - $210K/yr

You'll be evaluated on your speed at running thoughtful experiments. We're flexible on scheduling ... weekend to minimize impact on your weekday schedule. * Full-time offer 🎉 Compensation Range ...

New

Senior Speech AI Researcher

New York, NY · On-site +1

$180K - $300K/yr

You'll lead the development of AI assisted evaluation pipelines to characterize and validate the ... Expect roughly 60 hours per week, with occasional weekend work around launches and deadlines. We're ...

Technical Experimentation & Evaluation · Conduct benchmarking, model evaluation, prompt ... Standard business hours · Occasional after-hours or weekend work may be required based on business ...

Technical Experimentation & Evaluation · Conduct benchmarking, model evaluation, prompt ... Standard business hours · Occasional after-hours or weekend work may be required based on business ...

next page

Showing results 1-20

Weekend Ai Evaluator information

See salary details

$29.5K

$65.5K

$106.5K

How much do weekend ai evaluator jobs pay per year?

As of Aug 19, 2026, the average yearly pay for weekend ai evaluator in the United States is $65,471.00, according to ZipRecruiter salary data. Most workers in this role earn between $44,500.00 and $79,500.00 per year, depending on experience, location, and employer.

What is the difference between Weekend Ai Evaluator vs Data Annotator?

AspectWeekend Ai EvaluatorData Annotator
Required CredentialsHigh school diploma or equivalent; some roles may require basic technical skillsHigh school diploma or equivalent; attention to detail essential
Work EnvironmentRemote, flexible hours, part-timeRemote or on-site, flexible or fixed schedules
Industry UsageAI companies, tech startups, research firmsTech companies, AI development, machine learning projects
Common Search/ComparisonWeekend Ai Evaluator vs Data Annotator

The Weekend Ai Evaluator and Data Annotator roles both involve working with data in AI projects, often remotely and part-time. However, Weekend Ai Evaluators typically focus on evaluating AI outputs and providing feedback, while Data Annotators label and prepare data for machine learning models. Both roles require attention to detail and are common in AI and tech industries, but they serve different functions within the AI development process.

How much do Weekend AI Evaluators make?

Weekend AI Evaluators typically earn between $10 and $15 per hour, depending on the company and location. The role often involves evaluating AI outputs and may require basic language or analytical skills, with flexible weekend schedules.

How to become a Weekend AI Evaluator?

To become a Weekend AI Evaluator, you typically need strong language and analytical skills, attention to detail, and the ability to follow guidelines. Many roles require completing online training and passing quality assessments, with flexible weekend schedules often available. Relevant experience with AI or data annotation tools can be beneficial.

What cities are hiring for Weekend Ai Evaluator jobs?

Cities with the most Weekend Ai Evaluator job openings:

What are the most commonly searched types of Ai Evaluator jobs?

The most popular types of Ai Evaluator jobs are:

What states have the most Weekend Ai Evaluator jobs?

States with the most job openings for Weekend Ai Evaluator jobs include:

Infographic showing various Weekend Ai Evaluator job openings in the United States as of August 2026, with employment types broken down into 76% Full Time, 21% Part Time, and 3% Contract. Highlights an 64% Physical, 4% Hybrid, and 32% Remote job distribution, with an average salary of $65,471 per year, or $31.5 per hour.

Applied AI Data Analyst

Confido

New York, NY • On-site

$180K - $240K/yr

Full-time

Medical, Dental, Vision, Retirement, PTO

Posted 15 hours ago

Posted today


Job description

Confido is the AI infrastructure powering modern CPG - the platform that 200+ brands like OLIPOP, Simple Mills, Dr. Squatch, and Tropicana use to run everything from deductions to production planning. Finance, accounting, sales, and operations, unified in one system for the first time.
We're growing 5x year over year with a small team in New York City; the people who join now will shape the product, the culture, and the company itself.
If you want your work on shelves everywhere - and outsized ownership while you build - we'd love to meet you.
The Role
As an Applied AI Data Analyst, you will bridge the gap between our massive, unstructured financial and operational datasets to our core R&D team. You will own the evaluation platform of our agentic workflows forecasting pipelines, as well as improve customer facing insights leading to mission-critical decisions.
This role is part of the Applied AI Research department, with options to advance to senior data science and AI/ML engineering roles.
Location: New York, NY (In-Person | Relocation supported)
What you'll do
  • Algorithm & LLM Evaluation: Analyze edge cases, model predictions, hallucination rates, and extraction accuracy for our LLM-driven document processing systems and agentic workflows.
  • Root Cause Analysis: Dig deep into complex datasets (inconsistent financial documents, fragmented enterprise systems) to determine why an extraction or time-series model succeeded or failed.
  • Golden Dataset Curation: Define, curate, and maintain gold-standard evaluation datasets ("golden sets") for automated backtesting of AI agents and demand forecasting models.
  • Performance Monitoring & Evals: Build dashboards, data transformations, and evaluation harnesses to monitor system health, model drift, and AI performance metrics across release cycles.
  • Cross-Functional Collaboration: Partner directly with applied AI engineers, product managers and business operations managers to turn qualitative real-world problem definitions into quantitative product defining insights.

What we're looking for
Required
  • Education: Graduate degree (M.S. or Ph.D.) in Computer Science, Data Science, Statistics, Mathematics, Physics, or a related STEM field.
  • Experience: 2-4 years of experience in a quantitative data analyst or AI evaluation role, working closely alongside AI/ML R&D teams.
  • Programming & Agentic AI: Proficiency in Python and hands-on works with AI/ML models
  • Data Stack: Strong capabilities with SQL and enterprise data infrastructure (PostgreSQL, Snowflake, dbt, AWS).
  • Analytical Mindset: Exceptional ability to translate noisy, unstructured document data into rigorous evaluation insights for engineers.

Nice to have
  • Experience with LLM observability, tracing, and eval frameworks (LangSmith, Arize AI, or TruLens).
  • Proficiency with agentic workflows or multi-agent orchestration.
  • Familiarity with demand forecasting, time-series models, or statistical modeling.
  • Comfort in fast-moving, high-growth NYC startup environments.

Perks + Benefits
  • Equity - own a meaningful piece of the company you're helping build
  • Fully paid health coverage through Aetna - we cover 100% of employee premiums
  • Top-tier dental and vision coverage through Guardian
  • Carrot Fertility Pro - comprehensive fertility and family-forming support
  • 12 weeks of paid parental leave
  • Unlimited PTO - plus regular 4-day holiday weekends we actually take
  • 401(k) through Vestwell
  • Paid relocation support - we'll help you make the move to NYC
  • Fully equipped workspace from day one - laptop, monitor, keyboard, and a $200 stipend to personalize your setup
  • Team perks - catered Friday lunches, team dinners, and unlimited coffee + snacks featuring products from the brands we work with

Confido provides equal employment opportunities to all employees and applicants for employment and prohibits discrimination and harassment of any type without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state or local laws.