You will lead and grow a high-performing AI validation team and partner closely with engineering, product, and safety stakeholders to make AI quality and safety a competitive advantage. Role and ...
You will lead and grow a high-performing AI validation team and partner closely with engineering, product, and safety stakeholders to make AI quality and safety a competitive advantage. Role and ...
Systems Engineer, AI Validation
Sunnyvale, CA · Hybrid
$209K - $266K/yr
As a Systems Engineer, you will own a core part of the behavior validation for the Wayve AI Driver, from strategy to implementation and execution. Your contributions will enable the successful ...
Systems Engineer, AI Validation
Sunnyvale, CA · Hybrid
$209K - $266K/yr
As a Systems Engineer, you will own a core part of the behavior validation for the Wayve AI Driver, from strategy to implementation and execution. Your contributions will enable the successful ...
Systems Engineer, AI Validation
Sunnyvale, CA · On-site
$209K - $266K/yr
As a Systems Engineer, you will own a core part of the behavior validation for the Wayve AI Driver, from strategy to implementation and execution. Your contributions will enable the successful ...
Systems Engineer, AI Validation
Sunnyvale, CA · On-site
$209K - $266K/yr
As a Systems Engineer, you will own a core part of the behavior validation for the Wayve AI Driver, from strategy to implementation and execution. Your contributions will enable the successful ...
You will lead and grow a high-performing AI validation team and partner closely with engineering, product, and safety stakeholders to make AI quality and safety a competitive advantage. Role and ...
You will lead and grow a high-performing AI validation team and partner closely with engineering, product, and safety stakeholders to make AI quality and safety a competitive advantage. Role and ...
You will lead and grow a high-performing AI validation team and partner closely with engineering, product, and safety stakeholders to make AI quality and safety a competitive advantage. Role and ...
Quick apply
You will lead and grow a high-performing AI validation team and partner closely with engineering, product, and safety stakeholders to make AI quality and safety a competitive advantage. Role and ...
The role involves defining and driving the validation strategy for AI-assisted software delivery within the Macro Trading Technology team, ensuring trading integrity while accelerating releases ...
The role involves defining and driving the validation strategy for AI-assisted software delivery within the Macro Trading Technology team, ensuring trading integrity while accelerating releases ...
The AI Validation Engineer will define and drive the validation strategy for AI-assisted software delivery, ensuring trading integrity while accelerating releases and collaborating with various ...
The AI Validation Engineer will define and drive the validation strategy for AI-assisted software delivery, ensuring trading integrity while accelerating releases and collaborating with various ...
The AI Validation Engineer will define and drive validation strategies for AI-assisted software delivery, ensuring trading integrity while accelerating releases. Responsibilities : • Define and ...
The AI Validation Engineer will define and drive validation strategies for AI-assisted software delivery, ensuring trading integrity while accelerating releases. Responsibilities : • Define and ...
The AI Validation Engineer will define and drive the validation strategy for AI-assisted software delivery, collaborating with various stakeholders to ensure the integrity and efficiency of trading ...
The AI Validation Engineer will define and drive the validation strategy for AI-assisted software delivery, collaborating with various stakeholders to ensure the integrity and efficiency of trading ...
Define and drive the validation strategy for AI-assisted software delivery across our Macro Trading Technology team, building a scalable approach that protects trading integrity while accelerating ...
Define and drive the validation strategy for AI-assisted software delivery across our Macro Trading Technology team, building a scalable approach that protects trading integrity while accelerating ...
THE ROLE This role focuses on leading system-level, rack-level, and cluster-scale validation for next-generation AI and machine learning platforms. You will define validation strategies, drive ...
New
THE ROLE This role focuses on leading system-level, rack-level, and cluster-scale validation for next-generation AI and machine learning platforms. You will define validation strategies, drive ...
New
Define and drive the validation strategy for AI-assisted software delivery across our Macro Trading Technology team, building a scalable approach that protects trading integrity while accelerating ...
Define and drive the validation strategy for AI-assisted software delivery across our Macro Trading Technology team, building a scalable approach that protects trading integrity while accelerating ...
You will establish scalable AI Quality Engineering operating models, validation frameworks, runtime quality practices, synthetic data ecosystems, simulation-driven testing capabilities, and ...
You will establish scalable AI Quality Engineering operating models, validation frameworks, runtime quality practices, synthetic data ecosystems, simulation-driven testing capabilities, and ...
You will establish scalable AI Quality Engineering operating models, validation frameworks, runtime quality practices, synthetic data ecosystems, simulation-driven testing capabilities, and ...
You will establish scalable AI Quality Engineering operating models, validation frameworks, runtime quality practices, synthetic data ecosystems, simulation-driven testing capabilities, and ...
You will establish scalable AI Quality Engineering operating models, validation frameworks, runtime quality practices, synthetic data ecosystems, simulation-driven testing capabilities, and ...
You will establish scalable AI Quality Engineering operating models, validation frameworks, runtime quality practices, synthetic data ecosystems, simulation-driven testing capabilities, and ...
AI Quality Engineering Leadwith expertise inLLMs
Eden Prairie, MN · On-site
$73K - $94K/yr
Design and implement scalable AI validation frameworks, AI-assisted testing approaches, runtime quality controls, reusable testing accelerators, and workflow optimization capabilities supporting ...
AI Quality Engineering Leadwith expertise inLLMs
Eden Prairie, MN · On-site
$73K - $94K/yr
Design and implement scalable AI validation frameworks, AI-assisted testing approaches, runtime quality controls, reusable testing accelerators, and workflow optimization capabilities supporting ...
Sr. AI QA Engineer / Lead
Eden Prairie, MN · On-site
$90K - $122K/yr
Design and implement scalable AI validation frameworks, AI-assisted testing approaches, runtime quality controls, reusable testing accelerators, and workflow optimization capabilities supporting ...
Sr. AI QA Engineer / Lead
Eden Prairie, MN · On-site
$90K - $122K/yr
Design and implement scalable AI validation frameworks, AI-assisted testing approaches, runtime quality controls, reusable testing accelerators, and workflow optimization capabilities supporting ...
IT Validation Lead (GxP / CSV)
Libertyville, IL · Hybrid
$65 - $80/hr
IT Validation Lead (GxP / CSV / AI Systems) Location: Greater Chicago area, IL (hybrid -- 3 days/week onsite) Valspec--a global provider of system validation and lifecycle services--provides ...
Quick apply
IT Validation Lead (GxP / CSV)
Libertyville, IL · Hybrid
$65 - $80/hr
IT Validation Lead (GxP / CSV / AI Systems) Location: Greater Chicago area, IL (hybrid -- 3 days/week onsite) Valspec--a global provider of system validation and lifecycle services--provides ...
Chief AI Architect
Lanham, MD · On-site
This role will lead the development of scalable AI validation, monitoring, and governance capabilities supporting federal mission readiness and responsible AI deployment. The Chief AI Architect will ...
Chief AI Architect
Lanham, MD · On-site
This role will lead the development of scalable AI validation, monitoring, and governance capabilities supporting federal mission readiness and responsible AI deployment. The Chief AI Architect will ...
Senior AI Quality & Reliability Engineer
Edina, MN · On-site
$91K - $124K/yr
You will design and implement AI Quality Engineering practices, AI validation processes, AI-assisted testing approaches, runtime quality controls, and scalable testing frameworks that support the ...
New
Senior AI Quality & Reliability Engineer
Edina, MN · On-site
$91K - $124K/yr
You will design and implement AI Quality Engineering practices, AI validation processes, AI-assisted testing approaches, runtime quality controls, and scalable testing frameworks that support the ...
New
Ai Validation information
See salary details
$22.60 - $27.64
2% of jobs
$27.64 - $32.69
6% of jobs
$32.69 - $37.74
13% of jobs
$39.32 is the 25th percentile. Wages below this are outliers.
$37.74 - $42.79
13% of jobs
$42.79 - $47.84
11% of jobs
The median wage is $50.36 / hr.
$47.84 - $52.88
12% of jobs
$52.88 - $57.93
9% of jobs
$61.82 is the 75th percentile. Wages above this are outliers.
$57.93 - $62.98
13% of jobs
$62.98 - $68.03
13% of jobs
$68.03 - $73.08
6% of jobs
$73.08 - $78.13
3% of jobs
$22
$51
$78
How much do ai validation jobs pay per hour?
What is an AI Validation job?
An AI Validation job involves testing and verifying artificial intelligence models to ensure they function correctly, reliably, and ethically. This includes evaluating model accuracy, bias, robustness, and compliance with industry standards. AI validation specialists use various techniques such as data validation, performance testing, and adversarial testing to assess AI systems. Their work helps improve model quality, mitigate risks, and ensure AI applications are safe for real-world deployment.
What are some typical responsibilities of an AI Validation professional?
An AI Validation professional is responsible for evaluating the performance, reliability, and safety of AI models before deployment. This includes designing and executing test cases, analyzing outputs for bias and errors, and documenting results to inform further model refinement. They often collaborate closely with data scientists, engineers, and QA teams to ensure the AI system meets quality and compliance standards. Day-to-day tasks may involve working with datasets, preparing validation reports, and participating in cross-functional review meetings. This role is critical in ensuring that AI systems function as intended in real-world applications.
What are the key skills and qualifications needed to thrive in the Ai Validation position, and why are they important?
To thrive in AI Validation, you need strong analytical skills, familiarity with machine learning concepts, and typically a degree in computer science, data science, or a related field. Experience with tools such as Python, TensorFlow, PyTorch, and version control systems, as well as knowledge of data annotation and model evaluation frameworks, is highly valuable. Excellent attention to detail, problem-solving skills, and effective communication are important soft skills for success in this role. These competencies are essential to ensure that AI models are robust, accurate, and meet project requirements for deployment.

Other
Medical, Dental, Vision, Life, Retirement, PTO
Posted 8 days ago
Job description
Sonatus is looking for an experienced Senior Engineering Manager to build and lead our AI Validation function - the team responsible for how we test, evaluate, and govern the AI models and agentic capabilities embedded in our software-defined vehicle and cloud platforms, as well as our cloud-only AI and LLM-based products. You will deeply understand how AI is developed and deployed across Sonatus's embedded, cloud, and LLM/RAG-driven environments, identify gaps and friction in current validation practices, and turn those insights into scalable test strategy, evaluation frameworks, and governance mechanisms that let Sonatus ship trustworthy AI-driven features at automotive scale. You will lead and grow a high-performing AI validation team and partner closely with engineering, product, and safety stakeholders to make AI quality and safety a competitive advantage.
Role and Responsibility:- Define and drive Sonatus's AI validation strategy across embedded, in-vehicle, cloud-connected, and cloud-native AI systems, identifying gaps in model development, testing, deployment, and governance.
- Lead, hire, mentor, and grow a high-performing AI Validation organization, establishing scalable engineering processes, technical direction, and execution excellence.
- Own the end-to-end validation strategy for AI/ML models, LLMs, RAG pipelines, and agentic AI workflows-from data pipelines and model training through cloud services and in-vehicle deployment.
- Architect and operationalize scalable evaluation frameworks and benchmarking platforms for AI systems, including multi-step agentic workflows, using deterministic metrics, LLM-as-a-Judge methodologies, automated regression testing, and production feedback loops.
- Design and maintain evaluation harnesses for RAG and agentic systems that measure retrieval quality, grounding, citation accuracy, factual consistency, context relevance, safety, latency, reliability, and execution correctness.
- Evaluate, integrate, and optimize open-source and commercial AI validation technologies, driving build-versus-buy decisions for Sonatus's AI quality platform.
- Establish AI governance, Responsible AI practices, model lineage, safety guardrails, and compliance processes appropriate for automotive safety-critical systems and enterprise AI products.
- Act as the quality gatekeeper for AI-enabled releases, partnering with engineering, product, safety, and OEM stakeholders to identify risks, define release criteria, and ensure production readiness.
- Collaborate across engineering teams to define validation strategies for emerging AI capabilities, rapidly prototype new evaluation approaches, and standardize successful practices into reusable frameworks.
- Drive continuous improvement by tracking industry advances in AI evaluation, agentic AI, LLM validation, and RAG systems, translating them into scalable validation capabilities across Sonatus.
- Bachelor's degree in Computer Science, Computer Engineering, Electrical Engineering, or a related field required (MS preferred).
- 10+ years of experience in software or systems engineering-including embedded, cloud, networking, security, or automotive domains-with 3+ years leading high-performing engineering or QA organizations.
- Hands-on experience developing, deploying, testing, or operating AI/ML systems, with strong expertise in modern ML workflows, neural networks, and MLOps.
- Deep understanding of LLMs, RAG architectures, vector databases, embeddings, retrieval optimization, and agentic AI frameworks such as LangGraph or equivalent orchestration platforms.
- Proven experience designing and implementing scalable evaluation frameworks for AI systems, including multi-step agentic workflows, regression testing, benchmarking, and automated quality scoring.
- Strong expertise with hybrid evaluation methodologies combining deterministic validation (citation grounding, structural validation, exact matching) and probabilistic LLM-as-a-Judge techniques (faithfulness, answer relevance, context precision, task completion).
- Practical experience with RAG evaluation frameworks such as RAGAS, including evaluation tuning, embedding optimization, retrieval quality improvement, and production-scale LLM evaluation pipelines.
- Experience validating hallucination, grounding, citation accuracy, bias, fairness, toxicity, and factual consistency in production LLM applications.
- Experience designing systems that verify external knowledge claims and ensure responses are grounded in traceable citations and trusted data sources.
- Strong experience testing cloud-native platforms and cloud-managed embedded products, including end-to-end system validation.
- Experience establishing AI governance, safety, compliance, and Responsible AI practices for enterprise or safety-critical systems.
- Proficiency in Python, Linux, shell scripting, modern test frameworks (PyTest, Playwright, Behave), and engineering productivity tools such as Jenkins and JIRA.
- Experience validating AI or agentic systems in safety-critical or regulated industries (automotive, aerospace, medical).
- Track record building and scaling an AI test/evaluation platform or developer experience used by multiple teams (frameworks, reusable components, reference implementations).
- Demonstrated wins moving AI testing practices from ad hoc to standardized, organization-wide adoption, with measurable impact on cycle time, quality, or reliability.
- Experience implementing enterprise-grade AI governance (auditability, monitoring, policy enforcement) in production systems.
- Deep experience evaluating LLM and RAG systems at scale, including agentic workflows, RAGAS-based evaluation, citation verification, hallucination detection, groundedness, faithfulness, answer relevance, tool/task correctness, and automated regression testing across offline and online feedback loops.
Sunnyvale HQ Benefits & Perks Offered:
- Health care plan (Medical, Dental & Vision)
- Flexible and Dependent Care Expense program
- Retirement plan (401k)
- Life Insurance (Basic, Voluntary & AD&D)
- Unlimited paid time off per year, 14+ paid holidays
- Hybrid office work arrangement
- Complimentary lunches, snacks, and beverages during on-site working days
- Wellness benefit allowance
- Phone & Internet reimbursement
- Computer Accessory Allowance
About Sonatus
Sourced by ZipRecruiter
Industry
Software development
Company size
51 - 200 Employees
Headquarters location
Sunnyvale, CA, US
Year founded
2018