Designing and executing testing scenarios to identify how prompts or user inputs could be manipulated to generate harmful, misleading, or misaligned generative artificial intelligence outputs.
Designing and executing testing scenarios to identify how prompts or user inputs could be manipulated to generate harmful, misleading, or misaligned generative artificial intelligence outputs.
Designing and executing testing scenarios to identify how prompts or user inputs could be manipulated to generate harmful, misleading, or misaligned generative artificial intelligence outputs.
Designing and executing testing scenarios to identify how prompts or user inputs could be manipulated to generate harmful, misleading, or misaligned generative artificial intelligence outputs.
Software Development Manager, Amazon Music Artificial Intelligence & Personalization
Culver City, CA · On-site
$135K - $178K/yr
... Amazon Music's Artificial Intelligence & Personalization organization. We're transforming how ... testing, certification, and livesite operations - Experience partnering with product or program ...
Software Development Manager, Amazon Music Artificial Intelligence & Personalization
Culver City, CA · On-site
$135K - $178K/yr
... Amazon Music's Artificial Intelligence & Personalization organization. We're transforming how ... testing, certification, and livesite operations - Experience partnering with product or program ...
Senior Threat Intelligence Research Engineer
Sunnyvale, CA · On-site
$122K - $168K/yr
... testing, artificial intelligence and advanced malware concepts. Since this team engages externally, candidates must have some soft skills and be willing to speak, network and travel while applying ...
Senior Threat Intelligence Research Engineer
Sunnyvale, CA · On-site
$122K - $168K/yr
... testing, artificial intelligence and advanced malware concepts. Since this team engages externally, candidates must have some soft skills and be willing to speak, network and travel while applying ...
Senior Threat Intelligence Research Engineer
Sunnyvale, CA · Remote
$122K - $168K/yr
... testing, artificial intelligence and advanced malware concepts. Since this team engages externally, candidates must have some soft skills and be willing to speak, network and travel while applying ...
Senior Threat Intelligence Research Engineer
Sunnyvale, CA · Remote
$122K - $168K/yr
... testing, artificial intelligence and advanced malware concepts. Since this team engages externally, candidates must have some soft skills and be willing to speak, network and travel while applying ...
Senior Threat Intelligence Research Engineer
Sunnyvale, CA · Remote
$122K - $168K/yr
... testing, artificial intelligence and advanced malware concepts. Since this team engages externally, candidates must have some soft skills and be willing to speak, network and travel while applying ...
Senior Threat Intelligence Research Engineer
Sunnyvale, CA · Remote
$122K - $168K/yr
... testing, artificial intelligence and advanced malware concepts. Since this team engages externally, candidates must have some soft skills and be willing to speak, network and travel while applying ...
Intelligence Warehouse Systems
Fountain Valley, CA · On-site
$18.75 - $22.75/hr
... testing, validation, and post-launch performance tracking of automation solutions. Support ... Experience with artificial intelligence tools (Copilot, Ai Agents, etc.) is a plus. Work Schedule ...
Intelligence Warehouse Systems
Fountain Valley, CA · On-site
$18.75 - $22.75/hr
... testing, validation, and post-launch performance tracking of automation solutions. Support ... Experience with artificial intelligence tools (Copilot, Ai Agents, etc.) is a plus. Work Schedule ...
... testing. Conduct ELINT signals analyst of operational systems derived from radar emissions data ... Familiarity or experience with Artificial Intelligence/Mission Learning.
... testing. Conduct ELINT signals analyst of operational systems derived from radar emissions data ... Familiarity or experience with Artificial Intelligence/Mission Learning.
... testing. Conduct ELINT signals analyst of operational systems derived from radar emissions data ... Familiarity or experience with Artificial Intelligence/Mission Learning. Employment Type: Full Time
... testing. Conduct ELINT signals analyst of operational systems derived from radar emissions data ... Familiarity or experience with Artificial Intelligence/Mission Learning. Employment Type: Full Time
Intermediate AI and Decision Engineer with Security Clearance
San Diego, CA · On-site
$121K - $152K/yr
... AI testing and evaluation, multi-agent systems, knowledge-based reasoning, automated planning, semantics, and ontologies. As an Artificial Intelligence and Decision Engineer, you will have ...
Intermediate AI and Decision Engineer with Security Clearance
San Diego, CA · On-site
$121K - $152K/yr
... AI testing and evaluation, multi-agent systems, knowledge-based reasoning, automated planning, semantics, and ontologies. As an Artificial Intelligence and Decision Engineer, you will have ...
... testing methodologies to evaluate application performance • Build tools to automate workload ... Official account of Mohamed bin Zayed University of Artificial Intelligence. Dedicated to research ...
... testing methodologies to evaluate application performance • Build tools to automate workload ... Official account of Mohamed bin Zayed University of Artificial Intelligence. Dedicated to research ...
... testing (Python, Docker, Git). • Familiarity with data processing and pipeline orchestration ... Official account of Mohamed bin Zayed University of Artificial Intelligence. Dedicated to research ...
... testing (Python, Docker, Git). • Familiarity with data processing and pipeline orchestration ... Official account of Mohamed bin Zayed University of Artificial Intelligence. Dedicated to research ...
Artificial Intelligence solutions * Large Language Models (LLMs) * AI-enabled workflows and ... Strong understanding of automated testing, deployment, and release management processes.
Artificial Intelligence solutions * Large Language Models (LLMs) * AI-enabled workflows and ... Strong understanding of automated testing, deployment, and release management processes.
Lead Software & AI Engineer
San Diego, CA · On-site
$225K/yr
Artificial Intelligence solutions * Large Language Models (LLMs) * AI-enabled workflows and ... Strong understanding of automated testing, deployment, and release management processes.
Lead Software & AI Engineer
San Diego, CA · On-site
$225K/yr
Artificial Intelligence solutions * Large Language Models (LLMs) * AI-enabled workflows and ... Strong understanding of automated testing, deployment, and release management processes.
Lead Software & AI Engineer
San Diego, CA · On-site
$225K/yr
Artificial Intelligence solutions * Large Language Models (LLMs) * AI-enabled workflows and ... Strong understanding of automated testing, deployment, and release management processes.
Lead Software & AI Engineer
San Diego, CA · On-site
$225K/yr
Artificial Intelligence solutions * Large Language Models (LLMs) * AI-enabled workflows and ... Strong understanding of automated testing, deployment, and release management processes.
Machine Learning Infrastructure Engineer
Sunnyvale, CA · On-site
$125K - $164K/yr
... design, testing) • Proven multi-node experience (e.g., Slurm, Kubernetes, Ray) and debugging ... Official account of Mohamed bin Zayed University of Artificial Intelligence. Dedicated to research ...
Machine Learning Infrastructure Engineer
Sunnyvale, CA · On-site
$125K - $164K/yr
... design, testing) • Proven multi-node experience (e.g., Slurm, Kubernetes, Ray) and debugging ... Official account of Mohamed bin Zayed University of Artificial Intelligence. Dedicated to research ...
Machine Learning Infrastructure Engineer
Sunnyvale, CA · On-site
$125K - $164K/yr
... design, testing) • Proven multi-node experience (e.g., Slurm, Kubernetes, Ray) and debugging ... Official account of Mohamed bin Zayed University of Artificial Intelligence. Dedicated to research ...
Machine Learning Infrastructure Engineer
Sunnyvale, CA · On-site
$125K - $164K/yr
... design, testing) • Proven multi-node experience (e.g., Slurm, Kubernetes, Ray) and debugging ... Official account of Mohamed bin Zayed University of Artificial Intelligence. Dedicated to research ...
... testing (Python, Docker, Git). • Familiarity with data processing and pipeline orchestration ... Official account of Mohamed bin Zayed University of Artificial Intelligence. Dedicated to research ...
... testing (Python, Docker, Git). • Familiarity with data processing and pipeline orchestration ... Official account of Mohamed bin Zayed University of Artificial Intelligence. Dedicated to research ...
Data/Cloud Architect (AWS)
Irvine, CA · On-site
$81K - $109K/yr
... code, testing techniques, and support activities to enrich the knowledge base and assist other ... Preferred Skill and Experience Preferred exposure to Artificial Intelligence services, Large ...
Data/Cloud Architect (AWS)
Irvine, CA · On-site
$81K - $109K/yr
... code, testing techniques, and support activities to enrich the knowledge base and assist other ... Preferred Skill and Experience Preferred exposure to Artificial Intelligence services, Large ...
Data/Cloud Architect (AWS)
Irvine, CA · On-site
$69.75 - $88.75/hr
... code, testing techniques, and support activities to enrich the knowledge base and assist other ... Preferred Skill and Experience Preferred exposure to Artificial Intelligence services, Large ...
New
Data/Cloud Architect (AWS)
Irvine, CA · On-site
$69.75 - $88.75/hr
... code, testing techniques, and support activities to enrich the knowledge base and assist other ... Preferred Skill and Experience Preferred exposure to Artificial Intelligence services, Large ...
New
Artificial Intelligence Testing information
See California salary details
$39.38 - $41
16% of jobs
$41.58 is the 25th percentile. Wages below this are outliers.
$41 - $42.62
24% of jobs
The median wage is $43.38 / hr.
$42.62 - $44.23
21% of jobs
$44.23 - $45.85
10% of jobs
$45.85 - $47.47
0% of jobs
$48.99 is the 75th percentile. Wages above this are outliers.
$47.47 - $49.09
4% of jobs
$49.09 - $50.70
5% of jobs
$50.70 - $52.32
5% of jobs
$52.32 - $53.94
4% of jobs
$53.94 - $55.56
5% of jobs
$55.56 - $57.17
5% of jobs
$39
$45
$57
How much do artificial intelligence testing jobs pay per hour?
What are some common challenges faced by professionals in artificial intelligence testing?
Professionals in Artificial Intelligence Testing often encounter unique challenges, such as validating the unpredictable behavior of machine learning models and ensuring algorithmic fairness and accuracy. They must design comprehensive test cases to cover a wide variety of data inputs and potential edge cases, often in complex, rapidly evolving environments. Collaboration with data scientists, developers, and stakeholders is essential to understand model requirements and to interpret test results accurately. Staying up-to-date with advances in both AI and testing technologies is also key, as the field is continually evolving.
What are the key skills and qualifications needed to thrive in artificial intelligence testing?
To thrive in Artificial Intelligence Testing, candidates typically need a background in computer science, machine learning concepts, software testing methodologies, and knowledge of programming languages like Python or Java. Familiarity with AI testing frameworks, version control systems, and tools such as TensorFlow, PyTorch, or JUnit is highly valued, along with certifications in software testing or AI. Strong problem-solving ability, attention to detail, and effective communication skills are critical soft skills in this role. These qualifications ensure the tester can rigorously validate AI models, collaborate well with development teams, and maintain high-quality, reliable AI systems.
What is an artificial intelligence testing job?
An Artificial Intelligence Testing job involves evaluating and validating AI models, algorithms, and systems to ensure accuracy, reliability, and fairness. Testers design test cases, identify biases, detect errors, and assess model performance under different conditions. They use tools like automation frameworks, data validation techniques, and model debugging to improve AI functionality. The role requires knowledge of machine learning, programming, and testing methodologies to ensure AI systems perform as expected in real-world scenarios.
Is artificial intelligence testing a good career?
How do I become an artificial intelligence tester?

Full-time
Posted 19 days ago
Deloitte rating
8.2
Based on 92 frontline employees who took The Breakroom Quiz
45th of 150 rated financial services
Job description
Cyber Digital Trust and Online Safety Manager
The Digital Trust & Online Protection Professional will advise clients in developing, managing, and implementing policies, procedures, and strategies to ensure a safe, compliant, and trustworthy environment for our users. This individual will scale and mature digital trust and safety processes, including content compliance, user protection, and regulatory adherence across our platforms for our clients. Working closely with cross-functional stakeholders, this role will monitor regulatory changes, manage risks, and enhance our organization's approach to content safety, user trust, and online integrity.
Recruiting for this role ends on 12/31/3026.
Work you'll do
As a Manager, Strategy, Growth, and Transformation on the Deloitte Cyber team, you will be responsible for:
- Designing and executing testing scenarios to identify how prompts or user inputs could be manipulated to generate harmful, misleading, or misaligned generative artificial intelligence outputs.
- Researching emerging prompt injection, jailbreak, and adversarial testing techniques to evaluate model weaknesses, bias, factual inaccuracy, and misalignment with user intent.
- Assessing the effectiveness of content moderation systems in detecting unsafe outputs and documenting vulnerabilities, failure patterns, and potential misuse impacts.
- Recommending improvements to moderation policies, flagging mechanisms, training data, and governance controls based on testing findings.
- Collaborating with generative artificial intelligence development, content moderation, and cross-functional stakeholders to strengthen security, trust, safety, and responsible use outcomes.
- Developing multimodal test content and novel prompt manipulation methods to identify failure modes across text and other model inputs.
A successful candidate would possess these skills:
- Ability to work independently and collaborate as part of a team
- Effective written and verbal communication skills
- Meticulous attention to detail and quality of work product
- Ability to build and sustain professional relationships
- Ability to lead projects or workstreams
- Ability to manage and prioritize multiple tasks in a fast-paced and dynamic environment
- Strong interpersonal skills and professional demeanor
- Ability to meet deadlines
- Ability to mentor and provide clear guidance to others
The team
Enables trust and safety of online communications and digital products, protecting users, consumers, and patients from harm. Enables clients to provide consumer confidence in knowing with whom they are dealing and ensuring the integrity of access to data.
Qualifications
Required:
- Master's degree in Computer Science, Artificial Intelligence, Machine Learning, Data Science, Cybersecurity, Linguistics, Psychology, or a related field, or equivalent professional experience
- 10+ years of experience in threat modeling and simulation, prompt generation and analysis, novel testing, and reporting and improvement
- Demonstrated hands-on experience, portfolio work, publications, or research in prompt injection, jailbreak testing, model evaluation, adversarial machine learning, multimodal artificial intelligence safety, or generative artificial intelligence vulnerability assessment
- Ability to travel 25-50%, on average, based on the work you do and the clients and industries/sectors you serve.
- Limited immigration sponsorship may be available.
Preferred:
- Doctor of Philosophy (PhD) in Computer Science, Artificial Intelligence, Machine Learning, Data Science, Cybersecurity, Linguistics, Psychology, or a related field, or equivalent professional experience
- Specialized training or certifications in generative artificial intelligence red teaming, adversarial machine learning, artificial intelligence security, cybersecurity, responsible artificial intelligence, or artificial intelligence governance
- Experience designing and operationalizing trust and safety testing programs for large-scale consumer platforms, including escalation workflows, issue triage, and remediation tracking
- Experience working with product, legal, policy, and engineering stakeholders to translate risk findings into practical platform controls and governance improvements
The wage range for this role takes into account the wide range of factors that are considered in making compensation decisions including but not limited to skill sets; experience and training; licensure and certifications; and other business and organizational needs. The disclosed range estimate has not been adjusted for the applicable geographic differential associated with the location at which the position may be filled. At Deloitte, it is not typical for an individual to be hired at or near the top of the range for their role and compensation decisions are dependent on the facts and circumstances of each case. A reasonable estimate of the current range is $134,500 to $265,100.
You may also be eligible to participate in a discretionary annual incentive program, subject to the rules governing the program, whereby an award, if any, depends on various factors, including, without limitation, individual and organizational performance.
#CyberDTP27
Cyber Digital Trust and Online Safety Manager
The Digital Trust & Online Protection Professional will advise clients in developing, managing, and implementing policies, procedures, and strategies to ensure a safe, compliant, and trustworthy environment for our users. This individual will scale and mature digital trust and safety processes, including content compliance, user protection, and regulatory adherence across our platforms for our clients. Working closely with cross-functional stakeholders, this role will monitor regulatory changes, manage risks, and enhance our organization's approach to content safety, user trust, and online integrity.
Recruiting for this role ends on 12/31/3026.
Work you'll do
As a Manager, Strategy, Growth, and Transformation on the Deloitte Cyber team, you will be responsible for:
- Designing and executing testing scenarios to identify how prompts or user inputs could be manipulated to generate harmful, misleading, or misaligned generative artificial intelligence outputs.
- Researching emerging prompt injection, jailbreak, and adversarial testing techniques to evaluate model weaknesses, bias, factual inaccuracy, and misalignment with user intent.
- Assessing the effectiveness of content moderation systems in detecting unsafe outputs and documenting vulnerabilities, failure patterns, and potential misuse impacts.
- Recommending improvements to moderation policies, flagging mechanisms, training data, and governance controls based on testing findings.
- Collaborating with generative artificial intelligence development, content moderation, and cross-functional stakeholders to strengthen security, trust, safety, and responsible use outcomes.
- Developing multimodal test content and novel prompt manipulation methods to identify failure modes across text and other model inputs.
A successful candidate would possess these skills:
- Ability to work independently and collaborate as part of a team
- Effective written and verbal communication skills
- Meticulous attention to detail and quality of work product
- Ability to build and sustain professional relationships
- Ability to lead projects or workstreams
- Ability to manage and prioritize multiple tasks in a fast-paced and dynamic environment
- Strong interpersonal skills and professional demeanor
- Ability to meet deadlines
- Ability to mentor and provide clear guidance to others
The team
Enables trust and safety of online communications and digital products, protecting users, consumers, and patients from harm. Enables clients to provide consumer confidence in knowing with whom they are dealing and ensuring the integrity of access to data.
Qualifications
Required:
- Master's degree in Computer Science, Artificial Intelligence, Machine Learning, Data Science, Cybersecurity, Linguistics, Psychology, or a related field, or equivalent professional experience
- 10+ years of experience in threat modeling and simulation, prompt generation and analysis, novel testing, and reporting and improvement
- Demonstrated hands-on experience, portfolio work, publications, or research in prompt injection, jailbreak testing, model evaluation, adversarial machine learning, multimodal artificial intelligence safety, or generative artificial intelligence vulnerability assessment
- Ability to travel 25-50%, on average, based on the work you do and the clients and industries/sectors you serve.
- Limited immigration sponsorship may be available.
Preferred:
- Doctor of Philosophy (PhD) in Computer Science, Artificial Intelligence, Machine Learning, Data Science, Cybersecurity, Linguistics, Psychology, or a related field, or equivalent professional experience
- Specialized training or certifications in generative artificial intelligence red teaming, adversarial machine learning, artificial intelligence security, cybersecurity, responsible artificial intelligence, or artificial intelligence governance
- Experience designing and operationalizing trust and safety testing programs for large-scale consumer platforms, including escalation workflows, issue triage, and remediation tracking
- Experience working with product, legal, policy, and engineering stakeholders to translate risk findings into practical platform controls and governance improvements
The wage range for this role takes into account the wide range of factors that are considered in making compensation decisions including but not limited to skill sets; experience and training; licensure and certifications; and other business and organizational needs. The disclosed range estimate has not been adjusted for the applicable geographic differential associated with the location at which the position may be filled. At Deloitte, it is not typical for an individual to be hired at or near the top of the range for their role and compensation decisions are dependent on the facts and circumstances of each case. A reasonable estimate of the current range is $134,500 to $265,100.
You may also be eligible to participate in a discretionary annual incentive program, subject to the rules governing the program, whereby an award, if any, depends on various factors, including, without limitation, individual and organizational performance.
#CyberDTP27