1

Trust Safety Content Moderator Jobs in Spring, TX

... their trusted energy advisors. Our Mission Our mission is to lead the energy transition with ... Brand & Content * Own brand integrity across all materials, channels, and touchpoints, ensuring ...

next page

Showing results 1-20

Trust Safety Content Moderator information

See Spring, TX salary details

$26.3K

$103.8K

$114.8K

How much do trust safety content moderator jobs pay per year?

As of Aug 7, 2026, the average yearly pay for trust safety content moderator in Spring, TX is $103,775.00, according to ZipRecruiter salary data. Most workers in this role earn between $109,500.00 and $113,900.00 per year, depending on experience, location, and employer.

What are the key skills and qualifications needed to thrive as a Trust Safety Content Moderator, and why are they important?

To thrive as a Trust & Safety Content Moderator, you need strong analytical skills, attention to detail, and a solid understanding of platform policies, typically supported by relevant experience or a bachelor's degree. Familiarity with content management systems, moderation tools, and sometimes AI-powered review platforms is commonly required. Excellent judgment, emotional resilience, and effective communication are vital soft skills for handling sensitive or disturbing material and interacting with diverse users. These abilities ensure a safe online environment, compliance with regulations, and the protection of both users and the platform’s reputation.

What are some common challenges faced by Trust Safety Content Moderators, and how are they supported in handling sensitive content?

Trust & Safety Content Moderators often encounter challenging or distressing material as part of their daily responsibilities. To address this, many organizations provide structured support systems such as regular debriefings, access to counseling services, and resilience training. Additionally, teams typically work closely with supervisors and mental health professionals to ensure well-being, while workflow rotations help minimize prolonged exposure to sensitive topics. Open communication and a strong sense of teamwork are also encouraged to foster a supportive work environment.

What is a Trust Safety Content Moderator?

Trust & Safety Content Moderators are professionals responsible for reviewing user-generated content on digital platforms to ensure it complies with community guidelines, legal requirements, and platform policies. They help maintain a safe and respectful online environment by identifying and removing harmful, inappropriate, or illegal content. Their work often involves evaluating text, images, and videos, as well as responding to user reports of violations. Trust & Safety Content Moderators play a crucial role in protecting users and upholding the integrity of online communities.

What is the difference between Trust Safety Content Moderator vs Content Reviewer?

AspectTrust Safety Content ModeratorContent Reviewer
CredentialsHigh school diploma or equivalent; familiarity with platform policiesSimilar; often requires basic education and understanding of content guidelines
Work EnvironmentOnline, remote or office-based, monitoring digital contentOnline, remote or on-site, reviewing user-generated content
Industry UsageSocial media, online platforms, gaming, and community sitesMedia companies, social platforms, e-commerce sites
Search & Comparison IntentHigh overlap; both roles involve content monitoring and policy enforcement

Trust Safety Content Moderators and Content Reviewers both focus on monitoring and evaluating online content to ensure compliance with platform policies. While their titles differ, their roles often overlap in work environment, required skills, and industry usage, making them comparable in many contexts.

Do Trust Safety Content Moderators work from home?

Many Trust Safety Content Moderators work remotely, as companies often allow or require them to perform their duties from home using online tools and platforms. This setup provides flexibility and is common in the industry, especially for roles involving reviewing and moderating digital content. However, some positions may require on-site work depending on the company's policies and the nature of the content being moderated.
What are popular job titles related to Trust Safety Content Moderator jobs in Spring, TX? For Trust Safety Content Moderator jobs in Spring, TX, the most frequently searched job titles are:
What job categories do people searching Trust Safety Content Moderator jobs in Spring, TX look for? The top searched job categories for Trust Safety Content Moderator jobs in Spring, TX are:
What cities near Spring, TX are hiring for Trust Safety Content Moderator jobs? Cities near Spring, TX with the most Trust Safety Content Moderator job openings:
Infographic showing various Trust Safety Content Moderator job openings in Spring, TX as of August 2026, with employment types broken down into 1% As Needed, 79% Full Time, 17% Part Time, 2% Contract, and 1% Nights. Highlights an 99% Physical, and 1% Remote job distribution, with an average salary of $103,775 per year, or $49.9 per hour.

Cyber Digital Trust & Online Safety Manager

Deloitte

Houston, TX

Other

Posted 16 days ago


Deloitte rating

8.2

Company rating: 8.2 out of 10

Based on 92 frontline employees who took The Breakroom Quiz

45th of 150 rated financial services


Job description

Cyber Digital Trust and Online Safety Manager

The Digital Trust & Online Protection Professional will advise clients in developing, managing, and implementing policies, procedures, and strategies to ensure a safe, compliant, and trustworthy environment for our users. This individual will scale and mature digital trust and safety processes, including content compliance, user protection, and regulatory adherence across our platforms for our clients. Working closely with cross-functional stakeholders, this role will monitor regulatory changes, manage risks, and enhance our organization's approach to content safety, user trust, and online integrity.

Recruiting for this role ends on 12/31/3026.

Work you'll do

As a Manager, Strategy, Growth, and Transformation on the Deloitte Cyber team, you will be responsible for:

  • Designing and executing testing scenarios to identify how prompts or user inputs could be manipulated to generate harmful, misleading, or misaligned generative artificial intelligence outputs.
  • Researching emerging prompt injection, jailbreak, and adversarial testing techniques to evaluate model weaknesses, bias, factual inaccuracy, and misalignment with user intent.
  • Assessing the effectiveness of content moderation systems in detecting unsafe outputs and documenting vulnerabilities, failure patterns, and potential misuse impacts.
  • Recommending improvements to moderation policies, flagging mechanisms, training data, and governance controls based on testing findings.
  • Collaborating with generative artificial intelligence development, content moderation, and cross-functional stakeholders to strengthen security, trust, safety, and responsible use outcomes.
  • Developing multimodal test content and novel prompt manipulation methods to identify failure modes across text and other model inputs.

A successful candidate would possess these skills:

  • Ability to work independently and collaborate as part of a team
  • Effective written and verbal communication skills
  • Meticulous attention to detail and quality of work product
  • Ability to build and sustain professional relationships
  • Ability to lead projects or workstreams
  • Ability to manage and prioritize multiple tasks in a fast-paced and dynamic environment
  • Strong interpersonal skills and professional demeanor
  • Ability to meet deadlines
  • Ability to mentor and provide clear guidance to others

The team

Enables trust and safety of online communications and digital products, protecting users, consumers, and patients from harm. Enables clients to provide consumer confidence in knowing with whom they are dealing and ensuring the integrity of access to data.

Qualifications

Required:

  • Master's degree in Computer Science, Artificial Intelligence, Machine Learning, Data Science, Cybersecurity, Linguistics, Psychology, or a related field, or equivalent professional experience
  • 10+ years of experience in threat modeling and simulation, prompt generation and analysis, novel testing, and reporting and improvement
  • Demonstrated hands-on experience, portfolio work, publications, or research in prompt injection, jailbreak testing, model evaluation, adversarial machine learning, multimodal artificial intelligence safety, or generative artificial intelligence vulnerability assessment
  • Ability to travel 25-50%, on average, based on the work you do and the clients and industries/sectors you serve.
  • Limited immigration sponsorship may be available.

Preferred:

  • Doctor of Philosophy (PhD) in Computer Science, Artificial Intelligence, Machine Learning, Data Science, Cybersecurity, Linguistics, Psychology, or a related field, or equivalent professional experience
  • Specialized training or certifications in generative artificial intelligence red teaming, adversarial machine learning, artificial intelligence security, cybersecurity, responsible artificial intelligence, or artificial intelligence governance
  • Experience designing and operationalizing trust and safety testing programs for large-scale consumer platforms, including escalation workflows, issue triage, and remediation tracking
  • Experience working with product, legal, policy, and engineering stakeholders to translate risk findings into practical platform controls and governance improvements

The wage range for this role takes into account the wide range of factors that are considered in making compensation decisions including but not limited to skill sets; experience and training; licensure and certifications; and other business and organizational needs. The disclosed range estimate has not been adjusted for the applicable geographic differential associated with the location at which the position may be filled. At Deloitte, it is not typical for an individual to be hired at or near the top of the range for their role and compensation decisions are dependent on the facts and circumstances of each case. A reasonable estimate of the current range is $134,500 to $265,100.

You may also be eligible to participate in a discretionary annual incentive program, subject to the rules governing the program, whereby an award, if any, depends on various factors, including, without limitation, individual and organizational performance.

#CyberDTP27

Qualifications:

Cyber Digital Trust and Online Safety Manager

The Digital Trust & Online Protection Professional will advise clients in developing, managing, and implementing policies, procedures, and strategies to ensure a safe, compliant, and trustworthy environment for our users. This individual will scale and mature digital trust and safety processes, including content compliance, user protection, and regulatory adherence across our platforms for our clients. Working closely with cross-functional stakeholders, this role will monitor regulatory changes, manage risks, and enhance our organization's approach to content safety, user trust, and online integrity.

Recruiting for this role ends on 12/31/3026.

Work you'll do

As a Manager, Strategy, Growth, and Transformation on the Deloitte Cyber team, you will be responsible for:

  • Designing and executing testing scenarios to identify how prompts or user inputs could be manipulated to generate harmful, misleading, or misaligned generative artificial intelligence outputs.
  • Researching emerging prompt injection, jailbreak, and adversarial testing techniques to evaluate model weaknesses, bias, factual inaccuracy, and misalignment with user intent.
  • Assessing the effectiveness of content moderation systems in detecting unsafe outputs and documenting vulnerabilities, failure patterns, and potential misuse impacts.
  • Recommending improvements to moderation policies, flagging mechanisms, training data, and governance controls based on testing findings.
  • Collaborating with generative artificial intelligence development, content moderation, and cross-functional stakeholders to strengthen security, trust, safety, and responsible use outcomes.
  • Developing multimodal test content and novel prompt manipulation methods to identify failure modes across text and other model inputs.

A successful candidate would possess these skills:

  • Ability to work independently and collaborate as part of a team
  • Effective written and verbal communication skills
  • Meticulous attention to detail and quality of work product
  • Ability to build and sustain professional relationships
  • Ability to lead projects or workstreams
  • Ability to manage and prioritize multiple tasks in a fast-paced and dynamic environment
  • Strong interpersonal skills and professional demeanor
  • Ability to meet deadlines
  • Ability to mentor and provide clear guidance to others

The team

Enables trust and safety of online communications and digital products, protecting users, consumers, and patients from harm. Enables clients to provide consumer confidence in knowing with whom they are dealing and ensuring the integrity of access to data.

Qualifications

Required:

  • Master's degree in Computer Science, Artificial Intelligence, Machine Learning, Data Science, Cybersecurity, Linguistics, Psychology, or a related field, or equivalent professional experience
  • 10+ years of experience in threat modeling and simulation, prompt generation and analysis, novel testing, and reporting and improvement
  • Demonstrated hands-on experience, portfolio work, publications, or research in prompt injection, jailbreak testing, model evaluation, adversarial machine learning, multimodal artificial intelligence safety, or generative artificial intelligence vulnerability assessment
  • Ability to travel 25-50%, on average, based on the work you do and the clients and industries/sectors you serve.
  • Limited immigration sponsorship may be available.

Preferred:

  • Doctor of Philosophy (PhD) in Computer Science, Artificial Intelligence, Machine Learning, Data Science, Cybersecurity, Linguistics, Psychology, or a related field, or equivalent professional experience
  • Specialized training or certifications in generative artificial intelligence red teaming, adversarial machine learning, artificial intelligence security, cybersecurity, responsible artificial intelligence, or artificial intelligence governance
  • Experience designing and operationalizing trust and safety testing programs for large-scale consumer platforms, including escalation workflows, issue triage, and remediation tracking
  • Experience working with product, legal, policy, and engineering stakeholders to translate risk findings into practical platform controls and governance improvements

The wage range for this role takes into account the wide range of factors that are considered in making compensation decisions including but not limited to skill sets; experience and training; licensure and certifications; and other business and organizational needs. The disclosed range estimate has not been adjusted for the applicable geographic differential associated with the location at which the position may be filled. At Deloitte, it is not typical for an individual to be hired at or near the top of the range for their role and compensation decisions are dependent on the facts and circumstances of each case. A reasonable estimate of the current range is $134,500 to $265,100.

You may also be eligible to participate in a discretionary annual incentive program, subject to the rules governing the program, whereby an award, if any, depends on various factors, including, without limitation, individual and organizational performance.

#CyberDTP27

Education:Bachelor's DegreeEmployment Type:

What Deloitte employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom