1

Trust And Safety Agent Jobs (NOW HIRING)

... ML models, or agent moderation tools. What you should have * 2+ years of people management ... a Trust & Safety domain, such as fraud, spam, scams, account compromise or other high-volume ...

Experience with trust, safety, integrity, or policy enforcement platforms and the operational ... Demonstrated ongoing AI skill development (e.g., prompt/context engineering, agent orchestration ...

Expertise in designing and scaling agent training programs, including content development ... trust/safety. PREFERRED SKILLS AND EXPERIENCE: * Demonstrated ability to collect, analyze, and ...

... Trust & Safety, and Responsible AI. You will design the evaluations that catch unsafe behavior ... Agent safety is the primary focus of this role: you will help ensure that as our systems gain the ...

... Trust & Safety, and Responsible AI. You will design the evaluations that catch unsafe behavior ... Agent safety is the primary focus of this role: you will help ensure that as our systems gain the ...

We think about agent-native design in three areas: the interfaces agents need to operate, the human ↔ agent handoff where a person supervises and trusts agent work, and trust, safety, and ...

... on agent architecture and iteration cycles • Turn loose AI behavior into scalable product ... trust, safety, and governance • Design the product systems that allow AI agents to represent ...

Expertise in designing and scaling agent training programs, including content development ... trust/safety. PREFERRED SKILLS AND EXPERIENCE: * Demonstrated ability to collect, analyze, and ...

Expertise in designing and scaling agent training programs, including content development ... trust/safety. PREFERRED SKILLS AND EXPERIENCE: * Demonstrated ability to collect, analyze, and ...

Expertise in designing and scaling agent training programs, including content development ... trust/safety. PREFERRED SKILLS AND EXPERIENCE: * Demonstrated ability to collect, analyze, and ...

... agent workflows, tool-use orchestration, stateful reasoning patterns, and multi-step automation. Governance, Trust & Safety * Implement trust, safety, and governance controls including PII handling ...

Showing results 21-40

Trust And Safety Agent information

See salary details

$38K

$81.8K

$112K

How much do trust and safety agent jobs pay per year?

As of Aug 14, 2026, the average yearly pay for trust and safety agent in the United States is $81,805.00, according to ZipRecruiter salary data. Most workers in this role earn between $62,500.00 and $98,500.00 per year, depending on experience, location, and employer.

What is a trust and safety agent?

Trust and Safety Agents are professionals responsible for ensuring the safety, security, and integrity of online platforms or services. They review user-generated content, monitor for policy violations, handle reports of abuse, and enforce community guidelines to protect users from harm. Their work often involves investigating suspicious activities, preventing fraud, and collaborating with legal or law enforcement teams as needed. Trust and Safety Agents play a crucial role in maintaining a positive and secure online environment.

What are the key skills and qualifications needed to thrive as a trust and safety agent?

To thrive as a Trust and Safety Agent, you need strong analytical abilities, attention to detail, and a background in risk management or compliance, often supported by a relevant degree. Familiarity with content moderation platforms, case management systems, and sometimes certifications in cybersecurity or data privacy are typical requirements. Excellent judgment, resilience, and effective communication are vital soft skills for handling sensitive situations and interacting with diverse stakeholders. These skills ensure the protection of users, uphold platform integrity, and help organizations quickly respond to emerging risks.

What are some common challenges trust and safety agents face when addressing user reports, and how are these typically managed?

Trust and Safety Agents often encounter challenges such as handling a high volume of sensitive or distressing user reports, distinguishing between genuine and fraudulent claims, and making nuanced decisions while adhering to company policies. To manage these, agents rely on clear escalation protocols, ongoing training in content moderation, and robust support from colleagues and mental health resources. Regular team check-ins and access to up-to-date guidelines help ensure consistent and fair decision-making, while also supporting agent well-being.
More about Trust And Safety Agent jobs

What cities are hiring for Trust And Safety Agent jobs?

Cities with the most Trust And Safety Agent job openings:

What states have the most Trust And Safety Agent jobs?

States with the most job openings for Trust And Safety Agent jobs include:

What job categories do people searching Trust And Safety Agent jobs look for?

The top searched job categories for Trust And Safety Agent jobs are:

Infographic showing various Trust And Safety Agent job openings in the United States as of August 2026, with employment types broken down into 1% As Needed, 79% Full Time, 17% Part Time, 2% Contract, and 1% Nights. Highlights an 99% Physical, and 1% Remote job distribution, with an average salary of $81,805 per year, or $39.3 per hour.

Manager, Safety Analytics

Discord

San Francisco, CA • On-site

$220K - $275K/yr

Other

This job post has expired today. Applications are no longer accepted.


Job description

Discord exists to give people the power to create space to find belonging - to talk regularly with the people they care about and build genuine relationships with friends and communities close to home or around the world.

We're looking for an Analytics Manager to lead our Scaled Abuse Countermeasures and Research (SCAR) team - the team that safeguards Discord's platform integrity. SCAR detects, analyzes, and disrupts high-volume threats through a combination of automated systems, deep research, and active incident response. This role reports to the Head of Safety Intelligence and Automation.

What You'll Be Doing

  • Lead and mentor a team of data analysts, scientists, and researchers who investigate active threats, identify platform abuse vectors, and uncover adversarial patterns.
  • Define a strategic roadmap that prioritizes and disrupts the highest impact abuse operations through structured research, rigorous analyses, and live experimentation.
  • Collaborate closely with the safety machine learning team to improve models by identifying the threat signals that translate into long-term, automated countermeasures.
  • Partner cross-functionally with Product, Engineering, Data Science, Policy, Legal, and Revenue, influencing safety-by-design decisions upstream of abuse.
  • Influence capacity toward high-impact infrastructure, such as automated rule engines, ML models, or agent moderation tools.

What you should have

  • 2+ years of people management experience leading technical teams, including engineers, data scientists, analysts, applied researchers, or equivalent.
  • 4+ years of experience working in a Trust & Safety domain, such as fraud, spam, scams, account compromise or other high-volume threats.
  • Fluency in SQL and/or Python for data investigation and pattern analysis.
  • Strong analytical problem-solving skills, with expertise in measurement and experimentation.
  • Demonstrated ability to drive automation adoption within operations or incident response workflows that increase the team's leverage and reduce manual work.
  • Excellent communication and cross-functional collaboration skills, with an ability to translate complex technical topics into clear takeaways for senior leadership.

Bonus Points

  • Experience leveraging LLMs or AI agents for incident response, investigation automation, or signal triage.
  • Experience tackling scaled abuse at large social platforms, marketplaces, or other high-volume consumer products.
  • Threat intelligence research background, including understanding of internet infrastructure and the tools and techniques attackers use.
  • A strong passion for Discord and/or gaming, and an appreciation for the communities we serve.
  • A relevant degree in Computer Science, Machine Learning, Statistics, or a related quantitative field, or equivalent practical experience.

Candidates must reside in or be willing to relocate to the San Francisco Bay Area (Alameda, Contra Costa, Marin, Napa, San Francisco, San Mateo, Santa Clara, Solano, and Sonoma counties). Relocation assistance may be available.

The US base salary range for this full-time position is $220,000 to $275,000 + equity + benefits. Our salary ranges are determined by role and level. Within the range, individual pay is determined by additional factors, including job-related skills, experience, and relevant education or training. Please note that the compensation details listed in US role postings reflect the base salary only, and do not include benefits.