Skip to main content
Anthropic
Anthropic

Safeguards Enforcement Analyst, Integrity & Authenticity

RemoteUnited States, United Kingdom only
Published
Experience
Mid
Employment
Full-time
Company size
Startup
$285k–$330k/yr
Check eligibility

Open to US, GB only. Set where you work from to check your eligibility.

No BS summary

As a Safeguards Analyst focusing on Integrity & Authenticity, you will be responsible for building and executing enforcement workflows for our products and services, with a focus on detecting and mitigating attempts to misuse Anthropic's AI systems for coordinated inauthentic behavior, election manipulation, and targeting, tracking, and surveillance of individuals.

Required skills

SQLPython

Optional skills

open-source intelligence (OSINT)Python

What you'll do

  • Design and architect automated enforcement systems and review workflows that scale effectively while maintaining high accuracy
  • Partner with Engineering and Data Science teams to optimize detection models for policy violations and automated enforcement systems
  • Review flagged content to drive enforcement and policy improvements
  • Enforce usage policies with a focus on detecting and mitigating AI-enabled influence operations, coordinated inauthentic behavior, election interference, and targeting, tracking, or surveillance of individuals and groups
  • Support the Safeguards policy design team by providing detailed feedback on policy gaps based on real enforcement scenarios
  • Keep up to date with emerging AI policy enforcement best practices, evolving threat actor tactics, and the regulatory landscape around elections, privacy, and surveillance, using these to inform our decision-making and workflows

What they require

  • Experience in trust & safety, policy enforcement, threat intelligence, or a closely related field with a focus on one or more of: influence operations, disinformation, coordinated inauthentic behavior, election integrity, or privacy and surveillance harms
  • Experience standing up and scaling policy enforcement or content review workflows
  • Proficiency in SQL and/or other data analysis tools to draw insights from large datasets
  • Experience identifying emerging risks and threat actors, and communicating findings to a diverse set of stakeholders, such as Product, Policy, Engineering, and Legal teams
  • Experience working with generative AI products, including writing effective prompts for content review and enforcement
  • Understanding of the challenges involved in implementing product policies at scale, including in the content moderation space
  • Bachelor’s degree or an equivalent combination of education, training, and/or experience
  • A field relevant to the role as demonstrated through coursework, training, or professional experience
  • Years of experience required will correlate with the internal job level requirements for the position
  • In this position you may be exposed to and engage with explicit content spanning a range of topics, including those of a political, violent, or psychologically disturbing nature.
  • This role may require responding to escalations during weekends and holidays, particularly around major electoral events.

Benefits

  • competitive compensation and benefits
  • optional equity donation matching
  • generous vacation and parental leave
  • flexible working hours
  • a lovely office space in which to collaborate with colleagues

American artificial intelligence corporation

🇺🇸 United StatesArtificial IntelligenceStartupanthropic.com/

What people say about this company

5.0/ 5

Details

Visa sponsorshipYes
$285k–$330k/yr