HybridSan Francisco, CA | New York City, NY | Washington, DC
Safeguards Enforcement Analyst
A
Anthropic
Зарплата
$23,800–$27,500
Формат
Hybrid
О роли
Описание вакансии
About the company
Anthropic's mission is to create reliable, interpretable, and steerable AI systems.
A quickly growing team of researchers, engineers, policy experts, and business leaders building beneficial AI systems.
About the role
As a Safeguards Analyst focusing on Integrity & Authenticity, you will build and execute enforcement workflows for products and services, focusing on detecting and mitigating misuse of AI systems for coordinated inauthentic behavior, election manipulation, and targeting, tracking, and surveillance of individuals.
Work spans AI-enabled influence operations, disinformation campaigns, election interference, and use of AI for stalking, surveillance, profiling, and targeting.
Note: may be exposed to explicit content of a political, violent, or psychologically disturbing nature. May require responding to escalations during weekends and holidays, particularly around major electoral events.
Responsibilities
Design and architect automated enforcement systems and review workflows that scale while maintaining high accuracy
Partner with Engineering and Data Science to optimize detection models
Review flagged content to drive enforcement and policy improvements
Enforce usage policies focusing on AI-enabled influence operations, coordinated inauthentic behavior, election interference, and surveillance
Support the Safeguards policy design team with detailed feedback on policy gaps
Keep up to date with AI policy enforcement best practices, threat actor tactics, and the regulatory landscape around elections, privacy, and surveillance
Requirements
Experience in trust & safety, policy enforcement, threat intelligence, or a closely related field
Experience standing up and scaling policy enforcement or content review workflows
Proficiency in SQL and/or other data analysis tools
Experience identifying emerging risks and threat actors and communicating findings to cross-functional stakeholders
Experience working with generative AI products, including writing effective prompts
Understanding of implementing product policies at scale, including content moderation
Preferred qualifications
Experience conducting cross-platform investigations into influence operations or disinformation campaigns
Familiarity with OSINT techniques and tools for threat actor tracking and network analysis
Working knowledge of privacy law, surveillance technology, or data broker ecosystems
Experience with large language models
Familiarity with election security frameworks, campaign finance law, or electoral integrity standards
Experience navigating regulatory landscapes (e.g., DSA, EU AI Act, FEC regulations, GDPR)
Experience working with election bodies, civil society organizations, or government agencies
Proficiency in Python for data analysis and automation
Experience with dark web monitoring
Conditions
Annual salary: $285,000 – $330,000 USD
Location-based hybrid policy: at least 25% of the time in one of the offices
Visa sponsorship available
Minimum education: Bachelor's degree or equivalent combination of education, training, and/or experience