Anthropic logo

Safeguards Enforcement Lead (Cyber Harms)

Anthropic

RemoteFull timeMid level$285k – $330kPosted today
Apply with JobAssist

About the role

  • As an Enforcement Lead, you will be responsible for managing and executing enforcement actions across our products and services, with a focus on detecting and mitigating attempts to misuse Anthropic’s AI systems for malicious cyber operations
  • Your work will center on developing strategic enforcement frameworks for flagged activity related to cyberattacks, malware development, and offensive exploitation
  • Additionally, you will manage a team of Cyber Enforcement Analysts and contractors implementing this enforcement strategy
  • Safety is core to our mission, and you’ll help uphold policy enforcement so that our users can safely interact with and build on top of our products in a harmless, helpful, and honest way
  • Important context for this role: In this position you may be exposed to and engage with explicit content spanning a range of topics, including those of a violent, technical, or psychologically disturbing nature
  • This role may require responding to escalations during weekends and holidays
  • Manage a team of Cyber Enforcement Analysts and contractors, overseeing the vision of Cyber Enforcement strategy
  • Create strategies to detect and mitigate potential misuse of AI systems to facilitate cyberattacks, malware creation, exploitation tooling, and related harmful cyber operations
  • Collaborate with stakeholders regarding novel, ambiguous, or high-severity cases
  • Collaborate with the Safeguards Policy Design Team on policy gaps surfaced through real enforcement scenarios
  • Partner with Engineering and Data Science teams to ensure tooling and measurement support enforcement operations
  • Keep up to date with emerging AI policy enforcement best practices, threat actor tactics, and the evolving cyber threat landscape, using these to inform enforcement decisions

Benefits

  • Comprehensive health, dental, and vision insurance for you and your dependents
  • Inclusive fertility benefits via Carrot Fertility
  • 22 weeks of paid parental leave
  • Flexible paid time off and absence policies
  • Mental health support for you and your dependents
  • Competitive salary and equity packages
  • Optional equity donation matching at a 1:1 ratio, up to 25% of your equity grant
  • Retirement plans with competitive matching
  • Life and income protection plans
  • $500/month flexible wellness and time saver stipend
  • Commuter benefits
  • Annual education stipend
  • Home office stipends
  • Relocation support for those moving for Anthropic
  • Daily meals and snacks in the office- Proficiency in SQL and/or Python for data analysis and threat detection
  • Experience as a people manager
  • Experience identifying emerging risks and communicating findings to a diverse set of stakeholders, such as Product, Policy, Engineering, and Legal teams
  • Experience performing content review, abuse investigations, or policy enforcement at volume
  • Experience working with generative AI products, including writing effective prompts for content review and enforcement
  • Experience in cybersecurity, including knowledge of offensive techniques, exploit development, malware analysis, or vulnerability research
  • We encourage you to apply even if you do not believe you meet every single qualification
  • Understanding of the challenges involved in implementing product policies at scale, including in the content moderation space
  • Experience working with government agencies, regulated environments, or information sharing communities
  • Experience in trust & safety, abuse investigations, cybersecurity investigations, or threat intelligence in a technology or AI company
  • Experience with large language models and an understanding of how AI technology could be misused for cyber operations
  • Experience operating within abuse monitoring programs or enforcement review systems

Millions of jobs, with real people getting hired every day

20,000+
New jobs added daily
7,000,000+
Verified job listings
500,000+
Tailored applications submitted
FAQ

Questions, answered

Click "Apply with JobAssist" – we tailor your resume and application to this role and submit it for your approval.

Yes. This role at Anthropic was screened before publishing – we confirmed the employer before listing it.

The employer didn't disclose a salary range for this listing. JobAssist shows pay whenever it's available.

This position can be done from anywhere, with no in-office requirement.

Yes – every application is tailored from your profile and this job's requirements, and you can review and edit before it's sent.