Safeguards Enforcement Lead, Cyber Harms

Anthropic
Washington +2 more
Remote

Who this role is best for

Cybersecurity professionals with team leadership experience and technical expertise in AI misuse detection will find this strategic enforcement role in a mission-driven AI company.

Best fit for

  • Candidates with team leadership experience and deep cybersecurity knowledge, especially in offensive techniques and AI misuse.
    — “Experience as a people manager
  • Individuals who have managed high-volume policy enforcement and abuse investigations in AI or tech environments.
    — “Experience performing content review, abuse investigations, or policy enforcement at volume
  • Professionals who understand the intersection of AI and cyber threats, particularly in generative AI contexts.
    — “Experience working with generative AI products, including writing effective prompts for content review and enforcement

Things to consider

  • Exposure to explicit and potentially disturbing content is a direct requirement of the role.
    — “you may be exposed to and engage with explicit content spanning a range of topics
  • This role may require working during weekends and holidays due to escalation needs.
    — “responding to escalations during weekends and holidays
  • Candidates must be prepared to spend at least 25% of their time in office locations.
    — “expect all staff to be in one of our offices at least 25% of the time

How to stand out

  • Highlight experience in detecting AI-driven cyber threats and mitigating them through policy and enforcement.
    — “developing strategic enforcement frameworks for flagged activity related to cyberattacks
  • Emphasize your ability to translate complex technical findings into actionable insights for cross-functional teams.
    — “communicating findings to a diverse set of stakeholders
  • Showcase your background in managing enforcement teams and executing large-scale policy strategies.
    — “managing a team of Cyber Enforcement Analysts and contractors implementing this enforcement strategy
  • Demonstrate your experience with AI misuse scenarios and how they intersect with cybersecurity.
    — “understanding of how AI technology could be misused for cyber operations
  • Include specific examples of content review and prompt engineering for generative AI systems.
    — “writing effective prompts for content review and enforcement
Pace · Fast PacedCollaboration · HighAutonomy · HighDecision Impact · Company

Derived from job-description analysis by Serendipath's career intelligence engine.

What success looks like

  • Developed strategic enforcement frameworks for flagged activity related to cyberattacks
  • Managed a team of Cyber Enforcement Analysts and contractors
Typical background
Experience as a people managerExperience in cybersecurityExperience performing content review, abuse investigations, or policy enforcement at volume

Skills & requirements

Required

CybersecurityTeam ManagementPolicy EnforcementData AnalysisThreat Detection

Preferred

Trust & SafetyAbuse InvestigationsCybersecurity InvestigationsThreat Intelligence

Stack & domain

CybersecurityCyberattacksMalware CreationExploitation ToolingCyber OperationsAI SystemsAI Policy EnforcementSQLPythonData AnalysisThreat DetectionContent ReviewAbuse InvestigationsPolicy EnforcementProduct PoliciesContent ModerationGovernment AgenciesRegulated EnvironmentsInformation Sharing CommunitiesTeam ManagementStrategic PlanningCollaborationCommunicationProblem-solvingDecision MakingRisk IdentificationStakeholder EngagementTechnical WritingPolicy DevelopmentLegal ComplianceTechnical Expertise

About the role

Original posting from Anthropic via Greenhouse

About Anthropic

Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems.

About the role

As an Enforcement Lead, you will be responsible for managing and executing enforcement actions across our products and services, with a focus on detecting and mitigating attempts to misuse Anthropic's AI systems for malicious cyber operations. Your work will center on developing strategic enforcement frameworks for flagged activity related to cyberattacks, malware development, and offensive exploitation. Additionally, you will manage a team of Cyber Enforcement Analysts and contractors implementing this enforcement strategy. 

Safety is core to our mission, and you'll help uphold policy enforcement so that our users can safely interact with and build on top of our products in a harmless, helpful, and honest way.

Important context for this role: In this position you may be exposed to and engage with explicit content spanning a range of topics, including those of a violent, technical, or psychologically disturbing nature. This role may require responding to escalations during weekends and holidays.

Key responsibilities

Manage a team of Cyber Enforcement Analysts and contractors, overseeing the vision of Cyber Enforcement strategy 

Create strategies to detect and mitigate potential misuse of AI systems to facilitate cyberattacks, malware creation, exploitation tooling, and related harmful cyber operations

Collaborate with stakeholders regarding novel, ambiguous, or high-severity cases

Collaborate with the Safeguards Policy Design Team on policy gaps surfaced through real enforcement scenarios

Partner with Engineering and Data Science teams to ensure tooling and measurement support enforcement operations.

Keep up to date with emerging AI policy enforcement best practices, threat actor tactics, and the evolving cyber threat landscape, using these to inform enforcement decisions

Minimum qualifications

Experience as a people manager

Experience in cybersecurity, including knowledge of offensive techniques, exploit development, malware analysis, or vulnerability research

Experience performing content review, abuse investigations, or policy enforcement at volume

Proficiency in SQL and/or Python for data analysis and threat detection

Experience identifying emerging risks and communicating findings to a diverse set of stakeholders, such as Product, Policy, Engineering, and Legal teams

Experience working with generative AI products, including writing effective prompts for content review and enforcement

Preferred qualifications

Experience in trust & safety, abuse investigations, cybersecurity investigations, or threat intelligence in a technology or AI company

Experience with large language models and an understanding of how AI technology could be misused for cyber operations

Experience operating within abuse monitoring programs or enforcement review systems

Understanding of the challenges involved in implementing product policies at scale, including in the content moderation space

Experience working with government agencies, regulated environments, or information sharing communities

The annual compensation range for this role is listed below. 

For sales roles, the range provided is the role’s On Target Earnings ("OTE") range, meaning that the range includes both the sales commissions/sales bonuses target and annual base salary for the role.

Annual Salary:$285,000—$330,000 USDLogistics

Minimum education: Bachelor’s degree or an equivalent combination of education, training, and/or experience

Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience

Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position

Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices.

Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this.

We encourage you to apply even if you do not believe you meet every single qualification. Not all strong candidates will meet every single qualification as listed.  Research shows that people who identify as being from underrepresented groups are more prone to experiencing imposter syndrome and doubting the strength of their candidacy, so we urge you not to exclude yourself prematurely and to submit an application if you're interested in this work. We think AI systems like the ones we're building have enormous social and ethical implications. We think this makes representation even more important, and we strive to include a range of diverse perspectives on our team.

Your safety matters to us. To protect yourself from potential scams, remember that Anthropic recruiters only contact you from @anthropic.com email addresses. In some cases, we may partner with vetted recruiting agencies who will identify themselves as working on behalf of Anthropic. Be cautious of emails from other domains. Legitimate Anthropic recruiters will never ask for money, fees, or banking information before your first day. If you're ever unsure about a communication, don't click any links—visit anthropic.com/careers directly for confirmed position openings.

How we're different

We believe that the highest-impact AI research will be big science. At Anthropic we work as a single cohesive team on just a few large-scale research efforts. And we value impact — advancing our long-term goals of steerable, trustworthy AI — rather than work on smaller and more specific puzzles. We view AI research as an empirical science, which has as much in common with physics and biology as with traditional efforts in computer science. We're an extremely collaborative group, and we host frequent research discussions to ensure that we are pursuing the highest-impact work at any given time. As such, we greatly value communication skills.

The easiest way to understand our research directions is to read our recent research. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences.

Come work with us!

Anthropic is a public benefit corporation headquartered in San Francisco. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues. Guidance on Candidates' AI Usage: Learn about our policy for using AI in our application process.

Source: Anthropic careers (Greenhouse)

Similar roles