💼Jobs 📝Blog 🧮Salary Calc 🌍Cost of Living 📋Tax Guide
👤Sign In Post a Job — from $99
Anthropic

Safeguards Enforcement Lead, Cyber Harms

NewOtherContractLead
Location
Worldwide
Job Type
Contract
Experience
Lead
Apply Now

Job Description

ABOUT THE ROLE Anthropic is building reliable, interpretable, and steerable AI systems that benefit users and society. The Safeguards Enforcement Lead, Cyber Harms will spearhead the organization’s response to malicious use of our generative AI products. In this role you will shape and execute enforcement strategies that detect, investigate, and mitigate cyberattacks, malware development, and other harmful operations that leverage our technology. You will lead a team of Cyber Enforcement Analysts and contractors, working closely with policy, engineering, data science, and legal partners to ensure that Anthropic’s products remain safe and trustworthy for developers and end users alike. WHAT YOU'LL DO You will develop and refine enforcement frameworks for cyber‑related misuse of AI, translating policy into actionable detection and mitigation tactics. You will manage a high‑performing team, setting vision, priorities, and metrics for cyber‑enforcement operations. You will collaborate with stakeholders on high‑severity or ambiguous cases, ensuring that decisions are informed by the latest threat intelligence and policy gaps. You will partner with engineering and data science to build tooling and measurement systems that support scalable enforcement. You will stay current on emerging AI policy best practices, threat actor tactics, and the evolving cyber threat landscape, feeding insights back into policy and product design. You will oversee the review of content and abuse investigations at volume, ensuring consistent application of enforcement rules across all products. WHAT YOU'LL NEED You have proven experience managing people in a fast‑moving tech environment, ideally within a safety or trust & safety function. You possess deep knowledge of offensive cybersecurity techniques, including exploit development, malware analysis, or vulnerability research. You have a track record of performing content review, abuse investigations, or policy enforcement at scale. You are proficient in SQL and Python for data analysis, threat detection, and automation. You have experience identifying emerging risks and communicating findings to product, policy, engineering, and legal stakeholders. You have worked with generative AI products, crafting effective prompts for content review and enforcement. Preferred experience includes working in a technology or AI company’s abuse monitoring program, familiarity with large language models and their misuse potential, and engagement with government agencies or information‑sharing communities. WHY REMOTE Anthropic embraces a distributed, asynchronous culture that empowers teams to collaborate across time zones while maintaining high productivity. Remote work enables you to balance professional responsibilities with personal commitments, fostering a healthier work‑life rhythm. The company provides flexible scheduling, allowing you to respond to high‑severity incidents outside of traditional hours while ensuring that core team interactions remain timely and effective. By working remotely, you gain access to a global talent pool and the ability to contribute to Anthropic’s mission from anywhere in the world. BENEFITS Anthropic offers a competitive total compensation package that includes a base salary in the range of $285,000 to $330,000, along with performance‑based bonuses. Comprehensive health, dental, and vision coverage is provided for employees and their families. Generous paid time off, including paid holidays and personal days, supports recovery and personal growth. A remote work stipend covers home‑office setup and internet expenses. A dedicated learning budget and access to industry conferences, courses, and certifications support continuous professional development. Anthropic also offers equity, a 401(k) plan with company match, and a supportive culture that values diversity, equity, and inclusion.