Safeguards Enforcement Analyst, Cyber Harm
Menlo Ventures PortfolioRemote-Friendly, United States; San Francisco, CA | New York City, NY | Washington, DC · Remote · full time
$285,000 – $330,000
Listed on Menlo Ventures Portfolio’s own careers site. You apply with them directly — we never stand between you and the employer.
What this role is
This role involves reviewing content and enforcing policies against misuse of AI systems for cyber attacks and malware development at Anthropic. It suits cybersecurity professionals with content moderation or policy enforcement experience who want to work on AI safety at the intersection of security and trust & safety.
Our summary, not Menlo Ventures Portfolio’s wording. The full posting is on their site.
What they ask for
Required
- Cybersecurity experience including offensive techniques, exploit development, malware analysis, or vulnerability research
- Content review, abuse investigation, or policy enforcement experience at scale
- SQL and/or Python proficiency for data analysis and threat detection
- Ability to identify emerging risks and communicate findings to diverse stakeholders
- Experience with generative AI products and writing effective prompts for content review
Nice to have
- Trust & safety, abuse investigation, cybersecurity investigation, or threat intelligence experience in tech or AI
- Understanding of large language models and AI misuse for cyber operations
- Experience with abuse monitoring programs or enforcement review systems
- Understanding of challenges implementing policies at scale in content moderation
- Experience with government agencies, regulated environments, or information sharing communities