$320,000 – $485,000
Listed on Menlo Ventures Portfolio’s own careers site. You apply with them directly — we never stand between you and the employer.
What this role is
This role focuses on building safety and oversight systems to detect unwanted AI model behaviors and prevent misuse across Anthropic's products and API. It suits engineers with a background in abuse detection, trust & safety systems, or distributed infrastructure who want to work on real-time safeguarding at scale.
Our summary, not Menlo Ventures Portfolio’s wording. The full posting is on their site.
Skills this role names
Log in to see which of these are already on your profile.
What they ask for
Required
- Bachelor's degree in Computer Science, Software Engineering, or equivalent experience
- Proficiency in Python and TypeScript
- Ability to work across the stack
- Strong communication skills
Nice to have
- 8+ years of software engineering experience
- Experience with integrity, spam, fraud, or abuse detection and mitigation
- Experience building trust and safety detection mechanisms for AI/ML systems
- Experience with prompt engineering, jailbreak attacks, and adversarial inputs
- Experience building custom internal tooling with operational teams