$155,000 – $400,000
Listed on Sentry’s own careers site. You apply with them directly — we never stand between you and the employer.
What this role is
This role involves building evaluation infrastructure for Sentry's AI systems, designing datasets, benchmarks, and test harnesses that measure the accuracy and reliability of AI-powered debugging agents. It suits experienced software engineers who enjoy translating ambiguous AI behavior into concrete metrics and want to build foundational tools that enable faster, more confident AI development.
Our summary, not Sentry’s wording. The full posting is on their site.
Skills this role names
Log in to see which of these are already on your profile.
What they ask for
Required
- 5+ years of professional software engineering experience
- Bachelor's degree in computer science, machine learning, or related field
- Production-quality code in Python and TypeScript
- Experience building testing, evaluation, or data infrastructure for complex systems
- Experience with structured and unstructured datasets, labeling workflows, or data quality pipelines
- Familiarity with modern ML systems and evaluation techniques
Nice to have
- AI/ML infrastructure experience
- Experience evaluating LLMs or agentic systems
- Experience with AI-assisted developer tools