Skip to main content
CareerApp

Senior Research Scientist, Model Evaluation

Cohere

Toronto, San Francisco, Canada, Seattle, London, United States, New York · Remote · full time · Senior

CA$250,000 – CA$535,000

Listed on Cohere’s own careers site. You apply with them directly — we never stand between you and the employer.

What this role is

This role focuses on developing evaluation methods and benchmarks to measure large language model capabilities. It suits researchers who enjoy building prototypes, working with data quality, and advancing measurement science for AI systems.

Our summary, not Cohere’s wording. The full posting is on their site.

Skills this role names

Log in to see which of these are already on your profile.

What they ask for

Required

  • Rapid prototyping of LLM capability demonstrations
  • Extensive data and output review for quality assurance
  • Rigorous measurement of AI capabilities
  • Strong software engineering skills

Turn on analytics and we load Google Analytics: Google gets the pages you open and what you do here — searches, jobs you view, jobs you apply to — and sets its own cookies. Leave it off and the only cookies we set are your login, your theme, and this answer. Accept All also records a yes to advertising, which nothing uses yet. Privacy Policy.