Senior AI Model Evaluation Scientist; LLM Benchmarks
Listed on 2026-10-07
-
Research/Development
AI Evaluation -
IT/Tech
AI Evaluation, AI Engineer (Applied/Software), Machine Learning/ ML Engineer
Cohere is seeking a role focused on evaluating next-generation AI models, developing evaluation benchmarks and infrastructure to measure LLM progress. You will build ambitious benchmarks, work with cross-functional teams to translate model feedback into trustworthy evaluations, and advance state-of-the-art methods.
This remote-friendly role offers health and dental benefits, a lunch stipend, parental leave and education stipends, plus opportunities to work across our global offices and tools
This posting is for the Senior AI Model Evaluation Scientist (LLM Benchmarks) role at Cohere, based in Seattle, WA, United States.
Please review the full job details above before applying.
If your experience matches this role, we encourage you to apply.
All applications are reviewed carefully by our team.
The position is based in Seattle, WA, United States.
This opportunity is part of our work in IT & Technology, Other.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).