Biochemist AI Benchmark Tasks & Prompt Design
Listed on 2026-09-25
-
Research/Development
AI Evaluation, Data Annotation/ AI Labeling
Mercor is hiring PhD and Master’s scientists to author AI evaluation tasks for Sci Code. You will design original, executable research problems that today’s frontier models cannot solve.
Source your own material, write scientific prompts, build grading criteria, and calibrate against frontier models. Tasks ship only when strong models fail it more often than they succeed, with Docker runs and PR-based quality checks.
We are currently recruiting a Biochemist for AI Benchmark Tasks & Prompt Design for our team in San Francisco, CA, United States.
We would love to welcome a new Biochemist for AI Benchmark Tasks & Prompt Design to our organisation in San Francisco, CA, United States.
For the Biochemist for AI Benchmark Tasks & Prompt Design position at Mercor, we are reviewing applications now.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).