AI Benchmark Architect Scientific Computing
Listed on 2026-09-17
-
IT/Tech
Data Scientist, AI Evaluation, AI Business & Operations
Mercor is seeking researchers to author AI evaluation tasks and original, executable problems for frontier models. You will source material, write prompts, and define grading criteria across subdomains with a focus on two areas in mathematics. Engagement is six weeks, part-time, with start date immediate.
Experience with Python or R for scientific computing and familiarity with Git/Git Hub and Docker workflows are required to ensure reproducible runs and automated quality checks.
For the AI Benchmark Architect for Scientific Computing position at Obsidian, we are reviewing applications now.
Learn more about the AI Benchmark Architect for Scientific Computing role in the description above.
We appreciate your interest in this position.
Join Obsidian and contribute to our ongoing work.
Take a moment to read everything above and see whether this role is right for you.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).