Frontier AI Evaluation Scientist: Research & Benchmarks
Listed on 2026-09-30
-
Research/Development
AI Evaluation, Research Scientist
Hidden Jobs seeks a researcher to advance how frontier AI systems are evaluated, trained, and improved. You will tackle benchmark gaps, generate synthetic STEM data, and rigorously evaluate model reliability and calibration.
The role emphasizes hypothesis generation, experimentation, and publication across frontier STEM evaluation and data generation. The ideal candidate holds a PhD or equivalent and demonstrates strong experimental design, quantitative analysis, and knowledge of modern LLMs.
We are looking to fill the Frontier AI Evaluation Scientist:
Research & Benchmarks position at Hidden Jobs in United States.
Full responsibilities and requirements are described in the listing above.
Learn more about the Frontier AI Evaluation Scientist:
Research & Benchmarks role in the description above.
We appreciate your interest in this position.
Join Hidden Jobs and contribute to our ongoing work.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).