×
Register Here to Apply for Jobs or Post Jobs. X

STEM Researchers - Benchmark & First-Author Research

Remote / Online - Candidates ideally in
Worcester, Worcester County, Massachusetts, 01609, USA
Listing for: Gramian Consulting
Remote/Work from Home position
Listed on 2026-09-30
Job specializations:
  • Research/Development
    AI Evaluation, Research Scientist, AI Business & Operations, Data Scientist
Salary/Wage Range or Industry Benchmark: 30 - 80 USD Hourly USD 30.00 80.00 HOUR
Job Description & How to Apply Below
About Gramian

Gramian Consultancy is a boutique consultancy specializing in IT professional services and engineering talent solutions. With a strong background in software engineering and leadership, we help companies build high-performing teams by matching them with professionals who truly fit their needs.

About Gramian

Gramian Consultancy is a boutique consultancy specializing in IT professional services and engineering talent solutions. With a strong background in software engineering and leadership, we help companies build high-performing teams by matching them with professionals who truly fit their needs.

About

The Role

We're working with a highly specialized AI research lab building new benchmarks for how frontier AI systems perform real scientific research. They are looking for PhD students and postdocs across STEM fields to help reproduce papers, validate AI-generated research, and define what "good research" should look like for AI systems. We are looking for one Fellow per STEM field to help create a new benchmark for evaluating how well frontier AI models perform real scientific research.

The standout part of this opportunity: the benchmark will be released publicly with a research paper and open evaluation set, and Fellows will be named lead / first authors.

CONTRACT

Research fellowship / contractor

LOCATIONS

Remote, global

COMMITMENT

Flexible hours

COMPENSATION

$30-$80/hour

PROCESS

Application → research review → interview

Fields

Biology, Medicine, Neuroscience, Materials Science, Chemistry, Physics, Mechanical Engineering, Chemical Engineering.

Responsibilities
  • Validate AI-generated paper reproductions, simulations, and research results.
  • Map key research areas and taxonomy within your field.
  • Define what "good research" looks like for AI systems.
  • Create benchmark tasks and evaluation standards for frontier models.
  • Contribute directly to the benchmark paper and public eval release
Requirements
  • Current PhD student or postdoc in a relevant STEM field.
  • Based at a strong research university or institute.
  • At least one published paper in your field.
  • Able to critically review research and identify methodological or technical errors.
  • Currently using AI for Science in your own research.
  • Significant hands-on use of tools such as Claude Code, Codex, or similar AI agents
Benefits
  • Be a lead / first author on a public benchmark paper.
  • Help create one of the first research benchmarks in your scientific field.
  • Work directly with a niche AI research lab and frontier AI researchers.
  • Flexible remote work designed around academic commitments
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary