Causal Reasoning Model Evaluation Specialist
Listed on 2026-10-04
-
Research/Development
AI Evaluation, Data Annotation/ AI Labeling, Research Analyst, Data Scientist -
IT/Tech
AI Evaluation, Data Annotation/ AI Labeling, Data Scientist
Causal Reasoning Model Evaluation Specialist is a remote review track for evaluating AI outputs across causal reasoning model evaluation research review reasoning, calculations, and research workflows. Reviewers grade derivations and assumptions, reproduce key results, and document the correct method so the modeling team can train on it.
Category:
Math, Reasoning & Formal Methods
· Pay: $50 / hr
·
Location:
Remote — US-eligible
· Contractor
Causal Reasoning Model Evaluation Specialist is a remote review track for evaluating AI outputs across causal reasoning model evaluation research review reasoning, calculations, and research workflows.
About the roleCausal Reasoning Model Evaluation Specialist is a remote review track for evaluating AI outputs across causal reasoning model evaluation research review reasoning, calculations, and research workflows.
Causal Reasoning Model Evaluation research review models live or die on whether their derivations actually hold up under scrutiny.
Judge mathematical reasoning and theorem proving. Long proof chains included.
- Review AI outputs against current causal reasoning model evaluation research review methods, conventions, and prior work for Causal Reasoning Model Evaluation Specialist assignments.
- Reproduce or sanity-check key derivations, calculations, or experimental claims.
- Flag dimensional, methodological, and citation errors with structured severity tags.
- Capture the corrected reasoning or worked example so the modeling team can train on it.
Track STEM research review Work model Remote
· Independent specialist contractor Compensation Hourly rate confirmed after the interview process. Eligible from US
- Graduate-level training or equivalent applied experience in causal reasoning model evaluation research review or a closely related field for Causal Reasoning Model Evaluation Specialist work.
- Hands-on experience publishing, teaching, or advising on the topic at a professional level.
- Comfort applying multi-page rubrics consistently across long batches.
- Clear written reasoning that cites methods, papers, or worked examples.
- Reliable async availability for at least 10 hours per week.
- Reproduce a causal reasoning model evaluation research review derivation from a model output and flag any algebraic or dimensional errors.
- Grade a model's literature summary against the cited papers and rate the citation quality.
- Adjudicate a disputed answer between two reviewers using textbook methods.
- Audit a 25-row batch for rubric consistency and report drift to the program lead.
- PhD, postdoc, or industry research experience in the topic area.
- Prior work reviewing AI-assisted research tooling and its failure modes.
- Multilingual fluency for non-English papers and corpora.
Hourly rate confirmed after the interview process.
Expected arrangement: contractor , with program-defined task volume and review pacing. Placement depends on current program demand and reviewer confirmation.
- Scientific reasoning
- Method validation
- Citation review
- Quantitative analysis
- Causal Reasoning Model Evaluation research review
- Frontier evaluation
- Rubric calibration
- Failure analysis
- Causal
- Reasoning
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).