auraoneNew York, NY
Statistics Reasoning Reviewer
Statistics Reasoning Reviewer
Statistics Reasoning Reviewer
auraoneNew York, NY
yesterday
Other Scientific and Technical Consulting ServicesAll Other Professional, Scientific, and Technical ServicesOther Management Consulting Services
Apply for this role →Statistics Reasoning Reviewer is a remote review track for evaluating AI outputs across statistics reasoning, calculations, and research workflows. Reviewers grade derivations and assumptions, reproduce key results, and document the correct method so the modeling team can train on it.
Why this role matters Aura One uses scientific specialists to grade outputs the way a peer reviewer would — checking assumptions, reproducing key steps, and capturing the right method alongside the wrong one.
Responsibilities:
Review AI outputs against current statistics methods, conventions, and prior work for Statistics Reasoning Reviewer assignments.
Reproduce or sanity-check key derivations, calculations, or experimental claims.
Flag dimensional, methodological, and citation errors with structured severity tags.
Capture the corrected reasoning or worked example so the modeling team can train on it.
Adjudicate disputed answers against textbooks, papers, or community standards.
Maintain reviewer-quality scores in inter-rater calibration cycles.
Qualifications:
Graduate-level training or equivalent applied experience in statistics or a closely related field for Statistics Reasoning Reviewer work.
Hands-on experience publishing, teaching, or advising on the topic at a professional level.
Comfort applying multi-page rubrics consistently across long batches.
Clear written reasoning that cites methods, papers, or worked examples.
Reliable async availability for at least 10 hours per week.
Example tasks Reproduce a statistics derivation from a model output and flag any algebraic or dimensional errors.
Grade a model's literature summary against the cited papers and rate the citation quality.
Adjudicate a disputed answer between two reviewers using textbook methods.
Audit a 25-row batch for rubric consistency and report drift to the program lead.
Nice to have:
PhD, postdoc, or industry research experience in the topic area.
Prior work reviewing AI-assisted research tooling and its failure modes.
Multilingual fluency for non-English papers and corpora.
Skills Scientific reasoning
Method validation
Citation review
Quantitative analysis
Statistics
Formal reasoning
Proof review
Work model Remote — US-eligible. Remote · Independent specialist contractor.
Employment type:
CONTRACTOR. Applicants must be authorized to work from US.
Compensation Hourly rate confirmed after the interview process.
Also on the board Same function, level within a rung
Level
Mid
Location
New York, NY
Occupation
Statisticians
Industry
Other Scientific and Technical Consulting Services
Posted
yesterday