ARAC-Bench introduces a three-stage researcher-mimicking benchmark that scores autonomous research systems on proposal, experiment, and synthesis quality; evaluated systems reach at most 67.9 out of 100, and scores correlate with PhD rankings at 0.81 on average.
Accelerating scien- tific discovery with co-scientist.Nature, pages 1–3, 2026
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.AI 1years
2026 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
ARAC: Benchmarking Auto-Research's Alignment and Completeness on End-to-End Researchs
ARAC-Bench introduces a three-stage researcher-mimicking benchmark that scores autonomous research systems on proposal, experiment, and synthesis quality; evaluated systems reach at most 67.9 out of 100, and scores correlate with PhD rankings at 0.81 on average.