Pith. sign in

ScholarPeer: A Context-Aware Multi-Agent Framework for Automated Peer Review

5 Pith papers cite this work. Polarity classification is still indexing.

5 Pith papers citing it
abstract

The exponential growth of machine learning submissions has strained the traditional peer review process, resulting in slow feedback loops for authors and an immense burden on reviewers to rigorously audit technical soundness and verify literature. To address this, we introduce ScholarPeer, a multi-agent framework designed to operationalize the rigorous auditing workflow of a senior researcher. Rather than attempting to replace human judgment, ScholarPeer serves as a co-scientist: acting as a mentor for rapid author iteration prior to submission, and as an active verification assistant that augments human reviewers. The framework structurally decouples contextualization from critique by deploying a sub-domain historian to synthesize the field's trajectory, a baseline scout to proactively hunt for omitted state-of-the-art comparisons, and a multi-aspect Q&A engine that deeply audits technical soundness-scrutinizing internal logical consistency, experimental validity, and mathematical rigor-while cross-referencing claims against top-tier academic venues. We comprehensively evaluate ScholarPeer on ~1,800 ICLR submissions spanning 2020 through 2025. Our results show that ScholarPeer achieves significant win-rates against state-of-the-art fine-tuned models and search-augmented agentic baselines.

citation-role summary

background 1

citation-polarity summary

years

2026 5

roles

background 1

polarities

background 1

representative citing papers

ScientistOne: Towards Human-Level Autonomous Research via Chain-of-Evidence

cs.AI · 2026-05-25 · unverdicted · novelty 6.0

ScientistOne introduces Chain-of-Evidence and an audit system that achieves zero hallucinated references, perfect score verification, and top method-code alignment while matching or beating human experts on five frontier tasks and generalizing to six more.

VaseMuseum: Digital Intelligent Museum for Ancient Greek Pottery

cs.CV · 2026-07-07 · conditional · novelty 4.0

VaseMuseum is a training-free multimodal agent that combines DeepResearch-style retrieval, source/response reliability control, and best-of-K reranking to improve citation validity and reduce hallucination for museum VQA on ancient Greek pottery.

AI for Auto-Research: Roadmap & User Guide

cs.AI · 2026-05-18 · conditional · novelty 4.0

AI can generate research artifacts faster than it can verify them, so across all eight lifecycle stages the credible deployment mode is human-governed collaboration rather than full autonomy.

citing papers explorer

Showing 5 of 5 citing papers.

  • From Passive Generation to Investigation: A Proactive Scientific Peer Review Agent cs.CL · 2026-06-11 · unverdicted · none · ref 19 · internal anchor

    ProReviewer is an MDP-formulated proactive peer review agent trained with SFT and RL on an 8B model that outperforms larger frontier LLMs on review quality metrics.

  • ScientistOne: Towards Human-Level Autonomous Research via Chain-of-Evidence cs.AI · 2026-05-25 · unverdicted · none · ref 6 · internal anchor

    ScientistOne introduces Chain-of-Evidence and an audit system that achieves zero hallucinated references, perfect score verification, and top method-code alignment while matching or beating human experts on five frontier tasks and generalizing to six more.

  • Can AI Review Improve Paper Drafting? An Empirical Study on 20 Computer Architecture Submissions cs.AI · 2026-05-31 · unverdicted · none · ref 14 · internal anchor

    An empirical study on 20 architecture papers finds AI reviews capture a significant fraction of human-raised issues while also surfacing additional ones, using a released tool that clusters AI comments for comparison.

  • VaseMuseum: Digital Intelligent Museum for Ancient Greek Pottery cs.CV · 2026-07-07 · conditional · none · ref 34 · internal anchor

    VaseMuseum is a training-free multimodal agent that combines DeepResearch-style retrieval, source/response reliability control, and best-of-K reranking to improve citation validity and reduce hallucination for museum VQA on ancient Greek pottery.

  • AI for Auto-Research: Roadmap & User Guide cs.AI · 2026-05-18 · conditional · none · ref 55 · internal anchor

    AI can generate research artifacts faster than it can verify them, so across all eight lifecycle stages the credible deployment mode is human-governed collaboration rather than full autonomy.