REVIEW 4 cited by
Ai2 Scholar QA: Organized Literature Synthesis with Attribution
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Retrieval-augmented generation is increasingly effective in answering scientific questions from literature, but many state-of-the-art systems are expensive and closed-source. We introduce Ai2 Scholar QA, a free online scientific question answering application. To facilitate research, we make our entire pipeline public: as a customizable open-source Python package and interactive web app, along with paper indexes accessible through public APIs and downloadable datasets. We describe our system in detail and present experiments analyzing its key design decisions. In an evaluation on a recent scientific QA benchmark, we find that Ai2 Scholar QA outperforms competing systems.
Forward citations
Cited by 4 Pith papers
-
Idea2Plan: Exploring AI-Powered Research Planning
GPT-5 scores 62% on Idea2Plan, a new benchmark that grades AI-generated research plans against rubrics built from 200 post-cutoff ICML 2025 papers — the strongest result, with substantial headroom.
-
DeepWriter: A Fact-Grounded Multimodal Writing Assistant Based On Offline Knowledge Base
DeepWriter is a pipeline that generates fact-grounded, multimodal long-form documents from a curated offline knowledge base, but the reported evidence does not support the abstract's claim that it surpasses baselines.
-
Compare: A Framework for Scientific Comparisons
Compare is a RAG-based system that generates qualitative, citation-supported comparisons of scientific contributions at institution and publication granularity.
-
Literature-Grounded Novelty Assessment of Scientific Ideas
An LLM retrieval-augmented system that judges whether a research idea is novel by comparing it against reranked relevant papers and expert-labeled examples.
Discussion (0). Continue with ORCID to comment.