Pith. sign in

hub

Webexplorer: Exploreandevolvefortraininglong-horizonwebagents.arXivpreprint

17 Pith papers cite this work. Polarity classification is still indexing.

17 Pith papers citing it

hub tools

citation-role summary

background 2 baseline 1

citation-polarity summary

years

2026 16 2025 1

representative citing papers

Can AI Agents Synthesize Scientific Conclusions?

cs.AI · 2026-06-09 · unverdicted · novelty 7.0

A new benchmark and clean-room harness show frontier AI agents reach only 0.337 factual F1 when synthesizing conclusions from scientific evidence.

ResearchMath-14K: Scaling Research-Level Mathematics via Agents

cs.CL · 2026-05-27 · unverdicted · novelty 7.0

The authors release ResearchMath-14k, the largest dataset of research-level math problems, and demonstrate that agent-filtered reasoning trajectories from open models improve fine-tuned Qwen3 models by 9.2 points on average.

Argus: Evidence Assembly for Scalable Deep Research Agents

cs.CL · 2026-05-15 · unverdicted · novelty 6.0 · 2 refs

Argus coordinates a Navigator and multiple Searchers via an evidence graph for deep research, reporting average gains of 5.5 points with one Searcher and 12.7 points with eight parallel Searchers across eight benchmarks, reaching 86.2 on BrowseComp with 64 Searchers.

Learning to Retrieve from Agent Trajectories

cs.IR · 2026-03-30 · conditional · novelty 6.0

Retrievers trained on agent trajectories via the LRAT framework improve evidence recall, task success, and efficiency in agentic search benchmarks.

citing papers explorer

Showing 17 of 17 citing papers.