Pith. sign in

Paper Citation Record · LEDGER

ResearcherBench: Evaluating Deep AI Research Systems on the Frontiers of Scientific Inquiry

As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2507.16280.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.16280 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T14:43:55.822569Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T05:47:41.419232Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 093104ea-5231-4364-8110-36087bd5a93a · inbound

SafeSearch: Automated Red-Teaming of LLM-Based Search Agents cites this paper.

SafeSearch: Automated Red-Teaming of LLM-Based Search Agents ResearcherBench: Evaluating Deep AI Research Systems on the Frontiers of Scientific Inquiry

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-04T14:43:55.822569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:43:55.822569Z digest=sha256:6b753bf894cb2cdace5a396d74bc3e210788b98278a49cc0a2804487b1a7af4b

Observation 208889a2-d2bb-4ee8-a661-03b1e730bcc3 · inbound

AgencyBench: Benchmarking the Frontiers of Autonomous Agents in 1M-Token Real-World Contexts cites this paper.

AgencyBench: Benchmarking the Frontiers of Autonomous Agents in 1M-Token Real-World Contexts ResearcherBench: Evaluating Deep AI Research Systems on the Frontiers of Scientific Inquiry

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-16T14:12:58.706797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T14:11:14.123806Z digest=sha256:3c996502d4c1eaddff5bba034dca31a5955a4afced64b9ee2a786a8ccd656986

Observation acd4b857-3906-4289-97d5-7ed911cbef95 · inbound

FIRE-Bench: Evaluating AI Agents on the Rediscovery of Scientific Insights cites this paper.

FIRE-Bench: Evaluating AI Agents on the Rediscovery of Scientific Insights ResearcherBench: Evaluating Deep AI Research Systems on the Frontiers of Scientific Inquiry

Reference 8

Resolution
malformed identifier
no resolver link, observed 2026-08-03T05:17:18.399197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:17:18.399197Z digest=sha256:c0df8f3bff52786cabb2f1957b8f39305d4ae556b606e040826140222faf9ac8

Observation 16eec53e-f625-49f9-8cf6-884348be1f2f · inbound

SciPredict: Can LLMs Predict the Outcomes of Scientific Experiments in Natural Sciences? cites this paper.

SciPredict: Can LLMs Predict the Outcomes of Scientific Experiments in Natural Sciences? ResearcherBench: Evaluating Deep AI Research Systems on the Frontiers of Scientific Inquiry

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:36:03.520818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:55:34.768853Z digest=sha256:c0c35ce9b80ea1abef4ddb07c01aba81938f9e5c67a07f7df5852bd7977cc5e6

Observation e0e6c591-fa0f-4536-a350-66d34f38164b · inbound

Personalized Deep Research: A User-Centric Framework, Dataset, and Hybrid Evaluation for Knowledge Discovery cites this paper.

Personalized Deep Research: A User-Centric Framework, Dataset, and Hybrid Evaluation for Knowledge Discovery ResearcherBench: Evaluating Deep AI Research Systems on the Frontiers of Scientific Inquiry

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:46:31.422150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-12T04:55:22.596920Z digest=sha256:d569fbe117d370a6d8f09f51c9af36e7ee420831c2be65ee0c4db08320b0ce8a

Observation d23b6d33-4989-4f02-94dc-f03088c64355 · inbound

ASPI: Seeking Ambiguity Clarification Amplifies Prompt Injection Vulnerability in LLM Agents cites this paper.

ASPI: Seeking Ambiguity Clarification Amplifies Prompt Injection Vulnerability in LLM Agents ResearcherBench: Evaluating Deep AI Research Systems on the Frontiers of Scientific Inquiry

Reference 85

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T23:57:53.358041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-19T23:54:15.987953Z digest=sha256:65f5032c25c02872baf70fdc6393d6d7a94a65808f2192d5ec082b0cbd2b56ad

Observation 472f7f6e-c3db-4821-af0b-13d83a010732 · inbound

Evaluating Deep Research Agents on Expert Consulting Work: A Benchmark with Verifiers, Rubrics, and Cognitive Traps cites this paper.

Evaluating Deep Research Agents on Expert Consulting Work: A Benchmark with Verifiers, Rubrics, and Cognitive Traps ResearcherBench: Evaluating Deep AI Research Systems on the Frontiers of Scientific Inquiry

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-20T12:33:16.776361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-20T12:32:22.536994Z digest=sha256:616407e9a34504749b8fdb23d49a9c8da9544511fcc26b76defe9acc8d119987

Observation 7430d0fd-6313-4fa4-b62f-e027314db454 · inbound

AutoResearch AI: Towards AI-Powered Research Automation for Scientific Discovery cites this paper.

AutoResearch AI: Towards AI-Powered Research Automation for Scientific Discovery ResearcherBench: Evaluating Deep AI Research Systems on the Frontiers of Scientific Inquiry

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:50:21.711072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T04:46:43.679185Z digest=sha256:91b66a1f9c20dfd4cdfbe30a301c3c66823f411c2b280e5ad2a27c320393842b

Observation 3eed8187-4455-40aa-8cff-4db7a500ca1a · inbound

ScientistOne: Towards Human-Level Autonomous Research via Chain-of-Evidence cites this paper.

ScientistOne: Towards Human-Level Autonomous Research via Chain-of-Evidence ResearcherBench: Evaluating Deep AI Research Systems on the Frontiers of Scientific Inquiry

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-06-29T21:23:59.076511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-29T21:19:03.281629Z digest=sha256:52e3ede3a8e086026e8d57dce906be066a674e8d67034fc2c0d48ccb1b8e0840

Observation 30b91610-69e1-44dd-b054-56c5738fbe1d · inbound

Can AI Agents Synthesize Scientific Conclusions? cites this paper.

Can AI Agents Synthesize Scientific Conclusions? ResearcherBench: Evaluating Deep AI Research Systems on the Frontiers of Scientific Inquiry

Reference 133

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:47:41.420928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-27T13:05:07.718882Z digest=sha256:1fe0768562cec0197a682fa71b94fe306122294deee985c32b1e4c3362d0283b