Pith. sign in

Paper Citation Record · LEDGER

EXP-Bench: Can AI Conduct AI Research Experiments?

As of 3 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2505.24785.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.24785 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-03T06:30:56.289259+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-14T16:55:37.417202Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-10T12:15:01.137692Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 009a013c-5b33-421d-954b-5d6fae5e99c6 · inbound

AstaBench: Rigorous Benchmarking of AI Agents with a Scientific Research Suite cites this paper.

AstaBench: Rigorous Benchmarking of AI Agents with a Scientific Research Suite EXP-Bench: Can AI Conduct AI Research Experiments?

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T04:25:52.289457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-18T04:22:42.678252Z digest=sha256:88758b3535936a0a2c9a21379b7c9cf7349a5e6d5f8897b809a9ddd6384b6227

Observation c3e5a5ad-7bab-4e13-8f47-16a6ab38472c · inbound

SciPredict: Can LLMs Predict the Outcomes of Scientific Experiments in Natural Sciences? cites this paper.

SciPredict: Can LLMs Predict the Outcomes of Scientific Experiments in Natural Sciences? EXP-Bench: Can AI Conduct AI Research Experiments?

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:36:03.727079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-10T15:55:34.768853Z digest=sha256:97516868321f5676ecaf9ceb227ce6899401b5ad742bb13e01d4c14bdfb36a21

Observation bc9fa853-ecf8-495b-8a5f-637b2be7f596 · inbound

Experiment-as-Code Labs: A Declarative Stack for AI-Driven Scientific Discovery cites this paper.

Experiment-as-Code Labs: A Declarative Stack for AI-Driven Scientific Discovery EXP-Bench: Can AI Conduct AI Research Experiments?

Reference 112

Resolution
metadata mismatch
arxiv_id, observed 2026-05-08T20:04:07.156035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-08T17:28:41.217810Z digest=sha256:184be81eeb0e3c4de26a28621508440294098eda9f824160aa5ff02499df0b3e

Observation 492e38b0-fd08-4f14-9fb6-2eb49ec8adc4 · inbound

Experiment-as-Code Labs: A Declarative Stack for AI-Driven Scientific Discovery cites this paper.

Experiment-as-Code Labs: A Declarative Stack for AI-Driven Scientific Discovery EXP-Bench: Can AI Conduct AI Research Experiments?

Reference 112

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T00:13:52.165692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-21T00:13:07.546472Z digest=sha256:4a0b5acba4cde8e389e614f09820525c9a477acbfd0068e7c43f032c6b50fb37

Observation ace6bbb8-7263-4378-a8ef-f2e2d41412a4 · inbound

ASPI: Seeking Ambiguity Clarification Amplifies Prompt Injection Vulnerability in LLM Agents cites this paper.

ASPI: Seeking Ambiguity Clarification Amplifies Prompt Injection Vulnerability in LLM Agents EXP-Bench: Can AI Conduct AI Research Experiments?

Reference 74

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T23:57:53.470935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-19T23:54:15.987953Z digest=sha256:f7d6d307a76c68a5b75abac2a4c46d520e2f05dd7d34410cc5a708dd6e173a28

Observation 32acf25d-6524-48bc-a5c9-548105cc0539 · inbound

AutoResearch AI: Towards AI-Powered Research Automation for Scientific Discovery cites this paper.

AutoResearch AI: Towards AI-Powered Research Automation for Scientific Discovery EXP-Bench: Can AI Conduct AI Research Experiments?

Reference 115

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:50:21.791214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-25T04:46:43.679185Z digest=sha256:c825eca584dd2506a30aa168daf18baeb2dabda855c3d2201c8abf595b2f05e4

Observation 5eefbca1-1157-41dd-b023-9bd5f8764c0c · inbound

ScientistOne: Towards Human-Level Autonomous Research via Chain-of-Evidence cites this paper.

ScientistOne: Towards Human-Level Autonomous Research via Chain-of-Evidence EXP-Bench: Can AI Conduct AI Research Experiments?

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-06-29T21:23:59.101107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-06-29T21:19:03.281629Z digest=sha256:8d9b8b6546a64f46ae0dd16aedffe096c68d1f24bdc476259f1c7690a4fbafe7

Observation 4e858602-2958-4041-b6e6-0e5c84b2713a · inbound

InquiTree: Evaluating AI Agents in the Scientific Inquiry Loop with Paper-Derived Research Trees cites this paper.

InquiTree: Evaluating AI Agents in the Scientific Inquiry Loop with Paper-Derived Research Trees EXP-Bench: Can AI Conduct AI Research Experiments?

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T04:07:37.249691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-06-27T14:06:59.471772Z digest=sha256:2ab1d123784ffe1e80a5327436950d781ca57d529dc4e3976aedabe5487c4c48

Observation 1516a818-5c06-4ffd-a89f-061ebeb0d5c2 · inbound

One Reflection Is Not Enough: Self-Correcting Autonomous Research via Multi-Hypothesis Failure Attribution cites this paper.

One Reflection Is Not Enough: Self-Correcting Autonomous Research via Multi-Hypothesis Failure Attribution EXP-Bench: Can AI Conduct AI Research Experiments?

Reference 94

Resolution
verified exact
arxiv_id, observed 2026-07-01T05:45:25.135864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-07-01T05:44:51.299439Z digest=sha256:a5fa85e4aec32f728bc314d4903f8fa71cf46d0960a2019fe787a342dba1e419

Observation e8d3914a-2380-4743-8ddc-33e6b532b071 · inbound

FARS: A Fully Automated Research System Deployed at Scale cites this paper.

FARS: A Fully Automated Research System Deployed at Scale EXP-Bench: Can AI Conduct AI Research Experiments?

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:35:42.740047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-07-01T05:18:53.840963Z digest=sha256:8105acc8939eeed31b4417856121b3dffbbbb20818211a5c0561d65f8bde57bd

Observation 2bb3d404-5ad1-4dba-baf1-127af5b0d7f1 · inbound

FARS: A Fully Automated Research System Deployed at Scale cites this paper.

FARS: A Fully Automated Research System Deployed at Scale EXP-Bench: Can AI Conduct AI Research Experiments?

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-14T16:55:37.417202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T16:55:37.417202Z digest=sha256:91efba374b25fb9a08e7c07871dec0bddcef52f7a9ceaefd62e7de78e7766a8e