Pith. sign in

Paper Citation Record · LEDGER

SciReplicate-Bench: Benchmarking LLMs in Agent-driven Algorithmic Reproduction from Research Papers

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2504.00255.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.00255 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T10:55:15.114201Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T13:08:07.735939Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 4acaecfd-fe2e-4187-aa75-152e1afdd6c7 · inbound

From LLM Reasoning to Autonomous AI Agents: A Comprehensive Review cites this paper.

From LLM Reasoning to Autonomous AI Agents: A Comprehensive Review SciReplicate-Bench: Benchmarking LLMs in Agent-driven Algorithmic Reproduction from Research Papers

Reference 110

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:57:38.433500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T02:57:37.873567Z digest=sha256:5ab5ab3c69f1ba3260bc493dcab9256e2fa6df2e3936c3c5dc26d118902fe5b9

Observation e4069d0e-ac3c-40d2-b4e4-b5818d961e46 · inbound

RExBench: Can coding agents autonomously implement AI research extensions? cites this paper.

RExBench: Can coding agents autonomously implement AI research extensions? SciReplicate-Bench: Benchmarking LLMs in Agent-driven Algorithmic Reproduction from Research Papers

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-19T07:37:08.896507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T07:33:39.675929Z digest=sha256:761bb975b5125d751e3499ab585ce0aa8e9370e4f5c70068b6f282dd396cde82

Observation de566e82-297d-4a5f-bb48-4654f38dad21 · inbound

How Far Are AI Scientists from Changing the World? cites this paper.

How Far Are AI Scientists from Changing the World? SciReplicate-Bench: Benchmarking LLMs in Agent-driven Algorithmic Reproduction from Research Papers

Reference 183

Resolution
unresolved
no resolver link, observed 2026-08-06T10:55:15.114201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:55:15.114201Z digest=sha256:f5dadbd88cc530168f28eb182d6970a3ce84f13f986b341c0e21e742aedf1465

Observation 1d377818-4b16-46b2-92a3-bbe4e61933ba · inbound

Automated Table Reproduction via Code Generation cites this paper.

Automated Table Reproduction via Code Generation SciReplicate-Bench: Benchmarking LLMs in Agent-driven Algorithmic Reproduction from Research Papers

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-03T02:41:37.498146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:41:37.498146Z digest=sha256:3feeb56cc82693ef1b53221f50002f912096d4068684af014dfa8ce41de2175d

Observation c21c4961-1ec8-4e94-b511-a5a113b8fbbf · inbound

ReplicatorBench: Benchmarking LLM Agents for Replicability in Social and Behavioral Sciences cites this paper.

ReplicatorBench: Benchmarking LLM Agents for Replicability in Social and Behavioral Sciences SciReplicate-Bench: Benchmarking LLMs in Agent-driven Algorithmic Reproduction from Research Papers

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-16T02:02:06.904631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T02:01:07.555124Z digest=sha256:374026ccbba2112a032076f5fdd89d75ca6b5eac6f563d984fa7cd38bffe14da

Observation 741fd78f-74a9-45ad-a8fe-0ba6e7e5f457 · inbound

AutoSOTA: An End-to-End Automated Research System for State-of-the-Art AI Model Discovery cites this paper.

AutoSOTA: An End-to-End Automated Research System for State-of-the-Art AI Model Discovery SciReplicate-Bench: Benchmarking LLMs in Agent-driven Algorithmic Reproduction from Research Papers

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:00:56.862533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T18:43:51.525392Z digest=sha256:d0b092ce834f470d5ddf9436ee3e3eb5091c517941a128544005399de5341472

Observation 5e96322e-ac5c-4321-b3a5-c6eb0a5072e1 · inbound

AutoSOTA: An End-to-End Automated Research System for State-of-the-Art AI Model Discovery cites this paper.

AutoSOTA: An End-to-End Automated Research System for State-of-the-Art AI Model Discovery SciReplicate-Bench: Benchmarking LLMs in Agent-driven Algorithmic Reproduction from Research Papers

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-13T09:23:40.328143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T09:23:40.328143Z digest=sha256:dea2e5ffe32e36066c02b272da19f1bc796791666b422c9e142e61d74a61701e

Observation f49d13c1-7280-445b-b733-8e2cccab9811 · inbound

MLReplicate: Benchmarking Autonomous Research Systems for Machine Learning Reproducibility cites this paper.

MLReplicate: Benchmarking Autonomous Research Systems for Machine Learning Reproducibility SciReplicate-Bench: Benchmarking LLMs in Agent-driven Algorithmic Reproduction from Research Papers

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T20:03:44.090179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T19:59:40.519962Z digest=sha256:39314e6ada89d8c16621b5d46acac2c6f932a4a4a2754a5ef7c4afdfac1b01a8

Observation c952b031-2e68-43f2-a1e0-76e4516c2e54 · inbound

ArtifactLinker: Linking Scientific Artifacts for Automatic State-of-the-Art Discovery cites this paper.

ArtifactLinker: Linking Scientific Artifacts for Automatic State-of-the-Art Discovery SciReplicate-Bench: Benchmarking LLMs in Agent-driven Algorithmic Reproduction from Research Papers

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:32:45.495056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T20:29:27.398862Z digest=sha256:d23d0eb70f34be4a4606942c74263f112cd61da79815a9fe9dbcc6c46c4deb1f

Observation f6e7543c-29dc-4abe-9bb8-133c976c06ac · inbound

AI for Auto-Research: Roadmap & User Guide cites this paper.

AI for Auto-Research: Roadmap & User Guide SciReplicate-Bench: Benchmarking LLMs in Agent-driven Algorithmic Reproduction from Research Papers

Reference 225

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:33:12.443005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T10:30:50.256635Z digest=sha256:7a55a29771f8b7bddb631f6099cd39bf3086bc2516f54335a722c64144080acc

Observation e579bea5-ca37-4d48-a286-40cf00e5c174 · inbound

AI for Auto-Research: Roadmap & User Guide cites this paper.

AI for Auto-Research: Roadmap & User Guide SciReplicate-Bench: Benchmarking LLMs in Agent-driven Algorithmic Reproduction from Research Papers

Reference 224

Resolution
unresolved
no resolver link, observed 2026-08-02T13:43:50.375116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:43:50.375116Z digest=sha256:3003d791d1d383690ecf045d24d0df016b31f6b30bd9f5f81020c1a416906ae7

Observation 140dbc91-4057-4728-ae02-f51036cc599d · inbound

AutoResearch AI: Towards AI-Powered Research Automation for Scientific Discovery cites this paper.

AutoResearch AI: Towards AI-Powered Research Automation for Scientific Discovery SciReplicate-Bench: Benchmarking LLMs in Agent-driven Algorithmic Reproduction from Research Papers

Reference 116

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:50:21.718166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T04:46:43.679185Z digest=sha256:de44dc899373b61e9a773773b5d22e8b5dcefbde8d48247f2bac47a341b5911f

Observation 00e882c1-d35b-48fb-a843-93fb90c013b4 · inbound

From paper to benchmark: agentic, framework-based reproduction of under-specified methods in machine health intelligence cites this paper.

From paper to benchmark: agentic, framework-based reproduction of under-specified methods in machine health intelligence SciReplicate-Bench: Benchmarking LLMs in Agent-driven Algorithmic Reproduction from Research Papers

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-06-29T12:23:24.627423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T12:13:58.111299Z digest=sha256:a82dba953b0bb53081987bb0071476e77bdc0ea3ad82bd9de53ef66e715346f5

Observation 9c6a6831-9891-438f-9c8e-7590dbca4dc7 · inbound

Coding-agents can replicate scientific machine learning papers cites this paper.

Coding-agents can replicate scientific machine learning papers SciReplicate-Bench: Benchmarking LLMs in Agent-driven Algorithmic Reproduction from Research Papers

Reference 46

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T13:08:07.737869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-03T13:07:26.748722Z digest=sha256:cd79a9fa3d238ad52fcb04ac610f5ab693dc73d876950c22a7ed8ea678f62253

Observation 724e10b7-09c7-414d-a0d0-b4b676986196 · inbound

Coding-agents can replicate scientific machine learning papers cites this paper.

Coding-agents can replicate scientific machine learning papers SciReplicate-Bench: Benchmarking LLMs in Agent-driven Algorithmic Reproduction from Research Papers

Reference 46

Resolution
unresolved
no resolver link, observed 2026-07-13T07:13:13.354959Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T07:13:13.354959Z digest=sha256:c5e09c13f6080c40ce268eead91644780c2ac065d24344b43dac69822be2102f

Observation c7949f0a-f93f-4109-a5a3-ee40f1e937e1 · inbound

MADE: Belief-Driven Dual-Agent Coordination for Autonomous Model Deployment cites this paper.

MADE: Belief-Driven Dual-Agent Coordination for Autonomous Model Deployment SciReplicate-Bench: Benchmarking LLMs in Agent-driven Algorithmic Reproduction from Research Papers

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T00:31:12.049868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:31:12.049868Z digest=sha256:75a0c69587ffdffb0482e6b367981a43ecf01f2c051ef3dcb876b6b462006b19