Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2402.19450.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:05:07.629687Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T22:47:25.903971Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation da975d38-1082-4a7f-8cb6-a4ee7715d135 · inbound
LiveBench: A Challenging, Contamination-Limited LLM Benchmark Functional Benchmarks for Robust Evaluation of Reasoning Performance, and the Reasoning Gap
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation eab653d2-d2a8-4d6f-a349-0f105c367b7c · inbound
GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models Functional Benchmarks for Robust Evaluation of Reasoning Performance, and the Reasoning Gap
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f329c6bc-ee4e-4f40-a39c-393130f64d25 · inbound
ASyMOB: Algebraic Symbolic Mathematical Operations Benchmark Functional Benchmarks for Robust Evaluation of Reasoning Performance, and the Reasoning Gap
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3826905d-f638-4ed1-aa2b-97dc4f193be1 · inbound
Probing for Arithmetic Errors in Language Models Functional Benchmarks for Robust Evaluation of Reasoning Performance, and the Reasoning Gap
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fac53161-6d61-45da-b0ef-7080273f2620 · inbound
EngiBench: A Benchmark for Evaluating Large Language Models on Engineering Problem Solving Functional Benchmarks for Robust Evaluation of Reasoning Performance, and the Reasoning Gap
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ac9ef06e-0d3a-4efc-a2b4-f1446c679a46 · inbound
Riemann-Bench: A Benchmark for Moonshot Mathematics Functional Benchmarks for Robust Evaluation of Reasoning Performance, and the Reasoning Gap
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 36339f75-5eda-42d9-a948-c9fb76348a22 · inbound
Robust Reasoning Benchmark Functional Benchmarks for Robust Evaluation of Reasoning Performance, and the Reasoning Gap
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6374d269-33f9-4537-8891-e3976752897d · inbound
Robust Reasoning Benchmark Functional Benchmarks for Robust Evaluation of Reasoning Performance, and the Reasoning Gap
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3e85165a-702b-4235-8e8b-4d0b8e12e8b4 · inbound
Robust Reasoning Benchmark Functional Benchmarks for Robust Evaluation of Reasoning Performance, and the Reasoning Gap
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e40167f0-3073-4d43-b3d3-dcb104f12179 · inbound
LPDS: Evaluating LLM Robustness Through Logic-Preserving Difficulty Scaling Functional Benchmarks for Robust Evaluation of Reasoning Performance, and the Reasoning Gap
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 730dc534-79cb-4f5f-94b6-fd545a8a024e · inbound
Artificial Intelligence for Mathematical Reasoning: An Integrated Survey of Language Models, Neuro-symbolic Systems, and Verified Discovery Functional Benchmarks for Robust Evaluation of Reasoning Performance, and the Reasoning Gap
Reference 257
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9f4da5cd-79e4-435b-9d6a-4608fea40c7b · inbound
Artificial Intelligence for Mathematical Reasoning: An Integrated Survey of Language Models, Neuro-symbolic Systems, and Verified Discovery Functional Benchmarks for Robust Evaluation of Reasoning Performance, and the Reasoning Gap
Reference 259
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f402b3a-c51d-487d-ac1b-8d2bd36d255e · inbound
MirrorCode: AI can rebuild entire programs from behavior alone Functional Benchmarks for Robust Evaluation of Reasoning Performance, and the Reasoning Gap
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f775b9b4-0b01-49ef-b1c7-7cf72a1db984 · inbound
MirrorCode: AI can rebuild entire programs from behavior alone Functional Benchmarks for Robust Evaluation of Reasoning Performance, and the Reasoning Gap
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.