Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2312.17080.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T04:22:56.090625Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 4502f655-9c6c-40f9-939f-c6f49a5381a6 · inbound
LIMO: Less is More for Reasoning MR-GSM8K: A Meta-Reasoning Benchmark for Large Language Model Evaluation
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3352a403-d713-4bc4-8714-334ec1debcbf · inbound
VRBench: A Benchmark for Multi-Step Reasoning in Long Narrative Videos MR-GSM8K: A Meta-Reasoning Benchmark for Large Language Model Evaluation
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f78baf5-d5ec-4367-843b-8004c469db1f · inbound
Can You Trick the Grader? Adversarial Persuasion of LLM Judges MR-GSM8K: A Meta-Reasoning Benchmark for Large Language Model Evaluation
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a493ec8a-69c0-4af4-8a3d-1a0ba82680a9 · inbound
LLMs cannot spot math errors, even when allowed to peek into the solution MR-GSM8K: A Meta-Reasoning Benchmark for Large Language Model Evaluation
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9f1217e-f53c-4cb7-8020-a8fb7bf01c78 · inbound
Reward Modeling for Reinforcement Learning-Based LLM Reasoning: Design, Challenges, and Evaluation MR-GSM8K: A Meta-Reasoning Benchmark for Large Language Model Evaluation
Reference 121
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 676ed4a1-d5b1-4aa6-abde-c2d56aed9a2a · inbound
MedPRMBench: A Fine-grained Benchmark for Process Reward Models in Medical Reasoning MR-GSM8K: A Meta-Reasoning Benchmark for Large Language Model Evaluation
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9210e3f8-2089-426b-9df2-480b33761759 · inbound
DRIFTLENS: Measuring Memory-Induced Reasoning Drift in Personalized Language Models MR-GSM8K: A Meta-Reasoning Benchmark for Large Language Model Evaluation
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.