Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2410.03492.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:08:06.512628Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-21T08:14:03.236562Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 772c4ac6-2bcc-4f8e-9b10-16db216a2aa6 · inbound
Foundation Models for Geospatial Reasoning: Assessing Capabilities of Large Language Models in Understanding Geometries and Topological Spatial Relations Towards Reproducible LLM Evaluation: Quantifying Uncertainty in LLM Benchmark Scores
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7e1caf3-44f0-4c07-86a0-d19a3b528224 · inbound
From Queries to Criteria: Understanding How Astronomers Evaluate LLMs Towards Reproducible LLM Evaluation: Quantifying Uncertainty in LLM Benchmark Scores
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 464e1b2e-033f-4181-b6cd-8445e4645deb · inbound
Correcting Prompt Dependence in LLM Benchmarks: A Bayesian Hierarchical Model with Embedding-Space Clustering Towards Reproducible LLM Evaluation: Quantifying Uncertainty in LLM Benchmark Scores
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 985e0aaf-b029-4c94-806b-29d130e9e035 · inbound
ReasonBENCH: Benchmarking the (In)Stability of LLM Reasoning Towards Reproducible LLM Evaluation: Quantifying Uncertainty in LLM Benchmark Scores
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74b179ef-447e-4276-a3e8-e63b291a653c · inbound
AI Agent for Reverse-Engineering Legacy Finite-Difference Code and Translating to Devito Towards Reproducible LLM Evaluation: Quantifying Uncertainty in LLM Benchmark Scores
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59ea88af-9ed3-44f4-a313-28e2d64dd207 · inbound
Inspectable AI for Science: A Research Object Approach to Generative AI Governance Towards Reproducible LLM Evaluation: Quantifying Uncertainty in LLM Benchmark Scores
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9b3f1e80-46a1-4959-8f7a-6b08e0af7eff · inbound
How Compliant Are GitHub Actions Workflows? A Checklist-Based Study with LLM-Assisted Auditing Towards Reproducible LLM Evaluation: Quantifying Uncertainty in LLM Benchmark Scores
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 84a8b33b-3879-4803-bac8-d08a11bfb0a4 · inbound
QSTRBench: a New Benchmark to Evaluate the Ability of Language Models to Reason with Qualitative Spatial and Temporal Calculi Towards Reproducible LLM Evaluation: Quantifying Uncertainty in LLM Benchmark Scores
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 54f3f90f-0e1b-4d45-87d8-e5942a63277f · inbound
The Silent Hyperparameter: Quantifying the Impact of Inference Backends on LLM Reproducibility Towards Reproducible LLM Evaluation: Quantifying Uncertainty in LLM Benchmark Scores
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a2561c61-f772-4ee2-bdac-2a9207f60677 · inbound
The Silent Hyperparameter: Quantifying the Impact of Inference Backends on LLM Reproducibility Towards Reproducible LLM Evaluation: Quantifying Uncertainty in LLM Benchmark Scores
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d722c6cf-7d54-4b14-88a4-d94c794e860a · inbound
LEX-EC: A Lexical Evidence-Channel Audit Framework for Zero-Shot LLM Personality Classification in Black-Box Settings Towards Reproducible LLM Evaluation: Quantifying Uncertainty in LLM Benchmark Scores
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 332ff4db-a0c9-4568-97eb-ad4a60e7fae0 · inbound
LEX-EC: A Lexical Evidence-Channel Audit Framework for Zero-Shot LLM Personality Classification in Black-Box Settings Towards Reproducible LLM Evaluation: Quantifying Uncertainty in LLM Benchmark Scores
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.