Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T18:00:32.747694Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 1 inbound Pith citation observation for arXiv:2501.11721.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T18:00:32.747694Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:36:55.345947Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-15T20:36:55.445243Z
24 of 24 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9af8672c-e08a-41b2-91b7-49dbd3c8b608 · outbound
Explain-Query-Test: Self-Evaluating LLMs Via Explanation and Comprehension Discrepancy Introducing gemini: Google's multimodal ai model
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 604d3217-b67a-4864-83c9-53b324ee341a · outbound
Explain-Query-Test: Self-Evaluating LLMs Via Explanation and Comprehension Discrepancy Introducing claude: Anthropic's ai assistant
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 6c624207-06d6-40f1-b487-e8e488aa30e5 · outbound
Explain-Query-Test: Self-Evaluating LLMs Via Explanation and Comprehension Discrepancy Explainability in ai: A survey
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 4745ce75-e64c-4df0-bb58-1a58cf4b7fb3 · outbound
Explain-Query-Test: Self-Evaluating LLMs Via Explanation and Comprehension Discrepancy On the dangers of stochastic parrots: Can language models be too big? In Proceedings of the 2021 ACM Conference on Fairness, Accountability, and Transparency, pp.\ 610--623, 2021
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 38fa1fa5-51b2-4ff2-9a2c-49782fe47314 · outbound
Explain-Query-Test: Self-Evaluating LLMs Via Explanation and Comprehension Discrepancy On the Opportunities and Risks of Foundation Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f885034-6a26-45aa-a436-1f61cc575181 · outbound
Explain-Query-Test: Self-Evaluating LLMs Via Explanation and Comprehension Discrepancy Language models are few-shot learners
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e7c0ecb4-fe86-414c-949e-8c707a9f56c5 · outbound
Explain-Query-Test: Self-Evaluating LLMs Via Explanation and Comprehension Discrepancy Bert: Pre-training of deep bidirectional transformers for language understanding
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 05283e8a-614a-4685-a41b-4bad9c1d6467 · outbound
Explain-Query-Test: Self-Evaluating LLMs Via Explanation and Comprehension Discrepancy What if $\phi^4$ theory in 4 dimensions is non-trivial in the continuum?
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c537fa1b-b37e-464c-9e4f-0d61bb4d2692 · outbound
Explain-Query-Test: Self-Evaluating LLMs Via Explanation and Comprehension Discrepancy Prediction of solar wind speed by applying convolutional neural network to potential field source surface (PFSS) magnetograms
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a725d6df-df7f-4458-add2-f214bc0435c1 · outbound
Explain-Query-Test: Self-Evaluating LLMs Via Explanation and Comprehension Discrepancy Structure Factors for Hot Neutron Matter from Ab Initio Lattice Simulations with High-Fidelity Chiral Interactions
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4f9bfda-7183-431a-8d62-7ff38e768035 · outbound
Explain-Query-Test: Self-Evaluating LLMs Via Explanation and Comprehension Discrepancy XC-Cache: Cross-Attending to Cached Context for Efficient LLM Inference
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e7c513f4-769a-460a-86d0-0771b8c722e9 · outbound
Explain-Query-Test: Self-Evaluating LLMs Via Explanation and Comprehension Discrepancy RepLiQA: A Question-Answering Dataset for Benchmarking LLMs on Unseen Reference Content
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b06cca36-3cec-43ab-ba6d-052c9c216d86 · outbound
Explain-Query-Test: Self-Evaluating LLMs Via Explanation and Comprehension Discrepancy Gpt-4 technical report
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29a10000-aa0b-4db0-ba6b-4d80c3c78123 · outbound
Explain-Query-Test: Self-Evaluating LLMs Via Explanation and Comprehension Discrepancy Squad: 100,000+ questions for machine comprehension of text
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 94671660-d958-470d-8465-82b5acbbfaf9 · outbound
Explain-Query-Test: Self-Evaluating LLMs Via Explanation and Comprehension Discrepancy A Statistical Analysis of LLMs' Self-Evaluation Using Proverbs
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 152ce31f-476b-490f-b0de-7a40c0496714 · outbound
Explain-Query-Test: Self-Evaluating LLMs Via Explanation and Comprehension Discrepancy LLaMA: Open and Efficient Foundation Language Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8cfcdcf8-cdd4-4a73-9b45-91dcf06e0348 · outbound
Explain-Query-Test: Self-Evaluating LLMs Via Explanation and Comprehension Discrepancy On the fluid slip along a solid surface
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3250ac9a-1429-4ea8-a333-a1313e7004d2 · outbound
Explain-Query-Test: Self-Evaluating LLMs Via Explanation and Comprehension Discrepancy MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14b953e3-a8be-47dc-998b-8f641d11d0b5 · outbound
Explain-Query-Test: Self-Evaluating LLMs Via Explanation and Comprehension Discrepancy Teach me to explain: A review of machine learning interpretability through explanations
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 9cf92435-8d6a-446c-b62c-9041449ef5c9 · outbound
Explain-Query-Test: Self-Evaluating LLMs Via Explanation and Comprehension Discrepancy Language Models can Evaluate Themselves via Probability Discrepancy
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aac56f11-00f1-4206-b11f-996977b5673c · outbound
Explain-Query-Test: Self-Evaluating LLMs Via Explanation and Comprehension Discrepancy Evaluating paraphrase sensitivity in large language models
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation beac91c4-4df4-43ca-92f2-6504e68182b8 · outbound
Explain-Query-Test: Self-Evaluating LLMs Via Explanation and Comprehension Discrepancy @esa (Ref
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad4ab37a-ed28-4d22-bcd9-ffb4d37d84f9 · outbound
Explain-Query-Test: Self-Evaluating LLMs Via Explanation and Comprehension Discrepancy Unresolved cited work
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bded3b9a-c992-42de-ad0a-58f71e7b9702 · outbound
Explain-Query-Test: Self-Evaluating LLMs Via Explanation and Comprehension Discrepancy Unresolved cited work
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a41e5e2-253d-4009-bdaa-2fe060155d38 · inbound
Making Sense of the Unsensible: Reflection, Survey, and Challenges for XAI in Large Language Models Toward Human-Centered AI Explain-Query-Test: Self-Evaluating LLMs Via Explanation and Comprehension Discrepancy
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.