Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-01T02:41:53.849176Z
Paper Citation Record · LEDGER
As of 5 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 0 inbound Pith citation observations for arXiv:2607.26102.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-01T02:41:53.849176Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
24 of 24 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 16446460-0a45-4338-b6de-6041d2596f4d · outbound
A Reference-Free Score for Detecting Silent Reasoning Failures in Large Language Models Chain-of-thought prompting elicits reasoning in large language models,
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d31b836-828b-4711-9a36-b6d256bb1cc1 · outbound
A Reference-Free Score for Detecting Silent Reasoning Failures in Large Language Models Self-consistency improves chain of thought reasoning in language models,
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ef6c58b-d963-4bf6-b26a-90781f73acfb · outbound
A Reference-Free Score for Detecting Silent Reasoning Failures in Large Language Models Language models don’t always say what they think: Unfaithful explanations in chain-of- thought prompting,
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06bc27f3-5fd7-42c4-ba11-503531c8f602 · outbound
A Reference-Free Score for Detecting Silent Reasoning Failures in Large Language Models Federated generative intelligence for explainable and autonomous cyber defence in critical infrastructures,
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1ea8b0c-5f36-40ed-b395-37e7d3f453b4 · outbound
A Reference-Free Score for Detecting Silent Reasoning Failures in Large Language Models Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81e50c51-c0d7-49a5-83af-0160a179ab67 · outbound
A Reference-Free Score for Detecting Silent Reasoning Failures in Large Language Models Training Verifiers to Solve Math Word Problems
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb9e2ce9-471c-490d-ac5d-da24b6daa20c · outbound
A Reference-Free Score for Detecting Silent Reasoning Failures in Large Language Models Measuring mathematical problem solving with the MATH dataset,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a7da1d2-2edf-481c-8d62-2e0872c13e61 · outbound
A Reference-Free Score for Detecting Silent Reasoning Failures in Large Language Models A Careful Examination of Large Language Model Performance on Grade School Arithmetic
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f09d010b-6dda-448c-8409-984d68cec4cd · outbound
A Reference-Free Score for Detecting Silent Reasoning Failures in Large Language Models Agentic AI framework for autonomous and self-managing cloud services,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5189926f-6de6-4c5d-9a3d-c72bf80fb01e · outbound
A Reference-Free Score for Detecting Silent Reasoning Failures in Large Language Models On measuring faithfulness or self- consistency of natural language explanations,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76295650-45be-4ec4-a378-a6e94c3212b9 · outbound
A Reference-Free Score for Detecting Silent Reasoning Failures in Large Language Models ROSCOE: A suite of metrics for scoring step-by- step reasoning,
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe00edda-5930-4fe9-b34c-524fc2528c07 · outbound
A Reference-Free Score for Detecting Silent Reasoning Failures in Large Language Models DeBERTa: Decoding-enhanced BERT with disentangled attention,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ad813b1-e3bd-45cc-8dcb-7fa791cc1138 · outbound
A Reference-Free Score for Detecting Silent Reasoning Failures in Large Language Models ReCEval: Evaluating reasoning chains via correctness and informativeness,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c07f871-3937-45d9-8c67-17757a48e9ba · outbound
A Reference-Free Score for Detecting Silent Reasoning Failures in Large Language Models On authentication schemes using polynomials over non commutative rings,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b7716dde-f6c6-4e4e-966e-8158fbc0d3c5 · outbound
A Reference-Free Score for Detecting Silent Reasoning Failures in Large Language Models A benchmark for verifiers of reasoning chains,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 813cb8fc-300f-4cf8-9b68-dcd3c6aa2e23 · outbound
A Reference-Free Score for Detecting Silent Reasoning Failures in Large Language Models Direct evaluation of chain-of-thought in multi-hop reasoning with knowledge graphs,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d78d6be-d91a-4c28-8e1b-988c2ea84ad2 · outbound
A Reference-Free Score for Detecting Silent Reasoning Failures in Large Language Models Self-adjointness of semi-relativistic Pauli-Fierz Hamiltonian
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cadcf317-82bb-467d-a790-e4c1c12978f8 · outbound
A Reference-Free Score for Detecting Silent Reasoning Failures in Large Language Models Making reasoning matter: Measuring and improving faithfulness of chain-of-thought rea- soning,
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d9b508b-ab56-488a-95ac-76cd8a60a83d · outbound
A Reference-Free Score for Detecting Silent Reasoning Failures in Large Language Models Towards Better Chain-of-Thought: A Reflection on Effectiveness and Faithfulness
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44a42c15-a6d0-4a29-acbd-39bd6d3e271a · outbound
A Reference-Free Score for Detecting Silent Reasoning Failures in Large Language Models ProcessBench: Identifying Process Errors in Mathematical Reasoning
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6339b64f-3029-4e7d-951b-f1526013122c · outbound
A Reference-Free Score for Detecting Silent Reasoning Failures in Large Language Models Reasoning Models Don't Always Say What They Think
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7d0c5bc-7b1a-4d33-8a3a-a2e66e218315 · outbound
A Reference-Free Score for Detecting Silent Reasoning Failures in Large Language Models Chain-of-Thought Reasoning In The Wild Is Not Always Faithful
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e680009d-7091-4440-93d3-8ef717e4fa51 · outbound
A Reference-Free Score for Detecting Silent Reasoning Failures in Large Language Models DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e6c368a-13f9-498d-8a55-51c82260292f · outbound
A Reference-Free Score for Detecting Silent Reasoning Failures in Large Language Models DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.