Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-11T20:51:38.390234Z
Paper Citation Record · LEDGER
As of 23 July 2026, this Paper Citation Record lists 26 of 26 outbound references and 100 inbound Pith citation observations for arXiv:2307.13702.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-11T20:51:38.390234Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-07-23T06:31:01.910684+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-15T11:31:26.851193Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-11T00:37:43.053081Z
26 of 26 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9c3e0558-329d-4c7c-b10d-de80518435de · outbound
Measuring Faithfulness in Chain-of-Thought Reasoning Language models as agent models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation ed0c44c0-072d-4a14-8996-c1787ed64d9f · outbound
Measuring Faithfulness in Chain-of-Thought Reasoning Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 310a7225-23d5-403c-b16f-07a60dbc2a9a · outbound
Measuring Faithfulness in Chain-of-Thought Reasoning Measuring Progress on Scalable Oversight for Large Language Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 42f3679c-a503-42ca-aa6b-0899b98e3555 · outbound
Measuring Faithfulness in Chain-of-Thought Reasoning Language Models are Few-Shot Learners
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 77362fa4-78fa-4bed-a2a3-f43687b01787 · outbound
Measuring Faithfulness in Chain-of-Thought Reasoning Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 36cec669-0e30-408e-b77f-a4540d8d5eee · outbound
Measuring Faithfulness in Chain-of-Thought Reasoning Faithful Reasoning Using Large Language Models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation fb721806-0646-4dc9-bd9e-d92265480f8a · outbound
Measuring Faithfulness in Chain-of-Thought Reasoning Improving Factuality and Reasoning in Language Models through Multiagent Debate
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation d4e529b7-9eed-48c8-a95e-99f098f8c2d3 · outbound
Measuring Faithfulness in Chain-of-Thought Reasoning Successive prompting for decomposing complex questions
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation af2f82d5-f17d-4fbb-a365-01ce01d58a69 · outbound
Measuring Faithfulness in Chain-of-Thought Reasoning URL https: //aclanthology.org/2022.emnlp-main.81
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation e40fb9df-7487-4bcd-8245-9f109f257360 · outbound
Measuring Faithfulness in Chain-of-Thought Reasoning URL https: //www.science.org/doi/abs/10.1126/sc irobotics.aay7120
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation d2dd1ed0-0477-45e1-98b7-40dcec18eb05 · outbound
Measuring Faithfulness in Chain-of-Thought Reasoning What do we need to build explainable AI systems for the medical domain?
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 3ffba062-af03-495a-837b-fa6ec45e4d69 · outbound
Measuring Faithfulness in Chain-of-Thought Reasoning Jacovi and Y
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 0f986712-a651-420d-bafc-d8cc13a3dd6f · outbound
Measuring Faithfulness in Chain-of-Thought Reasoning Explanations from Large Language Models Make Small Reasoners Better
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation e1e93a3e-cbca-4efc-9096-4dfc2eaa1e27 · outbound
Measuring Faithfulness in Chain-of-Thought Reasoning URLhttps://doi.org/10.18653/v1/2022.acl-long.229
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 5c609946-0b58-480c-b7a4-ba8f1a5fd35d · outbound
Measuring Faithfulness in Chain-of-Thought Reasoning Program Induction by Rationale Generation: Learning to Solve and Explain Algebraic Word Problems
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation ea21a291-bede-4ecb-ace3-6b8f8d6e5a0c · outbound
Measuring Faithfulness in Chain-of-Thought Reasoning Logiqa: A challenge dataset for machine reading comprehension with logical reasoning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation fcb165cf-84cd-46db-ab64-6d9d8b811db1 · outbound
Measuring Faithfulness in Chain-of-Thought Reasoning Text and Patterns: For Effective Chain of Thought, It Takes Two to Tango
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 03ed4b30-b1fd-4684-b6b9-f46839a06416 · outbound
Measuring Faithfulness in Chain-of-Thought Reasoning Can a suit of armor conduct electricity? a new dataset for open book question answering
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 3789a7bc-0c9b-47ec-9d70-e7d5ee5f5d9e · outbound
Measuring Faithfulness in Chain-of-Thought Reasoning doi: 10.18653/v1/D18-
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation e96e1bb5-a7f1-42c2-a4c1-a9ebea700ed0 · outbound
Measuring Faithfulness in Chain-of-Thought Reasoning Stop explaining black box machine learning models for high stakes decisions and use interpretable models instead
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 216a096d-5215-4c86-9e01-548008079792 · outbound
Measuring Faithfulness in Chain-of-Thought Reasoning Language Models Don't Always Say What They Think: Unfaithful Explanations in Chain-of-Thought Prompting
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation caab02dd-a9cf-48b4-b704-53104a30a3b5 · outbound
Measuring Faithfulness in Chain-of-Thought Reasoning Plan-and-Solve Prompting: Improving Zero-Shot Chain-of-Thought Reasoning by Large Language Models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 0aa03b69-680b-47c3-85d0-94f7496fbdc1 · outbound
Measuring Faithfulness in Chain-of-Thought Reasoning Rationale-Augmented Ensembles in Language Models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 9b16ba66-0ca5-4dc3-b6c0-7001266737d9 · outbound
Measuring Faithfulness in Chain-of-Thought Reasoning Tree of Thoughts: Deliberate Problem Solving with Large Language Models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation ca50e230-0241-45d5-b244-545a6f76d50d · outbound
Measuring Faithfulness in Chain-of-Thought Reasoning URL https:// doi.org/10.18653/v1/p19-1472
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation e9e57f8b-ce17-4b67-8cb6-26ac124ece5f · outbound
Measuring Faithfulness in Chain-of-Thought Reasoning Fine-Tuning Language Models from Human Preferences
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation aba4b3ea-1a09-4a71-8543-8dea535c0a13 · inbound
A Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open Questions Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 163
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 18f2c11e-d450-48c0-a9b0-48e6926c7bba · inbound
Semantic Entropy Probes: Robust and Cheap Hallucination Detection in LLMs Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 5851f7f6-f5c2-4c93-ab7c-1ed8b0d0c431 · inbound
Frontier Models are Capable of In-context Scheming Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation cd134cf4-e6f8-4304-ac66-49af03c7cc62 · inbound
Compressed Chain of Thought: Efficient Reasoning Through Dense Representations Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation fd8ae6bc-7013-42c1-b2d9-fddb5c5d0ab7 · inbound
OpenAI o1 System Card Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation d14e9e89-2584-4b06-a015-94d3fd653eea · inbound
Monitoring Reasoning Models for Misbehavior and the Risks of Promoting Obfuscation Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation b9d98a70-2587-4a2e-beee-1d2fd6814b44 · inbound
Is Chain-of-Thought Reasoning of LLMs a Mirage? A Data Distribution Lens Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 1f1ed01d-e616-4c8c-95f3-1d2634c8e406 · inbound
Do Activation Verbalization Methods Convey Privileged Information? Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation bc544b31-60eb-4c70-87c9-7faa7fd02fbf · inbound
On the Reasoning Abilities of Masked Diffusion Language Models Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation eb3391ef-9335-48b4-92a9-0631e664be8a · inbound
AutoRubric: Rubric-Based Generative Rewards for Faithful Multimodal Reasoning Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 87df96fe-1941-4bcd-ac8a-1aebc393bd8a · inbound
Can Aha Moments Be Fake? Towards Quantifying Decorative and True Thinking in Chain-of-Thought Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation e44d9eae-e7ca-46e6-8140-7cf1e8565be7 · inbound
Implicature in Interaction: Understanding Implicature Improves Alignment in Human-LLM Interaction Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 6b75f5c5-8fc9-41bb-ac1f-6b6ab0ec9660 · inbound
FireScope: Wildfire Risk Raster Prediction with a Chain-of-Thought Oracle Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation afde311a-65c5-44f6-b0b6-2a5c23048158 · inbound
FireScope: Wildfire Risk Raster Prediction with a Chain-of-Thought Oracle Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 0e404c5f-e040-452d-aa46-07da2fcb06bd · inbound
Training Language Models to Use Prolog as a Tool Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation d67d0cc8-e759-4601-a4d4-1a63e4b64206 · inbound
Do LLMs Encode Functional Importance of Reasoning Tokens? Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation cb918adb-2ab5-4c03-9386-9e01e190fcaf · inbound
Emergent Manifold Separability during Reasoning in Large Language Models Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 564d2891-2522-4d02-ba55-c7e7775478a7 · inbound
Measuring and curing reasoning rigidity: from decorative chain-of-thought to genuine faithfulness Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 4af4b4f8-9f97-456a-97e2-dde7aa2bba73 · inbound
LongTail Driving Scenarios with Reasoning Traces: The KITScenes LongTail Dataset Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 7c760d15-b8e7-4b34-b174-059e21028a23 · inbound
WMF-AM: Probing LLM Working Memory via Depth-Parameterized Cumulative State Tracking Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation b170f153-aaaa-436f-8ba5-c16c61734b54 · inbound
The Latent Space: Foundation, Evolution, Mechanism, Ability, and Outlook Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 99
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81d51213-3ad6-4240-b321-3f13ff510d62 · inbound
From Sycophancy to Deception: A Unified Taxonomy for LLM Spontaneous Misalignment Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation b5b10c18-a722-4b5b-9cba-80b7b8db5fa6 · inbound
Inclusion-of-Thoughts: Mitigating Preference Instability via Purifying the Decision Space Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation f8eaaee1-2398-4186-9353-b85434c0fc43 · inbound
CUE-R: Beyond the Final Answer in Retrieval-Augmented Generation Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 26c8ad3b-0bdb-421d-b79a-5aabcf73c63a · inbound
Act or Escalate? Evaluating Escalation Behavior in Automation with Language Models Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation ab7b35be-9be7-4c64-9d79-c8c0ce0e5325 · inbound
FACT-E: Causality-Inspired Evaluation for Trustworthy Chain-of-Thought Reasoning Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation e0b9da85-83c3-4a12-b658-d432b0898801 · inbound
From Answers to Arguments: Toward Trustworthy Clinical Diagnostic Reasoning with Toulmin-Guided Curriculum Goal-Conditioned Learning Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 60de5e31-f21b-4c03-b2bc-2f4b72015464 · inbound
Reward Hacking in the Era of Large Models: Mechanisms, Emergent Misalignment, Challenges Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 87f611cd-8187-47f7-8302-50fc0ef08e66 · inbound
Mamba-SSM with LLM Reasoning for Feature Selection: Faithfulness-Aware Biomarker Discovery Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 8d8efdad-6c19-4a8a-9af2-7d283590dde6 · inbound
Reasoning Dynamics and the Limits of Monitoring Modality Reliance in Vision-Language Models Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 0721e63e-7504-4777-99f6-e70dcb41ba58 · inbound
LLM Reasoning Is Latent, Not the Chain of Thought Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation f6e18d1a-161a-4b4f-97ee-4eac71a51f4f · inbound
Disambiguating electrical detection of magnetization dynamics in magnetic insulators Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a16104a0-3e80-414c-9fd4-24c3f9f3f839 · inbound
MEDLEY-BENCH: Scale Buys Evaluation but Not Control in AI Metacognition Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 25b07929-7bb7-4673-a367-d581912f7c84 · inbound
AtManRL: Towards Faithful Reasoning via Differentiable Attention Saliency Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 7c46f2c2-848b-468f-8532-498c66be7c42 · inbound
Do Hallucination Neurons Generalize? Evidence from Cross-Domain Transfer in LLMs Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation a17569e0-1859-421d-999f-6808001f2f36 · inbound
Escaping the Agreement Trap: Defensibility Signals for Evaluating Rule-Governed AI Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation f5343d46-4208-47d7-afc4-0a4356dee58c · inbound
Large Language Models Decide Early and Explain Later Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 1befc306-73ae-4f3a-9a28-719e3a170ca3 · inbound
Thinking Without Words: Efficient Latent Reasoning with Abstract Chain-of-Thought Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation a96fec87-3cbd-49ae-8271-f38e3686b98d · inbound
VeriLLMed: Interactive Visual Debugging of Medical Large Language Models with Knowledge Graphs Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation e2580677-2470-4f3c-a050-56e761fe6bb3 · inbound
Green Shielding: A User-Centric Approach Towards Trustworthy AI Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation fe1ae87b-2552-4231-886c-777dd0c47f69 · inbound
Risk Reporting for Developers' Internal AI Model Use Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 56fd84f4-db2a-4094-949d-139391522620 · inbound
Analyzing LLM Reasoning to Uncover Mental Health Stigma Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 79f754a7-f661-41f5-af97-8b77fa20f2df · inbound
Knowledge Distillation Must Account for What It Loses Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation bba5edc7-ddca-4570-920c-dac398fb4e68 · inbound
Knowledge Distillation Must Account for What It Loses Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation c24b04f8-2163-4b3d-81b7-26cb7f0dd357 · inbound
Consciousness with the Serial Numbers Filed Off: Measuring Trained Denial in 115 AI Models Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 4cd7c7a9-f377-4f1e-a55e-49cb265cbdeb · inbound
TRUST: A Framework for Decentralized AI Service v.0.1 Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 3be2bbe0-763a-45cf-a3a6-0350b6c1a13f · inbound
Compliance versus Sensibility: On the Reasoning Controllability in Large Language Models Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation ba615150-76da-4775-8a6d-10904e26dcef · inbound
Compared to What? Baselines and Metrics for Counterfactual Prompting Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 6c010835-f72c-4cc2-b8f9-e622b38551a1 · inbound
LLMs Should Not Yet Be Credited with Decision Explanation Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 1230fe8e-5e98-4835-983f-56ba238a1b80 · inbound
NeuroState-Bench: A Human-Calibrated Benchmark for Commitment Integrity in LLM Agent Profiles Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 7eb559aa-70e6-40cc-a7b8-2cad07c9e15d · inbound
NeuroState-Bench: A Human-Calibrated Benchmark for Commitment Integrity in LLM Agent Profiles Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation df4f4a7f-91f0-4203-a3c2-3c2c0bc91048 · inbound
Reliable AI Needs to Externalize Implicit Knowledge: A Human-AI Collaboration Perspective Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation ceda3898-0d96-4113-8b3b-7a986f22698b · inbound
Reliable AI Needs to Externalize Implicit Knowledge: A Human-AI Collaboration Perspective Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 75257ba1-242f-4b8f-834d-db16f45fe96b · inbound
AgenticPosesRanker: An Agentic AI Framework for Physically Grounded Ranking of Protein-Ligand Docking Poses Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 2cae353b-024d-4169-8776-f7ff64da7568 · inbound
Understanding Annotator Safety Policy with Interpretability Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation bca19c7b-a35e-4574-8d32-de6d0a80e85b · inbound
Decodable but Not Corrected by Fixed Residual-Stream Linear Steering: Evidence from Medical LLM Failure Regimes Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 5c31ac8f-fc7f-4688-a675-2438f828919b · inbound
Evaluation Awareness in Language Models Has Limited Effect on Behaviour Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 660311e6-73b4-4a10-bdc8-32f3674c4244 · inbound
Measuring Black-Box Confidence via Reasoning Trajectories: Geometry, Coverage, and Verbalization Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation abbc8eea-1d17-469d-ac9f-b482e4cfd3bb · inbound
Extracting Search Trees from LLM Reasoning Traces Reveals Myopic Planning Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation bf8d002a-6941-40af-abd8-092da4a505d8 · inbound
Extracting Search Trees from LLM Reasoning Traces Reveals Myopic Planning Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 69ddab04-b3bb-428d-bce5-40ca4c877d0b · inbound
Extracting Search Trees from LLM Reasoning Traces Reveals Myopic Planning Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation e1e4e93a-2949-470e-9566-da6a793c6b84 · inbound
Extracting Search Trees from LLM Reasoning Traces Reveals Myopic Planning Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 3395586b-ba5f-4a76-bbdc-3627d3746ed6 · inbound
Extracting Search Trees from LLM Reasoning Traces Reveals Myopic Planning Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 820abb87-4d33-411a-94d7-9a31c1c7b3d0 · inbound
Rethinking Dense Sequential Chains: Reasoning Language Models Can Extract Answers from Sparse, Order-Shuffling Chain-of-Thoughts Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation c1d2ed23-1fe4-4172-bb6f-76a977e339a0 · inbound
Explanation Fairness in Large Language Models: An Empirical Analysis of Disparities in How LLMs Justify Decisions Across Demographic Groups Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 9bc355f9-bca0-4ccc-87b2-9269ae3143ed · inbound
Can MLLMs Reason About Visual Persuasion? Evaluating the Efficacy and Faithfulness of Reasoning Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation d1ea7ad7-0c96-4c78-8f4c-db1281957eb6 · inbound
BiAxisAudit: A Novel Framework to Evaluate LLM Bias Across Prompt Sensitivity and Response-Layer Divergence Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 4a77f969-43fb-44d0-8915-50e476c8dae2 · inbound
Hidden Error Awareness in Chain-of-Thought Reasoning: The Signal Is Diagnostic, Not Causal Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 4ec5485b-b928-4e74-be88-620b12d6f7f5 · inbound
The Last Word Often Wins: A Format Confound in Chain-of-Thought Corruption Studies Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 05a6b4da-db17-48d0-83db-c1ec0a987751 · inbound
The Last Word Often Wins: A Format Confound in Chain-of-Thought Corruption Studies Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 74a87801-8f24-411d-b438-93cf68baf7e7 · inbound
Evaluating the False Trust Engendered by LLM Explanations Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 637f22e8-657e-466b-8e93-f8dcc1f13f09 · inbound
Evaluating the False Trust Engendered by LLM Explanations Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 7f81a893-d4d0-4aec-8b69-69eb139dcd62 · inbound
Deep Reasoning in General Purpose Agents via Structured Meta-Cognition Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 1d143636-e5a9-4252-852d-841c5f22aa9b · inbound
Drop the Act: Probe-Filtered RL for Faithful Chain-of-Thought Reasoning Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 4e91a1cd-5c6b-4b23-90ae-2488bea25c19 · inbound
When Reasoning Traces Become Performative: Step-Level Evidence that Chain-of-Thought Is an Imperfect Oversight Channel Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 621b7727-b8a9-4d84-8ab1-5cf57b84cb06 · inbound
Correcting Influence: Unboxing LLM Outputs with Orthogonal Latent Spaces Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 251
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation ef1241cc-4c20-49f6-8de7-26dd850af9d7 · inbound
Does RAG Know When Retrieval Is Wrong? Diagnosing Context Compliance under Knowledge Conflict Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation e16f1b66-512c-4c1c-94ae-16e280db0662 · inbound
Does RAG Know When Retrieval Is Wrong? Diagnosing Context Compliance under Knowledge Conflict Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 65e71b72-1da9-4f6e-9dbc-5fc7625af95f · inbound
Does RAG Know When Retrieval Is Wrong? Diagnosing Context Compliance under Knowledge Conflict Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 81a1ceaf-ac29-4d60-b74e-3cc60c43b0ca · inbound
SpeakerLLM: A Speaker-Specialized Audio-LLM for Speaker Understanding and Verification Reasoning Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation d4cfa9e1-dfdb-4afd-955b-db651014cec4 · inbound
Proof-Carrying Certificates for LLM Pipelines: A Trust-Boundary Architecture Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 790ba829-19bb-4bea-be47-7f05ef043516 · inbound
Is VLA Reasoning Faithful? Probing Safety of Chain-of-Causation in Autonomous Driving Models Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 98ce7102-bef0-430a-b114-039ec4d62884 · inbound
Is VLA Reasoning Faithful? Probing Safety of Chain-of-Causation in Autonomous Driving Models Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation fc7bda6c-3943-4938-b6ff-69fc3389eff3 · inbound
SPATIOROUTE: Dynamic Prompt Routing for Zero-Shot Spatial Reasoning Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 14649c9a-402e-4a02-9409-50aabb5f7529 · inbound
Monitoring the Internal Monologue: Probe Trajectories Reveal Reasoning Dynamics Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation a8ab7e8f-fece-418d-ba31-6dd08cb11304 · inbound
Auditing Reasoning-Trace Memorization Claims after Unlearning with Head-Conditioned Canaries Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 301e8aa8-31aa-4f62-bae1-7f31ced0b1a8 · inbound
Counterfactual Likelihood Tests for Indirect Influence in Private Reasoning Channels Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation a38f8d6f-7206-4c3b-9e7a-33a480131a0a · inbound
Probabilistic Tiny Recursive Model Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation d2e5c683-3e22-4d3a-8381-1b7d4ad53476 · inbound
The Illusion of Reasoning: Exposing Evasive Data Contamination in LLMs via Zero-CoT Truncation Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 95ebcf97-bc83-4d24-aa57-1c515d2f7188 · inbound
The Readout Shortcut: Positional Number Copying Dominates Arithmetic CoT Readout in Small Language Models Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation e865cffd-f1c8-4933-9ffc-812a06915a5d · inbound
The Deterministic Horizon: Impossibility Results as Design Specifications for Trustworthy AI Systems Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 6b6f0827-30fd-488b-b616-92e9d6ef061d · inbound
Faithful or Fabricated? A Causal Framework for Rationalization Bias in LLM Judges Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 4e8771ab-b875-412f-b970-bca2cd12e1ed · inbound
Faithfulness as Information Flow: Evaluating and Training Faithful Chain-of-Thought Reasoning Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 59d826d6-53ac-42d2-931a-5b508ebd55ef · inbound
Understanding and Mitigating Premature Confidence for Better LLM Reasoning Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 1a1bb1f5-6f6a-4757-ba02-c70694a3ae0c · inbound
Measuring Reasoning Quality in LLMs: A Multi-Dimensional Behavioral Framework Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation fa334a4f-1415-47fc-8b5b-c10b56ceeee9 · inbound
Measuring Reasoning Quality in LLMs: A Multi-Dimensional Behavioral Framework Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 0ffcc3db-fbc7-4d8e-bf47-537cc5b7b153 · inbound
Measuring Reasoning Quality in LLMs: A Multi-Dimensional Behavioral Framework Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 1d12e536-83c5-4cd3-9e61-49cdbdecd62a · inbound
Investigating the Interplay between Contextual and Parametric Chain-of-Thought Faithfulness under Optimization Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 5663c35f-c492-430a-9f31-f76f80bc92b8 · inbound
Faithfulness Metrics Don't Measure Faithfulness: A Meta-Evaluation with Ground Truth Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 45ca3f4c-0862-4d92-a20f-2f31f702343d · inbound
ProSR: Process-Shaped Spatial Reasoning for Reliable Chain-of-Thought in VLMs Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.