Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2307.08678.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:14:41.316054Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
5
pith, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 98f75df4-e94b-4224-aa3e-00aa9fcf32a8 · inbound
New Faithfulness-Centric Interpretability Paradigms for Natural Language Processing Do Models Explain Themselves? Counterfactual Simulatability of Natural Language Explanations
Reference 214
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44df6470-a2a8-4fa4-80fa-1949b9dd7150 · inbound
Let your LLM generate a few tokens and you will reduce the need for retrieval Do Models Explain Themselves? Counterfactual Simulatability of Natural Language Explanations
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29df79b0-d079-4839-bd57-ba240b292bfa · inbound
The Science of Evaluating Foundation Models Do Models Explain Themselves? Counterfactual Simulatability of Natural Language Explanations
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5c304b5-0ab3-438d-ad9b-0814d8cde3eb · inbound
Measuring the Faithfulness of Thinking Drafts in Large Reasoning Models Do Models Explain Themselves? Counterfactual Simulatability of Natural Language Explanations
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54d38c93-7ef8-4d85-8794-6adf0cf2c45a · inbound
Teaching Models to Verbalize Reward Hacking in Chain-of-Thought Reasoning Do Models Explain Themselves? Counterfactual Simulatability of Natural Language Explanations
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9cad1910-5b6a-4cc6-ab7e-72aca740de40 · inbound
Beyond Correctness: Rewarding Faithful Reasoning in Retrieval-Augmented Generation Do Models Explain Themselves? Counterfactual Simulatability of Natural Language Explanations
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8eceb94-7d19-431f-ab49-a252e88c46d4 · inbound
From Features to Actions: Explainability in Traditional and Agentic AI Systems Do Models Explain Themselves? Counterfactual Simulatability of Natural Language Explanations
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a60abfb8-b0aa-4f5a-b5e0-a0bf1a849545 · inbound
Exposure is not manifestation: measurement target and output resolution jointly determine which behavioural-faithfulness evaluator wins Do Models Explain Themselves? Counterfactual Simulatability of Natural Language Explanations
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8f91e2e-1fb2-40cc-a98f-2f7011873347 · inbound
Training Large Language Models for Self-Explanation Faithfulness Do Models Explain Themselves? Counterfactual Simulatability of Natural Language Explanations
Reference 110
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.