Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:2401.13986.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T22:03:51.842152Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-11T08:01:00.206133Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 34c38ac9-afcd-4cee-aaa6-c721edec54f4 · inbound
Teaching Models to Verbalize Reward Hacking in Chain-of-Thought Reasoning Towards Consistent Natural-Language Explanations via Explanation-Consistency Finetuning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ff1e785-3ad6-44f2-a00f-3418d956fcfa · inbound
CLARity: Reasoning Consistency Alone Can Teach Reinforced Experts Towards Consistent Natural-Language Explanations via Explanation-Consistency Finetuning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf68ecdf-9504-4665-a4bb-28c94a2df181 · inbound
Free Energy-Driven Reinforcement Learning with Adaptive Advantage Shaping for Unsupervised Reasoning in LLMs Towards Consistent Natural-Language Explanations via Explanation-Consistency Finetuning
Reference 294
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation e35d50d0-9e79-45e9-8298-12772153471a · inbound
Adapt to Thrive! Adaptive Power-Mean Policy Optimization for Improved LLM Reasoning Towards Consistent Natural-Language Explanations via Explanation-Consistency Finetuning
Reference 279
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.