Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 2 inbound Pith citation observations for arXiv:2408.14874.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-12T21:50:34.927776Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-08T14:27:51.348717Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation b0129bce-d5ba-49c3-95c8-b328904045f5 · inbound
RED: Unleashing Token-Level Rewards from Holistic Feedback via Reward Redistribution Inverse-Q*: Token Level Reinforcement Learning for Aligning Large Language Models Without Preference Data
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef4dc255-3e0b-4c4f-8de8-5f10a5aa7577 · inbound
Exploring the Limit of Outcome Reward for Learning Mathematical Reasoning Inverse-Q*: Token Level Reinforcement Learning for Aligning Large Language Models Without Preference Data
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.