Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2502.16182.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T01:04:38.711041Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation bfd2a199-21d7-4047-b36c-e29037e716e8 · inbound
Exploring the Potential of Offline RL for Reasoning in LLMs: A Preliminary Study IPO: Your Language Model is Secretly a Preference Classifier
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 355554fc-20af-44fa-97b0-bd70159a2627 · inbound
Beyond Reasoning Gains: Mitigating General-Capability Forgetting in Large Reasoning Models IPO: Your Language Model is Secretly a Preference Classifier
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee57834e-e651-4d7b-b114-1e7c8fe64ea9 · inbound
CapTrack: Multifaceted Evaluation of Forgetting in LLM Post-Training IPO: Your Language Model is Secretly a Preference Classifier
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation d5df45eb-136f-427e-a3c2-527120413804 · inbound
Meta-Aligner: Bidirectional Preference-Policy Optimization for Multi-Objective LLMs Alignment IPO: Your Language Model is Secretly a Preference Classifier
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b1c78310-ee14-4559-a8d1-f26e58090f10 · inbound
Black-box, Adaptive, Efficient, Transferable, Harmful, Applicable... Attacks Are All You Need to Break LLMs IPO: Your Language Model is Secretly a Preference Classifier
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.