Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2503.11701.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:09:44.129023Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
2
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation e1c935f6-3a28-48d5-aa9a-b50a8efa91e9 · inbound
Consistent Paths Lead to Truth: Self-Rewarding Reinforcement Learning for LLM Reasoning A Survey of Direct Preference Optimization
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99d4cb26-b296-413f-a0f2-d56a0e88cb02 · inbound
Intra-Trajectory Consistency for Reward Modeling A Survey of Direct Preference Optimization
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a998fd7-a7b3-4929-9f05-f782f3d09b1b · inbound
A Novel Self-Evolution Framework for Large Language Models A Survey of Direct Preference Optimization
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8de070ee-5aae-42de-9bcb-a29fff768523 · inbound
Enhancing Speech Large Language Models through Reinforced Behavior Alignment A Survey of Direct Preference Optimization
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ecd42d58-be7e-4676-946a-a1142d2fc2ae · inbound
Efficiency vs. Alignment: Investigating Safety and Fairness Risks in Parameter-Efficient Fine-Tuning of LLMs A Survey of Direct Preference Optimization
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee5a3204-80e7-427e-b045-6d032292e33a · inbound
VC-Soup: Value-Consistency Guided Multi-Value Alignment for Large Language Models A Survey of Direct Preference Optimization
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1696a3ff-944e-435f-8c6c-275e5452166a · inbound
Large Language Model Post-Training: A Unified View of Off-Policy and On-Policy Learning A Survey of Direct Preference Optimization
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fd377ff2-8a5f-44e1-bfc0-9fec693a5a7f · inbound
Mobile GUI Agent Privacy Personalization with Trajectory Induced Preference Optimization A Survey of Direct Preference Optimization
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9c1d35da-0d6c-45d9-8dd0-cd8bd70ac6c5 · inbound
TokenRatio: Principled Token-Level Preference Optimization via Ratio Matching A Survey of Direct Preference Optimization
Reference 177
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6b978201-41a7-494c-915b-f4ea47e82fac · inbound
TokenRatio: Principled Token-Level Preference Optimization via Ratio Matching A Survey of Direct Preference Optimization
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 62e96638-a393-4a34-84f2-c538b10ce10c · inbound
TUX: Measuring Human--AI Tacit Understanding A Survey of Direct Preference Optimization
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 20ac0933-1d06-433f-bea2-65123a72cac4 · inbound
Reliable Neural-Codec Text-to-Speech by ASR Self-Verification and Distillation: Near-Zero Catastrophic Failures Across Models and Codecs A Survey of Direct Preference Optimization
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 35783b72-aa3d-4470-8121-0eaf885c9031 · inbound
SCOPE and SCION: A Benchmark and an Auditable Reference Pipeline for Schema Induction and Fusion from Text A Survey of Direct Preference Optimization
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f2b483e-719d-4b54-a9a6-702675fde444 · inbound
Rolling With Resistance: Preference-Optimized LLM Counselors Can Trade Goal Persistence for Relational Attunement in Motivational Interviewing A Survey of Direct Preference Optimization
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.