Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2503.01076.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T11:29:40.666387Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T01:09:19.284325Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 874a65c6-47f9-46fe-9fbb-95c87f096bdc · inbound
From Reviews to Dialogues: Active Synthesis for Zero-Shot LLM-based Conversational Recommender System Active Learning for Direct Preference Optimization
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20c9b007-853e-4034-8f15-64ffd0fee463 · inbound
Random Is Hard to Beat: Active Selection in online DPO with Modern LLMs Active Learning for Direct Preference Optimization
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation bc60356b-791d-404f-9916-eafa56dd8533 · inbound
Skill-CMIB: Multimodal Agent Skill for Consistent Action via Conditional Multimodal Information Bottleneck Active Learning for Direct Preference Optimization
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation b3acd522-1422-4ad3-8d40-7c69ce457b7a · inbound
MASS-DPO: Multi-negative Active Sample Selection for Direct Policy Optimization Active Learning for Direct Preference Optimization
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 7c84aea1-390e-4564-ba57-8bef1b9dec99 · inbound
OLIVIA: Online Learning via Inference-time Action Adaptation for Decision Making in LLM ReAct Agents Active Learning for Direct Preference Optimization
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a87b4c7a-2c10-4a3f-b686-5dffc95df2c1 · inbound
F-GRPO: Factorized Group-Relative Policy Optimization for Unified Candidate Generation and Ranking Active Learning for Direct Preference Optimization
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 85e32ca2-b8da-46d2-b942-4e2790446284 · inbound
Which Pairs to Compare for LLM Post-Training? Active Learning for Direct Preference Optimization
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.