Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2503.08942.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T22:50:33.166838Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-20T14:28:21.279200Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation d5d3765e-0dfd-4e57-8cae-406d66476daf · inbound
Game Theory Meets Large Language Models: A Systematic Survey with Taxonomy and New Frontiers Extragradient Preference Optimization (EGPO): Beyond Last-Iterate Convergence for Nash Learning from Human Feedback
Reference 137
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5cd7e6c0-792b-4999-8f7b-1ef62dfc01ff · inbound
Sign-SZPO: Provable Preference-based Reinforcement Learning with an Unknown Link Function Extragradient Preference Optimization (EGPO): Beyond Last-Iterate Convergence for Nash Learning from Human Feedback
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95a1eeb1-752b-4012-a402-27bab3384241 · inbound
Multiplayer Nash Preference Optimization Extragradient Preference Optimization (EGPO): Beyond Last-Iterate Convergence for Nash Learning from Human Feedback
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e730ae52-9436-42c0-bc19-1591dddadcc8 · inbound
Towards General Preference Alignment: Diffusion Models at Nash Equilibrium Extragradient Preference Optimization (EGPO): Beyond Last-Iterate Convergence for Nash Learning from Human Feedback
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 57272c39-ab3b-4311-af19-15397ae972bb · inbound
Transitivity Meets Cyclicity: Explicit Preference Decomposition for Dynamic Large Language Model Alignment Extragradient Preference Optimization (EGPO): Beyond Last-Iterate Convergence for Nash Learning from Human Feedback
Reference 108
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.