Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2505.10527.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:49:45.723481Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T00:07:27.940098Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation c6887f61-aa82-4a10-9f0e-16a07b78309b · inbound
RewardDance: Reward Scaling in Visual Generation WorldPM: Scaling Human Preference Modeling
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f423c084-1aae-4047-82a9-ee19963d9dc0 · inbound
AI Can Learn Scientific Taste WorldPM: Scaling Human Preference Modeling
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7f984f2-8052-45cb-abbe-747d8cfe1a63 · inbound
Beyond Overlap Metrics: Rewarding Reasoning and Preferences for Faithful Multi-Role Dialogue Summarization WorldPM: Scaling Human Preference Modeling
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d8ebd339-9560-40c2-9808-0fb0e0b52c20 · inbound
Leveraging Verifier-Based Reinforcement Learning in Image Editing WorldPM: Scaling Human Preference Modeling
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b54261bb-8481-4f67-9bc9-04a6944e3572 · inbound
Leveraging Verifier-Based Reinforcement Learning in Image Editing WorldPM: Scaling Human Preference Modeling
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2619ad0c-fc3c-45c9-8374-4ed6ab95b580 · inbound
RewardHarness: Self-Evolving Agentic Post-Training WorldPM: Scaling Human Preference Modeling
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7808e38c-29ba-4e74-a93b-4aad4a8ebd20 · inbound
Z-Reward: Beyond Scalar Rewards by Internalizing Reasoning into Score Distributions WorldPM: Scaling Human Preference Modeling
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2d25d118-c7e0-45f1-8ef5-4658c4a9af16 · inbound
Z-Reward: Beyond Scalar Rewards by Internalizing Reasoning into Score Distributions WorldPM: Scaling Human Preference Modeling
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e834e349-dfe1-4c35-b6a7-699bc4d1b514 · inbound
RRC: Unlocking Generative Reward Models in LLM Reinforcement Learning via Ranking-Based Reward Construction WorldPM: Scaling Human Preference Modeling
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.