Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T05:52:13.962991Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 1 of 1 outbound references and 3 inbound Pith citation observations for arXiv:2508.01174.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T05:52:13.962991Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-14T06:44:16.198117Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-06-30T19:05:00.410095Z
1 of 1 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation ba3c9b84-f6f2-4add-9c1a-b47eedb7755c · outbound
RSPO: Risk-Seeking Policy Optimization for Pass@k and Max@k Metrics in Large Language Models Advancing the Foundation Model for Music Understanding
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 16b6b28c-1a41-4a9c-b0fd-005498e23914 · inbound
Leveraging Error Diversity in Group Rollouts for Reinforcement Learning RSPO: Risk-Seeking Policy Optimization for Pass@k and Max@k Metrics in Large Language Models
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 522341bc-f461-4681-81fb-fff501ead888 · inbound
Leveraging Error Diversity in Group Rollouts for Reinforcement Learning RSPO: Risk-Seeking Policy Optimization for Pass@k and Max@k Metrics in Large Language Models
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation bc930beb-58cd-4ecc-a2ad-d273e66721c7 · inbound
Rank-Conditioned Sample Reuse for the Plackett--Luce Best-of-$K$ Objective RSPO: Risk-Seeking Policy Optimization for Pass@k and Max@k Metrics in Large Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.