Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T10:22:38.305872Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 0 inbound Pith citation observations for arXiv:2509.04372.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T10:22:38.305872Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
21 of 21 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation cd3082b0-8c74-4f96-9715-d129ba5474e2 · outbound
Connections between reinforcement learning with feedback,test-time scaling, and diffusion guidance: An anthology (2017).First-order methods in optimization
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c0a7c2e3-2b2c-4633-a31a-a452fca5ac6c · outbound
Connections between reinforcement learning with feedback,test-time scaling, and diffusion guidance: An anthology Training Diffusion Models with Reinforcement Learning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d7a47e9-e20b-47e4-a3af-42b94d32b2f8 · outbound
Connections between reinforcement learning with feedback,test-time scaling, and diffusion guidance: An anthology Classifier-Free Guidance is a Predictor-Corrector
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5316555-f017-4eb3-997b-825b26f673a9 · outbound
Connections between reinforcement learning with feedback,test-time scaling, and diffusion guidance: An anthology Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7cb71b9e-b355-49db-8a2a-5e754e480e46 · outbound
Connections between reinforcement learning with feedback,test-time scaling, and diffusion guidance: An anthology Large Language Monkeys: Scaling Inference Compute with Repeated Sampling
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2bc5fd5c-f9d7-43fd-801f-395a9864940d · outbound
Connections between reinforcement learning with feedback,test-time scaling, and diffusion guidance: An anthology Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d25f6725-d7f1-4125-aef6-f7c103f4a153 · outbound
Connections between reinforcement learning with feedback,test-time scaling, and diffusion guidance: An anthology and Nichol, A
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 52fd0fb8-f6fb-40c0-8e73-058ed5d255e2 · outbound
Connections between reinforcement learning with feedback,test-time scaling, and diffusion guidance: An anthology Unresolved cited work
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 87b7f893-9cbd-424d-822b-08be91ae1a17 · outbound
Connections between reinforcement learning with feedback,test-time scaling, and diffusion guidance: An anthology Reward-Directed Score-Based Diffusion Models via q-Learning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85aa4d1d-64c5-48f1-a16c-1f46a4dd569d · outbound
Connections between reinforcement learning with feedback,test-time scaling, and diffusion guidance: An anthology and Salimans, T
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a6e5b84c-2ae1-426f-a044-fad4ddd19f8a · outbound
Connections between reinforcement learning with feedback,test-time scaling, and diffusion guidance: An anthology and Martin, J
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation cb943ee2-cf26-40a3-b4b8-b7879a812195 · outbound
Connections between reinforcement learning with feedback,test-time scaling, and diffusion guidance: An anthology Unfamiliar Finetuning Examples Control How Language Models Hallucinate
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f72b04ac-3d8e-4a12-99b1-9aa8f46544ba · outbound
Connections between reinforcement learning with feedback,test-time scaling, and diffusion guidance: An anthology Unresolved cited work
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4add29e6-a4f2-4a2b-bfaf-c9620c485b75 · outbound
Connections between reinforcement learning with feedback,test-time scaling, and diffusion guidance: An anthology Unresolved cited work
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22f5eaa6-73ef-45f9-8a11-9d7b3ef3a84d · outbound
Connections between reinforcement learning with feedback,test-time scaling, and diffusion guidance: An anthology A Score-Based Density Formula, with Applications in Diffusion Generative Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 30c8b382-262b-445b-bdfe-55857c7d060b · outbound
Connections between reinforcement learning with feedback,test-time scaling, and diffusion guidance: An anthology s1: Simple test-time scaling
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 205eccfc-630f-4c06-9fda-1d0ef19f56d4 · outbound
Connections between reinforcement learning with feedback,test-time scaling, and diffusion guidance: An anthology Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 816f2ba0-a9dc-4797-a14e-7bedb121976a · outbound
Connections between reinforcement learning with feedback,test-time scaling, and diffusion guidance: An anthology Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4018f26-4601-459e-a03d-bbf657a72bf4 · outbound
Connections between reinforcement learning with feedback,test-time scaling, and diffusion guidance: An anthology Soft Best-of-n Sampling for Model Alignment
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54f756f7-d16c-4324-b4af-d85ed1f561e5 · outbound
Connections between reinforcement learning with feedback,test-time scaling, and diffusion guidance: An anthology Learning to Reason without External Rewards
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b663c7f-160e-49b9-bad5-fbae983d3217 · outbound
Connections between reinforcement learning with feedback,test-time scaling, and diffusion guidance: An anthology Fine-Tuning Language Models from Human Preferences
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.