Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T14:42:22.457966Z
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 62 of 62 outbound references and 3 inbound Pith citation observations for arXiv:2508.21016.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T14:42:22.457966Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-30T07:10:44.244790Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-06-30T07:14:20.959214Z
62 of 62 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation adfc3e4c-ff7a-411f-b5c9-795c3ef9f0f3 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance , " * write output.state after.block = add.period write newline
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7749846b-3fe2-4b42-adb7-26910a877101 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance write newline
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3458c617-ac39-4146-87a9-0dfe06b458d1 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6f991acc-92d2-4ab6-87ba-6bb34af8a437 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Training Diffusion Models with Reinforcement Learning
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed72c517-5e42-4777-91cf-93adc6dd41f1 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Classifier-Free Guidance is a Predictor-Corrector
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56785e19-4ba3-4acd-8221-afbd7bc3d6ac · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance PrefPaint: Enhancing Medical Image Inpainting through Expert Human Feedback
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 304f57c7-e897-46a5-b7c1-f449696b9cb4 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b737d5e-61b9-4f3b-9691-8ab6edf40637 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Masked-attention Mask Transformer for Universal Image Segmentation
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cddf071c-3a23-402e-b3d8-965bcde8eeaf · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c8f76b2-7df2-467d-ba51-cc68b1d6ada2 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance CFG++: Manifold-constrained Classifier Free Guidance for Diffusion Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f3cbfdb-5f44-4ee7-81a6-de6417e866a0 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Directly Fine-Tuning Diffusion Models on Differentiable Rewards
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5465019d-c0e2-4632-a8c3-1667283e239f · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Process Reinforcement through Implicit Rewards
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c02521e-4250-43a0-93f6-f6e08bd8a148 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c0a4618d-1253-49ef-b432-cdf07356bcb5 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Scaling Rectified Flow Transformers for High-Resolution Image Synthesis
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b73787a-0011-4fb3-8fcf-a88a0e1b9001 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Online Reward-Weighted Fine-Tuning of Flow Matching with Wasserstein Regularization
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8575a2c6-fc01-49ce-b6a8-64bd81386868 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance CFG-Zero*: Improved Classifier-Free Guidance for Flow Matching Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56e107ec-6e8c-4c85-a1b2-0b52646da785 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Diffusion Guidance Is a Controllable Policy Improvement Operator
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76ed0a85-9501-4c3c-abaf-e9390e43b717 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 221f89fb-3a73-4331-9671-19a81ecf03de · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance DiffuCoder: Understanding and Improving Masked Diffusion Models for Code Generation
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 768c7664-f57e-49fb-87f7-e3afe05f5220 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b05b307f-f8b5-4cab-87c5-c840d0a02599 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Classifier-Free Diffusion Guidance
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c97259dd-7afb-4d2e-ba2f-0afb4f919504 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 828132b4-1a46-49e5-a446-68bfdcad2680 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 61bcc3c4-df59-47f4-a972-b591474cea3b · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Pick-a-Pic: An Open Dataset of User Preferences for Text-to-Image Generation
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e1a7866-4c6d-4c1a-90a4-0b68584b0c53 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance HunyuanVideo: A Systematic Framework For Large Video Generative Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1195b3a-2b1b-4f4f-bce8-e7854c0f46bf · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Aligning Text-to-Image Models using Human Feedback
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e181c30-f54a-4486-95be-d5e9ec5dbe92 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e3f21ad-9081-4c4f-b264-8e22dd71db05 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Aesthetic Post-Training Diffusion Models from Generic Preferences with Step-by-step Preference Optimization
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2057f0dc-e59f-4592-9d2d-d65e6146a1fa · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Flow Matching for Generative Modeling
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 052041bf-f72c-4588-b694-ec3c62b8be08 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60d87b09-6a4d-4dd1-8425-f695b0ec1a0c · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Lumina-Video: Efficient and Flexible Video Generation with Multi-scale Next-DiT
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 281d1151-1fa6-4a4e-9e2b-8c1f98fa3d35 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Flow-GRPO: Training Flow Matching Models via Online RL
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b12ae850-7b16-4c10-805f-31fbfd4f1039 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e942faf-4539-4ea9-8ff8-5bc4e1c193c7 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c62698a-b136-4733-ab0f-6bb80427dfe0 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 74cf2b01-4e6b-4cbe-990c-f68593a31c32 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0e5c2df1-3b3f-40ed-9854-8c2889f00b69 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c5bd75d-e8e1-49e2-aa06-5ac443293669 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7cdceb0a-4e7e-4be4-8d08-d5f493990d01 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0c2b9ab9-deeb-41f9-b4e5-624a56116996 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Learning Transferable Visual Models From Natural Language Supervision
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 506981d4-0316-4bc6-be19-363c8232b831 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance From $r$ to $Q^*$: Your Language Model is Secretly a Q-Function
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 410a884b-1889-4fa0-8759-aa64e1820ae6 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance D.; Ermon, S.; and Finn, C
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f58168f-ffa2-4e1a-8301-abe13d006b47 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b21c10ea-7e63-4108-8bec-b8b4d627db30 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6aaf30ea-0ec4-44a1-9d12-f3031577a7fd · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2cee3b5a-d185-456f-922c-40af865bfa62 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a1d1980c-d34a-4fce-90d4-86515ab12e5a · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Proximal Policy Optimization Algorithms
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe7eb3a0-829a-4352-9bdd-6173339a1843 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1e67a8e-2637-4a3f-b8ca-b11796ff644a · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Feynman-Kac Correctors in Diffusion: Annealing, Guidance, and Product of Experts
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6863f946-2ed8-4327-879a-be101727b78e · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Denoising Diffusion Implicit Models
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6961d1a9-94a6-4f57-8a5c-a70ea13c91c4 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Score-Based Generative Modeling through Stochastic Differential Equations
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4e63a3a-d6be-4b11-9a66-f6f111ac4947 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 353e51e4-fb10-4da7-9abe-5c4b661d1f2c · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Improving and generalizing flow-based generative models with minibatch optimal transport
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation abe3e218-3c61-4a80-9874-7cd98d65b5ff · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance What is the Alignment Objective of GRPO?
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68f4abfb-8c0b-4976-8c74-286b39d2b6ef · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 62bb897e-36e3-40ae-8ca4-72e9363200ce · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Wan: Open and Advanced Large-Scale Video Generative Models
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44622b7b-0a22-457a-b400-959c20f5e6e8 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation cca8fd07-7dfe-4271-b53f-18e569c19dd2 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 8be7027c-db0b-463f-8eef-c1cfa47eaa20 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6fc77f98-4575-4218-8e62-c84e9a7b0553 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 8011e350-6a7b-424d-8960-13d85d075101 · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Guided Flows for Generative Modeling and Decision Making
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1bc722fa-6767-4acd-b577-a6dde6f7111c · outbound
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6674a156-2fbb-43f2-9c5f-0341df19c5b6 · inbound
DiffusionNFT: Online Diffusion Reinforcement with Forward Process Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a4fb9b96-bf1d-4ddb-bb27-01dc1c977e51 · inbound
Towards General Preference Alignment: Diffusion Models at Nash Equilibrium Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation fb3adc17-f5fd-4baa-a50b-90e33b0a143d · inbound
FlowAWR: Online Adaptive Flow Reinforcement via Advantage-Weighted Rectification Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.