Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:25:26.232105Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 4 inbound Pith citation observations for arXiv:2505.19196.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:25:26.232105Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:18:33.490433Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-06-29T08:03:13.983158Z
57 of 57 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 4a59cfb8-d6c1-4a4a-98d1-7b9a0c718506 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Deep unsupervised learning using nonequilibrium thermodynamics,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b9b42b1d-c934-4fe3-a11f-9245bc2f86dc · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Denoising diffusion probabilistic models,
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9f7afaf-0e9b-44a9-8d97-aa9f714e2f5e · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Denoising Diffusion Implicit Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 616f4244-65ae-41b0-af70-5fc41b127016 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Extrapolating and Decoupling Image-to-Video Generation Models: Motion Modeling is Easier Than You Think
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9ba1b6e-ed7a-4f4a-8446-0f66494ba7c5 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Generative Adversarial Networks
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36fc2aa5-ff92-4347-bb87-46614c818290 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Learning transferable visual models from natural language supervision,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a8cf128a-03e1-4295-ba87-01beadc4c0b3 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning BLIP: Bootstrapping language-image pre-training for unified vision-language understanding and generation,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c2c730e0-bca1-4d26-9780-15e0f95475fe · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Exploring the limits of transfer learning with a unified text-to-text transformer,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 31e3e112-8969-48c6-af7e-3303037df51c · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Microsoft COCO: Common Objects in Context
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5035b145-6b7c-4329-8388-b38f41b8f4b9 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Imagenet: A large-scale hierarchical image database,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5969777-9350-40a4-a3bd-dc39466c7730 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning High-resolution image synthe- sis with latent diffusion models,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 54a42102-2d77-47b5-891a-c44c17bdd6c9 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Improving image generation with better captions,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f7a42b79-f648-40cd-9bd7-d5abb12dc7c7 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Laion-5b: An open large-scale dataset for training next generation image-text models,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 73f0cfd5-16ce-492f-b66b-8a3922494028 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Training-free structured diffusion guidance for compositional text-to-image synthesis,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ac972a64-6cde-41a2-b6f0-db927337c602 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning TIFA: Accurate and Interpretable Text-to-Image Faithfulness Evaluation with Question Answering
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83dfc91e-b25e-4b5a-997d-134f4b12d7c2 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Human Preference Score v2: A Solid Benchmark for Evaluating Human Preferences of Text-to-Image Synthesis
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9ec2f17-2af9-4049-87ef-f4529c2dd380 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Imagereward: learning and evaluating human preferences for text-to-image generation,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b4ae061a-1de1-4a54-b92f-fc9145c7f647 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Pick-a-pic: an open dataset of user preferences for text-to-image generation,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9c0f8328-a04a-43ae-9372-84076f1c2057 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning OpenThinkIMG: Learning to Think with Images via Visual Tool Reinforcement Learning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffc72921-f5e6-4523-9723-d23778d31cd0 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Dpok: reinforcement learning for fine-tuning text-to-image diffusion models,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a49943ae-083e-4f04-924f-c5f98e89a191 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Training Diffusion Models with Reinforcement Learning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d82d039-20cc-42ab-9944-ced4ffa73f96 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Proximal Policy Optimization Algorithms
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5d1d4d8-172b-4707-b6b7-55870db39f66 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Direct prefer- ence optimization: Your language model is secretly a reward model,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0af8666d-0e7e-4f19-a634-d622111e675f · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Stimulating diffusion model for image denoising via adaptive embedding and ensembling,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation baa652e8-00ce-48a7-bf95-b584a9bdb855 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Blue noise for diffusion models,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7a764e69-28ab-4476-bac1-3702711a3a7c · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Boosting Diffusion Models with Moving Average Sampling in Frequency Domain
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80e01d5d-5291-4403-a1f0-853487a63e6a · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning FreSca: Scaling in Frequency Space Enhances Diffusion Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 293a2962-4691-4cb0-bce0-88f8b86fb19b · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning A dense reward view on aligning text-to-image diffusion with preference,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1f77c922-b895-41d0-a9a4-06d761f17ebe · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Confronting reward overoptimization for diffusion models: A perspective of inductive and primacy biases,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7ddc2152-916d-4b1e-b35e-103cb7023a17 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Aesthetic Post-Training Diffusion Models from Generic Preferences with Step-by-step Preference Optimization
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11dccaa8-3582-4b83-b462-f53b6771e70d · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Diffusion models beat gans on image synthesis,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 186e990f-a787-4d59-ab3d-7b884ffe7fd5 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Classifier-Free Diffusion Guidance
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edd2dd8b-19b6-4ce7-bea8-43db8f0eb1bd · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Aligning Text-to-Image Models using Human Feedback
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7795060c-6414-4eb9-821b-c97d09fe8840 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03783dd0-c41f-4932-9fd6-780499cd33e4 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Optimizing ddpm sampling with shortcut fine-tuning,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 44e5fb01-4654-4a95-ae6e-54baca475030 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Deep reward supervisions for tuning text-to-image diffusion models,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 44045a80-0c4d-4cde-b168-b89721392dd0 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Using human feedback to fine-tune diffusion models without any reward model,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4e8fbaae-74bc-41a7-8ce7-76a0602f52ed · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Diffusion model alignment using direct preference optimization,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 357d6f4e-48c6-430d-9274-161ff9e8f422 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Steps toward artificial intelligence,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d0ad225d-b9e1-4768-80e0-10a61a269b37 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Temporal credit assignment in reinforcement learning,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 65d6c08a-74bd-40f4-9270-b8aac917d6a7 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Learning guidance rewards with trajectory-space smooth- ing,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 33f72736-f8f7-4112-a81b-668036b499ca · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Harutyunyan, W
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3724d282-7afc-42bc-aa97-5da828ff567a · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Policy invariance under reward transformations: Theory and application to reward shaping,
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d0a81003-26d7-43ce-ba62-ae59e4f1d1e5 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Text2Reward: Reward Shaping with Language Models for Reinforcement Learning
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c99752e1-3a18-426a-8a2b-ebd7f2dcc812 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning DPO meets PPO: Reinforced token optimization for RLHF,
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7a6199c3-bf2b-4c68-ba3a-3e3b22308795 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning R3HF: Reward redistribution for enhancing reinforcement learning from human feedback,
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9469099a-82cb-40f7-a438-e3935c0b55f9 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Dense reward for free in reinforcement learning from human feedback,
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1be93a43-71a6-4aa1-8b09-4246ea680c60 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Simple statistical gradient-following algorithms for connectionist reinforcement learning,
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 69347cfb-4e3b-468c-9ed9-27beddb7a661 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Auto-Encoding Variational Bayes
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6951719-dd51-470b-898e-5a1b0c4726a2 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Emerging properties in self-supervised vision transformers,
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dabf1354-2074-4903-b190-7c30f48f6d6d · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning DiffSim: Taming Diffusion Models for Evaluating Visual Similarity
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92df4940-22bb-4894-8c43-65ab8291f10c · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning LoRA: Low-rank adaptation of large language models,
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1762ec38-ef64-4b33-9f98-a3eb686a4f25 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Policy gradient methods for rein- forcement learning with function approximation,
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f8186217-4635-4ead-9ac7-55d49981bf93 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning U-Net: Convolutional Networks for Biomedical Image Segmentation
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e00ca729-7998-4e5e-9493-2baae7aac08c · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Training deep nets with sublinear memory cost,
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f5acd2f4-921b-43af-b595-29fdb03e2953 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Training Deep Nets with Sublinear Memory Cost
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb825529-10da-4a0b-bf1e-6db2928fdaf5 · outbound
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning Available: https://openreview.net/forum?id=eyxVRMrZ4m
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 45e9e2b8-63b4-4484-a920-a61a0bdab5c2 · inbound
Draw ALL Your Imagine: A Holistic Benchmark and Agent Framework for Complex Instruction-based Image Generation Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb275552-9d13-432f-8f4c-003c320e670b · inbound
LatSearch: Latent Reward-Guided Search for Faster Inference-Time Scaling in Video Diffusion Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71e0a791-1444-4142-a587-66be70140e9f · inbound
LeapAlign: Post-Training Flow Matching Models at Any Generation Step by Building Two-Step Trajectories Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e2f44700-b3bf-4fee-b294-a22df8a5e414 · inbound
VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.