Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T14:58:39.053842Z
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 92 of 92 outbound references and 2 inbound Pith citation observations for arXiv:2411.15247.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T14:58:39.053842Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:18:52.995998Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-09T02:25:55.861273Z
92 of 92 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 4956b212-383f-4879-a6de-c4039a1d6f2f · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33e0345a-05c6-4dd6-9f27-b05b6d98d40a · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Training Diffusion Models with Reinforcement Learning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53ef57c1-d77f-42c6-9d61-cb2236dc3379 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Unresolved cited work
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04a93791-d50c-4285-9cee-48a4c7d2d15b · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Find: Fine- tuning initial noise distribution with policy optimization for diffusion models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 063509e9-44c3-4c28-9e40-c6c08a4b074d · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Diffusion policy: Visuomotor policy learning via action dif- fusion
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a784ceb9-4a1a-4439-9f82-86125c1e4242 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Simple Drop-in LoRA Conditioning on Attention Layers Will Improve Your Diffusion Model
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b9c41bb-c5f3-43a0-bac4-4a9d78235435 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Directly Fine-Tuning Diffusion Models on Differentiable Rewards
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b116dddd-8813-4d67-950a-98999f5a3685 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Prdp: Proximal reward difference prediction for large-scale reward finetuning of diffusion models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac517b83-c32f-413f-a180-cb050211f455 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6334cba3-c85c-43e0-9acc-8226be16a26f · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Sigmoid- weighted linear units for neural network function approxima- tion in reinforcement learning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc6cd460-0221-4b79-904b-ce95655e2ce3 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Optimizing DDPM Sampling with Shortcut Fine-Tuning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08a8f9e5-b91f-4ca4-b9b6-813c58c87cb3 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Dpok: Reinforcement learning for fine-tuning text-to-image diffu- sion models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ffd2d40-18c7-496a-a45b-9d4805ca6c59 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Re- inforcement learning for fine-tuning text-to-image diffusion models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c13d1e4a-0978-4357-9cea-4209d18f30d4 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Sharpness-Aware Minimization for Efficiently Improving Generalization
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb2a3cc9-bdf1-43e2-a0d2-b26849d51e96 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Scaling laws for reward model overoptimization
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62f168b6-b1ef-42f5-8e7f-0620b8f304d5 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Leveraging Reward Gradients For Reinforcement Learning in Differentiable Physics Simulations
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 37b759e8-fd66-4f24-91f1-77e73517f37f · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Jaddipal, Harish Prabhala, Sayak Paul, and Patrick V on Platen
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa30f412-c213-4817-ad9e-4387695947ad · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Gans trained by a two time-scale update rule converge to a local nash equilib- rium
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de9cc2bb-282c-4362-8563-82a79859099e · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Denoising dif- fusion probabilistic models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 807388fc-c788-4a6a-b8c3-5edeaefc7d63 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Imagen Video: High Definition Video Generation with Diffusion Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 970fb646-cd2f-4464-991d-f502e130a120 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Video dif- fusion models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d35f34fe-097d-462e-af24-155bc43b102f · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward LoRA: Low-Rank Adaptation of Large Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc3c9d5a-c480-49d6-8401-93bfda62b158 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward DiffTaichi: Differentiable Programming for Physical Simulation
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1c2fbee-8da1-49c6-8f14-3823f451c2b1 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward T2i-compbench: A comprehensive bench- mark for open-world compositional text-to-image genera- tion
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 70dbd66f-473a-4e3b-bf99-3a13252ab521 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward PlasticineLab: A Soft-Body Manipulation Benchmark with Differentiable Physics
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f6d658b-eef8-416e-bbb6-1dae83dc9999 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Planning with Diffusion for Flexible Behavior Synthesis
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c9eb66b-3a40-436a-a6f1-9068939ced38 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Information-theoretic local minima characterization and regularization
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 3f12abb2-e6e3-49ec-9729-75ed3d84c804 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Semantically robust unpaired image translation for data with unmatched seman- tics statistics
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 44def105-7c7f-453e-b089-2ac9bf6abd7c · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Pick-a-pic: An open dataset of user preferences for text-to-image generation
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 54ec5eb8-ed05-4cfd-bef9-bbe3ad5d93d2 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Actor-critic algorithms
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 26eca7f6-7cc3-4f95-8026-49c23f3786cc · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward DiffWave: A Versatile Diffusion Model for Audio Synthesis
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1219b57-2b0d-4b59-8ec6-774889cc45dc · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Plastic: Improving input and label plasticity for sample efficient reinforcement learning
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ed217d12-bf78-42c3-8444-ca800a329ca1 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Aligning Text-to-Image Models using Human Feedback
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 426e77d0-439d-4010-a3df-788bb5992f8b · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Playground v2.5: Three insights towards enhancing aesthetic quality in text-to-image genera- tion, 2024
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e952901-1710-4100-8cb8-c7fb0d04e1f6 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 833795ae-5fb0-4cce-b0a2-10acf5df95b5 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Reward Guided Latent Consistency Distillation
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69395d4b-3ae2-4f8f-8009-49bc795e32c9 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Aesthetic Post-Training Diffusion Models from Generic Preferences with Step-by-step Preference Optimization
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94da24d4-dcce-44aa-8fd0-74a001d3e454 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward SDXL-Lightning: Progressive Adversarial Diffusion Distillation
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12587c89-f915-47dd-ac5b-a89e8d3487ff · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Flow Matching for Generative Modeling
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2cb3db2-d51c-4d46-afbd-935e7d0320cf · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91c8a2b3-df27-4d2a-a253-fe2e6695f84f · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 893a082c-8042-4fcb-b202-2371cf8a315a · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Instaflow: One step is enough for high-quality diffusion- based text-to-image generation
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 52adc6f8-adda-49de-8ee4-42dd6a15c882 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Dpm-solver: A fast ode solver for diffusion probabilistic model sampling in around 10 steps
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22933794-341a-4fcb-a0b0-a6cc7baf5143 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Llmscore: Unveiling the power of large language models in text-to-image synthesis evaluation
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 67dae9df-1169-4987-abbb-677a4c81486a · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55790ba6-705b-4cae-bdca-178e57af356e · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Tuning Timestep-Distilled Diffusion Model Using Pairwise Sample Optimization
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a2114c3-789c-489d-9a83-280487036bac · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Playing Atari with Deep Reinforcement Learning
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7dc0934f-7e35-46b6-a7a3-7fc7ada010c4 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward RL for Consistency Models: Faster Reward Guided Text-to-Image Generation
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b9d41a1-f757-4a60-a84a-8fa0cf4895b5 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Reinforcement learning by reward-weighted regression for operational space control
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation eb00e920-4117-47c6-b4d9-4ab0d21b34bc · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2449c923-6564-4ff5-aba9-913a2621f1b0 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward DreamFusion: Text-to-3D using 2D Diffusion
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68159ef7-af13-4390-b128-ad85df3132b3 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Aligning Text-to-Image Diffusion Models with Reward Backpropagation
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04f259d4-dd69-478b-a65e-3d8d5242cd90 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Learning transferable visual models from natural language supervi- sion
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation db6d2560-42bb-4d7b-b08d-fba700087d6a · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Direct preference optimization: Your language model is secretly a 10 reward model
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 740b5d42-5f45-4e04-9d98-b7942b49e40f · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Zero-shot text-to-image generation
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 116e231f-ab18-4281-b566-b42848688ee3 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Hierarchical Text-Conditional Image Generation with CLIP Latents
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a84af830-1ccd-42b1-b4a1-78ea3bd01d05 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Blattmann, Dominik Lorenz, Patrick Esser, and Bj ¨orn Ommer
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f97ebbf-52cd-40c9-912d-d5bd3eeab4e8 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward High-resolution image synthesis with latent diffusion models
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8138986-6bf7-4ab4-824d-6e294cf4f68f · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Photorealistic text-to-image diffusion models with deep language understanding
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1fe4571-3ae4-4759-82c0-3580cfde86f0 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Progressive Distillation for Fast Sampling of Diffusion Models
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 980a99a2-8562-4d42-a2bf-3f5a99c467aa · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Adversarial Diffusion Distillation
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08982988-cccc-4d77-b933-3f9c5fc5145f · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Laion-5b: An open large-scale dataset for training next generation image-text models
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6623e559-72de-4988-8609-2a87407d092d · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward High-Dimensional Continuous Control Using Generalized Advantage Estimation
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d36fd569-e27d-4fe6-b63a-699e9472fe01 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Proximal Policy Optimization Algorithms
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa9157f3-50dd-4f0b-be06-dde56a0dcffe · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Defining and characterizing reward gam- ing
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 9aa05414-3768-4d14-a3b0-7dbcf81345e0 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Deep unsupervised learning using nonequilibrium thermodynamics
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 8605e395-a1dd-47d0-9ea7-f61f77fa2d01 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Denoising Diffusion Implicit Models
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab735257-0d42-42a1-9901-6451a7269602 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Generative modeling by esti- mating gradients of the data distribution
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 002e913c-ce96-4f55-8834-a5a1dbf00c85 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Score-Based Generative Modeling through Stochastic Differential Equations
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c766e8d-85f8-417c-98b8-21bd2df53122 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Consistency Models
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c638806f-4d08-43eb-8dad-f36c3ddba8c2 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Learning to predict by the methods of temporal differences
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 865461fe-4248-4167-9106-cba6d4133969 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Reinforcement learning: An introduction
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 0b139d73-91ad-4186-808e-8078d7e0a4b4 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Policy gradient methods for reinforcement learning with function approximation
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 6119596e-7044-47ad-96c7-16f6026bfc4c · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Evalalign: Supervised fine- tuning multimodal llms with human-aligned data for evalu- ating text-to-image models
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 1d12d40f-703f-492b-b32d-5c528e7b1993 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Understanding Reinforcement Learning-Based Fine-Tuning of Diffusion Models: A Tutorial and Review
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation efbe9370-c42f-40eb-9a2c-c4c4aef6c2b8 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Diffusion model align- ment using direct preference optimization
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 63f6d466-fee8-4295-9b02-48f90abdf1fe · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward DiffusionDB: A Large-scale Prompt Gallery Dataset for Text-to-Image Generative Models
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33b71460-13dd-45a0-8353-38374dca330c · outbound
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 23d2f4da-3bfe-4bab-86d1-dc457a15ee92 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Human Preference Score: Better Aligning Text-to-Image Models with Human Preference
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64ac4e93-5edc-4d61-b771-01cc32bbd108 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Human preference score: Better aligning text- to-image models with human preference
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 5307d521-1177-4f78-9caf-38f6c701cc11 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Deep Reward Supervisions for Tuning Text-to-Image Diffusion Models
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db15eef3-f273-453d-8e5f-b9c4f2b97d82 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Group normalization
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df6f2117-8297-4a6c-8d00-aa36262364f2 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Florence-2: Advancing a unified representation for a variety 11 of vision tasks
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a9a38e18-66b3-494c-95d4-5496f81f5525 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Imagere- ward: Learning and evaluating human preferences for text- to-image generation
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f3520083-701d-46da-bfa3-fc2aae3aa855 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Using human feedback to fine-tune diffusion models without any reward model
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 249e130b-fb6d-4d7a-9c99-f83594cd159c · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward LION: Latent Point Diffusion Models for 3D Shape Generation
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98d14ea4-6a43-473b-b6c7-a6bdf01c18dc · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward UniFL: Improve Latent Diffusion Model via Unified Feedback Learning
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26e3d646-eca1-4992-9bad-a0a625236a02 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward BERTScore: Evaluating Text Generation with BERT
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f633b17b-17db-446c-acbf-441bcae44a79 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Unresolved cited work
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation fab0be03-741d-4961-a996-eb993af0463f · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Unresolved cited work
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 6aed4586-243c-4a3e-8458-e75abd49783c · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Specifically, we use the CLIP-ViT-L/14 from the original CLIP paper and a similar BLIP model fine-tuned for VQA tasks (ViT-L for the vision backbone)
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f866bc04-e632-4dd5-85a4-f3710fb780c6 · outbound
Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward Note that, SDXL-Turbo mainly focuses on image generation of 5122 pixels and SDXL-Lightning only supports≥ 2 step generation
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 007c5f7c-0714-4c6f-bc71-fb2220e2149c · inbound
VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0d7ae0d-6bd3-4e43-b65f-4847ff0c983d · inbound
Selective Timestep Weighting and Advantage-Based Replay for Sample-Efficient Diffusion RLHF Reward Fine-Tuning Two-Step Diffusion Models via Learning Differentiable Latent-Space Surrogate Reward
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.