Pith. sign in

Paper Citation Record · LEDGER

Aligning Text-to-Image Diffusion Models with Reward Backpropagation

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 43 inbound Pith citation observations for arXiv:2310.03739.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2310.03739 v5

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 43 of 43 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:37:03.307786Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T16:59:58.107705Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 362d5f10-9edc-4426-ab6d-712020a60900 · inbound

Improving Video Generation with Human Feedback cites this paper.

Improving Video Generation with Human Feedback Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-13T15:30:02.776052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T15:30:02.578430Z digest=sha256:4a91fe32268611b6e65c483c79d5cb57150591ca791654d48a7451ae76e9b705

Observation 2f3eefa8-09b3-454a-b908-aa55c5d285ce · inbound

Flow-GRPO: Training Flow Matching Models via Online RL cites this paper.

Flow-GRPO: Training Flow Matching Models via Online RL Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:45:16.889213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T18:45:16.641012Z digest=sha256:2f361fdf3b3395b58424a12bc35200a0ed83930649c8e5869291cbf274cbcf4f

Observation a94bb423-231c-4428-a409-657d08cdb7bf · inbound

Sample and Computationally Efficient Continuous-Time Reinforcement Learning with General Function Approximation cites this paper.

Sample and Computationally Efficient Continuous-Time Reinforcement Learning with General Function Approximation Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T15:37:03.307786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:37:03.307786Z digest=sha256:40494f5abde24dc81d860b60251dc4c184d36dbb1eadc25714259ce74e85a909

Observation d56fca1a-9741-4051-9fe7-488528d479cf · inbound

$I^2G$: Generating Instructional Illustrations via Text-Conditioned Diffusion cites this paper.

$I^2G$: Generating Instructional Illustrations via Text-Conditioned Diffusion Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T15:03:40.138551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:03:40.138551Z digest=sha256:483876380ef3eab2d7a4267c3e16e7579e4ffe80b50f5fb061c287f9d89e35b4

Observation 33d88b77-13f3-4f3f-b26c-4b3a257eee9c · inbound

Alignment and Safety of Diffusion Models via Reinforcement Learning and Reward Modeling: A Survey cites this paper.

Alignment and Safety of Diffusion Models via Reinforcement Learning and Reward Modeling: A Survey Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T02:30:56.253451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T02:28:23.754561Z digest=sha256:04292b0858631a96d7ff06e59e0c6144035643ceb15ff1afdaddb21a747ae894

Observation 9bae9673-95ce-45d3-b8cc-9c491ca3ae17 · inbound

DiffusionReward: Enhancing Blind Face Restoration through Reward Feedback Learning cites this paper.

DiffusionReward: Enhancing Blind Face Restoration through Reward Feedback Learning Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:43:12.628795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:43:12.628795Z digest=sha256:9f63e874f169bc608634b6be3c442cc4c850301bea86df17007710e51cab8d39

Observation 965d89f6-03d8-410a-be39-57d56b62377d · inbound

ImageReFL: Balancing Quality and Diversity in Human-Aligned Diffusion Models cites this paper.

ImageReFL: Balancing Quality and Diversity in Human-Aligned Diffusion Models Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:40.872626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:09:40.872626Z digest=sha256:2aca471b19d9adb828db98103f18ba73018ddc1787177ef4e4c3e6d4807df884

Observation 5f86de4e-66d1-438e-89c0-1431fd1b70c6 · inbound

Rhetorical Text-to-Image Generation via Two-layer Diffusion Policy Optimization cites this paper.

Rhetorical Text-to-Image Generation via Two-layer Diffusion Policy Optimization Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T13:04:53.458718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:04:53.458718Z digest=sha256:267f673d38ba5c8942979aaf6a8e521b1b11061a89848e6e473b26423679fd88

Observation a11844ff-4d5b-4363-9403-43d511a4badf · inbound

Local Manifold Approximation and Projection for Manifold-Aware Diffusion Planning cites this paper.

Local Manifold Approximation and Projection for Manifold-Aware Diffusion Planning Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T12:03:34.513509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:03:34.513509Z digest=sha256:d3f44269ecf79b72233ba5e485949119c7324828d7f40cbd42dbf537d8362da3

Observation b9966f4a-cca6-4203-8db1-7c4ffaa55f5c · inbound

Enhancing Diffusion-based Unrestricted Adversarial Attacks via Adversary Preferences Alignment cites this paper.

Enhancing Diffusion-based Unrestricted Adversarial Attacks via Adversary Preferences Alignment Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:49.716814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:47:49.716814Z digest=sha256:2adc2cd4b496b26972bfb10576cf5bc32a9012d8ee0b81f745e5228dea284c6f

Observation 78489e33-44c9-44ff-984f-fcd44c666a83 · inbound

Smoothed Preference Optimization via ReNoise Inversion for Aligning Diffusion Models with Varied Human Preferences cites this paper.

Smoothed Preference Optimization via ReNoise Inversion for Aligning Diffusion Models with Varied Human Preferences Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T11:26:57.915179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:26:57.915179Z digest=sha256:d74ed5eca617fbe6f7cd252e963e8237bba9a828a01967607250580cfba3ff94

Observation b9bc263d-8b48-4040-a745-e44fbab45408 · inbound

Text2Stereo: Repurposing Stable Diffusion for Stereo Generation with Consistency Rewards cites this paper.

Text2Stereo: Repurposing Stable Diffusion for Stereo Generation with Consistency Rewards Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T13:26:13.814256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:26:13.814256Z digest=sha256:8f105ec35c4e2eda38f9a6bcd94f4f7425444943378769fc057b1eb541b6f8f6

Observation af975965-de9c-40e7-a545-2b1d53270c33 · inbound

AssetDropper: Asset Extraction via Diffusion Models with Reward-Driven Optimization cites this paper.

AssetDropper: Asset Extraction via Diffusion Models with Reward-Driven Optimization Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T10:20:01.173594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:20:01.173594Z digest=sha256:4a63457b3f6c9b752cf3084f0d882551ac4d795d8ca51c1ba06c689a2cf90e0a

Observation 69f41c5d-b01e-42e8-8ab7-65d145619497 · inbound

Diffusion Tree Sampling: Scalable inference-time alignment of diffusion models cites this paper.

Diffusion Tree Sampling: Scalable inference-time alignment of diffusion models Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T22:59:48.062413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:59:48.062413Z digest=sha256:2e19fa792442c8ecf7b19f7d6a21b35a9e17d20c94744d60e5d3c8a396837009

Observation 7e5c2d48-2831-46f9-941b-078d87e9a147 · inbound

ShortFT: Diffusion Model Alignment via Shortcut-based Fine-Tuning cites this paper.

ShortFT: Diffusion Model Alignment via Shortcut-based Fine-Tuning Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T11:32:42.332939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:32:42.332939Z digest=sha256:c32b32b70ed8e442fa2c545de724deeb8e88570596cebac707bf0180ae25e76b

Observation 7eafbac7-e96f-4597-a05f-a19098fdf3e5 · inbound

Instant Preference Alignment for Text-to-Image Diffusion Models cites this paper.

Instant Preference Alignment for Text-to-Image Diffusion Models Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T16:51:04.752840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:51:04.752840Z digest=sha256:682b1f5cd2a4883b23ee487b40a67dc5eb263ca00adaaf4dfb6ae970a7d81c59

Observation c2ced703-8af2-4cf7-afb6-d9e3a0805fd0 · inbound

A novel method and dataset for depth-guided image deblurring from smartphone Lidar cites this paper.

A novel method and dataset for depth-guided image deblurring from smartphone Lidar Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-04T19:29:20.881918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:29:20.881918Z digest=sha256:bb6b6ffc02fe018e6c4db291ddcf7c6998c42190758e53169f6267e51bdc8a16

Observation 63ae9124-ad20-4edb-9b33-ba687af16ae7 · inbound

Test-Time Alignment of Text-to-Image Diffusion Models via Null-Text Embedding Optimisation cites this paper.

Test-Time Alignment of Text-to-Image Diffusion Models via Null-Text Embedding Optimisation Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T20:11:26.841597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T20:11:26.841597Z digest=sha256:c8da282a42f5c3fd7a9662cb522c7193916a77443d4e6e2367a87576b2aa1dd7

Observation c13a0ffe-dbf4-4a85-83bb-4affe53f23ca · inbound

PSR: Scaling Multi-Subject Personalized Image Generation with Pairwise Subject-Consistency Rewards cites this paper.

PSR: Scaling Multi-Subject Personalized Image Generation with Pairwise Subject-Consistency Rewards Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:51:29.411384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T03:49:05.489626Z digest=sha256:0d56dc5936ac7ed09d2c39ba7e0802e8a1dfe0ce165bdda60ecec1bb56e3df63

Observation 39f8bac2-0db3-4499-bfcd-fa1625189ded · inbound

Direct Diffusion Score Preference Optimization via Stepwise Contrastive Policy-Pair Supervision cites this paper.

Direct Diffusion Score Preference Optimization via Stepwise Contrastive Policy-Pair Supervision Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-03T13:45:02.856954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:45:02.856954Z digest=sha256:40818b73e7d3b66706d74773a4473dd735bf8b525423b255edd31fe477d173ce

Observation a52b2a05-751a-4ddf-9a94-ef46be016991 · inbound

Dichotomous Diffusion Policy Optimization cites this paper.

Dichotomous Diffusion Policy Optimization Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-03T13:18:30.948259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:18:30.948259Z digest=sha256:2d3d4704a32331d1f7b44b97b953e14dc649fce1d91820e10e6033ed3a0881a5

Observation 422a6027-3b26-41b9-9977-77aba6521aa4 · inbound

Beyond VLM-Based Rewards: Diffusion-Native Latent Reward Modeling cites this paper.

Beyond VLM-Based Rewards: Diffusion-Native Latent Reward Modeling Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-25T06:45:26.127945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T06:41:33.493927Z digest=sha256:8280b1d49740086db33b8c8cd3e4d62010ffb2d6bc8fade83eeb9136aa39d6da

Observation d670fcd0-d0c0-4d5c-b6a1-0bbdc21f5e8e · inbound

Personalizing Text-to-Image Generation to Individual Taste cites this paper.

Personalizing Text-to-Image Generation to Individual Taste Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:30:56.087342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T18:07:31.236729Z digest=sha256:463a71a08d18853798334155d4570a5f524ebb9d34738c74d82d57f22d048512

Observation e6404207-6dd9-4711-92b5-a3955bcc7b52 · inbound

Adjoint Matching through the Lens of the Stochastic Maximum Principle in Optimal Control cites this paper.

Adjoint Matching through the Lens of the Stochastic Maximum Principle in Optimal Control Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-14T22:28:04.684338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T22:25:53.037164Z digest=sha256:e850ba97ddcd240d7c6511b4f27eeb8d40c71420850a4d10dc19d681ed815d46

Observation 1f61d26d-0033-46d0-9a79-5bc2c92bcb54 · inbound

Visual Generation in the New Era: An Evolution from Atomic Mapping to Agentic World Modeling cites this paper.

Visual Generation in the New Era: An Evolution from Atomic Mapping to Agentic World Modeling Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:16:29.010694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-07T06:38:04.459129Z digest=sha256:c18b725257b4b31c70dadc894ba599b3a8fcd4cd5c770a1c92b9bd8a0a8f049a

Observation bf842ae3-dd78-4859-baf6-f1178e2ebf38 · inbound

Reinforce Adjoint Matching: Scaling RL Post-Training of Diffusion and Flow-Matching Models cites this paper.

Reinforce Adjoint Matching: Scaling RL Post-Training of Diffusion and Flow-Matching Models Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:06:24.184520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-12T04:38:05.640892Z digest=sha256:6f635e5867faf790c36b0f2e0cf0debaa34a992b81e2842ba127f1d65ddd2292

Observation 6aa237e1-1ed0-4fe4-8de3-fdec4dd2e3e9 · inbound

Reinforce Adjoint Matching: Scaling RL Post-Training of Diffusion and Flow-Matching Models cites this paper.

Reinforce Adjoint Matching: Scaling RL Post-Training of Diffusion and Flow-Matching Models Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-20T22:29:09.789238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T22:24:45.862340Z digest=sha256:242b58d5792bd36b8a45a35351aec23c1c05f755c7b207bea39b553606cb29bc

Observation 40998a79-1f4e-4daa-8f98-6d2cc97bd482 · inbound

Efficient Adjoint Matching for Fine-tuning Diffusion Models cites this paper.

Efficient Adjoint Matching for Fine-tuning Diffusion Models Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:52:04.840735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T01:51:09.497417Z digest=sha256:d5fecb048e4b4e1b35b43d6f3bf95c48e87330bb5c2b04e52c091b10ed75e4d4

Observation 2b95bb4d-9932-4a6c-a7c1-f29b72a022f2 · inbound

Efficient Adjoint Matching for Fine-tuning Diffusion Models cites this paper.

Efficient Adjoint Matching for Fine-tuning Diffusion Models Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-20T22:09:07.033037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T22:08:21.562078Z digest=sha256:7e7bd6b465578d983d12fe18fd44dd802385d980d50a4f9fe3b13ba40fdc6358

Observation 998adf9c-7bc3-406d-abe3-4598238b86b9 · inbound

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution cites this paper.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:49:40.925892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:0bafb1a80e416d28ba2394ebba7676358373114460ab6afe9a84459f83ff6151

Observation a64017b0-2a58-48de-a8bd-7b78e7f43515 · inbound

AdvantageFlow: Advantage-Weighted Least Squares for RL in Flow Models cites this paper.

AdvantageFlow: Advantage-Weighted Least Squares for RL in Flow Models Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:54:01.604475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T22:44:25.152451Z digest=sha256:8f7bf535c96adebf7ef28693a0cdbabcee08263d7c81ef44dd090daea0d1a313

Observation ffc60efa-7beb-4ae4-8fa5-2f30ecc55940 · inbound

Parallel Tempering Initial Sampling in Inference-Time Reward Alignment cites this paper.

Parallel Tempering Initial Sampling in Inference-Time Reward Alignment Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-06-28T23:32:47.406224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T23:28:22.532621Z digest=sha256:f518b949ad47c36a35208e74e07b62f036d314ad3c79e7c7fdd7952f4239fff3

Observation 4a3d2124-8a1e-4e63-8e46-44d4eb9e3415 · inbound

Plug-and-Play Guidance for Discrete Diffusion Models via Gradient-Informed Logit Correction cites this paper.

Plug-and-Play Guidance for Discrete Diffusion Models via Gradient-Informed Logit Correction Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:16:56.665095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T02:20:35.909604Z digest=sha256:6c7ea7e9b7e73c44cd98906350fad7b9ee8197c07156048714896f304927b184

Observation e574da2a-b11b-4f6a-94ec-28853fa3a82d · inbound

Plug-and-Play Guidance for Discrete Diffusion Models via Gradient-Informed Logit Correction cites this paper.

Plug-and-Play Guidance for Discrete Diffusion Models via Gradient-Informed Logit Correction Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T04:56:02.912541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:56:02.912541Z digest=sha256:781514018248bd6619a2020af64935e712b26b032c48ab1165f7e6f5b38889b7

Observation bee45504-be20-4751-9c8b-0cc27344fd67 · inbound

STAR: SpatioTemporal Adaptive Reward Allocation for Text-to-Image RL Post-Training cites this paper.

STAR: SpatioTemporal Adaptive Reward Allocation for Text-to-Image RL Post-Training Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T21:18:58.590505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T00:45:39.167521Z digest=sha256:76b826f145fb904819adfb7494270bd8721c38e0b71d72cdda1ba80a8fd0d614

Observation e21b433d-5b7c-49eb-8171-c4530d208d3a · inbound

The Reward Was in Your Data All Along: Correcting Flow Matching with Discriminator-Guided RL cites this paper.

The Reward Was in Your Data All Along: Correcting Flow Matching with Discriminator-Guided RL Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-07-04T00:49:19.404735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-26T20:54:15.234897Z digest=sha256:a05ce082d92828688e6d665800db4ce7ae06b4ded07c0c3fdeafda44722c09ef

Observation 97b70a04-7bbb-4ea5-801e-8e99efe139c7 · inbound

DiT-Reward: Generative Representations for Text-to-Image Reward Modeling cites this paper.

DiT-Reward: Generative Representations for Text-to-Image Reward Modeling Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 109

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T10:39:45.910158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-26T08:35:04.038960Z digest=sha256:b39b70f9c07aecb4e7bd2ed9d7354a5046f28edc64d34e6e432941e599790a6e

Observation 0266f71f-d3d7-492c-80e3-e3390eefdfce · inbound

DiffusionBench: On Holistic Evaluation of Diffusion Transformers cites this paper.

DiffusionBench: On Holistic Evaluation of Diffusion Transformers Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 110

Resolution
verified exact
arxiv_id, observed 2026-07-04T16:59:58.109649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-26T00:06:11.951205Z digest=sha256:6930188b7e4a5d243d92e3f7cb4bf949a33c833c729e26ef7b7bd54c660a2871

Observation 35b7723d-7042-4f03-b467-b357392a4fa6 · inbound

DanceOPD: On-Policy Generative Field Distillation cites this paper.

DanceOPD: On-Policy Generative Field Distillation Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:49:51.351230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T04:55:42.018348Z digest=sha256:e5e2fca1a3bd42f92b02c0c447e3a089b73d4590f7f0044fe8ae8b09496cab51

Observation 5f1822cf-5348-4e7e-96a5-2a72a5d95a16 · inbound

DanceOPD: On-Policy Generative Field Distillation cites this paper.

DanceOPD: On-Policy Generative Field Distillation Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 76

Resolution
unresolved
no resolver link, observed 2026-07-12T11:44:54.717393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:44:54.717393Z digest=sha256:b470a2c1ecb5e473aadce4c233fd5a47a2ca0fe38f7a3704be6acbfde1b9cfa2

Observation bc3481a7-0945-4368-9053-ea524b3a6069 · inbound

FlowAWR: Online Adaptive Flow Reinforcement via Advantage-Weighted Rectification cites this paper.

FlowAWR: Online Adaptive Flow Reinforcement via Advantage-Weighted Rectification Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:14:20.974447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T07:10:44.244790Z digest=sha256:f247d6b8d90980ef90084380dc01fe136112c43de74169197ff0d43bc761f03b

Observation 2173b4b2-15b9-493d-9427-40229aeffded · inbound

WorldCycle: Self-Verifiable Reinforcement Learning for Long-Horizon Video World Models cites this paper.

WorldCycle: Self-Verifiable Reinforcement Learning for Long-Horizon Video World Models Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T12:52:07.433871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:52:07.433871Z digest=sha256:0dc02d387860d9de93385edd47d195b446f8e9ab505b71192cb28c3bfebeb249

Observation 525228fb-75d6-436c-8588-38ab1b9ead87 · inbound

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training cites this paper.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:31.267180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:31.267180Z digest=sha256:4e4ed2b61e2955ada7aca905581cde78f8180978d5558f26e1a5fbca4309c178