Pith. sign in

Paper Citation Record · LEDGER

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance

As of 6 August 2026, this Paper Citation Record lists 62 of 62 outbound references and 3 inbound Pith citation observations for arXiv:2508.21016.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.21016 v1

Coverage vector

measured 62 of 62 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T14:42:22.457966Z

measured 65 of 65 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-30T07:10:44.244790Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T07:14:20.959214Z

Reference resolution

62 of 62 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved61
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation adfc3e4c-ff7a-411f-b5c9-795c3ef9f0f3 · outbound

This paper cites , " * write output.state after.block = add.period write newline.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance , " * write output.state after.block = add.period write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.129723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.129723Z digest=sha256:d7a36f78f79db391a0132002b5b1921c9814025d2b58e8dc4462ce4e6f4dcba3

Observation 7749846b-3fe2-4b42-adb7-26910a877101 · outbound

This paper cites write newline.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.135394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.135394Z digest=sha256:a466f3b810d980b48efff986dacafe80404ff2978b5cdc63bb772623b666f11c

Observation 3458c617-ac39-4146-87a9-0dfe06b458d1 · outbound

This paper cites an unresolved cited work.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-05T14:42:23.668973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T14:42:22.143215Z digest=sha256:61f6f093ddd88954e882cf437132751c1928446aeae216cfcebe67f0236037a5

Observation 6f991acc-92d2-4ab6-87ba-6bb34af8a437 · outbound

This paper cites Training Diffusion Models with Reinforcement Learning.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Training Diffusion Models with Reinforcement Learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.148482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.148482Z digest=sha256:3314ed5b4a4fe45dcf19d80586d469006e0e6b13edf50f5c8211fa16e892158e

Observation ed72c517-5e42-4777-91cf-93adc6dd41f1 · outbound

This paper cites Classifier-Free Guidance is a Predictor-Corrector.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Classifier-Free Guidance is a Predictor-Corrector

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.154328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.154328Z digest=sha256:1e7bd07a24500d2c0fdcbb313cbb53c5ce6ac09bfc3ee91ec6601ad17d39e5e2

Observation 56785e19-4ba3-4acd-8221-afbd7bc3d6ac · outbound

This paper cites PrefPaint: Enhancing Medical Image Inpainting through Expert Human Feedback.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance PrefPaint: Enhancing Medical Image Inpainting through Expert Human Feedback

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-05T14:42:23.298058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T14:42:22.160551Z digest=sha256:2c12599ab83078242ba303a35f96e0d852c37f823716d453007e7fa4229bb02d

Observation 304f57c7-e897-46a5-b7c1-f449696b9cb4 · outbound

This paper cites an unresolved cited work.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.167273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.167273Z digest=sha256:4e9c2c2adaa3598177d22016c2bb807d7dc2c0195ba413e7bd38fe5151403528

Observation 7b737d5e-61b9-4f3b-9691-8ab6edf40637 · outbound

This paper cites Masked-attention Mask Transformer for Universal Image Segmentation.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Masked-attention Mask Transformer for Universal Image Segmentation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.172732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.172732Z digest=sha256:825fd592800c188c96df771b426e809078f4ad12ba152d307882d3142cc67dcf

Observation cddf071c-3a23-402e-b3d8-965bcde8eeaf · outbound

This paper cites an unresolved cited work.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.179641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.179641Z digest=sha256:ebac9370f4597a1157bacfb82ef4503634a9fbefc3a837bcc7328b80519f379a

Observation 4c8f76b2-7df2-467d-ba51-cc68b1d6ada2 · outbound

This paper cites CFG++: Manifold-constrained Classifier Free Guidance for Diffusion Models.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance CFG++: Manifold-constrained Classifier Free Guidance for Diffusion Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.185076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.185076Z digest=sha256:8c10ac2c76f4cea410a3b3822d8965ae8f4e22377456562caa6b64999a9a31a9

Observation 1f3cbfdb-5f44-4ee7-81a6-de6417e866a0 · outbound

This paper cites Directly Fine-Tuning Diffusion Models on Differentiable Rewards.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Directly Fine-Tuning Diffusion Models on Differentiable Rewards

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.194389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.194389Z digest=sha256:ba300884d329d6420f5f549a1d0e1f0b79a454ed2dfa59da322393509dbfa6d9

Observation 5465019d-c0e2-4632-a8c3-1667283e239f · outbound

This paper cites Process Reinforcement through Implicit Rewards.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Process Reinforcement through Implicit Rewards

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.199726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.199726Z digest=sha256:c86a63ee576a2606cc7a49b18e4c3ead3f954d4b5cbc605014a5cae267cc043c

Observation 7c02521e-4250-43a0-93f6-f6e08bd8a148 · outbound

This paper cites an unresolved cited work.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-05T14:42:23.638508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T14:42:22.204773Z digest=sha256:0d9ce4dc2ab088012b622e02135baac556e382a2afe73025800b46fa9956f3c4

Observation c0a4618d-1253-49ef-b432-cdf07356bcb5 · outbound

This paper cites Scaling Rectified Flow Transformers for High-Resolution Image Synthesis.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Scaling Rectified Flow Transformers for High-Resolution Image Synthesis

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.209359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.209359Z digest=sha256:7bc32d7cef76a0abb98d19c933379dbd2abb2af86312e5e20c7b2e9dbf4f5457

Observation 5b73787a-0011-4fb3-8fcf-a88a0e1b9001 · outbound

This paper cites Online Reward-Weighted Fine-Tuning of Flow Matching with Wasserstein Regularization.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Online Reward-Weighted Fine-Tuning of Flow Matching with Wasserstein Regularization

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.214222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.214222Z digest=sha256:68228c950c4bbf1b9681fdfc21930d8212afc7af5f986b792a39f59fa449a9df

Observation 8575a2c6-fc01-49ce-b6a8-64bd81386868 · outbound

This paper cites CFG-Zero*: Improved Classifier-Free Guidance for Flow Matching Models.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance CFG-Zero*: Improved Classifier-Free Guidance for Flow Matching Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.219088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.219088Z digest=sha256:4e72fcbf3bc63bdeb713fd5260af7a81ce1cb852b101bf3e63429cee23b6284f

Observation 56e107ec-6e8c-4c85-a1b2-0b52646da785 · outbound

This paper cites Diffusion Guidance Is a Controllable Policy Improvement Operator.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Diffusion Guidance Is a Controllable Policy Improvement Operator

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.223628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.223628Z digest=sha256:9204367b267bf13be2208e87c0b854196d5d5949ad98e63dcf5cc1166e437d70

Observation 76ed0a85-9501-4c3c-abaf-e9390e43b717 · outbound

This paper cites an unresolved cited work.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-05T14:42:23.620552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T14:42:22.228358Z digest=sha256:8a9ecb686ab1ae39320d5ce6f6d0f89c936d1e3da1e5d515bf837ada9eb35df0

Observation 221f89fb-3a73-4331-9671-19a81ecf03de · outbound

This paper cites DiffuCoder: Understanding and Improving Masked Diffusion Models for Code Generation.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance DiffuCoder: Understanding and Improving Masked Diffusion Models for Code Generation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.236761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.236761Z digest=sha256:e4f692367e049bdfb6f0a411995176b877d718be5241847320cbb0823217afd9

Observation 768c7664-f57e-49fb-87f7-e3afe05f5220 · outbound

This paper cites an unresolved cited work.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.241643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.241643Z digest=sha256:0bfba94e433c8e66aeb16488661dfc11f2dd8f6fda27f05672d38163fe1045d6

Observation b05b307f-f8b5-4cab-87c5-c840d0a02599 · outbound

This paper cites Classifier-Free Diffusion Guidance.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Classifier-Free Diffusion Guidance

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.246069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.246069Z digest=sha256:4b6663bcda585fb08459ac19e00c59791dea23fa89cf0d123dc80a3870ef0900

Observation c97259dd-7afb-4d2e-ba2f-0afb4f919504 · outbound

This paper cites an unresolved cited work.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-05T14:42:23.593229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T14:42:22.250855Z digest=sha256:c1b9971df54b5e55c1a8cfccbe15eeefc775dd866dda0b5628da911ea12e9809

Observation 828132b4-1a46-49e5-a446-68bfdcad2680 · outbound

This paper cites an unresolved cited work.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-05T14:42:23.575591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T14:42:22.255448Z digest=sha256:c79f8321586162b1262d98af9fde6d9deb9fad49f8e04a540528950070730036

Observation 61bcc3c4-df59-47f4-a972-b591474cea3b · outbound

This paper cites Pick-a-Pic: An Open Dataset of User Preferences for Text-to-Image Generation.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Pick-a-Pic: An Open Dataset of User Preferences for Text-to-Image Generation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.259928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.259928Z digest=sha256:f3f159a30a2c5de3fd96f45f836267f66d9cc0c9ef5830d25bca53d089bfc1fb

Observation 2e1a7866-4c6d-4c1a-90a4-0b68584b0c53 · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.264641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.264641Z digest=sha256:ae9436aa7225155d949a0808cbd2d9f9a2eae98da947cec8939e8bc291819514

Observation a1195b3a-2b1b-4f4f-bce8-e7854c0f46bf · outbound

This paper cites Aligning Text-to-Image Models using Human Feedback.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Aligning Text-to-Image Models using Human Feedback

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.270001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.270001Z digest=sha256:da50be55d96dcbe743cc8f71ff6a21d396627d6eae5c7155e29e9a13c9776ac1

Observation 0e181c30-f54a-4486-95be-d5e9ec5dbe92 · outbound

This paper cites BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.275222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.275222Z digest=sha256:b7f25b3d7b1f1e26324559b740fbaec20de4d31aa4060e7df37bfaea8da6a439

Observation 9e3f21ad-9081-4c4f-b264-8e22dd71db05 · outbound

This paper cites Aesthetic Post-Training Diffusion Models from Generic Preferences with Step-by-step Preference Optimization.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Aesthetic Post-Training Diffusion Models from Generic Preferences with Step-by-step Preference Optimization

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.280733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.280733Z digest=sha256:7b96b96200a1c32a5dc63274be24f2987ab97ccabcb1babd2cf09de2c3ec8491

Observation 2057f0dc-e59f-4592-9d2d-d65e6146a1fa · outbound

This paper cites Flow Matching for Generative Modeling.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Flow Matching for Generative Modeling

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.285749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.285749Z digest=sha256:9b50e21f318337cbb6e4780f552995d37cd0797df2ad3a22a8c775484a2b2025

Observation 052041bf-f72c-4588-b694-ec3c62b8be08 · outbound

This paper cites an unresolved cited work.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.291508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.291508Z digest=sha256:a53cdf8ffc9c2374710a010c854841ccfd877292b990ec6b38d7520092e3ada6

Observation 60d87b09-6a4d-4dd1-8425-f695b0ec1a0c · outbound

This paper cites Lumina-Video: Efficient and Flexible Video Generation with Multi-scale Next-DiT.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Lumina-Video: Efficient and Flexible Video Generation with Multi-scale Next-DiT

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.295895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.295895Z digest=sha256:fa2223a53cb4af28c7a4d7abd9c28d1d5a48b4b32ceadf4a09dacd1039543162

Observation 281d1151-1fa6-4a4e-9e2b-8c1f98fa3d35 · outbound

This paper cites Flow-GRPO: Training Flow Matching Models via Online RL.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Flow-GRPO: Training Flow Matching Models via Online RL

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.301547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.301547Z digest=sha256:502ad5de713c8a7dc328ef082ddfed498dd6131193a137e3ac7c6b70cd67be7b

Observation b12ae850-7b16-4c10-805f-31fbfd4f1039 · outbound

This paper cites Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.307315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.307315Z digest=sha256:7cf4d07d0e8b622e21d860fde1d726d3b76b342cb99a088713b9746b028b995d

Observation 9e942faf-4539-4ea9-8ff8-5bc4e1c193c7 · outbound

This paper cites an unresolved cited work.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.312178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.312178Z digest=sha256:e5a0650a494c94c92085a7ef154daf1641c6a851f2475c3b7b512bbafa582bba

Observation 0c62698a-b136-4733-ab0f-6bb80427dfe0 · outbound

This paper cites an unresolved cited work.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-05T14:42:23.548625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T14:42:22.317146Z digest=sha256:a8cd1c4de2439b5e13f803a4c17fa300b249f20ddeda1565e729c28d88d5152b

Observation 74cf2b01-4e6b-4cbe-990c-f68593a31c32 · outbound

This paper cites an unresolved cited work.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-05T14:42:23.531113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T14:42:22.321701Z digest=sha256:82d8ef7b8b1bc33e995dfb540e98f445808db996ba9d2e5834eeb125a471841a

Observation 0e5c2df1-3b3f-40ed-9854-8c2889f00b69 · outbound

This paper cites Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.326127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.326127Z digest=sha256:4ca643f4d258743f95a510f30e7cd3c0641c1301c29cdee7ec70b67bd02a3a18

Observation 4c5bd75d-e8e1-49e2-aa06-5ac443293669 · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.330799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.330799Z digest=sha256:86d4bfbb1e396cf708f97a4bdbbd0b4ffce51b54b7d951c28f158a682c3b70e3

Observation 7cdceb0a-4e7e-4be4-8d08-d5f493990d01 · outbound

This paper cites an unresolved cited work.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-05T14:42:23.514920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T14:42:22.335656Z digest=sha256:8f9360cc89e06c0c42086bbd5d86ee1eb13c645d4400770c1af6dbe5d5aa67da

Observation 0c2b9ab9-deeb-41f9-b4e5-624a56116996 · outbound

This paper cites Learning Transferable Visual Models From Natural Language Supervision.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Learning Transferable Visual Models From Natural Language Supervision

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.340041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.340041Z digest=sha256:5a050602a41c3e94049aa47a52440a0948a86feccdc1065ce2e2f97655a20ee7

Observation 506981d4-0316-4bc6-be19-363c8232b831 · outbound

This paper cites From $r$ to $Q^*$: Your Language Model is Secretly a Q-Function.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance From $r$ to $Q^*$: Your Language Model is Secretly a Q-Function

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.344947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.344947Z digest=sha256:1e42984966f2bac584ac774dfd92ac9c8d30c3f2384774b293ee3cd3f5de929d

Observation 410a884b-1889-4fa0-8759-aa64e1820ae6 · outbound

This paper cites D.; Ermon, S.; and Finn, C.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance D.; Ermon, S.; and Finn, C

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.349971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.349971Z digest=sha256:6fb64d3ea9b6e2a921095f121b051dbab31fda31448152da653e275ea4330507

Observation 7f58168f-ffa2-4e1a-8301-abe13d006b47 · outbound

This paper cites an unresolved cited work.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.354571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.354571Z digest=sha256:d1cb7a4f4159e65ee3b5339e5e73de8763ff64213d4e55669cabbb7f2a4dcb35

Observation b21c10ea-7e63-4108-8bec-b8b4d627db30 · outbound

This paper cites an unresolved cited work.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.359936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.359936Z digest=sha256:7f71ee391792ccd1f42adfe710879edec20a1ca2231f446f5e7eef735e509b1a

Observation 6aaf30ea-0ec4-44a1-9d12-f3031577a7fd · outbound

This paper cites an unresolved cited work.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-05T14:42:23.467023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T14:42:22.364590Z digest=sha256:21aa9c07f56182c97d26862a0c5e96070bad9aa3156613189325264b2b6586af

Observation 2cee3b5a-d185-456f-922c-40af865bfa62 · outbound

This paper cites an unresolved cited work.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-05T14:42:23.450161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T14:42:22.369687Z digest=sha256:39cd27fffdcbecb61c26f9df5bb1aececf416fbb7e9475ee8ee7d7762f0bfa62

Observation a1d1980c-d34a-4fce-90d4-86515ab12e5a · outbound

This paper cites Proximal Policy Optimization Algorithms.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Proximal Policy Optimization Algorithms

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.375167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.375167Z digest=sha256:6737daf651f0826cb9361ed8f0dedf64e3ef14c1941bc934e7d554945aacc156

Observation fe7eb3a0-829a-4352-9bdd-6173339a1843 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.381637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.381637Z digest=sha256:3c806b09afd49f73acc2c6055d43631739c93b902cea506f7b1f4eeea967b32c

Observation e1e67a8e-2637-4a3f-b8ca-b11796ff644a · outbound

This paper cites Feynman-Kac Correctors in Diffusion: Annealing, Guidance, and Product of Experts.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Feynman-Kac Correctors in Diffusion: Annealing, Guidance, and Product of Experts

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.387984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.387984Z digest=sha256:266b71fb7f59d03e904d0e210e8642922e3ee1674ca88e25b37bdd64892974cf

Observation 6863f946-2ed8-4327-879a-be101727b78e · outbound

This paper cites Denoising Diffusion Implicit Models.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Denoising Diffusion Implicit Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.393601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.393601Z digest=sha256:98133bc35e646a5fd5b743046dd277faaf5a394448dbd41ccd3dc9a2a81c3dbe

Observation 6961d1a9-94a6-4f57-8a5c-a70ea13c91c4 · outbound

This paper cites Score-Based Generative Modeling through Stochastic Differential Equations.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Score-Based Generative Modeling through Stochastic Differential Equations

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.399991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.399991Z digest=sha256:577d5b9e0d98fae23437c7771946b342f7484ea0ce7371460df407633729059c

Observation c4e63a3a-d6be-4b11-9a66-f6f111ac4947 · outbound

This paper cites F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.405894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.405894Z digest=sha256:18c5d5413a550c9a72fd510d93b4c45f2612ad9753e4223794f7926bdcb16dfb

Observation 353e51e4-fb10-4da7-9abe-5c4b661d1f2c · outbound

This paper cites Improving and generalizing flow-based generative models with minibatch optimal transport.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Improving and generalizing flow-based generative models with minibatch optimal transport

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.410951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.410951Z digest=sha256:6c58c9db24e728751f5725f367ad13eb37a8714f15328d45a13036a1de02a46a

Observation abe3e218-3c61-4a80-9874-7cd98d65b5ff · outbound

This paper cites What is the Alignment Objective of GRPO?.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance What is the Alignment Objective of GRPO?

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.416055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.416055Z digest=sha256:7be06592b6a9dabd971fd216bf4cdee7782229f9052799a6365562659ec39797

Observation 68f4abfb-8c0b-4976-8c74-286b39d2b6ef · outbound

This paper cites an unresolved cited work.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-05T14:42:23.433202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T14:42:22.421027Z digest=sha256:081dbca868b4419d206ad4f10740afbd199ee80ca461c657bf430e44aaa8f75e

Observation 62bb897e-36e3-40ae-8ca4-72e9363200ce · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Wan: Open and Advanced Large-Scale Video Generative Models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.425577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.425577Z digest=sha256:85c7c757e2d3f3f1ef038e3c2fa303e303ee7d2ac9ad5b91c2cde019c98cbc39

Observation 44622b7b-0a22-457a-b400-959c20f5e6e8 · outbound

This paper cites an unresolved cited work.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-05T14:42:23.417523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T14:42:22.430721Z digest=sha256:2a92fa4498f7ba274120a585e8b07bc6d397724aa02c54df22c41624e03ecc64

Observation cca8fd07-7dfe-4271-b53f-18e569c19dd2 · outbound

This paper cites an unresolved cited work.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-05T14:42:23.399598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T14:42:22.435770Z digest=sha256:613ab5255edf579899abb9ce058a19c8f03a0f1c93e9d8f2c475d8fe3a1d16d6

Observation 8be7027c-db0b-463f-8eef-c1cfa47eaa20 · outbound

This paper cites IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.441703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.441703Z digest=sha256:8cf8f73c933062a334f2c96412b06cd9075e6e18aa954f99aa986471e9292698

Observation 6fc77f98-4575-4218-8e62-c84e9a7b0553 · outbound

This paper cites an unresolved cited work.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-08-05T14:42:23.378995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T14:42:22.447275Z digest=sha256:08c9d2ff53c3486179a0ba2dfe18247347bc94c5d62a075dc665f80c112b0595

Observation 8011e350-6a7b-424d-8960-13d85d075101 · outbound

This paper cites Guided Flows for Generative Modeling and Decision Making.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Guided Flows for Generative Modeling and Decision Making

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:22.452258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:42:22.452258Z digest=sha256:6ec0f4f0c47403e99fd36b7a74d79d4b833801b73cef2435b7830598f9503ba1

Observation 1bc722fa-6767-4acd-b577-a6dde6f7111c · outbound

This paper cites an unresolved cited work.

Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance Unresolved cited work

Reference 62

Resolution
unresolved
raw_fallback, observed 2026-08-05T14:42:23.358216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T14:42:22.457966Z digest=sha256:529415dc31548b86c13cb5bf008ad44bc164b928d041161636be8e382076d337

Pith citing papers

Observation 6674a156-2fbb-43f2-9c5f-0341df19c5b6 · inbound

DiffusionNFT: Online Diffusion Reinforcement with Forward Process cites this paper.

DiffusionNFT: Online Diffusion Reinforcement with Forward Process Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-13T16:54:31.010581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T16:54:30.953199Z digest=sha256:b2191da27115cc4229c8f1f88574ed6d7f13692e3b6b5a875b8f2c47fef58bc0

Observation a4fb9b96-bf1d-4ddb-bb27-01dc1c977e51 · inbound

Towards General Preference Alignment: Diffusion Models at Nash Equilibrium cites this paper.

Towards General Preference Alignment: Diffusion Models at Nash Equilibrium Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:56:06.972632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T16:54:58.732444Z digest=sha256:9f0a17ea9f2dabb31eb9ae70eb5dd785b9a7e4ffa0a05cbdf3986e50a98bfcb0

Observation fb3adc17-f5fd-4baa-a50b-90e33b0a143d · inbound

FlowAWR: Online Adaptive Flow Reinforcement via Advantage-Weighted Rectification cites this paper.

FlowAWR: Online Adaptive Flow Reinforcement via Advantage-Weighted Rectification Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:14:20.961281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-30T07:10:44.244790Z digest=sha256:d86c6757939dd6d674d9fd7b86968ac9781842ccc456c238f259e38f895f2d35