Pith. sign in

Paper Citation Record · LEDGER

Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 58 inbound Pith citation observations for arXiv:2402.15194.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.15194 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 58 of 58 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T19:48:18.544359Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T06:14:38.555744Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5d8a776e-3d03-4c7d-98f3-7f5eed0ee754 · inbound

Enhancing Exploration with Diffusion Policies in Hybrid Off-Policy RL: Application to Non-Prehensile Manipulation cites this paper.

Enhancing Exploration with Diffusion Policies in Hybrid Off-Policy RL: Application to Non-Prehensile Manipulation Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T14:48:30.702335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:48:30.702335Z digest=sha256:eb95d458cf1212d57ff5c437d7d8d75efc67caf7251ecb4f95123c10e71f35c3

Observation df6887dc-8750-4568-a4cd-1e143c92e2e0 · inbound

Diffusion-VLA: Generalizable and Interpretable Robot Foundation Model via Self-Generated Reasoning cites this paper.

Diffusion-VLA: Generalizable and Interpretable Robot Foundation Model via Self-Generated Reasoning Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-11T22:38:32.747952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T22:38:32.747952Z digest=sha256:319acaaf88f200ed63afeb35e2fbafc8e6aa7bb1674cd0fd0fec9c3390afe490

Observation 2318703c-9446-4614-8758-70adfcf8b413 · inbound

Efficient Diversity-Preserving Diffusion Alignment via Gradient-Informed GFlowNets cites this paper.

Efficient Diversity-Preserving Diffusion Alignment via Gradient-Informed GFlowNets Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-11T18:35:18.610551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T18:35:18.610551Z digest=sha256:4e2d4f9a78726b8f055c2fb0e58c1d93d0793950e554d9009487a038c2d038ae

Observation 68fbedcb-4df2-4a7d-8e21-43ba93e5d1a7 · inbound

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning cites this paper.

Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-09T14:29:12.427168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T14:29:12.427168Z digest=sha256:887338a0ee7cec12f1d5602dad592b3c922199dd435b0ae561832cba003175cd

Observation 057b9782-9fbe-4cd3-91ef-09d2882e0091 · inbound

Direct Distributional Optimization for Provable Alignment of Diffusion Models cites this paper.

Direct Distributional Optimization for Provable Alignment of Diffusion Models Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-09T10:36:54.521232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:36:54.521232Z digest=sha256:a97edd653559411a47d76506d9bd7e5b52ed3d6874d5c3e40a531b48469b5207

Observation 4dc077ae-75a6-43dc-b3e4-6a18e360a941 · inbound

Statistical guarantees for continuous-time policy evaluation: blessing of ellipticity and new tradeoffs cites this paper.

Statistical guarantees for continuous-time policy evaluation: blessing of ellipticity and new tradeoffs Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-08T22:54:25.028899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T22:54:25.028899Z digest=sha256:e8d06114a526eb722d8288bf7c7579f1c805e05956a1b16d972e2378a12a3018

Observation 80ab42bc-d089-4cd8-84b9-351315ac080a · inbound

DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control cites this paper.

DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:48:48.840440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-14T19:48:48.725800Z digest=sha256:d90252e821d9ddd1b7f2212ba84ec9e8f32b762659b8cdb2803f93bdb8a7f0e0

Observation 2cae4956-7c19-4aa6-b38b-731be3b40efc · inbound

Outsourced diffusion sampling: Efficient posterior inference in latent spaces of generative models cites this paper.

Outsourced diffusion sampling: Efficient posterior inference in latent spaces of generative models Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T14:14:14.887804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T14:14:14.887804Z digest=sha256:688128130e22fa876e8ca7937f76ea7f6fb26ca67145ac84a4687b2fd5efc49c

Observation 8b3ec7d4-daf4-45d4-9a70-e76a44120b45 · inbound

A First-order Generative Bilevel Optimization Framework for Diffusion Models cites this paper.

A First-order Generative Bilevel Optimization Framework for Diffusion Models Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T23:43:38.026069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T23:43:38.026069Z digest=sha256:2a755a88df6a917ab5cc56cd8f4ed5392b49c6868f93512f0f00be6e34a7f820

Observation 638dc0d6-b238-4d53-9b57-91a7984e2bad · inbound

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL cites this paper.

VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 1999

Resolution
unresolved
no resolver link, observed 2026-08-07T15:18:55.480756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:18:55.480756Z digest=sha256:6ace3ad6228e1d3ff1dc05d2eb127670540b5176ed311a70c8979e30c6999e98

Observation bbda7e92-67ff-4d49-90e7-b10667e46bbd · inbound

Scaling Image and Video Generation via Test-Time Evolutionary Search cites this paper.

Scaling Image and Video Generation via Test-Time Evolutionary Search Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-07T14:49:48.583581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:49:48.583581Z digest=sha256:0b407abf2b762706392c83e1e51e212470666d5a90cfec189580c1bea2810429

Observation 33c41de6-8225-42ea-b84a-a0d2332fb78f · inbound

ChatVLA-2: Vision-Language-Action Model with Open-World Embodied Reasoning from Pretrained Knowledge cites this paper.

ChatVLA-2: Vision-Language-Action Model with Open-World Embodied Reasoning from Pretrained Knowledge Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T13:26:00.168043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:26:00.168043Z digest=sha256:bd1a16f61260c747b4adbbb39545420bdc995e1382ce6721ee85e57a169837f5

Observation b3aed93b-3141-4cd1-beeb-c588fbe8bd8f · inbound

Provable Maximum Entropy Manifold Exploration via Diffusion Models cites this paper.

Provable Maximum Entropy Manifold Exploration via Diffusion Models Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-15T19:48:18.544359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T19:48:18.544359Z digest=sha256:413014309891aeb38c72aac392c94ebf266f726cb29e414ef5e7bbfe996d660a

Observation 875a60c0-27df-45f7-a311-df60d57c3e25 · inbound

Nabla-R2D3: Effective and Efficient 3D Diffusion Alignment with 2D Rewards cites this paper.

Nabla-R2D3: Effective and Efficient 3D Diffusion Alignment with 2D Rewards Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:17.220344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:17.220344Z digest=sha256:23a1cb04fdef28bc21ca6b2ea56d6727ccf893e6d9f5c40648196ea35d217833

Observation e9e6fb94-0db1-4c0e-ada6-3b11fe78be1d · inbound

Diffusion Tree Sampling: Scalable inference-time alignment of diffusion models cites this paper.

Diffusion Tree Sampling: Scalable inference-time alignment of diffusion models Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T22:59:48.122120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:59:48.122120Z digest=sha256:df9620dabac6d7b866bc4f681732273f67d6b63cab81452238d3a216a55df343

Observation afb509c7-0c56-4222-98ed-3e87ec365fd1 · inbound

Posterior Inference in Latent Space for Scalable Constrained Black-box Optimization cites this paper.

Posterior Inference in Latent Space for Scalable Constrained Black-box Optimization Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-19T07:12:09.632499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-19T07:07:33.125667Z digest=sha256:e8fc84c0a0daedc01a3addcc369803e6220ec7f52d737fa2d4156e3db820e185

Observation 843b3a00-ae0c-41ef-a4af-6e14c93efd31 · inbound

Stein Diffusion Guidance: Training-Free Posterior Correction for Sampling Beyond High-Density Regions cites this paper.

Stein Diffusion Guidance: Training-Free Posterior Correction for Sampling Beyond High-Density Regions Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-22T00:14:27.755889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-22T00:13:51.383775Z digest=sha256:e6db9c6f30b75daac61ee7f15ed651366eda3690361d3d1aec0e5fa3617a1117

Observation cd7c3216-0edc-42c3-83a5-4032b4686ff3 · inbound

Inference-Time Scaling of Diffusion Language Models via Trajectory Refinement cites this paper.

Inference-Time Scaling of Diffusion Language Models via Trajectory Refinement Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-19T05:02:58.781541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-19T05:02:15.241418Z digest=sha256:3124e06b3c3f243a4a574af119d50ebee831556c0d821d5f82e1f434a666b55b

Observation a4f8b944-516f-4f3e-b7f5-34d9629ff514 · inbound

Noise Hypernetworks: Amortizing Test-Time Compute in Diffusion Models cites this paper.

Noise Hypernetworks: Amortizing Test-Time Compute in Diffusion Models Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-05T20:49:49.682338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:49:49.682338Z digest=sha256:606840dd1e670c60bbc99004ac628ba793a624d39ff15694fe86886140cdb48a

Observation f4541949-1eb7-4f74-bc9f-7d70b7a10957 · inbound

Towards Adaptive External Communication in Autonomous Vehicles: A Conceptual Design Framework cites this paper.

Towards Adaptive External Communication in Autonomous Vehicles: A Conceptual Design Framework Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 128

Resolution
unresolved
no resolver link, observed 2026-08-05T19:28:41.145544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:28:41.145544Z digest=sha256:43f49d48558ffc4cdb2583a1e203b16816e1a5c6dd67cbbc7a1f61bac5b48d96

Observation 6ee942e2-c912-4809-a2e2-485bb3e2e01f · inbound

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models cites this paper.

Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-15T16:44:36.803558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:44:36.803558Z digest=sha256:559a9438c5360576f37723d66644cfcc5065bc162b800133595f8b04a17048cd

Observation 0f858b1c-16a5-43c5-9962-38c106f29eaa · inbound

Calibrating Generative Models to Distributional Constraints cites this paper.

Calibrating Generative Models to Distributional Constraints Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-04T10:28:03.501643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T10:28:03.501643Z digest=sha256:c0466345c9fa3baaffd9758658bb5b7fbe92aeb23785fa6079b59e528a0d3b84

Observation bed39288-310a-4503-af62-939ed2137136 · inbound

Test-Time Alignment of Text-to-Image Diffusion Models via Null-Text Embedding Optimisation cites this paper.

Test-Time Alignment of Text-to-Image Diffusion Models via Null-Text Embedding Optimisation Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-03T20:11:27.940051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T20:11:27.940051Z digest=sha256:57eb424695499ca406059ddba6ae59ace0e8a4eac6baeb0429ce96798a39dfae

Observation c6b87cec-709d-4dd4-95be-219f2f5949b7 · inbound

Supervised Guidance Training for Infinite-Dimensional Diffusion Models cites this paper.

Supervised Guidance Training for Infinite-Dimensional Diffusion Models Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-16T10:30:50.423051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-16T10:30:01.994140Z digest=sha256:3d48bd06810aac132d37e5a75788154de96d21207d5c7d24ee829e29978f45a9

Observation 6f855de1-e8c0-4a8d-9efa-319a063bbd5d · inbound

ETS: Energy-Guided Test-Time Scaling for Training-Free RL Alignment cites this paper.

ETS: Energy-Guided Test-Time Scaling for Training-Free RL Alignment Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-16T09:40:49.007286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-16T09:37:57.120779Z digest=sha256:e148fedd05e124bc7e069ad344fd753db308b624b6714ad1803cf17ff988e140

Observation 6173065f-af2e-4d9c-9c8f-e4c4a21d425b · inbound

ETS: Energy-Guided Test-Time Scaling for Training-Free RL Alignment cites this paper.

ETS: Energy-Guided Test-Time Scaling for Training-Free RL Alignment Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-21T14:10:13.556700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T14:05:37.120262Z digest=sha256:f3046ef55dede99dbb1e9502d02570421a8859528e51f3b827bf527ce6692ca3

Observation 016e5c60-aad8-4a68-bbee-3e974a8ae35f · inbound

Conditional Diffusion Guidance under Hard Constraint: A Stochastic Analysis Approach cites this paper.

Conditional Diffusion Guidance under Hard Constraint: A Stochastic Analysis Approach Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-03T04:22:33.175337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T04:22:33.175337Z digest=sha256:bfa077ebca14d2cd2269ea681db55b44f66748fb32a8f89233efb48885af0d33

Observation f5ecd310-b1bf-4b2f-bc16-bf82e0f71f68 · inbound

Calibrated Test-Time Guidance for Bayesian Inference cites this paper.

Calibrated Test-Time Guidance for Bayesian Inference Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-02T20:48:40.728372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:48:40.728372Z digest=sha256:0e6b0f5d94796faef384fb9c91f6bf6020dd188d096f0974b98c32c8c293a949

Observation 401dede5-edd2-4c5f-b6f7-e5e73976230f · inbound

Reinforcement Learning for Diffusion LLMs with Entropy-Guided Step Selection and Stepwise Advantages cites this paper.

Reinforcement Learning for Diffusion LLMs with Entropy-Guided Step Selection and Stepwise Advantages Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-15T12:40:00.236337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-15T12:38:02.339424Z digest=sha256:ca489e798ce58c08358a0314e3b6aa3cd123112d64747f9e747094f8fb7ace81

Observation f7dd87b8-ad4e-45ce-a21a-618e1effffbe · inbound

Time-Reversed BSDEs for Accurate Gradient Estimation in Diffusion Models cites this paper.

Time-Reversed BSDEs for Accurate Gradient Estimation in Diffusion Models Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-13T21:32:09.389485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T21:32:09.389485Z digest=sha256:0f4813702c4d03345411161a664b9164b4534df886b93769b75ccc0ecd3181fc

Observation 6ec589c5-c8e7-4fd9-93c6-546aeb46e674 · inbound

Personalizing Text-to-Image Generation to Individual Taste cites this paper.

Personalizing Text-to-Image Generation to Individual Taste Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 53

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T05:30:56.356945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-10T18:07:31.236729Z digest=sha256:04aea62db14d682ac82574da66c3a26f7723f9822341f30f2a8de7f8d1a9ebf0

Observation 17a67241-cc20-4b4f-aee3-e0bb3a319e37 · inbound

Adjoint Matching through the Lens of the Stochastic Maximum Principle in Optimal Control cites this paper.

Adjoint Matching through the Lens of the Stochastic Maximum Principle in Optimal Control Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-14T22:28:04.677525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-14T22:25:53.037164Z digest=sha256:98b2aa6dbad05e74a95d260a43b70740d1e57b6aec0668bd3d258ecc6ce702ee

Observation 29ce6258-c0b0-41ed-a3b0-7a0a325a56d6 · inbound

Reward Score Matching: Unifying Reward-based Fine-tuning for Flow and Diffusion Models cites this paper.

Reward Score Matching: Unifying Reward-based Fine-tuning for Flow and Diffusion Models Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-10T05:36:01.639930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-10T05:34:01.064008Z digest=sha256:8d0c2e36b3338409870ee9aa00cc71fd9a37a2787f8e6e47eda4388a8b45d877

Observation 09fd1282-ee7d-464f-93ab-ce6e6fee9f8c · inbound

Reward Score Matching: Unifying Reward-based Fine-tuning for Flow and Diffusion Models cites this paper.

Reward Score Matching: Unifying Reward-based Fine-tuning for Flow and Diffusion Models Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 50

Resolution
verified exact
local_arxiv, observed 2026-07-05T17:51:14.859579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-07-05T17:45:49.335925Z digest=sha256:e67a23c85692dad5b4c4f00bdea8594d0e2b1b230aa7f8e072ed49f98fb9e695

Observation 86c8ef6d-c7d3-4a90-a9a2-ab740ae5e5e9 · inbound

Robust mean field control: stochastic maximum principle and variational mean field games cites this paper.

Robust mean field control: stochastic maximum principle and variational mean field games Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 96

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:41:13.084308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-09T21:18:36.137686Z digest=sha256:34f6c17d09ade469a914e47cd2bdbf82774a2def57984de15e5ccc77ff0d6909

Observation 1e3f96af-c386-4267-837f-ad7c9c504fa9 · inbound

How to Guide Your Flow: Few-Step Alignment via Flow Map Reward Guidance cites this paper.

How to Guide Your Flow: Few-Step Alignment via Flow Map Reward Guidance Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T09:46:26.622144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-07T09:24:42.020782Z digest=sha256:a2aa65650ce3ef8794117292b11f0b91173d1b3a60fb8790f835208c84ea640c

Observation 59476dde-0abf-4177-af80-9c8e3b42024c · inbound

How to Guide Your Flow: Few-Step Alignment via Flow Map Reward Guidance cites this paper.

How to Guide Your Flow: Few-Step Alignment via Flow Map Reward Guidance Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T09:29:57.139302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T09:25:19.439423Z digest=sha256:b0d734df5dabd8066e12a5cc70e39cd1d2d3d6cc56c34150396cc1aa32c1631f

Observation 9f6542dc-4681-4799-a8f7-5dca9d66937a · inbound

How to Guide Your Flow: Few-Step Alignment via Flow Map Reward Guidance cites this paper.

How to Guide Your Flow: Few-Step Alignment via Flow Map Reward Guidance Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T08:25:33.209884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-07-01T08:19:37.044714Z digest=sha256:f88356dc4fa029cbea47e0098ebd3d450c7703d5c19898ded01bcf4d5bd8619e

Observation 74f21517-faa7-4c4a-99e4-2cc33ba0fee6 · inbound

A unified perspective on fine-tuning and sampling with diffusion and flow models cites this paper.

A unified perspective on fine-tuning and sampling with diffusion and flow models Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:31:21.417360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-09T19:43:42.331642Z digest=sha256:d0cd6122b6df4facd1893dde7a2d356008883303b2fedb9c0f84b0d9a672af08

Observation 7e306ff6-adde-4f21-9e3e-0b9f8e892c6b · inbound

Improved techniques for fine-tuning flow models via adjoint matching: a deterministic control pipeline cites this paper.

Improved techniques for fine-tuning flow models via adjoint matching: a deterministic control pipeline Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:16:10.708476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-08T09:45:00.759474Z digest=sha256:35b6f8ed4b58ed95d45edc0e19191ef69a72c39e396294352a365b9f0b45d2f4

Observation fbe9da8e-259f-42e3-a10d-9b7938e6660a · inbound

Reinforce Adjoint Matching: Scaling RL Post-Training of Diffusion and Flow-Matching Models cites this paper.

Reinforce Adjoint Matching: Scaling RL Post-Training of Diffusion and Flow-Matching Models Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T06:06:24.201086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-12T04:38:05.640892Z digest=sha256:e60e2900a4a0a105e230e74d50b2c59364b2b0676bb249c1f4dcdec20f922da9

Observation 3579cf01-da84-4905-8f61-d061f622b06c · inbound

Reinforce Adjoint Matching: Scaling RL Post-Training of Diffusion and Flow-Matching Models cites this paper.

Reinforce Adjoint Matching: Scaling RL Post-Training of Diffusion and Flow-Matching Models Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-20T22:29:09.773226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-20T22:24:45.862340Z digest=sha256:70c072ce1bfd95b91e9c7f9c9cc22e7885275bbb47481e4dd9f0a76ce8087eb6

Observation 2f5efe89-f464-4dc5-8bf2-043413cbb7c1 · inbound

Gradient-Free Noise Optimization for Reward Alignment in Generative Models cites this paper.

Gradient-Free Noise Optimization for Reward Alignment in Generative Models Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:47:32.864145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-13T07:46:36.414833Z digest=sha256:9da58358f9e4fc93930dcc8eaede4de895e0c1a072c621979eb930627b1d6c83

Observation 0a46762f-9e6e-480b-88d8-c5a376dfb142 · inbound

Gradient-Free Noise Optimization for Reward Alignment in Generative Models cites this paper.

Gradient-Free Noise Optimization for Reward Alignment in Generative Models Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:38:00.027566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-14T21:37:15.137190Z digest=sha256:b40326b6ab6b92ecc576e54e0ed995a89a1a5ec8d61d04ca91c876b0d346feda

Observation e698f53e-9c4b-4d80-9de8-feb31f5d9a0b · inbound

Hierarchical Variational Policies for Reward-Guided Diffusion cites this paper.

Hierarchical Variational Policies for Reward-Guided Diffusion Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-22T08:51:18.363687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-22T08:47:54.424242Z digest=sha256:3a7c62687dff4b662a178f7d656d41863948acf594a4be8a616ac7f28f4e58ee

Observation 89f458d6-cc83-43e3-985d-e0e1bd574d60 · inbound

Aligning Few-Step Generative Models by Amortizing Sample-based Variational Inference cites this paper.

Aligning Few-Step Generative Models by Amortizing Sample-based Variational Inference Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 94

Resolution
verified exact
arxiv_id, observed 2026-06-29T19:23:54.050604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-29T19:17:17.682142Z digest=sha256:4ce071e926617110e5a069791fcead3788ff90bf056cc855c2e406e9f002fdb6

Observation c7e6fda4-6689-4a6f-8883-3345d3d6ee1a · inbound

Parallel Tempering Initial Sampling in Inference-Time Reward Alignment cites this paper.

Parallel Tempering Initial Sampling in Inference-Time Reward Alignment Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-06-28T23:32:47.417469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T23:28:22.532621Z digest=sha256:00745a93460a1bccc0c4ea0e3f0280067336b1e83bc8f67c2276344dd8e578f9

Observation 9eefbb82-2bb5-407a-8b24-aa3dbf34a86b · inbound

StressDream: Steering Video World Models for Robust Policy Evaluation and Improvement cites this paper.

StressDream: Steering Video World Models for Robust Policy Evaluation and Improvement Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 82

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:25:59.810025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T22:45:58.629263Z digest=sha256:253ea244856bd29917ab3289d519a21fac53adb41cb4d82c6efdd3cee4fe9fa8

Observation 155a235e-3a2a-4849-b568-4d9406f251f0 · inbound

Are we really tilting? The mechanics of reward guidance in flow and diffusion models cites this paper.

Are we really tilting? The mechanics of reward guidance in flow and diffusion models Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:26:18.143061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T15:20:52.980868Z digest=sha256:f72d71e8093b6aa8438dce505ecb069be9167d088a5a4f2d32ed5125a984bb6f

Observation 4fabda44-c3c2-4bf4-936c-27afac9965cd · inbound

Q-VGM: Q-Value-Gradient Matching for Off-Policy Reinforcement Learning of Flow-Matching VLA cites this paper.

Q-VGM: Q-Value-Gradient Matching for Off-Policy Reinforcement Learning of Flow-Matching VLA Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-07-02T21:07:23.765059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-27T19:57:25.010339Z digest=sha256:6ca520aad493b6d8137db69fe08a8412294b945b1e55870a03bc59b7c9b0a689

Observation 0ba5ba20-fce8-46aa-bc38-e03ef02b6d33 · inbound

Q-VGM: Q-Value-Gradient Matching for Off-Policy Reinforcement Learning of Flow-Matching VLA cites this paper.

Q-VGM: Q-Value-Gradient Matching for Off-Policy Reinforcement Learning of Flow-Matching VLA Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T12:13:46.640459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:13:46.640459Z digest=sha256:0e85d354db2673fef569c087095d199c7fa061e43292d469aa46906331d1df61

Observation b667e3bc-eb4c-4598-adde-fde60c3a68ba · inbound

A Markov Chain Approach to Preference Alignment cites this paper.

A Markov Chain Approach to Preference Alignment Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T09:09:42.549353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-26T10:35:25.956748Z digest=sha256:35d8893c89876a3fb5e8368b0e4ec7d6a7530a3a7aea01071171d3f179d3c6e8

Observation 0daedf99-a810-47f8-aab7-dcf39f0531db · inbound

PAPA: Online Personalized Active Preference Alignment cites this paper.

PAPA: Online Personalized Active Preference Alignment Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-02T16:17:08.251214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-07-02T16:16:48.255987Z digest=sha256:f70e893e714661f225893ac99a7531706ec3ab0954f06df9f3860ef54e8098e3

Observation 5315abf4-931b-4707-9087-b41ea4a3e38d · inbound

On the Design Space of Discrete Diffusion Online Adaptation for Molecular Optimization cites this paper.

On the Design Space of Discrete Diffusion Online Adaptation for Molecular Optimization Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-12T06:44:20.438683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:44:20.438683Z digest=sha256:435adfff148a3effe2289be936d0be7cc7d235f9de276b2744d3d3f593b8aa26

Observation 7e89d0cb-d1f6-47be-a6c6-68769884910b · inbound

TILDE: TILt-based Distributional Erasure for Concept Unlearning cites this paper.

TILDE: TILt-based Distributional Erasure for Concept Unlearning Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 38

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T06:14:38.557174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-07-08T06:06:17.561206Z digest=sha256:16474fba2254a3b58245bdf883c0796a4076a8ee6e682b1820e0c79e73adca0e

Observation 3be2a054-2ad4-4cfa-a279-26e45b44ddbc · inbound

VINE: Taming Generative Control Policies for Reinforcement Learning cites this paper.

VINE: Taming Generative Control Policies for Reinforcement Learning Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 57

Resolution
unresolved
no resolver link, observed 2026-07-14T12:17:04.321971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:17:04.321971Z digest=sha256:784b661cc0bed5cf934dd8ddc6bd26357553da6bf97bc9b86ba221877362e339

Observation 149ddf87-0428-43ff-96c4-425d309ca0f2 · inbound

Generalized Fine-Tuning of Diffusion Models via Stochastic Control and FBSDEs cites this paper.

Generalized Fine-Tuning of Diffusion Models via Stochastic Control and FBSDEs Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T09:47:35.240059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T09:47:35.240059Z digest=sha256:457f2f1ff325b9b8e6e151f643c9dab7f40a4a96d9ed9c34e9bcb287a15d1b04

Observation 16c6db0d-ca04-4022-96b2-351604c17d3c · inbound

CrystalGRPO: Target-Aligned and Coverage-Preserving Reinforcement Learning for Flow-Based Crystal Structure Prediction cites this paper.

CrystalGRPO: Target-Aligned and Coverage-Preserving Reinforcement Learning for Flow-Based Crystal Structure Prediction Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T04:23:54.840347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:23:54.840347Z digest=sha256:dbddbd667ccd5eafc84eba61ebee47442aaf360de17c875d308260975f2f66df