Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-13T16:54:30.953199Z
Paper Citation Record · LEDGER
As of 5 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 79 inbound Pith citation observations for arXiv:2509.16117.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-13T16:54:30.953199Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-04T16:29:31.429673Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-04T19:50:11.382174Z
31 of 31 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 139df65c-0428-47f3-8bd0-ccc5a605dcc9 · outbound
DiffusionNFT: Online Diffusion Reinforcement with Forward Process GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 662b90d9-79ac-4858-a45e-eb2cba7e9998 · outbound
DiffusionNFT: Online Diffusion Reinforcement with Forward Process Training Diffusion Models with Reinforcement Learning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 829aedec-ad5e-46e9-96db-d3cd034adb58 · outbound
DiffusionNFT: Online Diffusion Reinforcement with Forward Process Bridging supervised learning and reinforcement learning in math reasoning
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9774b4c3-e85f-47cd-b82d-dc3513ba3fb7 · outbound
DiffusionNFT: Online Diffusion Reinforcement with Forward Process Dpok: Reinforcement learning for fine-tuning text-to-image diffusion models.Advances in Neural Information Processing Systems, 36:79858–79885
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e09793a3-be45-46dd-ab6f-0509b118ac7c · outbound
DiffusionNFT: Online Diffusion Reinforcement with Forward Process Diffusion Guidance Is a Controllable Policy Improvement Operator
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 08f3f541-d2f6-4eca-ab02-ca8d6a3663ac · outbound
DiffusionNFT: Online Diffusion Reinforcement with Forward Process DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0d4ecbe7-5b4f-485d-97d3-e24d947f4914 · outbound
DiffusionNFT: Online Diffusion Reinforcement with Forward Process CLIPScore: A Reference-free Evaluation Metric for Image Captioning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 8ebfa2b7-940f-4a55-a70c-82fa718c2055 · outbound
DiffusionNFT: Online Diffusion Reinforcement with Forward Process Classifier-Free Diffusion Guidance
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6674a156-2fbb-43f2-9c5f-0341df19c5b6 · outbound
DiffusionNFT: Online Diffusion Reinforcement with Forward Process Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 666398f7-49e8-4918-b88e-196b1ddb07eb · outbound
DiffusionNFT: Online Diffusion Reinforcement with Forward Process Aligning Text-to-Image Models using Human Feedback
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d948cab9-2bd9-44fd-805e-896b35ce92aa · outbound
DiffusionNFT: Online Diffusion Reinforcement with Forward Process Reinforcement Learning and Control as Probabilistic Inference: Tutorial and Review
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1bc4cf71-7dea-4249-ad1b-17a6561c1363 · outbound
DiffusionNFT: Online Diffusion Reinforcement with Forward Process arXiv preprint arXiv:2507.07510 , year=
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e257db46-8f45-4a3e-8254-6a728a86e4ee · outbound
DiffusionNFT: Online Diffusion Reinforcement with Forward Process Flow Matching for Generative Modeling
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 954f5b36-d10f-4dea-b58e-32271cd7488b · outbound
DiffusionNFT: Online Diffusion Reinforcement with Forward Process Flow-GRPO: Training Flow Matching Models via Online RL
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 643ff579-85bd-4879-8c41-f5955b319e5e · outbound
DiffusionNFT: Online Diffusion Reinforcement with Forward Process Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5888bd0c-b5d3-4fdd-a454-06e36fae59f5 · outbound
DiffusionNFT: Online Diffusion Reinforcement with Forward Process DPM-Solver++: Fast Solver for Guided Sampling of Diffusion Probabilistic Models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f450e44f-0e3a-4bb5-ab2f-fa1866fa9d2b · outbound
DiffusionNFT: Online Diffusion Reinforcement with Forward Process Video Diffusion Alignment via Reward Gradients
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6bc7b6d4-3bf4-4ee8-9d86-82c3999b5add · outbound
DiffusionNFT: Online Diffusion Reinforcement with Forward Process Proximal Policy Optimization Algorithms
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b20570a1-271b-4f19-870e-3556dd227324 · outbound
DiffusionNFT: Online Diffusion Reinforcement with Forward Process DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a3da8bbf-4ade-40b7-8d08-a6220c04e582 · outbound
DiffusionNFT: Online Diffusion Reinforcement with Forward Process Denoising Diffusion Implicit Models
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation baa0c7fe-655c-47ed-8464-4450d7f27b24 · outbound
DiffusionNFT: Online Diffusion Reinforcement with Forward Process Coefficients-preserving sampling for reinforcement learning with flow matching.arXiv preprint arXiv:2509.05952
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d281ccfb-2a55-450b-82ff-588a5c6dcb03 · outbound
DiffusionNFT: Online Diffusion Reinforcement with Forward Process Unified Reward Model for Multimodal Understanding and Generation
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0e2c99e4-3cf7-4e2b-b9cc-5dfe8da4c49b · outbound
DiffusionNFT: Online Diffusion Reinforcement with Forward Process Human preference score: Better aligning text-to-image models with human preference
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3a903c09-9f27-48b5-996a-8758fa183abe · outbound
DiffusionNFT: Online Diffusion Reinforcement with Forward Process A Minimalist Approach to LLM Reasoning: from Rejection Sampling to Reinforce
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 13844d2c-8f9a-48e6-887f-d444a867f0ce · outbound
DiffusionNFT: Online Diffusion Reinforcement with Forward Process DanceGRPO: Unleashing GRPO on Visual Generation
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a10139b4-1190-4127-9708-6f829f981c1c · outbound
DiffusionNFT: Online Diffusion Reinforcement with Forward Process Fast Sampling of Diffusion Models with Exponential Integrator
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 67b9a9e0-0d82-4e10-b6b5-b1223c43c230 · outbound
DiffusionNFT: Online Diffusion Reinforcement with Forward Process Diffusion Bridge Implicit Models
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e58cc864-4d90-4083-b2cb-eea320d7e78f · outbound
DiffusionNFT: Online Diffusion Reinforcement with Forward Process Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c9c23f10-0754-4c29-a7d9-8a417337ed60 · outbound
DiffusionNFT: Online Diffusion Reinforcement with Forward Process We provide a simpler and more principled perspective based solely on the diffusion model framework
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e6292ee4-2bda-475b-86da-9a4f10183611 · outbound
DiffusionNFT: Online Diffusion Reinforcement with Forward Process B.3 INTUITION BEHIND THEFLOWGRPO OBJECTIVE We provide some insight into reverse-process diffusion RL by inspecting the FlowGRPO objective in a sampler-agnostic manner
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 434c7e89-71c2-4940-a5cb-9bda7f179e8e · outbound
DiffusionNFT: Online Diffusion Reinforcement with Forward Process Consult Doctor
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 041ca52f-8361-42d9-942b-e932235ab501 · inbound
Uniworld-V2: Reinforce Image Editing with Diffusion Negative-aware Finetuning and MLLM Implicit Feedback DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a0b86d4b-6801-473c-b07a-fd9ebc189bb4 · inbound
TAGRPO: Boosting GRPO on Image-to-Video Generation with Direct Trajectory Alignment DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b17e3dad-6831-4e73-a5ec-f797b1cb1685 · inbound
Bridging Information Asymmetry: A Hierarchical Framework for Deterministic Blind Face Restoration DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 738b0bb5-a1b9-4e84-aa95-0291161c62e0 · inbound
Optimizing Few-Step Generation with Adaptive Matching Distillation DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1296a26-47ef-49de-bdef-446b06f1333e · inbound
Improving Text-to-Image Generation with Intrinsic Self-Confidence Rewards DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4d7df37b-3eb7-435e-8da0-b18c71ec7f3d · inbound
LucidNFT: LR-Anchored Multi-Reward Preference Optimization for Flow-Based Real-World Super-Resolution DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b20ae49c-0cbb-4252-8d2d-f331c6cea680 · inbound
GeMPO: Generalized Measure Matching for Online Diffusion Reinforcement Learning DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06b307c4-bef8-4c60-994c-6850f400a37d · inbound
WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 91
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 797b00fe-d66e-42dd-9ddd-f72e7834049d · inbound
CellFluxRL: Biologically-Constrained Virtual Cell Modeling via Reinforcement Learning DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1683dc31-dbb9-4377-af63-bc6096b03661 · inbound
CellFluxRL: Biologically-Constrained Virtual Cell Modeling via Reinforcement Learning DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5d343de5-ace4-4bad-b25a-6ed8d4fe040c · inbound
FP4 Explore, BF16 Train: Diffusion Reinforcement Learning via Efficient Rollout Scaling DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a885f9a6-339f-460c-a16d-1a3b318db0c1 · inbound
FineEdit: Fine-Grained Image Edit with Bounding Box Guidance DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6bb206a3-1cfe-4c7a-8ecc-8f1e286891b0 · inbound
LeapAlign: Post-Training Flow Matching Models at Any Generation Step by Building Two-Step Trajectories DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation fe5a72f4-ea22-481c-8348-217ec93f57ed · inbound
Guiding Distribution Matching Distillation with Gradient-Based Reinforcement Learning DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9df95f4f-2e18-460d-93fa-26605d9311f8 · inbound
Generative Texture Filtering DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d04b458b-a707-475a-b149-8c0c0d9ca9d6 · inbound
Tstars-Tryon 1.0: Robust and Realistic Virtual Try-On for Diverse Fashion Items DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c7811369-8320-48e9-858c-0e8f26f2f982 · inbound
Tstars-Tryon 1.0: Robust and Realistic Virtual Try-On for Diverse Fashion Items DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 25154744-6234-4ffc-a512-b3314fd02279 · inbound
V-GRPO: Online Reinforcement Learning for Denoising Generative Models Is Easier than You Think DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9b4f6aea-76e1-4d66-8d94-02fb5b0ee72f · inbound
A Systematic Post-Train Framework for Video Generation DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6328fc2e-b632-4bea-b445-84cfa2a82106 · inbound
A unified perspective on fine-tuning and sampling with diffusion and flow models DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a76d68d1-a682-483e-927c-bfe222051121 · inbound
Mamoda2.5: Enhancing Unified Multimodal Model with DiT-MoE DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 403688aa-645a-47ac-97c9-e04ada8f1db9 · inbound
JoyAI-Image: Awaking Spatial Intelligence in Unified Multimodal Understanding and Generation DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 106
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation adfcc3d7-e345-44f0-8f06-31ae91d83dd8 · inbound
JoyAI-Image: Awaking Spatial Intelligence in Unified Multimodal Understanding and Generation DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 106
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e7772b12-e13e-44bf-a3ed-5346bcc1ef7d · inbound
Towards General Preference Alignment: Diffusion Models at Nash Equilibrium DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 237d4a6b-cade-40b9-ad0c-045c0646a4e1 · inbound
D-OPSD: On-Policy Self-Distillation for Continuously Tuning Step-Distilled Diffusion Models DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 117
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 134eeadf-ef05-466f-b0bc-05f79685d794 · inbound
D-OPSD: On-Policy Self-Distillation for Continuously Tuning Step-Distilled Diffusion Models DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 118
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ef38f4dc-789a-45cb-9c3b-75f5be85eebb · inbound
D-OPSD: On-Policy Self-Distillation for Continuously Tuning Step-Distilled Diffusion Models DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 118
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 193a290b-47fd-46fc-9e51-1226ac5b7e0c · inbound
Diffusion-APO: Trajectory-Aware Direct Preference Alignment for Video Diffusion Transformers DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b79facb1-3293-48c3-b980-44c21d77a8e0 · inbound
Flow-OPD: On-Policy Distillation for Flow Matching Models DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b9b0d17e-ffef-45fa-b783-993a5c145f0e · inbound
Flow-OPD: On-Policy Distillation for Flow Matching Models DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 92d520c6-469a-46ad-8e5d-c9964c7b58c4 · inbound
Flow-OPD: On-Policy Distillation for Flow Matching Models DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0918f842-0572-42fe-9ce9-d1f377f4cad1 · inbound
Flow-OPD: On-Policy Distillation for Flow Matching Models DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e8693666-5934-4d0a-bca5-8b80b71f7a5f · inbound
Flow-OPD: On-Policy Distillation for Flow Matching Models DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 37c8974f-2203-41d6-8cb1-0f07f3b734fb · inbound
Auto-Rubric as Reward: From Implicit Preferences to Explicit Multimodal Generative Criteria DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 98001f8c-bc21-4200-bfbb-ef193c9cd20c · inbound
RewardHarness: Self-Evolving Agentic Post-Training DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 25775d43-bfe8-4cd4-92d3-744b84d6990b · inbound
Qwen-Image-2.0 Technical Report DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c8826ad7-adfa-4e75-b081-4ba381ddbd2b · inbound
Power Reinforcement Post-Training of Text-to-Image Models with Super-Linear Advantage Shaping DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 98
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 779b8ea3-2730-435f-b0f3-94812f1f4c3f · inbound
When Policy Entropy Constraint Fails: Preserving Diversity in Flow-based RLHF via Perceptual Entropy DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation be15b09b-6310-4b81-bf0f-4d095500b7ce · inbound
Cutting rules in strong field QED with application to trident pair production DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d092edc2-7126-4bbd-8918-0df2b68d53bd · inbound
OmniNFT: Modality-wise Omni Diffusion Reinforcement for Joint Audio-Video Generation DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2c820e11-6079-49ce-b66a-00e5e31302ab · inbound
CreFlow: Corrective Reflow for Sparse-Reward Embodied Video Diffusion RL DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 53d44aaf-8457-4425-ba0c-3aa11d5ee317 · inbound
Unlocking Complex Visual Generation via Closed-Loop Verified Reasoning DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f755778f-cfea-4963-9c97-33f530a7167a · inbound
DiffusionOPD: A Unified Perspective of On-Policy Distillation in Diffusion Models DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 15756214-6d56-4c10-818e-d09873d117cc · inbound
Embedding-perturbed Exploration Preference Optimization for Flow Models DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 98
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation bbcddf5a-7642-4de7-b0c3-e3ab929ce4cd · inbound
Edit-GRPO: A Locality-Preserving Policy Optimization Framework for Image Editing DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6a6ca34e-16e8-42f3-9e41-15081b8997b0 · inbound
Lance: Unified Multimodal Modeling by Multi-Task Synergy DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 149
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b58532f2-2c9d-4342-9543-2a02ef384466 · inbound
Lance: Unified Multimodal Modeling by Multi-Task Synergy DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 150
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3fa1a167-6cd4-4348-a711-981a87f86153 · inbound
Lens: Rethinking Training Efficiency for Foundational Text-to-Image Models DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2f20a9ae-53fb-4663-897b-a00c3fbe8dbe · inbound
Contrastive Distribution Matching for Amortized Sequential Monte Carlo in Discrete Diffusion DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 36b93a51-861b-455c-8fdc-782f8b944f46 · inbound
GeoCycler: Reward-Aligned 3D Diffusion for Constraint-Conditioned Cyclic Peptide Design DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 22610273-e326-4f7a-8e9b-eea88aa558d4 · inbound
Precise: SDE-Consistent Stochastic Sampling for RL Post-Training of Flow-Matching Models DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6d82f758-7150-47f3-ae3b-571ab367f929 · inbound
AdvantageFlow: Advantage-Weighted Least Squares for RL in Flow Models DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2aafe461-0848-4f48-82f0-a23d5cd22fdd · inbound
Aligning Few-Step Generative Models by Amortizing Sample-based Variational Inference DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 112
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c7ec2534-5536-47a9-8c83-af5830c8357e · inbound
OSP-Next: Efficient High-Quality Video Generation with Sparse Sequence Parallelism, HiF8 Quantization, and Reinforcement Learning DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d1565b96-4356-4b95-aa4f-e0618eb1da5c · inbound
Guidance Contrastive Token Credit Assignment for Discrete Policy Optimization DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation cd8f7959-137b-4409-a7b4-1671459555cd · inbound
GDSD: Reinforcement Learning as Guided Denoiser Self-Distillation for Diffusion Language Models DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation dddcc74d-4b9c-436c-a787-bff677d262c8 · inbound
Parallel Tempering Initial Sampling in Inference-Time Reward Alignment DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 8c360d91-4310-4c02-b2ac-6041e8b15518 · inbound
SwanVoice: Expressive Long-Form Zero-Shot Speech Synthesis for Both Monologue and Dialogue DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d1eaa7c2-ba76-4a47-b845-8047652e5a6a · inbound
World Model Self-Distillation: Training World Models to Solve General Tasks DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0b5ccfd4-ed3d-4732-9302-09a5573ec089 · inbound
MaineCoon: Pursuing A Real-Time Audio-Visual Social World Model DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation dc6b5b1a-c076-47a7-9176-820446180ca1 · inbound
SeFi-Image: A Text-to-Image Foundation Model with Semantic-First Diffusion DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation cf856df9-db44-44e4-a34d-91dc2b3b6c25 · inbound
SeFi-Image: A Text-to-Image Foundation Model with Semantic-First Diffusion DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 75e6c51d-8734-482d-bfa7-4c1fef3704f5 · inbound
SeFi-Image: A Text-to-Image Foundation Model with Semantic-First Diffusion DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73e08bf2-5fd5-41f8-891b-a8edcb773476 · inbound
Causal-rCM: A Unified Teacher-Forcing and Self-Forcing Open Recipe for Autoregressive Diffusion Distillation in Streaming Video Generation and Interactive World Models DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 627034c1-f275-4fcf-9cbb-1462961fd612 · inbound
Qwen-Image-2.0-RL Technical Report DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9c42bcab-93e2-463a-91e2-adcfbe5819be · inbound
PerturbCellRL: Verifier-Guided Reinforcement Learning for Single-Cell Perturbation Prediction DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4ed72780-514f-45c5-8e1a-7036b14b7a17 · inbound
NormGuard: Reward-Preserving Norm Constraints in Flow-Matching Reinforcement Learning DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c45a2a27-10ba-4b16-9a45-21cd9585b40c · inbound
NormGuard: Reward-Preserving Norm Constraints in Flow-Matching Reinforcement Learning DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e36b6c4c-fe52-4459-aad9-8e924adf5df2 · inbound
NormGuard: Reward-Preserving Norm Constraints in Flow-Matching Reinforcement Learning DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01d12302-20d4-4289-9cd1-e126510b5729 · inbound
FlowAWR: Online Adaptive Flow Reinforcement via Advantage-Weighted Rectification DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d50855f6-fc2d-4607-a08b-b7488e4f771c · inbound
DetailAnywhere: Fashion Detail Generation via Cross-Modal Feature Alignment Distillation DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 498cfbbd-2378-43ff-8236-bd4292643df9 · inbound
Bridging Interleaved Multi-Modal Reasoning as a Unified Decision Process DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4fea896e-f570-404f-9d6a-fb5b465c2f4e · inbound
ABot-3DWorld 0: A Universal World Model to Explore Any 3D Space DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6ffa1e8-1b9b-43e3-9dcf-5b0da6407561 · inbound
PAVXploreRL: Physical-Action-Visual World Model Reinforcement Learning with Action Exploration DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98ddbe36-96fe-4644-9b66-9659f186cfbb · inbound
JAGG: Jacobian-Aggregated Group Gradient for Efficient GRPO Training of Diffusion Models DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 94
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2480606f-bc55-4b33-9b7b-ecea6f2b8686 · inbound
Mage-Flow: An Efficient Native-Resolution Foundation Model for Image Generation and Editing DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b12dfc64-932d-438d-8a3b-17a47a48d325 · inbound
Oxygen-TryOn: Fashion-Native Foundation Model for Any-item Virtual Try-On DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0010592-1585-4e5e-873e-e5296b7eb444 · inbound
FlowCTS: On-policy Continuous Trajectory Supervision of Flow Models DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d525ebd-deab-45ae-83b9-5f151015bda1 · inbound
SwanTale: Unified Multi-Speaker Speech and Audio Generation for Instruct and Zero-Shot Tasks DiffusionNFT: Online Diffusion Reinforcement with Forward Process
Reference 115
Source-reported events for the cited work
Unavailable: canonical work link unavailable.