Pith. sign in

Paper Citation Record · LEDGER

RewardDance: Reward Scaling in Visual Generation

As of 7 August 2026, this Paper Citation Record lists 72 of 72 outbound references and 30 inbound Pith citation observations for arXiv:2509.08826.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.08826 v1

Coverage vector

measured 72 of 72 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T20:09:01.822336Z

measured 102 of 102 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 30 of 30 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T21:47:20.112331Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T10:39:45.924030Z

Reference resolution

72 of 72 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved72
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 86905ea5-c022-43bb-8b50-c37819f87f6f · outbound

This paper cites Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models.

RewardDance: Reward Scaling in Visual Generation Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:55.396895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:55.396895Z digest=sha256:7a81bdc193b51623ecbf10a2713129da681017b8ca4de07417d4cd321693af4d

Observation ea855739-cf41-40b2-aa1e-1d10a6592d1f · outbound

This paper cites Improving image generation with better captions.Computer Science.

RewardDance: Reward Scaling in Visual Generation Improving image generation with better captions.Computer Science

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:55.529478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:55.529478Z digest=sha256:7a4f806948e2a29dd926f4a0acb77179b6007ff07554bb753a5ee2f0bd5971f3

Observation 132c6830-d12c-43f7-9156-dddaf4ed4d40 · outbound

This paper cites Training Diffusion Models with Reinforcement Learning.

RewardDance: Reward Scaling in Visual Generation Training Diffusion Models with Reinforcement Learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:55.650307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:55.650307Z digest=sha256:5fb44c1aef617dcb69401fbea954f5d42bbe5766110124852ff96296dd587faa

Observation 7efa567e-ef00-4989-a064-45e848d36bbf · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

RewardDance: Reward Scaling in Visual Generation Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:55.746553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:55.746553Z digest=sha256:90cfc7c120beced077a2a085fc3f6d5bc9b7cb10e7e2184e93909a457fa882d9

Observation 3abdd8b0-0899-4016-a961-6a5bcadc57df · outbound

This paper cites Rank analysis of incomplete block designs: I.

RewardDance: Reward Scaling in Visual Generation Rank analysis of incomplete block designs: I

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:55.865292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:55.865292Z digest=sha256:a0acd8069263fd892b0732331b03c7ad64d544d0223253dcd20310727605942d

Observation c3a37bb2-8a9b-4395-8d20-11ff4747b2f4 · outbound

This paper cites Video generation models as world simulators.OpenAI Blog, 1(8):1, 2024.

RewardDance: Reward Scaling in Visual Generation Video generation models as world simulators.OpenAI Blog, 1(8):1, 2024

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:55.944987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:55.944987Z digest=sha256:3d47e85dd01116689db4501748708f675406b280091faf05495f0c157fea1dfd

Observation 0f4bcd2d-bb19-4f13-b4b8-18379f110cfc · outbound

This paper cites PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis.

RewardDance: Reward Scaling in Visual Generation PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:56.030031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:56.030031Z digest=sha256:c2990a77e407d50a92f3eb78efb96aca03e24dd3cf09086177f4f92c520f7f42

Observation e91579b9-d6e5-4c0c-9e9d-818bfe04fc4c · outbound

This paper cites Control-a-video: Controllable text-to-video generation with diffusion models.CoRR, 2023.

RewardDance: Reward Scaling in Visual Generation Control-a-video: Controllable text-to-video generation with diffusion models.CoRR, 2023

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:56.108706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:56.108706Z digest=sha256:2907e82a3362ab85d176b20c3e60366995777acb923bfb602ff64c9337faa52a

Observation b272c883-8db7-4937-86eb-367ad3b86f30 · outbound

This paper cites The Accuracy Paradox in RLHF: When Better Reward Models Don't Yield Better Language Models.

RewardDance: Reward Scaling in Visual Generation The Accuracy Paradox in RLHF: When Better Reward Models Don't Yield Better Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:56.202728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:56.202728Z digest=sha256:90b08ab65bf96c6bbf498440013073f42594640f7b4e5bf169ba7063d73f2f3d

Observation 486d1163-e00a-4ea2-9d8a-5cf8cd809e93 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

RewardDance: Reward Scaling in Visual Generation Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:56.309210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:56.309210Z digest=sha256:f28d0960aef230a116ff60531976789693f7ca44cc1af7ec3a1671bd1830e517

Observation e2b42bf0-44f1-4691-ad5a-fc91b0429d1a · outbound

This paper cites RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment.

RewardDance: Reward Scaling in Visual Generation RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:56.415726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:56.415726Z digest=sha256:894fad93b24c8eba03e5d795fd6444696246d4ffd0669c20c842f3d2f07d98b0

Observation e8c40eb9-35e7-4a24-aa42-90d44b3ff2ab · outbound

This paper cites Scaling rectified flow transformers for high-resolution image synthesis.

RewardDance: Reward Scaling in Visual Generation Scaling rectified flow transformers for high-resolution image synthesis

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:56.497198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:56.497198Z digest=sha256:4745a98fb448771710c1ab932c33cf1ead77b17dccb03415f106abec41c0ffd8

Observation 1382ada2-b0c5-4221-8d9c-7225d5b6876a · outbound

This paper cites Reinforcement learning for fine-tuning text-to-image diffusion mod- els.

RewardDance: Reward Scaling in Visual Generation Reinforcement learning for fine-tuning text-to-image diffusion mod- els

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:56.572535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:56.572535Z digest=sha256:40ad1ba6233c50ae23935131528484efe0f4b8de24f34ab993070e1656ff94b7

Observation c4bbc6c6-5620-430b-94b2-27eedc40a862 · outbound

This paper cites Seedream 3.0 Technical Report.

RewardDance: Reward Scaling in Visual Generation Seedream 3.0 Technical Report

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:56.678650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:56.678650Z digest=sha256:ce0be35510d129ba3f5e4a4f266b8172b2235b57419cfd22798253117ebdbe13

Observation e1ed9b66-e8ca-4b34-9f27-5433927eb701 · outbound

This paper cites Seedance 1.0: Exploring the Boundaries of Video Generation Models.

RewardDance: Reward Scaling in Visual Generation Seedance 1.0: Exploring the Boundaries of Video Generation Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:56.802057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:56.802057Z digest=sha256:7547f39e401c6b5491daceeb3daac52990fc69d8b39dbb05994a4abf4832e213

Observation da4d94ab-2adb-47a2-8c1f-a6dcf2fc3c85 · outbound

This paper cites Seedream 2.0: A Native Chinese-English Bilingual Image Generation Foundation Model.

RewardDance: Reward Scaling in Visual Generation Seedream 2.0: A Native Chinese-English Bilingual Image Generation Foundation Model

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:56.884737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:56.884737Z digest=sha256:dd1091497f225998e49d385bb9bb668fc9cc71a19a485217940da2e1aab8e2d5

Observation d6e63974-c7c7-459b-a90b-de923f36774b · outbound

This paper cites Veo.https://deepmind.google/models/veo/, 2025.

RewardDance: Reward Scaling in Visual Generation Veo.https://deepmind.google/models/veo/, 2025

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:56.953258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:56.953258Z digest=sha256:e892c2cfc5c6b534b2693158edbd91dcd3f403cf70f7f69dd6d77c2db9cd3777

Observation 301b6b30-f1d6-4a81-8021-787bedd2734f · outbound

This paper cites Multi-Reward as Condition for Instruction-based Image Editing.

RewardDance: Reward Scaling in Visual Generation Multi-Reward as Condition for Instruction-based Image Editing

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:57.041134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:57.041134Z digest=sha256:62a30b9c9d23a53c8b42da8811dcedfc892117e587db569947e040601110c9bf

Observation 9093d8ed-5bfb-4fb1-b9d5-4663fcf14436 · outbound

This paper cites AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning.

RewardDance: Reward Scaling in Visual Generation AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:57.137515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:57.137515Z digest=sha256:539e9ee5b2790dfeb588949843625c75336f40d520f0f82e716f9376731e0a43

Observation 32d7c0f9-754f-47c4-b243-f3d6527613a3 · outbound

This paper cites A simple and effective reinforcement learning method for text-to-image diffusion fine-tuning.

RewardDance: Reward Scaling in Visual Generation A simple and effective reinforcement learning method for text-to-image diffusion fine-tuning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:57.228352Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:57.228352Z digest=sha256:bbb4523be077efb5ed23a0ace4c5f6c871cbcc348f23e6987f2ce6a6fe500735

Observation 31caed26-a32f-4c1d-a531-0594ffda0736 · outbound

This paper cites Denoising diffusion probabilistic models.

RewardDance: Reward Scaling in Visual Generation Denoising diffusion probabilistic models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:57.297593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:57.297593Z digest=sha256:9b40bdd9a9834edb4589db64063e20ee993c04bc495f9729d1b71de8ced6fa48

Observation 79c2c946-32a1-47a5-a3c1-8ee6c77af671 · outbound

This paper cites Ideogram.https://about.ideogram.ai/1.0., 2024.

RewardDance: Reward Scaling in Visual Generation Ideogram.https://about.ideogram.ai/1.0., 2024

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:57.397897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:57.397897Z digest=sha256:6f2bfec9a5c3aee878bb0df094d7305e5d3433cc30f4927dcf4b0053dd9d8410

Observation b88cd1fa-c9b6-41d9-aa71-bf3d89cbaa92 · outbound

This paper cites Pick-a-pic: An open dataset of user preferences for text-to-image generation.Advancesin neural information processing systems, 36:36652–36663, 2023.

RewardDance: Reward Scaling in Visual Generation Pick-a-pic: An open dataset of user preferences for text-to-image generation.Advancesin neural information processing systems, 36:36652–36663, 2023

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:57.483267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:57.483267Z digest=sha256:af7d09652fa4e7022dda56dc50f27b3e67650ddc288efd20134492db0513fc37

Observation 200e654a-6de4-4478-86be-48c758073b93 · outbound

This paper cites klingai.https://app.klingai.com/cn/, 2025.

RewardDance: Reward Scaling in Visual Generation klingai.https://app.klingai.com/cn/, 2025

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:57.585056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:57.585056Z digest=sha256:0d0b660cc1f636f05786d78515711b8bbda50b00f015b425c011154a5f3d0c73

Observation 836d2260-1b27-41f8-82c9-82b7380589d3 · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

RewardDance: Reward Scaling in Visual Generation HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:57.666163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:57.666163Z digest=sha256:5d1513ca27ae19c2a71bdcf2cf61dac7a9f1ae5a8820b80eb0966ba1e7d7e405

Observation af1ebe2a-2734-4c43-a81c-231af56b104d · outbound

This paper cites Flux: Official inference repository for flux.1 models, 2024.

RewardDance: Reward Scaling in Visual Generation Flux: Official inference repository for flux.1 models, 2024

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:57.744104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:57.744104Z digest=sha256:d23c0d94d09c912b853ec758bbea5b9e45c4881b2d8313df38ce24d9f5334ce8

Observation 94a153bb-e4c9-46f1-bbac-6a8c7f108c32 · outbound

This paper cites Flux.https://github.com/black-forest-labs/flux, 2024.

RewardDance: Reward Scaling in Visual Generation Flux.https://github.com/black-forest-labs/flux, 2024

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:57.838549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:57.838549Z digest=sha256:ca8a54ebe106d3280285582b968dc3e03f9e9c6c8382ea6a3d64e3e932809d9d

Observation 8e4c83b1-f2ed-44cd-91a9-df9a6b582506 · outbound

This paper cites Controlnet++: Improving conditional controls with efficient consistency feedback: Project page: liming-ai.

RewardDance: Reward Scaling in Visual Generation Controlnet++: Improving conditional controls with efficient consistency feedback: Project page: liming-ai

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:57.926797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:57.926797Z digest=sha256:ddf85cd68c36544a0ff5e0807dff75b34658fa563301b5d6461e04db93c61d50

Observation fba2b18b-871c-4917-8670-bbc8e802f8ee · outbound

This paper cites SuperEdit: Rectifying and Facilitating Supervision for Instruction-Based Image Editing.

RewardDance: Reward Scaling in Visual Generation SuperEdit: Rectifying and Facilitating Supervision for Instruction-Based Image Editing

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.003125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.003125Z digest=sha256:446357ab4b07b53313cc4b188cb1c0b8e2b61d410d36fb339d214bf0d76a6bb7

Observation 980525a9-64b4-4034-bc9c-1b258ed6ce7a · outbound

This paper cites Exploring How Generative MLLMs Perceive More Than CLIP with the Same Vision Encoder.

RewardDance: Reward Scaling in Visual Generation Exploring How Generative MLLMs Perceive More Than CLIP with the Same Vision Encoder

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.110097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.110097Z digest=sha256:816c97ae856faf59a515cf857e412b1d0df413e6e394ec8e650096a5238ca1de

Observation 40a2ebf7-fa29-4820-92fe-524f2545a508 · outbound

This paper cites An inverse scaling law for clip training.Advancesin Neural Information Processing Systems, 36:49068–49087, 2023.

RewardDance: Reward Scaling in Visual Generation An inverse scaling law for clip training.Advancesin Neural Information Processing Systems, 36:49068–49087, 2023

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.185503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.185503Z digest=sha256:98916df846a17afbe08fe71e0eb579d90c0833d8fe26b308ff736f0d7c0c6b06

Observation 9e56985d-a38c-4a4a-bc8a-387d8247ca8e · outbound

This paper cites Mogao: An Omni Foundation Model for Interleaved Multi-Modal Generation.

RewardDance: Reward Scaling in Visual Generation Mogao: An Omni Foundation Model for Interleaved Multi-Modal Generation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.270191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.270191Z digest=sha256:72af559e74c31f30a1f137d4de13a6125fed45932437417669193b7485897e06

Observation dbb474ed-47dd-4165-879f-85a61565a7b8 · outbound

This paper cites Flow Matching for Generative Modeling.

RewardDance: Reward Scaling in Visual Generation Flow Matching for Generative Modeling

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.373978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.373978Z digest=sha256:fae031d5f8d5f6893a5ed3d98e14476002f665c42d896c083f591a47e962842c

Observation 8fa26720-2d27-44f9-ab9f-4fc28f386f7a · outbound

This paper cites Flow-GRPO: Training Flow Matching Models via Online RL.

RewardDance: Reward Scaling in Visual Generation Flow-GRPO: Training Flow Matching Models via Online RL

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.469893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.469893Z digest=sha256:7540fc73c3ab20ee9f5e855520ad568ee6104621dd1fea6c15ecd12487a5cedb

Observation dcf4dd50-8905-4504-afa9-b61b94c65d46 · outbound

This paper cites Improving Video Generation with Human Feedback.

RewardDance: Reward Scaling in Visual Generation Improving Video Generation with Human Feedback

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.572097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.572097Z digest=sha256:bb1d483c764ec7488520f2c4ae0c39da8f3b8432919ead5061bff851e32adf24

Observation b90777e3-0f54-4c9f-b360-3e758cd57151 · outbound

This paper cites Inference-time scaling for generalist reward modeling.arXiv preprint arXiv:2504.02495, 2025.

RewardDance: Reward Scaling in Visual Generation Inference-time scaling for generalist reward modeling.arXiv preprint arXiv:2504.02495, 2025

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.678868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.678868Z digest=sha256:0eb1872db8adf7cd4e98e7020492cb46e170de41aa749fb2244e1b9fe041bf1e

Observation d26e027a-12f0-4535-b7df-455007611c6f · outbound

This paper cites lumalabs.https://lumalabs.ai/, 2024.

RewardDance: Reward Scaling in Visual Generation lumalabs.https://lumalabs.ai/, 2024

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.744034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.744034Z digest=sha256:b40c3841b37ee952dd1b566ed3c746d1a6ac6a6324af38d343e7ed9376f4e345

Observation e5445f7e-d9e2-48ea-89bf-662cbd1b64c2 · outbound

This paper cites Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model.

RewardDance: Reward Scaling in Visual Generation Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.832391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.832391Z digest=sha256:a7ebaad714443b3f1f668266b91ec2a0614928ece96ab104374753c43ae72e45

Observation 873f0fa2-ce1c-463e-81ea-5695252c4ed2 · outbound

This paper cites Inference-Time Scaling for Diffusion Models beyond Scaling Denoising Steps.

RewardDance: Reward Scaling in Visual Generation Inference-Time Scaling for Diffusion Models beyond Scaling Denoising Steps

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.890930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.890930Z digest=sha256:6c65039856c3f9fb1c2080e3b9f8d899e2a57a6f3ca6c5286fd812aaa3f3a56c

Observation 3930fdb3-dd8b-4d9a-abf8-1030b39d043a · outbound

This paper cites HPSv3: Towards Wide-Spectrum Human Preference Score.

RewardDance: Reward Scaling in Visual Generation HPSv3: Towards Wide-Spectrum Human Preference Score

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:58.995776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:58.995776Z digest=sha256:4750b9c70d25e938e0559fa05cb363f8846b16b56124226d4a479f18c8cc4868

Observation 69192038-94c5-445a-bf15-8916ef183f2f · outbound

This paper cites midjourney.https://www.midjourney.com/home, 2024.

RewardDance: Reward Scaling in Visual Generation midjourney.https://www.midjourney.com/home, 2024

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:59.094853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:59.094853Z digest=sha256:965b58380994d93b193df8071c79eea33a8982e9e1ddcc5b928da1b17d759b6b

Observation d0664f12-d885-48ec-a8de-318928b50deb · outbound

This paper cites Inference-time text-to-video alignment with diffusion latent beam search.arXiv preprint arXiv:2501.19252, 2025.

RewardDance: Reward Scaling in Visual Generation Inference-time text-to-video alignment with diffusion latent beam search.arXiv preprint arXiv:2501.19252, 2025

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:59.185710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:59.185710Z digest=sha256:fafab5246c4613890e9a0c448f6868f9888d103e44f022d1b02ead86386f4a80

Observation 668c5b5a-e40c-4444-9aaf-14c4c7247767 · outbound

This paper cites Training language models to follow instructions with human feedback.

RewardDance: Reward Scaling in Visual Generation Training language models to follow instructions with human feedback

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:59.294363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:59.294363Z digest=sha256:201a50a47c2f6283b503bad0f23242127787376111ad543f9390ac44f5e9120b

Observation 866901a6-8d1e-4f11-9f59-ecec90bfcced · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

RewardDance: Reward Scaling in Visual Generation SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:59.377336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:59.377336Z digest=sha256:f8784be833eebdedaaf8b0b448ac9ca2f35eea722ada4afc7df845bfe868b6ed

Observation 977bd8af-fed2-4656-ba94-2b087858437a · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

RewardDance: Reward Scaling in Visual Generation Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:59.470265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:59.470265Z digest=sha256:9a395bd413d4a53cce6627f592c17f0c290d4017d5298cc5207d3360e4635970

Observation efc721e4-fdb0-4f99-a5ea-fab919a1130e · outbound

This paper cites What makes a reward model a good teacher? an optimization perspective.arXiv preprint arXiv:2503.15477, 2025.

RewardDance: Reward Scaling in Visual Generation What makes a reward model a good teacher? an optimization perspective.arXiv preprint arXiv:2503.15477, 2025

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:59.549832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:59.549832Z digest=sha256:0414fff6c0db4c8b4d1167084125bf6e741ce79efe18a4f11ae08649b2294e47

Observation e3b317e8-9a99-42d8-87b3-6274f9ad808e · outbound

This paper cites recraft.https://www.recraft.ai/, 2024.

RewardDance: Reward Scaling in Visual Generation recraft.https://www.recraft.ai/, 2024

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:59.607627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:59.607627Z digest=sha256:3728511e29ad025472f31b6707e3453c9935db7d3d4b1d9c55109ac4ff78d191

Observation 81fb23ff-d3e2-4f00-99b9-c2f63a0f3a80 · outbound

This paper cites Byteedit: Boost, comply and accelerate generative image editing.

RewardDance: Reward Scaling in Visual Generation Byteedit: Boost, comply and accelerate generative image editing

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:59.703606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:59.703606Z digest=sha256:157503f111b4a18bbd568479022e339466b437af4611f7f8a8da778cb6ae267d

Observation 5da2af02-3f3f-409f-baf2-707e228638a6 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

RewardDance: Reward Scaling in Visual Generation High-resolution image synthesis with latent diffusion models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:59.800075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:59.800075Z digest=sha256:4df4cdfc9d583ba2a47080bed452dfb0de94d123820316720349219faa657c86

Observation f14a6ad6-6c8e-4c6e-bb67-41c035ef8863 · outbound

This paper cites Runway.https://runwayml.com/research/introducing-runway-gen-4, 2025.

RewardDance: Reward Scaling in Visual Generation Runway.https://runwayml.com/research/introducing-runway-gen-4, 2025

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:59.895610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:59.895610Z digest=sha256:72462a903d79a3a8db1e1362b49c80e1dfc4ba5f0c6ae2dfad3228f0f2271f9f

Observation 064fc98f-5828-48a3-ba5a-292bb297fa5d · outbound

This paper cites Seaweed-7B: Cost-Effective Training of Video Generation Foundation Model.

RewardDance: Reward Scaling in Visual Generation Seaweed-7B: Cost-Effective Training of Video Generation Foundation Model

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-04T20:08:59.968870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:08:59.968870Z digest=sha256:d9f2275429da01cdb3c3dbae6639b1f0d19e16f63c65976be47bdd950287a4b2

Observation d208184c-7791-4c8a-a987-fc85114cded9 · outbound

This paper cites Deep unsupervised learning using nonequilibrium thermodynamics.

RewardDance: Reward Scaling in Visual Generation Deep unsupervised learning using nonequilibrium thermodynamics

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:00.052186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:00.052186Z digest=sha256:2e6b858e68c2c4a9e585ae064e98843bcfb7f515d9c09920915958f67ff45138

Observation 83058bb5-07f5-4989-a8b5-0a2f1f9925b6 · outbound

This paper cites Score-Based Generative Modeling through Stochastic Differential Equations.

RewardDance: Reward Scaling in Visual Generation Score-Based Generative Modeling through Stochastic Differential Equations

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:00.112624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:00.112624Z digest=sha256:8951d05c2acf9846938a85e376c57185fa617b231505c539e5c38162577bcc46

Observation a6df593d-ba6d-4eeb-a27b-ee3828f84c4b · outbound

This paper cites Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation.

RewardDance: Reward Scaling in Visual Generation Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:00.245755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:00.245755Z digest=sha256:3414066c4c53f991fffd92ab8794c0dbeef953e28e7a151fd73c53e896c73499

Observation 4b61a4f2-c478-4e5b-bc2e-49d7706edeef · outbound

This paper cites Diffusion model alignment using direct preference optimization.

RewardDance: Reward Scaling in Visual Generation Diffusion model alignment using direct preference optimization

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:00.340485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:00.340485Z digest=sha256:a86db957334f9c88efef5081d43f20d04e957f3d58c27deafa4e218037c43755

Observation 02bedef1-7f2a-4ce3-bb51-59c74b8c2387 · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

RewardDance: Reward Scaling in Visual Generation Wan: Open and Advanced Large-Scale Video Generative Models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:00.409906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:00.409906Z digest=sha256:ef237830f8a62c17368530a0746d774bdc815de1814415ed318632df084587a1

Observation c6887f61-aa82-4a10-9f0e-16a07b78309b · outbound

This paper cites WorldPM: Scaling Human Preference Modeling.

RewardDance: Reward Scaling in Visual Generation WorldPM: Scaling Human Preference Modeling

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:00.504202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:00.504202Z digest=sha256:44414b5190b5c19948bd2746569b761f1423e3c80f7528f6a1d0b6797ca80d50

Observation de72a69d-f92c-4349-9115-e5f9f698a074 · outbound

This paper cites Emu3: Next-Token Prediction is All You Need.

RewardDance: Reward Scaling in Visual Generation Emu3: Next-Token Prediction is All You Need

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:00.588199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:00.588199Z digest=sha256:40a640024681fe1e8f737291d70d80e7594691c9a0b24ec550c8956c012240bc

Observation 6efd3f03-2a1b-4d8e-b84f-c6ee7772ab80 · outbound

This paper cites Unified multimodal chain-of-thought reward model through reinforcement fine-tuning.arXiv preprint arXiv:2505.03318, 2025.

RewardDance: Reward Scaling in Visual Generation Unified multimodal chain-of-thought reward model through reinforcement fine-tuning.arXiv preprint arXiv:2505.03318, 2025

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:00.663870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:00.663870Z digest=sha256:ecc9eb22ee4df40ae3daaab3e3e6a153281af13f61299b4d4debd0bc54e91eee

Observation efae6b3f-4095-4f72-8914-968501479d83 · outbound

This paper cites Rethinking Reward Model Evaluation: Are We Barking up the Wrong Tree?.

RewardDance: Reward Scaling in Visual Generation Rethinking Reward Model Evaluation: Are We Barking up the Wrong Tree?

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:00.786327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:00.786327Z digest=sha256:12ac591d234d772cfdffd3f8d4f1cee6b6d0339d81725a3884125e88e54d078e

Observation eb924bd3-99c4-4e38-a602-3e7f9fda43e7 · outbound

This paper cites Qwen-Image Technical Report.

RewardDance: Reward Scaling in Visual Generation Qwen-Image Technical Report

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:00.883913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:00.883913Z digest=sha256:20db45f79032191c231d4b747b83e6609e94ae38f4bcbc21e2de77349b16d4de

Observation 46e3a5f9-eca2-47eb-a390-548ed9974217 · outbound

This paper cites Human Preference Score: Better Aligning Text-to-Image Models with Human Preference.

RewardDance: Reward Scaling in Visual Generation Human Preference Score: Better Aligning Text-to-Image Models with Human Preference

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:00.976315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:00.976315Z digest=sha256:5c87a6f6fcdd5496d7658d280355e073e2c5200f24d7a3bc5cca55b75957d2a6

Observation 2d92b21d-ec74-43a2-b9b2-9b58ff1aa849 · outbound

This paper cites Human preference score: Better aligning text-to-image models with human preference.

RewardDance: Reward Scaling in Visual Generation Human preference score: Better aligning text-to-image models with human preference

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:01.048080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:01.048080Z digest=sha256:70eb44114732861204c818a81a346befd479a3dcc62009d1f31a866ed4015f00

Observation aa80b625-a1c6-4678-a710-15e30fb59008 · outbound

This paper cites Imagereward: Learning and evaluating human preferences for text-to-image generation.

RewardDance: Reward Scaling in Visual Generation Imagereward: Learning and evaluating human preferences for text-to-image generation

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:01.130666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:01.130666Z digest=sha256:a5ad69353e8660a859a879eda6a6f20054b6688d076d74187d178e5a3e1a349c

Observation 5101a8db-ce9f-4fd1-bb50-799474662b82 · outbound

This paper cites VisionReward: Fine-Grained Multi-Dimensional Human Preference Learning for Image and Video Generation.

RewardDance: Reward Scaling in Visual Generation VisionReward: Fine-Grained Multi-Dimensional Human Preference Learning for Image and Video Generation

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:01.179272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:01.179272Z digest=sha256:cb4a85d02546cfce8ab6be8c44693ef73a4dd539f0b8a72b23160712913bc646

Observation 96ff8a98-aaa1-4426-84c9-7d72a24e8711 · outbound

This paper cites A Unified Pairwise Framework for RLHF: Bridging Generative Reward Modeling and Policy Optimization.

RewardDance: Reward Scaling in Visual Generation A Unified Pairwise Framework for RLHF: Bridging Generative Reward Modeling and Policy Optimization

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:01.240055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:01.240055Z digest=sha256:90e8460da93b0a301bf079c815f2e1ed800c8e96b14d94f2bde3c8e15c2a537e

Observation 04c91260-2b23-467c-90c4-d1a0b87aea58 · outbound

This paper cites DanceGRPO: Unleashing GRPO on Visual Generation.

RewardDance: Reward Scaling in Visual Generation DanceGRPO: Unleashing GRPO on Visual Generation

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:01.340171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:01.340171Z digest=sha256:7cc11b781842d2329e509d4e133d5696d53f79762ced6fb753fa4852626bd63d

Observation b0c408f1-1a71-4588-a01e-4be385e58372 · outbound

This paper cites Schedule on the fly: Diffusion time prediction for faster and better image generation.

RewardDance: Reward Scaling in Visual Generation Schedule on the fly: Diffusion time prediction for faster and better image generation

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:01.438578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:01.438578Z digest=sha256:dc379db5bf784f3fdc172126992fa4cd42e4469f4a6865086d42470cf5a1bc43

Observation a019f125-4279-4470-a980-f33dec0ff0f8 · outbound

This paper cites Make pixels dance: High-dynamic video generation.

RewardDance: Reward Scaling in Visual Generation Make pixels dance: High-dynamic video generation

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:01.521145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:01.521145Z digest=sha256:a6a2c627309fd47123a92d6ebb2b82bc23a72fd55169dde7a6da8bf992c8710e

Observation 14d30f9e-c529-4d55-86ff-ac7fe409c8c4 · outbound

This paper cites Onlinevpo: Align video diffusion model with online video-centric preference optimization.arXiv preprint arXiv:2412.15159, 2024.

RewardDance: Reward Scaling in Visual Generation Onlinevpo: Align video diffusion model with online video-centric preference optimization.arXiv preprint arXiv:2412.15159, 2024

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:01.617649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:01.617649Z digest=sha256:dfa1aa2731860f32bf9d5e7f23adf5d311d9176a14708cc74a820df0bf239c70

Observation 844e159e-4617-4f71-98e5-b22ea819c96f · outbound

This paper cites Unifl: Improve latent diffusion model via unified feedback learning.Advances in Neural Information Processing Systems, 37:67355–67382, 2024.

RewardDance: Reward Scaling in Visual Generation Unifl: Improve latent diffusion model via unified feedback learning.Advances in Neural Information Processing Systems, 37:67355–67382, 2024

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:01.735214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:01.735214Z digest=sha256:d0c9fa0f003c9799743a42138f164c39f5f12a6f9c95c1c0d11b8e055fea0eb5

Observation a1fc5e6d-cd38-4165-878f-3d41f14e96dd · outbound

This paper cites Fine-Tuning Language Models from Human Preferences.

RewardDance: Reward Scaling in Visual Generation Fine-Tuning Language Models from Human Preferences

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:01.822336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:01.822336Z digest=sha256:97ce6a8b0d69a3f0c5f4e6f67de66a68f3b5c0fa11cc880cddf5a65317a012a8

Pith citing papers

Observation aa41ecd4-f187-471c-8c5d-4737894c39c5 · inbound

MixGRPO: Unlocking Flow-based GRPO Efficiency with Mixed ODE-SDE cites this paper.

MixGRPO: Unlocking Flow-based GRPO Efficiency with Mixed ODE-SDE RewardDance: Reward Scaling in Visual Generation

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-13T13:27:50.078729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T13:27:50.031781Z digest=sha256:e350da71cb3d9e7ea245ae70b4f20115e7935b5807840dedd2f2e7a4ccace0f4

Observation 0f9c6109-5bb3-4599-9e82-6fd95e2f7a29 · inbound

Seedream 4.0: Toward Next-generation Multimodal Image Generation cites this paper.

Seedream 4.0: Toward Next-generation Multimodal Image Generation RewardDance: Reward Scaling in Visual Generation

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-12T16:39:00.896443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-12T16:39:00.828442Z digest=sha256:4c31b25941363b690f0748ff85eaaa3abf0c079d3601f28de5c2ec51c3d20848

Observation 88b1eb00-efe6-485c-8474-5262e6456e6c · inbound

Uniworld-V2: Reinforce Image Editing with Diffusion Negative-aware Finetuning and MLLM Implicit Feedback cites this paper.

Uniworld-V2: Reinforce Image Editing with Diffusion Negative-aware Finetuning and MLLM Implicit Feedback RewardDance: Reward Scaling in Visual Generation

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-21T18:01:19.893657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-21T18:01:19.748677Z digest=sha256:81432186d5e78f7c4e8d45f89f2f02ab2d7367541ae01ff749c4da909cf5a937

Observation 5bc90bef-5c59-4c41-accd-9ced908fd7a5 · inbound

Distribution Matching Distillation Meets Reinforcement Learning cites this paper.

Distribution Matching Distillation Meets Reinforcement Learning RewardDance: Reward Scaling in Visual Generation

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-03T21:47:20.112331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:47:20.112331Z digest=sha256:65315072914277c6915417c51adfda7ad83ff125f66eca213c774987f42f5671

Observation 53be49cc-2695-4ae0-a27a-2e225b9edff9 · inbound

Reward Forcing: Efficient Streaming Video Generation with Rewarded Distribution Matching Distillation cites this paper.

Reward Forcing: Efficient Streaming Video Generation with Rewarded Distribution Matching Distillation RewardDance: Reward Scaling in Visual Generation

Reference 77

Resolution
verified exact
arxiv_id, observed 2026-05-16T18:17:55.054498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T18:17:54.943863Z digest=sha256:4c2610c004af40db5b94d66d4b82b4c9af076532fbfac29e6c570c81cc6f15d0

Observation 8e67035b-89db-48f8-9eda-19f252874dad · inbound

Seedance 1.5 pro: A Native Audio-Visual Joint Generation Foundation Model cites this paper.

Seedance 1.5 pro: A Native Audio-Visual Joint Generation Foundation Model RewardDance: Reward Scaling in Visual Generation

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-16T01:35:37.891530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T01:35:37.817082Z digest=sha256:9ce0441521c0fa5a3d9d4d4da7ff9431c60a62e79bf36bc6a9d8cdc295e3d6b6

Observation 38ee40f4-d9f5-422d-b436-c1dd5598246c · inbound

Beyond VLM-Based Rewards: Diffusion-Native Latent Reward Modeling cites this paper.

Beyond VLM-Based Rewards: Diffusion-Native Latent Reward Modeling RewardDance: Reward Scaling in Visual Generation

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-25T06:45:26.073726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-25T06:41:33.493927Z digest=sha256:db9f1cf86228753ef192a4992d6d1ff60ab5d8582b54c1815f6697c9f7e8f170

Observation ee3ef21c-0e70-4519-894f-780fdc8f522d · inbound

Seedance 2.0: Advancing Video Generation for World Complexity cites this paper.

Seedance 2.0: Advancing Video Generation for World Complexity RewardDance: Reward Scaling in Visual Generation

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:35:26.276205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T13:34:36.248186Z digest=sha256:5918362db311a9e4c6fa18c5dd60c795dad06e52de9e741d5c4d2092ff2ac4cd

Observation e7328192-0136-46d3-8da4-955bba962ac0 · inbound

A Systematic Post-Train Framework for Video Generation cites this paper.

A Systematic Post-Train Framework for Video Generation RewardDance: Reward Scaling in Visual Generation

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:26:20.184788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T16:58:33.014401Z digest=sha256:35458f079619c8415f6d8d2b01f560b6171d9bca8571d89a23a2bf48dfd77004

Observation 055871ed-053e-4653-a29e-8ef0bf1af01f · inbound

Leveraging Verifier-Based Reinforcement Learning in Image Editing cites this paper.

Leveraging Verifier-Based Reinforcement Learning in Image Editing RewardDance: Reward Scaling in Visual Generation

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:06:27.520658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T08:00:33.307429Z digest=sha256:c3035295336ad9991155bdc48001585c25b9c3169c3900a7d21a6c0126fb4484

Observation 5eaba5f4-7dea-4564-b3bc-8bb8268a148c · inbound

Leveraging Verifier-Based Reinforcement Learning in Image Editing cites this paper.

Leveraging Verifier-Based Reinforcement Learning in Image Editing RewardDance: Reward Scaling in Visual Generation

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-05-21T09:14:05.973733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-21T09:11:02.183133Z digest=sha256:49e2e9b7402bcf5d60d4209c888f808c37146bf69ae730451902483e8c26061b

Observation b1b1e6f1-ac7f-48aa-bf61-77fe883a442c · inbound

Think, then Score: Decoupled Reasoning and Scoring for Video Reward Modeling cites this paper.

Think, then Score: Decoupled Reasoning and Scoring for Video Reward Modeling RewardDance: Reward Scaling in Visual Generation

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:42:30.786573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T07:37:52.346280Z digest=sha256:954549559caf3392fec873eafc83e22c8eaed462141d2303178f54cf0302a6dc

Observation 18e9cbc2-15da-49eb-b302-ae263df069f0 · inbound

TMPO: Trajectory Matching Policy Optimization for Diverse and Efficient Diffusion Alignment cites this paper.

TMPO: Trajectory Matching Policy Optimization for Diverse and Efficient Diffusion Alignment RewardDance: Reward Scaling in Visual Generation

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:37:29.181018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T07:36:16.811765Z digest=sha256:a28d7d119534009ce97f302bf1f3f683dce81f2068b801047c5ba80b40458aca

Observation 89a90e18-c6af-43be-ab91-cb40c0d2ff94 · inbound

TMPO: Trajectory Matching Policy Optimization for Diverse and Efficient Diffusion Alignment cites this paper.

TMPO: Trajectory Matching Policy Optimization for Diverse and Efficient Diffusion Alignment RewardDance: Reward Scaling in Visual Generation

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-14T22:03:02.863836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T22:01:21.695270Z digest=sha256:1c54a397a38bed37f7de0a085127a9807f1201ad2e2508196bb161d1416fd7a5

Observation a3dbd52d-cab0-45e8-b875-8a9872158d8c · inbound

CaC: Advancing Video Reward Models via Hierarchical Spatiotemporal Concentrating cites this paper.

CaC: Advancing Video Reward Models via Hierarchical Spatiotemporal Concentrating RewardDance: Reward Scaling in Visual Generation

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:02:22.029828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T06:00:31.582714Z digest=sha256:ef968be201f7ef695a400018f974dab1b2a7c9af72089c192aea31a01b7eb9a9

Observation 68a52997-103e-422e-81c8-d98091d48fc1 · inbound

CaC: Advancing Video Reward Models via Hierarchical Spatiotemporal Concentrating cites this paper.

CaC: Advancing Video Reward Models via Hierarchical Spatiotemporal Concentrating RewardDance: Reward Scaling in Visual Generation

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-07-01T13:55:45.300168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-30T22:38:43.102769Z digest=sha256:70d201cadb3e3eae4f77eddce02e8e2f07884395067ad84c38873a24b7d8a716

Observation dadec1d0-7167-44be-bfb7-631db441c849 · inbound

DiffusionOPD: A Unified Perspective of On-Policy Distillation in Diffusion Models cites this paper.

DiffusionOPD: A Unified Perspective of On-Policy Distillation in Diffusion Models RewardDance: Reward Scaling in Visual Generation

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-07-01T14:25:46.853546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-30T21:20:23.420424Z digest=sha256:5a010a8e71bf5d564080741defa3c52dc5f2e75581780b7134b061055c8cf7e6

Observation cfaf8c48-414d-495c-b4fb-a7cb715e6ad4 · inbound

Edit-GRPO: A Locality-Preserving Policy Optimization Framework for Image Editing cites this paper.

Edit-GRPO: A Locality-Preserving Policy Optimization Framework for Image Editing RewardDance: Reward Scaling in Visual Generation

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:47:45.836625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-19T20:44:54.356658Z digest=sha256:28340b4c58e640c11e1c9b01b00f7d7f4be716335a7fa36106cf2ca96fdd4c91

Observation bc19bdb5-f87f-4e3f-8b2b-b1779a76eed8 · inbound

StressDream: Steering Video World Models for Robust Policy Evaluation and Improvement cites this paper.

StressDream: Steering Video World Models for Robust Policy Evaluation and Improvement RewardDance: Reward Scaling in Visual Generation

Reference 118

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:25:59.831028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T22:45:58.629263Z digest=sha256:96b71a8dc5b8a78ce3f214aa6b82d205b6493c58a9d01d79a2cc61a5b5d244a7

Observation 485db53f-e8ef-480a-bb21-91d596b0bdc5 · inbound

Improving Visual Representation Alignment Generation with GRPO cites this paper.

Improving Visual Representation Alignment Generation with GRPO RewardDance: Reward Scaling in Visual Generation

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T19:32:35.129428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T19:24:09.597127Z digest=sha256:62e83094fb40544441fd415652b1089bb39d91ac02e2c04256f0669d18dbe83b

Observation 876aee65-1061-4fbc-b3d5-b22c4f1ad8aa · inbound

Are we really tilting? The mechanics of reward guidance in flow and diffusion models cites this paper.

Are we really tilting? The mechanics of reward guidance in flow and diffusion models RewardDance: Reward Scaling in Visual Generation

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:26:18.153780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T15:20:52.980868Z digest=sha256:f4254abe8487d56607de3a4da788261c093c3f3eca932401af2ab729770e0a18

Observation 1c45d939-a4dd-476c-9e57-1990a9cbc875 · inbound

Z-Reward: Beyond Scalar Rewards by Internalizing Reasoning into Score Distributions cites this paper.

Z-Reward: Beyond Scalar Rewards by Internalizing Reasoning into Score Distributions RewardDance: Reward Scaling in Visual Generation

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:07:27.944548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T17:30:57.001021Z digest=sha256:3091451544e4b76bf1911bc217213695643499a421255282f1dcb0828b694e62

Observation 6e797adc-88a9-454b-89af-7a97da480d64 · inbound

Z-Reward: Beyond Scalar Rewards by Internalizing Reasoning into Score Distributions cites this paper.

Z-Reward: Beyond Scalar Rewards by Internalizing Reasoning into Score Distributions RewardDance: Reward Scaling in Visual Generation

Reference 58

Resolution
unresolved
no resolver link, observed 2026-07-15T10:53:37.186361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T10:53:37.186361Z digest=sha256:4b787135f676cf7f9f51f995d9f08991fed6c2499f2bd3f6955401b64d7c1c10

Observation 494cc52e-0a45-4538-8d3d-0487d70aa0d5 · inbound

Each Judge Its Own Yardstick: Discovering Per-VLM Taxonomies for Physical Video Evaluation cites this paper.

Each Judge Its Own Yardstick: Discovering Per-VLM Taxonomies for Physical Video Evaluation RewardDance: Reward Scaling in Visual Generation

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T10:09:45.413050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-26T09:04:23.965554Z digest=sha256:1bcd52d30e8c8194e6cfc4e7925476cc318a16f2b8ad05dc3efa34ccec72de56

Observation 0b5713ab-4f56-45d7-9d42-b547b47d54e2 · inbound

DiT-Reward: Generative Representations for Text-to-Image Reward Modeling cites this paper.

DiT-Reward: Generative Representations for Text-to-Image Reward Modeling RewardDance: Reward Scaling in Visual Generation

Reference 118

Resolution
verified exact
arxiv_id, observed 2026-07-04T10:39:45.925822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-26T08:35:04.038960Z digest=sha256:6bdb4263aea489ba65eac0b91b9a8934493c384c9c45ecbd78174d95e8cfc96b

Observation 47831f1c-db81-48f8-b061-18d80119d225 · inbound

NormGuard: Reward-Preserving Norm Constraints in Flow-Matching Reinforcement Learning cites this paper.

NormGuard: Reward-Preserving Norm Constraints in Flow-Matching Reinforcement Learning RewardDance: Reward Scaling in Visual Generation

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-06-29T19:03:51.757351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T05:00:04.037802Z digest=sha256:5b078eec491154f4f3e9f345335ee4a791da0983f9a6b7c69d67fb66bf75e197

Observation a16b5dda-0a53-42b1-9613-000a8c0c17a9 · inbound

NormGuard: Reward-Preserving Norm Constraints in Flow-Matching Reinforcement Learning cites this paper.

NormGuard: Reward-Preserving Norm Constraints in Flow-Matching Reinforcement Learning RewardDance: Reward Scaling in Visual Generation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-12T11:38:13.566600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:38:13.566600Z digest=sha256:9075d5c38ca6ffcc78bf6f0cd3aebbc6973a6be9d80265c1eebaf8fc9592a359

Observation c62e0581-9e0c-43ee-b72f-9b3456248803 · inbound

NormGuard: Reward-Preserving Norm Constraints in Flow-Matching Reinforcement Learning cites this paper.

NormGuard: Reward-Preserving Norm Constraints in Flow-Matching Reinforcement Learning RewardDance: Reward Scaling in Visual Generation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T09:56:25.929049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T09:56:25.929049Z digest=sha256:d80e14e5140330385ea90740ebfbd0033f864cd77d6bb22f489a6b71062c6fd0

Observation 99a15030-df7d-4c09-b32d-139e4a24111f · inbound

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation cites this paper.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation RewardDance: Reward Scaling in Visual Generation

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.594851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.594851Z digest=sha256:6640c84b69ed15efd2e7e4d71cf67ccb46f78d054d9abca71eacdcc982e8eea8

Observation 25d55f6b-63ee-43f3-b764-9312d1880f41 · inbound

SciForma: Structure-Faithful Generation of Scientific Diagrams cites this paper.

SciForma: Structure-Faithful Generation of Scientific Diagrams RewardDance: Reward Scaling in Visual Generation

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-01T16:12:23.899000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T16:12:23.899000Z digest=sha256:1c8f6db2488fcd8f301fe0769310bfface2995437ac1a3eea55797f446a5fca8