Pith. sign in

Paper Citation Record · LEDGER

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models

As of 11 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 2 inbound Pith citation observations for arXiv:2604.18518.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2604.18518 v4

Coverage vector

measured 50 of 50 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-05T11:32:38.636335Z

measured 52 of 52 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T04:57:36.109021Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

50 of 50 outbound references displayed

  • verified exact25
  • verified fuzzy19
  • unresolved2
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch4

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation cb0c44a9-7a71-473d-8e6d-cf8f71918e3e · outbound

This paper cites write newline.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models write newline

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-07-05T11:41:03.244801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:c414b51f77e635e17bb8c25c0fc005559ddb46f4e21a7f694dc35d8106999e57

Observation 9b9d4042-14be-4dea-b1c5-7cdd5635e7ed · outbound

This paper cites D., Ho, J., Tarlow, D., and Van Den Berg, R.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models D., Ho, J., Tarlow, D., and Van Den Berg, R

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-07-05T11:41:03.246692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:c408066fb798159ca30dccfe402351eae2de07ea968d6f195939ef90e59c621c

Observation ab1e2d61-29db-4d9a-95c5-b0c3d3027d31 · outbound

This paper cites Meissonic: Revitalizing masked generative transformers for efficient high-resolution text-to-image synthesis.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models Meissonic: Revitalizing masked generative transformers for efficient high-resolution text-to-image synthesis

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-07-05T11:41:03.249053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:5b1abbffe2292d128584a0291262ec692a9dde9e75e57e8759018f5e0c45b97e

Observation 6882a1a8-5c97-4e5e-a449-bd959d409a9b · outbound

This paper cites Training Diffusion Models with Reinforcement Learning.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models Training Diffusion Models with Reinforcement Learning

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-07-05T11:41:02.703395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:ce66fd953e39c25ab5d8404d1ef0fc65d01d6ece4dfd4afbaeb45b72d51eca98

Observation 22695962-c0ef-48c1-a953-bc06483c924c · outbound

This paper cites Video generation models as world simulators.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models Video generation models as world simulators

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-07-05T11:41:03.287586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:92f820a94df66d42a5a4ad382d4140d62d7d2174d6f61ab718c7c3ff0c7015aa

Observation 70fc3d91-0db1-4a37-90bc-f00a097725d3 · outbound

This paper cites an unresolved cited work.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-07-05T11:41:03.269624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:44e673e02cdb505f44446b0e94c55a8399eab7580b229f6bb14547549ee521df

Observation 7d97494e-1a26-412c-a643-e19525e82873 · outbound

This paper cites Muse: Text-To-Image Generation via Masked Generative Transformers.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models Muse: Text-To-Image Generation via Masked Generative Transformers

Reference 7

Resolution
metadata mismatch
local_arxiv, observed 2026-07-05T11:41:02.733153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:20c27b8f82f4c0f4a852c334a6a496f5b49008b094e34ed7e25582311127c7a5

Observation 0a6cd5dd-dda3-41c9-b163-87feb64903e3 · outbound

This paper cites Emu3.5: Native Multimodal Models are World Learners.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models Emu3.5: Native Multimodal Models are World Learners

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-07-05T11:41:02.738836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:ff1d23d4f8a89a3526545131282337357c35f86d31241a66a6158ac3e928aac9

Observation a776e8f9-68c0-4624-85ac-3f8a6cc338e1 · outbound

This paper cites Prdp: Proximal reward difference prediction for large-scale reward finetuning of diffusion models.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models Prdp: Proximal reward difference prediction for large-scale reward finetuning of diffusion models

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-07-05T11:41:03.263482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:e1d1feccf638fa9c70ecd73f2101e562d6dc318b19581857f9fa172a0add5e8a

Observation 00f966d5-e693-450b-955b-7caf90ae087f · outbound

This paper cites Autoregressive Video Generation without Vector Quantization.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models Autoregressive Video Generation without Vector Quantization

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-07-05T11:41:02.773900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:7e8c86fd584f610428fa09d1dccc091740c2db545bd42a1541aa5c3ec1fee33e

Observation 1f92f6cd-4fbd-4d96-a996-fcb1f73e3004 · outbound

This paper cites Uniform discrete diffusion with metric path for video generation.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models Uniform discrete diffusion with metric path for video generation

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-07-05T11:41:02.744358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:7bed7dd0f41e371463855108db76d48a708819737be15f630cb944a6520e6da8

Observation 6856cada-94e8-476f-a217-c4ab6a4b2050 · outbound

This paper cites Scaling rectified flow transformers for high-resolution image synthesis.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models Scaling rectified flow transformers for high-resolution image synthesis

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-07-05T11:41:03.259628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:ce411587fccee1506f1dd1e6eba079c70e3bebf6ca6621d4b83d77ccb8f09ecb

Observation 9249372a-a7bd-43fd-92bb-d43023588d16 · outbound

This paper cites T., Synnaeve, G., Adi, Y., and Lipman, Y.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models T., Synnaeve, G., Adi, Y., and Lipman, Y

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-07-05T11:41:03.261571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:caca9513c9364a9914ff5bc2679e70f338636b20f8665ed9138a05173055d709

Observation ed94b8f1-1c71-4248-a5c0-f7b4047cb54c · outbound

This paper cites Geneval: An object-focused framework for evaluating text-to-image alignment.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models Geneval: An object-focused framework for evaluating text-to-image alignment

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-07-05T11:41:03.265409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:d95382fd38c0e4c393aaa381fbd560b827936c7921726715f9f12fa2e2694a12

Observation 3a394d73-db75-477a-b3e2-451a84a9010c · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-07-05T11:41:02.709455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:445212f5ce5c793ed8d62fab0e51bdb1634f23c9b842d3f7afd70547ebe5907c

Observation a40fd1f7-f686-412b-8192-792a638a6738 · outbound

This paper cites TempFlow-GRPO: When Timing Matters for GRPO in Flow Models.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models TempFlow-GRPO: When Timing Matters for GRPO in Flow Models

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-07-05T11:41:02.736138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:edd512194d50afed7f8c1756a9e692800e6fc2e8cdfe23141cfc6667fd54559c

Observation a768d9a1-0079-4627-bcc7-22f10a16f5da · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilibrium.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models Gans trained by a two time-scale update rule converge to a local nash equilibrium

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-07-05T11:41:03.255688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:e58411823b17ca342306e849e2a8c7dde1c713dab59dcd404e80558fd46299ef

Observation 9b66c1c5-6a2a-4d4e-a7cb-0180b8814afd · outbound

This paper cites Denoising diffusion probabilistic models.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models Denoising diffusion probabilistic models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-07-05T11:41:03.257701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:b3f28b725fd69ff7399e32c2b06e6dfe9a51748590f597d5d0e20052227ddd49

Observation b32a3501-3f69-4e89-afdf-f150a9d66e55 · outbound

This paper cites CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-07-05T11:41:02.757009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:6108fa57857bc1d499c5630c155f33e3557e0bfb3bec27fa52dbb0fb7d7c81e2

Observation 77e73b82-03e3-4e0f-a6db-978ac209e8ca · outbound

This paper cites Argmax flows and multinomial diffusion: Learning categorical distributions.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models Argmax flows and multinomial diffusion: Learning categorical distributions

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-07-05T11:41:03.267736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:8840cb8fb7b7a1e157d2f0c88c2739b958efd8433f205736f355cde2c93bb7b3

Observation f0995c2f-587e-42f3-bdd3-97f8e854be63 · outbound

This paper cites Democratizing Text-to-Image Masked Generative Models with Compact Text-Aware One-Dimensional Tokens.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models Democratizing Text-to-Image Masked Generative Models with Compact Text-Aware One-Dimensional Tokens

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-07-05T11:41:02.696875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:1c6a092b7d0d9ad1bd5053af45686c98d8935acedd1e492fcf8a2433ad8f1d1d

Observation 911fe3f4-1e57-49cb-9b2b-234fa6b99c77 · outbound

This paper cites Pick-a-pic: An open dataset of user preferences for text-to-image generation.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models Pick-a-pic: An open dataset of user preferences for text-to-image generation

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-07-05T11:41:03.253269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:79fede2019afaf1b31bd4a6b321e06a995e9eab491a00ef828c5d1f3260b5173

Observation 4417b52a-f941-4960-9990-e6d9ab624258 · outbound

This paper cites an unresolved cited work.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-07-05T11:41:03.251085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:4e0ea8b43b8ddbe27ba5d9dda9aa5c5aff3dfbabded9c5df20b1799a2500313e

Observation b5c8e692-e264-402f-8f0a-a85f612b7fe0 · outbound

This paper cites Aligning Text-to-Image Models using Human Feedback.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models Aligning Text-to-Image Models using Human Feedback

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-07-05T11:41:02.691097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:2fda7bbe1baec5634ca7a49119ce9751493be8b96c5f9732d6c9d3effaabdec0

Observation b81c3fca-2bca-4437-979b-38122cd35747 · outbound

This paper cites MixGRPO: Unlocking Flow-based GRPO Efficiency with Mixed ODE-SDE.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models MixGRPO: Unlocking Flow-based GRPO Efficiency with Mixed ODE-SDE

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-07-05T11:41:02.741472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:eba722102fbe44f1430135affb03077039cdeb35d67538d194d0e3844de54c6c

Observation 0775978b-cf29-4203-8b46-d8ef832ab3e1 · outbound

This paper cites Flow Matching for Generative Modeling.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models Flow Matching for Generative Modeling

Reference 26

Resolution
metadata mismatch
local_arxiv, observed 2026-07-05T11:41:02.727340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:dc73a63dc74c4d558365a5d58284682b08edae45b661a7c784e215fbd748f673

Observation 96da687b-17c1-4a7e-bd1a-82ee1751759c · outbound

This paper cites Towards Out-Of-Distribution Generalization: A Survey.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models Towards Out-Of-Distribution Generalization: A Survey

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-07-05T11:41:02.688507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:a2488ca6687198df0f5473e5dc9e459c33305a374de2c59a964bd40f193d9cf4

Observation b640459f-ed0b-4348-bb53-b8f8b7e216bf · outbound

This paper cites Flow-grpo: Training flow matching models via online rl.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models Flow-grpo: Training flow matching models via online rl

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-07-05T11:41:03.283195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:e0cc10d2030a6ddc6612f331675f80dc37e27cba433fef08e559f4cfa46711b1

Observation 9869a29e-bc9e-4b39-8157-0264ecad0691 · outbound

This paper cites and Hutter, F.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models and Hutter, F

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-07-05T11:41:03.281262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:b0abefb8fc5995fe5e862155f71fa7ef907ce00c7838b1cea3b537c21cc49239

Observation b7f09390-328a-49e9-a48b-9fe86ef07b36 · outbound

This paper cites Next-omni: Towards any-to-any omnimodal foundation models with discrete flow matching.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models Next-omni: Towards any-to-any omnimodal foundation models with discrete flow matching

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-07-05T11:41:02.776969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:49fecc36c4a53fef51bac607f9ef83651f1d7821317758493de3a3370a8e37a1

Observation f15ea21e-23da-4c7a-8eb2-9fb0313dfc23 · outbound

This paper cites Reinforcement learning meets masked generative models: Mask- grpo for text-to-image generation.arXiv preprint arXiv:2510.13418.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models Reinforcement learning meets masked generative models: Mask- grpo for text-to-image generation.arXiv preprint arXiv:2510.13418

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-05T11:41:02.751239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:75a8e311f4dcba2321ce671c751264217a794d4eac78ad2470317cd91af7d09e

Observation 05f2e48e-353b-45d8-8ebe-620ec8419267 · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-07-05T11:41:02.768393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:7e39a73dba120632cc64ae9a9fb64993874391fec84122b673cfb2ce7ec85e32

Observation 53e87148-3260-458c-8c54-baa7c96f3e4e · outbound

This paper cites D., Ermon, S., and Finn, C.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models D., Ermon, S., and Finn, C

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-07-05T11:41:03.285164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:26a746b8c079a0d2511380bf21274049561a97ef205d36047ce3b7b400d0c08c

Observation 7d731e9b-1ba6-4632-9bc2-aa257d198dd7 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models High-resolution image synthesis with latent diffusion models

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-07-05T11:41:03.279112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:f0cc4301dc16da4c5c2af3c84aaec905ee44997f54f58b3a56642af16eda39ce

Observation 5eb561f2-ae11-49e1-9af3-8fad8d862a72 · outbound

This paper cites Proximal Policy Optimization Algorithms.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models Proximal Policy Optimization Algorithms

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-07-05T11:41:02.762955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:912091766c35beb038f0d2680b50c98c80cdc50ba7bae7be96735ba67d35e935

Observation be9bdce0-d43f-4a2b-acb1-2251160bf07a · outbound

This paper cites Seedream 4.0: Toward Next-generation Multimodal Image Generation.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models Seedream 4.0: Toward Next-generation Multimodal Image Generation

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-07-05T11:41:02.771268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:aadf01653d76492a980a745c08e40fd157f6e509ea550e3bdc4860c747af4208

Observation 486573f1-78a0-426d-a94f-28536a499113 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-07-05T11:41:02.747263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:07a3c6ac5977dc41929fa8f1316eab4aa19408790ad896c1e4fbf21bd864a3e5

Observation e27124f9-ff59-4d29-892c-4b872ea89dc4 · outbound

This paper cites Flow Matching with General Discrete Paths: A Kinetic-Optimal Perspective.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models Flow Matching with General Discrete Paths: A Kinetic-Optimal Perspective

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-07-05T11:41:02.754194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:59dc0834ac521f7dc481e94831e769584fd06a634358bd097b3cfd06ac3e598a

Observation 3abe7789-258f-47e5-9fb7-f9d6eecb4baa · outbound

This paper cites S., Barto, A.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models S., Barto, A

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-07-05T11:41:03.273719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:3dbb5694022f74596c9d291ab43078e59a2d28676fe406dfe9358d124c12be73

Observation 529d9cef-48ba-4771-be2f-38a21d06a531 · outbound

This paper cites Diffusion model alignment using direct preference optimization.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models Diffusion model alignment using direct preference optimization

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-07-05T11:41:03.271761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:3d83ac6d1df94ec8fab32b633149c425398974ffeb342934fa77843443e22625

Observation 782acd3b-c1c6-47c8-b995-e7a6163bd875 · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models Wan: Open and Advanced Large-Scale Video Generative Models

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-07-05T11:41:02.706343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:f0b85bf3de92f3fa46c6464dfe83ef1d3ff9b38a89fd53ce408c2bf9398e8eae

Observation f7f62af5-c6db-4057-85c9-ea7f551df780 · outbound

This paper cites FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-07-05T11:41:02.730254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:1c647cf4d4336e36f58471d8fe7a0bfa5617dbe9d671607636bca463437f1306

Observation 427688d5-baaa-4134-a99d-d42f8abae46b · outbound

This paper cites SimpleAR: Pushing the Frontier of Autoregressive Visual Generation through Pretraining, SFT, and RL.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models SimpleAR: Pushing the Frontier of Autoregressive Visual Generation through Pretraining, SFT, and RL

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-07-05T11:41:02.765614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:e0c64c0f9060451cdd6a911864a1647b805d560e96b0baff9a654d89651eef2a

Observation 2454d854-f445-4dcb-ac5b-daba494971b3 · outbound

This paper cites Emu3: Next-Token Prediction is All You Need.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models Emu3: Next-Token Prediction is All You Need

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-07-05T11:41:02.693604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:9a98edf9dbbe7feeab64dcc9c526bcc25369161b099226ef7e62319250e3e0fc

Observation 53d089ec-9a4c-47b7-8f73-98b919701ef1 · outbound

This paper cites SANA 1.5: Efficient Scaling of Training-Time and Inference-Time Compute in Linear Diffusion Transformer.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models SANA 1.5: Efficient Scaling of Training-Time and Inference-Time Compute in Linear Diffusion Transformer

Reference 45

Resolution
metadata mismatch
local_arxiv, observed 2026-07-05T11:41:02.724490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:6a6e843e618be756c7ba4164a347815c3a493ffb0d898a3afd5d753b98628b5d

Observation 8309d9e3-b323-4341-952d-e7fba4eed161 · outbound

This paper cites Show-o: One Single Transformer to Unify Multimodal Understanding and Generation.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models Show-o: One Single Transformer to Unify Multimodal Understanding and Generation

Reference 46

Resolution
metadata mismatch
local_arxiv, observed 2026-07-05T11:41:02.759648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:e375c137c9b79d697a502a3960691ac6072d32b41da0f9706731851c09a3b5da

Observation 714ec4cd-7f0a-4bf7-8eab-9095dc0d635a · outbound

This paper cites DanceGRPO: Unleashing GRPO on Visual Generation.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models DanceGRPO: Unleashing GRPO on Visual Generation

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-07-05T11:41:02.721740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:2fab3bacd47a22a622d66a6d9ae3ce59ff74397d93f67846d5d597bff3560311

Observation b3a36a80-9a83-44db-b2a4-c91951241ed6 · outbound

This paper cites GPT-ImgEval: A Comprehensive Benchmark for Diagnosing GPT4o in Image Generation.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models GPT-ImgEval: A Comprehensive Benchmark for Diagnosing GPT4o in Image Generation

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-07-05T11:41:02.717476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:a375b050da9846fe1f43126f40017b9de55c144fa7ced828786c4b1a4977c020

Observation 8d27081b-2c26-4c0d-a0fd-c0ebea6b92b7 · outbound

This paper cites A Dense Reward View on Aligning Text-to-Image Diffusion with Preference.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models A Dense Reward View on Aligning Text-to-Image Diffusion with Preference

Reference 49

Resolution
verified exact
local_arxiv, observed 2026-07-05T11:41:02.699888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:926ffc12f3d0ad59014bb91cf96bc537ea07a2b478547d20ed42e7d266d187ab

Observation 0c04e99a-1fe5-49d2-8239-1c83db1a03e4 · outbound

This paper cites G., Yang, M.-H., Hao, Y., Essa, I., et al.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models G., Yang, M.-H., Hao, Y., Essa, I., et al

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-07-05T11:41:03.275835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:130ac8cf39c457669809689f4045d1b4102625cff2f561fa58ca36d6e157019f

Pith citing papers

Observation 3d15d22c-22fd-45d9-a6d9-068d409238b0 · inbound

VGIF-Score: Interpretable and Diagnostic Evaluation of Spatio-Temporal Instruction Following in Video Generation cites this paper.

VGIF-Score: Interpretable and Diagnostic Evaluation of Spatio-Temporal Instruction Following in Video Generation UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T04:57:36.109021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:57:36.109021Z digest=sha256:7006a5f35c5bc1762903ff29913aeff2d38628f6964cdaf576fc49a05f0f30bf

Observation 50e745f9-0b7f-49ca-8b64-8b9f8b5e197f · inbound

JAGG: Jacobian-Aggregated Group Gradient for Efficient GRPO Training of Diffusion Models cites this paper.

JAGG: Jacobian-Aggregated Group Gradient for Efficient GRPO Training of Diffusion Models UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-01T17:41:02.166858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T17:41:02.166858Z digest=sha256:81ec32ff888a8882cadfb14bec7d463852d0d448786db8aae70170c379ccdd70