Pith. sign in

Paper Citation Record · LEDGER

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards

As of 7 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 3 inbound Pith citation observations for arXiv:2506.07963.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.07963 v3

Coverage vector

measured 66 of 66 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:30:27.240877Z

measured 69 of 69 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T09:54:24.937137Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T13:39:51.542556Z

Reference resolution

66 of 66 outbound references displayed

  • verified exact1
  • verified fuzzy2
  • unresolved63
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 34e31048-d3a1-4dbf-9d6a-5892834d9ee3 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.197768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:26.928846Z digest=sha256:038a05b01863537687afeebd295bdb2d19b4c1b35a0ab27cf25e76743429c1ba

Observation 030bf8b5-b31c-4f67-a578-be7f28a15ce5 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:26.934199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:26.934199Z digest=sha256:6fae306893b3b28a4fe45763bd7b283a52cb4380a7231c19abb106a1f03a7c1e

Observation 476e7ccf-7c3a-451d-a10d-6cbd6e7fccfb · outbound

This paper cites Qwen2.5-VL Technical Report.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Qwen2.5-VL Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:26.939490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:26.939490Z digest=sha256:fdaf21a60b8055a90f6820cf8bcbf66edef5ccc061f6a7d88cfa1d26305a3132

Observation bb092b5f-e680-4bca-ad89-a86b000b4bd7 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:26.944775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:26.944775Z digest=sha256:99d3d804e6e06d4e8298926f71f74f85e03d46c30de2485ee1ae72b748bf6de6

Observation 39545832-955e-442e-8d82-e5ce221c45f4 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.173054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:26.949807Z digest=sha256:5c74529c47f5b125b1878552af083f5ab4ced62f4ee57e994485eb9bdafde6b7

Observation f5695d3b-42fe-4661-9800-339c91ac2dba · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.158281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:26.954721Z digest=sha256:1f0834b153ec97610c3cb35d57143ed909ccdb1dfe3e4280c95f415fb99242bf

Observation f66c0f2e-a84a-4991-b9ed-9f51a6d12c29 · outbound

This paper cites Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:26.960872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:26.960872Z digest=sha256:de79f57ef66f216b38ea971c36298151a9fcae0b247a1057d7777b6a7fb518e1

Observation b9fda4c4-1c7e-457c-9ebd-83b4312d8a5a · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.144043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:26.965722Z digest=sha256:b750ed93201f5d797166067494abd4f2efe9b92d6ca4cd97871570b955b2b2be

Observation 6f8447af-42b3-4a09-9334-d6f56246f97a · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.129436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:26.971313Z digest=sha256:86c5277d467a2d024572f53007ed367d323c27fb83a5c534f3aed4eb7818d08b

Observation 5e2e6c19-fee8-4a54-aaa2-8aa0b24ca236 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.113226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:26.975940Z digest=sha256:ea0cb9558be74bab7f420a1abe5645f4782c2726ef68c33680564e9acd6482d7

Observation 66ffc0c5-a7e5-4aea-9ed5-6e0e1b0cb54a · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.096940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:26.980833Z digest=sha256:917713263cb69a93081538bb06f44ec5084be4f5561768baf69c8cbbdfab358f

Observation c05877b1-550a-4e8b-b27c-de10ba4401f1 · outbound

This paper cites Training-Free Structured Diffusion Guidance for Compositional Text-to-Image Synthesis.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Training-Free Structured Diffusion Guidance for Compositional Text-to-Image Synthesis

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:26.985271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:26.985271Z digest=sha256:639cf3fb688d88f88520523eacd97d86ea785f1883e907c46aa17fca6754d6b4

Observation 049a9ad0-4ba2-479a-a5f7-faee30561381 · outbound

This paper cites SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:26.990200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:26.990200Z digest=sha256:0d713d7cbe9459670b19c4230e0a269796917215d61bc7ac56ebab52b46f25e6

Observation 9549833c-2741-48fd-bcd2-f517171612cf · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.081911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:26.994948Z digest=sha256:01a4bee2cd38a6adb11ca198c51c8e4eaa2baa502b1bbb0fc391584f014ee934

Observation 152d9bfd-1b77-45e7-81d2-df9a777c97be · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.067245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.000145Z digest=sha256:60b98dc6284a0531c350aaab91420bd841ceb87e8f64651a4e590c3aa8cd3cea

Observation b8526ddb-73f2-469d-adfb-ae8cb4489e75 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.052371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.005060Z digest=sha256:99ee956ed563f1d4c229eb5b941838901db924f69d8378df239a5191e2f31de2

Observation 178dfa81-561c-42dd-8a3c-99f3cbd1dca0 · outbound

This paper cites T2I-R1: Reinforcing Image Generation with Collaborative Semantic-level and Token-level CoT.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards T2I-R1: Reinforcing Image Generation with Collaborative Semantic-level and Token-level CoT

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.014921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.014921Z digest=sha256:389223bb61f97e08e7cf577a2d0d258c8a33e581d0512a7c5f78dda94939fdc2

Observation e11f9d04-2253-495d-a522-fc5df98a9761 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.037690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.019805Z digest=sha256:142996aede335c8853d6356c6e47e157504d87be828baf29551f60735fa5b11a

Observation 1916b5cd-270c-4049-9fe7-74bbef8778db · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.024116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.024116Z digest=sha256:53563efa9e7b7b48c9df6d534cd0825c4743baa90aa0161a798d09e179c33ea7

Observation 0e05fe86-6b9a-404b-bac2-c61e4d0707ee · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.028909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.028909Z digest=sha256:fddca82173749900d1485cb50089cbd8a6b320d97cafdc38854f2d5ff04a6a88

Observation 01e38c16-6906-4a5d-964b-345271d45717 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards LLaVA-OneVision: Easy Visual Task Transfer

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.033266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.033266Z digest=sha256:fb1808e28eb12b15152e87c7b9544a1d2e283c25240b40b128a7c76cc289237f

Observation effd4c6d-cb0a-460d-abea-bc7c97d5b6e5 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.003881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.037679Z digest=sha256:bc053e2cfc7e7a3b45f7dfa3d2805433a222959a29923e460dd4322463b3ab1f

Observation b74913d7-a687-49da-bcfd-170f68407a83 · outbound

This paper cites Dual Diffusion for Unified Image Generation and Understanding.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Dual Diffusion for Unified Image Generation and Understanding

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.042171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.042171Z digest=sha256:109ac111eee79ff446718ced8006357c08a1e0ff8e9528bbc25b0c0f8820bfa9

Observation e153f12d-2664-414a-812b-58be4ac2903f · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.047061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.047061Z digest=sha256:2e0326c9ba2e417dfb7546896a22fc7aa27a41acc8d6d045c83eb6b221b5e199

Observation 62d1c41c-1cf1-4d21-a775-9b18b17991ff · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.051526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.051526Z digest=sha256:a538b035723b3e10e13bb05962031079d7d4b464a945c4da475d7e2cfa830b9b

Observation c98c7cb2-aeaa-4896-8cef-475a5211411e · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.056021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.056021Z digest=sha256:0e8963d14ffa2df94950282f4f27c6854b995d2265be94083c3d0da73f671209

Observation 5126711a-d8f7-42b9-a1ab-4f7d04482d9b · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.960574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.060265Z digest=sha256:7584f5fbac282cb9530c5ce54f183349b286d36243647d69ba22bb46e312351e

Observation 6cc3d5e4-6c41-4baa-b85a-b035356585b0 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.945919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.065128Z digest=sha256:ea572f9e08e198034b11780fb98d52647f7e80d0e02efd38ad0cb1f09d35b01c

Observation 83a12574-4bac-483a-94b9-b89428aeec13 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.931317Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.069704Z digest=sha256:226c6facb9c1ac285ddd1ae5918357451fe8a218631ea9c9c90a1e992adfa9ed

Observation 27be12df-04bf-49ef-a6dd-9370d2b9bd23 · outbound

This paper cites DeepSeek-VL: Towards Real-World Vision-Language Understanding.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards DeepSeek-VL: Towards Real-World Vision-Language Understanding

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.074181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.074181Z digest=sha256:6730af878d25452bd82e2a1eb903d8d4a758052e5343514558ffac4e7710efd6

Observation b416e6a3-2ec0-43c8-b9e5-f6ae9822a562 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.916308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.078965Z digest=sha256:b43fbb68677659e3b4dcf53e38d9f8b568e5606c38f39525f27f6e7738ccb9f3

Observation 72e653b7-f105-4d20-ba0d-a85145ae492d · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.083321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.083321Z digest=sha256:1860b112e89b4b92a015d70491485595d6379e67c3198406f137f4f8f36d8f0b

Observation 07f768bd-9e98-4005-90c1-90ed1f1de104 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.893081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.088748Z digest=sha256:5887690deb2ab418f99b8020d8bef6eb915c1fcb8f321887a34542d37b830097

Observation 3dff539c-e43e-4b39-a19c-e752418731de · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.878873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.093709Z digest=sha256:72e514ab9214f4d7fb96881e1cc392ea058edcb8504ae3c8f03a5eaf435ed009

Observation 1fd6c9af-f1c7-4e89-b4c3-9437bf9f51e2 · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.098652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.098652Z digest=sha256:abdec7c0a39c6af50bcd463057cf8862883bd471ed635a41ecf0c6498a64d5d4

Observation a6bdf839-3534-404d-a26d-b0076853d7a6 · outbound

This paper cites SILMM: Self-Improving Large Multimodal Models for Compositional Text-to-Image Generation.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards SILMM: Self-Improving Large Multimodal Models for Compositional Text-to-Image Generation

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:30:27.503745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.103280Z digest=sha256:15d1e784b63a2b9d6bef45a6132d3858f979d3de85aaa1f24418260807cebf87

Observation b677b4c3-48ab-4d87-847f-422ddf6996c3 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.864651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.107881Z digest=sha256:a2623a936d2731780bb1d46c5d7d2830ef9c6d5e950bee42f402b1fcaa9bbe2e

Observation 48cd6510-b10f-4df2-b7f5-8c3075e3e33c · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.112713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.112713Z digest=sha256:307173a20920d11cd2cf56b10a9f61c3b8b43587bacbd2e4e53ca6545602955f

Observation e28f5d81-0f6a-46fb-bf91-c583282c3836 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.117606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.117606Z digest=sha256:4a8102ad858aba24af6153a41da42f1e36d7d54e9445105c6beb1d6da5d3d31f

Observation 5fe7815c-4c28-4f19-9bb1-148063d24c0e · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.831519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.126900Z digest=sha256:3859ce0535c6c5b96651d0fd6c5f61473e13b97d01e179a2138eea86bb41790e

Observation 3628d465-384c-40c4-b3fb-ec01a9a60451 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.131912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.131912Z digest=sha256:1f3d4c16f2de845a6a274ef9fb2b0d4f47449e863327b8b13d070f52c1c14426

Observation d83006e3-c772-498c-a0c0-28521e5da2bc · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.816871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.137164Z digest=sha256:974d095ebbffe8f987ffd703c3b0f50d59745c621099a21b8d5e9bf1ac584495

Observation 4d50a96b-3af2-4a9c-834d-43bea0ecca57 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.801911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.142267Z digest=sha256:05f8e28febf4edfd1fb09054768cc1a1844e7d38cc0db0d75cf0d5f7c5531771

Observation 8903f20c-529f-4f42-a8b1-ed7715ef63b0 · outbound

This paper cites Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.146722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.146722Z digest=sha256:919a53c739b16e510b3e8b2516d6d4203b5d4d1648cde36535f2a0a36ac64b70

Observation d61d2a31-dcf0-4c6a-b518-06fcf2588a83 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.151759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.151759Z digest=sha256:f93cbc647a446549e8a9b17646eca570ab2a44540d35531167383ddb9bca6692

Observation 0e3b3c1a-db97-4757-affb-f25f766bf9bc · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.776772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.156139Z digest=sha256:791e0aecb45ae3236f776b56af145484fa9ca4551d548ec195c012fee60d4d3f

Observation 89edf2a4-b34e-4127-8ac3-2f4475933e0d · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.761861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.160476Z digest=sha256:9cd70a91a1035adaf5d2a9a436ac29da7fad28d59470e63b67a192debaf0a07e

Observation 55b058cf-fc10-4d8d-8c92-7e0ebf2249d5 · outbound

This paper cites ILLUME: Illuminating Your LLMs to See, Draw, and Self-Enhance.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards ILLUME: Illuminating Your LLMs to See, Draw, and Self-Enhance

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.164874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.164874Z digest=sha256:91b35f7c69aa54b6d557f12957b6e2c77b48d3e232a82347d0544a59ddead0f6

Observation 1df0a6a2-8f98-49df-a235-21116094180c · outbound

This paper cites SimpleAR: Pushing the Frontier of Autoregressive Visual Generation through Pretraining, SFT, and RL.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards SimpleAR: Pushing the Frontier of Autoregressive Visual Generation through Pretraining, SFT, and RL

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.169774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.169774Z digest=sha256:6200f6285e87822b6c7533e6d72dfcb59938c5eb561c81624884c6a7a66b08fa

Observation 3250fd9a-198b-4e96-ab3b-6c82dbca1af2 · outbound

This paper cites Emu3: Next-Token Prediction is All You Need.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Emu3: Next-Token Prediction is All You Need

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.174501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.174501Z digest=sha256:7e513d4bb13f3683ac92c844aa723d36f4239b0bd780a9663d32028f44d04747

Observation 8a602248-88e9-4179-af84-35c9e7f88507 · outbound

This paper cites Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.178910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.178910Z digest=sha256:1d952e18871225f7bb4c9bd095075d51aec77f3d2dbbb75ff7ba7f1a3bd33243

Observation ca31300d-1de9-4318-b41e-fbd39d46a64b · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.744102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.183828Z digest=sha256:80c6318b86877dd46389f81bb545776ac358b9a95b42d15e88183ba3e2b86586

Observation 94835a3c-4fb9-4b76-aa5d-61e8fcc7bc79 · outbound

This paper cites VILA-U: a Unified Foundation Model Integrating Visual Understanding and Generation.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards VILA-U: a Unified Foundation Model Integrating Visual Understanding and Generation

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.188204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.188204Z digest=sha256:3dc51c1616cca59c9c7b244d823a632397e5391ba4b5df8b7331350c141284f5

Observation 3b852fb8-cf91-4dbd-8c73-f0113944ef29 · outbound

This paper cites DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.192673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.192673Z digest=sha256:9b52893b2f835b2115c61fba2e911b35e9b9349cf2498adfeaccbddaf719fb62

Observation f81772fb-3b25-4026-8d2a-e621ad33bdac · outbound

This paper cites Show-o: One Single Transformer to Unify Multimodal Understanding and Generation.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Show-o: One Single Transformer to Unify Multimodal Understanding and Generation

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.197315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.197315Z digest=sha256:afa68e862e12ae098909354e34361e33df2cccdc931b682de595c61903ee0675

Observation 3ba61ad1-768a-456e-a447-82f18c302af3 · outbound

This paper cites X-VILA: Cross-Modality Alignment for Large Language Model.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards X-VILA: Cross-Modality Alignment for Large Language Model

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.201831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.201831Z digest=sha256:49567ac9eacade9bd730062dad6c412a595a3baf0982f639ca90fad5d671db1f

Observation 4124d818-2734-43a4-84d0-a5b0249efae7 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.729510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.207074Z digest=sha256:1f0ad7e6f3e15e8ca399ce87653aa24b88ef8a3896d3801d0d401756ca18c753

Observation 63e370f3-1875-452c-a273-bd578bf6d1e1 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.715018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.215968Z digest=sha256:d6dcd5d232f32db085f682b142b2d26da85a3f264ef50f849a99b06f4f4c37b7

Observation 2bf5185c-f9e2-4da1-8942-3f2e49ad5ced · outbound

This paper cites Ma, Simon Stepputtis, Deva Ra- manan, Russ Salakhutdinov, Louis-Philippe Morency, Katia P.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Ma, Simon Stepputtis, Deva Ra- manan, Russ Salakhutdinov, Louis-Philippe Morency, Katia P

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:30:27.699828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.220358Z digest=sha256:0d29df178cc5a461dba45d2c2bd0fe9075aa6ebd6deecdbb8bbd1f125b3b927a

Observation 0b9288bf-20e8-44f4-9836-6d28f7abcf2e · outbound

This paper cites d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.225291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.225291Z digest=sha256:b109a3ae755d08a74e9dc3ab5b50ca81ee3dd152509f7358635f88ea666a3e89

Observation ff90e7cc-102a-4ecf-8603-1ea2061cff47 · outbound

This paper cites Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.230749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.230749Z digest=sha256:1191db7a7a52e6dd640513f3ccead216b5f84639ab062f7f404aca1d5026ff0d

Observation 4b54aed6-a1ab-4e57-8c64-3aa90cfc3690 · outbound

This paper cites Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.235770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.235770Z digest=sha256:27585634158d90ec0385ece5c49e6229dee470de622f4ce7022438198ebf15ec

Observation cf990229-2d9f-4505-b12e-178a493f4849 · outbound

This paper cites M" and an.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards M" and an

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:30:27.684850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.240877Z digest=sha256:706454765caa7368caf3d3f8a988aad473c145b7e30f1bc5e25d4c316cc8aa61

Observation 8d0e7ef2-0718-41e0-8959-14284f5f9385 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.122335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.122335Z digest=sha256:fb3a665d98117f668484c40a28144028bb04c1dd83274ba81046ae85a2a3ad75

Observation 41965ca6-c991-4de9-88d5-55ec184c603d · outbound

This paper cites Align Anything: Training All-Modality Models to Follow Instructions with Language Feedback.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Align Anything: Training All-Modality Models to Follow Instructions with Language Feedback

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.009844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.009844Z digest=sha256:6696ec8c53b253c5579e1aa541d45502474c0d3c538dff94a122ce96168f34ee

Observation 1605fb19-019b-4012-bcce-2f1da7f47d0e · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.211480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.211480Z digest=sha256:752b048984d45d8fe8b4f99e445a11ca51dabd6a920cc66812be9bfa02413fcb

Pith citing papers

Observation 2fe94ad2-fdfa-4a8c-9c70-8019974e7869 · inbound

A Survey of Reinforcement Learning for Large Reasoning Models cites this paper.

A Survey of Reinforcement Learning for Large Reasoning Models SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards

Reference 197

Resolution
verified exact
arxiv_id, observed 2026-05-18T00:05:31.604172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-18T00:02:24.352947Z digest=sha256:1a85a66cad0a61726de6fa57554ca8a36a6d20b1b55a2b47a394e7765073fde3

Observation 7b72bc66-ff7d-4a96-9026-21567ceda278 · inbound

SRUM: Fine-Grained Self-Rewarding for Unified Multimodal Models cites this paper.

SRUM: Fine-Grained Self-Rewarding for Unified Multimodal Models SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-04T09:54:24.937137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:54:24.937137Z digest=sha256:1416a14c4d2cdb866f36bf3450fd77dc01c61ce88790a3547c6e75635711d8a1

Observation f617798e-60c8-40fd-88a7-fcdb4c02904d · inbound

Ask, Solve, Generate: Self-Evolving Unified Multimodal Understanding and Generation via Self-Consistency Rewards cites this paper.

Ask, Solve, Generate: Self-Evolving Unified Multimodal Understanding and Generation via Self-Consistency Rewards SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:39:51.544735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T04:58:15.891214Z digest=sha256:87430219eaa079fc2a42c730fc432b24c3ee48c9bd38795f57109454b1fcac91