Pith. sign in

Paper Citation Record · LEDGER

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards

As of 7 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 3 inbound Pith citation observations for arXiv:2506.07963.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.07963 v3

Coverage vector

measured 66 of 66 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:30:27.240877Z

measured 69 of 69 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T09:54:24.937137Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T13:39:51.542556Z

Reference resolution

66 of 66 outbound references displayed

  • verified exact1
  • verified fuzzy2
  • unresolved63
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 34e31048-d3a1-4dbf-9d6a-5892834d9ee3 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.197768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:26.928846Z digest=sha256:02b0d9f61bbbf7291cc2c96876f1f4c685d6c288c511eb689f6dc9b9c2d48145

Observation 030bf8b5-b31c-4f67-a578-be7f28a15ce5 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:26.934199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:26.934199Z digest=sha256:ed8132dda4b4231166bcdaf221a8be3e1b78d197be63f54f48f66e48cedb8b63

Observation 476e7ccf-7c3a-451d-a10d-6cbd6e7fccfb · outbound

This paper cites Qwen2.5-VL Technical Report.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Qwen2.5-VL Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:26.939490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:26.939490Z digest=sha256:4400099f7a180a1f62a385e3f60b260d5f838e82ea29043fd8135d740c36ef35

Observation bb092b5f-e680-4bca-ad89-a86b000b4bd7 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:26.944775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:26.944775Z digest=sha256:e40a21ef8890bd217919d310a7cf1ec9d3cf13dcd3468441d71d0ce4e7f56e29

Observation 39545832-955e-442e-8d82-e5ce221c45f4 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.173054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:26.949807Z digest=sha256:dc50ff2c798670a246b2d530f9eb17fe0ae558a2028e18425f9d879fa5551542

Observation f5695d3b-42fe-4661-9800-339c91ac2dba · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.158281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:26.954721Z digest=sha256:bd2c3f56e43cfaa4e8679fc2140dd13cb510c223129536ea91f6a88b278f452f

Observation f66c0f2e-a84a-4991-b9ed-9f51a6d12c29 · outbound

This paper cites Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:26.960872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:26.960872Z digest=sha256:fe1cbe4b489432297de095a54b91893d4a5482a71e0a1b6277ed0badfb593d80

Observation b9fda4c4-1c7e-457c-9ebd-83b4312d8a5a · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.144043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:26.965722Z digest=sha256:5421e8dbd41ef1b4468a0b938b820fe284e6d77e7ca35c5382ccfb4911035886

Observation 6f8447af-42b3-4a09-9334-d6f56246f97a · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.129436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:26.971313Z digest=sha256:08b32bd9e1c77fbe43c0aec321023c63e06ebe8a6bb6f16b5af2ea614ad24cf0

Observation 5e2e6c19-fee8-4a54-aaa2-8aa0b24ca236 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.113226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:26.975940Z digest=sha256:81d49ddebdf1e370d4f3fad650f84afac32fca5b6b5168ccf77a677c939aa831

Observation 66ffc0c5-a7e5-4aea-9ed5-6e0e1b0cb54a · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.096940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:26.980833Z digest=sha256:7b81b34cd31ca59236545f82552675fe7d706214e89d97fefea85c726f14743e

Observation c05877b1-550a-4e8b-b27c-de10ba4401f1 · outbound

This paper cites Training-Free Structured Diffusion Guidance for Compositional Text-to-Image Synthesis.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Training-Free Structured Diffusion Guidance for Compositional Text-to-Image Synthesis

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:26.985271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:26.985271Z digest=sha256:4dafb70cd52e5fcf605b40eb6592613768d7b79b95b3444e6c2e093caaf0c33f

Observation 049a9ad0-4ba2-479a-a5f7-faee30561381 · outbound

This paper cites SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:26.990200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:26.990200Z digest=sha256:c6ecca60c597dc618bc6f0683fa18ccf543f1987e3984f653144829a1a048ff6

Observation 9549833c-2741-48fd-bcd2-f517171612cf · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.081911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:26.994948Z digest=sha256:7b1cc93ec7b466246189d9bbdb8623d09ca6a38efcad0698ad0f97be33d1fcb0

Observation 152d9bfd-1b77-45e7-81d2-df9a777c97be · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.067245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.000145Z digest=sha256:77a82f12ec8eb579028e2ae0544eaac6086a0eec105fbd10e678cdc333f01c73

Observation b8526ddb-73f2-469d-adfb-ae8cb4489e75 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.052371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.005060Z digest=sha256:e8e6985ea06b5bdd6eacf407c2a3dcc570bff51bf764bd51175449c1a9d347dc

Observation 178dfa81-561c-42dd-8a3c-99f3cbd1dca0 · outbound

This paper cites T2I-R1: Reinforcing Image Generation with Collaborative Semantic-level and Token-level CoT.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards T2I-R1: Reinforcing Image Generation with Collaborative Semantic-level and Token-level CoT

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.014921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.014921Z digest=sha256:d67843f414e1dae80b4cfa7e67b9fb3b8fa448ecfaa192064bddece2e85a0ef2

Observation e11f9d04-2253-495d-a522-fc5df98a9761 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.037690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.019805Z digest=sha256:d48e81ef3cd24bc0cd635435ad3ac3752d0eb306b9f90d59c47c9ae71ac90d81

Observation 1916b5cd-270c-4049-9fe7-74bbef8778db · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.024116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.024116Z digest=sha256:6ebac25d4ba96b8bf294186f57971c2d057a9e50e746a903e141bc1087302799

Observation 0e05fe86-6b9a-404b-bac2-c61e4d0707ee · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.028909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.028909Z digest=sha256:32a7801885f25d228f611acc6d0a2958fa2d6a3a8c9396560bca98d18025e7d2

Observation 01e38c16-6906-4a5d-964b-345271d45717 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards LLaVA-OneVision: Easy Visual Task Transfer

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.033266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.033266Z digest=sha256:bd5c0631b8be20380869eba81a964f1d7a7f167d65b6e63ee719a607d0c0a0ae

Observation effd4c6d-cb0a-460d-abea-bc7c97d5b6e5 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.003881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.037679Z digest=sha256:d67ca3f6341b5fc463afabd8658e8b99525ac41a516bbe427cfbc2e19e705109

Observation b74913d7-a687-49da-bcfd-170f68407a83 · outbound

This paper cites Dual Diffusion for Unified Image Generation and Understanding.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Dual Diffusion for Unified Image Generation and Understanding

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.042171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.042171Z digest=sha256:59f3d2d69650077ec70efaf75b199bdea61e768bb9b950fb53b1cab081fe3fe8

Observation e153f12d-2664-414a-812b-58be4ac2903f · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.047061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.047061Z digest=sha256:6a35a54509ff7db8a46c30cf72105dc391209c413c34c064ff286206e05344f1

Observation 62d1c41c-1cf1-4d21-a775-9b18b17991ff · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.051526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.051526Z digest=sha256:eac43e542fa873173d3f716c1c1c69e3b994e7d3b7d9be050eb8d23a45126686

Observation c98c7cb2-aeaa-4896-8cef-475a5211411e · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.056021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.056021Z digest=sha256:5d51a74d2c68c7ffc08e3743720119a2b6e8d2e27fe97f152d4bf28d9275b2d6

Observation 5126711a-d8f7-42b9-a1ab-4f7d04482d9b · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.960574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.060265Z digest=sha256:e2499e8e2dc1eaf80f240416f7efd55f207e851575ce0a9a4684b161ebd2f2a8

Observation 6cc3d5e4-6c41-4baa-b85a-b035356585b0 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.945919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.065128Z digest=sha256:962a1a90a8ba7bb4f0ee7a3458e3c71f4fd51ea5eb5b39a37c326808b5887cec

Observation 83a12574-4bac-483a-94b9-b89428aeec13 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.931317Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.069704Z digest=sha256:bfd537e5efbc037f422fbf61e3ce915c75ed05ea2f2ebadf905d8da9b599abfc

Observation 27be12df-04bf-49ef-a6dd-9370d2b9bd23 · outbound

This paper cites DeepSeek-VL: Towards Real-World Vision-Language Understanding.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards DeepSeek-VL: Towards Real-World Vision-Language Understanding

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.074181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.074181Z digest=sha256:cf79551927d9f676aa2658b12c33826c528b4a24b6c776b9355606b1ab947bbe

Observation b416e6a3-2ec0-43c8-b9e5-f6ae9822a562 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.916308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.078965Z digest=sha256:9974b4bec97d42e73f740021a224b4b3d9003fce8967e2f91b05507046ed7ad6

Observation 72e653b7-f105-4d20-ba0d-a85145ae492d · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.083321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.083321Z digest=sha256:3ed24739c319821f2765df3df72817246c91744612968e8c21f4309528fd69a6

Observation 07f768bd-9e98-4005-90c1-90ed1f1de104 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.893081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.088748Z digest=sha256:1d5aeb8ca3c4ed1ca4bdc53908f20e83021c6426e0bb75031fbda0d12ec2d0a5

Observation 3dff539c-e43e-4b39-a19c-e752418731de · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.878873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.093709Z digest=sha256:d2d82f7bba102dad7356b4b11bd26834f7604ab524ea468b2d78e7dc02a76838

Observation 1fd6c9af-f1c7-4e89-b4c3-9437bf9f51e2 · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.098652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.098652Z digest=sha256:dc418f0e215e10f0a0994fcda4d037edd352cb403acee9f86dc7a67f8f37b572

Observation a6bdf839-3534-404d-a26d-b0076853d7a6 · outbound

This paper cites SILMM: Self-Improving Large Multimodal Models for Compositional Text-to-Image Generation.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards SILMM: Self-Improving Large Multimodal Models for Compositional Text-to-Image Generation

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:30:27.503745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.103280Z digest=sha256:1bd999880699ecdbf63c78801752140d37338e3f0257dd7e788184be125e0624

Observation b677b4c3-48ab-4d87-847f-422ddf6996c3 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.864651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.107881Z digest=sha256:c5a6fe78a0b955d1db4778853d77ec07ad57cedd6805efd2433102fd60724cd9

Observation 48cd6510-b10f-4df2-b7f5-8c3075e3e33c · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.112713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.112713Z digest=sha256:c5b60616289db6d2eb868d93f6f5f61dac7274f6891282ce8b1a70d4599bb20b

Observation e28f5d81-0f6a-46fb-bf91-c583282c3836 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.117606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.117606Z digest=sha256:464fe1417ab3224118c208eecc94416f70c0153c1a1be6ac2e88fcffc2c1ea1d

Observation 5fe7815c-4c28-4f19-9bb1-148063d24c0e · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.831519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.126900Z digest=sha256:e05618db69218229889ae0ae140e385f564931820d7245b2f432d95dda53948e

Observation 3628d465-384c-40c4-b3fb-ec01a9a60451 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.131912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.131912Z digest=sha256:9615fffff61e7e36e4ac798f25847685389aca62de6b08ac2031eb160f7a8019

Observation d83006e3-c772-498c-a0c0-28521e5da2bc · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.816871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.137164Z digest=sha256:102cb4fbd0763b730f274f1169b2931cb1c553020ae2dc57053110b1c6c2b466

Observation 4d50a96b-3af2-4a9c-834d-43bea0ecca57 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.801911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.142267Z digest=sha256:da2455b1d213fa6aad388c936df96aa0b52cce142f9cd4d67b6b13564e475135

Observation 8903f20c-529f-4f42-a8b1-ed7715ef63b0 · outbound

This paper cites Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.146722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.146722Z digest=sha256:512f42e364d8756138edfa9f71fb26481b52dea146aa743a42517890671aef0f

Observation d61d2a31-dcf0-4c6a-b518-06fcf2588a83 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.151759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.151759Z digest=sha256:2fd382037fd85ea613450aa63f00482fb32d0bdceaaf60ff5e5120a7ba11837a

Observation 0e3b3c1a-db97-4757-affb-f25f766bf9bc · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.776772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.156139Z digest=sha256:8f80b595408c9590596ce198d76c3c079d7057740399d64452cff304b80a3ab9

Observation 89edf2a4-b34e-4127-8ac3-2f4475933e0d · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.761861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.160476Z digest=sha256:c2c153d45d50d59c98505debf172c1b70708db5551f78cc2b95371271dee4601

Observation 55b058cf-fc10-4d8d-8c92-7e0ebf2249d5 · outbound

This paper cites ILLUME: Illuminating Your LLMs to See, Draw, and Self-Enhance.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards ILLUME: Illuminating Your LLMs to See, Draw, and Self-Enhance

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.164874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.164874Z digest=sha256:e20abc4a1e4e887e9ba764188a967bfe9f4bedb900294ff88fd172799d93e584

Observation 1df0a6a2-8f98-49df-a235-21116094180c · outbound

This paper cites SimpleAR: Pushing the Frontier of Autoregressive Visual Generation through Pretraining, SFT, and RL.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards SimpleAR: Pushing the Frontier of Autoregressive Visual Generation through Pretraining, SFT, and RL

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.169774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.169774Z digest=sha256:841152d8d3067137b3cb8e36b408a8873646f3c38ddf46254988390aff19d094

Observation 3250fd9a-198b-4e96-ab3b-6c82dbca1af2 · outbound

This paper cites Emu3: Next-Token Prediction is All You Need.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Emu3: Next-Token Prediction is All You Need

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.174501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.174501Z digest=sha256:2f7bb99533557097bbdf5dab4b90541d2e95979c1293adcae8d483badf09bbe7

Observation 8a602248-88e9-4179-af84-35c9e7f88507 · outbound

This paper cites Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.178910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.178910Z digest=sha256:433976ed8b9e4c7c54af78cb68b4b2da2e2ac484990806549aa2a6a1c9f291aa

Observation ca31300d-1de9-4318-b41e-fbd39d46a64b · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.744102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.183828Z digest=sha256:1d0bbcf2192235f82d38a7131aa2c062a849b4b2937b94d8828686cdf97083b7

Observation 94835a3c-4fb9-4b76-aa5d-61e8fcc7bc79 · outbound

This paper cites VILA-U: a Unified Foundation Model Integrating Visual Understanding and Generation.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards VILA-U: a Unified Foundation Model Integrating Visual Understanding and Generation

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.188204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.188204Z digest=sha256:df761f0de33566aea1d49ed48c1caa3637fc9a74b9b17b10b1bca0de119d3b31

Observation 3b852fb8-cf91-4dbd-8c73-f0113944ef29 · outbound

This paper cites DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.192673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.192673Z digest=sha256:3cff09c285cac5e839c10a237eb1e3c595e339a562140a732404f35e2255ece5

Observation f81772fb-3b25-4026-8d2a-e621ad33bdac · outbound

This paper cites Show-o: One Single Transformer to Unify Multimodal Understanding and Generation.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Show-o: One Single Transformer to Unify Multimodal Understanding and Generation

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.197315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.197315Z digest=sha256:5ad4819ca4fb86213eb4a6a041f13bb00439305187153806952ebb165ef87d69

Observation 3ba61ad1-768a-456e-a447-82f18c302af3 · outbound

This paper cites X-VILA: Cross-Modality Alignment for Large Language Model.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards X-VILA: Cross-Modality Alignment for Large Language Model

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.201831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.201831Z digest=sha256:f20c977ac1bb35bcb7dcd54a172638cdf087b9dfa2fd5819d13577c1ddf96865

Observation 4124d818-2734-43a4-84d0-a5b0249efae7 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.729510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.207074Z digest=sha256:d893dc549d3d48b03bbfd82de011f17b612eb96830796c053c571b22856e7aa9

Observation 63e370f3-1875-452c-a273-bd578bf6d1e1 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.715018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.215968Z digest=sha256:26637e5038c22216bd3cf3dfaa5e474c11b0fa75611afcd118a5da3371a16264

Observation 2bf5185c-f9e2-4da1-8942-3f2e49ad5ced · outbound

This paper cites Ma, Simon Stepputtis, Deva Ra- manan, Russ Salakhutdinov, Louis-Philippe Morency, Katia P.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Ma, Simon Stepputtis, Deva Ra- manan, Russ Salakhutdinov, Louis-Philippe Morency, Katia P

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:30:27.699828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.220358Z digest=sha256:2deb12cd60b3f90da6ab36dcf82d7092ec90b2ffc8039c87858d22251dcb0e18

Observation 0b9288bf-20e8-44f4-9836-6d28f7abcf2e · outbound

This paper cites d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.225291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.225291Z digest=sha256:9ae226c7d7b7f7593b4b5c6f76c712f2c08f962e391bccf70b074f267cecd81c

Observation ff90e7cc-102a-4ecf-8603-1ea2061cff47 · outbound

This paper cites Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.230749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.230749Z digest=sha256:4a20a9e88d5f76fa81ae441f84600699a2928249c5228abd61d10db64a667118

Observation 4b54aed6-a1ab-4e57-8c64-3aa90cfc3690 · outbound

This paper cites Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.235770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.235770Z digest=sha256:2565357c2730058fbfbeb0906092eebab4eb6ea09516c11b00c76f280ce1539f

Observation cf990229-2d9f-4505-b12e-178a493f4849 · outbound

This paper cites M" and an.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards M" and an

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:30:27.684850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:30:27.240877Z digest=sha256:77a11463c648b25c37edd21742fd2692b8f328deafb8c55e49f2c7be5ed07444

Observation 8d0e7ef2-0718-41e0-8959-14284f5f9385 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.122335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.122335Z digest=sha256:05eefb18c7e0fa9a0d72d111f68d4590f6b3e02d4a659e27c58c949a62657741

Observation 41965ca6-c991-4de9-88d5-55ec184c603d · outbound

This paper cites Align Anything: Training All-Modality Models to Follow Instructions with Language Feedback.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Align Anything: Training All-Modality Models to Follow Instructions with Language Feedback

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.009844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.009844Z digest=sha256:683ea6706b576776c32eb3c6a123f5e1d8f250b21cf87d2cb17b26d579e10eed

Observation 1605fb19-019b-4012-bcce-2f1da7f47d0e · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.211480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.211480Z digest=sha256:ddb2dbfc664c869550c61b648bc5930f487d8d78dc2bb901741d8b755a7710be

Pith citing papers

Observation 2fe94ad2-fdfa-4a8c-9c70-8019974e7869 · inbound

A Survey of Reinforcement Learning for Large Reasoning Models cites this paper.

A Survey of Reinforcement Learning for Large Reasoning Models SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards

Reference 197

Resolution
verified exact
arxiv_id, observed 2026-05-18T00:05:31.604172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-18T00:02:24.352947Z digest=sha256:fa63b5d7cd8c5b2abde50ec5a6636f0f285c1dd77c75091d7766155a4757d899

Observation 7b72bc66-ff7d-4a96-9026-21567ceda278 · inbound

SRUM: Fine-Grained Self-Rewarding for Unified Multimodal Models cites this paper.

SRUM: Fine-Grained Self-Rewarding for Unified Multimodal Models SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-04T09:54:24.937137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:54:24.937137Z digest=sha256:e2887a7facce0b773a4ee18788ea5d34123f5e69c9d73f54226b90664ada0c85

Observation f617798e-60c8-40fd-88a7-fcdb4c02904d · inbound

Ask, Solve, Generate: Self-Evolving Unified Multimodal Understanding and Generation via Self-Consistency Rewards cites this paper.

Ask, Solve, Generate: Self-Evolving Unified Multimodal Understanding and Generation via Self-Consistency Rewards SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:39:51.544735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T04:58:15.891214Z digest=sha256:df99a9de90482f522cbf8bde7f2cd94f3e3fe035de8e298e911e1547f06d007d