Pith. sign in

Paper Citation Record · LEDGER

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models

As of 5 August 2026, this Paper Citation Record lists 62 of 62 outbound references and 0 inbound Pith citation observations for arXiv:2604.25636.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2604.25636 v1

Coverage vector

measured 62 of 62 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-07T16:55:19.763050Z

measured 62 of 62 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

62 of 62 outbound references displayed

  • verified exact6
  • verified fuzzy33
  • unresolved2
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch21

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 66746c33-04a7-4097-b5ed-e6934f571539 · outbound

This paper cites GPT-4 Technical Report.

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models GPT-4 Technical Report

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-05-11T23:31:14.379637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:cb491c612e5a0f88188f5135f1bb294eb45faa53ae1e20255a2bb733baaa4d70

Observation 742fe87d-8177-4a53-881b-c38ce93678eb · outbound

This paper cites In: ICCV (2015).

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models In: ICCV (2015)

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:38:23.267962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:c4f5bf57b0885cda14248194101abf057c401131d99468abdea4d1943cd6a5e6

Observation 7cdc6705-f45e-4059-aee3-e3e53f4afbba · outbound

This paper cites Qwen2.5-VL Technical Report.

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models Qwen2.5-VL Technical Report

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-05-11T23:31:14.256993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:c204caebc70a503184553bdca7be250a38bd735f605e849fdc15f9ab972e2828

Observation 92f01916-d636-419b-b19f-2dab002d8772 · outbound

This paper cites OpenAI technical report (2023).

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models OpenAI technical report (2023)

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:38:23.280254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:e1abac0d073e7614f9f4bf5d2f6b239e3ca598739e990a79af2a1739d4079243

Observation 6e774047-4477-4da0-8c19-0e869a999b2e · outbound

This paper cites HunyuanImage 3.0 Technical Report.

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models HunyuanImage 3.0 Technical Report

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T02:02:32.945705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:2d18b56a0e2f12178cc4f05924e3859c16166047879c7e9acaf1f60a009a8e00

Observation b3efde80-7ebe-4c57-a0da-8bdcc47723b4 · outbound

This paper cites BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset.

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-11T23:31:14.365329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:6612a72b26ed769f9fcfe162a05f7aff5b9e4c620d717190180a4f0c36830bf4

Observation 542bfe51-e561-470f-ab2b-a78b7dcb5d7a · outbound

This paper cites Blip3o-next: Next frontier of native image generation.

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models Blip3o-next: Next frontier of native image generation

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T23:31:14.349163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:2907debe32c10557a8da7e28ec735bcb7938c7d70f524ff04e10254a9f0340ec

Observation d6b539a3-1f05-44bb-b13a-23e740583d2c · outbound

This paper cites In: ICLR (2024).

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models In: ICLR (2024)

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:38:23.366624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:a92957a8a46c3153fda1e921972d7a390cb3fe15a5646b877b593edf6f0cd0e5

Observation 338520bb-665a-48dc-ac89-5413ca7a6aa3 · outbound

This paper cites Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling.

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling

Reference 9

Resolution
metadata mismatch
local_arxiv, observed 2026-05-11T23:31:14.333401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:8c0c24410e55fc649a8477c9222850f22ec24a80024d2bd75709721b5d063985

Observation e8aa972d-b59b-4ae5-a43a-f5d79377e316 · outbound

This paper cites Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities.

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities

Reference 10

Resolution
metadata mismatch
local_arxiv, observed 2026-05-11T23:31:14.400267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:50537b19a93c4b6ac506c9c711760502a0ea6bda8138807d138a15a9e20bf916

Observation fa2aa916-f813-4e28-9b4f-34849fbf40f1 · outbound

This paper cites Emu: Enhancing Image Generation Models Using Photogenic Needles in a Haystack.

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models Emu: Enhancing Image Generation Models Using Photogenic Needles in a Haystack

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T23:31:14.407195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:f7ef14d06f6978026970052c07cb3c62812d0165be898b8197a5c3d7bfdb99f7

Observation 25903ecf-358c-447c-b16d-6a81c8632045 · outbound

This paper cites Emerging Properties in Unified Multimodal Pretraining.

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models Emerging Properties in Unified Multimodal Pretraining

Reference 12

Resolution
metadata mismatch
local_arxiv, observed 2026-05-11T23:31:14.374331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:fd3c7555d5ee519fed72a1dd6927ef64db86b1696ef8ed9e92cd5252da68f49a

Observation 466d7839-3afc-4190-9224-a2c3de5df50c · outbound

This paper cites In: ICLR (2021).

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models In: ICLR (2021)

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:38:23.375516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:dddbd76b36c2b9fc2111da6086327118fd2d49e682fc70b4b276f12d5e27f8fa

Observation 938cac5b-cde3-4185-a996-e1af56ef9baf · outbound

This paper cites The Llama 3 Herd of Models.

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models The Llama 3 Herd of Models

Reference 14

Resolution
metadata mismatch
local_arxiv, observed 2026-05-11T23:31:14.355301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:d6770b82fb41749e4dd3a85778028b06d57eeff595989ed03bcce4a8035ba0cc

Observation 412cd8c0-474f-4ade-8f72-5b8603e7b4ac · outbound

This paper cites In: ICML (2024).

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models In: ICML (2024)

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:38:23.303157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:2c8cbbc9b5e97cffc67062852763a96ce1deaee4eb690dcdb560858e99c5d566

Observation 886bb0bb-19dd-4f22-892b-51eb27cde6d6 · outbound

This paper cites SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation.

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T22:48:36.519178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:c763170ba420aa7bef449651e58b9ba6426a3d5bfbdfe6458dce60fa592cab55

Observation dd05ed90-a635-44bb-b98e-bd183fa3a6dc · outbound

This paper cites In: NeurIPS (2023).

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models In: NeurIPS (2023)

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:38:23.351888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:d28b650aa01be511a34832d9da2cb46e9787e126744f1c5732a7309f69f87936

Observation b07e7362-a084-4bdf-8dd6-e3bd13c2ada1 · outbound

This paper cites In: NeurIPS (2014).

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models In: NeurIPS (2014)

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:38:23.363965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:7fb9332ae0ad2f2ea0045b13825af1eb07470fc580af8d64aa571afd1057c9a0

Observation 2487f265-66d9-4e67-ac71-6a21dad5014b · outbound

This paper cites IEEE Transactions on Circuits and Systems for Video Technology (2023).

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models IEEE Transactions on Circuits and Systems for Video Technology (2023)

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:38:23.326571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:b57234272baa0f93bd135a08b84287dc166732e482e8a0da6d7cb95d8e9bc875

Observation eaaf3746-fa02-4baf-a500-7c18d397a83d · outbound

This paper cites In: CVPR (2024).

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models In: CVPR (2024)

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:38:23.372256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:808f6f1cbb5906817d86056e5ce576607b22a5f6a6a7c23d879631084b3c6746

Observation 90903038-222f-4397-91d0-f322573cb4f2 · outbound

This paper cites In: CVPR (2025).

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models In: CVPR (2025)

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:38:23.291167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:00449a31583e9494bd7df319797f93fab9a1bdf7348463fbbb99e8c25d430f74

Observation 621afa05-e237-4d8a-aa4a-0ef898191520 · outbound

This paper cites In: NeurIPS (2020).

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models In: NeurIPS (2020)

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:38:23.348799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:7ae72d4b2b406595f1af9474643dbb3806d711a3c9c6fa88a5996f2abe5e11bd

Observation 50b30e5e-492c-4209-bac3-f3929db4bc3f · outbound

This paper cites In: NeurIPS Workshops (2021).

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models In: NeurIPS Workshops (2021)

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:38:23.385717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:c8846dd103e914e07777cbdaba0564766ce13d9082940641109635837dd4bd09

Observation 7ee3b37f-0874-49c9-9bb4-9ff8ec711a75 · outbound

This paper cites ELLA: Equip Diffusion Models with LLM for Enhanced Semantic Alignment.

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models ELLA: Equip Diffusion Models with LLM for Enhanced Semantic Alignment

Reference 24

Resolution
metadata mismatch
local_arxiv, observed 2026-05-11T23:31:14.230321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:cf5bc209f19cbf09b2a225591fb856757cc788c5b47b42a7dbdaa400543f2bd9

Observation b890564b-ba3d-47d4-89d5-74e044db0147 · outbound

This paper cites Interleaving Reasoning for Better Text-to-Image Generation.

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models Interleaving Reasoning for Better Text-to-Image Generation

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T23:31:14.310799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:13b538a50376fe32aa5ef27519ecb9c31022a3491d85a45539d124c84cf48a9c

Observation 785b158f-3276-43fb-8bd1-a331380300dd · outbound

This paper cites GPT-4o System Card.

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models GPT-4o System Card

Reference 26

Resolution
metadata mismatch
local_arxiv, observed 2026-05-11T23:31:14.210187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:fb83d4f16d2d9c687ed85751880eb2eed29ebbef58a86c87378842420e5027ff

Observation 1b567b6b-60ff-44ce-8671-b56d2266e451 · outbound

This paper cites In: ICLR (2015).

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models In: ICLR (2015)

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:38:23.354733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:8ec2663037073acbf27a9fd5e665f76c4c5fb504cf9fc90296b0f642aba641ca

Observation baab5b39-edc2-4b57-b1a7-6e5a25c7668e · outbound

This paper cites an unresolved cited work.

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-05-27T01:38:23.334091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:776efdbd1d0b8e6f0b93ccc4ee7b97f60bbffe23389b13c7c58f701d6a2f9350

Observation 02aacbf4-f8fe-42cd-a890-bf1b6baf1e64 · outbound

This paper cites World Model on Million-Length Video And Language With Blockwise RingAttention.

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models World Model on Million-Length Video And Language With Blockwise RingAttention

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T06:36:57.375169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:462e5e87e68f0f0612e23470dc73c3d47d2d9fd67dab6e4b36d2c55ea8742990

Observation 9aa4180f-cbec-450b-8ded-38973bc4a312 · outbound

This paper cites In: NeurIPS (2024).

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models In: NeurIPS (2024)

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:38:23.318961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:2edd02cf7df9464019282e4a92c61f14838eec02a744f6dbead1719b83e23c37

Observation 898cf4d8-d3b4-4fa7-8d62-f51923af0c06 · outbound

This paper cites In: ICLR (2022).

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models In: ICLR (2022)

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:38:23.360675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:3c290cb33f654c424fc815485bf4c5ab7a39a0c2f01c8b68daf32d72c2a48faa

Observation 516cf467-f43d-48f2-8b81-8d699fd19a60 · outbound

This paper cites In: ICLR (2019).

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models In: ICLR (2019)

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:38:23.287490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:d62a63e482e30c939016f8d2a62bc2cf18619c97d9ce5067820d3b74f24b63d9

Observation a9eec926-8013-43a2-ac56-5cb58c334491 · outbound

This paper cites Understanding-in-generation: Reinforcing generative capability of unified model via infusing understanding into generation.

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models Understanding-in-generation: Reinforcing generative capability of unified model via infusing understanding into generation

Reference 33

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T23:31:14.422833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:69ba8130858290ae24339268f2743a130bd95d8de75723ec8d7467bb32b78683

Observation a56697e3-8d02-4b0b-9224-d0adfb1bcffd · outbound

This paper cites JanusFlow: Harmonizing Autoregression and Rectified Flow for Unified Multimodal Understanding and Generation.

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models JanusFlow: Harmonizing Autoregression and Rectified Flow for Unified Multimodal Understanding and Generation

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:31:14.274607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:46373812d57dd47f9e96e95ea6eef9870a6ac4634f0390baf379a0db396c39ca

Observation f4ca73c9-c430-42c3-86ef-de53a71b36c8 · outbound

This paper cites In: ICML (2022).

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models In: ICML (2022)

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:38:23.299110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:0ad4e57e860ecc529ff299faf1740684c9b37d07b00e5a5f42a1cb8168106dd9

Observation fab50a81-a0ea-4e40-9eb8-d3f0f37c4aa0 · outbound

This paper cites Transfer between Modalities with MetaQueries.

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models Transfer between Modalities with MetaQueries

Reference 36

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T22:49:23.311693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:15f62cc798c1e0e3f6650c8be50f38104ac0c3fe4861a4cedb44afb75b30f50b

Observation 895007c7-234e-4e51-b1e2-0ebe9eaf6ad2 · outbound

This paper cites In: ICCV (2023).

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models In: ICCV (2023)

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:38:23.314511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:5cc9a456cbd46ee9866d32e970f9bb150ec50973d865cd15942f15318e62be1d

Observation 6a8d0e67-f515-4266-82b1-82352a3e5713 · outbound

This paper cites In: ICLR (2023) Refinement via Regeneration 17.

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models In: ICLR (2023) Refinement via Regeneration 17

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:38:23.283665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:54154551e91105d07ce688e254d42ba876109895dbcb1128f484eb6b0b2d6cf4

Observation ed897660-6b89-4114-8526-5c051de54f7e · outbound

This paper cites Uni-cot: Towards unified chain-of-thought reasoning across text and vision.arXiv preprint arXiv:2508.05606.

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models Uni-cot: Towards unified chain-of-thought reasoning across text and vision.arXiv preprint arXiv:2508.05606

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:31:14.317354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:574e3cb151548905b33c6cd9b47168e4ca08da39526e1fb1fddbd570a995794f

Observation 92298813-f783-42cc-aef0-3b5957a413aa · outbound

This paper cites In: CVPR (2025).

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models In: CVPR (2025)

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:38:23.310439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:a9e961c23f24d2e9b757bfa151ab987411b23b626ac4d6feb9040ad5d843b2b3

Observation a7812e24-4cb8-4ff0-8b7e-b91ae36638ac · outbound

This paper cites In: ICML (2021).

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models In: ICML (2021)

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:38:23.379441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:73189a749d2e216bee6df1fcab51507a8034460c501cbb2c031d827a44e2fc43

Observation 78c3f9b5-5a92-40c4-ad66-6feaed42b426 · outbound

This paper cites JMLR (2020).

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models JMLR (2020)

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:38:23.271918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:1725366c9e2ff7896df37a38dbe206475466d7bff6d5bf9e0accadea82660518

Observation c13a2dd2-6e3e-4fc1-888c-a83931941cd8 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 44

Resolution
metadata mismatch
local_arxiv, observed 2026-05-11T23:31:14.283171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:2a46f67f60bb028b794ab0f44c95f0979b54f2680e9b311ad40de5da795a5948

Observation 88fba717-e53d-4491-ac62-e3124cc0e7c9 · outbound

This paper cites In: CVPR (2022).

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models In: CVPR (2022)

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:38:23.344224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:1ab2acfa540e94d252e4f5ba19058da7750ad56aeb96a0ef3b1965ffd513765a

Observation 92440000-655a-4ef1-a5be-2a5e7f7b8bf4 · outbound

This paper cites In: NeurIPS (2017).

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models In: NeurIPS (2017)

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:38:23.382510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:57aaae6087d570bafcc2cda2f8d922d16c73cf3393dc8a7cb0d7eee7792f6b03

Observation a787bcc2-ac01-48d2-8504-8f1d25ab28d5 · outbound

This paper cites Chameleon: Mixed-Modal Early-Fusion Foundation Models.

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models Chameleon: Mixed-Modal Early-Fusion Foundation Models

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-05-11T23:31:14.432128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:01f38f771f94d6be982dfd304e0e42fd6909d99d7f1d0c09eea3cf349edebb9e

Observation 0b137f40-9d08-45c3-92c8-7758a776d24d · outbound

This paper cites In: ICCV (2025).

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models In: ICCV (2025)

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:38:23.306938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:131b2789e144b24df781873df6f83f920f69be74afaca32b414c4c63504fc22d

Observation 58bfcd0f-040d-4bf0-bb8a-cbd367e0afac · outbound

This paper cites In: ICML (2025).

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models In: ICML (2025)

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:38:23.330377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:bc6a1ba35d8f2cada432dff439c431a29a702bf744e56251ce7d64bd36515cb1

Observation 28a2128e-f99d-4106-9dc2-b5ec8a30a8fd · outbound

This paper cites arXiv preprint:2509.04545 , year=.

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models arXiv preprint:2509.04545 , year=

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:31:14.298013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:68a76dea361816fca29ff77ab4c17e8b1848e3bc3030116803ba886eb3218eb2

Observation d8e6a60b-97cf-489a-8fcc-e3ff9c89a519 · outbound

This paper cites Emu3: Next-Token Prediction is All You Need.

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models Emu3: Next-Token Prediction is All You Need

Reference 51

Resolution
metadata mismatch
local_arxiv, observed 2026-05-11T23:31:14.291656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:c20ed5dff7184c36afad5d7958258defa84ec92df90a21310d00b18ffca54add

Observation f73691fb-24a9-494f-a8c1-f54689fd8e56 · outbound

This paper cites Unigenbench++: A unified semantic evaluation benchmark for text-to-image generation.

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models Unigenbench++: A unified semantic evaluation benchmark for text-to-image generation

Reference 52

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T23:31:14.224632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:fa77e767f738e206b2badbcc29f7b602f0cff9991a62287bc36580b627179d5a

Observation c9fbc888-e8a8-4844-a323-c2c72e31c3e4 · outbound

This paper cites Qwen-Image Technical Report.

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models Qwen-Image Technical Report

Reference 53

Resolution
metadata mismatch
local_arxiv, observed 2026-05-11T23:31:14.238859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:912cd34d8a2305e429ffc95f0bdf1eea8098bfe72642109d8a08c44034923476

Observation 8ddef5ec-63a0-4b63-8234-32ae3fdc7c84 · outbound

This paper cites Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation.

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation

Reference 54

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T22:09:16.660059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:cbe668370c5af50416284c76c9a896ec1e0154fba07dcd18f1991d9ca28dadcb

Observation 9901447c-408e-4778-a8b0-b70e44c7671f · outbound

This paper cites In: CVPR (2024).

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models In: CVPR (2024)

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:38:23.322876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:2395b9038a38da016916cd3ee4d9f066781f0b4529a70f21247290a4b7e490bd

Observation 727da076-e1c8-4ffa-a8b7-847924a8a4b8 · outbound

This paper cites In: ICLR (2025).

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models In: ICLR (2025)

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:38:23.357575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:d9135205db42cf9fc29114c43ce2f5e9b44c6c02587fd2d7e5183ff0183bb12a

Observation f444cbed-80ae-4098-88a7-9d2761f49d28 · outbound

This paper cites Show-o2: Improved Native Unified Multimodal Models.

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models Show-o2: Improved Native Unified Multimodal Models

Reference 57

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T18:51:16.424084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:f7ff9f42b5669865e1526dfda1ae0704a75bb457fef937c4245a261fc68e4602

Observation 5946345f-37d6-4d45-a0e8-04e96b1b7ce0 · outbound

This paper cites In: CVPR (2024) 18 J.

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models In: CVPR (2024) 18 J

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:38:23.295297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:53e3ec333ea1165a3970dfe62ed335f5ef8cfa725f08bdcd806fd8a7f10b8fc4

Observation 5b2bc67d-de4c-438c-abad-a07d1d24d66f · outbound

This paper cites In: ICML (2024).

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models In: ICML (2024)

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:38:23.264276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:69ba7c66190344b52185ae028a642795530ae81cab73235a8001b6a79054f57b

Observation 2a71cecb-7568-411d-af13-5b903aab4987 · outbound

This paper cites In: CVPR (2025).

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models In: CVPR (2025)

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:38:23.369243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:c2dfea9e628b10acf93d873820f8a2bc37c6958e5920b8330fcf4f32ae6374ab

Observation e6e589ea-a24a-4b02-966c-ee2bfdac3b4d · outbound

This paper cites In: NeurIPS (2023).

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models In: NeurIPS (2023)

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:38:23.341233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:ef6b3fb6fb29552710588b2bf3c798372e4bee86f5456c45fa5dfa164d3e9034

Observation 2a2ec9d8-df88-41aa-ac43-bea1a4525eb6 · outbound

This paper cites In: NeurIPS (2024).

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models In: NeurIPS (2024)

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:38:23.276763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:151fd968d3ef0e8c8d0a61fcb1d4073aa710458066d2367077255d1d26b861da

Observation 4d426fa5-d99a-40ef-9858-2fb0405647df · outbound

This paper cites an unresolved cited work.

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-05-27T01:38:23.337876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:55:19.763050Z digest=sha256:cae3685e6a21cb814484892b3e2f8bee2d2f5856ef772838a550d1408cb5222d

Pith citing papers

No inbound Pith citation observations are available.