Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-08T14:48:22.805268Z
Paper Citation Record · LEDGER
As of 4 August 2026, this Paper Citation Record lists 71 of 71 outbound references and 1 inbound Pith citation observation for arXiv:2605.05781.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-08T14:48:22.805268Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-01T18:56:58.714370Z
A source-named dated measurement, never combined with another source.
Source: cited_works
71 of 71 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 10bb8fb6-edeb-4e17-92d7-3e70b39f4d85 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5b152c83-e36c-416a-8a01-481ce1336fec · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Qwen2.5-VL Technical Report
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6bb45bf8-0403-4638-891c-dae03592dc85 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Black forest labs; frontier ai lab
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3eb67d1d-55e6-4c14-a0bf-5f36e1b0df3e · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Instructpix2pix: Learning to follow image editing instructions
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 55547654-a559-4f93-8355-ed0627988905 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation dd41f6e7-56e7-4d1a-9d75-6a982f0a017b · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b031ab72-d0c4-4ddd-93a2-57750730bffb · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Thinking with Generated Images
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 345cb79b-2f2b-487b-b7da-6d8d70ce0dda · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision EditMGT: Unleashing potentials of masked generative transformers in image editing
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ac2b1d96-5cd7-4cc1-856d-a176cdc78958 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Emerging Properties in Unified Multimodal Pretraining
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 25f4e5b8-dd67-4e20-b310-0009e062f9e7 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Dreamllm: Synergistic multimodal comprehension and creation
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7c6f48bd-bc1c-47b0-af80-664d9878c89c · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Scaling rectified flow trans- formers for high-resolution image synthesis
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation fe2be6f1-c4be-4b94-bb9e-1d382f0a71f6 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Mme: A comprehensive evaluation benchmark for multi- modal large language models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1d2bbd97-d671-4132-98c7-b617d66971a7 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4722ece3-867a-4179-ad36-e03911eace4e · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Geneval: An object-focused framework for evaluating text-to-image alignment.Advances in Neural Information Processing Systems, 36:52132–52152
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation db8c475d-c299-4cbd-8777-8a8342d33458 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Vq-va world: Towards high-quality visual question-visual answering
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3395922e-dcd6-44ce-b10c-7201a42e719a · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 835d9f7c-2081-4d8e-8261-9ccc71bd7f2a · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Infinity: Scaling bitwise autoregressive modeling for high-resolution image synthesis
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b5a41f77-f266-4959-815e-7a08135887f0 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision ELLA: Equip Diffusion Models with LLM for Enhanced Semantic Alignment
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 82b1dd93-f2a6-43dd-bf4b-7decbdbfd8af · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Anyedit: Edit any knowledge encoded in language models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6445a212-8b63-4809-8a5b-c1e80775daf9 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision yes" or
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9cbdd6be-e091-44c7-be74-ad6263659532 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Eq-vae: Equivariance regularized latent space for improved generative image modeling
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e1138c00-2eb0-4282-818d-b3b7af650329 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Nohumansrequired: Autonomous high-quality image editing triplet mining
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6b577421-5d82-4740-a92e-a7b9aefadf41 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation dff5991a-b75a-49ce-a82f-60ffae2bc574 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision FLUX.1 Kontext: Flow Matching for In-Context Image Generation and Editing in Latent Space
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4bbdc33e-89b1-425c-b4fb-26ad23fccd9d · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Repa-e: Unlocking vae for end-to-end tuning with latent diffusion transformers
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation af554a6c-01ac-4b7c-be23-96ac06311ceb · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Imagine while reasoning in space: Multimodal visualization-of-thought
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6c8f061d-d5b0-4af2-8f57-cf1830348974 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Onecat: Decoder-only auto-regressive model for unified understanding and generation
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d851a8fd-8365-4267-9b69-c12289c8e7a4 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Autoregressive image generation without vector quantization.Advances in Neural Information Processing Systems, 37:56424–56445
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 68ca3073-0111-4815-808a-da44b132fe78 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Mixture-of-Transformers: A Sparse and Scalable Architecture for Multi-Modal Foundation Models
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1ce28159-7216-4d13-b655-64a974fce6ee · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Mogao: An Omni Foundation Model for Interleaved Multi-Modal Generation
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8af7907f-c4a4-40a8-a6bf-dac4afd7dcad · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision UniWorld-V1: High-Resolution Semantic Encoders for Unified Visual Understanding and Generation
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9f026fb8-6fe4-44af-8459-a06fc3ba381d · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation fb46326a-f9fc-45bf-a142-7a6622433d1b · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Step1X-Edit: A Practical Framework for General Image Editing
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b29695d1-52a0-4c5f-a2fd-44d032af015e · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Janusflow: Harmonizing autoregression and rectified flow for unified multimodal understanding and generation
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 56ff0cf1-1875-4963-a979-70fe6d2aa5f9 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision WISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image Generation
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 492d389a-2974-4944-a700-9c47885e1172 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Transfer between Modalities with MetaQueries
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 87d28391-0fc1-40c1-b029-ab73b9eaf232 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Scalable diffusion models with transformers
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 14407db3-da7a-4ce4-a730-124670d26dc8 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Sdxl: Improving latent diffusion models for high-resolution image synthesis
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bde48f28-98ed-4f07-8e5f-03daad8248ee · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision High- resolution image synthesis with latent diffusion models
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2adbbfde-a18c-4541-a393-829bbbfa22a8 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Laion- 5b: An open large-scale dataset for training next generation image-text models.Advances in neural information processing systems, 35:25278–25294
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 59687e1d-c7e5-448e-aaab-aa817c4483af · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision SVG- T2I: Scaling up text-to-image latent diffusion model without variational autoencoder
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3e4e0ffd-964b-45d5-98c1-968bc43bf482 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Latent diffusion model without variational autoencoder
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 463a70e8-fdc5-489d-9f5f-e494804be603 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Journeydb: A benchmark for generative image understanding.Advances in neural information processing systems, 36:49659–49678
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8f089a44-edc8-48fc-9eea-36595f83fefa · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Generative multimodal models are in-context learners
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 57bc9132-c8bc-4d6a-b52b-d6bed248d7ab · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Exploring the deep fusion of large language models and diffusion transformers for text-to-image synthesis
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c11edd16-8839-4ce7-89d2-5ff81514a8c4 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Chameleon: Mixed-Modal Early-Fusion Foundation Models
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5cce2100-9b19-4717-9a7d-20fa9edbd5f3 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision NextStep-1: Toward Autoregressive Image Generation with Continuous Tokens at Scale
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c47e8962-5214-41ae-9763-ef8ba6967329 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision MetaMorph: Multimodal Understanding and Generation via Instruction Tuning
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a1b3da3c-d7d8-4e20-8b2c-98a41ef3bcab · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Scaling text-to-image diffusion transformers with representation autoencoders
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 76fc4755-77f5-45b3-840c-3c16c7a48942 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision LLaMA: Open and Efficient Foundation Language Models
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6daf6267-1aa6-4fa7-9662-65bfa5761ca7 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4a52d568-e384-4347-a93f-0a26a61eafd5 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Reconstructive visual instruction tuning
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3a9ca61c-c587-4eac-b716-b84075c2b862 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0c810bcd-36dd-4c19-acec-0f22177b4540 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Emu3: Next-Token Prediction is All You Need
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation dee93dbd-b567-4bab-9bdd-54cfba94765e · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Unigenbench++: A unified semantic evaluation benchmark for text-to-image generation
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6ea1ef48-e134-4a2b-852f-f60129561aa1 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Janus: Decoupling visual encoding for unified multimodal understanding and generation
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation eff83de5-2056-4c0d-ad4f-227e41458c57 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision OmniGen2: Towards Instruction-Aligned Multimodal Generation
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 467a9fa6-95b4-4b21-9f6e-c43f35fdd2ca · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Visual generation unlocks human-like reasoning through multimodal world models
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 07e9c6af-4246-4a18-bc7d-2f2ffbb24926 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Liquid: Language models are scalable and unified multi-modal generators
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 333d691b-d2e0-4bd4-9cd5-d6ba23abcb42 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision OpenUni: A Simple Baseline for Unified Multimodal Understanding and Generation
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 56957309-bc53-47df-b6f0-59e2b949f452 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision VILA-U: a Unified Foundation Model Integrating Visual Understanding and Generation
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7ce7be28-b763-4aa1-bc5c-c7508364adf8 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Kris-bench: Benchmarking next-level intelligent image editing models
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7e57cca6-0d17-4dc9-8172-6ed1e736f48e · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Omnigen: Unified image generation
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7af36f02-f6a2-47fd-a249-62fed1bf038c · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Reconstruction Alignment Improves Unified Multimodal Models
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 11931339-98f2-4a2f-81ed-6975f2a12951 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Show-o: One single transformer to unify multimodal understanding and generation
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f198310f-7118-497e-988c-b0a584cf4aa8 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Show-o2: Improved Native Unified Multimodal Models
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 88694500-2b08-4d55-9292-8d2b4b8d8545 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Qwen3 Technical Report
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d49e99d1-c05f-4a1d-bfec-ce135a40c3be · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Imgedit: A unified image editing dataset and benchmark
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2814af93-2e61-4341-be89-ab3f56295e66 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Representation alignment for generation: Training diffusion transformers is easier than you think
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation edd8b983-aa47-4415-a898-928e5a832812 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Magicbrush: A manually annotated dataset for instruction-guided image editing.Advances in Neural Information Processing Systems, 36:31428–31449
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ea7ae9cb-6703-4f27-946a-d11d66c078b0 · outbound
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision Diffusion Transformers with Representation Autoencoders
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 34c57d94-5c67-40a6-93a0-01c71264e2b1 · inbound
STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Steering Visual Generation in Unified Multimodal Models with Understanding Supervision
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.