Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 28 inbound Pith citation observations for arXiv:2307.08041.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-12T12:37:38.362436Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
10
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation d3fe3931-6c77-4fb0-ae5c-027a8e27dca6 · inbound
SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension Planting a SEED of Vision in Large Language Model
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 26d03b63-a1bc-4499-97de-22bfc5ebf43f · inbound
DragNUWA: Fine-grained Control in Video Generation by Integrating Text, Image, and Trajectory Planting a SEED of Vision in Large Language Model
Reference 258
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation d8eff912-36a9-4665-8a80-3d3ee5cdaeb6 · inbound
Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models Planting a SEED of Vision in Large Language Model
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 8bc9297e-4349-43c7-891b-fde40426852c · inbound
SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation Planting a SEED of Vision in Large Language Model
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation b9fe7687-ffd6-4a6c-b006-a12f68c18305 · inbound
Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs Planting a SEED of Vision in Large Language Model
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation e61c6e61-1175-439f-9f01-4d34cc6d17e3 · inbound
VILA-U: a Unified Foundation Model Integrating Visual Understanding and Generation Planting a SEED of Vision in Large Language Model
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 66db182b-a9ef-4077-b2f4-0460af504fe6 · inbound
Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation Planting a SEED of Vision in Large Language Model
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation de316236-edad-4bf0-86f1-c444ee35cb11 · inbound
MUSE-VL: Modeling Unified VLM through Semantic Discrete Encoding Planting a SEED of Vision in Large Language Model
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9154674f-9add-4bb6-9b97-678903d4b16d · inbound
Liquid: Language Models are Scalable and Unified Multi-modal Generators Planting a SEED of Vision in Large Language Model
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61e1d6ab-c3e5-4ca6-b4f9-94f8caa32e1e · inbound
Divot: Diffusion Powers Video Tokenizer for Comprehension and Generation Planting a SEED of Vision in Large Language Model
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ba20b08-59d5-4ff9-ae81-01a4de02db27 · inbound
EgoPlan-Bench2: A Benchmark for Multimodal Large Language Model Planning in Real-World Scenarios Planting a SEED of Vision in Large Language Model
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c83293ac-b572-44e4-a03e-2a2e4eec0f1a · inbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Planting a SEED of Vision in Large Language Model
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef9d18dc-8bc5-4634-bde8-aa445cfd9309 · inbound
MetaMorph: Multimodal Understanding and Generation via Instruction Tuning Planting a SEED of Vision in Large Language Model
Reference 216
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 8662debd-c011-4ec7-bd3e-628f45358232 · inbound
Next Token Prediction Towards Multimodal Intelligence: A Comprehensive Survey Planting a SEED of Vision in Large Language Model
Reference 131
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1cf65f1b-7790-4aa7-8ee8-aa3128b26611 · inbound
Dual Diffusion for Unified Image Generation and Understanding Planting a SEED of Vision in Large Language Model
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cef15181-dc25-4773-93a6-0df94599faa0 · inbound
HealthGPT: A Medical Large Vision-Language Model for Unifying Comprehension and Generation via Heterogeneous Knowledge Adaptation Planting a SEED of Vision in Large Language Model
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f3c46e5-a305-4c8f-8f23-2cdc7a32c81f · inbound
Transfer between Modalities with MetaQueries Planting a SEED of Vision in Large Language Model
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 3d2fe74f-9855-4096-83cc-2da854771b37 · inbound
Mogao: An Omni Foundation Model for Interleaved Multi-Modal Generation Planting a SEED of Vision in Large Language Model
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 9121c834-8d28-4554-bafa-d53507f3dc9c · inbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Planting a SEED of Vision in Large Language Model
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a314e9b2-206b-4574-adc3-c2b4999361bd · inbound
From Standalone LLMs to Integrated Intelligence: A Survey of Compound Al Systems Planting a SEED of Vision in Large Language Model
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 5244ed3a-2046-4a49-a4ee-77feb14e17cc · inbound
Instella-T2I: Pushing the Limits of 1D Discrete Latent Space Image Generation Planting a SEED of Vision in Large Language Model
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff4062fc-f51c-4b6a-8b42-aeb20bd31974 · inbound
DisCo: Towards Distinct and Coherent Visual Encapsulation in Video MLLMs Planting a SEED of Vision in Large Language Model
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81ff7b25-87d0-4a82-9adf-ca1f202ec84a · inbound
Colon-Bench: An Agentic Workflow for Scalable Dense Lesion Annotation in Full-Procedure Colonoscopy Videos Planting a SEED of Vision in Large Language Model
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c277ecd-dbf2-4959-a401-7b8862c8793e · inbound
DeepSight: Long-Horizon World Modeling via Latent States Prediction for End-to-End Autonomous Driving Planting a SEED of Vision in Large Language Model
Reference 206
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation adfc6a0b-763b-49c7-8317-ce7147a39521 · inbound
Learning to See What You Need: Gaze Attention for Multimodal Large Language Models Planting a SEED of Vision in Large Language Model
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 24696ebf-5e9e-48c5-9253-40d51874cfd6 · inbound
When Recovery Matters: The Blind Spot of Surrogate Privacy in MLLM Editing Planting a SEED of Vision in Large Language Model
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 4df426a2-fa2a-467a-9f90-1df208e42540 · inbound
Towards Physics of Multimodal Pretraining: Knowledge Flow, Modality Synergy, Early Unification, and Recipes Planting a SEED of Vision in Large Language Model
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c015a76d-6c53-43ef-b39f-fca19bfc22c3 · inbound
Towards Physics of Multimodal Pretraining: Knowledge Flow, Modality Synergy, Early Unification, and Recipes Planting a SEED of Vision in Large Language Model
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.