Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2506.18898.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T21:41:11.990825Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T10:09:44.750475Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation c3a0dc13-a72d-4be2-8da8-a5714a44191b · inbound
TBAC-UniImage: Unified Understanding and Generation by Ladder-Side Diffusion Tuning Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8379ad9-609e-4ff2-9c11-4af3053796be · inbound
Bridging Information Asymmetry: A Hierarchical Framework for Deterministic Blind Face Restoration Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16d3aa9f-6f7b-4885-a065-99e9b5a2a79f · inbound
InfoTok: Information-Theoretic Regularization for Capacity-Constrained Shared Visual Tokenization in Unified MLLMs Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d6f549d6-59d9-4b18-bb3e-de99e1335bb5 · inbound
Generative Refinement Networks for Visual Synthesis Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation bcb7a2aa-fdb1-4bb5-a13c-454254f918d4 · inbound
Generative Refinement Networks for Visual Synthesis Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f09bd54-8d22-4e48-9d5f-b6c7e89f3d47 · inbound
Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f9d5a11f-5ae7-40a6-b4ca-ea818e4428df · inbound
Tuna-2: Pixel Embeddings Beat Vision Encoders for Multimodal Understanding and Generation Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d48b7176-f7d3-438e-aec4-f7d2d4da5b79 · inbound
Tuna-2: Pixel Embeddings Beat Vision Encoders for Multimodal Understanding and Generation Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 15009bbc-4c51-493f-ad6c-cdb223faa208 · inbound
Beyond Accuracy: Benchmarking Cross-Task Consistency in Unified Multimodal Models Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 8c3754db-dd38-441f-b252-095f165ad71f · inbound
Visual Generation in the New Era: An Evolution from Atomic Mapping to Agentic World Modeling Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d5a69bff-ca17-448e-9ae3-f6bc8539c31e · inbound
MUSE: Resolving Manifold Misalignment in Visual Tokenization via Topological Orthogonality Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations
Reference 99
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 3687ca55-3ff9-4e3f-a0bc-71af85e59c14 · inbound
Residual Decoder Adapter: ID-Preserving Tokenizer Adaption for Autoregressive Text Rendering Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 941b87cd-4ce3-4a21-ae07-4551c6228197 · inbound
ARM: An AutoRegressive Large Multimodal Model with Unified Discrete Representations Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e09ca6db-7932-424b-9fbb-11e66550a125 · inbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations
Reference 208
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 3af3830b-45cc-44c3-9d49-aa06204b0ded · inbound
SPAR: Semantic-Pixel Self-Alignment and Adaptive Routing for Unified Multimodal Models Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 68c228fb-b342-4ecd-a447-7130a6b32ea9 · inbound
SPAR: Semantic-Pixel Self-Alignment and Adaptive Routing for Unified Multimodal Models Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6c72ffd8-1e11-45b2-ab65-c76324a081ef · inbound
Bridging Video Understanding and Generation in a Unified Framework Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 06e61899-6925-4113-a9e1-dfd30d2ed4e2 · inbound
dRAE: Representation Autoencoder with Hyper-Spherical Codes Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5376db81-4679-4749-aa0f-5ca47ef27479 · inbound
Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.