Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-27T07:01:07.362430Z
Paper Citation Record · LEDGER
As of 12 August 2026, this Paper Citation Record lists 100 of 281 outbound references and 2 inbound Pith citation observations for arXiv:2606.13289.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-27T07:01:07.362430Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-08T17:08:56.912092Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-06T11:55:28.362396Z
100 of 281 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 38bdd4f4-4ae6-4cbc-8a8e-6a8b1096aaa4 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation a99eaeac-f037-4288-b968-bc0c8fbd28f7 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers International conference on machine learning , pages=
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02e5cf2a-28bb-488d-ad08-6f83a9478409 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers 2023 , url=
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cdf36153-0d10-429d-b56b-7db745fd1a9e · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1914e40-d6c6-407d-9df5-c0e2b7c05c9a · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers OpenAI technical report , year=
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08509cb8-d874-47df-9501-1200f6c3ecf7 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 7d0ad57d-079d-4584-8e35-429289c7a6d9 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Data Filtering Networks
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 53ff37cc-e9a1-4912-9182-ffa657609327 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Proceedings of the IEEE/CVF international conference on computer vision , pages=
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc11bbb2-f9a0-4fad-b741-701085296d1e · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers BEiT: BERT Pre-Training of Image Transformers
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ff37c413-f814-4513-942d-bfb1945f2d36 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Proceedings of the IEEE/CVF international conference on computer vision , pages=
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78317259-bb70-4afc-9c69-fdc911f8532c · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers From Big to Small: Multi-Scale Local Planar Guidance for Monocular Depth Estimation
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 0d7bb7a9-fab0-4672-9fac-46779c9773a6 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edc35e0a-9445-4d58-bf6f-9bc4420c5c4b · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Benchmarking Detection Transfer Learning with Vision Transformers
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 2eebba51-a766-4f17-a24b-dc3f1061cae3 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Proceedings of the IEEE international conference on computer vision , pages=
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 381f1b43-cee1-4214-9fbd-71cc8c8626bd · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers ArXiv , year=
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e653307c-698c-489a-973d-bb1ded75a413 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers ArXiv , year=
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1efcd99-a458-4e1a-9a66-5ad02414ec2d · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers 2024 , eprint=
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41549d01-da07-44bf-8fbf-32f8c137c637 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Repa-e: Unlocking vae for end-to-end tuning with latent diffusion transformers
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 03c4caf3-515c-4b02-b175-1eaa7394f102 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Boosting latent diffusion models via disentangled representation alignment
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9d3abfb3-0d72-45f6-a516-78a74ab5dfc3 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Towards scalable pre-training of visual tokenizers for generation
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 3636ed34-81fc-46cd-be6e-e88d4a862b57 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Wan: Open and Advanced Large-Scale Video Generative Models
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f7d6ba4f-720c-45ab-9dfc-d85a1f4922e2 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Diffusion Transformers with Representation Autoencoders
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 582b5745-87fd-47e3-9de4-c8010932e8a3 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers The prism hypothesis: Harmonizing semantic and pixel representations via unified autoencoding
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 38ef4a00-688e-44c8-8d41-719bbf950efe · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8dab6d55-943d-4737-8d4b-44524abb0d83 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e115c19a-7ec6-487c-a166-d64469dd806d · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Proceedings of the IEEE/CVF international conference on computer vision , pages=
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4abdca51-3c1d-4f56-bf19-faf3386e173c · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Advances in Neural Information Processing Systems , volume=
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2338cc16-d921-461a-a725-387847b30ac5 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers ArXiv , year=
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be118554-1f50-4082-89ae-df5a631b0461 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers HunyuanImage 3.0 Technical Report
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation b7f1b6ff-2dcc-4560-aacb-430297665062 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Latent diffusion model without variational autoencoder
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d0b9e9da-8cd3-41c8-bef1-7dc9fa1b14e2 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Vector-quantized Image Modeling with Improved VQGAN
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c5f8e4c3-3d5f-4a0d-a217-7affa712167b · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Diffusion Autoencoders are Scalable Image Tokenizers
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 73b856cd-2489-4a98-a6b8-e50d9e6ec436 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers generation: Taming optimization dilemma in latent diffusion models , author=
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 294fda3b-4030-4564-84ef-2f2d40ecfee9 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb04f536-e5c6-4bcc-8909-919f7e7233aa · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Advances in Neural Information Processing Systems , volume=
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b47c3c07-ba25-4c7d-a6cb-35c54c83dfce · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers arXiv preprint arXiv:2507.08441 , year=
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 23721df6-51e7-48d1-9b31-dcc526261249 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Advances in neural information processing systems , volume=
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1929781b-9882-4d01-8de5-f2999afddaa6 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Advances in Neural Information Processing Systems , volume=
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7787fef2-8047-4ad3-a8ac-e50abd4f1246 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Latent denoising makes good visual tokenizers
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 49bd37e6-f4be-4269-b367-bccfadf8462f · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Advances in neural information processing systems , volume=
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f71baa52-ab77-49c2-b5b6-9e760ef08680 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Foundations and Trends
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1111e54b-6dcb-417e-b0fd-071e82cd40d2 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Image Processing Algorithms and Techniques II , volume=
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54209dcf-d4c5-420a-9e2a-0b3aa6947666 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Advances in Neural Information Processing Systems , volume=
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd2a13e7-4617-4f89-b924-0d3b765139af · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Open-MAGVIT2: An Open-Source Project Toward Democratizing Auto-regressive Visual Generation
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e9d36f68-0d8f-4812-9bec-57ab91a2194f · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 613719cd-5b2a-4053-9b77-3a7047568675 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers International Conference on Learning Representations , year=
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ecddbb8c-3993-45c6-a945-3b9b2ab7f84d · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Proceedings of the Computer Vision and Pattern Recognition Conference , pages=
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3b60c85-bd9d-43fc-9357-5d022a1a207b · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers ImageFolder: Autoregressive Image Generation with Folded Tokens
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c41b4ef4-93ed-4183-a671-cb9380c5377e · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Forty-second International Conference on Machine Learning , year=
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a93c7eb2-26e8-4d8a-9c5a-d0af8fbf165e · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers 2022 , eprint=
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c453ecf6-ccf9-48e1-a982-2321c425de44 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 092adf5e-e897-4641-8437-ffafa9b149cf · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Chameleon: Mixed-Modal Early-Fusion Foundation Models
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e6214d90-8861-4a4c-a0f4-2c1b5bfc5f4c · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers 2025 , eprint=
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd51b75a-c007-4e53-b58b-dd6166d1fa01 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers DreamLLM: Synergistic Multimodal Comprehension and Creation
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 68408783-f0e0-41fd-9149-ca83de540ac1 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Unified language-vision pretraining in llm with dynamic discrete visual tokenization
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9254e3b3-1c2e-4e2b-a23d-dc6f3da66391 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Proceedings of the Computer Vision and Pattern Recognition Conference , pages=
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1ce572f-4dd1-41f7-8b27-b0e4f573064b · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers arXiv preprint , year=
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7805d954-f801-4294-a566-f745f9568dd5 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Making LLaMA SEE and Draw with SEED Tokenizer
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 0c6aa083-e3d5-41ec-a214-a2276619f066 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Advances in Neural Information Processing Systems , volume=
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a6b70da-000d-4b30-934e-013547a0b354 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Emu3: Next-Token Prediction is All You Need
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 65458b22-33de-4bb8-850e-92fe0ca81b9a · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 977fd9ec-6dd2-431b-a9f4-f3480148eb34 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Proceedings of the IEEE/CVF International Conference on Computer Vision , year=
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3383bfe0-8243-4391-98a0-1699edfd775d · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers ILLUME: Illuminating Your LLMs to See, Draw, and Self-Enhance
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 0f0ae40e-923c-4008-b8ac-c0f24f7e5690 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Orthus: Autoregressive Interleaved Image-Text Generation with Modality-Specific Heads
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f02cbc2a-3ecd-4a24-9410-4721b4f3ccfc · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Proceedings of the Computer Vision and Pattern Recognition Conference , pages=
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08882222-8ee0-42b4-ae82-e8e2c6eb3999 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e32b757-65dd-4f02-a9ef-236b15d8ab85 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers 2023 , eprint=
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b7811ee-fe9d-4474-bfaa-5edb250b0eef · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers 2025 , eprint=
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d93c88c-7d14-4a5d-ae65-d42304fbf072 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers QLIP: Text-Aligned Visual Tokenization Unifies Auto-Regressive Multimodal Understanding and Generation
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 56505dc3-693d-40eb-bc29-692615926173 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers OpenUni: A Simple Baseline for Unified Multimodal Understanding and Generation
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 0c738ec6-7764-4a1d-a5fd-22cfe2530c27 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Liquid: Language Models are Scalable and Unified Multi-modal Generators
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 0cebaca9-115c-4a46-9df5-a92158d2c38a · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Unitok: A unified tokenizer for visual generation and understanding.arXiv preprint arXiv:2502.20321, 2025a
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f6597786-6122-437f-9948-ce371a6ddebf · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Unilip: Adapting clip for unified multimodal understanding, generation and editing
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 0677f3a0-72b9-48e3-9c23-e0e8fa80aea9 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 750863bb-62c3-4cdf-815e-4f3281272cf1 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers ILLUME+: Illuminating Unified MLLM with Dual Visual Tokenization and Diffusion Refinement
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 802d7e2d-d8dd-44e7-b3de-2ddda65cc6c1 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers XQ-GAN: An Open-source Image Tokenization Framework for Autoregressive Generation
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 7eb61c66-8c0d-41de-8f2e-6b6208abac1b · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Finite Scalar Quantization: VQ-VAE Made Simple
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 2f200e47-713e-4aa5-beb2-522225c5c174 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Robust Latent Matters: Boosting Image Generation with Sampling Error Synthesis
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 2702eb36-e7e9-45bf-a32f-d60ad9f63b47 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Proceedings of the IEEE/CVF International Conference on Computer Vision , year=
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4cb9248f-9406-4295-ad51-3e7438e6479f · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Selftok: Discrete Visual Tokens of Autoregression, by Diffusion, and for Reasoning
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 17636725-4e5a-4026-b180-1e1a1c2987e4 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Forty-second International Conference on Machine Learning , year=
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2f6432a-7602-4fae-ae8c-d99b66e7788e · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1be63069-928f-4b15-8b98-f36a20dfec22 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Proceedings of the IEEE conference on computer vision and pattern recognition , pages=
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84b5e0b9-3509-4e9d-bca9-9d2a67a885eb · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers 2025 , url=
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7221f5a7-2264-4f00-8acf-956dbc3cc119 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0032e54-cc03-4315-9e48-79806020c9e8 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13c1b387-1eb9-41e9-b899-4dab3c353bb0 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers European conference on computer vision , pages=
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c50e2deb-5af4-4a25-a40d-015970efef42 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Evaluating Object Hallucination in Large Vision-Language Models
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation cd070a70-ef08-4dfa-924b-ca50fadc5394 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers nature , volume=
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2bb8a214-b480-438d-8ad8-f116101d8372 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Proceedings of the European conference on computer vision (ECCV) , pages=
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afc4c253-6112-4d28-911f-b7793737ffa0 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Proceedings of the Computer Vision and Pattern Recognition Conference , pages=
Reference 94
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc4ef71f-64a8-4405-8b4f-590c9b026413 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers European conference on computer vision , pages=
Reference 96
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1e95262-2d95-41aa-a8c3-8d0ace513f77 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers International conference on machine learning , pages=
Reference 97
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e683a672-62fb-4d26-b850-ea2426fbc961 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers CoCa: Contrastive Captioners are Image-Text Foundation Models
Reference 98
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 64630893-3eba-400a-ba49-bb6c866a70ea · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Internvideo-next: Towards general video foundation models without video-text supervision
Reference 99
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 36fac3a4-3f89-4464-afef-b3b4c1f6c9bc · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers DINOv2: Learning Robust Visual Features without Supervision
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 24e8eaff-199f-4922-a142-f9ee555bd31b · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Advances in Neural Information Processing Systems , volume=
Reference 103
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd201d84-672c-4b9e-bd3f-8edb0fb4ad87 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Proceedings of the IEEE/CVF international conference on computer vision , pages=
Reference 104
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6948834e-a774-4b49-8d4d-7ada40ba6c8b · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers DualToken: Towards Unifying Visual Understanding and Generation with Dual Visual Vocabularies
Reference 105
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 470607fd-e8a9-4ccc-bfa1-82f4ba8a0c71 · outbound
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers arXiv preprint arXiv:2503.06764 (2025) 4, 7, 9, 10, 1
Reference 106
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 4a166ade-3560-4c34-9537-15b3524dff4e · inbound
Towards Physics of Multimodal Pretraining: Knowledge Flow, Modality Synergy, Early Unification, and Recipes HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers
Reference 154
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation a2edc933-347d-4858-8181-67aadbfbbd76 · inbound
Towards Physics of Multimodal Pretraining: Knowledge Flow, Modality Synergy, Early Unification, and Recipes HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers
Reference 154
Source-reported events for the cited work
Unavailable: canonical work link unavailable.