Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T22:27:13.689128Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 1 inbound Pith citation observation for arXiv:2501.01709.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T22:27:13.689128Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T23:57:21.526961Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-06T23:57:29.885926Z
53 of 53 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation cf6a0da2-a87d-43cc-8679-884247b53839 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72bc8d5d-2859-488b-9811-fdd6089aacd4 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Flamingo: a visual language model for few-shot learning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 394a6094-ea0f-44c6-aa33-56a8d98c6341 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Does Combining Parameter-efficient Modules Improve Few-shot Transfer Accuracy?
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29eadd2a-b321-4116-aaba-f565bcd3dc3b · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Qwen Technical Report
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1af4069-2089-434c-a895-c190c413f8ee · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2687125c-24e1-48de-ab2e-96b715d75515 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders DeepSeek LLM: Scaling Open-Source Language Models with Longtermism
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e50f52a4-ca56-4c32-a693-2cd55c8f7e59 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2fa7aa1-aa03-43d2-87b6-742b494381e7 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4034b046-be82-4ccd-ac3f-6a160a373862 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Vision transformers need registers
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 8aec9829-59fc-488c-b5c1-76ea0e4368cd · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Eva-02: A visual representation for neon genesis
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65d36236-5e96-4cbd-9627-66da1a147166 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 578d87cc-5f4c-4c24-8401-dedc8188ec17 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Dat- acomp: In search of the next generation of multimodal datasets
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation daca6a56-0b16-47e8-b833-314ee6d83281 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Making the v in vqa matter: Elevating the role of image understanding in visual question answer- ing
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bba2ffc5-ee89-464a-9fbc-86b4d9ca3449 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Vizwiz grand challenge: Answering visual questions from blind people
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 488b43f6-0e74-4076-bee2-cd45e28ef5b9 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Distilling the Knowledge in a Neural Network
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5271c9cd-2ae0-461f-85cc-35d4c849efe7 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders LoRA: Low-Rank Adaptation of Large Language Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5f3b0c7-f067-4180-a8e8-8e2f027fb0d0 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Masked Distillation with Receptive Tokens
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f62af99d-6761-4c4f-8e45-5d29316f57ed · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders GQA: A new dataset for real-world visual reasoning and compositional question answering
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 87fc07ba-5945-4995-863d-0ebb0777647d · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Adaptive mixtures of local experts.Neu- ral computation, 3(1):79–87, 1991
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6154a4ab-ccde-4993-8d97-e860da380d1c · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Lift3d foun- dation policy: Lifting 2d large-scale pretrained models for robust 3d robotic manipulation, 2024
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation abd0a356-d95e-490c-9744-21e05817f992 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Segment any- thing
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c29da05-1f08-4f24-a27b-e8095cba03dd · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Evaluating Object Hallucination in Large Vision-Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb110152-8f13-4916-a61a-3729b7c2f7d1 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Uni-MoE: Scaling Unified Multimodal LLMs with Mixture of Experts
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 101b2808-9905-4246-8aaf-4da0ed31e7c7 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2d3d16b-74e6-4b1f-8485-73d85e27d9b2 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders MoE-LLaVA: Mixture of Experts for Large Vision-Language Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b8063a9-da3c-4c10-9bca-b042bab36ae5 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Draw-and-Understand: Leveraging Visual Prompts to Enable MLLMs to Comprehend What You Want
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f68b4fa-2d80-436e-85a4-0f58ef3f9bdc · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Improved Baselines with Visual Instruction Tuning
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2974143e-4c3f-48c8-b346-855f5fb5070f · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Llava-next: Im- proved reasoning, ocr, and world knowledge, 2024
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50f0df7c-440d-4b04-9783-e3aab29c2115 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Visual instruction tuning
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 7eb9eb38-33d2-4e57-b0f5-ef1c3e9b7ff7 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders MMBench: Is Your Multi-modal Model an All-around Player?
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cbafa25b-627d-46b0-85b8-438fd0be0cda · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders A convnet for the 2020s
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 205b5671-f1d1-48e7-9646-48d2498a0787 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Learn to explain: Multimodal reasoning via thought chains for science question answering
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80498a9c-ada8-4563-95bd-d9a85ce414c6 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Llm as dataset ana- lyst: Subpopulation structure discovery with large language model
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 2607c897-9f4b-40d7-b82c-a561a42ab248 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders DINOv2: Learning Robust Visual Features without Supervision
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12b5d433-b490-4924-927a-714bfe6a51f1 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Learning transferable visual models from natural language supervi- sion
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation fdfdba85-9aec-4478-9b2c-c5f8a7ad38ee · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Am-radio: Agglomerative vision foundation model reduce all domains into one
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation b8b51a70-6c67-4186-bf91-29f9eda3a3ca · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders When do we not need larger vision models? In European Conference on Computer Vision, pages 444–462
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation cea5852e-cdf8-4cd4-90d7-0adf3091d588 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Eagle: Exploring The Design Space for Multimodal LLMs with Mixture of Encoders
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1fb87df2-bf6a-4d24-86d2-b06f143106d1 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders LLaVA-MoD: Making LLaVA Tiny via MoE Knowledge Distillation
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation feab3b34-00d0-4463-af1b-f2373b6e3400 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Towards VQA models that can read
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 537f842c-2f51-4b1f-a934-71619aa3f266 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Gemini: A Family of Highly Capable Multimodal Models
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06dddca7-3899-4452-a179-9dfe1bb073dd · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc810408-1203-4a73-b879-54a5a7f7a875 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders LLaMA: Open and Efficient Foundation Language Models
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6fd74cd9-03fe-46e4-b28b-161e108eac19 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Contrastive Learning Rivals Masked Image Modeling in Fine-tuning via Feature Distillation
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f461571d-9743-41b5-a747-724556a9ebd8 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Beyond full fine-tuning: Harnessing the power of lora for multi-task instruction tuning
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation aeb84c7c-b67b-4163-8095-7d3114642c16 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders One Student Knows All Experts Know: From Sparse to Dense
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 360e5dd2-430c-497c-b871-c338e1d837f6 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Baichuan 2: Open Large-scale Language Models
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 739e32ef-aa33-4050-b254-3065d0eade26 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders mPLUG-Owl: Modularization Empowers Large Language Models with Multimodality
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70e96a49-6bc9-4699-b345-d73b97b094ee · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Learn- ing from multiple teacher networks
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 99d02520-c2a3-488b-b4a0-cf37515833be · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Yi: Open Foundation Models by 01.AI
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44a85115-9640-4250-aeab-06739e43206d · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders SparseVLM: Visual Token Sparsification for Efficient Vision-Language Model Inference
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df1d01a3-4cf9-421d-9494-3812d1d04ec7 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Beyond LLaVA-HD: Diving into High-Resolution Large Multimodal Models
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5c23181-1923-488d-a002-443d8b833862 · outbound
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders Judging llm-as-a-judge with mt-bench and chatbot arena.Advances in Neural Information Processing Systems, 36:46595–46623, 2023
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 67cb66a0-1f7c-4564-97a1-625fa7ed7c47 · inbound
GenRecal: Generation after Recalibration from Large to Small Vision-Language Models MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.