Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T06:04:22.130258Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 79 of 79 outbound references and 7 inbound Pith citation observations for arXiv:2411.19628.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T06:04:22.130258Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T11:46:40.764054Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T11:28:04.162152Z
79 of 79 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation a7fa0714-7f7f-4c6c-8c14-ea1ed9f12fcb · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e933751-9f44-416f-b4dd-be77566ed270 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Qwen Technical Report
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20d4be6e-e68a-4069-9c1c-db8d9afd5f8a · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9bd0c36-ec62-4fcc-9594-8c8cd5dcc18f · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Qwen2.5-VL Technical Report
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1437d00c-11ba-4f2d-a941-c1dcc9780652 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Token Merging: Your ViT But Faster
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8842a2e0-28a1-4ecf-9c35-a9257a9fd0ee · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Language Models are Few-Shot Learners
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f5753ea-e6f4-4547-a809-2f18a63c6689 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings An Image is Worth 1/2 Tokens After Layer 2: Plug-and-Play Inference Acceleration for Large Vision-Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f74320d-d2ef-475c-a61f-065ec70a59bb · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Diffrate: Differentiable compression rate for efficient vision transformers
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 30eab4de-4038-458b-a653-c6cc1e24f9bd · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings EE-LLM: Large-Scale Training and Inference of Early-Exit Large Language Models with 3D Parallelism
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9cb2905a-ccdc-41ca-bbe7-2fabeec38825 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b6f44f4-dbc4-42b1-aa8e-5703f5d01691 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6949731c-6b21-4127-a66a-fe773eb2b603 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5272414e-490f-4617-8024-66431c39a580 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Heatvit: Hardware-efficient adaptive token pruning for vision transformers
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 5b6809dc-15dc-453c-aa96-308612bf2668 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings InternLM-XComposer2: Mastering Free-form Text-Image Composition and Comprehension in Vision-Language Large Model
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68666eef-9be2-41c9-b6f6-7c7402ca0b14 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Internlm-xcomposer2-4khd: A pioneering large vision- language model handling resolutions from 336 pixels to 4k hd
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 40a00deb-fdca-48a4-a7fa-e33528c8e097 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings The Llama 3 Herd of Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 424c65b5-9b19-45bd-820c-d9a27478b5b0 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Not All Layers of LLMs Are Necessary During Inference
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 478adf23-ba37-49e2-8642-a7fcb4620345 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Adaptive token sampling for efficient vision transformers
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 1e08c585-456a-44b4-aee9-b6f8dea135d2 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Deecap: Dynamic early exiting for efficient image captioning
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation df016c43-767a-4315-a598-5f8403efb6a4 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Mme: A comprehensive evaluation benchmark for multimodal large language models
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8940cc22-bc4e-452e-a394-25ff7c716919 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Frameexit: Conditional early exiting for efficient video recognition
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a56632bd-b690-4206-aae0-249a999b5b41 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Making the v in vqa matter: Elevating the role of image understanding in visual question answering
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a9aab22-7e9c-4f66-a838-41f9c525b287 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Gqa: A new dataset for real-world visual reasoning and compositional question answering
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 0c679286-540e-463f-8042-5b10796a3c58 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Categorical Reparameterization with Gumbel-Softmax
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8697c61-d8d2-4148-b85b-8079e5afc944 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Mixtral of Experts
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34b38d83-6c62-4c96-b68a-760f1e5e1bee · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings What Kind of Visual Tokens Do We Need? Training-free Visual Token Pruning for Multi-modal Large Language Models from the Perspective of Graph
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b0dc344-5387-49c1-9d37-7bee09ec74ed · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Spvit: Enabling faster vision transformers via latency-aware soft token pruning
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 5211d79a-491c-4951-9946-bb5dd6b3947d · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings On information and sufficiency
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ab43473-6cb9-4410-80f8-f00ce523c587 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings LLaVA-OneVision: Easy Visual Task Transfer
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b3189d7-0583-4edf-9937-d7f94e36997c · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Seed-bench: Benchmarking multimodal llms with generative comprehension
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 105a0086-71cf-4f37-a1c3-291c38ce80bd · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5a7df4a-714f-4097-893a-1e71631a0529 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings TokenPacker: Efficient Visual Projector for Multimodal LLM
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 600d70d4-c2c0-4ec5-af37-6af14ea74c78 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Evaluating object hallucination in large vision-language models
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 88c28a23-9dac-4efe-94ec-f9420798bfa3 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings MoE-LLaVA: Mixture of Experts for Large Vision-Language Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43faf140-3256-4be7-8d68-3f5a65ec93d5 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings VILA: On Pre-training for Visual Language Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a26d361b-458d-46ad-8683-25451e33e88f · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Boosting Multimodal Large Language Models with Visual Tokens Withdrawal for Rapid Inference
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 650dff22-6f06-47c2-8247-6d95ecf79a50 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Looking Beyond The Top-1: Transformers Determine Top Tokens In Order
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0da02878-3dfa-4a2a-bde0-2093aeff14e8 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Llava-next: Improved reasoning, ocr, and world knowledge, January 2024
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3ce24cf-b297-4372-9d14-6d3835a357de · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Visual instruction tuning.Advances in neural information processing systems, 36, 2024
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a0da20e-0e42-4499-8a25-64c5e48374d4 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Mmbench: Is your multi-modal model an all-around player? Computing Research Repository (CoRR), 2023
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation caf056e2-52c7-46b2-a82c-646e45cc0ffa · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Learn to explain: Multimodal reasoning via thought chains for science question answering
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 90d7aa42-0a91-4048-bd20-d1b3611c86cf · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Feast Your Eyes: Mixture-of-Resolution Adaptation for Multimodal Large Language Models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c93317a-d433-433c-b558-e12423e4ee24 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Infographicvqa
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0147851e-7164-4655-a069-d35197dbf8c3 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Docvqa: A dataset for vqa on document images
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d4e526e-5e47-4348-acca-dfa99769d574 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4739534a-6456-4b2d-b560-a050ad465524 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Llava-prumerge: Adaptive token reduction for efficient large multimodal models
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 520bca3a-9828-4f0a-8aa0-be5507544d76 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Eagle: Exploring The Design Space for Multimodal LLMs with Mixture of Encoders
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff6630a0-aae5-4100-82b5-96e4e0137e79 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Towards vqa models that can read
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8033177-2b91-49e7-918c-f0cce9e83620 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings A Mechanistic Interpretation of Arithmetic Reasoning in Language Models using Causal Mediation Analysis
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9985e08-4041-479c-87a0-897586267db7 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings You need multiple exiting: Dynamic early exiting for accelerating unified vision language model
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08cf5ee7-7419-4783-ae7d-c1f6fab53703 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Qwen2.5: A party of foundation models, September 2024
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79c47bf7-ee05-4909-b592-60bfe0b79867 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings LLaMA: Open and Efficient Foundation Language Models
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d04d3bfc-a572-4ab3-b771-538ee0dbf1fb · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b23538f6-64d4-4fc2-831f-f564543689e8 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Zero-tprune: Zero-shot token pruning through leveraging of the attention graph in pre-trained transformers
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 23960389-7de0-4cbe-a5aa-d2090aa7c592 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f9f16fc-2bce-44a8-948c-969f0bdebf39 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Joint token pruning and squeezing towards more aggressive compression of vision transformers
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8a20ff2a-cece-4070-bc03-06edff368251 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Routing Experts: Learning to Route Dynamic Experts in Multi-modal Large Language Models
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75827c6d-13a3-48fa-b147-4e8e45f3a131 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Parameter and computation efficient transfer learning for vision-language pre-trained models
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 6b3c2619-29df-49d0-a8f8-2f8fefc4e496 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings PyramidDrop: Accelerating Your Large Vision-Language Models via Pyramid Visual Redundancy Reduction
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 393329a3-b647-4b1a-8236-92e1c3c38c42 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Fit and Prune: Fast and Training-free Visual Token Pruning for Multi-modal Large Language Models
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e12da4e-5ad2-4e0a-8d5b-c84464ae0f64 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Mm-vet: Evaluating large multimodal models for integrated capabilities
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 2cef6003-40d7-47cf-9eae-088364b11489 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Llava-mini: Efficient image and video large multimodal models with one vision token
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 0bba889d-0024-422e-a317-4e22f0bcf10c · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings TinyLLaVA: A Framework of Small-scale Large Multimodal Models
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d7b68f4-2d66-4079-b56b-29c5bcf58245 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Guidelines: • The answer NA means that the abstract and introduction do not include the claims made in the paper
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 9a82735b-f8ee-485f-9835-1e7060909915 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Limitations
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation b6d86275-0ed9-4426-8559-654e80b2ffaf · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Guidelines: • The answer NA means that the paper does not include theoretical results
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c62e8254-8b34-46a5-a549-397f140a401f · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Guidelines: • The answer NA means that the paper does not include experiments
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 1e40ba61-aa22-43df-8c85-bd997ae91d3b · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Guidelines: • The answer NA means that paper does not include experiments requiring code
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 221209bd-72ff-40bb-be5d-155ee7c16017 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Guidelines: • The answer NA means that the paper does not include experiments
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 9f01c4e8-b7a7-43a1-b1c5-bbe616b58a1c · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Guidelines: • The answer NA means that the paper does not include experiments
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 1cd81772-c7b1-43e2-8fe9-76077bd87df8 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Guidelines: • The answer NA means that the paper does not include experiments
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 6f92737b-a167-4fca-b32c-b8db77214605 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings 22 Guidelines: • The answer NA means that the authors have not reviewed the NeurIPS Code of Ethics
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 16c13f32-6630-4ecd-8e4e-307463c81beb · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Guidelines: • The answer NA means that there is no societal impact of the work performed
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 2124bd1f-02d6-46b6-9fc1-51a1bc1425ab · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Guidelines: • The answer NA means that the paper poses no such risks
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 145f4e33-8a3b-4749-bfbb-917199901370 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Guidelines: • The answer NA means that the paper does not use existing assets
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation f25ec9c9-58a6-4300-9235-aa723ca96499 · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Guidelines: • The answer NA means that the paper does not release new assets
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a1ac8fbb-c5e9-4718-9ed8-0b305978fb0b · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Guidelines: • The answer NA means that the paper does not involve crowdsourcing nor research with human subjects
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation f046147d-d26f-485f-b3a9-4ec0a8b82eeb · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Guidelines: • The answer NA means that the paper does not involve crowdsourcing nor research with human subjects
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 200a829f-cd32-4381-a14e-8127df4d31fc · outbound
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings Answer: [No] Justification: LLM, as a part of MLLM, is the object of our study
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8b26692e-6562-4150-969d-72e6e6bad9ad · inbound
Growing a Multi-head Twig via Distillation and Reinforcement Learning to Accelerate Large Vision-Language Models Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 0ed183cc-559e-4e27-a08d-8b033531194e · inbound
Video-MMLU: A Massive Multi-Discipline Lecture Understanding Benchmark Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings
Reference 134
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 673f7275-980c-40d1-98c7-f3809026eda4 · inbound
ForestPrune: High-ratio Visual Token Compression for Video Multimodal Large Language Models via Spatial-Temporal Forest Modeling Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e11b46e0-8a0b-400e-842e-a2919bcf1509 · inbound
Counting to Four is still a Chore for VLMs Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 86468620-b557-4ddc-86e4-0b67b7a43ae4 · inbound
Reroute, Don't Remove: Recoverable Visual Token Routing for Vision-Language Models Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 1d859195-f4e4-4553-8e8e-ade08e6c65ee · inbound
Seeing the End at Step Zero: Accelerating Diffusion MLLMs via MLP Sparsity-Aware Truncation Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 289f2c93-275c-42ca-99dd-849180301cb3 · inbound
Learning to Predict Middle-Layer Attention in MLLMs for Visual Token Prunin Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.