Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-03T03:36:43.892124Z
Paper Citation Record · LEDGER
As of 5 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 4 inbound Pith citation observations for arXiv:2602.07574.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-03T03:36:43.892124Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T18:16:46.738649Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-04T08:19:44.559138Z
36 of 36 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c941f91f-458d-424b-bfee-57c79342a20f · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66d283e1-b9c9-4aac-a3a6-5ebdbb55a070 · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention From llms to lrms: Rethinking pruning for reasoning-centric models.arXiv preprint arXiv:2601.18091,
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e581471b-eaaa-4f14-bc2b-e3565ed59700 · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention Visipruner: Decoding discontinuous cross-modal dynamics for efficient multimodal llms
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba9bde6f-dc17-470b-b9ca-27ae5a9e6181 · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention The framework tax: Disparities between inference effi- ciency in NLP research and deployment
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a00e5172-ce7a-4dab-a7b9-6aceeea84fb6 · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 725bcd9b-1c9d-45b9-85f1-f21bfefa4233 · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention Filter, correlate, com- press: Training-free token reduction for mllm accelera- tion.arXiv preprint arXiv:2411.17686,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6ac5bc9-df02-4fb8-b581-a2bcfa666296 · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention Training-free token pruning via zeroth-order gradient estimation in vision-language models.arXiv preprint arXiv:2509.24837,
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2698dd5f-f538-4696-9fc2-925bfeb145e0 · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention Crafting papers on machine learning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe6a8878-5f93-4c2b-a242-d594f641f58d · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention Unresolved cited work
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f79e5802-4c2a-456f-93e8-e2229b6e0f5a · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention Unresolved cited work
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 344c15a1-56d7-4455-85a7-9c432869e3f9 · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention The few govern the many: Unveiling few-layer dominance for time series models.arXiv preprint arXiv:2511.07237,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6a59e34-f22e-4c8d-9df7-618e24ce7a7d · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention Growing a Multi-head Twig via Distillation and Reinforcement Learning to Accelerate Large Vision-Language Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f6e8c13-3300-4526-90f7-e31565de98f1 · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention WeLM: A Well-Read Pre-trained Language Model for Chinese
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1fd97f3b-7b79-4986-bc5d-e941bd1876c4 · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention Unraveling the Mystery of Scaling Laws: Part I
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e46ced7-cd4e-4db9-b348-b7ac0ab0b99d · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention Unresolved cited work
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 422fd897-39a2-45c6-b90b-6c39a3d0f062 · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention TokenCarve: Information-Preserving Visual Token Compression in Multimodal Large Language Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d4548d6-a087-41d5-86cf-b33d5ca5d617 · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention Flowcut: Rethinking redundancy via information flow for efficient vision-language models.arXiv preprint arXiv:2505.19536,
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77fa743d-5dae-40b2-8c33-67135aaea41d · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention LLaMA: Open and Efficient Foundation Language Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4a79494-0980-49de-93da-1059d4310440 · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a379abfc-05fe-4aee-bb5f-0ce88744d63f · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 018b7785-79ee-4344-a710-5e45add4c883 · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention Stop Looking for Important Tokens in Multimodal Language Models: Duplication Matters More
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d0b4ba3-5918-4c10-af00-09dbfa938a6a · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention PyramidDrop: Accelerating Your Large Vision-Language Models via Pyramid Visual Redundancy Reduction
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 068ed5f8-001a-4c3a-b897-6ec3b1f2a7d8 · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention Shortv: Efficient multi- modal large language models by freezing visual tokens in ineffective layers.arXiv preprint arXiv:2504.00502,
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70140cca-3983-45bf-b5b3-eec4c05b5e23 · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention and Shukla, D
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e035451b-2ac5-4803-9ea6-bea64010fbf2 · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention Skip-Vision: Efficient and Scalable Acceleration of Vision-Language Models via Adaptive Token Skipping
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe600739-bc70-4b9b-b5f5-0881241d255b · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention Vscan: Rethinking visual token reduction for efficient large vision-language models.arXiv preprint arXiv:2505.22654, 2025a
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f8eb748-1179-4bab-a8f9-329a9567700c · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention SkipGPT: Dynamic Layer Pruning Reinvented with Token Awareness and Module Decoupling
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d029859-003d-4829-b021-a177564b018d · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention Don’t just chase” highlighted tokens” in mllms: Revisiting visual holistic context reten- tion.arXiv preprint arXiv:2510.02912,
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c9823b9-2a33-4884-8775-97d688063185 · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention Unresolved cited work
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eba19160-294e-465f-af7b-e273d149be6e · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention (b) Our method under eager attention, where visual tokens are frozen via explicit masking and participate in attention only as key–value representations
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e48dae1a-7083-4a06-9cc9-8fe29af1ac0b · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention LLaVA-OneVision: Easy Visual Task Transfer
Reference 2000
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ff8ca33-d940-433b-be82-67ebea37e28f · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention Informed routing in llms: Smarter token- level computation for faster inference.arXiv preprint arXiv:2510.13831,
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e25429e2-c7ad-4414-a600-650fcb048802 · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention BLIP-2: Boot- strapping language-image pre-training with frozen image encoders and large language models
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91f02fb3-bf7b-443b-a0cb-d07aa6f705b0 · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42a9394c-6e02-47db-8123-83d92b80befd · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention Dynamic-LLaVA: Efficient Multimodal Large Language Models via Dynamic Vision-language Context Sparsification
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba07eca3-772f-4565-9d43-a4718bd14d41 · outbound
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention Qwen2.5-VL Technical Report
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36894a2f-169f-4e22-aa11-c9d715c1990f · inbound
SkipOPU: An FPGA-based Overlay Processor for Large Language Models with Dynamically Allocated Computation ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ff9dad2-ab9e-4d4a-9b72-d7405e708a29 · inbound
ViCoStream: Streaming VideoLLMs Can Run Beyond 100 FPS with Stage-Wise Coordinated Inference ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5477009a-0d09-4044-911e-9ec5cb869b84 · inbound
From Recognition to Understanding: Unlocking Cognitive Time Series Reasoning with LLMs ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a404a56c-7951-4876-828e-90c1de1ccc84 · inbound
Intrinsic and Triangulation-Agnostic Attention: A Simple and Powerful Approach for Learning on Meshes ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.