Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T13:42:29.264825Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 67 of 67 outbound references and 5 inbound Pith citation observations for arXiv:2509.00419.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T13:42:29.264825Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T05:03:26.782904Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T15:59:57.422464Z
67 of 67 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation e39a8bf1-6a96-4633-a840-16d2d58e8397 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression online" 'onlinestring :=
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9af7304-6f1e-4b75-93f5-cf7be541e311 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression write newline
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f5f2323-fbf7-48d4-b866-c3214d6797bf · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d7f06c1f-40ff-48c0-80e1-ca68a43e3c98 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression HiRED: Attention-Guided Token Dropping for Efficient Inference of High-Resolution Vision-Language Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 603fcc49-fa2b-4b00-8795-4c2462b8697c · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Qwen2.5-VL Technical Report
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 347f759e-8492-48c5-84ef-6dfbb74b682a · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19f6daa0-b112-4d0b-8739-d3cb0e016320 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Are We on the Right Way for Evaluating Large Vision-Language Models?
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df1e7d22-0ce9-428d-aadb-65a68c83b5fa · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7dd90ed1-4f5f-402c-b270-bec831610849 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b71f6f9e-8bbd-488c-af92-ebd5aa486989 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression The Llama 3 Herd of Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 560f4c9c-079f-44b7-8539-577a27236f78 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f23f3775-03fd-4e32-a21e-5b146cf594ea · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1147cb59-830a-408f-99b9-19c28faace18 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 71527a54-85a4-408a-a8cc-59a865b404be · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a885e856-5489-4b10-a5b8-8f1246dbcfe6 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ae53f67d-1f4c-43a2-b450-9aa43406ddc4 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c74e6b7b-d87b-4c18-a689-7a8d5c6b64e9 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2dc21c2c-f757-4c8f-8ff8-2dacefe3fbdc · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 15e204f8-d967-4db2-8f4e-5215cf798120 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 88adabbc-f3bd-4964-881c-21c49781875d · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7942e703-e04c-4a86-9248-d90e23cb6765 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Mixtral of Experts
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a371c47-94e7-4010-9ec6-7bcc63c35797 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d9b0595-6538-432c-8b88-2af5e8c6db83 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Length-Adaptive Transformer: Train Once with Length Drop, Use Anytime with Search
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9184c06d-1e16-43a5-a5a6-8298c04848e3 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d544c8a8-2608-4435-bd7b-f86a3f2ebff6 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression A Study on Token Pruning for ColBERT
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80d32214-cd9b-4680-a3e2-3236d9848557 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression LLaVA-OneVision: Easy Visual Task Transfer
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1fcfaf2a-5b8b-41b2-990d-0e07150b2d22 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a4a0ee39-230b-4234-bb61-943f16198377 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 64bdbd4b-cafd-471f-b876-bb7f4949788d · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87b720a0-3bce-4fe3-9172-8b2ce2411744 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6c4ce98d-5f94-448c-8c52-98e5e06645e4 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8f897563-0faf-4608-9fef-cee0c1953d83 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3cfce5f-6fc0-463c-b5ae-ba57ba05aea9 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df2cab8c-e828-45da-a04b-041ab567f309 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b3f3b9b4-3496-4f5f-88e8-c0ad221b67bb · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8602b363-94c4-44e2-92b9-90a272b24cf9 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8fff648c-754a-4a65-abf4-58b9a9a03746 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7c1019a9-df2d-4491-9132-b0c54f8e5909 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 54933ae2-6ad6-44e0-ac61-f8078488fb40 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 572a5fd9-9bb5-413f-b21d-fc320bcca9a9 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 268c6568-f2c4-4d78-b141-00cc5df95e19 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6310c7f4-3ba8-4fc0-bcd0-e43273fb5e01 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5a75df77-d674-49af-8aa5-1a7007c24541 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1858da30-ad44-4644-b2d6-33a296c7664e · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e33b9e5b-8a14-45c9-91e5-185a5831f4fc · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5f53bc73-532b-48fd-a632-d703e01c71ba · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b94b5cbe-cc0b-4558-9749-fdc79c3181a6 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a5556458-4547-43b2-8622-19b6cc050ee1 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression MuirBench: A Comprehensive Benchmark for Robust Multi-image Understanding
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2891eb46-dd02-4ca6-83e6-12703098abad · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0f5980a-123e-4b83-b2c3-a3d69785e456 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5d9b632-2d99-4c96-bc62-58df3fb18d29 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 10561fb1-d74c-4e84-a724-da5e72d258e5 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression LVLM-eHub: A Comprehensive Evaluation Benchmark for Large Vision-Language Models
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45b5c767-8842-4af4-8e06-31814c83e787 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bd5ad8e8-fc81-48ee-a34e-24501979196a · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Qwen2.5 Technical Report
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e564b3d-cf4c-499e-ac67-fe18c1b5b0d0 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 89509a5e-4f84-4017-9308-07e483a219ee · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression LiDAR-LLM: Exploring the Potential of Large Language Models for 3D LiDAR Understanding
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 049c161e-8a49-4b47-9290-ff795abedf9a · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6f26fe0-4e32-421e-a57d-cf143ef66f24 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6531b8d-3542-4a86-9833-35c90dc321a1 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2c558228-3dce-432e-9d48-c7a5414d3206 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5538dd61-e472-4646-a6fe-73bfb96e193a · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression MMMU-Pro: A More Robust Multi-discipline Multimodal Understanding Benchmark
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb9953f2-7e5b-4906-b8aa-4cfbb8fb392a · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a8b6add4-2c32-4340-9e76-eb965d4975ed · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Beyond Text-Visual Attention: Exploiting Visual Cues for Effective Token Pruning in VLMs
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d467945-0623-4172-9ce8-608ec0c12b99 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Token-level Correlation-guided Compression for Efficient Multimodal Document Understanding
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 559511f1-0a75-4938-ab81-a3d6c7986e18 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5e0251c3-181a-43d3-8741-1daedf4f0dba · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression Unresolved cited work
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2f0c0531-01f2-4bf4-8299-2aec7fa82cd7 · outbound
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression MLVU: Benchmarking Multi-task Long Video Understanding
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3a4f5ed-3447-4948-8f42-5d316adb27df · inbound
HybridKV: Hybrid KV Cache Compression for Efficient Multimodal Large Language Model Inference LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cfbce84e-0766-4d32-b12b-649646c1da17 · inbound
CIVIC: End-to-End Sequence Compactness for Efficient Vision-Language Models LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f36012ed-4480-4707-9ac7-e2c97e48ee7b · inbound
Accelerating Multimodal Large Language Models with Prior-Corrected Token Reduction LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation be980c68-f51c-42f2-9eb3-a29cbae6d681 · inbound
Spectral Evolution-Guided Token Pruning in Multimodal Large Language Models LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d13bbfb4-1f49-43db-b16c-80d33edd3730 · inbound
Attention-Free and Lightweight Token Reduction for Efficient Vision-Language Models LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.