Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T14:46:02.691635Z
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 4 inbound Pith citation observations for arXiv:2501.15021.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T14:46:02.691635Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-09T14:35:10.578399Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-20T19:58:59.018078Z
32 of 32 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 2907b0fc-0182-4702-b433-c119f298753e · outbound
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models A survey on deep multimodal learning for computer vision: advances, trends, applications, and datasets,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation aa69e52e-4501-4857-a2d6-2ed43cc9a2b7 · outbound
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models Multimodal intelligence: Representation learning, information fusion, and applica- tions,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f5a8b587-f08c-47c2-bf1e-b4f8f583f375 · outbound
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models A Survey on Multimodal Large Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36c7e950-bdf8-4211-860b-43f7ecd34277 · outbound
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models Vision- language models for vision tasks: A survey,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e5535967-32e9-4051-8d77-8d402f89e2af · outbound
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34a4dc67-f6ff-43a0-9299-eaaa39fe5e07 · outbound
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models Attention is all you need,
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a579d485-6400-495b-bae2-795a7872a962 · outbound
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models KIVI: A Tuning-Free Asymmetric 2bit Quantization for KV Cache
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7365c393-920c-4980-8ba4-da75b8ee00b6 · outbound
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models ZipCache: Accurate and Efficient KV Cache Quantization with Salient Token Identification
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 313df24d-61b0-4af9-bc3a-58f2b96edaef · outbound
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models SKVQ: Sliding-window Key and Value Cache Quantization for Large Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc4dbbfe-97da-4ac6-ace5-5beb16544490 · outbound
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 941b6830-364a-4660-bf14-9df404ed0d77 · outbound
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models Visual instruction tuning,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 422ee90a-9303-45ac-9d62-6c3b679dc066 · outbound
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models IntactKV: Improving Large Language Model Quantization by Keeping Pivot Tokens Intact
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8ef2a7d-6b2a-4539-a203-bfa076ed1026 · outbound
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models Efficient Streaming Language Models with Attention Sinks
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e98dbfdb-7b16-4efd-8aca-1176e80b66db · outbound
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models KVQuant: Towards 10 Million Context Length LLM Inference with KV Cache Quantization
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44d2c809-e1f0-4945-88ee-dfa0debcab89 · outbound
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models QuaRot: Outlier-Free 4-Bit Inference in Rotated LLMs
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aff4718b-c0a3-4b33-bcaf-1f30955985b8 · outbound
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models QuIP#: Even Better LLM Quantization with Hadamard Incoherence and Lattice Codebooks
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d01b6e7-0e06-4a96-9212-907b78e2c062 · outbound
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models Flashattention: Fast and memory-efficient exact attention with io- awareness,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 1683751b-9abd-411c-ac42-b2a68008c86d · outbound
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models Pointer Sentinel Mixture Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e23be23a-fca7-4ccf-9384-7ed714e1cc6e · outbound
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models Microsoft COCO Captions: Data Collection and Evaluation Server
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80883f06-b8de-4bdc-b98e-2f219fa9b5f7 · outbound
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e8d4ff5-4a2e-47c9-96c1-6650b7141b24 · outbound
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models Improved baselines with visual instruction tuning,
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73b22388-3389-48dc-a8ef-d364a7654621 · outbound
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bcbb5214-0551-4032-bdb9-eb7d2ad6792c · outbound
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models No Token Left Behind: Reliable KV Cache Compression via Importance-Aware Mixed Precision Quantization
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 036eb6bc-19f3-4c1e-879b-2877d18b95b1 · outbound
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models Massive Activations in Large Language Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2cff7cdb-d691-4be8-b6e2-7f2fcb76685b · outbound
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f04b087-f4aa-4ad4-8392-fd41af31a49a · outbound
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models ZipVL: Efficient Large Vision-Language Models with Dynamic Token Sparsification
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7d14ab3-8f70-4bdb-92cb-756da1020f2f · outbound
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models SpinQuant: LLM quantization with learned rotations
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61d6b4fe-7c93-47bc-85c8-1258a6985e36 · outbound
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models Roformer: Enhanced transformer with rotary position embedding,
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 573fe177-decd-4d7f-a325-eeef130edf6f · outbound
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models Unified matrix treatment of the fast walsh-hadamard transform,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 0fa639e4-067c-4126-9e24-8803931dd96e · outbound
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks,
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e23df800-0085-451a-ba68-199ae662ff8b · outbound
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models MileBench: Benchmarking MLLMs in Long Context
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aaab400c-d280-497f-830b-38469e01e32d · outbound
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models SmoothQuant: Accurate and efficient post-training quantization for large language models,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 6a3f4b03-3f62-4b49-a654-dd4d3dab1e07 · inbound
Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75706047-98c0-465f-8482-4571787ff875 · inbound
SnapMLA: Efficient Long-Context MLA Decoding via Hardware-Aware FP8 Quantized Pipelining AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 66c56521-712f-4369-a62d-1b327c02f325 · inbound
WindowQuant: Mixed-Precision KV Cache Quantization based on Window-Level Similarity for VLMs Inference Optimization AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 349a6c55-f8c7-4b09-9350-1f1978497543 · inbound
KVCapsule: Efficient Sequential KV Cache Compression for Vision-Language Models with Asymmetric Redundancy AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.