Pith. sign in

Paper Citation Record · LEDGER

VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration

As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 15 inbound Pith citation observations for arXiv:2410.23317.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.23317 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T23:09:25.101407Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T01:36:44.125857Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b031d2b5-81c3-4933-942f-d478c5a449db · inbound

FrameFusion: Combining Similarity and Importance for Video Token Reduction on Large Vision Language Models cites this paper.

FrameFusion: Combining Similarity and Importance for Video Token Reduction on Large Vision Language Models VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T23:09:25.101407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:09:25.101407Z digest=sha256:8472071a50bff12f734578ad3c95de77a49e33a8cd9816b4b806cc30f6a32d27

Observation 90c47060-c235-431c-96c6-da62c80e40bc · inbound

EffiVLM-BENCH: A Comprehensive Benchmark for Evaluating Training-Free Acceleration in Large Vision-Language Models cites this paper.

EffiVLM-BENCH: A Comprehensive Benchmark for Evaluating Training-Free Acceleration in Large Vision-Language Models VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T12:09:08.527752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:09:08.527752Z digest=sha256:5b513e9399274ea9321455e4291f2e27516c1568fd3dd58a700a6406367fb8dc

Observation 5640b0b4-4a97-406b-bbbe-6fce424be439 · inbound

Blink: Dynamic Visual Token Resolution for Enhanced Multimodal Understanding cites this paper.

Blink: Dynamic Visual Token Resolution for Enhanced Multimodal Understanding VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-03T17:09:37.841750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T17:09:37.841750Z digest=sha256:00312743c6550ec2da3cd74236db83103bf0e522ec1c40738eebd0f20a51fd52

Observation 0eaae785-0c23-4276-af5c-453416f6e029 · inbound

FreqCache: Accelerating Embodied VLN Models with Adaptive Frequency-Guided Token Caching cites this paper.

FreqCache: Accelerating Embodied VLN Models with Adaptive Frequency-Guided Token Caching VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:16:14.386310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-08T03:10:29.477937Z digest=sha256:d2061e132371ca9357999c27e57a76cabae1eb8a8fc8ddf89277c71151c8c642

Observation 3badd050-e228-468c-b125-9876c8b2ff81 · inbound

Prefill-Time Intervention for Mitigating Hallucination in Large Vision-Language Models cites this paper.

Prefill-Time Intervention for Mitigating Hallucination in Large Vision-Language Models VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:31:16.372723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-07T16:50:26.044591Z digest=sha256:e317df241a1cdb7aaa2dbc70fb541020fa36f2b88e546630c0e4bd5a3496a054

Observation 2542ebda-0081-4fd9-b658-cd54ead99e4f · inbound

Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction cites this paper.

Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:41:26.551067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-12T05:02:25.513351Z digest=sha256:ca1ada3daa725846acf710afcd6dd270dd7cc159eecabfc6953de94e51e8c6da

Observation 6046ae1d-bb82-41cc-8c7a-e866a4f8f604 · inbound

KVCapsule: Efficient Sequential KV Cache Compression for Vision-Language Models with Asymmetric Redundancy cites this paper.

KVCapsule: Efficient Sequential KV Cache Compression for Vision-Language Models with Asymmetric Redundancy VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-20T19:58:59.047015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-20T19:55:35.051834Z digest=sha256:7a29b269e378783d8622ed3ecbd343077687cb3215422b789c1adc7b993c4114

Observation f29b59e7-a302-4fff-965d-7748e7d2375b · inbound

Rotation-Aligned Key Channel Pruning for Efficient Vision-Language Model Inference cites this paper.

Rotation-Aligned Key Channel Pruning for Efficient Vision-Language Model Inference VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:43:23.805107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-20T07:43:18.828740Z digest=sha256:2a1c279069b33c3bfa03c089a62bb78ec2c22841b475d0d1d2bde39f801b46ee

Observation 82acc9df-1f2a-40cf-bd23-5dad3beb21df · inbound

CIVIC: End-to-End Sequence Compactness for Efficient Vision-Language Models cites this paper.

CIVIC: End-to-End Sequence Compactness for Efficient Vision-Language Models VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-06-29T12:13:26.457127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T12:12:51.867760Z digest=sha256:0367385df7587469facbf2858d447560d7eca151eaf0020aba492f9cb747ee89

Observation 7b166343-3228-4f89-8eb1-e3d06546c103 · inbound

AsymVLM: Asymmetric Token Pruning for Efficient Vision-Language Model Inference cites this paper.

AsymVLM: Asymmetric Token Pruning for Efficient Vision-Language Model Inference VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:53:15.966617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T08:48:58.583657Z digest=sha256:3e6016226d64870d47508ccd1cb88f3389022f733a163adec7d3cf857d19e3fa

Observation 7e68414f-3ad7-4120-bf96-bf65e130301d · inbound

AURA: Action-Gated Memory for Robot Policies at Constant VRAM cites this paper.

AURA: Action-Gated Memory for Robot Policies at Constant VRAM VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-07-01T23:16:24.296427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-28T14:30:49.006075Z digest=sha256:a0e7f9ac77778852a291ce80a2d3d943d8d30b85d90e93d9af408ce7a99854fa

Observation a76f5f52-7d7a-4b1d-a24b-8f239a7457e6 · inbound

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents cites this paper.

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration

Reference 115

Resolution
verified exact
local_arxiv, observed 2026-07-10T01:36:44.126991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-07-10T01:26:59.421158Z digest=sha256:458194f0f98057a9b82889bba3379ecfa9ba8db8627101b506bcd0169c699eaa

Observation a8c0ecb3-1caa-4907-9a68-1e4757205293 · inbound

Reflex: Real-Time VLA Control through Streaming Inference cites this paper.

Reflex: Real-Time VLA Control through Streaming Inference VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T01:26:09.793123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:26:09.793123Z digest=sha256:82f287a4a081d0b56e63a59fafa7fb7dcff0f34871fa3f1788edb2965bbb489c

Observation a074e4b0-1be7-4cee-9531-71a8ab901f67 · inbound

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs cites this paper.

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-04T13:43:56.846247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T13:43:56.846247Z digest=sha256:bc84a062ed7772283df8c0c77c4503d3946e3f57a5adc462e87d8f81cc7916dd

Observation 0f2b95f0-7e76-412d-a8bf-6831a46cd8b1 · inbound

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs cites this paper.

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-07T00:11:49.128823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:11:49.128823Z digest=sha256:52ed0665389d3c8146c5ac1413aa6943b42fcc30e5cd65f962b846a334d98a33