Pith. sign in

Paper Citation Record · LEDGER

Similarity-Aware Token Pruning: Your VLM but Faster

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2503.11549.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.11549 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:20:31.175344Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T15:59:57.424957Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 8e4ee5fe-b3d4-4fb5-864d-64b5088904fe · inbound

Beyond Attention or Similarity: Maximizing Conditional Diversity for Token Pruning in MLLMs cites this paper.

Beyond Attention or Similarity: Maximizing Conditional Diversity for Token Pruning in MLLMs Similarity-Aware Token Pruning: Your VLM but Faster

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-07T04:20:31.175344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:20:31.175344Z digest=sha256:98d806d538220aefb2ff43d9660997734aa51b2a67cb81d44f17903b2c877882

Observation 024ed357-3714-4cb4-b720-08028238635e · inbound

HAWK: Head Importance-Aware Visual Token Pruning in Multimodal Models cites this paper.

HAWK: Head Importance-Aware Visual Token Pruning in Multimodal Models Similarity-Aware Token Pruning: Your VLM but Faster

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:01:06.116080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T17:20:19.653806Z digest=sha256:bd2c762e0254808de52068c1cbc3c5c1b0ed15e6d75180a8f4d70ae0b3c242bf

Observation d5a38039-ef7a-4bb5-92cf-1a675bbca74c · inbound

RetentiveKV: State-Space Memory for Uncertainty-Aware Multimodal KV Cache Eviction cites this paper.

RetentiveKV: State-Space Memory for Uncertainty-Aware Multimodal KV Cache Eviction Similarity-Aware Token Pruning: Your VLM but Faster

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:26:02.138347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T15:29:17.567557Z digest=sha256:b33f595d1d5f60174971c801c249e8e13631859918cb629de2124ee5d7a0598f

Observation ca5aa4e9-997a-4e54-a2b7-e1df500d4ff6 · inbound

VisMMOE: Exploiting Visual-Expert Affinity for Efficient Visual-Language MoE Offloading cites this paper.

VisMMOE: Exploiting Visual-Expert Affinity for Efficient Visual-Language MoE Offloading Similarity-Aware Token Pruning: Your VLM but Faster

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:36:06.620346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-09T16:10:22.588945Z digest=sha256:8ad64ee5643bf6e99a34dc5374b71d7e98e35f93b7a6d3bbc6610a33e9541ac9

Observation e8bb49e9-c983-417c-b75d-cb3f87f50185 · inbound

DynaTok: Temporally Adaptive and Positional Bias-Aware Token Compression for Video-LLMs cites this paper.

DynaTok: Temporally Adaptive and Positional Bias-Aware Token Compression for Video-LLMs Similarity-Aware Token Pruning: Your VLM but Faster

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-20T06:58:06.004736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T06:55:01.619441Z digest=sha256:9a16d35ce4a5ff3a0b5fc15bdce5c11c98e1e2f2879db3e06f0aaae65c23f2e4

Observation e37c1600-d6d6-4683-8888-1a566f0c04f9 · inbound

Beyond Surrogate Gradients: Fully Differentiable Token Pruning for Vision-Language Models cites this paper.

Beyond Surrogate Gradients: Fully Differentiable Token Pruning for Vision-Language Models Similarity-Aware Token Pruning: Your VLM but Faster

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:13:27.411246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T13:07:21.789901Z digest=sha256:5ea081a2de77a5c5ad70c8f95075785e80e592611886a2d9449dbc2b9b3f38d8

Observation 23ea4db1-ac06-4d3f-9fbd-9107811bef42 · inbound

RAPID: Layer-Wise Redundancy-Aware Pruning and Importance-Driven Token Merging for Efficient ViT cites this paper.

RAPID: Layer-Wise Redundancy-Aware Pruning and Importance-Driven Token Merging for Efficient ViT Similarity-Aware Token Pruning: Your VLM but Faster

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-02T21:27:24.487816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T19:46:23.188930Z digest=sha256:774805f9e9a0f23d78f71bed0aff596f309a3a729f32182a3f89990baadcb4a7

Observation a1e3b4b9-1721-4539-9413-73f061516819 · inbound

AVIS: Adaptive Test-Time Scaling for Vision-Language Models cites this paper.

AVIS: Adaptive Test-Time Scaling for Vision-Language Models Similarity-Aware Token Pruning: Your VLM but Faster

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-07-03T08:17:45.886204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T10:47:41.183211Z digest=sha256:9935b2ec68a333d51535dfb279bb16b6c4bdc06c3752367b660e2b59c4b26205

Observation 096ab6c4-a4bb-4e65-92ae-d45dcd9c0bb0 · inbound

Accelerating Multimodal Large Language Models with Prior-Corrected Token Reduction cites this paper.

Accelerating Multimodal Large Language Models with Prior-Corrected Token Reduction Similarity-Aware Token Pruning: Your VLM but Faster

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T15:59:56.757240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T01:10:00.373247Z digest=sha256:b1e08362a91927780e5256a9dbd797d80192fe47aef5381bce9cd7d0b78d11a8

Observation 8d76fb6d-49a5-45e0-8371-e81fbe9c6771 · inbound

Spectral Evolution-Guided Token Pruning in Multimodal Large Language Models cites this paper.

Spectral Evolution-Guided Token Pruning in Multimodal Large Language Models Similarity-Aware Token Pruning: Your VLM but Faster

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T15:59:57.426487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T01:04:20.802531Z digest=sha256:0b15a7f6946c9b40dbd9ec84e39a44edadc3b3d76b87ee2a1e0c59a5930c1736

Observation e38bd0a8-f4a2-43fc-8418-9115e89b8aa6 · inbound

TOPS: First-Principles Visual Token Pruning via Constructing Token Optimal Preservation Sets for Efficient MLLM Inference cites this paper.

TOPS: First-Principles Visual Token Pruning via Constructing Token Optimal Preservation Sets for Efficient MLLM Inference Similarity-Aware Token Pruning: Your VLM but Faster

Reference 35

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T14:09:53.804242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T04:24:38.917137Z digest=sha256:176f7089e5a9f3b2ad630c928306ef11d0c6f59827ece6a2d037f5bedb200c99

Observation 9774a8cb-bf6f-4697-814e-07c4097cb51e · inbound

Attention-Free and Lightweight Token Reduction for Efficient Vision-Language Models cites this paper.

Attention-Free and Lightweight Token Reduction for Efficient Vision-Language Models Similarity-Aware Token Pruning: Your VLM but Faster

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T05:03:27.630859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:03:27.630859Z digest=sha256:44b41c73f8c87af1f2611447113c861d1b042bb271fff7f11ec8ceee776e3a2c

Observation 6db57410-b9af-47f3-b175-37c0f9f2f26a · inbound

C-PTQ: Fisher-weighted Channel-wise Sensitivity for Post-training Quantization of MLLMs cites this paper.

C-PTQ: Fisher-weighted Channel-wise Sensitivity for Post-training Quantization of MLLMs Similarity-Aware Token Pruning: Your VLM but Faster

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-01T08:35:55.672675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T08:35:55.672675Z digest=sha256:322aca404176117dd08b6f41d3dcecf7382e938b2e60cd81c05e072c1b1e7fc6

Observation acbe806f-1b20-4e4b-a766-b8adb443d60e · inbound

DiffPrune: differentiable information throttling for token pruning in vision-language models cites this paper.

DiffPrune: differentiable information throttling for token pruning in vision-language models Similarity-Aware Token Pruning: Your VLM but Faster

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T17:18:16.095066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:18:16.095066Z digest=sha256:259b37c7e8f9713b90acca3ed8dccb4f61d2349cc4c9bdfb11cd5ff4d7b2c740