Pith. sign in

Paper Citation Record · LEDGER

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning

As of 7 August 2026, this Paper Citation Record lists 33 of 33 outbound references and 4 inbound Pith citation observations for arXiv:2506.21873.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.21873 v1

Coverage vector

measured 33 of 33 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:22:38.056203Z

measured 37 of 37 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T04:26:51.521243Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T11:28:04.115693Z

Reference resolution

33 of 33 outbound references displayed

  • verified exact1
  • verified fuzzy14
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 87ff0b8a-ca3b-4887-b17f-c9c0cec093d5 · outbound

This paper cites Learning Transferable Visual Models From Natural Language Supervision.

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning Learning Transferable Visual Models From Natural Language Supervision

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T22:22:34.222152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:22:34.222152Z digest=sha256:1f1ef4c0d78e9ff6cc11ade9e4d2e8a122f120963a552d8ac0048dae4eae9f19

Observation 6a4d00e1-d3d4-4eff-8769-671d4229e0c0 · outbound

This paper cites Lmms-eval: Accelerating the development of large multimoal models, 2024.

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning Lmms-eval: Accelerating the development of large multimoal models, 2024

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:22:40.797436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:22:34.319536Z digest=sha256:c7c004bec792290b444685b59c7b079b61fdf76629e58b20bbcec3fe66efd89e

Observation 13662c94-8ae8-41ee-ae62-f8865c5086aa · outbound

This paper cites an unresolved cited work.

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T22:22:34.453915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:22:34.453915Z digest=sha256:55ab3582b2ff45241f6f02ee716ab71ac368087bedaca15b3fdcec3fb5868b0d

Observation ed13873d-7303-41c2-887f-f70ad36352f5 · outbound

This paper cites PuMer: Pruning and merging tokens for efficient vision language models.

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning PuMer: Pruning and merging tokens for efficient vision language models

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:22:40.638759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:22:34.502806Z digest=sha256:4330a19603c3a75a15ad100b23f4b0c853666610a46b66ae8c860ccf90255ae1

Observation 3f9feba8-b212-40bb-b00d-6e25488e19e9 · outbound

This paper cites Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic.

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T22:22:34.600551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:22:34.600551Z digest=sha256:39b2282b3bd77f9e26554fd4c879127a94fd3110c229b1892680337752c46a3e

Observation 5af470b7-381b-4a2d-9d38-e97d8fc76450 · outbound

This paper cites Chasing sparsity in vision transform- ers: An end-to-end exploration.

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning Chasing sparsity in vision transform- ers: An end-to-end exploration

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:22:40.493296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:22:36.588326Z digest=sha256:8c63368b4f1a2718a5a6d23cd5a94d6105b308af0133af46e4ce268092cb33aa

Observation 7977b23f-3ab8-4e3b-a190-dbfb2a7e998a · outbound

This paper cites Recoverable Compression: A Multimodal Vision Token Recovery Mechanism Guided by Text Information.

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning Recoverable Compression: A Multimodal Vision Token Recovery Mechanism Guided by Text Information

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T22:22:36.892337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:22:36.892337Z digest=sha256:ecc1c6b50dc6bf16f5788f687662276b9092632326094c8a16eae24ec7e1e384

Observation fc0b8107-ea83-4856-9830-afdc21830e16 · outbound

This paper cites Bert: Pre-training of deep bidirectional trans- formers for language understanding, 2019.

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning Bert: Pre-training of deep bidirectional trans- formers for language understanding, 2019

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:22:40.317527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:22:36.947010Z digest=sha256:6690a6f6782c225a04caba1c168abd36b461529fb4f66218a947e216c2e553c1

Observation 9d0203be-a9ae-4267-b819-47f2116e220d · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T22:22:37.063835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:22:37.063835Z digest=sha256:d8c80db9b7746d909e3ccd3c1bfb445493fe0a049afa98aafbe48c040aa60864

Observation 80d3a7ef-ddd5-48d5-8a6d-db11f437122b · outbound

This paper cites Mme: A compre- hensive evaluation benchmark for multimodal large language models, 2024.

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning Mme: A compre- hensive evaluation benchmark for multimodal large language models, 2024

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:22:40.112200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:22:37.158952Z digest=sha256:82172af088acbcbd3245707e587e40f56caeeab6c987a8c8a628787a41c47db5

Observation 6252318a-6096-4ab7-8ba4-8f2660f663eb · outbound

This paper cites Stangl, Anhong Guo, Chi Lin, Kristen Grauman, Jiebo Luo, and Jeffrey P.

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning Stangl, Anhong Guo, Chi Lin, Kristen Grauman, Jiebo Luo, and Jeffrey P

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:22:39.929819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:22:37.240849Z digest=sha256:4c9586c3173b18ae3d80192a693c2dd1716439591766d6d55018464a8a5159f9

Observation 96308886-7034-4e99-b06c-ec9193ca46fc · outbound

This paper cites Ferret: Refer and Ground Anything Anywhere at Any Granularity.

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning Ferret: Refer and Ground Anything Anywhere at Any Granularity

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T22:22:37.288321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:22:37.288321Z digest=sha256:06e552b07208678dd69114427b0296fdb6f79a9cda4aa8a9093143be1faf889b

Observation 94758521-e0ad-4fef-9396-4cb57a0b7a2c · outbound

This paper cites Hudson and Christopher D.

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning Hudson and Christopher D

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:22:39.759716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:22:37.322374Z digest=sha256:1c05d62577f3c12e72eddd52a1eed4391a3f5310a063cd4586bf79e1b5276f14

Observation 7ffded63-3567-47ec-910b-4f75f5548ef0 · outbound

This paper cites MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning.

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T22:22:37.354640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:22:37.354640Z digest=sha256:e76039353f358c385638d9ddb408f633f0c85f65b10d30dad0c3121516de3a6d

Observation e5177977-f620-44b3-af16-da781b15a257 · outbound

This paper cites MDETR -- Modulated Detection for End-to-End Multi-Modal Understanding.

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning MDETR -- Modulated Detection for End-to-End Multi-Modal Understanding

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T22:22:37.398957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:22:37.398957Z digest=sha256:d490e201f7768e14b9d512711cd184620645ebef7e754a8408d5a917dacc9ceb

Observation 1047d42b-4035-4686-8f2f-34b0c62f05e4 · outbound

This paper cites ReferItGame: Referring to objects in pho- tographs of natural scenes.

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning ReferItGame: Referring to objects in pho- tographs of natural scenes

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:22:39.581887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:22:37.439860Z digest=sha256:c6c28a416a2fe5890595b6b9dac185b4c16f4f51627049f2a80a3cfe98f4bc05

Observation 6aaf2107-1780-41a2-9582-3bfcd228cdca · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning LLaVA-OneVision: Easy Visual Task Transfer

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T22:22:37.486139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:22:37.486139Z digest=sha256:9c4efd4a4714472408a03798ebe5308c6f8204189937e38ebbd5d218b0a70a7c

Observation 6b850b5a-68b8-422d-8942-b13c090b933f · outbound

This paper cites Improved baselines with visual instruction tuning, 2023.

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning Improved baselines with visual instruction tuning, 2023

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:22:39.392107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:22:37.532926Z digest=sha256:7da6789bbedda3c4e4709209ca4794f7d8ff4ff3aab9d1eba2aed42b2a362899

Observation 68059e97-b956-4043-b2dd-cd8c386fdc96 · outbound

This paper cites Visual instruction tuning.

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning Visual instruction tuning

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:22:39.211131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:22:37.566606Z digest=sha256:64a88afd264e1fb8955a1d26d9e5020db47e3386d3e5d7b7ec73f9228585f32f

Observation a98f1f51-a427-4ee7-8008-049e208faa8f · outbound

This paper cites Llava-next: Im- proved reasoning, ocr, and world knowledge, 2024.

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning Llava-next: Im- proved reasoning, ocr, and world knowledge, 2024

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:22:39.048593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:22:37.595557Z digest=sha256:0020abb4a2e73b17480088e99f280b6b757c0a49008a60e234f51575ab2c1f0b

Observation 56f3c6af-f16a-4ade-a470-9398eb3fad83 · outbound

This paper cites Ok-vqa: A visual question answer- ing benchmark requiring external knowledge.

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning Ok-vqa: A visual question answer- ing benchmark requiring external knowledge

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:22:38.868981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:22:37.636569Z digest=sha256:7aa5a27b0b64a4af87dd28e3d1573c7cce0d1cd5b4744fe053d4dea0b96a5955

Observation 54baaea2-b91c-4979-b628-4e52c11da981 · outbound

This paper cites Llava-prumerge: Adaptive token reduction for efficient large multimodal models.

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning Llava-prumerge: Adaptive token reduction for efficient large multimodal models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T22:22:37.656404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:22:37.656404Z digest=sha256:d8757eb3c3298aa0772d55a6e916f81d64bc0e507bed5deeee1fbbdeca4c64de

Observation 503b3c1d-7e40-4cf9-b445-c27e3b6e1c85 · outbound

This paper cites Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection.

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T22:22:37.694996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:22:37.694996Z digest=sha256:95997a036507819947b6624e65aa2a2f61e86cf3dfe2df2675f600dc99d237cd

Observation 8632f363-36a1-41db-8844-8a1c699956a1 · outbound

This paper cites Less is More: A Simple yet Effective Token Reduction Method for Efficient Multi-modal LLMs.

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning Less is More: A Simple yet Effective Token Reduction Method for Efficient Multi-modal LLMs

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T22:22:37.715782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:22:37.715782Z digest=sha256:33b45e4d78f867f7a86136791bfdb660f57b7f17927ef43a37d9642c46fe579e

Observation 7f131960-fa44-40a5-9d6e-9b4f5b04fd33 · outbound

This paper cites SmartTrim: Adaptive Tokens and Attention Pruning for Efficient Vision-Language Models.

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning SmartTrim: Adaptive Tokens and Attention Pruning for Efficient Vision-Language Models

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-06T22:22:38.253638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:22:37.743642Z digest=sha256:dc6f967657e04d2e500f4fd8f285661f79fafdbad8d539c0d7f333aa8eafc09e

Observation 9006af8c-53da-4f81-ac21-12430f38a059 · outbound

This paper cites CogVLM: Visual Expert for Pretrained Language Models.

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning CogVLM: Visual Expert for Pretrained Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T22:22:37.779534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:22:37.779534Z digest=sha256:d37af9c2a452d6847bcdc436c0083fd017add8a6011068b5f7bf916bd54ee058

Observation 740a8bef-e1ff-46b1-89b5-71d8a266bd68 · outbound

This paper cites Sigmoid Loss for Language Image Pre-Training.

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning Sigmoid Loss for Language Image Pre-Training

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T22:22:37.804075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:22:37.804075Z digest=sha256:1b608e07ba36827ab5030f3428a261bc9714056ba77870b9a7efd3119b2c3603

Observation f0abc5a8-7da7-4a9b-b116-4f52f17b07b3 · outbound

This paper cites Referring Expression Comprehension: A Survey of Methods and Datasets.

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning Referring Expression Comprehension: A Survey of Methods and Datasets

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T22:22:37.839061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:22:37.839061Z digest=sha256:10b27550916d5957097414f7b78047ba5530a4d254f952e68742accb420c06be

Observation 00fd528a-580c-476a-9dff-37e14e2d1699 · outbound

This paper cites Sparsity Meets Similarity: Leveraging Long-Tail Distribution for Dynamic Optimized Token Representation in Multimodal Large Language Models.

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning Sparsity Meets Similarity: Leveraging Long-Tail Distribution for Dynamic Optimized Token Representation in Multimodal Large Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T22:22:37.891208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:22:37.891208Z digest=sha256:04c2d8f275e9367b5d3791606f0a114c80960cb689c97c72e827b9f847c30a8e

Observation f291d98b-bdaa-49bb-bf10-0afed6d4abc4 · outbound

This paper cites Mm-vet: Evaluating large multimodal models for integrated capabilities, 2023.

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning Mm-vet: Evaluating large multimodal models for integrated capabilities, 2023

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:22:38.683766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:22:37.932357Z digest=sha256:d0697485d5dceaa189a9b354a2f285421e47807b63f0895f50fdeaa6505e1e6a

Observation a7432697-617b-4d76-9b68-62afca620f25 · outbound

This paper cites EVA: Exploring the Limits of Masked Visual Representation Learning at Scale.

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning EVA: Exploring the Limits of Masked Visual Representation Learning at Scale

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T22:22:37.974207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:22:37.974207Z digest=sha256:dc46c91fb63d6b6bbc3203f3f664ef3bceb662b30bcfa128cef9c4d054cc51bb

Observation 9769bad8-5242-44f0-8796-3fb811209059 · outbound

This paper cites Beyond Text-Visual Attention: Exploiting Visual Cues for Effective Token Pruning in VLMs.

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning Beyond Text-Visual Attention: Exploiting Visual Cues for Effective Token Pruning in VLMs

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T22:22:38.007759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:22:38.007759Z digest=sha256:cdccfe5c91b03227bc484eb319308baec6da7bb2a7094025c8614edf7779577d

Observation 8406bbc7-9c0c-4615-b864-0bf5b81d436a · outbound

This paper cites Sparsevlm: Visual token sparsification for efficient vision- language model inference.

Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning Sparsevlm: Visual token sparsification for efficient vision- language model inference

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:22:38.533842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:22:38.056203Z digest=sha256:318678033493ec4baee13765200659e520317959cbc5af1d2a7685059ffaeba4

Pith citing papers

Observation 1d10623c-83fb-4f45-95cd-23b3146a6d3d · inbound

Reroute, Don't Remove: Recoverable Visual Token Routing for Vision-Language Models cites this paper.

Reroute, Don't Remove: Recoverable Visual Token Routing for Vision-Language Models Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-03T11:28:04.117867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T09:35:24.118536Z digest=sha256:4f2f221d66398f567408fd8ba16e07687d8fd12615824f53ea8d32f834961133

Observation dd37fb2c-b09e-4a78-b543-4a85336b2d2f · inbound

Token-Sparse Medical Multimodal Reasoning via Dual-Stream Reinforcement Learning cites this paper.

Token-Sparse Medical Multimodal Reasoning via Dual-Stream Reinforcement Learning Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning

Reference 26

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:15:45.167041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-07-01T05:36:43.609602Z digest=sha256:16f0260ef16f0454d531bc575f3dcbccd5457cd5210b2cb7c25d3096cd2066f0

Observation 9580cabb-5320-4a7c-ba02-205e08d07265 · inbound

Beyond Accuracy: Auditing Spatial Provenance in Visual Token Pruning for OCR-Critical MLLM Inference cites this paper.

Beyond Accuracy: Auditing Spatial Provenance in Visual Token Pruning for OCR-Critical MLLM Inference Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-04T01:29:54.348125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T01:29:54.348125Z digest=sha256:04e8ce8bdcf43d48b8969843dd2a7f7bf458684325e1aa44dc88118be276cc99

Observation 1208ac5a-4bec-4880-8f3e-70d05e81322a · inbound

Beyond Accuracy: Auditing Spatial Provenance in Visual Token Pruning for OCR-Critical MLLM Inference cites this paper.

Beyond Accuracy: Auditing Spatial Provenance in Visual Token Pruning for OCR-Critical MLLM Inference Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-05T04:26:51.521243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T04:26:51.521243Z digest=sha256:e06e109661406ecbc2e3a8382cf01cbc35880d2db3b8349a06c66911830690cc