Pith. sign in

Paper Citation Record · LEDGER

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models

As of 14 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 0 inbound Pith citation observations for arXiv:2608.03112.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.03112 v1

Coverage vector

measured 34 of 34 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T00:55:50.162416Z

measured 34 of 34 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

34 of 34 outbound references displayed

  • verified exact0
  • verified fuzzy9
  • unresolved25
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6aa95c04-c4fd-424c-bf4e-5f080bb33dc9 · outbound

This paper cites Divprune: Diversity-based visual token pruning for large multimodal models.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models Divprune: Diversity-based visual token pruning for large multimodal models

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:55:50.724195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T00:55:50.036588Z digest=sha256:85a3c306dfc3303a6e6e66e467ac9954abba57d56affa7e88fa0c6c191a1c6c1

Observation 42b15421-c221-4477-b3a9-e3a5bfdbbe09 · outbound

This paper cites Qwen2.5-VL Technical Report.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models Qwen2.5-VL Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T00:55:50.040897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:55:50.040897Z digest=sha256:c5ceca834d82914bc84e9fddcf05ed01a59ec123c97df1a6ccc185b6ce06cc1f

Observation 0ec23800-fd5b-4476-8bc9-bb6b9304a4b7 · outbound

This paper cites LLaVA-KD: A Framework of Distilling Multimodal Large Language Models.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models LLaVA-KD: A Framework of Distilling Multimodal Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T00:55:50.045305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:55:50.045305Z digest=sha256:e4581851e4e53d1d8a46e6fa8da32771fde2379d03f475948222057bcbbf0822

Observation 3a00e5f9-79d9-4cc1-b3a2-4b63c60a11b1 · outbound

This paper cites An image is worth 1/2 tokens after layer 2: Plug-and-play inference acceleration for large vision-language models.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models An image is worth 1/2 tokens after layer 2: Plug-and-play inference acceleration for large vision-language models

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:55:50.713888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T00:55:50.049443Z digest=sha256:87f743e2de33d7b4eb648319844e0ad8a0581604b90c3a4368f6f79853453733

Observation 6aaaff18-9677-42cd-9ac8-98930dc07965 · outbound

This paper cites VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T00:55:50.053363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:55:50.053363Z digest=sha256:0b68251aaa238e77236b5a8bab4e1749cab553c85f298527ffb609540fec2e15

Observation e04ecb77-5faa-4c3d-a1c5-ab31c2c5dd2b · outbound

This paper cites Instructblip: Towards general-purpose vision- language models with instruction tuning.Advances in neural information processing systems, 36:49250–49267, 2023.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models Instructblip: Towards general-purpose vision- language models with instruction tuning.Advances in neural information processing systems, 36:49250–49267, 2023

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T00:55:50.057359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:55:50.057359Z digest=sha256:11c06e8040db2340be046c4f1ab42c0e7cf0aa85d87507419057ed0de1cff9a9

Observation 43fb8f65-9a80-46fd-9c21-fc66ef589dfb · outbound

This paper cites Video-mme: The first-ever comprehensive evaluation benchmark of multi-modal llms in video analysis.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models Video-mme: The first-ever comprehensive evaluation benchmark of multi-modal llms in video analysis

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T00:55:50.061082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:55:50.061082Z digest=sha256:4f37549643d750b128b36207e6727e129a6108990bd0eb533f3b5e68eabae50e

Observation 126021bf-1271-4442-a0ca-2f14caebe633 · outbound

This paper cites Attention Score is not All You Need for Token Importance Indicator in KV Cache Reduction: Value Also Matters.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models Attention Score is not All You Need for Token Importance Indicator in KV Cache Reduction: Value Also Matters

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T00:55:50.064934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:55:50.064934Z digest=sha256:16a2a5badd2d1e0b993e0aced52a9f7b5c5600dce4716dfd8750f8e555df2b46

Observation 97536288-417c-4bb4-a920-90d2bcc1e097 · outbound

This paper cites Filter, correlate, compress: Training-free to- ken reduction for mllm acceleration.arXiv preprint arXiv:2411.17686, 2024.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models Filter, correlate, compress: Training-free to- ken reduction for mllm acceleration.arXiv preprint arXiv:2411.17686, 2024

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-08T00:55:50.068691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:55:50.068691Z digest=sha256:aa67e7fb36fb38a7f6860b99bd666affccac38e035be1ced73e183704b0a8730

Observation 7e115841-d1f7-4f52-aae3-970b7392d232 · outbound

This paper cites Efficient Multimodal Learning from Data-centric Perspective.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models Efficient Multimodal Learning from Data-centric Perspective

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T00:55:50.072148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:55:50.072148Z digest=sha256:1d3c521efa4ad97cbdcc1a43c3059fa2b29c2cd51882c62229d98f0ca6d4da51

Observation 9b88ced9-c6c7-4e39-8b1f-f92e6d2fd0e6 · outbound

This paper cites Ivtp: Instruction-guided visual token pruning for large vision-language models.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models Ivtp: Instruction-guided visual token pruning for large vision-language models

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:55:50.691372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T00:55:50.076060Z digest=sha256:447f70862ebd47c4e63584cab1ef0da3eb49a41b13d3c1ee71faecc2f19e16fa

Observation 475ef900-2adf-4dfa-aa48-3fb750537e2c · outbound

This paper cites Fast pruning using principal components.Advances in neural information processing systems, 6, 1993.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models Fast pruning using principal components.Advances in neural information processing systems, 6, 1993

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:55:50.680940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T00:55:50.079301Z digest=sha256:726faa2dba9abbf06de82e4e6dbb9458a49ecf9d77442d1f823565a057539bb6

Observation 17d30c23-5d1b-4b62-b66b-5d3912721494 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models LLaVA-OneVision: Easy Visual Task Transfer

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T00:55:50.082897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:55:50.082897Z digest=sha256:a7866b210654f69647565999b8c735394300ba55e0b9fd06e2af6e746e46c7eb

Observation a2efe416-7b05-43a0-938b-57e1118d90c7 · outbound

This paper cites Llama-vid: An image is worth 2 tokens in large language models.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models Llama-vid: An image is worth 2 tokens in large language models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T00:55:50.086574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:55:50.086574Z digest=sha256:5d039344b5bbef9a8eac99b02e3425bbb847c32d60fbb8aaae31ae1b6c55b141

Observation a2fbb7ef-1df6-4de6-a5ce-e13c9d235442 · outbound

This paper cites Video-XL-Pro: Reconstructive Token Compression for Extremely Long Video Understanding.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models Video-XL-Pro: Reconstructive Token Compression for Extremely Long Video Understanding

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T00:55:50.089925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:55:50.089925Z digest=sha256:3d38da5135179db1c36313e8c9d5d74cb506a75161c1809b57a8fb7b23e9df7d

Observation d0860654-ac66-4fdc-9b8d-f70b7d351ac5 · outbound

This paper cites Video detail caption.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models Video detail caption

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:55:50.665223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T00:55:50.093552Z digest=sha256:7e37c12f71143ff6fa4dd2c5e2020b6a3b5babedae287a56d43accb1b494f132

Observation c51946b4-8667-4d8d-8930-2d124f6c9292 · outbound

This paper cites an unresolved cited work.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-08T00:55:50.654611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T00:55:50.096902Z digest=sha256:2be2fec5a8cc84584b89db5b0e885ee81d166fa51fb43e7e71a5baad53d20fc3

Observation 08fe127c-bdcf-4157-b0ee-2d503c0556d9 · outbound

This paper cites Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-08T00:55:50.100827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:55:50.100827Z digest=sha256:2a9491ed9b53e70cf0615c5bbdefd6a3201afb260a89cc08c60343af9b83456e

Observation 581c3805-4d45-4d1e-bd00-1602d8e60b17 · outbound

This paper cites Per- ception test: A diagnostic benchmark for multimodal video models.Advances in Neural Information Processing Sys- tems, 36:42748–42761, 2023.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models Per- ception test: A diagnostic benchmark for multimodal video models.Advances in Neural Information Processing Sys- tems, 36:42748–42761, 2023

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T00:55:50.104780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:55:50.104780Z digest=sha256:91c22f391ad088c4faeacd130875e35df000ce771e8db3edb4982a70a6c1d8dc

Observation 30bbac77-b536-4487-b583-6426e3c42043 · outbound

This paper cites Llava-prumerge: Adaptive token reduction for efficient large multimodal models.arXiv preprint arXiv:2403.15388,.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models Llava-prumerge: Adaptive token reduction for efficient large multimodal models.arXiv preprint arXiv:2403.15388,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-08T00:55:50.108437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:55:50.108437Z digest=sha256:314f581085db4f53223028d59c0a57ffb978fb3e2f83542b7de61a95e1160546

Observation 1cad450f-7298-4506-beb1-ab398a6982d8 · outbound

This paper cites Imp: Highly capable large multimodal models for mobile devices.IEEE Transactions on Multime- dia, 2025.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models Imp: Highly capable large multimodal models for mobile devices.IEEE Transactions on Multime- dia, 2025

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:55:50.638126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T00:55:50.112464Z digest=sha256:5514e43d42978598dd5fe6b580862570f8c43de42744eff8e7916e1ca1a6be0a

Observation 346c318a-9fc1-49a1-8dd2-d59fbe86ab65 · outbound

This paper cites LongVU: Spatiotemporal Adaptive Compression for Long Video-Language Understanding.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models LongVU: Spatiotemporal Adaptive Compression for Long Video-Language Understanding

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-08T00:55:50.116104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:55:50.116104Z digest=sha256:bcb79e9c56033716d36a8e52c697d91c65d99907c2059c95e2e2b282e67dea51

Observation afb7bdd5-f0d4-4314-9070-1171eb84c3f4 · outbound

This paper cites LLaVA-MoD: Making LLaVA Tiny via MoE Knowledge Distillation.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models LLaVA-MoD: Making LLaVA Tiny via MoE Knowledge Distillation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-08T00:55:50.120232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:55:50.120232Z digest=sha256:0a6a0380791171795421f9c2f5ebe2f33391545bd79f385702b88f73283be82d

Observation 636baae2-9a92-4e79-84c1-c37fdd4ea012 · outbound

This paper cites LLaVA-Scissor: Token Compression with Semantic Connected Components for Video LLMs.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models LLaVA-Scissor: Token Compression with Semantic Connected Components for Video LLMs

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-08T00:55:50.124410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:55:50.124410Z digest=sha256:67074486f191f1e8777170240c26d27a3d5738b511c459769ccf8927960443b2

Observation 5449c859-83cf-4de4-87c8-6b5172039234 · outbound

This paper cites Dynamic-VLM: Simple Dynamic Visual Token Compression for VideoLLM.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models Dynamic-VLM: Simple Dynamic Visual Token Compression for VideoLLM

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-08T00:55:50.128159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:55:50.128159Z digest=sha256:6197fa52485816599d7f529d0b5800a4255894664dcb43aca3041bbc687d0583

Observation 10cb08f1-eda7-487a-89fd-613ca4a412ab · outbound

This paper cites VideoLLaMB: Long Streaming Video Understanding with Recurrent Memory Bridges.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models VideoLLaMB: Long Streaming Video Understanding with Recurrent Memory Bridges

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-08T00:55:50.132180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:55:50.132180Z digest=sha256:32cde440bff01d5aee4475c6c8bee2cc2a9fd45f397099ea1a66906b8f38c1ae

Observation 637995de-1493-4280-9a81-4a9e02102bf0 · outbound

This paper cites Next-qa: Next phase of question-answering to explaining temporal actions.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models Next-qa: Next phase of question-answering to explaining temporal actions

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-08T00:55:50.136531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:55:50.136531Z digest=sha256:f82152dc5de2948da86266f2a3881d106ecfab4ffccd4a3aa29f4b885f866708

Observation 5c40ef14-ccdd-40e2-ba5e-56e629d10789 · outbound

This paper cites Topv: Compatible token pruning with infer- ence time optimization for fast and low-memory multimodal vision language model.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models Topv: Compatible token pruning with infer- ence time optimization for fast and low-memory multimodal vision language model

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:55:50.620506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T00:55:50.140040Z digest=sha256:aef98d42e064824b350f075df1321e0aa686fa4688b721f6cfc5c763d20b22de

Observation 4f98ddbd-2456-4d52-bcca-2b69b771b108 · outbound

This paper cites Atp-llava: Adaptive token pruning for large vision language models.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models Atp-llava: Adaptive token pruning for large vision language models

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:55:50.609139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T00:55:50.144028Z digest=sha256:31e50b8dd0f675ebd42f32c85552a10f72f6d3cbf952f57873adc195433249ce

Observation ed439344-4c14-4165-ae58-c54fe0510026 · outbound

This paper cites LLaVA-Video: Video Instruction Tuning With Synthetic Data.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models LLaVA-Video: Video Instruction Tuning With Synthetic Data

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-08T00:55:50.147700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:55:50.147700Z digest=sha256:370d4639231f5607c4bda622373a88690cbd5f3b4a41b581e59de9866fdbaf1c

Observation f2b898cd-650a-4113-89c2-8fac10f34d16 · outbound

This paper cites TinyLLaVA: A Framework of Small-scale Large Multimodal Models.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models TinyLLaVA: A Framework of Small-scale Large Multimodal Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-08T00:55:50.151584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:55:50.151584Z digest=sha256:be1aa214e71beb4984bfa8259892d1518d189dbc42b174faff6dba1d8fad7d7b

Observation 89dabbd2-75bb-43a6-9a2e-664507f1caa0 · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-08T00:55:50.155447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:55:50.155447Z digest=sha256:3e732025dee3683defb2e30d64de262e8bf64c1af899168121cf720839df47c9

Observation e854eb82-62e2-4e56-8079-a8cc27b8091f · outbound

This paper cites Also, we use beam size of 1, and the number of maximum new to- kens is capped to 1024.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models Also, we use beam size of 1, and the number of maximum new to- kens is capped to 1024

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:55:50.597966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T00:55:50.159058Z digest=sha256:2932119de15e26bb9993e40d83364f17583065467b747e9490a08dc3e0ffeb1e

Observation f25dd4a7-d912-47b3-91ad-b1909b26b579 · outbound

This paper cites an unresolved cited work.

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-08T00:55:50.586532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-08T00:55:50.162416Z digest=sha256:a4f6171e0892c4028acea75867a270c7e7b5bc91921356293e9217c406f20fbd

Pith citing papers

No inbound Pith citation observations are available.