Pith. sign in

Paper Citation Record · LEDGER

Think in Sets for Streaming Video Token Compression

As of 19 August 2026, this Paper Citation Record lists 16 of 16 outbound references and 0 inbound Pith citation observations for arXiv:2608.01169.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.01169 v1

Coverage vector

measured 16 of 16 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T00:31:40.869774Z

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

16 of 16 outbound references displayed

  • verified exact0
  • verified fuzzy4
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a5d372ed-84a9-4298-b130-3eafc0ad2260 · outbound

This paper cites Feige, U.

Think in Sets for Streaming Video Token Compression Feige, U

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T00:31:39.766617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:31:39.766617Z digest=sha256:c1dc2208b9a7aac0c3c5937c981e88090f0ccc8d6e56c12a1926d363a5e0ffa2

Observation c820a7ff-80a7-4b8e-a532-5ad41824e789 · outbound

This paper cites InICASSP 2026-2026 IEEE Interna- tional Conference on Acoustics, Speech and Signal Process- ing (ICASSP), 12147–12151.

Think in Sets for Streaming Video Token Compression InICASSP 2026-2026 IEEE Interna- tional Conference on Acoustics, Speech and Signal Process- ing (ICASSP), 12147–12151

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:31:42.179603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T00:31:40.050765Z digest=sha256:3549d17f5e2419112a1a0116396e01233b756b4a29a9353bb8cec4a6394e4e92

Observation d49abe4c-8398-4526-b529-33bf2889eb50 · outbound

This paper cites StreamChat: Chatting with Streaming Video.

Think in Sets for Streaming Video Token Compression StreamChat: Chatting with Streaming Video

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T00:31:40.113571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:31:40.113571Z digest=sha256:74e8235ddaa0857d699b3574ae72ce678608a2eeb4f2ca75e732c05faf602b4c

Observation 3021d23c-a052-4835-b8f5-6f398ec8abfb · outbound

This paper cites InProceedingsofthe2025 ConferenceonEmpiricalMethodsinNaturalLanguagePro- cessing, 1910–1924.

Think in Sets for Streaming Video Token Compression InProceedingsofthe2025 ConferenceonEmpiricalMethodsinNaturalLanguagePro- cessing, 1910–1924

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:31:42.032678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T00:31:40.221854Z digest=sha256:dfb73aa2e72020fb205924efd77f27a9ed87b0e0f3774901185c6bb2bf0df958

Observation 38dd63f1-5e84-4aa2-a60b-d5e7b1bf0f86 · outbound

This paper cites LiveVLM: Efficient Online Video Understanding via Streaming-Oriented KV Cache and Retrieval.

Think in Sets for Streaming Video Token Compression LiveVLM: Efficient Online Video Understanding via Streaming-Oriented KV Cache and Retrieval

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T00:31:40.388286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:31:40.388286Z digest=sha256:6d164a5c2666e21d213a39029286f7a1c8e7fae8d944286bb765918255a27fa8

Observation 8343e107-5362-41a2-8b14-543c887f05a3 · outbound

This paper cites Ren,S.;Chen,S.;Li,S.;Sun,X.;andHou,L.2023.

Think in Sets for Streaming Video Token Compression Ren,S.;Chen,S.;Li,S.;Sun,X.;andHou,L.2023

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:31:41.595601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T00:31:40.464213Z digest=sha256:604054a70144e5350588a26b739c70bacb86ee47927fbb7e1a8b88732170e086

Observation e8cc92f2-e08c-48c4-a469-0c38834bf70e · outbound

This paper cites StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling.

Think in Sets for Streaming Video Token Compression StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T00:31:40.601037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:31:40.601037Z digest=sha256:a15b138666c1138527cb9605528207decdca893e9b2d913871782ac44e3c6e7e

Observation 86841120-403d-4a15-9817-168607e12b99 · outbound

This paper cites Yang,S.;Chen,Y.;Tian,Z.;Wang,C.;Li,J.;Yu,B.;andJia, J.

Think in Sets for Streaming Video Token Compression Yang,S.;Chen,Y.;Tian,Z.;Wang,C.;Li,J.;Yu,B.;andJia, J

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T00:31:40.688248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:31:40.688248Z digest=sha256:688a449678a18c66e9f07bcec6f3d980edcbcf3bca797f38023d81925da74b6a

Observation 4be5c251-8887-4fa8-8c21-c79ef554a716 · outbound

This paper cites VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding.

Think in Sets for Streaming Video Token Compression VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T00:31:40.780394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:31:40.780394Z digest=sha256:90ef7a407bac92c4d54ecde9af64100ed177001c9aaebfe0874c2238f7a79fb8

Observation e094ca18-d2e0-4d3e-8437-d653e139b5d3 · outbound

This paper cites LLaVA-Video: Video Instruction Tuning With Synthetic Data.

Think in Sets for Streaming Video Token Compression LLaVA-Video: Video Instruction Tuning With Synthetic Data

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T00:31:40.869774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:31:40.869774Z digest=sha256:8285cbd408360127f4f11615bb473448e5fa96f9109b4447b61da02212390d10

Observation 90034daa-eba1-4fcb-976d-f84f935654c8 · outbound

This paper cites an unresolved cited work.

Think in Sets for Streaming Video Token Compression Unresolved cited work

Reference 1998

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:31:42.452559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T00:31:39.838004Z digest=sha256:01c41e85fe7b2703a33199c76fc547d979bb65b27de64264088615160d8a945a

Observation ac560be2-623d-42f1-83e5-80d01c5921d6 · outbound

This paper cites InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling.

Think in Sets for Streaming Video Token Compression InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-06T00:31:40.530514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:31:40.530514Z digest=sha256:6de2661ac56fd3b4ecb541a396085196372a80e3bb51e5792f026a875c36c8cf

Observation 3328c7ea-0883-4a95-833e-b9e1d93d4550 · outbound

This paper cites Nemhauser,G.L.;Wolsey,L.A.;andFisher,M.L.1978.

Think in Sets for Streaming Video Token Compression Nemhauser,G.L.;Wolsey,L.A.;andFisher,M.L.1978

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:31:41.815286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T00:31:40.288189Z digest=sha256:4bd0d1e1cbeaa3a86c047ebf42bbf91e1d1decc5530f828725bb362b6869e44b

Observation 93842de1-11d0-4924-ab15-a4af9ce683a7 · outbound

This paper cites InEuropean Conference on Computer Vision (ECCV), 19–35.

Think in Sets for Streaming Video Token Compression InEuropean Conference on Computer Vision (ECCV), 19–35

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T00:31:39.647181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:31:39.647181Z digest=sha256:d7727c380d65c5df1edf7d04627bd8a8eebe716b0fba38a0f52cc0ab65f3c307

Observation 7a6dfe20-6b7d-4808-9ea1-bb9f7a6bb7c5 · outbound

This paper cites Qwen2.5-VL Technical Report.

Think in Sets for Streaming Video Token Compression Qwen2.5-VL Technical Report

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T00:31:39.567984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:31:39.567984Z digest=sha256:eb56f6d8068bba42fa2d2b37c2ab01b4ae22cef5faaa978aa601886779d4353c

Observation e6144178-5a93-448c-b4e6-19e8bdefe58d · outbound

This paper cites Moving Beyond Diversity: Visual Token Pruning as Subspace Reconstruction for Efficient VLMs.

Think in Sets for Streaming Video Token Compression Moving Beyond Diversity: Visual Token Pruning as Subspace Reconstruction for Efficient VLMs

Reference 2026

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T00:31:41.236232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T00:31:39.923680Z digest=sha256:24680b43805b53669c541a204483496e994909bf316f7b9cd2f17020bd13b11d

Pith citing papers

No inbound Pith citation observations are available.