Pith. sign in

Paper Citation Record · LEDGER

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering

As of 5 August 2026, this Paper Citation Record lists 62 of 62 outbound references and 0 inbound Pith citation observations for arXiv:2605.22269.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.22269 v1

Coverage vector

measured 62 of 62 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-22T07:16:33.817183Z

measured 62 of 62 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

62 of 62 outbound references displayed

  • verified exact25
  • verified fuzzy35
  • unresolved1
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f48a51dc-6644-4799-abfc-ded2c76cc6ac · outbound

This paper cites Flamingo: a visual language model for few-shot learning.NeurIPS, 35: 23716–23736.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Flamingo: a visual language model for few-shot learning.NeurIPS, 35: 23716–23736

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.518829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:41d2373753b8a9f2b5d838d085166c6f772abf06b7ad631779dc54f4df307ce3

Observation d79c84ca-d52d-41a6-8ff8-d6cf64f156f7 · outbound

This paper cites Qwen3-VL Technical Report.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Qwen3-VL Technical Report

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-22T07:21:13.171108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:f8b454aee687f1b32c4eae590f06e1325548d24079b0987bbb6e7e407b9e2761

Observation 7313532e-d499-4156-833d-4fa4aebf9593 · outbound

This paper cites Qwen2.5-VL Technical Report.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Qwen2.5-VL Technical Report

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-05-22T07:21:13.211109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:d19c3f6ace9a10167da3940543ae76ec2a463e4ca83efa93f78dce72afb5fcfc

Observation 67f14717-b47b-48f1-bace-e605ee57cb6b · outbound

This paper cites Memory-efficient Streaming VideoLLMs for Real-time Procedural Video Understanding.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Memory-efficient Streaming VideoLLMs for Real-time Procedural Video Understanding

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:21:13.234731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:f81258f92eacbe6e8c11231c9fef47fee5aeedc8253347a92f12cc6f96664113

Observation 08211e91-7187-46fa-98b3-04232d970c51 · outbound

This paper cites An image is worth 1/2 tokens after layer 2: Plug-and-play inference acceleration for large vision-language models.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering An image is worth 1/2 tokens after layer 2: Plug-and-play inference acceleration for large vision-language models

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.635686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:679fa183d0b8bb9b6279186008ec4912152efd7e30692d564c9ab1acab3a93ef

Observation 8556364b-d2cb-49f6-ab4f-836644a08bed · outbound

This paper cites Stream- ingtom: Streaming token compression for efficient video un- derstanding.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Stream- ingtom: Streaming token compression for efficient video un- derstanding

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:21:13.157633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:17ff176bdf938165966e3b8db8b3324650a8c026df64d5a4c239e774a183fc71

Observation e9abc9ae-0c01-40ee-bf20-a09172d12d3a · outbound

This paper cites Gonzalez, Ion Stoica, and Eric P.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Gonzalez, Ion Stoica, and Eric P

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.541807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:07ed86e4dcc996f18471f358776da037cb0f5a3ea0f40ef2b492eb1da2d548ff

Observation bb0f27fe-7b0d-41db-a602-924e7b6910d0 · outbound

This paper cites An algorithm for the machine calculation of complex fourier series.Mathematics of computation, 19(90):297–301.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering An algorithm for the machine calculation of complex fourier series.Mathematics of computation, 19(90):297–301

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.545879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:f461b589792d2db1fed46bcf534d86abeaf6f6d10ec73af5c2950610f0881f32

Observation 5aae5c93-d068-4aee-83e3-c8afed73c690 · outbound

This paper cites Streaming video question-answering with in-context video kv-cache retrieval.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Streaming video question-answering with in-context video kv-cache retrieval

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.650695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:afd224e624a93688d8a6722844178f937b3b7bde984cb8fb0f43e8b56b6c92df

Observation 786c6920-8039-4858-b5a3-56a8ab67c4d3 · outbound

This paper cites The llama 3 herd of models.arXiv e-prints, pages arXiv–2407.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering The llama 3 herd of models.arXiv e-prints, pages arXiv–2407

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.534445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:8b40128695831de6e513858e07522371748cda92034ac8bbd28f38fd0b1ce296

Observation a1511583-0714-4d33-8a52-8e03951bab0d · outbound

This paper cites Videoagent: A memory-augmented mul- timodal agent for video understanding.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Videoagent: A memory-augmented mul- timodal agent for video understanding

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.530357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:e408dc5bf6af3199c4743291c3fa2964d316a52863b94b79f57ee7e4ef85e722

Observation 9afe74f7-d43d-481e-be94-175d877b771d · outbound

This paper cites an unresolved cited work.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Unresolved cited work

Reference 12

Resolution
parse uncertain
raw_fallback, observed 2026-05-22T07:21:14.538100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:f417bc0c0f559f1a8eca826598476e69db019dbc2ce958428ae21aba2506191e

Observation 7348c682-6026-4180-b4d9-4720b17de7a6 · outbound

This paper cites Video-mme: The first-ever comprehensive evaluation benchmark of multi-modal llms in video analysis.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Video-mme: The first-ever comprehensive evaluation benchmark of multi-modal llms in video analysis

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.550053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:946da232ca825c5c03abdb871b6eb88a023396ff0717d3cc94b3b917ae049add

Observation fd5bebd7-5cbe-43d4-8f70-f5aacd5cea0c · outbound

This paper cites Not all heads matter: A head-level kv cache compression method with integrated retrieval and reasoning.ICLR.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Not all heads matter: A head-level kv cache compression method with integrated retrieval and reasoning.ICLR

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.522477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:e581d4a373ff7ce65d659f31c62310215b44ed08264092c24571dacbef5ec58e

Observation 66ff5f31-3f37-4b67-bec1-002ccbfd0d21 · outbound

This paper cites Kvquant: Towards 10 million context length llm inference with kv cache quantization.Advances in Neural Information Processing Systems, 37:1270–1303.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Kvquant: Towards 10 million context length llm inference with kv cache quantization.Advances in Neural Information Processing Systems, 37:1270–1303

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.526751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:647c46be07a38cab40fbe4b12679dff009332fe3cc412f3f17ed8f1a9e3b14a6

Observation fdcbb041-873e-42a3-ad48-4a58e48cb870 · outbound

This paper cites GPT-4o System Card.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering GPT-4o System Card

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-05-22T07:21:13.253524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:741454395a90a72003cb7da2af28f241f958310b8875cc085210a3f4a50445af

Observation d7b9526d-fd9b-4e8f-bc47-5e37cdc0777b · outbound

This paper cites FastKV: Decoupling of Context Reduction and KV Cache Compression for Prefill-Decoding Acceleration.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering FastKV: Decoupling of Context Reduction and KV Cache Compression for Prefill-Decoding Acceleration

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-22T07:21:13.141948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:69d1466b7abb9de8c256682520d92be0a78cbd3f5858eafbf448be9cbe65ca4b

Observation 1198858a-1275-4c0f-8ccb-ebcea1cc33e3 · outbound

This paper cites Freqkv: Frequency domain key- value compression for efficient context window extension.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Freqkv: Frequency domain key- value compression for efficient context window extension

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:21:13.147333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:de8cbd53d229172fd8ecec0d42ce90788021b1596d2168449fd1c499ecd3eee2

Observation c32ad5e0-704f-41ed-9201-ea322d728f4b · outbound

This paper cites Infinipot-v: Memory-constrained kv cache compres- sion for streaming video understanding.NeurIPS.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Infinipot-v: Memory-constrained kv cache compres- sion for streaming video understanding.NeurIPS

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.553283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:67312bc410ec90537bbb6d377614cf68f1cdb53ad0dfc724bb06d3d1b07ba301

Observation 8c3d3d4d-88e4-4f4f-9529-cf6f9424fe98 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering LLaVA-OneVision: Easy Visual Task Transfer

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-05-22T07:21:13.166616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:6bef505badb14cfb721cba298155cc3d3e898efb25024010ac78723af25305d6

Observation c5799eed-7eb5-429c-a6ed-600f4c0a464f · outbound

This paper cites Mvbench: A comprehensive multi-modal video understand- ing benchmark.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Mvbench: A comprehensive multi-modal video understand- ing benchmark

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.556925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:6a95e8dc5ac8150dd83b996df65a9326d88189b7f4c5f2612f5cf9eec5a81c1a

Observation 4900905f-d1b9-4830-9a55-451bf3df74f9 · outbound

This paper cites Video-LLaVA: Learning United Visual Representation by Alignment Before Projection.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Video-LLaVA: Learning United Visual Representation by Alignment Before Projection

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-05-22T07:21:13.205540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:db6649cd9dde5262909e6743aa9e43c96467c21c2129f10c27154251ec8b0474

Observation ff030c80-ab6b-4b45-ba45-c3975a3dddef · outbound

This paper cites StreamingBench: Assessing the Gap for MLLMs to Achieve Streaming Video Understanding.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering StreamingBench: Assessing the Gap for MLLMs to Achieve Streaming Video Understanding

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:21:13.153199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:f5931435eba80b0cf234e5af6e624999cd65a06a3c2487a9301567c8eb882a12

Observation 20bf36fa-d7e9-49b5-b497-d4e95c941b42 · outbound

This paper cites Freekv: Boosting kv cache retrieval for efficient llm inference.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Freekv: Boosting kv cache retrieval for efficient llm inference

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:21:13.191012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:fa518bdab0f8729575869a7196fa93dcf5d00363dc7bf6b3a2bad835f7d0f54d

Observation 8ddbfde8-7c09-4cdb-8d23-7e42aa394cc4 · outbound

This paper cites Cachegen: Kv cache com- pression and streaming for fast large language model serv- ing.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Cachegen: Kv cache com- pression and streaming for fast large language model serv- ing

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.643410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:e9c28967b9697548c99fe18330278cbb84dc9e561c837a7a8ae26d3b26231f9e

Observation 7969c3f8-41b1-48e7-8a5e-eca94aee438c · outbound

This paper cites Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-05-22T07:21:13.243711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:e805e0704582d6aa788afb9a3791721bc1b4b7b00c5bf86ec7dc11980bceac6b

Observation 6a612dde-440f-486e-af3c-80e3864da606 · outbound

This paper cites Egoschema: A diagnostic benchmark for very long- form video language understanding.Advances in Neural In- formation Processing Systems, 36:46212–46244.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Egoschema: A diagnostic benchmark for very long- form video language understanding.Advances in Neural In- formation Processing Systems, 36:46212–46244

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.631955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:03cc114bd274d93a029e7444180fb31107d9d45e54324b87c8c8be6d10f86fed

Observation 285586a3-6d54-43cf-a36a-5e7003746a77 · outbound

This paper cites Morevqa: Exploring modular reason- ing models for video question answering.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Morevqa: Exploring modular reason- ing models for video question answering

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.628333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:81add82c935a9a25eaa1502f421527537932137312afff34ec918e13931059e8

Observation 9d48b7c7-e9dc-43fc-9759-88459a5e170f · outbound

This paper cites LiveVLM: Efficient Online Video Understanding via Streaming-Oriented KV Cache and Retrieval.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering LiveVLM: Efficient Online Video Understanding via Streaming-Oriented KV Cache and Retrieval

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-05-22T07:21:13.215686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:912ed31bf44be7c58264e5fee649b5e3986497aab803e49810388f4ceb3c41fd

Observation b7e7ee0a-073b-43ac-813f-daa36f9de4f7 · outbound

This paper cites Streaming long video understanding with large language models.NeurIPS, 37: 119336–119360.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Streaming long video understanding with large language models.NeurIPS, 37: 119336–119360

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.560478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:cff32f50625a510e57deff91ef5dc0d71db95890e2cfd5ca2bd658ce22512e91

Observation d1aca159-7f11-4ca3-8591-21c2e03aa2f3 · outbound

This paper cites Dispider: Enabling video llms with active real-time interaction via dis- entangled perception, decision, and reaction.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Dispider: Enabling video llms with active real-time interaction via dis- entangled perception, decision, and reaction

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.621147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:22a73c4c82e9b7af4283866159bbb2d5f67c4733e5da15dba49efdda77366a28

Observation 43b82120-628a-41fc-8396-d8de75b65ee2 · outbound

This paper cites Question- answering dense video events.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Question- answering dense video events

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.624663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:8b05a884caf44ee3d5dc3af10bf81392ce548773778879b8dbfa72861f91abf0

Observation f4a8f6fd-ce67-428a-9ae4-8bd9fe2216e0 · outbound

This paper cites LongVU: Spatiotemporal Adaptive Compression for Long Video-Language Understanding.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering LongVU: Spatiotemporal Adaptive Compression for Long Video-Language Understanding

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-05-22T07:21:13.220539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:1f74a88c8009f0c2b41e538aa9c8ab4a80975e5e89eb68d619291d0815825ac2

Observation f33e1d2f-db16-4f3f-b3d5-1382c9f93e9e · outbound

This paper cites Video-xl: Extra-long vision language model for hour-scale video understanding.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Video-xl: Extra-long vision language model for hour-scale video understanding

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.606203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:8199b7fafc4b1b767277aa67c9337d1411aa506e8668c7b36150e678b45cb6a8

Observation 0af27a94-9b11-4271-95be-303332d1520b · outbound

This paper cites Moviechat+: Question-aware sparse memory for long video question answering.IEEE Transac- tions on Pattern Analysis and Machine Intelligence.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Moviechat+: Question-aware sparse memory for long video question answering.IEEE Transac- tions on Pattern Analysis and Machine Intelligence

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.610239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:52b0d4ece17aebe9ef9ff050e618863e345aa4afd84675a444179c53519172cb

Observation be3c5e96-0714-44b6-a02a-6f6bdeb68a4d · outbound

This paper cites Razorattention: Ef- ficient kv cache compression through retrieval heads.ICLR.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Razorattention: Ef- ficient kv cache compression through retrieval heads.ICLR

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.617433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:49df9fff11a829878d2764b6c12bd973585d7e7c6fa901e5bbfd3cf8bb17656a

Observation 56721c51-539e-4a84-83a1-7d1801205009 · outbound

This paper cites Fier: Fine-grained and efficient kv cache retrieval for long-context llm inference.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Fier: Fine-grained and efficient kv cache retrieval for long-context llm inference

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:21:13.225391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:f12cd8c038ddf46a0bd1b3561d0f996f9c80801757d81798120756fbbdf81677

Observation cf529906-fa32-4aab-b958-28592168f03a · outbound

This paper cites InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-05-22T07:21:13.239247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:ee80213f47cfc179d798e43ad2eebbe6b6adb5633ac08d16d1996b45401729ff

Observation 2dda01c6-f905-474d-b6e8-4ed73230c145 · outbound

This paper cites Videoagent: Long-form video understanding with large language model as agent.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Videoagent: Long-form video understanding with large language model as agent

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.613815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:b9541b51b50f3559216d6cdfbdc1652d5c562d4a7da2dad5b547bb8bf7597e95

Observation bf36e788-09c0-43e7-8528-2b51fd61e1ed · outbound

This paper cites InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-05-22T07:21:13.185135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:6e6dcc37e0417940839426adf354a6108c2b4276fe0e39fdab9ed4720f745661

Observation fcb46db8-ca8e-4b40-8810-bbc67439016b · outbound

This paper cites Videochat-a1: Thinking with long videos by chain-of-shot reasoning.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Videochat-a1: Thinking with long videos by chain-of-shot reasoning

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:21:13.248852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:c746520eb049d731e2ef77284d4701a9152bf91d1a7975ccdae992671c964abb

Observation c011934e-e4a6-4bcb-b1c5-839b29a33695 · outbound

This paper cites Videotree: Adaptive tree-based video representation for llm reasoning on long videos.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Videotree: Adaptive tree-based video representation for llm reasoning on long videos

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.602442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:eb129f10d05d131f466fb97c078ea2b6f56e2f0204c48301392d270258d18e09

Observation d0d419aa-88e2-4df5-af3e-3e58681b2410 · outbound

This paper cites Longvlm: Efficient long video understand- ing via large language models.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Longvlm: Efficient long video understand- ing via large language models

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.598984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:0405244eebc5c660e3948fa3988bb4b52264ae63232a1ab8ef635ff3a08907c0

Observation 69ea27b5-249d-4e79-93ec-380aaa7089fd · outbound

This paper cites Efficient streaming language models with attention sinks.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Efficient streaming language models with attention sinks

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.589088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:6dd15826e2d20d716dc08a6993c93a9a5ca9b4ed9dec0c687977695718ed32e1

Observation 97b736e7-2cfd-4019-a811-bd6572a6b774 · outbound

This paper cites Next-qa: Next phase of question-answering to explaining temporal actions.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Next-qa: Next phase of question-answering to explaining temporal actions

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.585766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:54630252647bf2b2785e64f21b4f2a48c26328be26da64ed3547e385981a5c09

Observation 551801b3-427e-4510-8928-626f655cc50f · outbound

This paper cites Videoqa in the era of llms: An empirical study.International Journal of Computer Vi- sion, 133(7):3970–3993.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Videoqa in the era of llms: An empirical study.International Journal of Computer Vi- sion, 133(7):3970–3993

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.592604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:556c5986a368cc78244838ca8abe5facd4be18a12104005b50a666a5391b645c

Observation 291d701a-1627-4e1f-af12-8960d57ab3e1 · outbound

This paper cites Unleashing the power of llms for medical video answer localization.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Unleashing the power of llms for medical video answer localization

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.578509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:81434436b03e2f8bf5d4c5a9e4c75404e2f09b2547bc7b97986b3cf1251470d3

Observation b49c3d7f-a602-4830-a94c-533543d45598 · outbound

This paper cites Video question answer- ing via gradually refined attention over appearance and mo- tion.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Video question answer- ing via gradually refined attention over appearance and mo- tion

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.573793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:e3e3263ec539c1ae6e4b8300a1b7846cad8b693268af4f6bf2586288d00ccefe

Observation 55609626-adcd-46cd-88ff-eb6783f7d384 · outbound

This paper cites PLLaVA : Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering PLLaVA : Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning

Reference 49

Resolution
verified exact
local_arxiv, observed 2026-05-22T07:21:13.229626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:2761e61ced98ed3b261b286811ce3005ea7a05696b7b4805465828042c6edf3a

Observation cd37c81d-61b9-4404-a76b-cf719321fb94 · outbound

This paper cites StreamingVLM: Real-Time Understanding for Infinite Video Streams.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering StreamingVLM: Real-Time Understanding for Infinite Video Streams

Reference 50

Resolution
verified exact
local_arxiv, observed 2026-05-22T07:21:13.201067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:ab9e7e795121ff8ca5ce2e8f0bb60ed379f5d7a75fd67c3f355c78ea430bf4f1

Observation 99402878-9b5f-4874-aa84-faf21aaf5fc9 · outbound

This paper cites Qwen3 Technical Report.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Qwen3 Technical Report

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-05-22T07:21:13.180275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:464a39c9e0b8ea924d6845e5585a2b1b67e542fae2bc861c34d4a55e1c92c349

Observation 6628c57b-9336-400a-b86a-95d6822793db · outbound

This paper cites StreamMem: Query-Agnostic KV Cache Memory for Streaming Video Understanding.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering StreamMem: Query-Agnostic KV Cache Memory for Streaming Video Understanding

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:21:13.162351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:4d1de6eddc055847976fb464e08c62c9c7c8ab5b6df597e8c0ca8be197d8556d

Observation 1e15930b-668f-4629-b693-018775e1bacf · outbound

This paper cites Activitynet-qa: A dataset for understanding complex web videos via question answering.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Activitynet-qa: A dataset for understanding complex web videos via question answering

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.570323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:4466029ade744359ed48fdab8831c8d3d0dc779d979a6546553f93763651c901

Observation 51f1d8a0-4f7e-4669-b215-4aa338b27296 · outbound

This paper cites Socratic models: Composing zero-shot multimodal reasoning with language.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Socratic models: Composing zero-shot multimodal reasoning with language

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.581986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:9d091b7e8d02557c707dbdebaa210876bbdc5ae8929e49b88a8c1589820a3be9

Observation a4b2eb10-8254-4afe-9eb6-bf7a245d332a · outbound

This paper cites VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding

Reference 55

Resolution
verified exact
local_arxiv, observed 2026-05-22T07:21:13.175600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:6d88f580c5339b1cfc574f26f7fddad6bcb377c66d77b9900faf020a4507eed6

Observation 3bacab6a-b5c9-4eed-8c17-e716b8d675c1 · outbound

This paper cites A simple llm framework for long-range video question-answering.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering A simple llm framework for long-range video question-answering

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.566770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:7dbbdc915f85ff3eb63c696ad5730982a2ef5adbdbf20133cc7f96e5df64ca58

Observation 212bcabf-15f2-42fc-9d1f-28eae45cda23 · outbound

This paper cites Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-05-22T07:21:13.195740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:667961cd9f648e4a07af51a1bb6326a829d984f0c1d1e393fd4cca14b0a2ef04

Observation d250c4d7-474d-49d6-9da9-2c345d28a825 · outbound

This paper cites Flash-vstream: Memory- based real-time understanding for long video streams.ICCV.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Flash-vstream: Memory- based real-time understanding for long video streams.ICCV

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.595603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:c073d9a98601c28cee5e5a5fb6fb5ad2ace5539da2d83ce34361822f7fc12c36

Observation cacf7c82-db82-4615-8625-563b629ee5cb · outbound

This paper cites Long Context Transfer from Language to Vision.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Long Context Transfer from Language to Vision

Reference 59

Resolution
verified exact
local_arxiv, observed 2026-05-22T07:21:13.259289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:39a7246ba3e52700a63ffcfdea5e094e4cb53d614bc5d0845d50113cda03988e

Observation ad2eac76-8dcb-4c7f-a80c-2cb14fce30c7 · outbound

This paper cites Mlvu: Benchmarking multi-task long video understanding.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Mlvu: Benchmarking multi-task long video understanding

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.655074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:211a12362924442fd79126d7c9990cd9fdd7b2844051192797cf6018a1d7e555

Observation 461620b9-6954-48d2-98df-5faed5f8a491 · outbound

This paper cites What is the person holding right now?.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering What is the person holding right now?

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T07:21:14.647143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:c079f4d486371712d4307a2860cdd42ef7e4187a0a4c8451a309021e54564f77

Observation 8602b304-783d-4946-b1a4-d8ef17a4e380 · outbound

This paper cites an unresolved cited work.

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering Unresolved cited work

Reference 62

Resolution
unresolved
raw_fallback, observed 2026-05-22T07:21:14.563576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T07:16:33.817183Z digest=sha256:7dd5862c1b109a9541d6d096b7ec62f52dd9a9681243731bdb47b750cea129fa

Pith citing papers

No inbound Pith citation observations are available.