Pith. sign in

Paper Citation Record · LEDGER

Visual Context Window Extension: A New Perspective for Long Video Understanding

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2409.20018.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2409.20018 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T17:46:46.750179Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T10:48:03.220493Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ba361a7a-a74e-498e-8ba7-1aab43771bf1 · inbound

VideoChat-Flash: Hierarchical Compression for Long-Context Video Modeling cites this paper.

VideoChat-Flash: Hierarchical Compression for Long-Context Video Modeling Visual Context Window Extension: A New Perspective for Long Video Understanding

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-18T04:02:43.552982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T04:02:43.261543Z digest=sha256:0320bf3008ffc30ef0648423e73ca1eb6e06e2754fb90689e549ea8e4341b24c

Observation 0dde0fbc-b0b1-4d05-bb63-26e47ffb689c · inbound

StreamMem: Query-Agnostic KV Cache Memory for Streaming Video Understanding cites this paper.

StreamMem: Query-Agnostic KV Cache Memory for Streaming Video Understanding Visual Context Window Extension: A New Perspective for Long Video Understanding

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T17:46:46.750179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:46:46.750179Z digest=sha256:e5fb51ea9e73cf9229d24443761b859eba74af49b6ad808c4fc467083b5c363a

Observation 22e1696a-3703-4675-9506-ef2c6cae2c87 · inbound

DATE: Dynamic Absolute Time Enhancement for Long Video Understanding cites this paper.

DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Visual Context Window Extension: A New Perspective for Long Video Understanding

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T19:28:34.247158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:28:34.247158Z digest=sha256:f7dac86aa37cb3af828b30d1072c55afcb1d1a3b982337b3018ad5655637a83c

Observation fa2f9987-8ba9-487d-9619-d60f58622ae3 · inbound

Event-Causal RAG: A Retrieval-Augmented Generation Framework for Long Video Reasoning in Complex Scenarios cites this paper.

Event-Causal RAG: A Retrieval-Augmented Generation Framework for Long Video Reasoning in Complex Scenarios Visual Context Window Extension: A New Perspective for Long Video Understanding

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:06:13.303063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-08T10:15:15.129358Z digest=sha256:5c73cef05fe0d9596caca71e226ed3ef7a5c96d5fa06a91b1bb08ba6ba00a204

Observation 01534f90-515a-4f47-9d64-a5cc70a0920a · inbound

UNIVID: Unified Vision-Language Model for Video Moderation cites this paper.

UNIVID: Unified Vision-Language Model for Video Moderation Visual Context Window Extension: A New Perspective for Long Video Understanding

Reference 58

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T16:07:09.291604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-27T22:56:26.674841Z digest=sha256:79c62759294b3d6588664a121103a0a6ffa60af29113d99979e3310da522229c

Observation b990bd02-1e6f-434b-bab0-9d981d3092d9 · inbound

From Content to Knowledge: Lightning Fast Long-Video Understanding with Neural Knowledge Representations cites this paper.

From Content to Knowledge: Lightning Fast Long-Video Understanding with Neural Knowledge Representations Visual Context Window Extension: A New Perspective for Long Video Understanding

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:17:57.400734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-27T10:10:37.233160Z digest=sha256:e4450a39904479328c06ae1668d7002263a362d42ac8dac971129865058ad7bd

Observation 0946f4d4-2194-425d-819c-85305bd23e68 · inbound

InternVideo3: Agentify Foundation Models with Multimodal Contextual Reasoning cites this paper.

InternVideo3: Agentify Foundation Models with Multimodal Contextual Reasoning Visual Context Window Extension: A New Perspective for Long Video Understanding

Reference 284

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T10:48:03.221935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-27T09:48:27.652901Z digest=sha256:65741afe466c05dca13ddf1d7c521448f6f8afd3899a8835cc6b9666001ecc61