Pith. sign in

Paper Citation Record · LEDGER

Flash-VStream: Efficient Real-Time Understanding for Long Video Streams

As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2506.23825.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.23825 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T00:03:35.208400Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-10T12:15:01.137692Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5d958aba-4e18-443b-bce7-01c2b8959940 · inbound

Empowering Nanoscale Connectivity through Molecular Communication: A Case Study of Virus Infection cites this paper.

Empowering Nanoscale Connectivity through Molecular Communication: A Case Study of Virus Infection Flash-VStream: Efficient Real-Time Understanding for Long Video Streams

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-06T00:03:35.208400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:03:35.208400Z digest=sha256:650a6fc90879ffab7b44843386404c48c3eafb5425b3f0d689fe7a25dcd41ab6

Observation 011ffc0d-af10-4df5-b428-909d0d7ea02e · inbound

Delayed Bidirectional Alignment via Disentangled Audio Semantics for Audio-Visual Segmentation cites this paper.

Delayed Bidirectional Alignment via Disentangled Audio Semantics for Audio-Visual Segmentation Flash-VStream: Efficient Real-Time Understanding for Long Video Streams

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-03T14:32:00.456065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T14:32:00.456065Z digest=sha256:7dbe93dee4a69e27b07c10f49b339ba55eac75468815c00970a8936996e17631

Observation 6f3dff6f-4043-41f6-b592-79e133eaebbc · inbound

Mosaic: Cross-Modal Clustering for Efficient Video Understanding cites this paper.

Mosaic: Cross-Modal Clustering for Efficient Video Understanding Flash-VStream: Efficient Real-Time Understanding for Long Video Streams

Reference 26

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T16:10:34.388472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T16:07:36.133404Z digest=sha256:6f552f7f964758d2ba609671b41ade035e01ccd57ae7c17f9226c98ac7675cd4

Observation ce12786c-d594-42aa-8810-0d56971e0a6c · inbound

POINTS-Long: Adaptive Dual-Mode Visual Reasoning in MLLMs cites this paper.

POINTS-Long: Adaptive Dual-Mode Visual Reasoning in MLLMs Flash-VStream: Efficient Real-Time Understanding for Long Video Streams

Reference 112

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:41:03.709919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T15:23:08.671342Z digest=sha256:0c13ea8e361e79731de5262787761e9682f1887f3b45ff5a35798f46f05aec5a

Observation 78041cb7-9628-4162-a2e3-c47795275266 · inbound

OASIS: On-Demand Hierarchical Event Memory for Streaming Video Reasoning cites this paper.

OASIS: On-Demand Hierarchical Event Memory for Streaming Video Reasoning Flash-VStream: Efficient Real-Time Understanding for Long Video Streams

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:56:47.780231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T06:51:52.861981Z digest=sha256:366f8a9c525aa24c8e5328e00ac908beec401b3ce9f31cabc501185c58d00b2a

Observation 44ceb307-4b36-4750-bc72-23d3b53ade81 · inbound

Response-G1: Explicit Scene Graph Modeling for Proactive Streaming Video Understanding cites this paper.

Response-G1: Explicit Scene Graph Modeling for Proactive Streaming Video Understanding Flash-VStream: Efficient Real-Time Understanding for Long Video Streams

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:20:56.393925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T02:30:55.939351Z digest=sha256:18b7d24485c57e648a180709fa88adf69005141eb941e3ffc4c1d5d14cafef45

Observation cf4b5ff5-efc3-4be8-997b-d36a18201cae · inbound

Response-G1: Explicit Scene Graph Modeling for Proactive Streaming Video Understanding cites this paper.

Response-G1: Explicit Scene Graph Modeling for Proactive Streaming Video Understanding Flash-VStream: Efficient Real-Time Understanding for Long Video Streams

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:01:17.733473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-12T03:00:34.728880Z digest=sha256:209fb894ffa8539a59931745e6babbb66625ec79358c82213a941acdbd088f43

Observation 54e3e141-b2f2-4945-be55-c0e5ffa8023f · inbound

MemoryCard: Topic-Aware Multi-Modal Clue Compression for Long-Video Question Answering cites this paper.

MemoryCard: Topic-Aware Multi-Modal Clue Compression for Long-Video Question Answering Flash-VStream: Efficient Real-Time Understanding for Long Video Streams

Reference 67

Resolution
verified exact
arxiv_id, observed 2026-06-28T02:01:29.198189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-28T01:52:33.768494Z digest=sha256:5ceb88ba9b6681bbecf336b0963478fe471cfe2ee3213e4999969c6173b9e56f

Observation 061c197b-c067-4d9a-bdf5-0e5971b2fd3a · inbound

Harnessing Streaming Video in the Wild cites this paper.

Harnessing Streaming Video in the Wild Flash-VStream: Efficient Real-Time Understanding for Long Video Streams

Reference 69

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T22:37:25.733342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T18:47:55.910417Z digest=sha256:a225fbdaf2068b725fab8d3aff0630183aa8dd6f84412ff87a4b9a096f7d63cf