Pith. sign in

Paper Citation Record · LEDGER

Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 43 inbound Pith citation observations for arXiv:2406.08085.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.08085 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 43 of 43 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T08:12:47.458517Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T20:00:08.182505Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 16ce91fa-49dc-4fb9-85aa-a081ec43c810 · inbound

LongVILA: Scaling Long-Context Visual Language Models for Long Videos cites this paper.

LongVILA: Scaling Long-Context Visual Language Models for Long Videos Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:51:25.463933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-17T03:51:25.396887Z digest=sha256:a40b86f014cf1ff6abae71bdd0ad29e35360f879e0ebf3f4ffa24cff0b7a9d08

Observation fed152d1-acea-4a91-8536-678bbe328e50 · inbound

PPLLaVA: Varied Video Sequence Understanding With Prompt Guidance cites this paper.

PPLLaVA: Varied Video Sequence Understanding With Prompt Guidance Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-23T17:33:15.672665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T17:31:59.030963Z digest=sha256:11ba3d1305994e554e77daaf4ccf233de59a3f917dd81d3fa7347c4b47b3b4d2

Observation 80098db2-8922-4721-9090-bfaa8ada79d8 · inbound

VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding cites this paper.

VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 168

Resolution
verified exact
arxiv_id, observed 2026-05-11T01:20:00.236233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-11T01:19:59.603343Z digest=sha256:792b4e9cddc829df252f881e762c204bcfa5dab82516ab01a04b0b96f41cea43

Observation 9c3d523b-1f4c-4892-95fa-71444da65956 · inbound

LiveVLM: Efficient Online Video Understanding via Streaming-Oriented KV Cache and Retrieval cites this paper.

LiveVLM: Efficient Online Video Understanding via Streaming-Oriented KV Cache and Retrieval Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-22T14:31:40.813023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T14:26:59.015559Z digest=sha256:279436fb015b558edf6c5fa76610facd4bbaa60c637ea26f32be67392fa5be9e

Observation f614c2f9-c1ee-4bd9-92c9-241b1d2d38bb · inbound

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning cites this paper.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 87

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T12:55:40.390683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:0d6db536f7a72b89479dc9846c1317b94eab8add63a038f6fe6a9127fc05c06b

Observation 946eac70-1f6c-4433-aad0-c82da8c1daea · inbound

Vision-Language Memory for Spatial Reasoning cites this paper.

Vision-Language Memory for Spatial Reasoning Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-03T20:15:36.429995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T20:15:36.429995Z digest=sha256:5118e9e67955ca0ec34fecef022c6c71b25c803cab96b59a852e027c32fcff38

Observation 14008280-962a-4104-a8cc-86e1bb31ca92 · inbound

Can Multi-Modal LLMs Provide Live Step-by-Step Task Guidance? cites this paper.

Can Multi-Modal LLMs Provide Live Step-by-Step Task Guidance? Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:39:06.167340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-17T05:36:09.208754Z digest=sha256:4bb57831039cec69c18b5157a33a779826fd1bc44f17f7262d976c98b43092ba

Observation 394e8922-e627-4b86-b997-907e05b8b432 · inbound

Streaming Video Instruction Tuning cites this paper.

Streaming Video Instruction Tuning Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-16T19:48:21.792961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T19:44:11.032898Z digest=sha256:e1eaef07e122b064c14afca174d94042e8ab3b6b36622adacc7e59c200522d38

Observation 22a96c46-ffa6-459c-a1f9-b95b727ad2cb · inbound

HERMES: KV Cache as Hierarchical Memory for Efficient Streaming Video Understanding cites this paper.

HERMES: KV Cache as Hierarchical Memory for Efficient Streaming Video Understanding Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:57:53.864829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T12:55:04.564442Z digest=sha256:0a1eed1a94bd3a77a1b7fbcfbb6c4c50e65cd374819a4deac16b01990b261ab3

Observation 0bf7a826-66ac-4eb0-a31b-6e0dc62ccb89 · inbound

LongVideo-R1: Smart Navigation for Low-cost Long Video Understanding cites this paper.

LongVideo-R1: Smart Navigation for Low-cost Long Video Understanding Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:01:33.505455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T20:01:31.129959Z digest=sha256:9939a3b27228de53349316231ac49f17d1cea5639f8ca61a3159966cf9aff9f6

Observation 5bc8b24f-ac61-4f4b-a48e-e571ca4389d6 · inbound

From Verbatim to Gist: Distilling Pyramidal Multimodal Memory via Semantic Information Bottleneck for Long-Horizon Video Agents cites this paper.

From Verbatim to Gist: Distilling Pyramidal Multimodal Memory via Semantic Information Bottleneck for Long-Horizon Video Agents Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-15T18:10:13.107496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T18:09:59.236030Z digest=sha256:eab8216b7202bd42d55c172a6083e92d39ae3e26b3fdc54d617965f0d652457a

Observation 104221c3-5d21-4f4e-a341-9336cef722da · inbound

An Updated SynthPop Model for Microlensing Simulations I: Model Description & Evaluation cites this paper.

An Updated SynthPop Model for Microlensing Simulations I: Model Description & Evaluation Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-14T22:25:30.596294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T22:25:30.596294Z digest=sha256:51974fb0d53367a42c65d21d9b13afdb01151fb4245bdbc31dcf8754eafd273d

Observation df9ac8cf-62b0-49a1-b05c-90746584bb94 · inbound

Seeing the Scene Matters: Revealing Forgetting in Video Understanding Models with a Scene-Aware Long-Video Benchmark cites this paper.

Seeing the Scene Matters: Revealing Forgetting in Video Understanding Models with a Scene-Aware Long-Video Benchmark Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-05-14T22:08:04.448097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-14T22:05:07.326202Z digest=sha256:3bb0845f020659a1132fa3bceba0c787c1db6ae20e6cef4e4879d0c9ecb79b97

Observation 2583ddc4-2dc7-4abc-af36-05ebbeaae521 · inbound

VSAS-Bench: Real-Time Evaluation of Visual Streaming Assistant Models cites this paper.

VSAS-Bench: Real-Time Evaluation of Visual Streaming Assistant Models Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:25:58.031174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T17:39:30.731688Z digest=sha256:cd59d9bcd57eba7a85c0966cc8d2354f9b62da04a9d40f19e2d062385036e6be

Observation 8cb884b5-c0e4-4c76-b400-7aa8cb670cf2 · inbound

StreamMeCo: Long-Term Agent Memory Compression for Efficient Streaming Video Understanding cites this paper.

StreamMeCo: Long-Term Agent Memory Compression for Efficient Streaming Video Understanding Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:45:50.436845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T18:19:41.543165Z digest=sha256:95669abc117e4567f5aada02f189de8d80dd0fdfd853d28fcd912b25bb403525

Observation 266777e1-7e61-40fb-aea8-674cb00bbeca · inbound

OASIS: On-Demand Hierarchical Event Memory for Streaming Video Reasoning cites this paper.

OASIS: On-Demand Hierarchical Event Memory for Streaming Video Reasoning Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:56:47.788325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T06:51:52.861981Z digest=sha256:b048117257c595f1e4a354b2c8e0b5c4beb1f5e6e7de1d783e9d7db6ade17c5a

Observation 13efc00b-3054-4d1d-8fd7-92eba1529fe7 · inbound

Don't Pause! Every prediction matters in a streaming video cites this paper.

Don't Pause! Every prediction matters in a streaming video Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:41:18.019791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T04:32:01.379605Z digest=sha256:7db4b0f3cab3736cc80abe56e6eaf5918b71728df9355266f377272440306bec

Observation 09c9c0bb-1fff-461f-aa8e-83141dc06e75 · inbound

Decouple and Cache: KV Cache Construction for Streaming Video Understanding cites this paper.

Decouple and Cache: KV Cache Construction for Streaming Video Understanding Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:31:03.562532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:47:54.917408Z digest=sha256:5dd2cbc75bb0417a943f37973b27a4aaeefac4b93b46b979b01ac6e18bbde725

Observation 5cdcc589-c2b1-4b8e-8faa-0b46fd2d4312 · inbound

From Priors to Perception: Grounding Video-LLMs in Physical Reality cites this paper.

From Priors to Perception: Grounding Video-LLMs in Physical Reality Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:21:08.334591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T17:41:23.233366Z digest=sha256:7d1b806e12b32e470f34c09707d786bcb3e51a55c0fcfb94eb11d0fd2eb89590

Observation 212f5093-1618-4c58-a2ad-0068af94c72a · inbound

LATERN: Test-Time Context-Aware Explainable Video Anomaly Detection cites this paper.

LATERN: Test-Time Context-Aware Explainable Video Anomaly Detection Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 39

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T14:25:47.016973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T21:19:12.655706Z digest=sha256:c5ef448e37e2ab155191e81bbf84bb387a8f9bf7d0712a4c7be1a1fdfdb98042

Observation eb91be19-68ab-41e9-ba79-cff4ed9f0561 · inbound

PyraVid: Hierarchical Multimodal Memory for Long-Horizon Video Reasoning cites this paper.

PyraVid: Hierarchical Multimodal Memory for Long-Horizon Video Reasoning Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-20T15:13:24.798248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-20T15:12:00.408851Z digest=sha256:0671084634540ec9b0a9e7d3bc38b8b89c5121a57022b3a4b1ed6ba9bd720a6a

Observation 0c263db3-57e5-4683-92be-a303128f8bbc · inbound

OProver: A Unified Framework for Agentic Formal Theorem Proving cites this paper.

OProver: A Unified Framework for Agentic Formal Theorem Proving Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 80

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T14:48:23.506925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-20T14:43:46.517807Z digest=sha256:23f09a8357243fcf1f6a466766c7d936f3ad875522ffb0478bbe7b8a27af7e0a

Observation 96f99809-c799-4cf6-8451-346d8e4bcfc1 · inbound

Omni-DuplexEval: Evaluating Real-time Duplex Omni-modal Interaction cites this paper.

Omni-DuplexEval: Evaluating Real-time Duplex Omni-modal Interaction Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-20T13:38:19.169529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-20T13:36:44.071188Z digest=sha256:4aea4511434a75734d16e250339b095f2c05c86320626b3248bbae1558b8327e

Observation aaf4586a-8910-4338-80b9-98681305850c · inbound

Omni-DuplexEval: Evaluating Real-time Duplex Omni-modal Interaction cites this paper.

Omni-DuplexEval: Evaluating Real-time Duplex Omni-modal Interaction Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-04T01:19:20.326643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-04T01:11:42.073993Z digest=sha256:5bec1f4922b09b1e5cdb7418df8fd39e190cf8967d8bf917940db65fa4237d51

Observation 6743be0a-b7b7-40fd-9707-80343302ea06 · inbound

IPIBench: Evaluating Interactive Proactive Intelligence of MLLMs under Continuous Streams cites this paper.

IPIBench: Evaluating Interactive Proactive Intelligence of MLLMs under Continuous Streams Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:33:51.028063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T18:24:57.881644Z digest=sha256:f325bb3bd5122a061d021df86d2f00f6805609b9d3906ea21b6b3a8bd16622c2

Observation 2b18fe52-a846-4236-b189-fc792d0a4fb2 · inbound

Linear Scaling Video VLMs for Long Video Understanding cites this paper.

Linear Scaling Video VLMs for Long Video Understanding Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 79

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T23:02:46.258016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T23:00:11.246232Z digest=sha256:1fd21afb35d123c71637439736778c80555d4519030b3aa1b26201d28726cab6

Observation 7356d735-315d-4c30-af27-76028d7db25c · inbound

Zero-Shot 3D Question Answering via Hierarchical View-to-Token Transportation cites this paper.

Zero-Shot 3D Question Answering via Hierarchical View-to-Token Transportation Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 81

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T02:26:27.065722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-28T10:54:02.188634Z digest=sha256:b5df5c38e387acab578dc9a1667fecd20bbeff70811e2f7a1c5f13a8b67c3d2f

Observation cb654ab0-e384-4fd2-9d22-52721b12820f · inbound

OVO-S-Bench: A Hierarchical Benchmark for Streaming Spatial Intelligence in Multimodal LLMs cites this paper.

OVO-S-Bench: A Hierarchical Benchmark for Streaming Spatial Intelligence in Multimodal LLMs Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 67

Resolution
verified exact
arxiv_id, observed 2026-07-02T02:16:27.100649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T11:02:07.122615Z digest=sha256:791d5a30b368b3ee5b14f54bf0944072cb954f7c040c734036d825f2b201ff92

Observation f17c9624-dcc6-4b32-9028-2268977d50fc · inbound

Don't Pause: Streaming Video-Language Synchrony for Online Video Understanding cites this paper.

Don't Pause: Streaming Video-Language Synchrony for Online Video Understanding Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:07:12.826114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T22:11:01.690237Z digest=sha256:a842285eb2b8f86dd44fea0167d541db43495f8841d933cb125b38e1b39d871c

Observation ace06fdb-019e-44c5-aecf-8cefd999604a · inbound

Watch, Remember, Reason: Human-View Video Understanding with MLLMs cites this paper.

Watch, Remember, Reason: Human-View Video Understanding with MLLMs Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:27:15.497535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T22:00:28.350003Z digest=sha256:1b1bda40f89ddc8e7d4e858958ef0749f421ab844a5db4a1633545251d5609ad

Observation 8b545684-c27d-47da-8667-b0cb0d11a8eb · inbound

Streaming Interventions: Can Video Large Language Models Correct Mistakes as They Occur? cites this paper.

Streaming Interventions: Can Video Large Language Models Correct Mistakes as They Occur? Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:57:30.647830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T16:52:22.811857Z digest=sha256:5b77ec35be5b6e7b12a0b133fa0047f5ef6eb54286c44d32bb1326fda3568ef9

Observation 8b069b59-ddbf-4ffd-8919-b3520863a203 · inbound

LiveStarPro: Proactive Streaming Video Understanding with Hierarchical Memory for Long-Horizon Streams cites this paper.

LiveStarPro: Proactive Streaming Video Understanding with Hierarchical Memory for Long-Horizon Streams Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:38:56.143340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T01:12:46.295455Z digest=sha256:f3d4df3b94fe1a1e736b19ca65835bf9d64544c65d24d6c0f9865cbd02610fd6

Observation 587d4ed6-2de8-4cf1-aff5-54b31c752bf1 · inbound

ViCoStream: Streaming VideoLLMs Can Run Beyond 100 FPS with Stage-Wise Coordinated Inference cites this paper.

ViCoStream: Streaming VideoLLMs Can Run Beyond 100 FPS with Stage-Wise Coordinated Inference Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T03:19:29.905981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-26T18:17:53.013043Z digest=sha256:bef703db66d7941c0eb332d377ae8ad9a9dd974bc3323641d1ed128ee1662b6c

Observation 25d70545-f7d4-4401-ad83-bd7f31670d73 · inbound

How Well Can Your Video Model Remember? Measuring Memory-Budget Trade-offs in Long Video Understanding cites this paper.

How Well Can Your Video Model Remember? Measuring Memory-Budget Trade-offs in Long Video Understanding Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T23:39:04.710461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-26T21:51:04.050833Z digest=sha256:3ca25f78c18749a02ed012ab8bcf9e076a3f90f989bf225257b4941930263b1d

Observation 16e9037b-f4fe-44d4-8fa8-a2fee29d1c6c · inbound

video-SALMONN-R$^3$: Learning to ReWatch, ReAsk, and ReAnswer for Efficient Video Understanding cites this paper.

video-SALMONN-R$^3$: Learning to ReWatch, ReAsk, and ReAnswer for Efficient Video Understanding Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-07-04T16:39:58.343073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T00:19:26.153682Z digest=sha256:f118d7a76e5cf1a5b39cdcc3352481a2ba73ec978b6ad4dc2335c043c5485e7e

Observation bf1fbb9e-037d-40dd-9ba6-a6f5309f6bd5 · inbound

Towards a Dynamic and Fixed-budget Memory Bank for Efficient Streaming Video Understanding cites this paper.

Towards a Dynamic and Fixed-budget Memory Bank for Efficient Streaming Video Understanding Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-07-04T20:00:08.184260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-25T20:54:49.319252Z digest=sha256:074a8b0db6424a73230685dbe027b9c483bcd4e298527f604c92592cf92b23bf

Observation 97a1906b-585e-4bfb-9a68-31c654f7fd22 · inbound

Light-Omni: Reflex over Reasoning in Agentic Video Understanding with Long-Term Memory cites this paper.

Light-Omni: Reflex over Reasoning in Agentic Video Understanding with Long-Term Memory Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 53

Resolution
unresolved
no resolver link, observed 2026-07-11T06:35:35.951554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T06:35:35.951554Z digest=sha256:0b1c1288d5fc4e3f8c7f27c843bc94ff744bb1d33e110fe1530aa1375e2503a0

Observation 8226f807-a8ea-4980-a197-0d07ec3918a5 · inbound

Vinci2: Providing Proactive Assistance in Continuous Egocentric Videos cites this paper.

Vinci2: Providing Proactive Assistance in Continuous Egocentric Videos Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 70

Resolution
unresolved
no resolver link, observed 2026-07-14T05:01:06.200663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:01:06.200663Z digest=sha256:89f59c77341797d389725335e885e1f4d7dd5f448f5839315d810e5d5566f0be

Observation 7e5440df-9915-40ca-8717-78f8973438bf · inbound

VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding cites this paper.

VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T00:44:41.589936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:44:41.589936Z digest=sha256:5664340d5c6ae068df456233a6cfa401f65264714ec12a5c81115ff5af24e898

Observation d760ee98-5e90-420d-9bca-e53aecc7cd85 · inbound

Mage-VL: An Efficient Codec-Native Streaming Multimodal Foundation Model cites this paper.

Mage-VL: An Efficient Codec-Native Streaming Multimodal Foundation Model Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 74

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:13.884368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:13.884368Z digest=sha256:70a2a47d0bd59b446c7d390b96be3d6c6664beb1a309ef20a4d951b4551e3936

Observation 1b043da0-8aaa-49d4-99f7-8bb3ff0b33e3 · inbound

ObjectStream: Latent Objects as Memory Anchors for Streaming Video Understanding cites this paper.

ObjectStream: Latent Objects as Memory Anchors for Streaming Video Understanding Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-04T03:21:48.375332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T03:21:48.375332Z digest=sha256:6e0075ad0c190f1028a9b3dd207b3ba2fcbab1f16b0ba5c3d2b3683cca16c74a

Observation 9cea1237-dc36-4c46-8a46-6741e1a482ff · inbound

ViSAGE: Constructing Self-Correcting Memories for Long-Form Video Understanding cites this paper.

ViSAGE: Constructing Self-Correcting Memories for Long-Form Video Understanding Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-03T00:45:21.931265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T00:45:21.931265Z digest=sha256:ce1c68aa1dd376026c6ccc1cf3fc59ed6d2a1bdf72851891f207acfea442150c

Observation 259dbeea-6a1e-41a5-89ad-a27f7e1e4aaf · inbound

GROVE: Growing and Reasoning over Temporally Stratified Memory from Streaming Video Experience cites this paper.

GROVE: Growing and Reasoning over Temporally Stratified Memory from Streaming Video Experience Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T08:12:47.458517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:12:47.458517Z digest=sha256:4448e97d8b1b7cc9bd939d99206a43f944e5bd549af0255aeaf5a58f2246b357