Pith. sign in

Paper Citation Record · LEDGER

MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 23 inbound Pith citation observations for arXiv:2404.17176.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.17176 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 23 of 23 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:46:40.695883Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T20:38:56.117133Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c968ea4c-2577-487c-92db-2822ddbffd9b · inbound

InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output cites this paper.

InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 136

Resolution
verified exact
arxiv_id, observed 2026-05-17T10:46:28.683009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-17T10:46:28.447347Z digest=sha256:f622a8e0d0dedf423c70c05a4ea474354ae697a1a14704550691300b7bd91ff1

Observation c59ad5fa-7152-4191-9dd5-9fa4c2acdc0d · inbound

TS-LLaVA: Constructing Visual Tokens through Thumbnail-and-Sampling for Training-Free Video Large Language Models cites this paper.

TS-LLaVA: Constructing Visual Tokens through Thumbnail-and-Sampling for Training-Free Video Large Language Models MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T19:03:32.209929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:03:32.209929Z digest=sha256:5d6ebdde4c0502e52443b6838bdb2d688ea9f3fd7558b244d6d869c96a4fe91f

Observation 58efed60-9c42-4c88-8dcb-34eebf9ef1ea · inbound

SAMURAI: Adapting Segment Anything Model for Zero-Shot Visual Tracking with Motion-Aware Memory cites this paper.

SAMURAI: Adapting Segment Anything Model for Zero-Shot Visual Tracking with Motion-Aware Memory MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T18:46:30.629868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:46:30.629868Z digest=sha256:ac5e0e2306667b54e347d2880a86861263b217ab09b0579e47b63a2d25bdda7c

Observation adf57b1c-2f88-4999-9e6e-b3a5261cf747 · inbound

Lucia: A Temporal Computing Platform for Contextual Intelligence cites this paper.

Lucia: A Temporal Computing Platform for Contextual Intelligence MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T17:46:01.885992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:46:01.885992Z digest=sha256:33b2475b0dcc37a68d532962d33557e959d29770088a776602ba97a2096bb0d2

Observation 0818c33c-afa7-4eca-9840-9cc51092e5ad · inbound

SEAL: Semantic Attention Learning for Long Video Representation cites this paper.

SEAL: Semantic Attention Learning for Long Video Representation MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T01:00:10.653819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T01:00:10.653819Z digest=sha256:e5701c4307e77e16d1a8513fe5fdfd9b5fe647af8de1fb28447133eaed2eb865

Observation 47c3e7f5-047f-441f-9cfd-9fe183439f7b · inbound

InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System for Long-term Streaming Video and Audio Interactions cites this paper.

InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System for Long-term Streaming Video and Audio Interactions MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-11T16:56:06.396297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:56:06.396297Z digest=sha256:f28ec071f08a77ff9e3b05c404ab3849ecc7708150e7b036aefc3902bde241eb

Observation 0e915a04-0f0e-4302-bbfd-d9e68fd29c5d · inbound

HumanVBench: Probing Human-Centric Video Understanding in MLLMs with Automatically Synthesized Benchmarks cites this paper.

HumanVBench: Probing Human-Centric Video Understanding in MLLMs with Automatically Synthesized Benchmarks MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-23T07:05:29.177851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-23T07:05:08.716223Z digest=sha256:96f919d2efe42f137cf27b92c3bf03000c0d0b9c252de218b7df8e1bbfcdc9d7

Observation 3789079c-55e0-4dda-adc8-338588fb0e78 · inbound

Online Video Understanding: OVBench and VideoChat-Online cites this paper.

Online Video Understanding: OVBench and VideoChat-Online MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-10T22:57:40.136078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:57:40.136078Z digest=sha256:bfbd57a6c13ea40f230b9f7f5373423a51b7533d3c1456aee7ce980ba3f4cdfa

Observation a4bee17c-1f8f-44a3-8887-3919a4db3b45 · inbound

VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding cites this paper.

VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T01:20:00.203242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-11T01:19:59.603343Z digest=sha256:c0b3d3bf110f332f6ed4c83be35387d1fe88fb4b83c6f2cb7193fb5f5a6ddbf5

Observation 38282128-a542-43ec-880c-c90d314deca6 · inbound

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation cites this paper.

$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-09T21:21:45.766382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T21:21:45.766382Z digest=sha256:dd89da1875056a48459794ed7da3b34729c8e2aac6121fd73c4e3985edeb74e7

Observation 94ccdc71-a984-4057-88bd-0bc16fce18ab · inbound

Video-MMLU: A Massive Multi-Discipline Lecture Understanding Benchmark cites this paper.

Video-MMLU: A Massive Multi-Discipline Lecture Understanding Benchmark MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 113

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:40.695883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:46:40.695883Z digest=sha256:75936d612fa6f6f2a827f16a7a542aae4693f9d4be6c1069b3f2714e634aa5d2

Observation f9824825-2ac7-473d-8253-02560a1e0e0c · inbound

MASR: Self-Reflective Reasoning through Multimodal Hierarchical Attention Focusing for Agent-based Video Understanding cites this paper.

MASR: Self-Reflective Reasoning through Multimodal Hierarchical Attention Focusing for Agent-based Video Understanding MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-16T10:50:21.357351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:50:21.357351Z digest=sha256:3cb54524943c7dbeab244b58ab51e1d7400949500fbd45f163a95e49bbb69ac1

Observation c21d2be4-e3e3-4d33-915e-e43881b92819 · inbound

TEMPURA: Temporal Event Masked Prediction and Understanding for Reasoning in Action cites this paper.

TEMPURA: Temporal Event Masked Prediction and Understanding for Reasoning in Action MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-16T04:17:40.247960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:17:40.247960Z digest=sha256:64a56f33b65600ae16a008dd25a1ad94ca617972ee6a95513511801eb8044170

Observation b73d6246-93f3-4368-88c4-970bc6c5cfa7 · inbound

RAVU: Retrieval Augmented Video Understanding with Compositional Reasoning over Graph cites this paper.

RAVU: Retrieval Augmented Video Understanding with Compositional Reasoning over Graph MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-16T00:01:14.765351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:01:14.765351Z digest=sha256:77e220ddd25550747959f86fe4f26580021d1c6b90a665ff073ae330ac83ec37

Observation 0b115c45-2019-43e6-ad60-3a1f2add900f · inbound

LiveVLM: Efficient Online Video Understanding via Streaming-Oriented KV Cache and Retrieval cites this paper.

LiveVLM: Efficient Online Video Understanding via Streaming-Oriented KV Cache and Retrieval MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-22T14:31:40.772346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-22T14:26:59.015559Z digest=sha256:bb0a8e34ebccb077a3b3702f42e6bca9e9fc008a91e32e2f7fde91bfc01d58c5

Observation a72b4ca9-7698-4895-9ef0-3b6547706f43 · inbound

Movie Facts and Fibs (MF$^2$): A Benchmark for Long Movie Understanding cites this paper.

Movie Facts and Fibs (MF$^2$): A Benchmark for Long Movie Understanding MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:56.970276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:56.970276Z digest=sha256:5800edcd079215a991132a82d62a8a71a81a33b03ddafeedbe160de6fdb017a4

Observation 89b8f273-1564-497c-bb65-37c560ab0e08 · inbound

MANTA: Cross-Modal Semantic Alignment and Information-Theoretic Optimization for Long-form Multimodal Understanding cites this paper.

MANTA: Cross-Modal Semantic Alignment and Information-Theoretic Optimization for Long-form Multimodal Understanding MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T22:00:07.170081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:00:07.170081Z digest=sha256:26e98f5563f62939cf92880300d8f17d53e7db932c4f411367174c15747fc910

Observation c4ebeb15-4443-4c06-9dcd-0a56d5531c28 · inbound

AuroraLong: Bringing RNNs Back to Efficient Open-Ended Video Understanding cites this paper.

AuroraLong: Bringing RNNs Back to Efficient Open-Ended Video Understanding MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:54.353632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:54.353632Z digest=sha256:dd13ddfcf5a10535eb7c438630c8386b9a946ccec2b08e4fcec9468ccac192f4

Observation bcea2d96-461a-428d-bce5-d7f7a744bc27 · inbound

InterAct-Video: Reasoning-Rich Video QA for Urban Traffic cites this paper.

InterAct-Video: Reasoning-Rich Video QA for Urban Traffic MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:52.956161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:53:52.956161Z digest=sha256:09daf7e465261c25b0f073156903d745e7df406f80033b8b7342383768631959

Observation 46d85ce4-206a-4de0-b05c-25ed74047074 · inbound

StreamMem: Query-Agnostic KV Cache Memory for Streaming Video Understanding cites this paper.

StreamMem: Query-Agnostic KV Cache Memory for Streaming Video Understanding MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T17:46:46.573054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:46:46.573054Z digest=sha256:1e4fb304326bb038a88178c9adefebd172febfda5146243b0874f8f8994edeea

Observation 5f6696c8-d6b7-40a4-a677-b56b1c03cd55 · inbound

AdsQA: Towards Advertisement Video Understanding cites this paper.

AdsQA: Towards Advertisement Video Understanding MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-04T20:20:36.871295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:20:36.871295Z digest=sha256:bcae299e5d63f9f5cf8fc1875e580f063154a0ba364906377be3db3938833b5f

Observation f5be6b5e-1180-4d3a-8fdf-bff484473869 · inbound

Don't Pause: Streaming Video-Language Synchrony for Online Video Understanding cites this paper.

Don't Pause: Streaming Video-Language Synchrony for Online Video Understanding MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:07:12.889758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-27T22:11:01.690237Z digest=sha256:b6bb928e24fa2d8d7e956f74a8afb6122ffd5d1ba21d360a83995b652e48ddaf

Observation b17d68cf-e15c-4649-9cce-ea65ece727a3 · inbound

LiveStarPro: Proactive Streaming Video Understanding with Hierarchical Memory for Long-Horizon Streams cites this paper.

LiveStarPro: Proactive Streaming Video Understanding with Hierarchical Memory for Long-Horizon Streams MovieChat+: Question-aware Sparse Memory for Long Video Question Answering

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:38:56.118616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-27T01:12:46.295455Z digest=sha256:49980682dbacc1045ce85e3df8c76827e83256813d13bcfe4a5c73bb34582b3f