Pith. sign in

Paper Citation Record · LEDGER

A Simple LLM Framework for Long-Range Video Question-Answering

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2312.17235.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2312.17235 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:57:53.167691Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T10:48:03.011665Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation df1358d4-cea1-4665-8cf1-34ee1e4e3d41 · inbound

PLLaVA : Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning cites this paper.

PLLaVA : Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning A Simple LLM Framework for Long-Range Video Question-Answering

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:21:58.044985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T20:21:57.873354Z digest=sha256:c7d9c2154b21d2efbc3743cabba7d68c46f46273fd9d0fd85c3b3fee722cd64e

Observation 5fbec168-da9c-4c0f-8ff8-5d4de93b152a · inbound

InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output cites this paper.

InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output A Simple LLM Framework for Long-Range Video Question-Answering

Reference 172

Resolution
verified exact
arxiv_id, observed 2026-05-17T10:46:28.857614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T10:46:28.447347Z digest=sha256:7434089a0bb53af852cb0d1bce680c8991d12913adeb1ba4b87c4df39765d1bc

Observation d88c3ee7-3d28-4528-8fc2-6a4f70996126 · inbound

Four Eyes Are Better Than Two: Harnessing the Collaborative Potential of Large Models via Differentiated Thinking and Complementary Ensembles cites this paper.

Four Eyes Are Better Than Two: Harnessing the Collaborative Potential of Large Models via Differentiated Thinking and Complementary Ensembles A Simple LLM Framework for Long-Range Video Question-Answering

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:53.167691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:57:53.167691Z digest=sha256:81f85f28788138d863057cdc8f7503ac552c45c0cd130f6ccf52de8b1923ff26

Observation 89b6721f-3457-49b3-88e1-24ab7fba52b6 · inbound

HCQA-1.5 @ Ego4D EgoSchema Challenge 2025 cites this paper.

HCQA-1.5 @ Ego4D EgoSchema Challenge 2025 A Simple LLM Framework for Long-Range Video Question-Answering

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T13:52:47.594214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:52:47.594214Z digest=sha256:00ca591838d1ee374fe9ac991ed21c9192dcc1ed0c54add99af5c9553c0c9049

Observation 0cd89e7c-11c0-45cd-8bc6-4f53a320ddf1 · inbound

MUPA: Towards Multi-Path Agentic Reasoning for Grounded Video Question Answering cites this paper.

MUPA: Towards Multi-Path Agentic Reasoning for Grounded Video Question Answering A Simple LLM Framework for Long-Range Video Question-Answering

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T23:29:26.732647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:29:26.732647Z digest=sha256:0ee39020ca41d6db309704dca7812850f76fad5162889afca21ea0b654a6e9be

Observation d6452b5a-89f2-4f23-9ca3-fe3a857482c8 · inbound

AuroraLong: Bringing RNNs Back to Efficient Open-Ended Video Understanding cites this paper.

AuroraLong: Bringing RNNs Back to Efficient Open-Ended Video Understanding A Simple LLM Framework for Long-Range Video Question-Answering

Reference 106

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:56.959356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:56.959356Z digest=sha256:9fa04c52646e1426a5f4c0b28b6a7e8722c32e9268481df738d63a4cb24aba20

Observation cd904f4d-844d-430f-a3a9-23449a5f5372 · inbound

LeAdQA: LLM-Driven Context-Aware Temporal Grounding for Video Question Answering cites this paper.

LeAdQA: LLM-Driven Context-Aware Temporal Grounding for Video Question Answering A Simple LLM Framework for Long-Range Video Question-Answering

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:36.153446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:53:36.153446Z digest=sha256:60f7787eb465141d3c39f9fe34b081785c6aea9fa252f1afc0b177d33bd7f61a

Observation 34e11361-29a3-4833-90e7-58ddcf13c67e · inbound

VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering cites this paper.

VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering A Simple LLM Framework for Long-Range Video Question-Answering

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T04:49:42.312571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:49:42.312571Z digest=sha256:a729d60cd9e39ca1cd36e42328456448134b84f6630116f600d63194ae68e768

Observation 4c6df0f6-e239-4eee-b50c-4bf0e9b1d78f · inbound

CAViAR: Critic-Augmented Video Agentic Reasoning cites this paper.

CAViAR: Critic-Augmented Video Agentic Reasoning A Simple LLM Framework for Long-Range Video Question-Answering

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T21:27:33.767446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T21:27:33.767446Z digest=sha256:6cbbedcf2196e006a4587e2edd40a89747a6958ed7661f0e8f9107b7a1487a85

Observation cea25f6f-a7db-49e6-bb95-189a99b560d1 · inbound

Perceive, Verify and Understand Long Video: Multi-Granular Perception and Active Verification via Interactive Agents cites this paper.

Perceive, Verify and Understand Long Video: Multi-Granular Perception and Active Verification via Interactive Agents A Simple LLM Framework for Long-Range Video Question-Answering

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-18T12:41:22.621623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-18T12:40:19.544260Z digest=sha256:f3c8b43d23d92fbc5f97592654b15345c984fb91423684177a94124aa1633ac6

Observation b693d7da-96f3-4119-aeba-636ba2d90386 · inbound

VIDEOP2R: Video Understanding from Perception to Reasoning cites this paper.

VIDEOP2R: Video Understanding from Perception to Reasoning A Simple LLM Framework for Long-Range Video Question-Answering

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-05-17T22:25:22.682531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T22:24:41.760120Z digest=sha256:790f43e56cacc826ad941f31ca8b86ee524a019a821750576884189104d2b3c5

Observation 39f82adb-aefa-44aa-b2f5-9b44fd964678 · inbound

Towards Sparse Video Understanding and Reasoning cites this paper.

Towards Sparse Video Understanding and Reasoning A Simple LLM Framework for Long-Range Video Question-Answering

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-02T23:33:16.311655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:33:16.311655Z digest=sha256:0a1b45f20c3389ea0ff431d86de4fe46e6ada70ff1fef5d9becd1c3c347a6486

Observation 290ab08c-93d4-4de8-85c2-b2ab3dc664c7 · inbound

Progressive Video Condensation with MLLM Agent for Long-form Video Understanding cites this paper.

Progressive Video Condensation with MLLM Agent for Long-form Video Understanding A Simple LLM Framework for Long-Range Video Question-Answering

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:43:14.889680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T20:40:41.380829Z digest=sha256:a8c6499c10ed675978e5e0f263f3a23ceca86d12ea1c21aa8a283361620b2d10

Observation 488288cd-eb57-4fc0-a317-92bbea7afcb2 · inbound

UpstreamQA: A Modular Framework for Explicit Reasoning on Video Question Answering Tasks cites this paper.

UpstreamQA: A Modular Framework for Explicit Reasoning on Video Question Answering Tasks A Simple LLM Framework for Long-Range Video Question-Answering

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:31:13.224993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T08:41:42.061219Z digest=sha256:ff6806c30382729286d51666de6a0dedada75e90cb468e2e5f6d21304fcc4afc

Observation 058c212c-0e18-4f04-9a59-49c33874108f · inbound

UNIVID: Unified Vision-Language Model for Video Moderation cites this paper.

UNIVID: Unified Vision-Language Model for Video Moderation A Simple LLM Framework for Long-Range Video Question-Answering

Reference 52

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T16:07:09.283746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-27T22:56:26.674841Z digest=sha256:949d235b04c7a4d67cab694f7c3d7a4e604071742bb2b05a9c3c45b7f05714e6

Observation 4760a439-14da-41c5-b5d7-935380a77d67 · inbound

InternVideo3: Agentify Foundation Models with Multimodal Contextual Reasoning cites this paper.

InternVideo3: Agentify Foundation Models with Multimodal Contextual Reasoning A Simple LLM Framework for Long-Range Video Question-Answering

Reference 170

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T10:48:03.012988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-27T09:48:27.652901Z digest=sha256:b3d4b6f9ae7adfe37777a44dbbd6a625cb996c2a3fe7ba64a3da2e45ae6f7319

Observation 5edf0bcf-564c-4d65-9551-ccf929071349 · inbound

Agent-Computer Observation Interfaces Enable Dynamic Computer Use cites this paper.

Agent-Computer Observation Interfaces Enable Dynamic Computer Use A Simple LLM Framework for Long-Range Video Question-Answering

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:04:21.233670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T06:59:20.295818Z digest=sha256:e3de2dad26de0cfcdb6ce691116b556005342c75d2e732f18603cee2aee5378c