Pith. sign in

Paper Citation Record · LEDGER

VideoAgent2: Enhancing the LLM-Based Agent System for Long-Form Video Understanding by Uncertainty-Aware CoT

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2504.04471.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.04471 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:20:13.504824Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T09:39:46.773095Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1bfc5510-31fb-46b2-a18e-27f3ea93819f · inbound

DIVE: Deep-search Iterative Video Exploration A Technical Report for the CVRR Challenge at CVPR 2025 cites this paper.

DIVE: Deep-search Iterative Video Exploration A Technical Report for the CVRR Challenge at CVPR 2025 VideoAgent2: Enhancing the LLM-Based Agent System for Long-Form Video Understanding by Uncertainty-Aware CoT

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T22:20:13.504824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:20:13.504824Z digest=sha256:3bdd36946c27fb7adade12e970ec589f1c9b0c73c5bf932a188ae4025f0427ba

Observation a2adb14f-9d42-4d38-b7e0-922e7efaf176 · inbound

GLANCE: A Global-Local Coordination Multi-Agent Framework for Music-Grounded Non-Linear Video Editing cites this paper.

GLANCE: A Global-Local Coordination Multi-Agent Framework for Music-Grounded Non-Linear Video Editing VideoAgent2: Enhancing the LLM-Based Agent System for Long-Form Video Understanding by Uncertainty-Aware CoT

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:20:51.141817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T19:13:30.400059Z digest=sha256:41fa2dd94938578e523fc4b04515aed5a1fc4fc5511b688335e5d13608c0daf8

Observation 665262fc-14de-4e6b-8a45-0ff88db0a8e8 · inbound

SVAgent: Storyline-Guided Long Video Understanding via Cross-Modal Multi-Agent Collaboration cites this paper.

SVAgent: Storyline-Guided Long Video Understanding via Cross-Modal Multi-Agent Collaboration VideoAgent2: Enhancing the LLM-Based Agent System for Long-Form Video Understanding by Uncertainty-Aware CoT

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:00:48.959183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T20:20:08.590407Z digest=sha256:b4d0a003755df23ff1903dae17284dff021ca853a8c448daba0b3c4f4a419e8a

Observation f6717910-b0a6-45b5-aab0-a00dfcb494e7 · inbound

SagaQA: A Multi-hop Reasoning Benchmark for Long-form Narrative Understanding in TV Series cites this paper.

SagaQA: A Multi-hop Reasoning Benchmark for Long-form Narrative Understanding in TV Series VideoAgent2: Enhancing the LLM-Based Agent System for Long-Form Video Understanding by Uncertainty-Aware CoT

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T03:16:33.481200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T10:14:48.103490Z digest=sha256:8d400a4ef61cb06de55520cd56f24dd8d89cb606791c1d2d3a1f0e09a9e6c3b3

Observation 7c727cf3-ca7f-48f2-8f7b-6cc87e0d11ea · inbound

Watch, Remember, Reason: Human-View Video Understanding with MLLMs cites this paper.

Watch, Remember, Reason: Human-View Video Understanding with MLLMs VideoAgent2: Enhancing the LLM-Based Agent System for Long-Form Video Understanding by Uncertainty-Aware CoT

Reference 182

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:27:15.702666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T22:00:28.350003Z digest=sha256:b31e72847130331ae7cd5493eacacdd7a0d17db1c71346db9fa7e4a71cd6f4f9

Observation 96c1d961-ebf1-443c-916e-2f4b09575ae8 · inbound

CoVStream: Edge-Cloud Collaboration for Understanding of Long Video Streams cites this paper.

CoVStream: Edge-Cloud Collaboration for Understanding of Long Video Streams VideoAgent2: Enhancing the LLM-Based Agent System for Long-Form Video Understanding by Uncertainty-Aware CoT

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T09:39:46.774494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T09:42:58.861599Z digest=sha256:1a40c60003ac5a6da755fd092f47945ca832a3927b9c91cddaff39be00b1a511

Observation 7c456548-7f5e-42d0-bab2-7fc368bcd735 · inbound

EventCoT: Event-centric Video Chain-of-thought for Reasoning Temporal Localization cites this paper.

EventCoT: Event-centric Video Chain-of-thought for Reasoning Temporal Localization VideoAgent2: Enhancing the LLM-Based Agent System for Long-Form Video Understanding by Uncertainty-Aware CoT

Reference 58

Resolution
unresolved
no resolver link, observed 2026-07-11T12:10:52.100410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T12:10:52.100410Z digest=sha256:3d56213a48ebd4a2435912b2ee1243c32910bc57f03436fcb3ada5fcfd9d4162

Observation 9845b1b6-cfa6-4b06-80cf-186d8d6f6ff4 · inbound

AgenticVAU: Multi-Agent Explore-Verify Reasoning for Video Anomaly Understanding cites this paper.

AgenticVAU: Multi-Agent Explore-Verify Reasoning for Video Anomaly Understanding VideoAgent2: Enhancing the LLM-Based Agent System for Long-Form Video Understanding by Uncertainty-Aware CoT

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T12:48:10.918578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T12:48:10.918578Z digest=sha256:658c43851536178ee4d8982efb8ba820075d0b9e8c207f589b53ef53aa07447f