Pith. sign in

Paper Citation Record · LEDGER

HourVideo: 1-Hour Video-Language Understanding

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2411.04998.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.04998 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:45:26.874122Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-19T16:03:08.061136Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation cbe8f332-5b99-4c76-8ab7-33c1155b89e5 · inbound

VideoChat-Flash: Hierarchical Compression for Long-Context Video Modeling cites this paper.

VideoChat-Flash: Hierarchical Compression for Long-Context Video Modeling HourVideo: 1-Hour Video-Language Understanding

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-18T04:02:43.666934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-18T04:02:43.261543Z digest=sha256:9bc5b97bdb482f47a50ca1fe47afcb3f238286b76eb8b03b3175bbae6ba6ba5f

Observation c04387ad-e5a2-4f71-b3a1-cb412a049b00 · inbound

InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling cites this paper.

InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling HourVideo: 1-Hour Video-Language Understanding

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-17T02:52:20.694510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T02:52:20.643070Z digest=sha256:aaf9f65b28b09aa9d766c02aebcd522722e3586026ec0daaa1f38532f550f1ad

Observation 414f0402-51a4-494a-af35-088dcae167d2 · inbound

WorldSense: Evaluating Real-world Omnimodal Understanding for Multimodal LLMs cites this paper.

WorldSense: Evaluating Real-world Omnimodal Understanding for Multimodal LLMs HourVideo: 1-Hour Video-Language Understanding

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:53:26.359538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T05:53:26.066674Z digest=sha256:969082e8582e188dd158ecf850772cc20ecf6f6156d33af2df1df20fcf24ff47

Observation ed00242b-b453-4eac-8022-3632c09b33c5 · inbound

CausalVQA: A Physically Grounded Causal Reasoning Benchmark for Video Models cites this paper.

CausalVQA: A Physically Grounded Causal Reasoning Benchmark for Video Models HourVideo: 1-Hour Video-Language Understanding

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T04:44:58.871782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:44:58.871782Z digest=sha256:988c31e7e8f8da3104494a759cc074fd085ba2ec107dcdd4fe6c459951735b30

Observation 1be4414a-8ae4-4d38-8c76-e76aaa204ed4 · inbound

VideoConviction: A Multimodal Benchmark for Human Conviction and Stock Market Recommendations cites this paper.

VideoConviction: A Multimodal Benchmark for Human Conviction and Stock Market Recommendations HourVideo: 1-Hour Video-Language Understanding

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T10:45:26.874122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:45:26.874122Z digest=sha256:91abefca021dee5a1a4388ee647085d46d27e443a5f66adf3f684656a5290458

Observation 01bcc5ab-e0c7-4408-a60e-48774759a8fe · inbound

Infinite Video Understanding cites this paper.

Infinite Video Understanding HourVideo: 1-Hour Video-Language Understanding

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T18:09:18.211447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:09:18.211447Z digest=sha256:8a8f5fb80c10423b57ec8cea84153e16e6fb2edc084df5861b59c9705a2b4551

Observation 4f25828d-e529-40a3-a226-d630579dc041 · inbound

EgoEverything: A Benchmark for Human Behavior Inspired Long Context Egocentric Video Understanding in AR Environment cites this paper.

EgoEverything: A Benchmark for Human Behavior Inspired Long Context Egocentric Video Understanding in AR Environment HourVideo: 1-Hour Video-Language Understanding

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T06:36:00.359960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T17:35:34.760659Z digest=sha256:e1435c100b5d17df5b2ea9dcd6e166a2902b28993191f9238aa858e1ca7cd1a3

Observation 7d037126-b102-4049-b78a-c9a19bdd683b · inbound

EgoEverything: A Benchmark for Human Behavior Inspired Long Context Egocentric Video Understanding in AR Environment cites this paper.

EgoEverything: A Benchmark for Human Behavior Inspired Long Context Egocentric Video Understanding in AR Environment HourVideo: 1-Hour Video-Language Understanding

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-03T00:18:16.657829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T00:18:16.657829Z digest=sha256:ea98405878cfea24a0a3ced2637242056ab9e037407ae95a8415845ef2d8b0d4

Observation 864588dc-4e06-40b1-b981-5da984992834 · inbound

Minerva-Ego: Spatiotemporal Hints for Egocentric Video Understanding cites this paper.

Minerva-Ego: Spatiotemporal Hints for Egocentric Video Understanding HourVideo: 1-Hour Video-Language Understanding

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-19T16:03:08.063406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T16:02:53.887605Z digest=sha256:ed447666455a44abf5b41295c223421664910e352aa8fb5db16546d907374fb8