Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T04:49:42.909084Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 0 inbound Pith citation observations for arXiv:2508.03039.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T04:49:42.909084Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
53 of 53 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d0fc53e2-d681-4262-9f49-feb7351126b9 · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering The IKEA ASM Dataset: Understanding People Assembling Furniture through Actions, Objects and Pose
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8916b5d6-5696-4ab6-85bb-0c4570017a2e · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dbbd0beb-e5af-453c-862f-f1b19e196efe · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering A Short Note on the Kinetics-700 Human Action Dataset
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 457862a8-afb3-48f7-85d2-efa803b70330 · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering ShareGPT4Video: Improving Video Understanding and Generation with Better Captions
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ed2ed85-0e9d-4dc1-a83a-a32078302cf0 · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering Enhancing Long Video Understanding via Hierarchical Event-Based Memory
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e691ea34-fc65-4152-ac4c-de2035cca43e · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a74b6568-0c14-4b40-bcfd-65d75c08ab60 · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 59dd60d2-be20-446d-ac9c-a41ce768a327 · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering Unresolved cited work
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d8855dac-cc08-4c80-8e64-ec425fe59203 · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39f00fc1-9517-4c66-b420-af0ad379cf9d · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c12f38d-35bd-4389-bdcc-01ab01d93e16 · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering Unresolved cited work
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c69bb0d7-baac-4ee6-80ee-9b06a3803f67 · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering Video ReCap: Recursive Captioning of Hour-Long Videos
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ada452eb-d963-4838-8da9-cf70e5235c73 · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering BIMBA: Selective-Scan Compression for Long-Range Video Question Answering
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4bcf83e-1057-4518-8e37-70fa5b0f5f89 · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering Unresolved cited work
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3abd9c83-6526-4618-8dd2-b7e698badfed · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering LLaVA-OneVision: Easy Visual Task Transfer
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb82cd92-7085-45ee-9194-c19c7c338fa8 · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering Unresolved cited work
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6aafc0f-31fb-4274-a25b-06f213b47ac4 · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering Unresolved cited work
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 215cecf1-67bb-41ca-83fc-6a30a3038aa5 · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering VideoChat-Flash: Hierarchical Compression for Long-Context Video Modeling
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 880ab0e9-887c-47b5-bbbd-b542464beaca · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering Unresolved cited work
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4be1116f-7d98-4a42-97ed-97c8dac97951 · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering Commonsense Video Question Answering through Video-Grounded Entailment Tree Reasoning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1fe38cbe-5629-477a-b57f-8a8f54bea77d · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering Unresolved cited work
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43064c23-18b9-4948-b091-01583c16cbb1 · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d55e937b-66a2-44c1-97c4-9820b5cc216c · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f6d41882-8c87-4c39-8e1f-71d88ac17d0f · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering FineAction: A Fine-Grained Video Dataset for Temporal Action Localization
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c9ffc015-b9e8-4de1-b45f-f0574edc4c8a · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering DrVideo: Document Retrieval Based Long Video Understanding
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f84d07b-d150-4a22-a3e6-d5af65ef3d1d · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering Unresolved cited work
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 085a55b7-2b21-43a7-8a36-80c70466fbb1 · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering Foundation Models for Video Understanding: A Survey
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6dd47bf-e0be-4ea9-9140-bfcfb4aac0f3 · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering Unresolved cited work
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58c488f5-236a-4915-8b1e-458f1a1cce65 · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering InProceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (ACL 2024)
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e049e15f-6f1d-43c2-b1ff-b5fc0c38de89 · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering Unresolved cited work
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48bf7a61-c99d-485a-a7e5-114d49c266a4 · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering Qasim, R
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bfe8e4e3-9465-4421-9f4b-1e4d1c9e580b · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering Video-Bench: A Comprehensive Benchmark and Toolkit for Evaluating Video-based Large Language Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0404bfd9-8da6-4e13-a968-9f43d86969a0 · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering TV-TREES: Multimodal Entailment Trees for Neuro-Symbolic Video Reasoning
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 384eec76-6808-4fd5-8fea-1a8de3ccfbc5 · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering Hollywood in Homes: Crowdsourcing Data Collection for Activity Understanding
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d5c72c7-0faa-494f-b3f4-6453f13a14c2 · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering Unresolved cited work
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb9e33e5-fe94-4d8d-824e-90f52617c15d · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering Unresolved cited work
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cdb02b9b-67d9-4830-be24-be0b0b3b54f8 · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering InternVideo: General Video Foundation Models via Generative and Discriminative Learning
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0dd44e4c-e53b-44f3-be17-484d8fd9ff9b · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering ChatVideo: A Tracklet-centric Multimodal and Versatile Video Understanding System
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72be860a-0026-4722-bb6d-14242836a9c7 · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering Unresolved cited work
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 28fd947c-bfe8-4d0b-8954-e19b2f9e491f · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering LongVideoBench: A Benchmark for Long-context Interleaved Video-Language Understanding
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ac3d27f-7c44-4e0a-be78-6d8ba40ee92e · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20751638-ba9e-4f60-84aa-533399259549 · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering LLaVA-Critic: Learning to Evaluate Multimodal Models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7633f68b-5142-46fb-be42-0dcf7ba2cfcf · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f9d426d-bf5f-4d1d-a70a-ea7ea25b8dd6 · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering Unresolved cited work
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 34e11361-29a3-4833-90e7-58ddcf13c67e · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering A Simple LLM Framework for Long-Range Video Question-Answering
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e28f971e-4ef0-4741-b281-96adc07bbee8 · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a16f30d-b100-4137-849c-bc3eb44ebdfa · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering Self-Chained Image-Language Model for Video Localization and Question Answering
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9efbab19-1769-48cb-9f6a-b522e115e48d · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering HACS: Human Action Clips and Segments Dataset for Recognition and Temporal Localization
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22477e2d-58ab-4c4a-bd9a-878b1ce5b9da · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering Unresolved cited work
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 274ddc68-04d3-4e29-b61a-c2eefc215d8d · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering Unresolved cited work
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb1fe118-e6d6-4eb4-a643-551b0be153ba · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering Unresolved cited work
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ea88ecab-50cd-4258-909f-1cbeec8dc364 · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering https://api.semanticscholar.org/CorpusID:1710722
Reference 2015
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7d166b95-e89a-4822-9a73-8309863d4e2a · outbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering VideoVista: A Versatile Benchmark for Video Understanding and Reasoning
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.