Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T11:14:20.027276Z
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 100 of 148 outbound references and 4 inbound Pith citation observations for arXiv:2504.16082.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T11:14:20.027276Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-28T06:33:32.090913Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T07:56:47.335540Z
100 of 148 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c1092f00-5e5c-4fb7-a37a-d7f1c007282f · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33642b88-0fab-4c1c-919c-3fa238d837a8 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Qwen2.5-VL Technical Report
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 647c0f4c-cd02-47b7-9f82-f91f92bb0ec4 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding AuroraCap: Efficient, Performant Video Detailed Captioning and a New Benchmark
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 646c31ea-ce45-47d4-a619-77cc32fd244c · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding An image is worth 1/2 tokens after layer 2: Plug-and-play inference acceleration for large vision-language models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53603afc-4afc-44c7-8c83-35fec21371ac · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding TimeMarker: A Versatile Video-LLM for Long and Short Video Understanding with Superior Temporal Localization Ability
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40466cbc-85d2-454c-8611-bf7245309421 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding InternVL: Scaling up vision foundation models and aligning for generic visual-linguistic tasks
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 044f8aa0-8947-47cd-a3af-430818a1af6e · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding MapReduce: simplified data processing on large clusters
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25d57e1b-059c-4663-b110-7506926c0ef9 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding VideoAgent: A memory-augmented mul- timodal agent for video understanding
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 920db2e5-7d9e-4302-8dee-ba661337c856 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b507845-0c48-40e7-b30c-10e946c75054 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f58cfd34-1c92-42f1-a391-341a333d5e34 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Visual program- ming: Compositional visual reasoning without training
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d1fc84e-f9af-4a15-9e36-c6e51ece875a · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Perceiver: General perception with iterative attention
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f686851-14c1-4309-b03d-55fa62707c2c · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Perceiver IO: A general architecture for structured inputs & outputs
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 130fe568-ca69-44f7-823f-3b2c2e6c4b16 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding SWE- Bench: Can language models resolve real-world github is- sues? In ICLR, 2024
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3db1e49b-7a5c-4c55-93f7-4ff2c2d4f706 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding LLaVA-OneVision: Easy Visual Task Transfer
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4b7b2d6-89d2-4c76-809a-f8873361e3d4 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e07fd8e3-a97e-40e5-9e84-3ad949312a3f · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding BLIP: Bootstrapping language-image pre-training for uni- fied vision-language understanding and generation
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f325a6f-7e85-4d76-bd63-4401f0c36c38 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding BLIP-2: Bootstrapping language-image pre-training with frozen image encoders and large language models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e14c2bd4-2ab7-4d13-8ad6-b9816ba29287 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding VideoChat: Chat-Centric Video Understanding
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a09e42ab-705b-4d78-91d8-6d93e72fbbaa · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Temporal Preference Optimization for Long-Form Video Understanding
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 565c56fc-a2a4-4cb9-9369-3c2a3140fe2e · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding VideoChat-Flash: Hierarchical Compression for Long-Context Video Modeling
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ff42e3f-7bae-4540-8ace-e38ba110efef · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Llama-VID: An image is worth 2 tokens in large language models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5bb5563-1585-412d-823c-6fa687f70f84 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Video-LLaV A: Learning united visual rep- resentation by alignment before projection
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f71adb37-4c22-44ea-8147-9622b7c3fa46 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding MM-Embed: Universal multimodal retrieval with multimodal LLMs
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd158bb0-d796-461a-bc07-2bf36b37103a · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Improved Baselines with Visual Instruction Tuning
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4895218d-029d-4923-9f96-4481941b6cf2 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Visual instruction tuning
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36346025-23e5-4f5d-be8c-d2f61fe60eed · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding LLaV A-NeXT: Im- proved reasoning, ocr, and world knowledge, 2024
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5c13e66-7bb2-4e3e-b177-f9af96ab17eb · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding NVILA: Efficient Frontier Visual Language Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06dcb4a4-afc2-4527-9f21-f89c9179bd94 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Oryx MLLM: On-demand spatial- temporal understanding at arbitrary resolution
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4a5e991-59a3-493e-9e92-685b3f934419 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding EgoSchema: A diagnostic benchmark for very long- form video language understanding
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 330ec540-4a94-4708-b257-b0d74d2a36fb · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding TimeChat: A time-sensitive multimodal large language model for long video understanding
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 872ad005-aa6b-4ccb-90a8-45ff89ab1737 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding LLaV A-PruMerge: Adaptive token reduc- tion for efficient large multimodal models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cee9a261-f156-4da4-b1e5-65296ba20d83 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding LongVU: Spatiotemporal Adaptive Compression for Long Video-Language Understanding
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92cd8083-6679-4c6c-89a6-bc79d22f6b21 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding MovieChat: From dense token to sparse memory for long video understanding
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d39cd2f-a0be-4c6f-a916-3c842c5be408 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding ViperGPT: Visual inference via python execution for reasoning
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe553232-1f03-471e-b811-0b0615b95bf9 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Gemini: A Family of Highly Capable Multimodal Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e46b2d20-0deb-4627-8f67-74315e2db408 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7d04396-d7b3-44ea-892d-a20693979b5b · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding LVBench: An Extreme Long Video Understanding Benchmark
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8298e034-14e3-4253-b1f1-3ba82ca623d7 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding ReTaKe: Reducing Temporal and Knowledge Redundancy for Long Video Understanding
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0cfc33df-fa4d-4996-83c1-34b9513aa4f0 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding LongLLaV A: Scaling multi-modal LLMs to 1000 images efficiently via a hybrid architecture
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d300838f-fb52-4546-a345-221d53434497 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding VideoAgent: Long-form video understanding with large language model as agent
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd4a64fa-0b05-4b3d-b302-095a289f5579 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b40d3b04-17b8-4128-930d-09555e3d088c · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding VideoTree: Adaptive Tree-based Video Representation for LLM Reasoning on Long Videos
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32a72efd-32a9-4b5e-a15c-c514fbb23d26 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Chain-of-thought prompting elicits reasoning in large lan- guage models
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0881e717-e72e-4fc6-a042-58c45af04b66 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding LongVLM: Efficient long video understand- ing via large language models
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb3598c7-5a5b-45be-b697-da995de695ae · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Longvideobench: A benchmark for long-context interleaved video-language understanding
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54199bcc-917c-4bfa-8915-248812d00afc · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Pyramiddrop: Accelerating your large vision-language models via pyramid visual redundancy re- duction
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33f83d22-0245-48ea-86b1-1551a5849916 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding SlowFast-LLaVA: A Strong Training-Free Baseline for Video Large Language Models
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4de1b057-12c3-4f1e-b9a5-9a0ebe6bb377 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding LongVILA: Scaling Long-Context Visual Language Models for Long Videos
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d766a48-3598-4139-954b-14a64ca56cf3 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Qwen2.5 Technical Report
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7212029-79a4-4d1f-89ae-bb0c8a72fb50 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding SWE- Agent: Agent-computer interfaces enable automated soft- ware engineering
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91b4fb99-caa2-4321-91bc-fdaa74ab55fd · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c15a3738-63ea-440d-8e9f-9a3f1d1c60da · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding VCA: Video Curious Agent for Long Video Understanding
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c7e8a2a-baa2-477f-b0a8-5ed4ca5d3938 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding ReAct: Synergizing reasoning and acting in language models
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d275e97a-2fce-4806-9c91-5f9d6900bd55 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding mPLUG- OWL3: Towards long image-sequence understanding in multi-modal large language models
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59742282-5e13-4c3b-ab04-f50be511f0ff · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37451518-0140-4205-bde1-471fbf9a7401 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding A simple LLM framework for long-range video question-answering
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88f4134a-6376-4a23-84e8-649c40e784c7 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding LLaV A-NeXT: A strong zero-shot video understanding model, 2024
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92d9400d-f343-4c4c-b24e-a6fb2824f3c0 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Language agent tree search unifies reasoning acting and planning in language models
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21db97ff-f3fb-41ae-8e51-a1146d116a98 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding MLVU: Benchmarking Multi-task Long Video Understanding
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05b07131-905d-4b5f-9ab1-88579892eaed · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Apollo: An Exploration of Video Understanding in Large Multimodal Models
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd6a3520-631c-423d-b5a4-ca589a53efa3 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Scene Merging
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation acc6f42e-bf58-43e1-9831-85da297a9585 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding The prompts are in Table C
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec5cd9dd-512c-40ed-8ddb-7160f7bee4dd · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Reduce: Consistent Characters and Objects
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85fbb2ac-147b-4767-bc6b-551b92346321 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Unresolved cited work
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b12c936a-2ff6-44c2-ba6b-bd61fb2a6bae · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Unresolved cited work
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3751bd4-06f4-4bb4-899f-8af4c29b9153 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding [1. Description]:
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69c2a5c7-0be9-4c59-b1e2-f057e589a776 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Is this video segment a single scene or a combination of multiple scenes?
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01e668f1-f9ef-4662-ab1a-1455fe5c613c · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding no", please provide the index of frame(s) separating the scenes from the given frame. Your answer should come with a header:
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 437d2a4f-b3e6-4824-acdf-5148dc9a9b1b · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Try to be rigorous and faithful to the video without making assumptions
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9728f37-3494-44b6-8c4c-9ceb3022b2b4 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Unresolved cited work
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0819d016-f84c-4d7f-a745-1c11ca072cac · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding person a
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab82b672-d59c-4cd2-ae9c-49c053dec7d4 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Unresolved cited work
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a17f75b7-cc26-4323-8a35-4a5238b16621 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding It could be a person in the movie, an animal in the documentary or cartoon, etc
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ae66a10-2ecf-40f7-af46-3009069ff4e3 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Please keep the strings in identical formattings to ensure smooth post-processing
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e826a54e-340a-4a7d-b9e5-1550d71b75ec · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Character Selection
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ccebd5ff-889b-4e5a-816b-12db6a89b9fe · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Unresolved cited work
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb6e644d-be11-4cd5-b166-f08e1fac7d17 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Unresolved cited work
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15af2956-cbdb-48e6-a04e-239ce4a1ba3f · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Unresolved cited work
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a77dbbc9-6a2b-4899-8c84-582a5aec5d53 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Unresolved cited work
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 723026b6-4e2b-4be6-91f4-a96f3cd73642 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding If so, please list their name out
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d186c8d0-9601-459f-a9be-f017ba50f10d · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Some more detailed tips:
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f4da4f3-e0c3-4159-a9ce-f1c7afe602e9 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding The goal is that a human should read your captions and feel like watching a continuous video
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2942e76c-f48d-4c7a-9328-80d8ac7d220a · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Unresolved cited work
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd50afd3-36df-460d-8dc5-482d90cfeb37 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding person a
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3506fac-bf4a-4e1a-aa43-6455aa342b87 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Unresolved cited work
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65767748-ecc0-41c8-860e-0b67ea170fee · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Dense Captioning
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b58c023-235f-4897-aa09-9aec933aca4f · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Unresolved cited work
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 547a60e2-4328-4ee2-a6e0-18accb8c50e1 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Unresolved cited work
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2fb64f7b-888f-4a57-9211-2df026864d44 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Unresolved cited work
Reference 91
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0f22290-9772-4c17-b022-002571c640c4 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Unresolved cited work
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2c06448-54ba-4e7e-bd37-caa872a49c2a · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Unresolved cited work
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 542db1c6-6e2c-408a-917f-609eda5eef0a · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Unresolved cited work
Reference 94
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fda50274-fd07-45cb-abc7-4f13e3f645e9 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Unresolved cited work
Reference 95
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3204551e-a098-4002-8e83-83705bb670bf · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Unresolved cited work
Reference 96
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b707b233-79b9-4ebc-ac22-0a85f4b11191 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Character Merging
Reference 97
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0bb0c0d4-f311-4a52-8945-a45db7c2ba4f · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Unresolved cited work
Reference 98
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8cf94b18-4fd2-4e7d-a21e-95f7a5c1db5b · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Output: Your output should be the modified description of the video clip strictly following the original format and contents, only with names changed
Reference 99
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c69f0c2b-e5f5-4303-9048-0f631db62514 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Unresolved cited work
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 7dd82448-1b05-4a3f-b431-64901403ac32 · outbound
MR. Video: "MapReduce" is the Principle for Long Video Understanding Unresolved cited work
Reference 101
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 9eaf1eeb-ce67-45d9-bd88-2d53b82834e9 · inbound
LVBench: An Extreme Long Video Understanding Benchmark MR. Video: "MapReduce" is the Principle for Long Video Understanding
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e2a32a97-bee5-4846-89e4-383a7787ce4e · inbound
Graph-to-Frame RAG: Visual-Space Knowledge Fusion for Training-Free and Auditable Video Reasoning MR. Video: "MapReduce" is the Principle for Long Video Understanding
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e7919ba8-883e-46b6-9af3-a94637f69d8b · inbound
Scaling Video Understanding via Compact Latent Multi-Agent Collaboration MR. Video: "MapReduce" is the Principle for Long Video Understanding
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b020e007-054f-4fde-a454-cbb756619e1f · inbound
GOPAgen: Motion-Aware and Efficient Agentic Long-Video Understanding with Structural Memory and Hierarchical Reasoning MR. Video: "MapReduce" is the Principle for Long Video Understanding
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.