Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-14T00:32:41.059558Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 62 of 62 outbound references and 100 inbound Pith citation observations for arXiv:2501.13826.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-14T00:32:41.059558Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:01:23.214406Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
62 of 62 outbound references displayed
External citation measurements
0
pith, observed 2026-08-05T02:28:24.338817Z
Observation d1731bd0-fc5a-4283-9187-75745f8a10c7 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Claude Team
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 0103abb2-85af-42fe-b43a-7fd3de990e89 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos A systematic classification of knowl- edge, reasoning, and context within the ARC dataset
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 8728683d-c1dd-4cae-8913-cb685f76a829 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos TemporalBench: Benchmarking Fine-grained Temporal Understanding for Multimodal Video Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 68652cac-f1a8-4700-8de1-5f779442999b · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos AuroraCap: Efficient, Performant Video Detailed Captioning and a New Benchmark
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation adcbb1ca-eb72-4683-9220-a3db179f0bc8 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos AutoEval-Video: An Automatic Benchmark for Assessing Large Vision Language Models in Open-Ended Video Question Answering
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 499f417f-6454-400c-b216-2e2203827930 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ec1e6ef3-fe23-4b98-b54c-6247ef07148a · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos MMBench-Video: A Long-Form Multi-Shot Benchmark for Holistic Video Understanding
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation b211095d-f3a6-450a-995b-ace5fa0509a1 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Bloom’s taxonomy
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6a3db641-70ca-4bbe-8e76-45f13f9b81ef · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 634b36d2-6e64-41b8-8fa4-8a709e07d1f2 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Knowit vqa: Answering knowledge-based questions about videos
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e40648bb-8a65-4324-8efc-706e63b9e0b0 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Exploring the video-based learning research: A review of the literature
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 5995dafd-53b9-4ea0-bb20-296df47c3427 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Mammoth-vl: Eliciting multimodal reasoning with instruction tuning at scale
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a70dfd21-e07c-497d-9544-5d3012bedbcc · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Measuring massive multitask language understanding
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 56bea1ca-668e-4db5-862a-4ad84f6da898 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos How Good is my Video LMM? Complex Video Reasoning and Robustness Evaluation Suite for Video-LMMs
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation bc0b2524-7f91-4dc2-ba18-60771d57269b · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos LLaVA-OneVision: Easy Visual Task Transfer
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c5f32c4e-ddd6-47aa-8090-8408a836f07f · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Aria: An Open Multimodal Native Mixture-of-Experts Model
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 915aa26a-9de4-43ef-8e4d-801e92bafa3a · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Mvbench: A comprehensive multi- modal video understanding benchmark
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 59ad81f0-88ca-41c0-8374-acbe6a98c6cb · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos VITATECS: A Diagnostic Dataset for Temporal Concept Understanding of Video-Language Models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 588643b7-9054-49ca-8bae-8de9a918f330 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos VILA: On Pre-training for Visual Language Models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3a2498ba-9a00-451f-a231-ac83e85b75f2 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos TempCompass: Do Video LLMs Really Understand Videos?
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation b0e2c3a6-1603-4f3d-ab28-3152b6d0830d · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Learn to explain: Multimodal reasoning via thought chains for science question answering
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 7cdd6559-5e27-4bc0-aa65-028447b70a19 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Egoschema: A diagnostic benchmark for very long- form video language understanding
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 430b8443-435c-442c-8373-0ff0ab6d710d · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Llama 3.2: Revolutionizing Edge AI and Vision with Open, Customizable Models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 4083582e-e888-40b4-ae0b-4be05929bbf3 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Position: Levels of AGI for operational- izing progress on the path to AGI
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c06a4d2c-520e-4fbb-aff2-933b4f80fb63 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Video-Bench: A Comprehensive Benchmark and Toolkit for Evaluating Video-based Large Language Models
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 2eb31417-79ae-49f7-9ffc-2a8a39dcc7c6 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Introducing openai o1
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 1ffafa23-a8d3-4939-848d-3f5d1006a3de · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Hello gpt4-o
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 9404d1a3-bec6-4ce9-9a44-35ce9b13fc2c · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Per- ception test: A diagnostic benchmark for multimodal video models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 07987b6a-79e5-4c70-adfd-fdbf8e344ddf · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Robust Speech Recognition via Large-Scale Weak Supervision
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation eededdb2-bd5f-4664-9a1e-58e7b16b11a6 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Video- based learning (vbl)—past, present and future: An overview of the research published from 2008 to 2019
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c87f1033-5839-4a62-b3a7-7730ec1d7bda · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos MovieChat: From Dense Token to Sparse Memory for Long Video Understanding
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ffd3e082-84bc-491c-8f1e-16c2db43c57c · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 0f0ff1f7-be7f-42cb-bc7a-f9921acf35de · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a65d6845-9f16-4523-aea4-6ced7e032767 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos LVBench: An Extreme Long Video Understanding Benchmark
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e394b953-b258-4a47-aed4-17f5205795cb · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Vatex: A large-scale, high-quality multilingual dataset for video-and-language research
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 95cf7e11-30a1-4de7-bf0b-3c86ec5b9bdc · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 73d0fb67-fa4f-4517-ba83-ea220dd15072 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos LongVideoBench: A Benchmark for Long-context Interleaved Video-Language Understanding
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 9c7ea373-073a-4463-8849-0d687316e734 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Next-qa: Next phase of question-answering to explaining tem- poral actions
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 09cfd761-6a95-423e-a405-10bc3d937db2 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Funqa: Towards surprising video comprehension
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 78f3424c-fce0-4833-813d-99694014b486 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Video question answering via gradually refined attention over appearance and motion
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e6d6e0df-a896-43d1-804d-d341cc1e4878 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Msr-vtt: A large video description dataset for bridging video and language
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c4a27a04-9bb7-4f9a-bad5-ab798e268567 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Tenenbaum
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 18ecb92a-f00a-4f39-8dbb-25d5a20f45ae · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos The state of video-based learning: A review and future perspectives
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 5919365b-02e9-4e07-b197-b421e87e2b1a · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Activitynet-qa: A dataset for understanding complex web videos via question answering
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 80b0e1f4-62f5-4e60-a75e-ed7a37b914cd · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c4e8f526-35d7-4824-83e4-1d5975a085a1 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos MMMU-Pro: A More Robust Multi-discipline Multimodal Understanding Benchmark
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 8525a6ab-88cd-4869-90a4-a9bd1a23a7bd · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos LMMs-Eval: Reality Check on the Evaluation of Large Multimodal Models
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c28d9e8e-ea83-486e-a407-c8508b62acbc · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Long Context Transfer from Language to Vision
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 0e5a4491-57c8-48bc-8d32-d175db17ac2d · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos LLaVA-Video: Video Instruction Tuning With Synthetic Data
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation d9882c54-9125-4896-870d-a8a50bc149c4 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos WorldQA: Multimodal World Knowledge in Videos through Long-Chain Reasoning
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation b625a5af-0660-47b0-bf93-357a2637536e · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos AGIEval: A human-centric benchmark for evalu- ating foundation models
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 783b2a2c-8d75-4270-a132-88339aad9d51 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos MLVU: Benchmarking Multi-task Long Video Understanding
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6135b1d8-db26-465f-8541-3de4210e528d · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Towards au- tomatic learning of procedures from web instructional videos
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 76f97e8b-fff3-4f23-9609-cdf66b036536 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Subjects categorized under six disciplines
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 1bdaac07-7ced-44f0-b9d6-6a24267829eb · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Unresolved cited work
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ba349dcd-e51c-4ada-bcf2-16c57c3a6072 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos The ∆knowledge metric reveals a gap between human ex- perts and models, particularly in their ability to learn new information from videos
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 604b140f-ea30-4fba-9eed-e66aa98bab85 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos We introduce the prompt as shown in Fig
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 43e47588-beb8-4e79-a037-a249a5b4c3b4 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Unresolved cited work
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation dbbdfb37-9ea2-494a-8272-128dbdafa546 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos Unresolved cited work
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation eb692f5b-9469-4d6f-8f12-39cff4aae646 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos We begin by examining errors made by Claude-3.5-Sonnet [1] in the Adaptation track
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a7a853d7-670f-44e2-a0fe-13c299240a80 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos 17 and Fig
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ef256c8d-5598-4cde-9142-ab530c034b35 · outbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos reason"
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation efd0b836-425d-4aa3-b5a7-ea57af6c3468 · inbound
Qwen2.5-VL Technical Report Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 9fc59b67-919d-457e-843e-70be6809261f · inbound
Video-R1: Reinforcing Video Reasoning in MLLMs Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 8b90fa82-393e-438c-9a3a-e2da6583651c · inbound
Seed1.5-VL Technical Report Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 2bf9ee16-4a31-4cbb-a89b-bbbe64aa3707 · inbound
Seeing is Believing, but How Much? A Comprehensive Analysis of Verbalized Calibration in Vision-Language Models Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf0f1754-3b0e-4d04-9148-d9a3824fac12 · inbound
Video-Holmes: Can MLLM Think Like Holmes for Complex Video Reasoning? Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation cf25da3c-0d03-4627-8fd6-25d567b1309b · inbound
ReFoCUS: Reinforcement-guided Frame Optimization for Contextual Understanding Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69b421ce-8c08-4917-abe0-bdfc3944b2de · inbound
ReAgent-V: A Reward-Driven Multi-Agent Framework for Video Understanding Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a3b0e9d-49b5-48eb-889e-a442eaca222c · inbound
VideoCap-R1: Enhancing MLLMs for Video Captioning via Structured Thinking Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d211c42b-33cd-43bf-a248-ebabe66b7dc9 · inbound
MiMo-VL Technical Report Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9737389c-37da-4dc9-96bb-e87c317bf88e · inbound
EOC-Bench: Can MLLMs Identify, Recall, and Forecast Objects in an Egocentric World? Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6155ea1-b6fa-4f25-a4e4-e83f48302a82 · inbound
VideoMathQA: Benchmarking Mathematical Reasoning via Multimodal Understanding in Videos Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83866043-2fe4-421c-baad-5889f5fab71d · inbound
Movie Facts and Fibs (MF$^2$): A Benchmark for Long Movie Understanding Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3afe48cc-0bfa-4fef-b495-e4872d0f038f · inbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23b5364e-be60-48e2-8e6d-77a1ca0006f8 · inbound
VRBench: A Benchmark for Multi-Step Reasoning in Long Narrative Videos Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e2d0813-3723-406c-9f37-99613b9ef861 · inbound
GRPO-CARE: Consistency-Aware Reinforcement Learning for Multimodal Reasoning Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb3617e0-da3f-4432-a830-7d9fd7b958eb · inbound
HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ec607c9-c843-4993-88d0-adafbf7c95ac · inbound
Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation acb10c26-4472-4002-b820-f8ac1d9702d1 · inbound
ExpStar: Towards Automatic Commentary Generation for Multi-discipline Scientific Experiments Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6fbbd43-c890-4ca8-a9d3-fa6583e59791 · inbound
Position: Reasoning After Perception Means Reasoning Without Vision Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf07f444-28eb-4b9c-a345-6ac8b2ac5981 · inbound
CausalStep: A Benchmark for Explicit Stepwise Causal Reasoning in Videos Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc331370-fca2-4f22-95be-38d86ae9cd9d · inbound
Empowering Nanoscale Connectivity through Molecular Communication: A Case Study of Virus Infection Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d32cec37-b735-4122-b313-e80965f80cf5 · inbound
FineBadminton: A Multi-Level Dataset for Fine-Grained Badminton Video Understanding Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11e48a79-c945-4f94-a423-e6c4edf28133 · inbound
HumanPCR: Probing MLLM Capabilities in Diverse Human-Centric Scenes Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bac3de43-4733-4cfc-8917-fae8087cf7ea · inbound
LLaVA-Critic-R1: Your Critic Model is Secretly a Strong Policy Model Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58847d44-65cc-47b0-83c6-4c87767cd765 · inbound
Kwai Keye-VL 1.5 Technical Report Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b57281fa-1794-4517-9b12-31f4fb6d763a · inbound
NeMo: Needle in a Montage for Video-Language Understanding Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43d1e47b-5fec-4a95-8a97-da2b40f971b2 · inbound
Video Reasoning without Training Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81837d7b-8245-4805-bf7d-2fa9af697a9f · inbound
Cambrian-S: Towards Spatial Supersensing in Video Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 58b52126-6b6d-4af6-b7c0-a2b0dafc457f · inbound
VIDEOP2R: Video Understanding from Perception to Reasoning Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6d4f1304-964d-4c0f-8f18-bc54829ac23d · inbound
Boosting Reasoning in Large Multimodal Models via Activation Replay Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e7c55f0d-303b-4a64-9a88-1f4fcf08e9cf · inbound
LongVT: Incentivizing "Thinking with Long Videos" via Native Tool Calling Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation fb5536a7-0eee-4369-8936-f4e7a9f9c40a · inbound
OneThinker: All-in-one Reasoning Model for Image and Video Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation bc9bb639-42bc-455a-93fc-696a0943f177 · inbound
Mimic Human Cognition, Master Multi-Image Reasoning: A Meta-Action Framework for Enhanced Visual Understanding Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46230b7d-a056-4d80-981a-a5af8cdc0ee5 · inbound
CASHEW: Stabilizing Multimodal Reasoning via Iterative Trajectory Aggregation Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1502abd8-4c8b-4009-9db7-290ceeb31924 · inbound
Kimi K2.5: Visual Agentic Intelligence Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e30a6bc3-0ff6-4fab-9f6f-95773074f92d · inbound
Multimodal Fact-Level Attribution for Verifiable Reasoning Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 975ed005-70c3-454a-a2f0-0753e32adda6 · inbound
Seed1.8 Model Card: Towards Generalized Real-World Agency Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c8c948a5-b8cd-42d7-b9cc-1ab14bd6d24d · inbound
STRIVE: Structured Spatiotemporal Exploration for Reinforcement Learning in Video Question Answering Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation df48e1ea-5eec-4fa0-aefa-fdebe8c86817 · inbound
Graph-to-Frame RAG: Visual-Space Knowledge Fusion for Training-Free and Auditable Video Reasoning Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 95b9441b-58d2-4316-86fb-d2b139bd910c · inbound
Reinforce to Learn, Elect to Reason: A Dual Paradigm for Video Reasoning Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 99e6c9fe-8f53-4edf-a8c4-5fe02ae8b085 · inbound
Video-MME-v2: Towards the Next Stage in Benchmarks for Comprehensive Video Understanding Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 32406f91-108a-4100-9698-0db4ae057762 · inbound
Watch Before You Answer: Learning from Visually Grounded Post-Training Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f13f9e76-f634-456d-8e11-b5626ab11d3a · inbound
Mastering PokeGym: Graph-Guided Multimodal Evolution at Test Time Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 8d4db040-b2d4-4437-8065-0f0eaf386051 · inbound
Mastering PokeGym: Graph-Guided Multimodal Evolution at Test Time Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5860d76e-5fd8-4000-bce0-48763bc471b7 · inbound
GeoMMBench and GeoMMAgent: Toward Expert-Level Multimodal Intelligence in Geoscience and Remote Sensing Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 662af7b6-fcdd-4eb7-80f5-c21a804fea49 · inbound
EvoDiagram: Agentic Editable Diagram Creation via Design Expertise Evolution Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 0e8368ba-52af-4f81-8084-cf5a7813dbe4 · inbound
Video-ToC: Video Tree-of-Cue Reasoning Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 30bcae99-70f8-48b1-86e5-4f95d26b9c51 · inbound
The category of Whittaker modules over the Cartan Type Lie algebra $\bar{S}_2$ Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 0fbe53cc-b066-405a-83d5-fcc41ea0d682 · inbound
FCMBench-Video: Benchmarking Document Video Intelligence Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a3e841cf-fbc4-4ee8-a112-989bbaea4e71 · inbound
Valley3: Scaling Omni Foundation Models for E-commerce Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 399d77b8-1975-4024-b06f-daaecd98fc2f · inbound
Beyond Perceptual Shortcuts: Causal-Inspired Debiasing Optimization for Generalizable Video Reasoning in Lightweight MLLMs Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 670b33bb-aab2-4f71-a3b2-3a34d47f9dd0 · inbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ecd96d5f-e3e1-4aee-a47e-1c5169d31208 · inbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation b8fe4e0c-8b91-44a0-bdcc-1a174110770a · inbound
MACS: Modality-Aware Capacity Scaling for Efficient Multimodal MoE Inference Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation d65a0186-1bde-4011-b6ab-6b082174943f · inbound
VISD: Enhancing Video Reasoning via Structured Self-Distillation Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 37032c69-546a-4511-a58a-9e247519c44f · inbound
VISD: Enhancing Video Reasoning via Structured Self-Distillation Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6902eb56-7ec2-4c8d-9aaf-51e804f7ec20 · inbound
VISD: Enhancing Video Reasoning via Structured Self-Distillation Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 04079913-6695-4578-ae1e-593894db49a6 · inbound
VISD: Enhancing Video Reasoning via Structured Self-Distillation Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3bbf2dbb-27cc-47c6-be66-704261458a7b · inbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation fdbd658b-4ca4-4051-bf46-7245cbe26f63 · inbound
EgoMemReason: A Memory-Driven Reasoning Benchmark for Long-Horizon Egocentric Video Understanding Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 77e403aa-7e0a-4617-9a56-89e70375d6f7 · inbound
EchoPrune: Interpreting Redundancy as Temporal Echoes for Efficient VideoLLMs Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 09d2012b-c9ad-4fea-a7cf-e3810d48b0cd · inbound
AdaFocus: Adaptive Relevance-Diversity Sampling with Zero-Cache Look-back for Efficient Long Video Understanding Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation eaf66829-84aa-4c51-a8f6-51102c737900 · inbound
Video-Zero: Self-Evolution Video Understanding Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 5d37020e-f4a6-4ed8-9879-7241181bc7dd · inbound
GRASP: Learning to Ground Social Reasoning in Multi-Person Non-Verbal Interactions Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 11c768d2-96b1-4b5c-9b61-77d97b6e8cbf · inbound
OProver: A Unified Framework for Agentic Formal Theorem Proving Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 86dc8a94-eebf-4a07-91d8-daf4fbc58408 · inbound
Self-Evolving Spatial Reasoning in Vision Language Models via Geometric Logic Consistency Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 35f983cf-b3fd-4ab5-9d4c-ed5e032f54a5 · inbound
EvoVid: Temporal-Centric Self-Evolution for Video Large Language Models Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f9f54f77-1808-4579-9437-5cf3ad7adc8b · inbound
Cambrian-P: Pose-Grounded Video Understanding Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 02c75af7-b6cb-4797-8f26-a77bd34fa8d1 · inbound
Cambrian-P: Pose-Grounded Video Understanding Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ebcfad3e-b98f-46bf-872c-884bcdf6d89b · inbound
VideoOdyssey: A Benchmark for Ultra-Long-Context and Omni-Modal Video Understanding Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 0d39a4c6-621c-4053-b4df-56760c7414c9 · inbound
MetaphorVU: Towards Metaphorical Video Understanding Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 242c2a57-807e-4572-a409-cb13252dbdfd · inbound
Moment-Video: Diagnosing Temporal Fidelity of Video MLLMs on Momentary Visual Events Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 020b3d45-544e-4925-82be-d3faa3709af9 · inbound
Benchmarking Visual State Tracking in Multimodal Video Understanding Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3d47b554-d500-42a7-8e76-8ac6d3715538 · inbound
VideoKR: Towards Knowledge- and Reasoning-Intensive Video Understanding Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 87186e59-cf0d-40a7-8c35-802d3cf697f9 · inbound
StoryVideoQA: Scaling Deep Video Understanding with a Large-Scale, Multi-Genre and Auto-Generated Dataset Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 26d3b2c8-06c6-43cd-af18-ba2ad3dee28f · inbound
Watch, Remember, Reason: Human-View Video Understanding with MLLMs Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 251
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 23dd6458-9750-4df2-98a5-95f316ea0833 · inbound
Test-Time Scaling in Multimodal Foundation Models: A Comprehensive Survey of Generation and Reasoning Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e382ad80-428b-4bd1-a2a9-d79a528c69ce · inbound
When No Answer Is Correct: Diagnosing Absent Answer Detection for MLLMs in Video Understanding Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation b8a80bee-0064-452e-82b0-0978aaafad45 · inbound
Harnessing Streaming Video in the Wild Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 40e94dc2-374e-4ba9-bc44-28a3d2e0ae44 · inbound
Streaming Interventions: Can Video Large Language Models Correct Mistakes as They Occur? Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 60d05fc2-1132-4415-b50f-0555ceb720df · inbound
Mitigating Manifold Departure: Uncertainty-Aware Subspace Rectification for Trustworthy MLLM Decoding Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 8f22d88c-4c28-489c-ad91-d4cc8f183447 · inbound
Kwai Keye-VL-2.0 Technical Report Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 477122f4-1492-47ec-b288-ce6ac919f074 · inbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 26f4f85a-2d54-4a07-a1e8-9e3a4eb3b666 · inbound
MultiToP: Learning to Patch Visual Tokens to Mitigate Hallucinations in Video Large Multimodal Models Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3afcec70-e702-4724-bc73-5f78ec9efa3b · inbound
Zone of Proximal Policy Optimization: Teacher in Prompts, Not Gradients Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 154
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f0e9a8d7-384a-4f12-b05f-22c7a349164b · inbound
Reasoning as Intersection: Consensus-Frame Alignment for Visual Focus in Video-MLLMs Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a7df9cba-fea1-4bf8-b833-80e21872a515 · inbound
CARE: Competence-Aware Reward Shaping for Adaptive Reasoning Length in Video-MLLMs Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 096ab594-a4d1-4d3c-81cd-e78aea2f8b8d · inbound
CineCap: Structured Reasoning with Spatio-Temporal Anchors for Cinematographic Video Captioning Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 86253e11-0555-40ca-a8c1-1c3749753fe3 · inbound
ReasonCLIP-58M: Visually Grounded Commonsense Reasoning Supervision for CLIP Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 75bcbb7a-7456-4d0d-b18a-f5a9cfb67c53 · inbound
MuseBench: Benchmarking Intent-Level Audiovisual Arts Understanding in MLLMs Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 9a1e924e-edb1-4bec-b26b-a1526b84a879 · inbound
Seed2.0 Model Card: Towards Intelligence Frontier for Real-World Complexity Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 7a5f67c0-d40c-47cc-9019-72f1c7e0bf0d · inbound
EFlow: Learning Evidence Flow for Long-Video Reasoning with Adaptive Reflection Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation d9c714e0-1507-4d88-956b-847ee6678d29 · inbound
EFlow: Learning Evidence Flow for Long-Video Reasoning with Adaptive Reflection Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f704073d-3c4e-4b54-8fc9-4a3843aa86e1 · inbound
EFlow: Learning Evidence Flow for Long-Video Reasoning with Adaptive Reflection Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56ae4074-eea7-41a1-9e60-87eb220447d9 · inbound
DELTAVID: Enhancing Fine-Grained Spatiotemporal Perception with Cross-Video Differences Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79b7c0e5-ac12-46d6-bdf4-d8a4afaae848 · inbound
Latent Visual Cache for Video Reasoning Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87446d1e-82fd-4a2b-a961-4bb2e12e55a9 · inbound
S-EMBER: A Large-Scale Benchmark for Streaming Egocentric Memory Retrieval Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 403c2c3f-1b87-4da6-a15a-2e3900ab221a · inbound
VideoSearcher: Empowering Video Deep Research with Multi-Tool Agentic Reasoning via Reinforcement Learning Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b511d71-d1ed-44d5-8104-b1faa0a5c943 · inbound
TimeThink: Reasoning with Time for Video LLMs Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a0e508e-08bc-41c5-a574-c1a893454cb8 · inbound
VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.