Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-12T01:43:34.898639Z
Paper Citation Record · LEDGER
As of 5 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 0 inbound Pith citation observations for arXiv:2605.03276.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-12T01:43:34.898639Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
47 of 47 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 36c2242c-c6f4-493b-9fc5-c85c78507edc · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing Qwen2.5-VL Technical Report
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 046bbe7a-7f19-48b8-99b0-d3d23e3ab290 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing LongVILA: Scaling Long-Context Visual Language Models for Long Videos
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9766c89d-524a-4dbe-9045-7ceadbe02802 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing Video-Holmes: Can MLLM Think Like Holmes for Complex Video Reasoning?
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 181d5190-dbb2-436b-afbf-31755d7a4053 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing Routledge
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6a2909b4-03d3-4426-b70c-775ab0f8442e · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing Motion-grounded video reasoning: Understanding and perceiving motion at pixel level
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 49f06259-5d2f-4df0-8174-49fdfd75bff5 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing arXiv preprint arXiv:2510.08559 , year=
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ba6fc461-36a2-49dd-9d33-13637a3d96ff · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing Gemini models: Gemini 2.5 pro
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation aaf8f905-9c22-4d1f-8a28-cf0f8565531f · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing 2, 4, 6, 7, 8, 12, 17, 18, 19
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2e1f3bec-4f6d-4347-98b7-d978410731ea · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing Rout- ledge
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 56898397-9ff3-4fbc-8501-74f686ab5c1a · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ddcbbf63-5dd0-4327-a7c6-54e35000ab50 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing VITA-1.5: Towards GPT-4o Level Real-Time Vision and Speech Interaction
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1eef172e-fe64-428e-8567-954c569f4741 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing Edit3K: Universal Representation Learning for Video Editing Components
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2080c02b-a044-42fb-ae7b-427c50401297 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing Activitynet: A large-scale video benchmark for human activity understanding
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ecd96d5f-e3e1-4aee-a47e-1c5169d31208 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 17c7c706-eea0-4983-9498-7d7dbfba9c59 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing B-script: Transcript-based b- roll video editing with recommendations
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a0735e0e-082a-4940-b3d7-7480a72c8923 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing Routledge
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1279e9d8-e6c8-4d53-b329-a2fcc9510f0c · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing The Kinetics Human Action Video Dataset
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 52f57a4c-073a-46a2-8636-465ae3e7ae87 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing Veu-bench: Towards comprehensive under- standing of video editing
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d971314e-1153-44c3-9631-166148f9a111 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing Omnivideobench: Towards audio-visual understanding evaluation for omni mllms
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c05576d7-0c16-485c-820f-9b631f10e9f6 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a8c54e84-a4df-4cca-897c-47b0584dc843 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing From representa- tion to reasoning: Towards both evidence and commonsense reasoning for video question-answering
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 7288ce7f-15ce-4d77-8d38-0c9ba640ea99 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e2f2234b-8652-4a36-b7e1-3feb5e406316 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing Shotbench: Expert-level cinematic understanding in vision-language models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 54dfbac0-7998-4141-9c1f-9cebbd314ab0 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing Video-chatgpt: Towards detailed video understanding via large vision and language models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3559ed95-7c5e-44b8-bd63-e7504db47c62 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing University of Chicago Press
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 32806bd2-610e-4563-ab90-eccaea93760b · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing Gpt-4o: Openai’s newest multimodal model
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1a1d5f99-c8eb-4af1-b18c-a8005a997050 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing Perazzi, J
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation eb9960d6-39dc-45a2-9a1d-54c5718145b7 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9f6ddd63-1eef-4b50-ad80-bc0c7dc857cb · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing Qwen3-vl.https : / / github
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 19345d41-4596-4226-96ae-3d7696e51159 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing Sage Publica- tions
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e8833042-c6e1-44b2-a39e-8742d9f4b0c8 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing Video-mmlu: A massive multi- discipline lecture understanding benchmark
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 39a509f4-61e7-4cc0-b868-bfae9ddd69a0 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3cb3f822-1bb4-48fa-99e6-ac825bf317a5 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing Qwen3 technical report
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ee453567-c7e2-470a-b981-c97471120e93 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing Research on the application of nonlinear edit- ing technology in vlog production
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3df129eb-f5ae-4e41-afde-f1e7fb8bc5d7 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing CineTechBench: A Benchmark for Cinematographic Technique Understanding and Generation
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e7890ddb-aaee-42f8-8a92-00b72e438295 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing Longvideobench: A benchmark for long-context interleaved video-language understanding
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 323eaa52-a13e-4814-a781-90ad3a54291c · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing Beyond raw videos: Understanding edited videos with large multimodal model
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 19766d70-a65c-4db4-90d9-718cd387a5f0 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing YouTube-VOS: A Large-Scale Video Object Segmentation Benchmark
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1ed5d487-9e1b-4e07-a2fc-e37cd6298617 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing Qwen2.5 Technical Report
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d890b3e6-1d6d-45fe-b98f-6abcdd6cfe48 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 42e9010d-71af-4681-9788-50b528eb2db4 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing Long Context Transfer from Language to Vision
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3014223d-0f95-496e-8e00-a295897c150a · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing Towards automatic learning of procedures from web instructional videos
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c6394ad0-b7b2-4769-bb38-c383566883f1 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing When does the{CUT TYPE}occur in the video?
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 44a6e817-0593-4731-ad88-62052e2024d2 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing high", "medium
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d5b16fa2-6ad3-4e0f-8ac8-67c724b9861b · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing Yes"/"No
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 331a5521-99df-4ee9-be25-1d0b1ec7116e · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing Unresolved cited work
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 7df02a14-7bf2-48c1-aba5-0d5f5b8af252 · outbound
VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing timestamp_start
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
No inbound Pith citation observations are available.