Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-03T02:34:01.740152Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 81 of 81 outbound references and 7 inbound Pith citation observations for arXiv:2603.16870.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-03T02:34:01.740152Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-01T11:24:28.926512Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-01T22:26:17.260644Z
81 of 81 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 0f37a201-64e9-40b5-8115-2a32baea8c3d · outbound
Demystifying Video Reasoning Advances in neural information processing systems 35, 23716–23736 (2022)
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eaff6a74-4695-4296-8363-0a1950ad5250 · outbound
Demystifying Video Reasoning Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4820378-d8fd-49c6-b65c-495db2a7bb93 · outbound
Demystifying Video Reasoning Neuron 100(2), 490–509 (2018)
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 097e7648-d32c-4efd-8214-3d0523b5fa13 · outbound
Demystifying Video Reasoning BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9dd078b8-1151-42d5-85c8-cc2c9a113731 · outbound
Demystifying Video Reasoning arXiv preprint arXiv:2602.12279 (2026)
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39eb3e3e-cf45-4af9-b7e1-0628fe0c2829 · outbound
Demystifying Video Reasoning In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2339c8b-b782-4231-8d29-031878628030 · outbound
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd3ad8d6-45e0-42fb-8ad1-61a466a4cdcd · outbound
Demystifying Video Reasoning Emerging Properties in Unified Multimodal Pretraining
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f70ffa70-5579-41ee-a828-8fb71f5b77ba · outbound
Demystifying Video Reasoning In: International conference on machine learning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 516faf33-70ee-4909-9683-b156ffc9db19 · outbound
Demystifying Video Reasoning GoT-R1: Unleashing Reasoning Capability of MLLM for Visual Generation with Reinforcement Learning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11b88191-a3d4-4dfa-985f-636b32b79294 · outbound
Demystifying Video Reasoning Vchitect-2.0: Parallel Transformer for Scaling Up Video Diffusion Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffd9ecd0-296e-4f19-8bc1-37937673d194 · outbound
Demystifying Video Reasoning arXiv preprint arXiv:2510.27684 (2025)
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3766ccda-d2e8-4b89-a5ea-a612eab982a5 · outbound
Demystifying Video Reasoning GoT: Unleashing Reasoning Capability of Multimodal Large Language Model for Visual Generation and Editing
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9ae9374-51f2-4a29-943f-6e9afd7eb16a · outbound
Demystifying Video Reasoning SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad251a02-08dc-4a6a-8bc9-d18bc9b24c14 · outbound
Demystifying Video Reasoning Technical Report Veo 3.1, Google DeepMind (January 2026), https: //blog.google/innovation-and-ai/technology/ai/veo-3-1-ingredients-to-video/ , released January 13, 2026
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17c05c7b-cb86-4bb6-a657-19a2bcbeb692 · outbound
Demystifying Video Reasoning Can We Generate Images with CoT? Let's Verify and Reinforce Image Generation Step by Step
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b14d321-d112-415f-a26f-9371d35a4e4a · outbound
Demystifying Video Reasoning LTX-2: Efficient Joint Audio-Visual Foundation Model
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7546311-49ee-460a-88b2-9e59e75daff6 · outbound
Demystifying Video Reasoning Training Large Language Models to Reason in a Continuous Latent Space
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01f0d2eb-bc92-44c4-b0f5-3acaccabad6b · outbound
Demystifying Video Reasoning arXiv preprint arXiv:2512.02622 (2025)
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5763056b-9a86-49f8-b9b8-c70d511e0706 · outbound
Demystifying Video Reasoning In: Larochelle, H., Ranzato, M., Hadsell, R., Balcan, M., Lin, H
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dae0de17-10b7-422b-885e-d6148340e0fc · outbound
Demystifying Video Reasoning In: Proceedings of the 2023 conference on empirical methods in natural language processing
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7ce3380-e8bc-4589-92ef-baf681f3e1f3 · outbound
Demystifying Video Reasoning In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (2024)
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9036b58d-d12f-4cf9-ae27-f4630177d170 · outbound
Demystifying Video Reasoning VChain: Chain-of-Visual-Thought for Reasoning in Video Generation
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37feeb0c-ab25-4076-be2b-32d2b6ff2ec8 · outbound
Demystifying Video Reasoning IEEE Transactions on Pattern Analysis and Machine Intelligence (2025)
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e924498-084f-4499-a413-563005529d2f · outbound
Demystifying Video Reasoning T2I-R1: Reinforcing Image Generation with Collaborative Semantic-level and Token-level CoT
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a35fe085-c00a-4c49-a1ad-29dee50fef5a · outbound
Demystifying Video Reasoning Auto-Encoding Variational Bayes
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af5925bf-b8b4-4ed8-a628-f116730c4493 · outbound
Demystifying Video Reasoning NeurIPS (2023)
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b4af0e4-6043-42d9-99c4-a6f6edff39b5 · outbound
Demystifying Video Reasoning HunyuanVideo: A Systematic Framework For Large Video Generative Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0278788-d2a9-490f-8f71-bfbe3a30a828 · outbound
Demystifying Video Reasoning In: International conference on machine learning
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70886151-777a-419f-a5b6-14b5dd8acc6e · outbound
Demystifying Video Reasoning simultaneous audio-visual generation
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e17a3ed1-3199-4e5e-92d3-77967eb97d3b · outbound
Demystifying Video Reasoning Imagine while Reasoning in Space: Multimodal Visualization-of-Thought
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96511202-ec72-4d57-9479-e51dd39acf60 · outbound
Demystifying Video Reasoning In: International conference on machine learning
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21d33951-39b9-4f9b-a8c4-8db51d955fc8 · outbound
Demystifying Video Reasoning UniFork: Exploring Modality Alignment for Unified Multimodal Understanding and Generation
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e78c0c14-6d6a-4dd8-9330-ccad39e82f4f · outbound
Demystifying Video Reasoning Mogao: An Omni Foundation Model for Interleaved Multi-Modal Generation
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb69a5f8-d580-49f2-ac6d-90bc26e5857e · outbound
Demystifying Video Reasoning arXiv preprint arXiv:2512.11464 (2025)
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5584e8f9-1c0f-4618-9749-6390c1de95e7 · outbound
Demystifying Video Reasoning Unresolved cited work
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 264f299b-d416-4343-aaba-85d72ae9d937 · outbound
Demystifying Video Reasoning Unresolved cited work
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e63bac1c-283b-44a5-b421-70e161b0795b · outbound
Demystifying Video Reasoning arXiv preprint arXiv:2511.16668 (2025)
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 89997916-1492-496e-ae4e-ab84913000cd · outbound
Demystifying Video Reasoning Unresolved cited work
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1163cc56-2a48-4352-904f-9974574aa65f · outbound
Demystifying Video Reasoning Neuron 110(6), 914–934 (2022)
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6c8bd32-0cdd-45c3-84d6-221d60f0de36 · outbound
Demystifying Video Reasoning Unresolved cited work
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2947cdd5-766c-4992-b2cb-55f7e3e03625 · outbound
Demystifying Video Reasoning Transfer between Modalities with MetaQueries
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b3b89fb-608a-4898-9555-3de5ba8fe668 · outbound
Demystifying Video Reasoning In: Proceedings of the IEEE/CVF international conference on computer vision
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df04f4f8-7114-49d9-b9a6-f30f0b6197b6 · outbound
Demystifying Video Reasoning Nature 497(7447), 74–79 (2013)
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12acc4b0-9372-4e7f-8a69-0d04b3b6be29 · outbound
Demystifying Video Reasoning arXiv preprint arXiv:2508.05606 (2025)
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37ff14c8-08d2-4a33-988a-48a3b7e55eba · outbound
Demystifying Video Reasoning In: Proceedings of the Computer Vision and Pattern Recognition Conference
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3fb10181-f827-4d19-a156-0722e0220711 · outbound
Demystifying Video Reasoning In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c5a1038-46f4-4a9f-9d37-48d223a30356 · outbound
Demystifying Video Reasoning Unresolved cited work
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83fc4cf3-1d08-439c-acfc-809263045dce · outbound
Demystifying Video Reasoning arXiv preprint arXiv:2509.24791 (2025)
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f54f030-7751-43f7-9d1a-645e1db41827 · outbound
Demystifying Video Reasoning LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42abb59e-6135-4e0b-97bd-1be2c051949d · outbound
Demystifying Video Reasoning arXiv preprint arXiv:2510.14958 (2025)
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d11b36a1-2c3f-4a61-af49-4c4ce76fc6da · outbound
Demystifying Video Reasoning Unresolved cited work
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1a96eba-cd25-4557-b2ee-d3583052f493 · outbound
Demystifying Video Reasoning arXiv preprint arXiv:2507.06119 (2025)
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 670869ea-7046-448a-a5a5-ed58481567aa · outbound
Demystifying Video Reasoning Chameleon: Mixed-Modal Early-Fusion Foundation Models
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f471e0dc-1dea-4e03-9dc0-e2fe01729b3d · outbound
Demystifying Video Reasoning Thinking with Video: Video Generation as a Promising Multimodal Reasoning Paradigm
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25c15500-b632-4d7c-ac3a-9bdf58c806f8 · outbound
Demystifying Video Reasoning In: Proceedings of the IEEE/CVF International Conference on Computer Vision
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9051897-54b9-4e15-a5ba-b5f46a8e0f90 · outbound
Demystifying Video Reasoning Wan: Open and Advanced Large-Scale Video Generative Models
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11eb632c-a796-451b-8407-2cc90c0e87b7 · outbound
Demystifying Video Reasoning arXiv preprint arXiv:2602.20159 (2026), https://arxiv.org/abs/2602.20159
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af1c5d49-5e6d-4a16-9327-8967dda424b5 · outbound
Demystifying Video Reasoning International Journal of Computer Vision 133(5), 3059–3078 (2025)
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1102ae90-26a7-48bd-b9d1-0cf1f919283e · outbound
Demystifying Video Reasoning Emergent Abilities of Large Language Models
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d1ab11c-f965-4e39-8430-6c37eb09ee6e · outbound
Demystifying Video Reasoning Advances in neural information processing systems 35, 24824–24837 (2022)
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 235c9488-a005-442c-aae0-b67a4bbaab26 · outbound
Demystifying Video Reasoning Video models are zero-shot learners and reasoners
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45021f51-674c-4b4d-8f3d-d5b1ba22c13d · outbound
Demystifying Video Reasoning In: International conference on machine learning
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 187208af-d9ec-47d9-9672-b487604587ea · outbound
Demystifying Video Reasoning Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9df42cb4-f0c1-4cdf-9ea2-3e0d77a4cf38 · outbound
Demystifying Video Reasoning OmniGen2: Towards Instruction-Aligned Multimodal Generation
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12d425a0-b7e2-4e3f-8cf5-3146297b1b50 · outbound
Demystifying Video Reasoning arXiv preprint arXiv:2510.04290 (2025)
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0967b82-5fb2-430d-ae00-a9f2fe1aabca · outbound
Demystifying Video Reasoning MindOmni: Unleashing Reasoning Generation in Vision Language Models with RGPO
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf877170-9014-46ea-b23a-59efc1ded141 · outbound
Demystifying Video Reasoning Unresolved cited work
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7cbc9ef9-354b-4349-8e04-04e7a71963fc · outbound
Demystifying Video Reasoning Understanding Aha Moments: from External Observations to Internal Mechanisms
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a963288-41e8-4616-ab0d-fe9e3d9022d2 · outbound
Demystifying Video Reasoning CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf894df8-e668-4f5a-9d44-adf6bd97ecc9 · outbound
Demystifying Video Reasoning In: Oh, A., Naumann, T., Globerson, A., Saenko, K., Hardt, M., Levine, S
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18660236-bce0-4e81-8c7b-69634b7a9b28 · outbound
Demystifying Video Reasoning In: International Conference on Learning Representations (ICLR) (2023)
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e93bda6-b364-42ba-8712-53a4c8278e1f · outbound
Demystifying Video Reasoning arXiv preprint arXiv:2511.08585 (2025)
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1fd5c3ea-c979-4587-a900-29ed70942d0a · outbound
Demystifying Video Reasoning Robotic Control via Embodied Chain-of-Thought Reasoning
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10b44e85-d8ec-4886-aa4e-cb42eca79b8c · outbound
Demystifying Video Reasoning FutureSightDrive: Thinking Visually with Spatio-Temporal CoT for Autonomous Driving
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33b67750-7ebb-4e76-9b76-f56f83e82902 · outbound
Demystifying Video Reasoning In: Proceedings of the Computer Vision and Pattern Recognition Conference
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3d61ab9-1ca5-40dd-8f67-f100d397e96e · outbound
Demystifying Video Reasoning Advances in Neural Information Processing Systems 37, 12847–12871 (2024)
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0871f3ab-e617-4dea-8835-4fdf9977c74b · outbound
Demystifying Video Reasoning VBench-2.0: Advancing Video Generation Benchmark Suite for Intrinsic Faithfulness
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62405ab9-1e9b-4b96-8ca9-86eeba6799b6 · outbound
Demystifying Video Reasoning Unresolved cited work
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e988bc26-b644-4d96-aa89-0e5d9a344453 · outbound
Demystifying Video Reasoning VARGPT: Unified Understanding and Generation in a Visual Autoregressive Multimodal Large Language Model
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db262d51-c76b-4fd2-b7c1-d22937148fd1 · outbound
Demystifying Video Reasoning Uni-MMMU: A Massive Multi-discipline Multimodal Unified Benchmark
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41a988b5-909d-4a01-87d5-ba45aa6f9e3f · inbound
Do multimodal models imagine electric sheep? Demystifying Video Reasoning
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ae335b51-2285-4a55-8cf5-7468b845d319 · inbound
Video Models Can Reason with Verifiable Rewards Demystifying Video Reasoning
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation bdde8ebd-63b9-4e90-a23c-a3ec92fa1b4e · inbound
VLMs are Good Teachers for Video Reasoning via Adaptive Test-Time Optimization Demystifying Video Reasoning
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 522a6245-66ba-4160-b948-fb59ee29d99c · inbound
VLMs are Good Teachers for Video Reasoning via Adaptive Test-Time Optimization Demystifying Video Reasoning
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 92abe9cd-18cb-4901-9e05-57beaba17782 · inbound
Diffusion ReRoll: Revisable Denoising for Robotic Sequential Prediction Demystifying Video Reasoning
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation baafe26f-2357-4050-b7fb-99c5e9acca6e · inbound
Bridging Compute- and Data-Optimal Pretraining Demystifying Video Reasoning
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1e35699-3f7b-4289-b025-7f4af4bed506 · inbound
Visual prompt engineering for video models Demystifying Video Reasoning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.