Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:25:41.790587Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 98 of 98 outbound references and 1 inbound Pith citation observation for arXiv:2505.19000.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:25:41.790587Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-27T22:00:28.350003Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T17:27:15.521253Z
98 of 98 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 1d7e8ef0-f616-4233-b6b8-7c5546b3fb49 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Vivit: A video vision transformer, 2021
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76d4e32c-d59d-4411-b43d-484e9c0b9034 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Qwen2.5-vl technical report, 2025
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0961278c-cf23-4573-8a72-ec11bc76853c · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Hadzic, Taran Kota, Jimming He, Cristobal Eyzaguirre, Zane Durante, Manling Li, Jiajun Wu, and Fei-Fei Li
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 409da519-2955-42e7-b6f6-3deb200c6084 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Mecd: Unlocking multi-event causal discovery in video reasoning, 2024
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2cf40d0-030b-4afe-b70b-34471a2baf45 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization On the suitability of reinforcement fine-tuning to visual tasks, 2025
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04afaa31-a2d4-4e3e-acfd-b8b2e62b02fa · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Videovista-culturallingo: 360◦ horizons-bridging cultures, languages, and domains in video comprehension, 2025
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5266d1d5-e3f8-41fa-b581-b017a02346c8 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Visrl: Intention-driven visual perception via reinforced reasoning, 2025
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37bae379-aac5-41b5-906a-b7e3ccdcd4af · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Expanding performance boundaries of open-source multimodal models with model, data, and test-time scaling, 2025
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9753eec-9007-448c-9e2d-a954ce290fb3 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Videollama 2: Advancing spatial- temporal modeling and audio understanding in video-llms, 2024
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b603b5b1-3815-42a2-9a04-e5e9eb6ee7b9 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Skywork r1v2: Multimodal hybrid reinforcement learning for reasoning, 2025
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b4c2fd8-0f4c-4296-8aa9-21ea56e7bcc9 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Instructblip: Towards general-purpose vision-language models with instruction tuning, 2023
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ee485d6-b41d-423c-ba79-8f6a363217c0 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Mm-spatial: Exploring 3d spatial understanding in multimodal llms, 2025
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c6f3ce1-2ded-4a01-acd7-42caf7e37e5d · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Deepseek-r1: Incentivizing reasoning capability in llms via reinforcement learning, 2025
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d556e72-8013-424b-a2f7-f47a33501972 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Boosting the generalization and reasoning of vision language models with curriculum reinforcement learning, 2025
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3951efe9-2156-475d-b43b-2fae23be11df · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Insight-v: Exploring long-chain visual reasoning with multimodal large language models, 2025
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 045a8a13-2a17-4340-b703-f407c3577525 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization An image is worth 16x16 words: Transformers for image recognition at scale, 2021
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb9e0ad6-7a21-4b74-8675-62abdd22a36a · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Video-of-thought: Step-by-step video reasoning from perception to cognition, 2024
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4df0315d-e9ad-40f6-a897-404bfd87ced8 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Video-r1: Reinforcing video reasoning in mllms, 2025
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eefe869a-ff52-4113-b12c-7a877552bf95 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62ae0e71-efab-4eed-8c25-3071c665ac49 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Ampo: Active multi-preference optimization, 2025
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 06263630-fe84-4bef-acbe-3daaedff91e5 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Video-mmmu: Evaluating knowledge acquisition from multi-discipline professional videos, 2025
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6de313b-5d04-4d0e-9610-853f4aa3ec4a · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Vision-r1: Incentivizing reasoning capability in multimodal large language models, 2025
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6c110c8a-34eb-4a67-8c7c-b71ba96b09e9 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization GPT-4o System Card
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f81678d-94c2-46d6-8cb3-a7c0ceb43465 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Llava-onevision: Easy visual task transfer, 2024
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47f11427-e3c6-4d4e-b5f8-320f644cef56 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization mplug: Effective and efficient vision-language learning by cross-modal skip-connections, 2022
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 694d051c-050d-4572-a1b7-d061a7db60cb · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Videochat: Chat-centric video understanding, 2024
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28608a34-9347-49bc-8472-e0d4921349ce · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Mvbench: A comprehensive multi-modal video understanding benchmark, 2023
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2606f9a0-c124-4410-befb-498f6dd6566e · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Videochat-r1: Enhancing spatio-temporal perception via reinforcement fine-tuning, 2025
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ca29190a-592f-4596-bf5c-8469b8adeb75 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Lmeye: An interactive perception network for large language models
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 36f6cafb-21df-4428-85ea-3a801d322d36 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Uni-moe: Scaling unified multimodal llms with mixture of experts
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5564ec1e-1a3c-4ca7-9f33-586185119adb · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Perception, Reason, Think, and Plan: A Survey on Large Multimodal Reasoning Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41962da8-18cf-4446-99e1-a953da3e28de · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization STI-Bench: Are MLLMs Ready for Precise Spatial-Temporal World Understanding?
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c074983b-78b9-48a4-85ae-111466a8ce31 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Vila: On pre-training for visual language models, 2024
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9c71c0c2-5ca9-4ed6-82d6-f067018ce985 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Spatialcot: Advancing spatial reasoning through coordinate alignment and chain-of-thought for embodied task planning, 2025
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7ed38859-3287-459d-b6bc-99df5e5c01e4 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization TempCompass: Do Video LLMs Really Understand Videos?
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a487b60-f8b8-41db-9bdb-01b659bc5e59 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Videomind: A chain-of-lora agent for long video reasoning, 2025
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 13772a12-cafc-4c28-8b44-6cbb865b1596 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Seg-zero: Reasoning-chain guided segmentation via cognitive reinforcement, 2025
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5eeef245-9d51-4420-a4cd-79593a5113c5 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Video swin transformer, 2021
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5a8ded3c-ef20-4f05-bd47-5f30ccd8638e · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Visual-rft: Visual reinforcement fine-tuning, 2025
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f3e6f78-c2a0-48b0-bc17-8dd518d17540 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Othink- mr1: Stimulating multimodal generalized reasoning capabilities via dynamic reinforcement learning, 2025
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 20223307-5c25-4dfc-bbae-66b86c8b6fce · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Gui-r1 : A generalist r1-style vision-language action model for gui agents, 2025
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a6cce647-dddc-463f-b3e6-9ef9099aa604 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Video-chatgpt: Towards detailed video understanding via large vision and language models, 2024
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e152623-b3b3-4979-8f71-49db77abb6f1 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Mm-eureka: Exploring the frontiers of multimodal reasoning with rule-based reinforcement learning, 2025
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4fe77fa-5fdc-4ef2-a39f-24ec1e725581 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Video transformer network, 2021
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2e798a55-98d7-456b-91d9-1a14b0132818 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Dinov2: Learning robust visual features without supervision, 2024
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16920c04-7cd3-4975-90fb-6ea98b4cf632 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Spatial-r1: Enhancing mllms in video spatial reasoning, 2025
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5ba9fb0b-7020-4977-acc0-70fdc2ecbb05 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Unresolved cited work
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 274d75a4-99c2-4fae-b82f-8c8ea323620a · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Lmm-r1: Empowering 3b lmms with strong reasoning abilities through two-stage rule-based rl, 2025
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ae60161-9a1d-4baf-983b-5b9857f9dbec · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Learning transferable visual models from natural language supervision, 2021
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d41aa557-de70-43b5-9dc6-904fbcf18037 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Manning, and Chelsea Finn
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20d7cbe9-da84-4ea7-bb10-c5d45ff3169c · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Plummer, Ranjay Krishna, Kuo-Hao Zeng, and Kate Saenko
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2ac9b6b0-24dd-445c-bd92-afbc4ea05eae · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Proximal policy optimization algorithms, 2017
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37d9daa0-a665-4770-9ce8-60933f1e9014 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Tomato: Assessing visual temporal reasoning capabilities in multimodal foundation models, 2024
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f98fa328-6db1-4d87-9835-ddb4c4c3bf3f · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Unresolved cited work
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31de27d0-f23a-4a18-bda0-eaba13153a2c · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Efficient reinforcement finetuning via adaptive curriculum learning, 2025
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b28d21e-c5ed-4f3f-b71e-5b423f72ca50 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Mm-verify: Enhancing multimodal reasoning with chain-of-thought verification, 2025
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c7cc9cbe-8314-4493-aaf9-6f00923459ea · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Reason-rft: Reinforcement fine-tuning for visual reasoning, 2025
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d3d5cbb0-b7a8-462e-8de6-9ae7dfdde0bd · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Game-theoretic regularized self-play alignment of large language models, 2025
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b14dd7fb-4bb3-45f1-8dcd-927f57756db1 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context, 2024
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 81d046a8-3033-4a98-aabe-f111467af0b0 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Gemma 3 technical report, 2025
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b0b7e909-c990-4a8b-a390-46ecc82fa521 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Kimi-vl technical report, 2025
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 458ec069-71a3-4c01-8b75-3725b593243a · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Model cards & prompt formats-llama 3.2, 2024
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4a57df1e-44aa-491b-a32c-55f7f36a3b64 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Vila: On pre-training for visual language models, 2024
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 56c6dde5-cc56-4c18-8bd5-efddf967b176 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Goucher, et al
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation edf34835-998e-4cf8-85e2-4d5ff31e0076 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Qwen3: Think deeper, act faster, April 2025
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3a4ccaf4-f190-434d-98f5-00bde03988c8 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization VL-Rethinker: Incentivizing Self-Reflection of Vision-Language Models with Reinforcement Learning
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de1a0619-fe07-467a-b6b7-99fac19ae44e · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Piecing it all together: Verifying multi-hop multimodal claims, 2024
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation df2376cd-cd36-41ec-9c1e-943cdf29682d · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Lvbench: An extreme long video understanding benchmark, 2024
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e47208b-c819-471e-8bc0-f136c1983262 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Sota with less: Mcts-guided sample selection for data-efficient visual reasoning self-improvement, 2025
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 93bdbdc0-9ed4-437a-86c3-d2a7a9d46d85 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Internvideo2.5: Empowering video mllms with long and rich context modeling, 2025
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d6b0d26-3bed-48fa-be71-bbf7d253e879 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Unified multimodal chain-of-thought reward model through reinforcement fine-tuning, 2025
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6cbbd6c0-91cf-4e2c-8721-c2d947882dca · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Chain-of-thought prompting elicits reasoning in large language models, 2023
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aaa784da-4c0f-4057-9298-2f5431b7b7fa · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Videorope: What makes for good video rotary position embedding?, 2025
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d1d8dda5-cc5b-4249-9587-6db55b0b4828 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Longvideobench: A benchmark for long-context interleaved video-language understanding, 2024
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af16f7f3-7145-4de7-aa47-b97cb4c42464 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization St-think: How multimodal large language models reason about 4d worlds from ego-centric videos, 2025
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 58361d60-a467-4c6a-84df-43188927f974 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Self-play preference optimization for language model alignment, 2024
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4122e80-3584-4c60-b78a-044ff0624539 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Deepseek-vl2: Mixture-of-experts vision-language models for advanced multimodal understanding, 2024
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 887025b9-a251-40eb-892e-22deded07598 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Atomthink: A slow thinking framework for multimodal mathematical reasoning, 2024
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7005e81d-407c-4fb1-bc4c-e4cd25d974e4 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Echoink-r1: Exploring audio-visual reasoning in multimodal llms via reinforcement learning, 2025
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b0a3ded2-742d-49cf-997f-23e8cd3e3487 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Redstar: Does scaling long-cot data unlock better slow-reasoning systems?, 2025
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f1148cd-e94b-4c22-9c6c-bc7ba63037d4 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Thinking in Space: How Multimodal Large Language Models See, Remember, and Recall Spaces
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94135652-4ba3-47a5-9413-8ca9ab4c4bbd · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization R1-onevision: Advancing generalized multimodal reasoning through cross-modal formalization, 2025
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1422e568-a62c-4f38-ab79-dae7400bd33d · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization mplug-owl3: Towards long image-sequence understanding in multi-modal large language models, 2024
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc33e49c-0dbd-4179-8060-09fd3d69d3ef · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Dapo: An open-source llm reinforcement learning system at scale, 2025
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4af47f56-f40f-45d5-a768-769402dbf9aa · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Unresolved cited work
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8be89472-ba36-48a2-acaf-6c8964894f8e · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Videollama 3: Frontier multimodal foundation models for image and video understanding, 2025
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f365eae-08de-4075-bbac-827c029c8241 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Video-llama: An instruction-tuned audio-visual language model for video understanding, 2023
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d92d34fe-6d3f-4d77-a3ce-7707b6eb1004 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization From flatland to space: Teaching vision-language models to perceive and reason in 3d
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d02863f-399b-481a-9883-b7c9f4745e9e · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Long context transfer from language to vision, 2024
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b05ae24-3ab6-4989-8d53-d93c74eedfa3 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Tinyllava-video-r1: Towards smaller lmms for video reasoning, 2025
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 87dd1c2c-1b86-46ef-b568-3ddeb924ad6a · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Video instruction tuning with synthetic data, 2024
Reference 91
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1ac7842-aec5-467a-a165-9a68c7a038f4 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Openrft: Adapting reasoning foundation model for domain-specific tasks with reinforcement fine-tuning, 2024
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a3773b89-0653-4ec5-83c5-ff8ee6856a14 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization MME-RealWorld: Could Your Multimodal LLM Challenge High-Resolution Real-World Scenarios that are Difficult for Humans?
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0588aaa-2156-472e-bc48-20aefa853e37 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Multi- modal chain-of-thought reasoning in language models, 2024
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 213df899-8b29-40ec-bc12-e26386b42818 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization R1-omni: Explainable omni-multimodal emotion recognition with reinforcement learning, 2025
Reference 95
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a98b2a9-a395-4ccf-9c72-3c192bb9fa41 · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Mmvu: Measuring expert-level multi-discipline video understanding, 2025
Reference 96
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 367c941e-1057-4472-9b44-5bb2de7f7d4a · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Villa: Video reasoning segmentation with large language model, 2025
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2ccaa982-07eb-45cb-9427-ff88ce35dd4c · outbound
VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization MLVU: Benchmarking Multi-task Long Video Understanding
Reference 98
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f0f584d-4899-4432-99e9-981f89d29755 · inbound
Watch, Remember, Reason: Human-View Video Understanding with MLLMs VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization
Reference 186
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.