Pith. sign in

Paper Citation Record · LEDGER

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization

As of 10 August 2026, this Paper Citation Record lists 98 of 98 outbound references and 1 inbound Pith citation observation for arXiv:2505.19000.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.19000 v1

Coverage vector

measured 98 of 98 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:25:41.790587Z

measured 99 of 99 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T22:00:28.350003Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T17:27:15.521253Z

Reference resolution

98 of 98 outbound references displayed

  • verified exact0
  • verified fuzzy37
  • unresolved61
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1d7e8ef0-f616-4233-b6b8-7c5546b3fb49 · outbound

This paper cites Vivit: A video vision transformer, 2021.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Vivit: A video vision transformer, 2021

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:33.753694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:33.753694Z digest=sha256:18a6a0065930f7dea08cb7d6eed603fc9429af1e23b5df7fa0232aeeac1fdd00

Observation 76d4e32c-d59d-4411-b43d-484e9c0b9034 · outbound

This paper cites Qwen2.5-vl technical report, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Qwen2.5-vl technical report, 2025

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:33.825918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:33.825918Z digest=sha256:88e6df56a7a9d7841d18a811fb8e9b8617e538612a72f03e758fa586997aa966

Observation 0961278c-cf23-4573-8a72-ec11bc76853c · outbound

This paper cites Hadzic, Taran Kota, Jimming He, Cristobal Eyzaguirre, Zane Durante, Manling Li, Jiajun Wu, and Fei-Fei Li.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Hadzic, Taran Kota, Jimming He, Cristobal Eyzaguirre, Zane Durante, Manling Li, Jiajun Wu, and Fei-Fei Li

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:33.894037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:33.894037Z digest=sha256:6685d5691b8ac5a8eb61f4ac03440229c979e3aa6b1d4c76a33f928347267dd8

Observation 409da519-2955-42e7-b6f6-3deb200c6084 · outbound

This paper cites Mecd: Unlocking multi-event causal discovery in video reasoning, 2024.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Mecd: Unlocking multi-event causal discovery in video reasoning, 2024

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:33.962297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:33.962297Z digest=sha256:784b63bf9393ee6329e7688a61f6808e21ad801ad7b469d83f5c9e8e61c0a2ae

Observation e2cf40d0-030b-4afe-b70b-34471a2baf45 · outbound

This paper cites On the suitability of reinforcement fine-tuning to visual tasks, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization On the suitability of reinforcement fine-tuning to visual tasks, 2025

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:34.037141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:34.037141Z digest=sha256:a6236a2bda170326392159f89cf207393836ccd3352a7baff036fea07447ec14

Observation 04afaa31-a2d4-4e3e-acfd-b8b2e62b02fa · outbound

This paper cites Videovista-culturallingo: 360◦ horizons-bridging cultures, languages, and domains in video comprehension, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Videovista-culturallingo: 360◦ horizons-bridging cultures, languages, and domains in video comprehension, 2025

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:34.112293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:34.112293Z digest=sha256:f9a59f334d40e6866c915e4ecc350c678d56225e2ec6a4147c7670106560c3d2

Observation 5266d1d5-e3f8-41fa-b581-b017a02346c8 · outbound

This paper cites Visrl: Intention-driven visual perception via reinforced reasoning, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Visrl: Intention-driven visual perception via reinforced reasoning, 2025

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:34.190497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:34.190497Z digest=sha256:a857d9bc1d7ce73e0bb2de0535ec36a9589703a076d66470bd639643858e454d

Observation 37bae379-aac5-41b5-906a-b7e3ccdcd4af · outbound

This paper cites Expanding performance boundaries of open-source multimodal models with model, data, and test-time scaling, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Expanding performance boundaries of open-source multimodal models with model, data, and test-time scaling, 2025

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:34.286812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:34.286812Z digest=sha256:f6fe56a8dccb23e2228ffe1091b94635375c3eef7570f1ad75ab13a17bd0adb4

Observation b9753eec-9007-448c-9e2d-a954ce290fb3 · outbound

This paper cites Videollama 2: Advancing spatial- temporal modeling and audio understanding in video-llms, 2024.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Videollama 2: Advancing spatial- temporal modeling and audio understanding in video-llms, 2024

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:34.362878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:34.362878Z digest=sha256:76440644bd9ad7bf9ebc506a36e9111f5855fb5580eb421f8cf4fbeecfb0ddb8

Observation b603b5b1-3815-42a2-9a04-e5e9eb6ee7b9 · outbound

This paper cites Skywork r1v2: Multimodal hybrid reinforcement learning for reasoning, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Skywork r1v2: Multimodal hybrid reinforcement learning for reasoning, 2025

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:34.444821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:34.444821Z digest=sha256:3c40f8fa9b142148330c08cb816592b59a9b7caab92b3ec69b7bd8cd18e9a0df

Observation 6b4c2fd8-0f4c-4296-8aa9-21ea56e7bcc9 · outbound

This paper cites Instructblip: Towards general-purpose vision-language models with instruction tuning, 2023.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Instructblip: Towards general-purpose vision-language models with instruction tuning, 2023

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:34.524495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:34.524495Z digest=sha256:f88e3ea0ed7c4ca122d4f53b9f262113592293db3748bb21daa0da0a3d3b3af5

Observation 5ee485d6-b41d-423c-ba79-8f6a363217c0 · outbound

This paper cites Mm-spatial: Exploring 3d spatial understanding in multimodal llms, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Mm-spatial: Exploring 3d spatial understanding in multimodal llms, 2025

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:34.598577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:34.598577Z digest=sha256:cc8d3c4b29727a5dba124b31ab6f0644f7099009a6a0b9c20e266e433bb96d04

Observation 3c6f3ce1-2ded-4a01-acd7-42caf7e37e5d · outbound

This paper cites Deepseek-r1: Incentivizing reasoning capability in llms via reinforcement learning, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Deepseek-r1: Incentivizing reasoning capability in llms via reinforcement learning, 2025

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:34.672404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:34.672404Z digest=sha256:6e76ede72b23ede8a8ea813da2aeeff34f1e2ada6ba61e004b3e4001d8799055

Observation 1d556e72-8013-424b-a2f7-f47a33501972 · outbound

This paper cites Boosting the generalization and reasoning of vision language models with curriculum reinforcement learning, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Boosting the generalization and reasoning of vision language models with curriculum reinforcement learning, 2025

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:34.746770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:34.746770Z digest=sha256:8fe7a02d88ca7ee35cccf604301a792fa8965900a3b18b9d127fad9fd9fd7e01

Observation 3951efe9-2156-475d-b43b-2fae23be11df · outbound

This paper cites Insight-v: Exploring long-chain visual reasoning with multimodal large language models, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Insight-v: Exploring long-chain visual reasoning with multimodal large language models, 2025

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:34.820578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:34.820578Z digest=sha256:d0c889a0c7c00db1cc491344587019931f7c422f63fa5f76959c4972c74fc193

Observation 045a8a13-2a17-4340-b703-f407c3577525 · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale, 2021.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization An image is worth 16x16 words: Transformers for image recognition at scale, 2021

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:34.899833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:34.899833Z digest=sha256:12eb13479172bb1005ab30f0cb98c2750703f981ebd884782e829c5789b1cc14

Observation bb9e0ad6-7a21-4b74-8675-62abdd22a36a · outbound

This paper cites Video-of-thought: Step-by-step video reasoning from perception to cognition, 2024.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Video-of-thought: Step-by-step video reasoning from perception to cognition, 2024

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:34.948834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:34.948834Z digest=sha256:3c32e44e4d3cfe183dac8084d9df50cad7d1285dddb91eb119df0d28a934ffd3

Observation 4df0315d-e9ad-40f6-a897-404bfd87ced8 · outbound

This paper cites Video-r1: Reinforcing video reasoning in mllms, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Video-r1: Reinforcing video reasoning in mllms, 2025

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:35.012190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:35.012190Z digest=sha256:9835d41d09e5769bb59efb33d5c270c0a004c6445d5b6516a3973b323e9b0a6e

Observation eefe869a-ff52-4113-b12c-7a877552bf95 · outbound

This paper cites Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:35.129702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:35.129702Z digest=sha256:761c095d5518dd725b193e409e85d3bb07ad0417dbc9b285f04d68f4e77cc676

Observation 62ae0e71-efab-4eed-8c25-3071c665ac49 · outbound

This paper cites Ampo: Active multi-preference optimization, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Ampo: Active multi-preference optimization, 2025

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:44.568271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:35.229600Z digest=sha256:e012d906de938b54267ddf5929fda3bd8c302ef346fa9fd22c4e7a347ef43b22

Observation 06263630-fe84-4bef-acbe-3daaedff91e5 · outbound

This paper cites Video-mmmu: Evaluating knowledge acquisition from multi-discipline professional videos, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Video-mmmu: Evaluating knowledge acquisition from multi-discipline professional videos, 2025

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:35.301478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:35.301478Z digest=sha256:077df37d9f82f12290a9cc4a53f987a40287cf2b784a0f79e3ed152ce6aee533

Observation d6de313b-5d04-4d0e-9610-853f4aa3ec4a · outbound

This paper cites Vision-r1: Incentivizing reasoning capability in multimodal large language models, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Vision-r1: Incentivizing reasoning capability in multimodal large language models, 2025

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:44.542641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:35.426624Z digest=sha256:ec32383d02c3a712465dc0358cc0c126c83e163403535be123acc2560bdd12e0

Observation 6c110c8a-34eb-4a67-8c7c-b71ba96b09e9 · outbound

This paper cites GPT-4o System Card.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization GPT-4o System Card

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:35.522280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:35.522280Z digest=sha256:91ce4815edbbf52203e9a7a14144c0b6bbbcd96913ec1483325d85beb7719944

Observation 4f81678d-94c2-46d6-8cb3-a7c0ceb43465 · outbound

This paper cites Llava-onevision: Easy visual task transfer, 2024.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Llava-onevision: Easy visual task transfer, 2024

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:35.624195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:35.624195Z digest=sha256:82dbf39cde2218c001a3007d4dab0d1d7d3b600a1374aa0e953ac8b520ca8eb9

Observation 47f11427-e3c6-4d4e-b5f8-320f644cef56 · outbound

This paper cites mplug: Effective and efficient vision-language learning by cross-modal skip-connections, 2022.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization mplug: Effective and efficient vision-language learning by cross-modal skip-connections, 2022

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:44.514345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:35.702600Z digest=sha256:29ab832850066dd3f16b887ba40b74da11e9c5ccbf55d67e492173f87615badf

Observation 694d051c-050d-4572-a1b7-d061a7db60cb · outbound

This paper cites Videochat: Chat-centric video understanding, 2024.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Videochat: Chat-centric video understanding, 2024

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:35.785498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:35.785498Z digest=sha256:0e66785c457e7d824c10b8666863583281628d064ed3c6f099a8acb61238f40f

Observation 28608a34-9347-49bc-8472-e0d4921349ce · outbound

This paper cites Mvbench: A comprehensive multi-modal video understanding benchmark, 2023.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Mvbench: A comprehensive multi-modal video understanding benchmark, 2023

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:35.876567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:35.876567Z digest=sha256:9f891563612a890bf7c206546e547d9665acf15cdb077d6ac90f9186599a2420

Observation 2606f9a0-c124-4410-befb-498f6dd6566e · outbound

This paper cites Videochat-r1: Enhancing spatio-temporal perception via reinforcement fine-tuning, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Videochat-r1: Enhancing spatio-temporal perception via reinforcement fine-tuning, 2025

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:44.475598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:35.985818Z digest=sha256:2e6d2031b5a72efc6dce255595669e4fb01a951c60d18b6d583e9457e51e66ee

Observation ca29190a-592f-4596-bf5c-8469b8adeb75 · outbound

This paper cites Lmeye: An interactive perception network for large language models.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Lmeye: An interactive perception network for large language models

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:44.458668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:36.059418Z digest=sha256:930d62af13af16c7ad47600df3b85a425aaf8e163fef487e78616b7a9175d3d4

Observation 36f6cafb-21df-4428-85ea-3a801d322d36 · outbound

This paper cites Uni-moe: Scaling unified multimodal llms with mixture of experts.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Uni-moe: Scaling unified multimodal llms with mixture of experts

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:44.439968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:36.131342Z digest=sha256:e367e04ad1784a8e968e266a61e66166c51f25558baad1728aae252629f37c26

Observation 5564ec1e-1a3c-4ca7-9f33-586185119adb · outbound

This paper cites Perception, Reason, Think, and Plan: A Survey on Large Multimodal Reasoning Models.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Perception, Reason, Think, and Plan: A Survey on Large Multimodal Reasoning Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:36.210901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:36.210901Z digest=sha256:93980dbf19dd223ef54c856c7d900660ba430ed727ea2c50374342a1797e35fb

Observation 41962da8-18cf-4446-99e1-a953da3e28de · outbound

This paper cites STI-Bench: Are MLLMs Ready for Precise Spatial-Temporal World Understanding?.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization STI-Bench: Are MLLMs Ready for Precise Spatial-Temporal World Understanding?

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:36.284175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:36.284175Z digest=sha256:0e13df64a12699f6f0379f91552ec73e7411c9bbdb842a83f2c8b39e84af4b75

Observation c074983b-78b9-48a4-85ae-111466a8ce31 · outbound

This paper cites Vila: On pre-training for visual language models, 2024.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Vila: On pre-training for visual language models, 2024

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:44.420911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:36.381623Z digest=sha256:a03a992f81121bc73b911175b68571a44d4cd8c71f38f59ad15a07f086854362

Observation 9c71c0c2-5ca9-4ed6-82d6-f067018ce985 · outbound

This paper cites Spatialcot: Advancing spatial reasoning through coordinate alignment and chain-of-thought for embodied task planning, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Spatialcot: Advancing spatial reasoning through coordinate alignment and chain-of-thought for embodied task planning, 2025

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:44.401740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:36.433148Z digest=sha256:499f8c526e579d6cbf443bf00604544fbc64832e993dfa4e38cedd6a6a3d27c5

Observation 7ed38859-3287-459d-b6bc-99df5e5c01e4 · outbound

This paper cites TempCompass: Do Video LLMs Really Understand Videos?.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization TempCompass: Do Video LLMs Really Understand Videos?

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:36.477596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:36.477596Z digest=sha256:74804c844fd4384c3fadcea27f90d46cc60af619fd4eaccf89d7137f46472fb9

Observation 7a487b60-f8b8-41db-9bdb-01b659bc5e59 · outbound

This paper cites Videomind: A chain-of-lora agent for long video reasoning, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Videomind: A chain-of-lora agent for long video reasoning, 2025

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:44.380424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:36.570295Z digest=sha256:278e8860e11f8b4fb6033c9041a72d2300086f8381258015c9a5dd2f7fd6c304

Observation 13772a12-cafc-4c28-8b44-6cbb865b1596 · outbound

This paper cites Seg-zero: Reasoning-chain guided segmentation via cognitive reinforcement, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Seg-zero: Reasoning-chain guided segmentation via cognitive reinforcement, 2025

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:44.362075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:36.627352Z digest=sha256:fcfad6e9f8b9c03bafafaac3583d923a575eb64958cd1de55bea2f725e6c7904

Observation 5eeef245-9d51-4420-a4cd-79593a5113c5 · outbound

This paper cites Video swin transformer, 2021.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Video swin transformer, 2021

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:44.341147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:36.678713Z digest=sha256:8109e4b40b46566e6c9ab00562439ce221546cb274cff035576e8601e2ff673f

Observation 5a8ded3c-ef20-4f05-bd47-5f30ccd8638e · outbound

This paper cites Visual-rft: Visual reinforcement fine-tuning, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Visual-rft: Visual reinforcement fine-tuning, 2025

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:36.726888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:36.726888Z digest=sha256:11be4b8f6c65506a3abfb82f3ce666282add41dc0278343cee104cb09885c408

Observation 5f3e6f78-c2a0-48b0-bc17-8dd518d17540 · outbound

This paper cites Othink- mr1: Stimulating multimodal generalized reasoning capabilities via dynamic reinforcement learning, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Othink- mr1: Stimulating multimodal generalized reasoning capabilities via dynamic reinforcement learning, 2025

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:44.311845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:36.822010Z digest=sha256:d246eb3f5fc633a223c5248cac8141143ec56dd3633a234488875eb80eda61c5

Observation 20223307-5c25-4dfc-bbae-66b86c8b6fce · outbound

This paper cites Gui-r1 : A generalist r1-style vision-language action model for gui agents, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Gui-r1 : A generalist r1-style vision-language action model for gui agents, 2025

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:44.293008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:36.900953Z digest=sha256:2c0921daf7ddb5d2ef8122549f9a324b2efed31533a7a88c2f232e00da5bd223

Observation a6cce647-dddc-463f-b3e6-9ef9099aa604 · outbound

This paper cites Video-chatgpt: Towards detailed video understanding via large vision and language models, 2024.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Video-chatgpt: Towards detailed video understanding via large vision and language models, 2024

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:36.947665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:36.947665Z digest=sha256:0dd57df5da20fd61260ffe7709fa013adc0488dbf3ffcefc865e3a679cde3feb

Observation 5e152623-b3b3-4979-8f71-49db77abb6f1 · outbound

This paper cites Mm-eureka: Exploring the frontiers of multimodal reasoning with rule-based reinforcement learning, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Mm-eureka: Exploring the frontiers of multimodal reasoning with rule-based reinforcement learning, 2025

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:36.992181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:36.992181Z digest=sha256:de6b95f8bdfd20485e7b7c4fa26a0bdf4d9b3f1b3548a1535e2980857c14d4f5

Observation b4fe77fa-5fdc-4ef2-a39f-24ec1e725581 · outbound

This paper cites Video transformer network, 2021.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Video transformer network, 2021

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:44.256306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:37.087623Z digest=sha256:2de06348c9fdf39028e24508dcdc43ff042239c4eb8548916744ab985ebf5b13

Observation 2e798a55-98d7-456b-91d9-1a14b0132818 · outbound

This paper cites Dinov2: Learning robust visual features without supervision, 2024.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Dinov2: Learning robust visual features without supervision, 2024

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:37.161812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:37.161812Z digest=sha256:dad93babe9aeacd0a4aa51e9fdc9f666836c8985601329f53a089d49ff751e4c

Observation 16920c04-7cd3-4975-90fb-6ea98b4cf632 · outbound

This paper cites Spatial-r1: Enhancing mllms in video spatial reasoning, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Spatial-r1: Enhancing mllms in video spatial reasoning, 2025

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:44.229929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:37.218409Z digest=sha256:97060fcd41d9d68e9ec8b4ab695526386e79023747c9c7166ec84ff8a9009e16

Observation 5ba9fb0b-7020-4977-acc0-70fdc2ecbb05 · outbound

This paper cites an unresolved cited work.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:25:44.212313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:37.291636Z digest=sha256:8f5423d3dd8c31a3d4dfb008f2d939bda245af528184c42d0bc6e022fca9c726

Observation 274d75a4-99c2-4fae-b82f-8c8ea323620a · outbound

This paper cites Lmm-r1: Empowering 3b lmms with strong reasoning abilities through two-stage rule-based rl, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Lmm-r1: Empowering 3b lmms with strong reasoning abilities through two-stage rule-based rl, 2025

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:37.358642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:37.358642Z digest=sha256:e27f520cb813c46dacd5af8c8b55592aaa72238d76ef9114fb6733632c61924f

Observation 6ae60161-9a1d-4baf-983b-5b9857f9dbec · outbound

This paper cites Learning transferable visual models from natural language supervision, 2021.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Learning transferable visual models from natural language supervision, 2021

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:37.441654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:37.441654Z digest=sha256:43cbbf847780ca4fd3cdbb589a7d8dedabdfcfd198c307f9f67183476b9195ff

Observation d41aa557-de70-43b5-9dc6-904fbcf18037 · outbound

This paper cites Manning, and Chelsea Finn.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Manning, and Chelsea Finn

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:37.508376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:37.508376Z digest=sha256:11edbe122fcfba936dbea5d538cd4f2a2795f402dace27c2897fdc9cded0c316

Observation 20d7cbe9-da84-4ea7-bb10-c5d45ff3169c · outbound

This paper cites Plummer, Ranjay Krishna, Kuo-Hao Zeng, and Kate Saenko.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Plummer, Ranjay Krishna, Kuo-Hao Zeng, and Kate Saenko

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:44.164453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:37.578740Z digest=sha256:9e3a4d3fbb7b0d09212703787643f08b5e9fc835a11d901ab1cd629ed7198d26

Observation 2ac9b6b0-24dd-445c-bd92-afbc4ea05eae · outbound

This paper cites Proximal policy optimization algorithms, 2017.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Proximal policy optimization algorithms, 2017

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:37.624830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:37.624830Z digest=sha256:9b90190deab50ad781158aaf2b0e0d0cef723dd48fca31d73ec9a06f73624857

Observation 37d9daa0-a665-4770-9ce8-60933f1e9014 · outbound

This paper cites Tomato: Assessing visual temporal reasoning capabilities in multimodal foundation models, 2024.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Tomato: Assessing visual temporal reasoning capabilities in multimodal foundation models, 2024

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:44.137898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:37.677996Z digest=sha256:d3e28b07f1aabbe6ff8da4813e8ae1735b740da826f0f7f1fe41609a5e7daff2

Observation f98fa328-6db1-4d87-9835-ddb4c4c3bf3f · outbound

This paper cites an unresolved cited work.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:37.777362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:37.777362Z digest=sha256:4a365fbdd6d8a5347f6c1a2f0fdd93eeb7f33e0ecb0a2aec6304b1754b697171

Observation 31de27d0-f23a-4a18-bda0-eaba13153a2c · outbound

This paper cites Efficient reinforcement finetuning via adaptive curriculum learning, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Efficient reinforcement finetuning via adaptive curriculum learning, 2025

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:37.839585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:37.839585Z digest=sha256:c3412d97fbb8d6fc8285ff453081fc04be95962a7e49b83db8605f391fa4a4e5

Observation 2b28d21e-c5ed-4f3f-b71e-5b423f72ca50 · outbound

This paper cites Mm-verify: Enhancing multimodal reasoning with chain-of-thought verification, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Mm-verify: Enhancing multimodal reasoning with chain-of-thought verification, 2025

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:44.100649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:37.907304Z digest=sha256:6d6319af11e5d34552951c6efafa75bf8ea6caf1ea6bc0e2326264eefe7fdb20

Observation c7cc9cbe-8314-4493-aaf9-6f00923459ea · outbound

This paper cites Reason-rft: Reinforcement fine-tuning for visual reasoning, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Reason-rft: Reinforcement fine-tuning for visual reasoning, 2025

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:44.082649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:37.979083Z digest=sha256:a62d28c01d3788f66642875375310d2f50c659dea0077c88758d713925a54abb

Observation d3d5cbb0-b7a8-462e-8de6-9ae7dfdde0bd · outbound

This paper cites Game-theoretic regularized self-play alignment of large language models, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Game-theoretic regularized self-play alignment of large language models, 2025

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:44.065611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:38.070074Z digest=sha256:f3f47932125b163f0d4b690d8bc2e1c8bf745fd9777eaa7293b87ebe4616bedb

Observation b14dd7fb-4bb3-45f1-8dcd-927f57756db1 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context, 2024.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context, 2024

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:44.047418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:38.149756Z digest=sha256:6687ec4ed76a44f9da119be8941750e56be915da982e49300dd1965aa3396e9a

Observation 81d046a8-3033-4a98-aabe-f111467af0b0 · outbound

This paper cites Gemma 3 technical report, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Gemma 3 technical report, 2025

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:44.030350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:38.227195Z digest=sha256:618d2db6ffe621dd3b9c8210e7d286d43ca865934f0cf6887712aeccd011dfe1

Observation b0b7e909-c990-4a8b-a390-46ecc82fa521 · outbound

This paper cites Kimi-vl technical report, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Kimi-vl technical report, 2025

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:44.013279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:38.315286Z digest=sha256:1a3474e5001b9469511d0b78e6657614d89acbb6a1b93d36443272ff4969a7af

Observation 458ec069-71a3-4c01-8b75-3725b593243a · outbound

This paper cites Model cards & prompt formats-llama 3.2, 2024.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Model cards & prompt formats-llama 3.2, 2024

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:43.995469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:38.408579Z digest=sha256:98cdaf1912906893cfba1c549b8db13dbb429592dca0871bcd87700679ba975e

Observation 4a57df1e-44aa-491b-a32c-55f7f36a3b64 · outbound

This paper cites Vila: On pre-training for visual language models, 2024.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Vila: On pre-training for visual language models, 2024

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:43.979505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:38.508413Z digest=sha256:827b7baabac133c2cd4e35027699a03311510a91a2121c3faa058cd40d352a16

Observation 56c6dde5-cc56-4c18-8bd5-efddf967b176 · outbound

This paper cites Goucher, et al.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Goucher, et al

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:43.961920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:38.586583Z digest=sha256:aef8f57e0605a500278d0e668929d864417a8b72730bcba9f54be1e4f71926e5

Observation edf34835-998e-4cf8-85e2-4d5ff31e0076 · outbound

This paper cites Qwen3: Think deeper, act faster, April 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Qwen3: Think deeper, act faster, April 2025

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:43.943223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:38.661955Z digest=sha256:4e8c64908ff66da41169d029c9286fa7b8d08aef8cef21ebd343fc0bc0c17d65

Observation 3a4ccaf4-f190-434d-98f5-00bde03988c8 · outbound

This paper cites VL-Rethinker: Incentivizing Self-Reflection of Vision-Language Models with Reinforcement Learning.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization VL-Rethinker: Incentivizing Self-Reflection of Vision-Language Models with Reinforcement Learning

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:38.751659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:38.751659Z digest=sha256:1d4b8fbc37c22fd49b4785f4a94a3189b705be9a3a21d42f0ce53912b165073a

Observation de1a0619-fe07-467a-b6b7-99fac19ae44e · outbound

This paper cites Piecing it all together: Verifying multi-hop multimodal claims, 2024.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Piecing it all together: Verifying multi-hop multimodal claims, 2024

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:43.926036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:38.845495Z digest=sha256:3efcf491ec6ba90e48c734f68b9c69aca5f3b0ece1fba1058febd5312415433f

Observation df2376cd-cd36-41ec-9c1e-943cdf29682d · outbound

This paper cites Lvbench: An extreme long video understanding benchmark, 2024.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Lvbench: An extreme long video understanding benchmark, 2024

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:38.942661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:38.942661Z digest=sha256:ef1c90d64b32b3df72a12abe3d35484c4f93ff766d54f15d8ef522e0ab57fcc7

Observation 1e47208b-c819-471e-8bc0-f136c1983262 · outbound

This paper cites Sota with less: Mcts-guided sample selection for data-efficient visual reasoning self-improvement, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Sota with less: Mcts-guided sample selection for data-efficient visual reasoning self-improvement, 2025

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:43.899023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:39.057589Z digest=sha256:975898818ec34371f787f295b95efeabd5c98512b22a242648df72f80fc4a6af

Observation 93bdbdc0-9ed4-437a-86c3-d2a7a9d46d85 · outbound

This paper cites Internvideo2.5: Empowering video mllms with long and rich context modeling, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Internvideo2.5: Empowering video mllms with long and rich context modeling, 2025

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:39.162833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:39.162833Z digest=sha256:0269f2f03cbce37cf4047f2989abea0c4b50b31d1af5dfde86e1f6b2539a2729

Observation 1d6b0d26-3bed-48fa-be71-bbf7d253e879 · outbound

This paper cites Unified multimodal chain-of-thought reward model through reinforcement fine-tuning, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Unified multimodal chain-of-thought reward model through reinforcement fine-tuning, 2025

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:39.290885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:39.290885Z digest=sha256:d203a7b793cddfddf73289a0c57194f6320cb181078e94608618f77fe0867597

Observation 6cbbd6c0-91cf-4e2c-8721-c2d947882dca · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models, 2023.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Chain-of-thought prompting elicits reasoning in large language models, 2023

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:39.407423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:39.407423Z digest=sha256:e6edd577840ef3c08e61d75c78a7d74b050545dde0a1582debc4525587388712

Observation aaa784da-4c0f-4057-9298-2f5431b7b7fa · outbound

This paper cites Videorope: What makes for good video rotary position embedding?, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Videorope: What makes for good video rotary position embedding?, 2025

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:43.848334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:39.524418Z digest=sha256:6bc1122a63cc4486edd42f56579b4a104907c01a767b455adb02c388667ef20b

Observation d1d8dda5-cc5b-4249-9587-6db55b0b4828 · outbound

This paper cites Longvideobench: A benchmark for long-context interleaved video-language understanding, 2024.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Longvideobench: A benchmark for long-context interleaved video-language understanding, 2024

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:39.600130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:39.600130Z digest=sha256:78a28c87b772425706364b5916dcef2f36ce7b388f59fc1776dedf79fffa9745

Observation af16f7f3-7145-4de7-aa47-b97cb4c42464 · outbound

This paper cites St-think: How multimodal large language models reason about 4d worlds from ego-centric videos, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization St-think: How multimodal large language models reason about 4d worlds from ego-centric videos, 2025

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:43.613616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:39.720983Z digest=sha256:5f83ff41cec98e435a9faa7e4e7c6aa98b4334993cb4628380d8be6fc569dde2

Observation 58361d60-a467-4c6a-84df-43188927f974 · outbound

This paper cites Self-play preference optimization for language model alignment, 2024.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Self-play preference optimization for language model alignment, 2024

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:39.783059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:39.783059Z digest=sha256:4c86cd355356a3ec2d1b411ff89090f6731ede0fe0447de615d3b989c0e7660f

Observation e4122e80-3584-4c60-b78a-044ff0624539 · outbound

This paper cites Deepseek-vl2: Mixture-of-experts vision-language models for advanced multimodal understanding, 2024.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Deepseek-vl2: Mixture-of-experts vision-language models for advanced multimodal understanding, 2024

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:39.828082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:39.828082Z digest=sha256:03793df526b2ec9f1dd7d85cffcfb3ffb8a10b36261c2943aa5d7857a0ec8bf0

Observation 887025b9-a251-40eb-892e-22deded07598 · outbound

This paper cites Atomthink: A slow thinking framework for multimodal mathematical reasoning, 2024.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Atomthink: A slow thinking framework for multimodal mathematical reasoning, 2024

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:43.465569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:39.949967Z digest=sha256:cc4a554bb8fa4219e99b892c20f3c9aaa80b3e7a9edf6dda094755d38ad31ea5

Observation 7005e81d-407c-4fb1-bc4c-e4cd25d974e4 · outbound

This paper cites Echoink-r1: Exploring audio-visual reasoning in multimodal llms via reinforcement learning, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Echoink-r1: Exploring audio-visual reasoning in multimodal llms via reinforcement learning, 2025

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:43.228647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:40.067849Z digest=sha256:db3566bc7fa80dfe3b8410fc2fd22e5eac4915497963fb4a14abb89f4a64d2c7

Observation b0a3ded2-742d-49cf-997f-23e8cd3e3487 · outbound

This paper cites Redstar: Does scaling long-cot data unlock better slow-reasoning systems?, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Redstar: Does scaling long-cot data unlock better slow-reasoning systems?, 2025

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:40.138654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:40.138654Z digest=sha256:35dedb6a2a8f8727f3b5b8d58a5d548b145df9f84b00adeff5029c7f4b6d2e81

Observation 2f1148cd-e94b-4c22-9c6c-bc7ba63037d4 · outbound

This paper cites Thinking in Space: How Multimodal Large Language Models See, Remember, and Recall Spaces.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Thinking in Space: How Multimodal Large Language Models See, Remember, and Recall Spaces

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:40.216644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:40.216644Z digest=sha256:0a648b90b37d6bda6f05dfcaf2e043b5b22b31da8463c74f7c1f830ffa4903a6

Observation 94135652-4ba3-47a5-9413-8ca9ab4c4bbd · outbound

This paper cites R1-onevision: Advancing generalized multimodal reasoning through cross-modal formalization, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization R1-onevision: Advancing generalized multimodal reasoning through cross-modal formalization, 2025

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:40.264180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:40.264180Z digest=sha256:69848012bf919ed02f124c526f2ad654a679d00d72f83fda7a2f9c20a7418551

Observation 1422e568-a62c-4f38-ab79-dae7400bd33d · outbound

This paper cites mplug-owl3: Towards long image-sequence understanding in multi-modal large language models, 2024.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization mplug-owl3: Towards long image-sequence understanding in multi-modal large language models, 2024

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:40.339854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:40.339854Z digest=sha256:814bd9a5e2d98be49166626ea35caf4589ed7ae5998e0b52a071ca8b69cd21cb

Observation fc33e49c-0dbd-4179-8060-09fd3d69d3ef · outbound

This paper cites Dapo: An open-source llm reinforcement learning system at scale, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Dapo: An open-source llm reinforcement learning system at scale, 2025

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:40.447842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:40.447842Z digest=sha256:dccfd0a93a5d36ad3a47aff3eb31431347d943cfd707eb025ea4f0b836699a39

Observation 4af47f56-f40f-45d5-a768-769402dbf9aa · outbound

This paper cites an unresolved cited work.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Unresolved cited work

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:40.548668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:40.548668Z digest=sha256:a338747f8cd7a4c15880e2e047ae5fc6fe646daa5f87f80d9c89f11162c94d36

Observation 8be89472-ba36-48a2-acaf-6c8964894f8e · outbound

This paper cites Videollama 3: Frontier multimodal foundation models for image and video understanding, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Videollama 3: Frontier multimodal foundation models for image and video understanding, 2025

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:40.719622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:40.719622Z digest=sha256:3a54405c5084c84dba1474ee7e00cff54939f0a5a5bc0b314a1f68a0318ddd34

Observation 7f365eae-08de-4075-bbac-827c029c8241 · outbound

This paper cites Video-llama: An instruction-tuned audio-visual language model for video understanding, 2023.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Video-llama: An instruction-tuned audio-visual language model for video understanding, 2023

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:40.763798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:40.763798Z digest=sha256:fc0ad0c9fb77c23375d0ff44a04f3a35750d3792f2ba343fe74300b173963674

Observation d92d34fe-6d3f-4d77-a3ce-7707b6eb1004 · outbound

This paper cites From flatland to space: Teaching vision-language models to perceive and reason in 3d.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization From flatland to space: Teaching vision-language models to perceive and reason in 3d

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:40.821209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:40.821209Z digest=sha256:b47baa712457cbd6621ab39072ffcd7882be7bbd54ec768d337d9a6b1531b03e

Observation 5d02863f-399b-481a-9883-b7c9f4745e9e · outbound

This paper cites Long context transfer from language to vision, 2024.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Long context transfer from language to vision, 2024

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:40.877573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:40.877573Z digest=sha256:85f3b99db6932cc2f84d4d158a7559d08bed3602b92a1df87e8d068f7d964a4d

Observation 3b05ae24-3ab6-4989-8d53-d93c74eedfa3 · outbound

This paper cites Tinyllava-video-r1: Towards smaller lmms for video reasoning, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Tinyllava-video-r1: Towards smaller lmms for video reasoning, 2025

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:42.901334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:40.994352Z digest=sha256:e08d009137b7b262ccca4d68635c858ecc6db607fb0266ad088a2cc0df4d0e20

Observation 87dd1c2c-1b86-46ef-b568-3ddeb924ad6a · outbound

This paper cites Video instruction tuning with synthetic data, 2024.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Video instruction tuning with synthetic data, 2024

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:41.057201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:41.057201Z digest=sha256:eb94a2543fab02fab15e05a392519d76ba6dc871d792bdce1c0976ca729e1369

Observation b1ac7842-aec5-467a-a165-9a68c7a038f4 · outbound

This paper cites Openrft: Adapting reasoning foundation model for domain-specific tasks with reinforcement fine-tuning, 2024.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Openrft: Adapting reasoning foundation model for domain-specific tasks with reinforcement fine-tuning, 2024

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:42.697324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:41.091513Z digest=sha256:53dd81aee0d71c79268534c0e44e40db8aed2248048d0a4abbde481022e175b6

Observation a3773b89-0653-4ec5-83c5-ff8ee6856a14 · outbound

This paper cites MME-RealWorld: Could Your Multimodal LLM Challenge High-Resolution Real-World Scenarios that are Difficult for Humans?.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization MME-RealWorld: Could Your Multimodal LLM Challenge High-Resolution Real-World Scenarios that are Difficult for Humans?

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:41.182463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:41.182463Z digest=sha256:97be51dbef3d786c23a50c580d5d2b88623c23d173d2a0f10e1a0a4b8a6b9653

Observation f0588aaa-2156-472e-bc48-20aefa853e37 · outbound

This paper cites Multi- modal chain-of-thought reasoning in language models, 2024.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Multi- modal chain-of-thought reasoning in language models, 2024

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:42.505700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:41.332337Z digest=sha256:7f14703025ae08f02f7066f317ce34cccf63d12f1244bbe4c6067777a8701740

Observation 213df899-8b29-40ec-bc12-e26386b42818 · outbound

This paper cites R1-omni: Explainable omni-multimodal emotion recognition with reinforcement learning, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization R1-omni: Explainable omni-multimodal emotion recognition with reinforcement learning, 2025

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:41.463000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:41.463000Z digest=sha256:6f94a0e4962daab3dcefa8d0f5adaa9c157e6382cd40cb0114f6ae4ebfbb3e95

Observation 7a98b2a9-a395-4ccf-9c72-3c192bb9fa41 · outbound

This paper cites Mmvu: Measuring expert-level multi-discipline video understanding, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Mmvu: Measuring expert-level multi-discipline video understanding, 2025

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:41.566529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:41.566529Z digest=sha256:074ab67c1016a6c327bbe3354a077f59da202477c61d9010960e46acc8e7bc58

Observation 367c941e-1057-4472-9b44-5bb2de7f7d4a · outbound

This paper cites Villa: Video reasoning segmentation with large language model, 2025.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization Villa: Video reasoning segmentation with large language model, 2025

Reference 97

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:25:42.255405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:25:41.666069Z digest=sha256:dec588ae26003d724f686fe2405e1131f51b21158f0e368578715ea658840f11

Observation 2ccaa982-07eb-45cb-9427-ff88ce35dd4c · outbound

This paper cites MLVU: Benchmarking Multi-task Long Video Understanding.

VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization MLVU: Benchmarking Multi-task Long Video Understanding

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:41.790587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:41.790587Z digest=sha256:ef6330eb6e2af133f4f2ace0df4a2ad25e109fdcd5b5f7143fe146da758dd6dc

Pith citing papers

Observation 7f0f584d-4899-4432-99e9-981f89d29755 · inbound

Watch, Remember, Reason: Human-View Video Understanding with MLLMs cites this paper.

Watch, Remember, Reason: Human-View Video Understanding with MLLMs VerIPO: Cultivating Long Reasoning in Video-LLMs via Verifier-Gudied Iterative Policy Optimization

Reference 186

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:27:15.522720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T22:00:28.350003Z digest=sha256:14b0d20deb5a5f14c8b67ba9a3ac7159414858f214e2ae88bd0fd62ba99765b1