Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 12 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 44 inbound Pith citation observations for arXiv:2504.07934.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-11T00:36:01.053166Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T16:29:57.248099Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 884e815e-c5c4-4874-8ba5-349dc1358242 · inbound
OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b95d8041-f1e2-4f76-a90b-ab6dbf56b8f8 · inbound
Adaptive Chain-of-Focus Reasoning via Dynamic Visual Search and Zooming for Efficient VLMs SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 5a912158-fb1e-4d4a-a7ea-2af8962ac3b9 · inbound
R1-ShareVL: Incentivizing Reasoning Capability of Multimodal Large Language Models via Share-GRPO SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 967f09ed-0582-4c9f-8691-c1c707ee491e · inbound
Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 97
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a8b3ec2-4343-4921-9602-85df3d133902 · inbound
Point-RFT: Improving Multimodal Reasoning with Visually Grounded Reinforcement Finetuning SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2522fcd-2a76-43c9-970d-72cd89da76a5 · inbound
More Thinking, Less Seeing? Assessing Amplified Hallucination in Multimodal Reasoning Models SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39f7e458-7dd4-4b3b-9084-63c8f4085552 · inbound
Advancing Multimodal Reasoning via Reinforcement Learning with Cold Start SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9d85935-da37-4c4a-a56d-c37c6088d6a0 · inbound
ReAgent-V: A Reward-Driven Multi-Agent Framework for Video Understanding SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04625a6a-c12b-4741-a4d4-b573b244768e · inbound
SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73127110-abe6-4cc4-9e98-5bb17a5fec37 · inbound
What makes Reasoning Models Different? Follow the Reasoning Leader for Efficient Decoding SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9559091-c7ca-459c-8da7-146301ceae8d · inbound
ViCrit: A Verifiable Reinforcement Learning Proxy Task for Visual Perception in VLMs SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa107860-b77c-457c-bd26-239bb6829ccf · inbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 305f264d-089d-4270-bb78-93469a4b5cfe · inbound
VL-Cogito: Progressive Curriculum Reinforcement Learning for Advanced Multimodal Reasoning SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f055fe9b-4553-40eb-bc57-d58d410b375a · inbound
MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47e2937b-eda5-4503-9d59-4ba53a5bd6b7 · inbound
LLaVA-Critic-R1: Your Critic Model is Secretly a Strong Policy Model SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e60e136a-4e8a-40ab-bd9b-f467fdabf639 · inbound
Beyond Reasoning Gains: Mitigating General-Capability Forgetting in Large Reasoning Models SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 98
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 588ea2e5-6f30-46fa-9f66-4532d18477e5 · inbound
DeepEyesV2: Toward Agentic Multimodal Model SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 1478b99e-35d9-45ef-b481-a9237ec9ca09 · inbound
Learning Self-Correction in Vision-Language Models via Rollout Augmentation SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2d32b3c-dea8-4a39-80a6-782201174bce · inbound
ReMoT: Reinforcement Learning with Motion Contrast Triplets SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 106
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c71b368f-6819-4534-9280-31d87dd2d01d · inbound
Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 9d1ad665-f26d-4244-bee5-b3d8153020c9 · inbound
Understanding the Role of Hallucination in Reinforcement Post-Training of Multimodal Reasoning Models SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 70b3a7de-b31a-4d89-aa60-a6bcf3c4d78d · inbound
Act Wisely: Cultivating Meta-Cognitive Tool Use in Agentic Multimodal Models SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a1e423e2-ed19-4823-9544-cf39f1c1b1e9 · inbound
Cognitive Pivot Points and Visual Anchoring: Unveiling and Rectifying Hallucinations in Multimodal Reasoning Models SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a02e4845-b209-41b1-9f77-774585094b34 · inbound
Cognitive Pivot Points and Visual Anchoring: Unveiling and Rectifying Hallucinations in Multimodal Reasoning Models SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a9263d9-5e58-49e1-970e-4bfcfa85a00d · inbound
DR-MMSearchAgent: Deepening Reasoning in Multimodal Search Agents SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 4bbacae6-4a7b-4f3b-aaeb-2cb579d8f937 · inbound
SSL-R1: Self-Supervised Visual Reinforcement Post-Training for Multimodal Large Language Models SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f693a19e-16f1-44d6-b921-05860080a572 · inbound
Measure Twice, Click Once: Co-evolving Proposer and Visual Critic via Reinforcement Learning for GUI Grounding SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 06990ab1-b734-4eab-9a9e-73a7c6db852f · inbound
Building a Precise Video Language with Human-AI Oversight SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 39021625-41ec-4258-b314-47fc36e33e18 · inbound
CGC: Compositional Grounded Contrast for Fine-Grained Multi-Image Understanding SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d3f0cc98-2210-40c5-8975-c22e86850d8b · inbound
Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 7e640e53-610c-4f6a-bfda-62784bf90ac2 · inbound
Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation df048a35-a7fe-4496-b6f0-40e77bff5d6d · inbound
Perceptual Flow Network for Visually Grounded Reasoning SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 4d5b7c2b-cedb-443c-8a5a-f1b444c17e44 · inbound
CAVE: A Structured Credit Assignment Approach for Fragmented Visual Evidence Reasoning SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f7bfaa8b-890e-4b97-96d6-22cb3912dfb7 · inbound
Are Tools Always Beneficial? Learning to Invoke Tools Adaptively for Dual-Mode Multimodal LLM Reasoning SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 69a5d072-97fc-4bbb-8a42-0a05eb72b807 · inbound
Learning Spatiotemporal Sensitivity in Video LLMs via Counterfactual Reinforcement Learning SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation fd0d599e-858d-4314-b105-19810a7dee72 · inbound
AnE: Pushing the Reasoning Frontier of Multimodal LLMs via Anchor Evolution SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 76f19e51-4722-4db9-b7e5-7bdd6e629898 · inbound
Smart Picks in the Dark: Towards Efficient RLVR for Reasoning via Tracing Metacognitive Pivots SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 054e8c7e-f438-48cc-b2d1-32f7f80c9388 · inbound
DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d864721f-d0b0-4ffc-b656-4f7331d7ed9a · inbound
VeriEvol: Scaling Multimodal Mathematical Reasoning via Verifiable Evol-Instruct SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 38cde0c3-4f48-4539-a80e-1b65579130f1 · inbound
Latent Visual States for Efficient Multimodal Reasoning SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 6a07455b-2c78-496e-a08d-1d641ba59446 · inbound
No Place to Hide: Benchmarking Video Hallucination with Background-Controlled Pairs SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation bf766f73-9c7e-47ba-8f1b-5d285e2986d9 · inbound
Trace: A Taxonomy-Guided Environment for Multidomain Visual Reasoning SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b90d450a-95dd-4a70-accb-41ccb8724fa9 · inbound
Aligning Large Vision-Language Models at Test Time: A Trajectory-Guided Structured Sampling Approach SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21dea8d1-a349-4d73-afc9-6085a711eb7e · inbound
Multi-Branch Policy Optimization for Multimodal Large Language Models SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.