Pith. sign in

Paper Citation Record · LEDGER

SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

As of 12 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 44 inbound Pith citation observations for arXiv:2504.07934.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.07934 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 44 of 44 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T00:36:01.053166Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T16:29:57.248099Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 884e815e-c5c4-4874-8ba5-349dc1358242 · inbound

OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles cites this paper.

OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-05-19T06:59:03.299659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-19T06:59:03.112252Z digest=sha256:a0554811637ab6b60060d68b6f93460d4d2b820144eb69e3909baa063cc476d6

Observation b95d8041-f1e2-4f76-a90b-ab6dbf56b8f8 · inbound

Adaptive Chain-of-Focus Reasoning via Dynamic Visual Search and Zooming for Efficient VLMs cites this paper.

Adaptive Chain-of-Focus Reasoning via Dynamic Visual Search and Zooming for Efficient VLMs SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:35:13.201395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-17T05:35:13.118221Z digest=sha256:cf08b575ce99702a95a2ec745dad212ea3147a0f727846e1d79ee8de537ab3c8

Observation 5a912158-fb1e-4d4a-a7ea-2af8962ac3b9 · inbound

R1-ShareVL: Incentivizing Reasoning Capability of Multimodal Large Language Models via Share-GRPO cites this paper.

R1-ShareVL: Incentivizing Reasoning Capability of Multimodal Large Language Models via Share-GRPO SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T15:01:38.528553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:01:38.528553Z digest=sha256:7f13893f3306e4c7e0929b4043afbe0601c61fbcb9eb93db566dd5a50e4a5551

Observation 967f09ed-0582-4c9f-8691-c1c707ee491e · inbound

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models cites this paper.

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-07T14:31:17.317950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:31:17.317950Z digest=sha256:7642fb7a504a92cb7692f52fb181690c4d225e10943f4f4866ff03707e2276f7

Observation 9a8b3ec2-4343-4921-9602-85df3d133902 · inbound

Point-RFT: Improving Multimodal Reasoning with Visually Grounded Reinforcement Finetuning cites this paper.

Point-RFT: Improving Multimodal Reasoning with Visually Grounded Reinforcement Finetuning SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:05.877103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:05.877103Z digest=sha256:438b933f73ff8ecfb534b4d9f592fe43d33700481406b8a9f9b54d03e2e76c25

Observation b2522fcd-2a76-43c9-970d-72cd89da76a5 · inbound

More Thinking, Less Seeing? Assessing Amplified Hallucination in Multimodal Reasoning Models cites this paper.

More Thinking, Less Seeing? Assessing Amplified Hallucination in Multimodal Reasoning Models SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:35.404157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:35.404157Z digest=sha256:22de52a6a89dfc9e6cf587431b3db50cc3d4b07571b5335375046e2870730482

Observation 39f7e458-7dd4-4b3b-9084-63c8f4085552 · inbound

Advancing Multimodal Reasoning via Reinforcement Learning with Cold Start cites this paper.

Advancing Multimodal Reasoning via Reinforcement Learning with Cold Start SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T13:12:57.363121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:12:57.363121Z digest=sha256:263cee80e59d020b83016ce23a9588b6bd192432670252fb177aa93172d22d95

Observation f9d85935-da37-4c4a-a56d-c37c6088d6a0 · inbound

ReAgent-V: A Reward-Driven Multi-Agent Framework for Video Understanding cites this paper.

ReAgent-V: A Reward-Driven Multi-Agent Framework for Video Understanding SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T11:52:04.646937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:52:04.646937Z digest=sha256:a8c66d58a15c178feb6555401483bf7ba389cf21ae50996f4cfebedfbcb37790

Observation 04625a6a-c12b-4741-a4d4-b573b244768e · inbound

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis cites this paper.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:47.087104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:47.087104Z digest=sha256:d8a0bdadc4e4a4c51c39d85bf09656bdd07e31e37a33bd6f7971d0437c5a41e4

Observation 73127110-abe6-4cc4-9e98-5bb17a5fec37 · inbound

What makes Reasoning Models Different? Follow the Reasoning Leader for Efficient Decoding cites this paper.

What makes Reasoning Models Different? Follow the Reasoning Leader for Efficient Decoding SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T05:52:18.654116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:52:18.654116Z digest=sha256:967d8723a0b3a2d922c1a10b818eae6bde87c10881c93c33c3f34666ee6ee228

Observation a9559091-c7ca-459c-8da7-146301ceae8d · inbound

ViCrit: A Verifiable Reinforcement Learning Proxy Task for Visual Perception in VLMs cites this paper.

ViCrit: A Verifiable Reinforcement Learning Proxy Task for Visual Perception in VLMs SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T04:40:11.821726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:40:11.821726Z digest=sha256:b903c4c04f5bc56a8d7a65e34855b5499f70ca4fae16147a326d6b975966bc0a

Observation fa107860-b77c-457c-bd26-239bb6829ccf · inbound

MiCo: Multi-image Contrast for Reinforcement Visual Reasoning cites this paper.

MiCo: Multi-image Contrast for Reinforcement Visual Reasoning SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T22:10:22.226382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:10:22.226382Z digest=sha256:42227a187a3e83745edbecfc3b770409a2864b444a96629af8539dc0311ee653

Observation 305f264d-089d-4270-bb78-93469a4b5cfe · inbound

VL-Cogito: Progressive Curriculum Reinforcement Learning for Advanced Multimodal Reasoning cites this paper.

VL-Cogito: Progressive Curriculum Reinforcement Learning for Advanced Multimodal Reasoning SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-06T11:35:16.945350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:35:16.945350Z digest=sha256:01dc51e5d96344d4a4f906640380a9d87b8af743bcb905611aaaef9311a51092

Observation f055fe9b-4553-40eb-bc57-d58d410b375a · inbound

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models cites this paper.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.133767Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.133767Z digest=sha256:9d14433df61deab5b3266866b7a29ec504c9316adadbb55141d69db881a53b50

Observation 47e2937b-eda5-4503-9d59-4ba53a5bd6b7 · inbound

LLaVA-Critic-R1: Your Critic Model is Secretly a Strong Policy Model cites this paper.

LLaVA-Critic-R1: Your Critic Model is Secretly a Strong Policy Model SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-05T13:24:39.870088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T13:24:39.870088Z digest=sha256:347555c081f472b2a256bd82c533128613260ea4c570616334f8466ba38710e7

Observation e60e136a-4e8a-40ab-bd9b-f467fdabf639 · inbound

Beyond Reasoning Gains: Mitigating General-Capability Forgetting in Large Reasoning Models cites this paper.

Beyond Reasoning Gains: Mitigating General-Capability Forgetting in Large Reasoning Models SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-04T08:15:57.524813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T08:15:57.524813Z digest=sha256:211e59edc41eeebd260a5e6740a2f933895a5546d90f98111317260087714ceb

Observation 588ea2e5-6f30-46fa-9f66-4532d18477e5 · inbound

DeepEyesV2: Toward Agentic Multimodal Model cites this paper.

DeepEyesV2: Toward Agentic Multimodal Model SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-16T05:32:29.455296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-16T05:32:29.266583Z digest=sha256:f0d3631edadc1faeb76b435fa7fd685b607b872d5b55458946f4b70776c6b125

Observation 1478b99e-35d9-45ef-b481-a9237ec9ca09 · inbound

Learning Self-Correction in Vision-Language Models via Rollout Augmentation cites this paper.

Learning Self-Correction in Vision-Language Models via Rollout Augmentation SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-03T03:21:13.921539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T03:21:13.921539Z digest=sha256:f3ed756099a83ac0de9274004dd14efb48eb587c08f5069d7728978cf3391f64

Observation e2d32b3c-dea8-4a39-80a6-782201174bce · inbound

ReMoT: Reinforcement Learning with Motion Contrast Triplets cites this paper.

ReMoT: Reinforcement Learning with Motion Contrast Triplets SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 106

Resolution
unresolved
no resolver link, observed 2026-08-02T19:58:27.284572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T19:58:27.284572Z digest=sha256:e9a339df0aec1ec82556d3d5a609ea81abd532a4038a0bedb8d678a64af5f615

Observation c71b368f-6819-4534-9280-31d87dd2d01d · inbound

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models cites this paper.

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-21T12:00:04.251113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-21T11:56:29.743281Z digest=sha256:0b06272dfb8462eee8b6a8e8303b44556f87ad07231d6f25737856a0c6179373

Observation 9d1ad665-f26d-4244-bee5-b3d8153020c9 · inbound

Understanding the Role of Hallucination in Reinforcement Post-Training of Multimodal Reasoning Models cites this paper.

Understanding the Role of Hallucination in Reinforcement Post-Training of Multimodal Reasoning Models SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:53:16.231096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-13T20:48:52.130130Z digest=sha256:7025a6eb2691bd1c6522289409b6f33641b3c186c468f242b0bc11a8e34e9cf6

Observation 70b3a7de-b31a-4d89-aa60-a6bcf3c4d78d · inbound

Act Wisely: Cultivating Meta-Cognitive Tool Use in Agentic Multimodal Models cites this paper.

Act Wisely: Cultivating Meta-Cognitive Tool Use in Agentic Multimodal Models SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:20:53.915067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-10T18:35:21.514502Z digest=sha256:f4e3002db146a16fdc4c4d3274df160ce84a6b267b4ba343d620ee28462abb6b

Observation a1e423e2-ed19-4823-9544-cf39f1c1b1e9 · inbound

Cognitive Pivot Points and Visual Anchoring: Unveiling and Rectifying Hallucinations in Multimodal Reasoning Models cites this paper.

Cognitive Pivot Points and Visual Anchoring: Unveiling and Rectifying Hallucinations in Multimodal Reasoning Models SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 81

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:25:59.078095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-10T16:03:15.222571Z digest=sha256:7436c82d27b64a40327a2d3a628df91c3d874b419b05fb3e36cbdedba407f10d

Observation a02e4845-b209-41b1-9f77-774585094b34 · inbound

Cognitive Pivot Points and Visual Anchoring: Unveiling and Rectifying Hallucinations in Multimodal Reasoning Models cites this paper.

Cognitive Pivot Points and Visual Anchoring: Unveiling and Rectifying Hallucinations in Multimodal Reasoning Models SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 92

Resolution
unresolved
no resolver link, observed 2026-07-12T22:48:45.647588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T22:48:45.647588Z digest=sha256:2b55279171fb1ba3f6cb78d860e77eeedba47526ad7b2174b23efc2a412189b4

Observation 6a9263d9-5e58-49e1-970e-4bfcfa85a00d · inbound

DR-MMSearchAgent: Deepening Reasoning in Multimodal Search Agents cites this paper.

DR-MMSearchAgent: Deepening Reasoning in Multimodal Search Agents SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 74

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T12:41:01.163336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-10T03:21:30.732925Z digest=sha256:cc448303e409c375b974821a87fd34183b5e4ed065fadf7616741452b664e7b1

Observation 4bbacae6-4a7b-4f3b-aaeb-2cb579d8f937 · inbound

SSL-R1: Self-Supervised Visual Reinforcement Post-Training for Multimodal Large Language Models cites this paper.

SSL-R1: Self-Supervised Visual Reinforcement Post-Training for Multimodal Large Language Models SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:36:08.856150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-10T01:23:32.849326Z digest=sha256:eec71fcd1fd716baabd394b3fe5a4b202cd99ff377883f5797b62971c8c6d31b

Observation f693a19e-16f1-44d6-b921-05860080a572 · inbound

Measure Twice, Click Once: Co-evolving Proposer and Visual Critic via Reinforcement Learning for GUI Grounding cites this paper.

Measure Twice, Click Once: Co-evolving Proposer and Visual Critic via Reinforcement Learning for GUI Grounding SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 91

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T14:16:06.142584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-09T23:05:05.251150Z digest=sha256:3167c89de05526d00380cf60e1972b49660cd310d8c3d34622f20eee14a65a61

Observation 06990ab1-b734-4eab-9a9e-73a7c6db852f · inbound

Building a Precise Video Language with Human-AI Oversight cites this paper.

Building a Precise Video Language with Human-AI Oversight SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:46:04.410651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-10T00:37:31.858728Z digest=sha256:c9e6dc37a98de4493e6371138c3ae2c5c2f4fea989ca06e5dbf8b9892bcef674

Observation 39021625-41ec-4258-b314-47fc36e33e18 · inbound

CGC: Compositional Grounded Contrast for Fine-Grained Multi-Image Understanding cites this paper.

CGC: Compositional Grounded Contrast for Fine-Grained Multi-Image Understanding SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:16:08.355431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-08T12:26:01.568507Z digest=sha256:e3a2d680d839d111fd1c69207cb5d4349ed666585f741406cbc1df0da92fe8b9

Observation d3f0cc98-2210-40c5-8975-c22e86850d8b · inbound

Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs cites this paper.

Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 79

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:01:22.419357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-09T18:53:06.494640Z digest=sha256:eb5e2d680acdab7b0efb6159c31c513adbc4e1a72b23416a3c3852edac709eb0

Observation 7e640e53-610c-4f6a-bfda-62784bf90ac2 · inbound

Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs cites this paper.

Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 79

Resolution
verified exact
arxiv_id, observed 2026-05-11T01:50:51.588858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-11T01:49:15.136031Z digest=sha256:631a5fcbf9a56b9a2866ab1d45c2917db049cc690911a0f175a58a100ddd91fb

Observation df048a35-a7fe-4496-b6f0-40e77bff5d6d · inbound

Perceptual Flow Network for Visually Grounded Reasoning cites this paper.

Perceptual Flow Network for Visually Grounded Reasoning SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:15:38.894927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-08T18:40:55.753827Z digest=sha256:cf6eafcbedd3271fd0d9c2763a355322458cb419d685c440f0d39519d0ac906d

Observation 4d5b7c2b-cedb-443c-8a5a-f1b444c17e44 · inbound

CAVE: A Structured Credit Assignment Approach for Fragmented Visual Evidence Reasoning cites this paper.

CAVE: A Structured Credit Assignment Approach for Fragmented Visual Evidence Reasoning SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-20T21:13:44.763007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-20T21:10:56.533649Z digest=sha256:6287f22a9095c3ec69638a77526d1330f7f441b1b5cf674d63b4be73de437977

Observation f7bfaa8b-890e-4b97-96d6-22cb3912dfb7 · inbound

Are Tools Always Beneficial? Learning to Invoke Tools Adaptively for Dual-Mode Multimodal LLM Reasoning cites this paper.

Are Tools Always Beneficial? Learning to Invoke Tools Adaptively for Dual-Mode Multimodal LLM Reasoning SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 46

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T06:18:05.253285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-20T06:16:47.650748Z digest=sha256:96416007dfd3ef26b98f7c57a30796c923c1a3af532733511c61236363151dda

Observation 69a5d072-97fc-4bbb-8a42-0a05eb72b807 · inbound

Learning Spatiotemporal Sensitivity in Video LLMs via Counterfactual Reinforcement Learning cites this paper.

Learning Spatiotemporal Sensitivity in Video LLMs via Counterfactual Reinforcement Learning SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:46:14.794643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-22T07:45:56.473188Z digest=sha256:374e50f1e40d9da65cebe90bdaba1ea67adca4082c316062aeeb948397592497

Observation fd0d599e-858d-4314-b105-19810a7dee72 · inbound

AnE: Pushing the Reasoning Frontier of Multimodal LLMs via Anchor Evolution cites this paper.

AnE: Pushing the Reasoning Frontier of Multimodal LLMs via Anchor Evolution SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-06-30T00:14:04.632391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-29T22:56:39.504430Z digest=sha256:fc3b548e5b78808580e6a07d8493fbea4c091ae53610d7dea7a5f7a97746e996

Observation 76f19e51-4722-4db9-b7e5-7bdd6e629898 · inbound

Smart Picks in the Dark: Towards Efficient RLVR for Reasoning via Tracing Metacognitive Pivots cites this paper.

Smart Picks in the Dark: Towards Efficient RLVR for Reasoning via Tracing Metacognitive Pivots SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-07-02T07:26:46.252644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-28T06:55:09.927034Z digest=sha256:e621584d9cba97820ebb3e3812d36048bbf6619cc477e5bd5e3ddaab3643eab7

Observation 054e8c7e-f438-48cc-b2d1-32f7f80c9388 · inbound

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning cites this paper.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:47:23.075925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:41ab5c8eab1415d3c7eaf00dc2f00c37eab0f8472a04f4b09ee0c07ea1e9b20e

Observation d864721f-d0b0-4ffc-b656-4f7331d7ed9a · inbound

VeriEvol: Scaling Multimodal Mathematical Reasoning via Verifiable Evol-Instruct cites this paper.

VeriEvol: Scaling Multimodal Mathematical Reasoning via Verifiable Evol-Instruct SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-07-04T10:39:46.112526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-26T08:34:27.719022Z digest=sha256:33ea8babee5a9c9bf30ac7d720f57ce4b8a8f58e0131c3a0d54046a3eeb01e08

Observation 38cde0c3-4f48-4539-a80e-1b65579130f1 · inbound

Latent Visual States for Efficient Multimodal Reasoning cites this paper.

Latent Visual States for Efficient Multimodal Reasoning SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-07-04T16:29:57.249633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-26T00:38:11.619574Z digest=sha256:36e03e9e2ebd20e3ca10377a9261b1b15cc03dff5cc42c03410b016cf3379288

Observation 6a07455b-2c78-496e-a08d-1d641ba59446 · inbound

No Place to Hide: Benchmarking Video Hallucination with Background-Controlled Pairs cites this paper.

No Place to Hide: Benchmarking Video Hallucination with Background-Controlled Pairs SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 70

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:25:41.282559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-07-01T05:35:08.200219Z digest=sha256:39ddc91ee0036171772e16ce56c2dfd97b7bbb877dec8733fa8f430630e076a3

Observation bf766f73-9c7e-47ba-8f1b-5d285e2986d9 · inbound

Trace: A Taxonomy-Guided Environment for Multidomain Visual Reasoning cites this paper.

Trace: A Taxonomy-Guided Environment for Multidomain Visual Reasoning SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-01T11:47:27.535628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:47:27.535628Z digest=sha256:a0b297ca7ef12285530fdd537c5acef27cf270a3a99cf286c705237a4d2d0085

Observation b90d450a-95dd-4a70-accb-41ccb8724fa9 · inbound

Aligning Large Vision-Language Models at Test Time: A Trajectory-Guided Structured Sampling Approach cites this paper.

Aligning Large Vision-Language Models at Test Time: A Trajectory-Guided Structured Sampling Approach SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T23:40:44.227213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:40:44.227213Z digest=sha256:0061247910c78c43d360e80c3ff989294226873920a5e608cad43f61a1896ce1

Observation 21dea8d1-a349-4d73-afc9-6085a711eb7e · inbound

Multi-Branch Policy Optimization for Multimodal Large Language Models cites this paper.

Multi-Branch Policy Optimization for Multimodal Large Language Models SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T00:36:01.053166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:36:01.053166Z digest=sha256:9767f97c5180d08d4313637b0b4d168fd5f0c01fbb346549353e5bb90152a954