Pith. sign in

Paper Citation Record · LEDGER

Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

As of 22 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 68 inbound Pith citation observations for arXiv:2412.18319.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.18319 v2

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T04:53:11.881661Z

measured 98 of 98 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 68 of 68 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T05:39:41.133050Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

30 of 30 outbound references displayed

  • verified exact0
  • verified fuzzy2
  • unresolved28
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 81eb2a25-5998-4329-b68c-5ea524fe735c · outbound

This paper cites Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models.

Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T04:53:11.765441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:53:11.765441Z digest=sha256:71d16c519b4e0e46f14733fe6c11e1eb3fa80daf16497ed5ff882c2c67f1ec93

Observation 1a9b162e-0f4c-4f29-8cd9-2ab06e06629c · outbound

This paper cites GPT-4o System Card.

Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search GPT-4o System Card

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T04:53:11.785803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:53:11.785803Z digest=sha256:e18700438b73fdb61c8e66ce73b5ea538eceb24febee7374ffe87391ba8edc63

Observation d285ebb5-6d30-485e-8903-df7e028a2b79 · outbound

This paper cites FigureQA: An Annotated Figure Dataset for Visual Reasoning.

Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search FigureQA: An Annotated Figure Dataset for Visual Reasoning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T04:53:11.790199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:53:11.790199Z digest=sha256:ee2d6a2059826727734fca3585c890093b6616fd0c343679faabd9887514e202

Observation 46180e2c-6fcb-42e5-8063-9c7ba423f5a7 · outbound

This paper cites Li, Z., Wang, X., Stengel-Eskin, E., Kortylewski, A., Ma, W., Van Durme, B., and Yuille, A.

Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search Li, Z., Wang, X., Stengel-Eskin, E., Kortylewski, A., Ma, W., Van Durme, B., and Yuille, A

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:53:12.285186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T04:53:11.803394Z digest=sha256:a3666bf5a7019624a99180f2e8a8c90b45b9e7199505ed5454af4605704f67f3

Observation 5274f275-fe13-4bfc-ac1f-5685911b1364 · outbound

This paper cites CLEVR-Math: A Dataset for Compositional Language, Visual and Mathematical Reasoning.

Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search CLEVR-Math: A Dataset for Compositional Language, Visual and Mathematical Reasoning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T04:53:11.807652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:53:11.807652Z digest=sha256:f04d06ef5cfaf181e738e3b928241d8fe081a012bdcdc24ec6f119f5fe93e5b1

Observation e54dd486-895b-4c2f-97b7-bec701bf0d1a · outbound

This paper cites DeepSeek-VL: Towards Real-World Vision-Language Understanding.

Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search DeepSeek-VL: Towards Real-World Vision-Language Understanding

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T04:53:11.811964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:53:11.811964Z digest=sha256:8d9dd1d138629bba6c481769e775f88b3551ec027e517222663461fbcffeb94d

Observation 8d457759-ff86-4225-81b8-b30da5703dce · outbound

This paper cites Improve Mathematical Reasoning in Language Models by Automated Process Supervision.

Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search Improve Mathematical Reasoning in Language Models by Automated Process Supervision

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T04:53:11.816269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:53:11.816269Z digest=sha256:6d549cf1bf27172a8d472ae76c4e2cf5bc70d0af132863c654f2d3df678b70dd

Observation 666598a6-b623-4e77-8fa6-9dd3c559b2df · outbound

This paper cites ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning.

Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T04:53:11.820533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:53:11.820533Z digest=sha256:edd7c58eea462f77a1997cea4a4e0970cc9dc613017564c06ceb04c9c400985e

Observation 0f9efeb2-6bdd-4959-b142-e174385ebb26 · outbound

This paper cites Solving geometry problems: Combining text and diagram interpretation.

Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search Solving geometry problems: Combining text and diagram interpretation

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:53:12.271352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T04:53:11.824617Z digest=sha256:973230ec3033acf78a6350c8df505d3a1570b96b6ace1356aaebf3fa3d2ad8fe

Observation a8a23b19-74db-4885-9b32-ec481d98c792 · outbound

This paper cites PhyPlan: Compositional and Adaptive Physical Task Reasoning with Physics-Informed Skill Networks for Robot Manipulators.

Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search PhyPlan: Compositional and Adaptive Physical Task Reasoning with Physics-Informed Skill Networks for Robot Manipulators

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T04:53:11.837943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:53:11.837943Z digest=sha256:e05ede41cb3023caff96bb6789af1f72e4616a2ecab96d4bca36698e11897b82

Observation 20c68557-2c2d-4123-ba51-e3d50c2e887e · outbound

This paper cites Measuring Multimodal Mathematical Reasoning with MATH-Vision Dataset.

Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search Measuring Multimodal Mathematical Reasoning with MATH-Vision Dataset

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T04:53:11.842388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:53:11.842388Z digest=sha256:d163164e58c9106cd6e99dffee5391c1fc1397823763909a9f4faa6a7f3c5a22

Observation edc18909-d823-484d-bd94-9334a14f1dd3 · outbound

This paper cites DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding.

Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T04:53:11.846702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:53:11.846702Z digest=sha256:d4293072c93b2408e4d5a0461457ee763d5a6d9190f190f41e6a9c72ea8b7e11

Observation 88b8f79a-7ea5-42ed-aba9-38b42bd20d11 · outbound

This paper cites Monte Carlo Tree Search Boosts Reasoning via Iterative Preference Learning.

Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search Monte Carlo Tree Search Boosts Reasoning via Iterative Preference Learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T04:53:11.851150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:53:11.851150Z digest=sha256:5dcd59210999ace50e9a5429cab53fb53b505f17f7f41e733b3fa46607ca6088

Observation f336bb95-8206-4a23-948e-20d069caaece · outbound

This paper cites LLaVA-CoT: Let Vision Language Models Reason Step-by-Step.

Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search LLaVA-CoT: Let Vision Language Models Reason Step-by-Step

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T04:53:11.855660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:53:11.855660Z digest=sha256:34bf2453c78424aee77e4f67892e37f3f8b1bb44182157f290bc2fc649f4f31f

Observation 1db25a7c-6ab2-41a6-97c8-a16b66e46da6 · outbound

This paper cites An Integrated Framework Integrating Monte Carlo Tree Search and Supervised Learning for Train Timetabling Problem.

Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search An Integrated Framework Integrating Monte Carlo Tree Search and Supervised Learning for Train Timetabling Problem

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T04:53:11.859993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:53:11.859993Z digest=sha256:bd3d5170c7cc02796259f7bcfc9d1f53732c1b30ab05bdf77665bf9b4b557fb2

Observation 2420f65f-0a02-4be7-b9dc-0ea2495fe35a · outbound

This paper cites Dense Connector for MLLMs.

Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search Dense Connector for MLLMs

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T04:53:11.864340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:53:11.864340Z digest=sha256:8e0cea67aa5ddbf2bc2a1c7c827b389da920253b25b8675d4465a497265badd7

Observation 140be8b1-b78b-4d2f-b878-247733deb9f2 · outbound

This paper cites MMMU-Pro: A More Robust Multi-discipline Multimodal Understanding Benchmark.

Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search MMMU-Pro: A More Robust Multi-discipline Multimodal Understanding Benchmark

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T04:53:11.868800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:53:11.868800Z digest=sha256:b1756bf80f3a95328cc6480c20b3b696344c3d5af265699e3727076ef51cc7c6

Observation 6f2e1174-af77-4c1d-9cf2-9f6d4ef505a5 · outbound

This paper cites ReST-MCTS*: LLM Self-Training via Process Reward Guided Tree Search.

Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search ReST-MCTS*: LLM Self-Training via Process Reward Guided Tree Search

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T04:53:11.873185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:53:11.873185Z digest=sha256:eb198d0215a35cd9122e472980daf46b2d67187f5a63e46e0fe4081ee5f64705

Observation 9860984a-0cc5-4a06-84b4-1f1e3741a20e · outbound

This paper cites MultiHiertt: Numerical Reasoning over Multi Hierarchical Tabular and Textual Data.

Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search MultiHiertt: Numerical Reasoning over Multi Hierarchical Tabular and Textual Data

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T04:53:11.877393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:53:11.877393Z digest=sha256:cd8ff14d146fc981af2dfd446a507f2578b75c85d016591320de63ac041645c6

Observation d2108ee9-3515-40f0-8f6b-26083aa9dba7 · outbound

This paper cites an unresolved cited work.

Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-11T04:53:12.256443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T04:53:11.881661Z digest=sha256:adb7cdd0c232e29e27a60eec5f8dd2a237af34011e99b34cc2c151e2532a2a4d

Observation 91801d1a-5dd6-4d6c-a5ee-027a0dcff59e · outbound

This paper cites GeoQA: A Geometric Question Answering Benchmark Towards Multimodal Numerical Reasoning.

Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search GeoQA: A Geometric Question Answering Benchmark Towards Multimodal Numerical Reasoning

Reference 1998

Resolution
unresolved
no resolver link, observed 2026-08-11T04:53:11.750387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:53:11.750387Z digest=sha256:637ef06b0a604af4be421bf1ebdfc77a7da02ed728a6496a27cb469b317e81db

Observation 84b70e61-b89e-403f-ac4e-bd605e7ec43c · outbound

This paper cites Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs.

Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs

Reference 2011

Resolution
unresolved
no resolver link, observed 2026-08-11T04:53:11.833403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:53:11.833403Z digest=sha256:82b69c195aa2c234871e8d999223a56a1a487f69b698cc27f8eea7fe21439cf8

Observation dab83655-1716-4dc4-9ce1-41746ae4996e · outbound

This paper cites Math-LLaVA: Bootstrapping Mathematical Reasoning for Multimodal Large Language Models.

Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search Math-LLaVA: Bootstrapping Mathematical Reasoning for Multimodal Large Language Models

Reference 2015

Resolution
unresolved
no resolver link, observed 2026-08-11T04:53:11.828780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:53:11.828780Z digest=sha256:cba2cd0053a941f1b3356c96cfee435cc7bcd9f71118fe3c58a26ef30f802ed2

Observation 232e2a99-5574-490f-9ea4-cba47188c600 · outbound

This paper cites G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model.

Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-11T04:53:11.775687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:53:11.775687Z digest=sha256:b96ce515b9439e4e85b4e37e2f8dd25b6f4085859a269d373d37519d53bce86c

Observation 35174df6-e4c8-412c-a46c-9e505155b2f1 · outbound

This paper cites GeomVerse: A Systematic Evaluation of Large Models for Geometric Reasoning.

Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search GeomVerse: A Systematic Evaluation of Large Models for Geometric Reasoning

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-11T04:53:11.794660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:53:11.794660Z digest=sha256:4fe84fdc6e4667572da219460808232cdf21e057213455fdda77566981ebc2a4

Observation 37a6830a-9e8e-43ea-953f-aec62da1ffe3 · outbound

This paper cites A Survey on Evaluation of Multimodal Large Language Models.

Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search A Survey on Evaluation of Multimodal Large Language Models

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-11T04:53:11.781191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:53:11.781191Z digest=sha256:0d3bca6330796e039daeba7d4080aaee0325eab930a58d2c5b80e93df8592af7

Observation 8b741d1b-1012-4f4f-b367-94477b568d92 · outbound

This paper cites UniGeo: Unifying Geometry Logical Reasoning via Reformulating Mathematical Expression.

Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search UniGeo: Unifying Geometry Logical Reasoning via Reformulating Mathematical Expression

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-11T04:53:11.755682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:53:11.755682Z digest=sha256:08cd1a5d41e57398781aeacb29803b88f4ad50125cff0bc757fecdd864ec1e67

Observation e1e86f5f-52a9-4d06-ba6d-8adc1f5c89ff · outbound

This paper cites How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites.

Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-11T04:53:11.760494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:53:11.760494Z digest=sha256:be5553079656998256efbd64257c06e1bc50348f4d05521688bbc0ea19d32a0c

Observation eb4d3653-19db-4139-b6e9-6f77334002e5 · outbound

This paper cites A diagram is worth a dozen images.

Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search A diagram is worth a dozen images

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-11T04:53:11.799132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:53:11.799132Z digest=sha256:be82523b9423b82b829940d6d1a59180dbf25eb900d526d106f2d38f58824bca

Observation d545414e-7000-4cbd-b9c8-548049589512 · outbound

This paper cites The Llama 3 Herd of Models.

Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search The Llama 3 Herd of Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-11T04:53:11.770720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:53:11.770720Z digest=sha256:53cb17eb7fbd12d23c554950e7be1ebb761acfe1a9d927a87f2548e300cea536

Pith citing papers

Observation b10fc253-f090-4bb4-b52a-9276718d1b17 · inbound

ZoomEye: Enhancing Multimodal LLMs with Human-Like Zooming Capabilities through Tree-Based Image Exploration cites this paper.

ZoomEye: Enhancing Multimodal LLMs with Human-Like Zooming Capabilities through Tree-Based Image Exploration Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T13:41:49.762335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:41:49.762335Z digest=sha256:90c8ebe6bbeb46d5eeb60997f836ab2f263a070d479dc6e02948ddd1a41271c7

Observation ad423973-0710-4592-aede-ac1814388c2f · inbound

FCMR: Robust Evaluation of Financial Cross-Modal Multi-Hop Reasoning cites this paper.

FCMR: Robust Evaluation of Financial Cross-Modal Multi-Hop Reasoning Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T14:00:26.830532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:00:26.830532Z digest=sha256:25e003ded436373ec0fb8154dca371939905e8882296d584937034a6ceaf2284

Observation b580e726-7449-4b26-bbf6-24935a0bd077 · inbound

Search-o1: Agentic Search-Enhanced Large Reasoning Models cites this paper.

Search-o1: Agentic Search-Enhanced Large Reasoning Models Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:36:27.624658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-13T17:36:27.515468Z digest=sha256:c0e80a43347b0e033b9e900435664739b884463918ef37334b21b3cb37991d7d

Observation 0c1bcde0-415f-47b3-a589-c2117be0ae93 · inbound

A Survey on LLM Test-Time Compute via Search: Tasks, LLM Profiling, Search Algorithms, and Relevant Frameworks cites this paper.

A Survey on LLM Test-Time Compute via Search: Tasks, LLM Profiling, Search Algorithms, and Relevant Frameworks Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-10T19:27:23.501890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T19:27:23.501890Z digest=sha256:7586ba72cd5a0c39829e84d1722a9adf29ec0b8269a30c955f48fec1a5b86d89

Observation 23c6a007-5724-4046-9993-79b620aaf323 · inbound

O1-Pruner: Length-Harmonizing Fine-Tuning for O1-Like Reasoning Pruning cites this paper.

O1-Pruner: Length-Harmonizing Fine-Tuning for O1-Like Reasoning Pruning Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T17:09:06.849411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:09:06.849411Z digest=sha256:667629a3adcd85ad4754724e3848291f0f9781455de352d48120959d2b806e22

Observation 2445863c-5c1f-4b83-836c-443a5ad29832 · inbound

Position: Multimodal Large Language Models Can Significantly Advance Scientific Reasoning cites this paper.

Position: Multimodal Large Language Models Can Significantly Advance Scientific Reasoning Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 235

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:32:32.914733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-23T04:30:38.804702Z digest=sha256:7c8484eb91c023de6f2fc4ba9e5024641c559af3074d53ed928c1850c7893d1d

Observation 05e96a44-673c-498a-9e44-a5c222bd5e75 · inbound

From System 1 to System 2: A Survey of Reasoning Large Language Models cites this paper.

From System 1 to System 2: A Survey of Reasoning Large Language Models Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 234

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:36:24.386960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-13T01:36:23.845366Z digest=sha256:d4c87f37bb4c5cf00a2dc83d5cfa6c473440e1fadaa2ed98640be1420e5d4e93

Observation 0168748b-34d2-4f34-931e-69076eab310a · inbound

LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL cites this paper.

LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 85

Resolution
verified exact
arxiv_id, observed 2026-05-16T15:15:46.332540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-16T15:15:46.255296Z digest=sha256:fcf97bd4aa14a10b7885f7e59b3e5f39ce225d12d422d932e3d828bf19e8b142

Observation f0a57240-9dbe-4dc8-9a2c-892def0145cc · inbound

R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization cites this paper.

R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-16T00:19:20.521597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-16T00:19:20.462455Z digest=sha256:670357220cc6897a6b05dbbfe1bce79bdad88b500461e9d3b6a64dea90d0280f

Observation c6be422e-5438-4bd1-af12-e449029a4a2a · inbound

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey cites this paper.

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 94

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:18:53.762696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-15T17:18:52.996467Z digest=sha256:49c1f2bb3743a6b065fcd44c3072c459471273930244f4f397c21c0e79e4aa07

Observation 8b2e5172-eb26-4aab-bb35-1dfb5b7b0bde · inbound

R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization cites this paper.

R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-16T15:04:22.772654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-16T15:04:22.690503Z digest=sha256:89ddde9c5089eb1e03c45c2228152d61e1e2fd8730d77a5651e64150623905fe

Observation e845a83c-c550-4af2-8f1d-78b2529bbb72 · inbound

OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles cites this paper.

OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 81

Resolution
verified exact
arxiv_id, observed 2026-05-19T06:59:03.316807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-19T06:59:03.112252Z digest=sha256:c786bda59ec7898b34404c3f54a5d4cd01b5f3ca18f8242a61724f4d6313f6bc

Observation b12f219c-6aff-409e-9b29-c0416c8f240e · inbound

FLARE: Fully Integration of Vision-Language Representations for Deep Cross-Modal Understanding cites this paper.

FLARE: Fully Integration of Vision-Language Representations for Deep Cross-Modal Understanding Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 71

Resolution
verified exact
arxiv_id, observed 2026-05-22T19:52:01.898750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-22T19:49:00.961388Z digest=sha256:73c2368556326b2fa7c8dc139103f8ad06feef8173ac7ab9790e156a0ad88870

Observation 0502a755-aa09-4b7f-8a9a-8768539c88f2 · inbound

Weaving Context Across Images: Improving Vision-Language Models through Focus-Centric Visual Chains cites this paper.

Weaving Context Across Images: Improving Vision-Language Models through Focus-Centric Visual Chains Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-16T05:39:41.133050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:39:41.133050Z digest=sha256:368eb460ec9d938f693c1f41fe8cd1c9054f6fbeca064a2668b33b4b9983ce21

Observation 61adf9fe-006a-4855-b3de-58f0b2a44cfd · inbound

Reinforced MLLM: A Survey on RL-Based Reasoning in Multimodal Large Language Models cites this paper.

Reinforced MLLM: A Survey on RL-Based Reasoning in Multimodal Large Language Models Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 130

Resolution
unresolved
no resolver link, observed 2026-08-16T05:12:18.654182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:12:18.654182Z digest=sha256:9424dad29e2a91171cffe3846af285560ca020772fa9ab027d5d66755f437cb3

Observation ff77fb0c-e3ed-4bc3-b77b-f2eaaa223465 · inbound

A Survey of Slow Thinking-based Reasoning LLMs using Reinforced Learning and Inference-time Scaling Law cites this paper.

A Survey of Slow Thinking-based Reasoning LLMs using Reinforced Learning and Inference-time Scaling Law Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 142

Resolution
unresolved
no resolver link, observed 2026-08-16T00:48:19.665826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:48:19.665826Z digest=sha256:b13671cbb4ce409f640031fef9be368e6e159952f50d916f4d2c56a52477e469

Observation 5fb6049a-a965-430f-a92e-3813bd5b6a78 · inbound

UniGen: Enhanced Training & Test-Time Strategies for Unified Multimodal Understanding and Generation cites this paper.

UniGen: Enhanced Training & Test-Time Strategies for Unified Multimodal Understanding and Generation Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-07T15:34:55.877402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:34:55.877402Z digest=sha256:c37a3b4b4066c7605779c5b951959f93adfb9b3a03dae15874a6b5e7c9f9f96f

Observation 9d634656-5c72-4ccb-9890-89369531f8f8 · inbound

Training-Free Reasoning and Reflection in MLLMs cites this paper.

Training-Free Reasoning and Reflection in MLLMs Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T15:13:52.634235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:13:52.634235Z digest=sha256:d3a1296991d8cecf6807813e954f14678bca7bd9a8c2690bfa4db30b536654fb

Observation 90baaafb-bf9e-450e-bf43-6004bb7b83b6 · inbound

R1-ShareVL: Incentivizing Reasoning Capability of Multimodal Large Language Models via Share-GRPO cites this paper.

R1-ShareVL: Incentivizing Reasoning Capability of Multimodal Large Language Models via Share-GRPO Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T15:01:34.575094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:01:34.575094Z digest=sha256:4a0c62286e42e82672d7428d9b751cc31f734d353cc1aad09956d69db0fba719

Observation 9ecc574c-ada4-42b0-9d5d-6131135f8e9a · inbound

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models cites this paper.

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T14:31:14.568519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:31:14.568519Z digest=sha256:e7a04d7d8d632294546a45e251660a5e4243789d008163ab1bb890b1cf15fd15

Observation 47598c7c-5742-4d49-b975-a4a009bdb9c0 · inbound

Unveiling the Compositional Ability Gap in Vision-Language Reasoning Model cites this paper.

Unveiling the Compositional Ability Gap in Vision-Language Reasoning Model Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:19:04.371188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:19:04.371188Z digest=sha256:da88afbc7e5b308381267bef0d1b04e3fefb890c08d1533ec9b2f80f729516c4

Observation 73db5779-c516-4837-9876-314466996ab7 · inbound

MME-Reasoning: A Comprehensive Benchmark for Logical Reasoning in MLLMs cites this paper.

MME-Reasoning: A Comprehensive Benchmark for Logical Reasoning in MLLMs Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T13:36:01.466524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:36:01.466524Z digest=sha256:771ffb3b04b9c216b2b3af3537058759af1151d0343aa3dd106ad920159c5349

Observation daa75c69-f23d-4a59-a3ab-8ab7d728b3e0 · inbound

More Thinking, Less Seeing? Assessing Amplified Hallucination in Multimodal Reasoning Models cites this paper.

More Thinking, Less Seeing? Assessing Amplified Hallucination in Multimodal Reasoning Models Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:35.946043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:35.946043Z digest=sha256:2e8495f02e623b7f0219ec4c96adf3469cd74e3440095b127d8521a1ce10fc95

Observation c410146c-2b0d-4cf1-9aec-ba6358eaa55b · inbound

Advancing Multimodal Reasoning via Reinforcement Learning with Cold Start cites this paper.

Advancing Multimodal Reasoning via Reinforcement Learning with Cold Start Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T13:12:58.218444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:12:58.218444Z digest=sha256:121b87921e3436b5be08815ed02da057c53b60fa479223f381ba094799123ac0

Observation 3e0959b3-67cf-4dbc-b435-144b138447c1 · inbound

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models cites this paper.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:22.779575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:22.779575Z digest=sha256:5db4d3598298bd6add95f31845ccbf8d7ebfa86594e90bc1a7231f21aab817d3

Observation 78b947a2-9833-44fc-9df2-237ae3597012 · inbound

Grounded Reinforcement Learning for Visual Reasoning cites this paper.

Grounded Reinforcement Learning for Visual Reasoning Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 80

Resolution
verified exact
arxiv_id, observed 2026-05-22T01:05:52.100628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-22T01:05:18.801388Z digest=sha256:aa16ee2f7e0ae726d2f074c2586fd974129bfe9252a83944e26c0cd13df5e9bd

Observation 0b48fbce-4c79-415a-aeac-4e3e872724ae · inbound

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM cites this paper.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:12.258432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:12.258432Z digest=sha256:87d21ab91e049a8ffff67abea1b6fab2e6b7d8bb0ebe72b16549d3b43bc238bb

Observation 126c7a05-0e26-4e47-af71-db622ff377b7 · inbound

GThinker: Towards General Multimodal Reasoning via Cue-Guided Rethinking cites this paper.

GThinker: Towards General Multimodal Reasoning via Cue-Guided Rethinking Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T11:55:14.280040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:55:14.280040Z digest=sha256:505cd914295afed54f3911d6cf1a5cabef8fcd4b7808658050a187bc16620509

Observation 11382a79-e212-483d-acee-0e4633d8e18a · inbound

VideoCap-R1: Enhancing MLLMs for Video Captioning via Structured Thinking cites this paper.

VideoCap-R1: Enhancing MLLMs for Video Captioning via Structured Thinking Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T11:40:55.873734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:40:55.873734Z digest=sha256:ad0bbac765ca2b7040f1dd55e555778b68b22059dce906a87d8f78574ca63495

Observation e11d0494-ee93-48e0-b805-ec5b346c0859 · inbound

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis cites this paper.

SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T11:35:47.633710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:35:47.633710Z digest=sha256:5a3acf2ff376334eb9c8064277c6cb2766a2dbe5080b880f82b3852412bcd2a3

Observation 45f4b524-8bd6-478b-9f8e-112eb596ffea · inbound

MINT-CoT: Enabling Interleaved Visual Tokens in Mathematical Chain-of-Thought Reasoning cites this paper.

MINT-CoT: Enabling Interleaved Visual Tokens in Mathematical Chain-of-Thought Reasoning Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:53.914971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:53.914971Z digest=sha256:086ad07024cd220333ffcb6f5cac60f77533808c2b43b21d33e107baa7b478c9

Observation eb26364b-9673-4dc3-92d1-342ae56409e4 · inbound

SIGMA: Refining Large Language Model Reasoning via Sibling-Guided Monte Carlo Augmentation cites this paper.

SIGMA: Refining Large Language Model Reasoning via Sibling-Guided Monte Carlo Augmentation Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T06:01:49.661304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:01:49.661304Z digest=sha256:4fe616e763d18e668f357e7ffe78ca56c4bd75c0d9fb7c8762a5e26fde544a81

Observation 087b4ef5-85d7-4553-bb5f-ca2c6e4f167e · inbound

Socratic-MCTS: Test-Time Visual Reasoning by Asking the Right Questions cites this paper.

Socratic-MCTS: Test-Time Visual Reasoning by Asking the Right Questions Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:20.446667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:01:20.446667Z digest=sha256:b2d2e0aa0ef22f800e63faa78fc9b23afba4c8aae2b6c2440c63f54e810f3d03

Observation f848019f-2edc-47ec-a455-f1c34aded2ed · inbound

VFaith: Do Large Multimodal Models Really Reason on Seen Images Rather than Previous Memories? cites this paper.

VFaith: Do Large Multimodal Models Really Reason on Seen Images Rather than Previous Memories? Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T04:07:34.668498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:07:34.668498Z digest=sha256:c967060036b4e62eb684818d12e1d6897634aa154c38688635e97ec6b5fb93ff

Observation 7dcb0f31-3bf2-4cea-bf6c-bd8fbc64ad49 · inbound

PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning cites this paper.

PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T00:13:38.425045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:13:38.425045Z digest=sha256:f15f6955006c15f28af92de327652f1445798f4cd56a202558a450284eeda616

Observation ee98e613-7690-4797-9504-94d2f83e0d75 · inbound

APO: Enhancing Reasoning Ability of MLLMs via Asymmetric Policy Optimization cites this paper.

APO: Enhancing Reasoning Ability of MLLMs via Asymmetric Policy Optimization Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T22:29:28.485026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:29:28.485026Z digest=sha256:a4126bd76020adab9a158ee4adaf392d31a82bdb2a9e34585e95b6426f8231bf

Observation fec71ebb-790b-405a-aa09-33befdc3f34a · inbound

MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI cites this paper.

MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T21:41:40.304579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:41:40.304579Z digest=sha256:16c7c22261722851e591d7666fab5a13f74fdf4e1d2d7d1810921731b8f8a527

Observation 4dd3c650-550b-4868-80f4-384c5615805b · inbound

Multimodal Mathematical Reasoning with Diverse Solving Perspective cites this paper.

Multimodal Mathematical Reasoning with Diverse Solving Perspective Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:24.323882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:24.323882Z digest=sha256:de722447de1658c038d762617c4d0c8234a3c4c89928c905df37d3ee3f56f616

Observation 2d984caf-423b-429f-a797-ea9f406ecd3e · inbound

Ascending the Infinite Ladder: Benchmarking Spatial Deformation Reasoning in Vision-Language Models cites this paper.

Ascending the Infinite Ladder: Benchmarking Spatial Deformation Reasoning in Vision-Language Models Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-06T21:22:27.289983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:22:27.289983Z digest=sha256:cc170dcfad8a55253ffbb995c77754c855bf1c8b74ee7bdff3895320593caec4

Observation fb6c8088-67a7-449e-a9d4-7d72bc5773c8 · inbound

EduFlow: Advancing MLLMs' Problem-Solving Proficiency through Multi-Stage, Multi-Perspective Critique cites this paper.

EduFlow: Advancing MLLMs' Problem-Solving Proficiency through Multi-Stage, Multi-Perspective Critique Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T18:04:55.194413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:04:55.194413Z digest=sha256:4f0f2921ca1c67e7284286e27c654731e9c8e1a57815ec60c5bac3a7f8c375c0

Observation 4a463d6b-3352-4a14-aa1c-214b488ba78f · inbound

VisionThink: Smart and Efficient Vision Language Model via Reinforcement Learning cites this paper.

VisionThink: Smart and Efficient Vision Language Model via Reinforcement Learning Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-06T16:34:00.236317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:34:00.236317Z digest=sha256:2e081c07fda2069115926aecb427485698555a8d420c163dc8a3d2e80b6de468

Observation 9e9f9bd2-7984-4338-86fd-750d3c07a831 · inbound

VL-Cogito: Progressive Curriculum Reinforcement Learning for Advanced Multimodal Reasoning cites this paper.

VL-Cogito: Progressive Curriculum Reinforcement Learning for Advanced Multimodal Reasoning Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-06T11:35:17.014678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:35:17.014678Z digest=sha256:8580fefc61bc9240d96047229d6b1133ee791003c2cf656e96ba295c33b5d9bb

Observation cca519d0-326f-488d-8942-db1710d26895 · inbound

The Emotional Baby Is Truly Deadly: Does your Multimodal Large Reasoning Model Have Emotional Flattery towards Humans? cites this paper.

The Emotional Baby Is Truly Deadly: Does your Multimodal Large Reasoning Model Have Emotional Flattery towards Humans? Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T00:59:43.431479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:59:43.431479Z digest=sha256:4b04207fc1f4f49776930327c95d0c6d79775291192d2e3432fd517974de8286

Observation 22b6c7f7-abb2-4156-85da-64c777a29856 · inbound

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models cites this paper.

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-05T10:39:02.195906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:39:02.195906Z digest=sha256:9bdcc995196e2389903def4db9cd3a7403a29c2dfe5f2727f9be896788dadd93

Observation 4d257517-7fcc-4fae-ba0e-77b5f9c88cd8 · inbound

Measuring Epistemic Humility in Multimodal Large Language Models cites this paper.

Measuring Epistemic Humility in Multimodal Large Language Models Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T18:49:39.220967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:49:39.220967Z digest=sha256:ea70e860f6e91f9fd9a1edac4979018023b7da7ff55b16e788a5cb683314604c

Observation 1901e20f-0587-4281-904a-fc08c4dbeab0 · inbound

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning cites this paper.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:36:25.218707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:e3f517d01e4ce7f8ac52f095b3db3e448ca11fe1c3bad37c9ea2fc45aba18b4c

Observation dcb294ac-1729-4a20-957a-1b69187f1710 · inbound

AutoRubric: Rubric-Based Generative Rewards for Faithful Multimodal Reasoning cites this paper.

AutoRubric: Rubric-Based Generative Rewards for Faithful Multimodal Reasoning Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-18T06:25:58.662923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-18T06:22:51.194258Z digest=sha256:7ff82a9dad805f4554516b27d0abb45646ce114f04f16fd38119b28d039cdfcb

Observation eeae1f07-b5ad-4c89-b6b7-b8cbdaed8421 · inbound

From Hindsight to Foresight: Self-Encouraged Hindsight Distillation for Knowledge-based Visual Question Answering cites this paper.

From Hindsight to Foresight: Self-Encouraged Hindsight Distillation for Knowledge-based Visual Question Answering Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-03T22:19:23.737176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:19:23.737176Z digest=sha256:2b6696a2758ddf4da097a82249035632ef1a3da879d6ac84865bf3d942ee57b9

Observation a6e8d204-9495-40ae-8eeb-3eef15c4d873 · inbound

See, Think, Learn: A Self-Taught Multimodal Reasoner cites this paper.

See, Think, Learn: A Self-Taught Multimodal Reasoner Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-03T19:04:28.472931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T19:04:28.472931Z digest=sha256:c1b9dd38d54a282ebde8db980f98854c32fee04923d1661f0e12ceca4b068b0d

Observation ed8b0df9-e4fc-4c3e-b2b3-a98afd06df6f · inbound

Boosting RL-Based Visual Reasoning with Selective Adversarial Entropy Intervention cites this paper.

Boosting RL-Based Visual Reasoning with Selective Adversarial Entropy Intervention Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-03T17:15:17.705468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T17:15:17.705468Z digest=sha256:be3c66482bc75b48daadff366789ab80720d63664253869f5f4ffe2c0b1c4e11

Observation 06be3bcc-8d74-4ced-b42d-4ba96826ed5f · inbound

OmniDrive-R1: Reinforcement-driven Interleaved Multi-modal Chain-of-Thought for Trustworthy Vision-Language Autonomous Driving cites this paper.

OmniDrive-R1: Reinforcement-driven Interleaved Multi-modal Chain-of-Thought for Trustworthy Vision-Language Autonomous Driving Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-16T22:38:37.768730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-16T22:34:00.895252Z digest=sha256:8bb351fab64ea91e6f23dfad6313c633245d4f367c8244b0c80d2d69c9612157

Observation 85d50452-a5a5-40d9-a4a2-c57091229151 · inbound

Mimic Human Cognition, Master Multi-Image Reasoning: A Meta-Action Framework for Enhanced Visual Understanding cites this paper.

Mimic Human Cognition, Master Multi-Image Reasoning: A Meta-Action Framework for Enhanced Visual Understanding Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-03T11:13:16.953162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T11:13:16.953162Z digest=sha256:bf00249eea1785b1248d184155935b09128cbbf53465f8efd4e9bc517b2d121e

Observation 2742394c-f1b6-4286-9112-e5884e5e2ce1 · inbound

CodePercept: Code-Grounded Visual STEM Perception for MLLMs cites this paper.

CodePercept: Code-Grounded Visual STEM Perception for MLLMs Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 57

Resolution
unresolved
no resolver link, observed 2026-07-14T23:22:13.847876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T23:22:13.847876Z digest=sha256:051b8f13851b6e9a0b99afcb3677ea33fb0e63f6218b07bc615e3a747a514a8b

Observation 872e0a96-19b6-4882-b7c2-c9591849c3dc · inbound

Understanding the Role of Hallucination in Reinforcement Post-Training of Multimodal Reasoning Models cites this paper.

Understanding the Role of Hallucination in Reinforcement Post-Training of Multimodal Reasoning Models Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:53:16.233964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-13T20:48:52.130130Z digest=sha256:2073c0f97e31a35350a4328aabe0bacf501b85fd6fe3d3c7871c1d76714ce617

Observation 2fe2d175-6bc3-4591-afde-d9e5046d62f7 · inbound

EgoMind: Activating Spatial Cognition through Linguistic Reasoning in MLLMs cites this paper.

EgoMind: Activating Spatial Cognition through Linguistic Reasoning in MLLMs Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:43:22.880169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-13T22:41:09.840792Z digest=sha256:e9a8442915c154c2708da51f3c85021304806bb0f860f55a43bb6c43ba4d25ce

Observation ef9652c7-9ee0-4bf7-90b9-b8bcf885574c · inbound

EgoMind: Activating Spatial Cognition through Linguistic Reasoning in MLLMs cites this paper.

EgoMind: Activating Spatial Cognition through Linguistic Reasoning in MLLMs Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 52

Resolution
unresolved
no resolver link, observed 2026-07-13T14:39:55.177552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T14:39:55.177552Z digest=sha256:101f6e6faf6a760c9147b9ed5143b79206c3c482ff513eeae5167d39b3ef036e

Observation 29ab38f9-d3ae-4998-8379-e10253f297d4 · inbound

Saliency-R1: Enforcing Interpretable and Faithful Vision-language Reasoning via Saliency-map Alignment Reward cites this paper.

Saliency-R1: Enforcing Interpretable and Faithful Vision-language Reasoning via Saliency-map Alignment Reward Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 82

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:20:47.814497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T19:59:19.379119Z digest=sha256:b3a5b876db32e89b639f9c9be5b94f38ec7e6a1c15e814d573250cff5d5c6acf

Observation b1f3277b-1dbf-4699-b405-70d78dc7fd39 · inbound

See Further, Think Deeper: Advancing VLM's Reasoning Ability with Low-level Visual Cues and Reflection cites this paper.

See Further, Think Deeper: Advancing VLM's Reasoning Ability with Low-level Visual Cues and Reflection Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:36:17.429128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-08T04:46:16.497585Z digest=sha256:ac834a2efd37aedce97915a84f5299cf85d8fb68d44f2b002137fc640f13f3bc

Observation fe21a0c5-c498-4073-81d0-d37b724abdfa · inbound

Improving Medical VQA through Trajectory-Aware Process Supervision cites this paper.

Improving Medical VQA through Trajectory-Aware Process Supervision Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T07:06:06.106083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T17:18:17.846140Z digest=sha256:1bcb220e2583e72aa10b66ab72b4e26f9e18ac8eb3a0b7710588799b2307545e

Observation 4d9b3300-e2e6-4d78-811f-73a3c07cdc0b · inbound

Reflection Anchors for Propagation-Aware Visual Retention in Long-Chain Multimodal Reasoning cites this paper.

Reflection Anchors for Propagation-Aware Visual Retention in Long-Chain Multimodal Reasoning Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:41:26.743697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-12T02:24:12.349405Z digest=sha256:1ca6362104fa0bc714ff560385532eb381b2c13cb4fbbad3f3dd909b543984d5

Observation eeb130d9-0f29-43d0-9bfd-e1f0120f3efb · inbound

AnE: Pushing the Reasoning Frontier of Multimodal LLMs via Anchor Evolution cites this paper.

AnE: Pushing the Reasoning Frontier of Multimodal LLMs via Anchor Evolution Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-06-30T00:14:04.640763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-29T22:56:39.504430Z digest=sha256:cbcfb064e33037b0629de81f61bdf02f7303a1e4fc94d77af73efab50453dedf

Observation ab338af3-4a90-4700-88a7-e8dd47e69556 · inbound

Test-Time Scaling in Multimodal Foundation Models: A Comprehensive Survey of Generation and Reasoning cites this paper.

Test-Time Scaling in Multimodal Foundation Models: A Comprehensive Survey of Generation and Reasoning Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 105

Resolution
verified exact
arxiv_id, observed 2026-07-02T21:37:25.300780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-27T19:36:57.231932Z digest=sha256:d2b33283ba83bd507f8d6272de81c2e37a036145475b3a220c37ec53cd2acd83

Observation c315e3cf-a431-4d02-aa05-953eb9477c07 · inbound

TVI-CoT: Text-Visual Interleaved Chain-of-Thought Reasoning for Multimodal Understanding cites this paper.

TVI-CoT: Text-Visual Interleaved Chain-of-Thought Reasoning for Multimodal Understanding Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-02T22:27:25.919050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-27T18:54:34.353940Z digest=sha256:b1ea31a0efdd61a302c9c07f88397891aba1376b80e00f320d4231f0dac231e9

Observation d5c49abe-4f0d-4486-a1c8-496ccb92d828 · inbound

MathVis-Fine: Aligning Visual Supervision with Necessity via Progressive Dependency-Guided Training for Multimodal Mathematical Reasoning cites this paper.

MathVis-Fine: Aligning Visual Supervision with Necessity via Progressive Dependency-Guided Training for Multimodal Mathematical Reasoning Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T20:18:57.680997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-27T01:23:40.564561Z digest=sha256:ab3e7ba0968c85a203d81e63b5b59464eb0577de5c798ed5d0ef9a13f071fb04

Observation 73568d86-d25c-4e8c-b17b-b33d43b41570 · inbound

MetaPS: Adaptive Programmatic Strategy Selection for Market Agents cites this paper.

MetaPS: Adaptive Programmatic Strategy Selection for Market Agents Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 122

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T08:39:42.611122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-26T11:06:28.690956Z digest=sha256:51cda17f765462ad71f4913b065d6a064a96c20f0d34ae430d16eee996a63408

Observation 53b5e208-032d-458c-91c4-6c7daec29b65 · inbound

H-OPD: Confidence Aware Heterogeneous Multi-Teacher Multimodal On-policy Distillation cites this paper.

H-OPD: Confidence Aware Heterogeneous Multi-Teacher Multimodal On-policy Distillation Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-12T09:29:14.105502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-12T09:29:14.105502Z digest=sha256:a2eeabeddd768071a0e770e9266970c0ce5485694cb35a6f21a1e300c8aa5009

Observation 3fe60863-c310-4158-80a8-9a02861da75b · inbound

CritiqueDriveVLM: From Verifier-Guided Reinforcement Learning to Latent Thought Distillation for Autonomous Driving cites this paper.

CritiqueDriveVLM: From Verifier-Guided Reinforcement Learning to Latent Thought Distillation for Autonomous Driving Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 59

Resolution
unresolved
no resolver link, observed 2026-07-11T21:09:01.431209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T21:09:01.431209Z digest=sha256:d26916f953f0dec9fccbf9ccf4211c2990f0f7c9c5a3da28061bc6534e9266eb

Observation 5e933b81-ad73-442f-8fe3-4765e11e0c6c · inbound

Model Guides You How to Draw: Adaptive Visual Gating for Unified Multimodal Reasoning cites this paper.

Model Guides You How to Draw: Adaptive Visual Gating for Unified Multimodal Reasoning Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-14T01:07:06.861298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T01:07:06.861298Z digest=sha256:adbd7120e65c6a5ce6a8bad94b133865846eaa008fde6cc823622a5c7df45af7