Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:10:14.202049Z
Paper Citation Record · LEDGER
As of 22 August 2026, this Paper Citation Record lists 65 of 65 outbound references and 20 inbound Pith citation observations for arXiv:2505.16192.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:10:14.202049Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T23:21:12.926143Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
65 of 65 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation 0950a071-901a-45c6-a06a-8ba43b563435 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought https://deepmind.google/technologies/gemini/
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation de2a7e6e-7fe8-4598-8bdc-e4b044e5c007 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought https://openai.com/index/introducing-o3-and-o4- mini/
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation a989b189-95c9-4c95-9d01-a4190fa85214 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought https://qwenlm.github.io/blog/qvq-72b-preview/
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation a5c7d6b5-0e62-42ef-a7cf-7ac074e493a5 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought https://huggingface.co/datasets/TheEighthDay/SeekWorld
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 29727806-d59a-4088-9f2e-31af322b8723 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Flamingo: a visual language model for few-shot learning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation f82d7041-8c4e-4e48-8aa9-d6e71c94b091 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Gemini: A Family of Highly Capable Multimodal Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43fa7503-a1e3-44dc-b529-0cb4945c3b7f · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84c52e37-6fd1-4bec-b5ce-fdede12d09c4 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Qwen2.5-VL Technical Report
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c98685ab-94a2-4e9f-8982-ff40d49a9a6a · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought M 3cot: A novel benchmark for multi-domain multi-step multi-modal chain-of-thought, 2024
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 4afef9e2-3862-46db-8b3f-ac67746f1a82 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Unresolved cited work
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 512219ac-acd7-4f08-82bd-2822370f9ec0 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Expanding performance boundaries of open-source multimodal models with model, data, and test-time scaling, 2025
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 450f395a-44f9-4b55-bca0-8b2b8bf3bbc2 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75eaa120-3b25-458f-ac2c-201dc42adbf5 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Virgo: A preliminary exploration on reproducing o1-like mllm, 2025
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation ba5c080c-43ad-43e0-b9ba-ceb6db7b3118 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6131cccc-c161-40d4-85b9-7360e6fa1a28 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Hallusion- bench: An advanced diagnostic suite for entangled language hallucination and visual illusion in large vision-language models, 2024
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 2d87f4ec-d619-4cb2-b5c5-a6634f812c64 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought REINFORCE++: Stabilizing Critic-Free Policy Optimization with Global Advantage Normalization
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 423304cc-522e-4e76-985c-534d7140d033 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Gqa: A new dataset for real-world visual reasoning and compositional question answering
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a111b37-51f5-41c5-8693-1d7eec2198d8 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought MathPrompter: Mathematical Reasoning using Large Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ce2b8b3-de00-4422-ab69-f0ead01be8b9 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Tab-cot: Zero-shot tabular chain of thought
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 36ec47c7-fa3a-46cd-8b34-f9cefd8e4966 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Imagine while reasoning in space: Multimodal visualization-of-thought, 2025
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3be5777e-2352-4c7a-a174-753e4a5195bf · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Llava-next-interleave: Tackling multi-image, video, and 3d in large multimodal models, 2024
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6394dd6-7423-497d-8f53-349fcbf14425 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 031441a8-32e5-4bf6-98e9-0c8f5558a271 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Let’s verify step by step, 2023
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46af13df-a9ab-4ea2-bd91-0c9967fed09b · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought DeepSeek-V3 Technical Report
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b34ef889-8ed0-4c08-bfcc-4db52ab25a86 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Visual spatial reasoning, 2023
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c310283-a77a-4a0b-8a95-3c006dbed489 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Improved Baselines with Visual Instruction Tuning
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c429e97-4e44-4af1-99f7-8c07904cb77b · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Seg-zero: Reasoning-chain guided segmentation via cognitive reinforcement, 2025
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation d16f2635-af50-4e65-b935-ab1a39e61149 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Visual-rft: Visual reinforcement fine-tuning, 2025
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98631811-9405-49ef-bfe7-6677c9eaf95c · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Mathvista: Evaluating mathematical reasoning of foundation models in visual contexts, 2024
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8b818d9-0a87-457c-b8d7-64dc11dad799 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Learn to explain: Multimodal reasoning via thought chains for science question answering, 2022
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d254833d-c3cf-412e-89b0-e1a3cb72c310 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought V Jawahar
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 596b13ad-e256-46ac-addb-77152f353952 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Unresolved cited work
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0562d9e9-61e4-4f1f-9d53-cc899a91a916 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Visual-O1: Understanding Ambiguous Instructions via Multi-modal Multi-turn Chain-of-thoughts Reasoning
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3caf8933-d096-4cbe-9b05-67d07173d7ce · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Skeleton-of-Thought: Prompting LLMs for Efficient Parallel Generation
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6cf2ee6d-e3ed-41ce-b7ce-18f13abab1ed · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 439815e4-6c98-4276-8837-22a22b0ddd60 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Unresolved cited work
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation e42b4b21-f613-43c1-8dae-7d5ef8b0ebf5 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Gpt-4v(ision) system card
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8d91474-9220-4939-a7d9-feb4bd2b6d51 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Unresolved cited work
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6cdfce82-46d1-41ca-97ae-3d97c0ffd5e0 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Lmm-r1: Empowering 3b lmms with strong reasoning abilities through two-stage rule-based rl, 2025
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 088d5aa3-919b-412f-bf99-72f47975f052 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e73181cc-24de-4bc7-a1ef-0206e9c48b12 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Deepspeed: System optimizations enable training deep learning models with over 100 billion parameters
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 824d62e9-3f98-438e-9cc0-bbd5e9de4f34 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c91f2f17-9755-406c-b8f4-543711379eda · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Proximal Policy Optimization Algorithms
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9394e39f-b78c-4e6a-99c2-16d01cbda5d4 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Visual cot: Advancing multi-modal language models with a comprehensive dataset and benchmark for chain-of-thought reasoning, 2024
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 2dda3982-92c7-4584-a460-adedfb2bc661 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Unresolved cited work
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d7b0369-51d0-4b37-92e5-34cc502a2485 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Vlm-r1: A stable and generalizable r1-style large vision-language model, 2025
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 8b1b5965-1ea6-45b3-ad43-b39e1c67ed30 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Towards vqa models that can read
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 770be980-d70c-4be1-8953-22287adf3c29 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0867eecc-0353-42d4-b8f0-7fe6c80e3aed · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought A Survey of Reasoning with Foundation Models
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f1c7586-bf8e-4c45-ba24-25554f6cb3b5 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Mm-verify: Enhancing multimodal reasoning with chain-of-thought verification, 2025
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22ad4768-6ba8-472b-94d3-56213fab3cf9 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Kimi k1.5: Scaling Reinforcement Learning with LLMs
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fdb46ffc-3c0d-443e-9558-9af7b145e9b2 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Llamav-o1: Rethinking step-by-step visual reasoning in llms, 2025
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15e7fecf-8aac-42b7-b5da-dc2bc16fffd6 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Measuring multimodal mathematical reasoning with math-vision dataset, 2024
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation bb8395e6-d08f-4a8e-a7ff-49fd77e1e389 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Self-Consistency Improves Chain of Thought Reasoning in Language Models
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45b885bb-646b-43c7-a571-e45c7916ff37 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Chain-of-thought prompting elicits reasoning in large language models, 2023
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df3c4660-9ecc-48a2-ab0a-29eea0d7cab9 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Boosting multimodal reasoning with mcts-automated structured thinking.arXiv preprint arXiv:2502.02339, 2025
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2fbea10-09db-4734-9a55-a9cbe4d558f2 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Llava-cot: Let vision language models reason step-by-step, 2025
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4420ba65-9895-43a3-9c82-cbc4f9dfc36d · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought R1-onevision: Advancing generalized multimodal reasoning through cross-modal formalization, 2025
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c4a7202-d320-4948-bafb-d89e722b0d78 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Mulberry: Empowering mllm with o1-like reasoning and reflection via collective monte carlo tree search, 2024
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa6a6c8a-356e-40b2-b24f-64276641fa8d · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Tree of Thoughts: Deliberate Problem Solving with Large Language Models
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc02ea5e-e333-4105-9fa9-899492236544 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought mPLUG-Owl: Modularization Empowers Large Language Models with Multimodality
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d15f8ad1-dd4b-4635-be59-43d437fd5024 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi, 2024
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 2907a747-20df-4ae8-b728-efaaae2a8939 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Star: Bootstrapping reasoning with reasoning.Advances in Neural Information Processing Systems, 35:15476–15488, 2022
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44fa5f25-11c6-4fad-a0ca-ef63fd42e4b0 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Multimodal Chain-of-Thought Reasoning in Language Models
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35e4a926-a3aa-49b2-b595-c63477df9e60 · outbound
VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Reflection of Thought: Inversely Eliciting Numerical Reasoning in Language Models via Solving Linear Systems
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 42c6c1bb-1db9-4ab9-9877-dd4a870d0fb4 · inbound
Perception, Reason, Think, and Plan: A Survey on Large Multimodal Reasoning Models VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought
Reference 192
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3eebcfac-f75a-4bb9-ba5d-69145ee7db7c · inbound
DeepEyesV2: Toward Agentic Multimodal Model VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 90e92e39-f412-4ebf-90c2-fb05c19a2c0e · inbound
PhyDetEx: Detecting and Explaining the Physical Plausibility of T2V Models VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 6b6f1835-d10b-4e24-ae05-668478a595b6 · inbound
OmniDrive-R1: Reinforcement-driven Interleaved Multi-modal Chain-of-Thought for Trustworthy Vision-Language Autonomous Driving VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 7bd2ed84-da36-4a67-97fb-fe33e2faaae2 · inbound
Zoom-IQA: Image Quality Assessment with Reliable Region-Aware Reasoning VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a120669a-8048-4583-8d48-6b584cac864b · inbound
Imagination Helps Visual Reasoning, But Not Yet in Latent Space VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d20484e1-aae1-4b27-946d-95ca8299eeeb · inbound
CharTool: Tool-Integrated Visual Reasoning for Chart Understanding VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 697bb1cf-2eea-47d3-9e26-3cff66cbf941 · inbound
See Further, Think Deeper: Advancing VLM's Reasoning Ability with Low-level Visual Cues and Reflection VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 8243bba0-66b1-4c28-95a8-0c40596e4190 · inbound
Perceptual Flow Network for Visually Grounded Reasoning VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 262f28c4-bf5f-4a80-a590-b54756f9abbb · inbound
MHPR: Multidimensional Human Perception and Reasoning Benchmark for Large Vision-Languate Models VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 60ec0b92-9072-4c86-9b6b-d5661b39d939 · inbound
When Relations Break: Analyzing Relation Hallucination in Vision-Language Model Under Rotation and Noise VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 3f53fdc3-3fd8-4364-a2ed-1beb43ac3a8d · inbound
When Relations Break: Analyzing Relation Hallucination in Vision-Language Model Under Rotation and Noise VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation ed2eb08a-8cf8-4e43-a016-e37db3fc75c4 · inbound
Self-Consistent Latent Reasoning: Long Latent Sequence Reasoning for Vision-Language Model VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 09482764-0ac1-4214-b55c-2cbba6b5b528 · inbound
Self-Consistent Latent Reasoning: Long Latent Sequence Reasoning for Vision-Language Model VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 9f1803a0-0cc1-4b0c-bd37-dce56e429f34 · inbound
Position Rebinding Cache Reuse: Replay-Free Visual Revisiting for Interleaved Multimodal Reasoning VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 00ebaf5f-2631-4fd4-a108-c12080a2a6ef · inbound
How LLMs See Creativity: Zero-Shot Scoring of Visual Creativity with Interpretable Reasoning VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 6595ff6c-e86a-42fe-aaa4-b29ab7409254 · inbound
How LLMs See Creativity: Zero-Shot Scoring of Visual Creativity with Interpretable Reasoning VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation b360c749-1614-45a1-8676-f81404cd65e4 · inbound
Beyond Zooming: Learning Multi-Tool Visual Reasoning for Ultra-High-Resolution Remote Sensing VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0608c6b9-0449-4965-9822-d85057b798a4 · inbound
OPLD: On-Policy Latent Distillation for Multimodal Reasoning VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81ef07af-f41b-4322-ad2f-254890b8419c · inbound
InSight-doc: Agentic Visual Perception for Long-Document Understanding VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought
Reference 146
Source-reported events for the cited work
Unavailable: canonical work link unavailable.