Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T21:41:41.103071Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 3 inbound Pith citation observations for arXiv:2506.23563.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T21:41:41.103071Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-04T16:07:45.693853Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-16T15:04:22.765546Z
51 of 51 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation be9ea60a-6bb1-447f-9f50-bc46dd5b167e · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI Vqa: Visual question answering
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de0012f4-cfd3-446e-bb25-da7f0447338f · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI Qwen2.5-VL Technical Report
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f03e35c7-15c6-4412-a186-6ba364de3669 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI Are We on the Right Way for Evaluating Large Vision-Language Models?
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36ffbdb7-f622-4f64-8ec5-c54f06370650 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI R1-v: Reinforcing super generalization ability in vision- language models with less than $3
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7730835c-196e-46e4-8e49-79513dffde51 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI M 3cot: A novel benchmark for multi- domain multi-step multi-modal chain-of-thought
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2c7e9642-a714-4ae7-96d4-65bf223d4959 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI Vlmevalkit: An open-source toolkit for evaluating large multi-modality models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a59ed741-d1c7-4846-9d49-55986827a34e · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI The Llama 3 Herd of Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ebdc35f-8e27-473e-997c-4698bd2eadc0 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI Mme: A compre- hensive evaluation benchmark for multimodal large language models, 2024
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 93ab7eea-85ef-4eb5-9497-ead8a6c49b62 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI Hallusionbench: an advanced diagnos- tic suite for entangled language hallucination and visual il- lusion in large vision-language models
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a9715380-7d87-4d83-986f-97ae974160d1 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77e758fd-1c0b-4b84-ad84-7a408c596d97 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb28a5cd-421c-445f-9007-ff80bac185c9 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b0541e0-2129-48eb-9b5d-c0aaf865050a · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI Gqa: A new dataset for real-world visual reasoning and compositional question answering
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a995200-d266-4af7-b4d1-4c749bb2f0a5 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI GPT-4o System Card
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1e9aacb-9249-43fa-a880-9ded340d290c · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI OpenAI o1 System Card
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d2fc3ba-87a4-408e-b6d2-28629dc2b5da · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI MME-CoT: Benchmarking Chain-of-Thought in Large Multimodal Models for Reasoning Quality, Robustness, and Efficiency
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce92bb88-2148-40fe-9df0-df9730d36ce9 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI LLaVA-OneVision: Easy Visual Task Transfer
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e47b29ba-02f9-44ef-b388-744b0b1acbe1 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI Evaluating Object Hallucination in Large Vision-Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29fc1abd-5a76-480e-a5ca-aa060e71d691 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI From System 1 to System 2: A Survey of Reasoning Large Language Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc37b8db-903b-4034-bbb2-e83c62082ef2 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI DeepSeek-V3 Technical Report
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c142dbff-ded5-415c-9b20-b7d4a1992ca8 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI X-Reasoner: Towards Generalizable Reasoning Across Modalities and Domains
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0000f601-0756-4c2f-8a0f-7a2471a0e18f · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI Mmbench: Is your multi-modal model an all-around player? In European conference on computer vi- sion, pages 216–233
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0e6e344c-68b5-455a-9592-b135010773f9 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf81a703-ce5f-46bb-910c-e78427fd777b · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5144df1c-c94b-4336-8a75-5c547156bb85 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI Visual cot: Advancing multi-modal language models with a com- prehensive dataset and benchmark for chain-of-thought rea- soning
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0f2f5a40-523a-4187-a9dd-eff88897d922 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f337445-6d0b-4852-9546-0e8596ac8d34 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI Claude 3.7 sonnet, 2025
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 38f60056-f59a-4e94-941d-da25dc558f65 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 864c0936-66c9-4841-a69e-43aa9db98b84 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI Kimi k1.5: Scaling Reinforcement Learning with LLMs
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33880566-25cf-42b6-8ffc-4764fa08ec7b · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI Qvq: To see the world with wisdom, 2024
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b36a3b7a-3683-4148-9a7a-dd9fd47ab33c · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI Qwq: Reflect deeply on the boundaries of the unknown, 2024
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 944a66d3-3182-42f5-bbbd-ff2df9985f9f · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc1e6874-59e1-4d60-ac44-28e7c5356027 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI Open-r1-video
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d176d38e-f478-4a60-859b-b22c4b999294 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI VisuoThink: Empowering LVLM Reasoning with Multimodal Tree Search
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d763cf7c-4589-4b5c-be57-5cf424486580 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI Charxiv: Charting gaps in realistic chart understanding in multimodal llms
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1e22d933-8446-44f4-b425-e1ea0b1ed865 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI Boosting mul- timodal reasoning with mcts-automated structured thinking
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47746207-6c7c-4ee6-8b2c-07e227d36c45 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58b2ef09-753d-40cc-9c41-ae15c01fabd7 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI Logic-RL: Unleashing LLM Reasoning with Rule-Based Reinforcement Learning
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea48561d-de45-475c-a596-68d921cd0686 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI LLaVA-CoT: Let Vision Language Models Reason Step-by-Step
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fec71ebb-790b-405a-aa09-33befdc3f34a · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2aacdaa0-ba2e-4fb9-901f-e0efd17ce28d · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI R1-ShareVL: Incentivizing Reasoning Capability of Multimodal Large Language Models via Share-GRPO
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3559689b-0da9-42f6-94d4-496b6fe9deb6 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI MiniCPM-V: A GPT-4V Level MLLM on Your Phone
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ea9b439-fd49-40a5-9db9-5e1f2a8f0b91 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI MMT-Bench: A Comprehensive Multimodal Benchmark for Evaluating Large Vision-Language Models Towards Multitask AGI
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf7dae34-fce7-45c6-bffe-2d2f5dfb778b · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19eed217-3bfb-42ce-86d9-feddd144e924 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for ex- pert agi
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4d8d3328-49e0-4ce5-b281-1c03ffe00383 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI MMMU-Pro: A More Robust Multi-discipline Multimodal Understanding Benchmark
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe7fe619-ec37-48b2-b76d-529d226dc326 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI MM1.5: Methods, Analysis & Insights from Multimodal LLM Fine-tuning
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee42f373-7c73-4b4d-b7db-39c4af95d126 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1fcd04ec-7e36-4357-9dbc-10dbab8f295e · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI Mathverse: Does your multi-modal llm truly see the diagrams in visual math problems? In European Conference on Computer Vision, pages 169–186
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9ca8fd4c-e733-4ed2-ab5d-dc3e9595b0a2 · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI Improve Vision Language Model Chain-of-thought Reasoning
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68b418f4-f233-4f31-b187-9a27584609bf · outbound
MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41cb0031-bb58-4119-abee-969cf25cdbb7 · inbound
R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 54e58cba-4bf8-433d-a37c-d4188b30aaea · inbound
Reinforcement Learning Meets Large Language Models: A Survey of Advancements and Applications Across the LLM Lifecycle MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI
Reference 224
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c7ca8d1-37a3-4a88-9c36-c51ee1825ca8 · inbound
OMIBench: Benchmarking Olympiad-Level Multi-Image Reasoning in Large Vision-Language Model MMReason: An Open-Ended Multi-Modal Multi-Step Reasoning Benchmark for MLLMs Toward AGI
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.