Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:38:46.508251Z
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 5 inbound Pith citation observations for arXiv:2505.12432.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:38:46.508251Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T22:36:14.896841Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-18T19:21:48.411112Z
25 of 25 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation a505e807-a861-4c0b-98d9-8e1720f2b6fc · outbound
Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1da66c8c-df16-4f25-9390-592bbbca1659 · outbound
Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78eb0315-3194-4bc4-a610-f33084568247 · outbound
Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01ce5095-d6a1-4c50-abb3-1c032426e652 · outbound
Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning OpenAI o1 System Card
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3de2749a-c9ef-48ca-8a11-c16cdaf572ac · outbound
Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning SceMQA: A Scientific College Entrance Level Multimodal Question Answering Benchmark
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd9a489e-3d24-478e-8cbe-8ca7ff4f0224 · outbound
Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3a6211f-9b00-45ea-92a2-ad3aadac7c63 · outbound
Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning s1: Simple test-time scaling
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 735177b5-6545-4778-bf20-2bc984e92764 · outbound
Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning GPT-4o System Card
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24df8388-fb31-43b5-8374-2447c667c5d1 · outbound
Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning?
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 673daf41-b42f-45df-bc96-62a88921f2fc · outbound
Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning Dast: Difficulty-adaptive slow-thinking for large reasoning models.arXiv preprint arXiv:2503.04472,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b05100e-6e67-42f7-bccf-94e02c4e7266 · outbound
Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning Math-LLaVA: Bootstrapping Mathematical Reasoning for Multimodal Large Language Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e115d9fa-dfea-44cb-baf5-0a547975168d · outbound
Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning LlamaV-o1: Rethinking Step-by-step Visual Reasoning in LLMs
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f599363-13ac-4945-8cbc-1c801bfcf3c3 · outbound
Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning Self-Consistency Improves Chain of Thought Reasoning in Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6ed0821-378e-4180-a4dd-925e76b47924 · outbound
Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning LLaVA-CoT: Let Vision Language Models Reason Step-by-Step
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db346497-f6f2-465a-b9db-98e7eb84f9c7 · outbound
Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ba20f66-acb9-4b26-9a8d-c9b26048fe75 · outbound
Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning DAPO: An Open-Source LLM Reinforcement Learning System at Scale
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51e63674-0375-40a4-ac0e-fe2c5a6716fe · outbound
Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba12231c-599e-4856-8da1-a59661153281 · outbound
Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems?
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8ece693-0bbd-448a-8bda-e65ecbde9a7d · outbound
Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6514ab55-bde0-4c99-8342-b8488a7798b1 · outbound
Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning Proximal Policy Optimization Algorithms
Reference 2015
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b8bf07e-fdf7-4fd9-84f2-e45c3aaa964f · outbound
Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12cb828b-d076-4d14-a2f5-f8e6f59452d9 · outbound
Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning Logic-RL: Unleashing LLM Reasoning with Rule-Based Reinforcement Learning
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b74eb5e-bf75-46a1-9f50-7334544e6f31 · outbound
Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ed0cbac-6ab7-47ad-802e-2f06fb49c99c · outbound
Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning Polymath: A Challenging Multi-modal Mathematical Reasoning Benchmark
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b59c9d22-f03f-4e1d-bcb0-92607a77b814 · outbound
Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning Qwen2.5-VL Technical Report
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ccc0b7ab-2e2b-4db4-ad07-1df3b15e635b · inbound
HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5fdd68c9-9932-4818-b4ee-3ff99ee3e845 · inbound
APO: Enhancing Reasoning Ability of MLLMs via Asymmetric Policy Optimization Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21d7b7ee-9adf-4259-bc08-f4f631342e59 · inbound
The Landscape of Agentic Reinforcement Learning for LLMs: A Survey Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning
Reference 241
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 083c6a0b-ebfc-4fbd-8cd4-7ace85e552bb · inbound
HART: High-Resolution Annotation-Free Reasoning Technique through a Closed-loop Framework Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77d9b4ec-6780-4bfe-96a7-700501be8b6b · inbound
LenGuard-GPC: Length Guarding with Guided-Prompt Consistency for Spatial Reasoning Reinforce Learning Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.