Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 48 inbound Pith citation observations for arXiv:2504.07954.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T05:12:18.669805Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-05T11:41:02.675605Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation ae8d8474-6395-471d-9ed2-dad6e7b3dcb0 · inbound
Reinforced MLLM: A Survey on RL-Based Reasoning in Multimodal Large Language Models Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 134
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3363e78c-dbd3-49b5-a566-eea2ce310682 · inbound
UniVG-R1: Reasoning Guided Universal Visual Grounding with Reinforcement Learning Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 131b1d12-7e74-488f-9639-5a66260a5684 · inbound
RePrompt: Reasoning-Augmented Reprompting for Text-to-Image Generation via Reinforcement Learning Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54ee01f6-ba4b-4c0a-bea4-e51e7d7f9665 · inbound
Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 512576cf-8bdd-4828-88db-3d113f7ce37f · inbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d68c3269-b1d9-48f9-8ace-48cd4b1d7387 · inbound
Advancing Multimodal Reasoning via Reinforcement Learning with Cold Start Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a2ae015-0241-454b-8d78-6ab933a81cb4 · inbound
AlphaOne: Reasoning Models Thinking Slow and Fast at Test Time Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6575182-2a03-4c9a-a160-4fad195fd962 · inbound
Rex-Thinker: Grounded Object Referring via Chain-of-Thought Reasoning Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d120e7ae-903b-489a-aa89-14f38d5d8a2f · inbound
WeThink: Toward General-purpose Vision-Language Reasoning via Reinforcement Learning Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2f7a081-d9ec-4414-bb4c-2ecd80e87187 · inbound
Efficient Medical VIE via Reinforcement Learning Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 318971f5-3c8e-4986-9e14-17882b99da2e · inbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5c2021d-9678-414a-904f-2e5d2f55041b · inbound
Improving the Reasoning of Multi-Image Grounding in MLLMs via Reinforcement Learning Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation ca86a00a-d62e-4365-a261-53a173b04924 · inbound
StructVRM: Aligning Multimodal Reasoning with Structured and Verifiable Reward Models Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05bd08eb-7222-4cb6-880d-87c3c8d19e68 · inbound
An Explainable Machine Learning Framework for Railway Predictive Maintenance using Data Streams from the Metro Operator of Portugal Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b47c1fe3-3739-4914-b810-0b8a7b30ce6a · inbound
Beyond Reasoning Gains: Mitigating General-Capability Forgetting in Large Reasoning Models Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 107
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09e4aa05-6496-48db-a02d-4fc0d6edfcf2 · inbound
VIDEOP2R: Video Understanding from Perception to Reasoning Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 2865d547-2a01-4168-afe4-567ee1a15d04 · inbound
OneThinker: All-in-one Reasoning Model for Image and Video Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 2dd413a9-b3c8-492c-8eac-b7f26b238764 · inbound
RefBench-PRO: Perceptual and Reasoning Oriented Benchmark for Referring Expression Comprehension Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71fe76fd-1a14-4f4b-bde2-9582ca997613 · inbound
ProAgent: Harnessing On-Demand Sensory Contexts for Proactive LLM Agent Systems in the Wild Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 7a2cbe96-a7ea-458a-8649-21e4cc364f6a · inbound
Zoom-IQA: Image Quality Assessment with Reliable Region-Aware Reasoning Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d764c15-923c-4e85-99a5-88fa76f4428e · inbound
Can Textual Reasoning Improve the Performance of MLLMs on Fine-grained Visual Classification? Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 134255b6-36c8-4b8f-ab2e-40551bd5d901 · inbound
CamReasoner: Reinforcing Camera Movement Understanding via Structured Spatial Reasoning Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 1453a838-9af6-498b-a1f3-d2f3aca6e466 · inbound
PaLMR: Towards Faithful Visual Reasoning via Multimodal Process Alignment Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 653f93b5-ec73-4dff-97e5-21730460aae8 · inbound
CodePercept: Code-Grounded Visual STEM Perception for MLLMs Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83b88eaf-9ed7-4802-aeb5-e0b6c180f8db · inbound
V-Reflection: Transforming MLLMs from Passive Observers to Active Interrogators Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 03360d8f-d0ee-411e-aba6-b93696797b5d · inbound
Act Wisely: Cultivating Meta-Cognitive Tool Use in Agentic Multimodal Models Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 471b1c4a-5dc7-4667-8d48-6679a0729125 · inbound
Steadily moving semi-infinite fracture in plane poroelasticity Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 110
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 1b3b8c61-45ed-4b6e-b32b-58230ffa91b3 · inbound
XEmbodied: A Foundation Model with Enhanced Geometric and Physical Cues for Large-Scale Embodied Environments Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 110
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 013d3648-2590-407f-9943-9e8bc8b7be78 · inbound
SSL-R1: Self-Supervised Visual Reinforcement Post-Training for Multimodal Large Language Models Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 754de6ce-fdbc-4d95-8492-d22ba9e68312 · inbound
CGC: Compositional Grounded Contrast for Fine-Grained Multi-Image Understanding Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 3bc4e86f-2a6c-4127-9d6a-d4392f44e1d4 · inbound
Injecting Distributional Awareness into MLLMs via Reinforcement Learning for Deep Imbalanced Regression Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 9cc4566b-bd7b-40b5-8c5a-6c53122a417a · inbound
Injecting Distributional Awareness into MLLMs via Reinforcement Learning for Deep Imbalanced Regression Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation a9c3f915-f5ba-46bf-b7f4-de634d8f2efa · inbound
Self-Consistent Latent Reasoning: Long Latent Sequence Reasoning for Vision-Language Model Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation aede9686-4b49-444d-97bf-70e3bd2a780f · inbound
Self-Consistent Latent Reasoning: Long Latent Sequence Reasoning for Vision-Language Model Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 739c202d-54a7-48c9-a7a5-0a7ededea585 · inbound
CurveBench: A Benchmark for Exact Topological Reasoning over Nested Jordan Curves Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation c1c34f59-9dc7-4120-afc6-d10ba06562ba · inbound
CurveBench: A Benchmark for Exact Topological Reasoning over Nested Jordan Curves Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 048032d9-4f39-43d3-b943-7d2ca08caca8 · inbound
VideoSeeker: Incentivizing Instance-level Video Understanding via Native Agentic Tool Invocation Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 772dcefb-41e5-4bf0-9feb-5315db61efca · inbound
Reasoning Portability: Guiding Continual Learning for MLLMs in the RLVR Era Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation c023d8d4-1963-497e-8158-a53ef0397e09 · inbound
CaMo: Camera Motion Grounded Evaluation and Training for Vision-Language Models Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 8c666529-1689-4752-93a3-7448513c3cd5 · inbound
Guidance Contrastive Token Credit Assignment for Discrete Policy Optimization Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 3c63fc7a-baf4-4113-8d65-f3e6dcbb0f01 · inbound
From Structure to Synergy: A Survey of Vision-Language Perception Paradigm Evolution in Multimodal Large Language Models Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 171
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 00dc9687-0d54-45e4-9029-c3c156ae5d36 · inbound
Perceive-to-Reason: Decoupling Perception and Reasoning for Fine-Grained Visual Reasoning Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 35d974e9-fbad-4f29-9edf-19989a11d0e1 · inbound
DELTAVID: Enhancing Fine-Grained Spatiotemporal Perception with Cross-Video Differences Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54621a2c-7471-45b4-94fc-9b72a41f2ae0 · inbound
Stop Thinking, Start Looking: Efficient Post-Training for Multimodal Document Question Answering via Reasoning-Free Alignment Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 141dcf71-7f39-4c77-a1a5-63b7bb52163e · inbound
IoU-PD: IoU-Aware Privileged Distillation for Visual Grounding with Multimodal Large Language Models Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 120
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8cb68f21-faca-4452-8d73-7fbfea0e70dc · inbound
IoU-PD: IoU-Aware Privileged Distillation for Visual Grounding with Multimodal Large Language Models Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 120
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86241d39-3a7a-48cf-b4fd-83c5afcb81af · inbound
Hi-Token: Hierarchical Coordinate Tokenization for Generative Visual Grounding Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 118
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4189cd70-4d04-4c72-8181-5dc7d919a096 · inbound
SCOUT: Unlocking Enhanced Spatial Reasoning via Structured Chain-of-Thought and Multi-Objective Process Reward Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.