Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T00:13:39.624185Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 1 inbound Pith citation observation for arXiv:2506.14907.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T00:13:39.624185Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-12T02:24:12.349405Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-12T07:41:27.623339Z
54 of 54 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 7092c52b-5bca-473c-9826-8149a8b766be · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a6445511-1020-4d41-bf05-90d35ac66653 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a406960-5b41-49e5-ab7d-faf1c48482fa · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning Flamingo: a visual language model for few-shot learning.Advances in neural information processing systems, 35:23716–23736, 2022
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99ab7755-a867-4a3f-a010-852e7f934c84 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning OpenFlamingo: An Open-Source Framework for Training Large Autoregressive Vision-Language Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edb2f5a2-d2e2-418c-9825-522082cb27e5 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning Qwen2.5-VL Technical Report
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28e8f5b2-f541-40ab-bddc-7ac53985a7e9 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning R1-v: Reinforcing super generaliza- tion ability in vision-language models with less than $3.https://github.com/Deep-Agent/ R1-V, 2025
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 184742c6-bfff-4dff-b576-0814a270fcff · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9654d880-b9d4-40cc-9421-ba067ee95d51 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning How far are we to gpt-4v? closing the gap to commercial multimodal models with open-source suites.Science China Information Sciences, 67(12):220101, 2024
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81d95e07-7c81-4516-aee5-534b3670bf48 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e8e6fc8-a836-430b-a708-30f8d4e9fa8c · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f041fc3e-2fd4-49b4-ac1e-f3b8011fc45b · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning Vlmevalkit: An open-source toolkit for evaluating large multi-modality models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1433762e-0e14-48f1-9b58-26362d305067 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning Blink: Multimodal large language models can see but not perceive
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de473be6-df96-4f94-8e09-42e1161ac5ac · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ba9abc4-c2b4-418a-97f2-d52f1800de0d · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 478156f5-7241-4773-a1a2-e247ff7c9487 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning MANTIS: Interleaved Multi-Image Instruction Tuning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec0b58ea-ebb6-455b-8d00-5f57b13f99a5 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning Remi: A dataset for reasoning with multiple images.Advances in Neural Information Processing Systems, 37:60088– 60109, 2024
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1e62c27f-6b19-485b-bdc4-9715a2074995 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning MIMIC-IT: Multi-Modal In-Context Instruction Tuning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e53a2749-ad2a-4335-a6f2-d026a31666f3 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning LLaVA-OneVision: Easy Visual Task Transfer
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa1eb206-ca7b-4363-b8fd-77b9c0e6ef69 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2274912a-62dd-498c-9c8a-40ea3cd393f8 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning Migician: Revealing the Magic of Free-Form Multi-Image Grounding in Multimodal Large Language Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e90f3af-d9cf-4891-82ab-e32038325b4b · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning From System 1 to System 2: A Survey of Reasoning Large Language Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a1a9c86-15d6-4581-8f42-f0c60376f761 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning Improved baselines with visual instruction tuning
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c824f2f8-60bc-48b9-b6c8-08237dc06a80 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning Noisyrollout: Reinforcing visual reasoning with data augmentation
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4ca0f18-ef2c-4ad0-934b-34a7904cc1df · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning Visual-RFT: Visual Reinforcement Fine-Tuning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ff7774c-4b31-439e-9014-b250ba33e0d9 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e8f4daf-9202-42bd-bce7-898792bd8cbc · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bdadbe14-78f8-4313-9062-86a97b7c5f15 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning MMIU: Multimodal Multi-image Understanding for Evaluating Large Vision-Language Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d624d141-7b69-4ce4-92d1-14900e371b1d · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning Compositional chain-of- thought prompting for large multimodal models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b003eea8-f02f-44ca-b787-9f27e783d4f9 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62e16628-6758-4883-802e-bb8abce82333 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90c1b662-c989-420d-b621-aa1b4ad3a9b6 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b60d486f-3e83-4dfc-82a1-882f978f8b58 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning HybridFlow: A Flexible and Efficient RLHF Framework
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fbe8e408-d38e-4d83-b997-b3745f1fca6d · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning Generative multimodal models are in-context learners
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ce6b951-cd47-4683-884f-15b36ff612f0 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning Reason-rft: Reinforcement fine-tuning for visual reasoning.arXiv preprint arXiv:2503.20752, 2025
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ab728e6-e5c3-49aa-9bf2-89d59bb4acb3 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning Gemini: A Family of Highly Capable Multimodal Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd78d89e-0a1d-4c24-9d28-09fb75ad6012 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning Identifying and Mitigating Position Bias of Multi-image Vision-Language Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba28c4ed-87bd-4f28-8a1e-0dc9d73bc252 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning MuirBench: A Comprehensive Benchmark for Robust Multi-image Understanding
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1964cd2d-2e29-4943-8abb-0bb44abc25d2 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning VL-Rethinker: Incentivizing Self-Reflection of Vision-Language Models with Reinforcement Learning
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f441ceca-fd9d-4f93-be5f-703d0c7dc510 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning Measuring multimodal mathematical reasoning with math-vision dataset
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d44726a-0687-4b8f-8626-30727eb0e571 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning MV-MATH: Evaluating Multimodal Math Reasoning in Multi-Visual Contexts
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1a54b7a3-06ed-4e51-9bac-347642c2cbbc · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b17b943e-74b1-4da0-b8b0-fa9175c1f749 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a55003c-15a0-438a-849e-98e537465ebb · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7dcb0f31-3bf2-4cea-bf6c-bd8fbc64ad49 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 318971f5-3c8e-4986-9e14-17882b99da2e · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8660b85e-6440-41bf-bfad-4132c46f346c · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e96dd684-1d67-49e9-ade6-dc48f2146cc4 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning Weaving Context Across Images: Improving Vision-Language Models through Focus-Centric Visual Chains
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04c97516-3cfc-4ad1-b5ff-7318ad84af4b · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning Mathverse: Does your multi-modal llm truly see the diagrams in visual math problems? InEuropean Conference on Computer Vision, pages 169–186
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d101a7d7-f6ae-4f1b-b221-7b6c73884ae8 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning R1-Zero's "Aha Moment" in Visual Reasoning on a 2B Non-SFT Model
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71e1be06-232f-4421-a9e0-465fa8439e73 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning <image> <image> <image>
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 658e9f6c-80b5-417d-ae13-dc24caa3edd0 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning A", "B",
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0489e2ff-37dc-4717-83b3-87d0ffaefcc4 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning question
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e6222173-1915-4f3d-bf24-b3597aaff662 · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning Unresolved cited work
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 49b31685-98cd-4e34-ac50-81f53b919c9b · outbound
PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning should_change
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b6ea3834-aab7-4c19-9a34-08b612d1218e · inbound
Reflection Anchors for Propagation-Aware Visual Retention in Long-Chain Multimodal Reasoning PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.