Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:24:54.668238Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 16 inbound Pith citation observations for arXiv:2505.22019.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:24:54.668238Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T13:23:34.984840Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T15:09:55.135392Z
54 of 54 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 69531200-78c5-409b-9075-705dee559482 · outbound
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8ccdf16-eff3-4b06-8bc0-390b31df102e · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning Qwen2.5-VL Technical Report
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 268315af-74ce-4a4d-8877-217f37f1bfc5 · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning Benchmarking large language models in retrieval-augmented generation
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5baac46d-a7d7-4111-b649-57ccaf6cddc8 · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning R1-v: Reinforcing super generalization ability in vision-language models with less than \ 3
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation cb69defb-b57e-41c1-bb83-ccf06e98986f · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0af78db5-411e-4aa0-afe7-7ceefecb0d9d · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning Mindsearch: Mimicking human minds elicits deep ai searcher
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 762efeb2-6d0c-4bec-97a0-d036d833bbeb · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning Agent-FLAN: Designing Data and Methods of Effective Agent Tuning for Large Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9790918-9649-4cc5-ab71-331c48752194 · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d32f2860-0616-4982-b3a8-d69e8cd7924b · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8925cca-2976-4e05-b1cd-19f0c477861c · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning PP-OCR: A Practical Ultra Lightweight OCR System
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation acb7fd87-8eb6-47a1-bd8b-4ed82d8b5d9e · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning Colpali: Efficient document retrieval with vision language models
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b48724e1-3d34-4973-8f85-220a4d7e13b9 · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning Retrieval-Augmented Generation for Large Language Models: A Survey
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07fc34a4-9d1a-4da1-a76d-6a940277e99b · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ff22933-ee18-42ad-9a6e-3307aa72aa7c · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00eeeb99-9d5a-4aeb-9c23-ed590f381a76 · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning OpenAI o1 System Card
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00e16907-51c7-4867-a3b2-3f372a8b503a · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning MMSearch: Benchmarking the Potential of Large Models as Multi-modal Search Engines
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7548be0-6977-4d9b-8ed3-e540525ea30f · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning DeepRetrieval: Hacking Real Search Engines and Retrievers with Large Language Models via Reinforcement Learning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0914aadd-44d0-486a-a3d4-d761414d506c · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning Long-Context LLMs Meet RAG: Overcoming Challenges for Long Inputs in RAG
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ead5704-8345-4c87-852f-5f9e97d29a78 · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e205582-82cd-4269-8724-f25b68dd7ad2 · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning Reinforcement learning: A survey
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6d2d3ac9-3765-48a5-8255-04acfd34cbbf · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning NV-Embed: Improved Techniques for Training LLMs as Generalist Embedding Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2d12a3a-869d-4955-84ef-bb3ddd92a9cb · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning u ttler, Mike Lewis, Wen-tau Yih, Tim Rockt \
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0219256-bd67-4df6-9a69-389dde6c5c69 · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning Search-o1: Agentic Search-Enhanced Large Reasoning Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11459a76-2481-48f2-b128-0ef82e496bae · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning Benchmarking Multimodal Retrieval Augmented Generation with Dynamic VQA Dataset and Self-adaptive Planning Agent
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b9ceca8-29b6-4506-b59b-b18646e7e1e1 · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning Towards General Text Embeddings with Multi-stage Contrastive Learning
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8acbfd5-bd5a-4648-b7b9-5e14039ded2f · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning Improved baselines with visual instruction tuning
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0f17fc1-8a65-485b-ae63-665bbe9a6844 · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning LlamaIndex , 11 2022
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02dc48b8-0155-4e5a-a008-61cb42d6d12e · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning Visual-RFT: Visual Reinforcement Fine-Tuning
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0ed3f2c-aa1f-435a-90bb-bee0c89944d7 · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning MMLongBench-Doc: Benchmarking Long-context Document Understanding with Visualizations
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d634e7bc-32f8-43f8-a76e-bec59d2d6849 · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c2ea2cf-8642-469e-924b-799c5c19984d · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning Simpo: Simple preference optimization with a reference-free reward
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c8c0d4e-b367-4663-93eb-1233a01ea3dd · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning NV-Retriever: Improving text embedding models with effective hard-negative mining
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4dd90c39-b902-4ffc-b08b-f7206b3b0c2e · outbound
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fcafe132-9e4f-4759-bc04-037d07ee81cb · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning Introducing gemini 2.0: our new ai model for the agentic era, 2024
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d251e266-f212-41fb-bcc0-cba2aec026cf · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning Direct preference optimization: Your language model is secretly a reward model
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f33cf807-ebaa-485e-92ef-8f9286a003c8 · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning Proximal Policy Optimization Algorithms
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65ebfb42-2e5d-49bc-9dfe-29506b85d40c · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning Visual cot: Advancing multi-modal language models with a comprehensive dataset and benchmark for chain-of-thought reasoning
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d74a9cfd-9ee9-48f7-b84d-946c85862326 · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning HybridFlow: A Flexible and Efficient RLHF Framework
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b86812e1-643b-4b23-82d6-b42c127aa112 · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning Reinforcement learning
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ebfbe6e8-0a78-475e-9803-6fd86710633c · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning Slidevqa: A dataset for document visual question answering on multiple images
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 056425ae-5ace-4508-a89e-5f31bb5c7e35 · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning ViDoRAG: Visual Document Retrieval-Augmented Generation via Dynamic Iterative Reasoning Agents
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ebaea24-4713-4d9a-a30b-d63e2cd79fc3 · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning RAGEN: Understanding Self-Evolution in LLM Agents via Multi-Turn Reinforcement Learning
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5bd4bdc1-02f8-4d2f-98c2-88b1d53c7328 · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning Simple statistical gradient-following algorithms for connectionist reinforcement learning
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f15666ee-bdbc-4691-9a14-5246d2f099a3 · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning WebWalker: Benchmarking LLMs in Web Traversal
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e328ef64-c72b-47c8-8eda-6451b5932802 · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning Unfolding the Headline: Iterative Self-Questioning for News Retrieval and Timeline Summarization
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04b2de5d-2172-4ee4-baca-17c71b3a49fa · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning Rule: Reliable multimodal rag for factuality in medical vision language models
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3ec533b4-47bb-425c-bc96-4f0808b4965f · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning Qwen2.5 Technical Report
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 161fa363-a68a-42c3-8582-5de71460e329 · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning React: Synergizing reasoning and acting in language models
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 512576cf-8bdd-4828-88db-3d113f7ce37f · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning Perception-R1: Pioneering Perception Policy with Reinforcement Learning
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59e799c9-745e-4f80-8775-79913f385c70 · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning Introducing Visual Perception Token into Multimodal Large Language Model
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bdcf5e24-caf3-41f1-8888-177af150b37d · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning VisRAG: Vision-based Retrieval-augmented Generation on Multi-modality Documents
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23622d0a-3561-49b8-9458-c3817c6b7f97 · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning LlamaFactory: Unified Efficient Fine-Tuning of 100+ Language Models
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d0bfe84-fb56-4288-a145-74e854cc139d · outbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning , " * write output.state after.block = add.period write
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d2679c0-1349-4812-9f51-95df8bf0e25d · outbound
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4dc91ff-51d2-4752-aa5a-47c8c4b80e89 · inbound
Learning Only with Images: Visual Reinforcement Learning with Reasoning, Rendering, and Visual Feedback VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf6bdab8-5b24-4a69-9e98-363b25af0fbb · inbound
M2IO-R1: An Efficient RL-Enhanced Reasoning Framework for Multimodal Retrieval Augmented Multimodal Generation VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a8a7c1b-2d4e-46f7-b0c4-35065cf68ee2 · inbound
DeepEyesV2: Toward Agentic Multimodal Model VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2967d069-f60e-4f57-8581-64b9eed36b05 · inbound
GeoBrowse: A Geolocation Benchmark for Agentic Tool Use with Expert-Annotated Reasoning Traces VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 97ad0c2a-e9dc-4d15-8928-993a8a37b856 · inbound
ReAlign: Optimizing the Visual Document Retriever with Reasoning-Guided Fine-Grained Alignment VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation bec402f7-7a52-459e-b962-8d50ea8d82cf · inbound
VISOR: Agentic Visual Retrieval-Augmented Generation via Iterative Search and Over-horizon Reasoning VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 36ea29d2-7175-46ed-95d2-4625e83a559c · inbound
POINTS-Seeker: An Open Recipe for Multimodal Search Agents with Visual Memory Management VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ca24bd0d-f4ba-4938-9eea-6f75e42bbbf8 · inbound
POINTS-Seeker: An Open Recipe for Multimodal Search Agents with Visual Memory Management VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37e43759-ef2d-4700-bc61-27bd5aa2e355 · inbound
CC-OCR V2: Benchmarking Large Multimodal Models for Literacy in Real-world Document Processing VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ce49c680-a71e-42ed-919f-fd2c54d9a638 · inbound
SMMBench: A Benchmark for Source-Distributed Multimodal Agent Memory VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 39c126c9-dc8c-43cd-8c4d-2adf8647af8b · inbound
From Structure to Synergy: A Survey of Vision-Language Perception Paradigm Evolution in Multimodal Large Language Models VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning
Reference 215
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6ad6b6c6-2c8d-42bc-8c53-fde35b53ed31 · inbound
SimpleSearch-VL: A Simple Recipe for Multimodal Agentic Deep Search VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c51fda0a-6a8b-47ae-ae27-03d8243df2b8 · inbound
BVS: Bayesian Visual Search with Multimodal Large Language Model for Fine-grained Perception VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning
Reference 163
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb841e23-0e1c-448e-9519-cdee7432d478 · inbound
Enhancing Large Multimodal Models in Key Information Extraction via Scene-Aware Document Synthesis VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b77f11b1-eb43-4aaa-a7ba-0da6aa9f7156 · inbound
HierDoc: Hierarchical Page-to-Region Evidence Routing for Long-Document Visual Question Answering VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07c32dab-92fe-471c-a56d-9a84f9aad946 · inbound
Video-DeepResearch: Towards the Next-Generation Multimodal Deepresearch Agent VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.