Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T22:10:23.293868Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 7 inbound Pith citation observations for arXiv:2506.22434.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T22:10:23.293868Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-04T02:47:38.676727Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T15:09:54.937856Z
46 of 46 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation a8b68a97-0e72-4a9c-9ad8-c454e6497c4d · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4c6f8d8-18de-4fd4-b427-3fca9ac4deae · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning Qwen2.5-VL Technical Report
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e0fbd76-dce3-4de0-8570-97c3b8910b53 · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning Emerging properties in self-supervised vision transformers
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 24113702-aa5e-43d8-9559-237ebc7b9ba6 · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning SFT or RL? An Early Investigation into Training R1-Like Reasoning Large Vision-Language Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 295ed156-2397-45a4-8b99-322dc2aef68e · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation febea6a5-eb8a-4871-b738-a26a0ff716d5 · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning Are we on the right way for evaluating large vision-language models?NeurIPS,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1ee3d4c4-9c61-42d0-b067-2fad28420e2d · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning A simple framework for contrastive learning of visual representations
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation efe40d4f-71f4-48da-bbb3-9fe78dd576c5 · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 35e6b1cf-aba6-497f-a8ab-26cd02e18cad · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning Vlmevalkit: An open-source toolkit for evaluating large multi-modality models
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f806f609-79e3-4de7-9241-2314288d1859 · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning Blink: Multimodal large language models can see but not perceive
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation df7af8c2-2dd1-4a87-9d21-0f599b0a5c6a · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning Hallusionbench: an advanced diagnostic suite for entangled language hallucination and visual illusion in large vision-language models
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6ab5d93a-f684-4a9e-b584-7c5e5b073eec · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ad90e29-c3eb-4a72-b5e7-8228ec987354 · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning Masked autoencoders are scalable vision learners
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a1702eba-3960-450d-8ef9-9564231be04d · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning Momentum contrast for unsupervised visual representation learning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 30c8b7a1-2087-4abf-9a58-30f83bf48927 · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning GPT-4o System Card
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e64c7b2-2243-4638-a3ce-b61883e05687 · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning OpenAI o1 System Card
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cff9c650-f687-4f86-9f3d-cdb099491a55 · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning Llava-onevision: Easy visual task transfer.TMLR, 2025
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 491e6990-29ec-4934-a82d-bdedd27f0979 · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 694b5750-16b2-49ab-b776-38b59ea66a5d · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning Visual instruction tuning
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dd08eda1-bfeb-49df-bcf4-1b3c9cfd4720 · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning Noisyrollout: Reinforcing visual reasoning with data augmentation.arXiv:2504.13055, 2025
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d8b729d-bc4d-4c58-800b-669a16e0a869 · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning Mmdu: A multi-turn multi-image dialog understanding benchmark and instruction-tuning dataset for lvlms
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6def041d-365c-42d8-a439-367f88f6ffc7 · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning Mathvista: Evaluating mathematical reasoning of foundation models in visual contexts
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c2810585-c011-4c7d-8ccb-bc2427e70979 · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 508e4ff3-66f7-41f2-a4d8-e6703505eadd · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning Mmiu: Multimodal multi-image understanding for evaluating large vision- language models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 41ca8a4b-5b82-4ff3-a3b4-f32a305a2bd7 · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning Skywork R1V: Pioneering Multimodal Reasoning with Chain-of-Thought
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f566777-8f5c-4265-890a-d60523701d0c · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74b28711-97ec-463d-9e80-aa6dcca57847 · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning Seed-thinking-v1
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 429e076a-6e03-4d5e-b0b4-22dd59d2653b · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82cc9a97-fd6c-4e9d-84df-525ab92e2c2c · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning Reason-rft: Reinforcement fine-tuning for visual reasoning.arXiv:2503.20752, 2025
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f858776c-a47b-478c-b409-976b86053609 · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning VidGen-1M: A Large-Scale Dataset for Text-to-video Generation
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5ccabef-a66a-42c6-a438-ff63745e2410 · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning Kimi k1.5: Scaling Reinforcement Learning with LLMs
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d9b5079-b267-42ad-8466-aa1573e257a6 · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning Muirbench: A comprehensive benchmark for robust multi-image understanding
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 65027c1c-7b21-48d4-9e46-8f1b6da2f3e1 · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa107860-b77c-457c-bd26-239bb6829ccf · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 89425d77-f32b-436d-8a79-d4d75a0b76e4 · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning Omniedit: Building image editing generalist models through specialist supervision
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation aa15e761-d7e7-4dfb-bd56-d4c133797abc · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning Chain-of-thought prompting elicits reasoning in large language models
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f311559c-40f0-497b-bd9e-f8c5f4fb16d1 · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning Towards open-ended visual quality comparison
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7a41f29b-b701-4235-b643-88ca56801731 · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning mplug-owl2: Revolutionizing multi-modal large language model with modality collaboration
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 90f9bd71-b043-4599-ace4-89f419051ef8 · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0e367bb4-2cbc-4d97-ac8b-e3719455b251 · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning VLM2-Bench: A Closer Look at How Well VLMs Implicitly Link Explicit Matching Visual Cues
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ed0821a-b305-4cb4-a1fe-aa014113d41e · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning Weaving Context Across Images: Improving Vision-Language Models through Focus-Centric Visual Chains
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ac77033f-768b-4197-b7c0-04d0923444c2 · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning Long Context Transfer from Language to Vision
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6e43932-032a-497b-a74f-48c42b77f00f · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning LLaVA-Video: Video Instruction Tuning With Synthetic Data
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ba2266b-47ea-4432-87f9-5ebfd913950d · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning Benchmarking Multi-Image Understanding in Vision and Language Models: Perception, Knowledge, Reasoning, and Multi-Hop Reasoning
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 626392d1-1c8c-4849-8999-a0fbb297730b · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning Ultraedit: Instruction-based fine-grained image editing at scale
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cab53f23-af9c-4c36-83a5-25d31c6b7cc9 · outbound
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning Are They the Same? Exploring Visual Correspondence Shortcomings of Multimodal LLMs
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56bbfb6a-5856-4629-9ef6-5dd15c4e0698 · inbound
ReMoT: Reinforcement Learning with Motion Contrast Triplets MiCo: Multi-image Contrast for Reinforcement Visual Reasoning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eac79354-ed08-4554-9934-83e1f5b37637 · inbound
CGC: Compositional Grounded Contrast for Fine-Grained Multi-Image Understanding MiCo: Multi-image Contrast for Reinforcement Visual Reasoning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4ec5dc58-95fe-426b-8a51-454a61ac8dc6 · inbound
DataClaw0: Agentic Tailoring Multimodal Data from Raw Streams MiCo: Multi-image Contrast for Reinforcement Visual Reasoning
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ef3e5ca9-6892-4534-8b80-e6d8b0a38be8 · inbound
DataClaw0: Agentic Tailoring Multimodal Data from Raw Streams MiCo: Multi-image Contrast for Reinforcement Visual Reasoning
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0a7c20a-8dbe-416a-9703-83022d97f974 · inbound
From Structure to Synergy: A Survey of Vision-Language Perception Paradigm Evolution in Multimodal Large Language Models MiCo: Multi-image Contrast for Reinforcement Visual Reasoning
Reference 193
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5d3575bf-4f7f-445a-a8df-ca754d0c73fa · inbound
RSICCLLM: A Multimodal Large Language Model for Remote Sensing Image Change Captioning MiCo: Multi-image Contrast for Reinforcement Visual Reasoning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7c81efc5-ace3-4125-b140-de5c81811780 · inbound
MARS-RA: Rank Aggregation for Credit Assignment via Multimodal Comparisons in Embodied Multi-Agent Cooperation MiCo: Multi-image Contrast for Reinforcement Visual Reasoning
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.