Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:43:44.008852Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 1 inbound Pith citation observation for arXiv:2506.07202.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:43:44.008852Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-29T08:36:21.863358Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-06-29T08:43:15.762081Z
42 of 42 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation ca4abf32-e965-4329-8ddb-2d501f7288dd · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 461a6f64-d357-4671-838d-54ffc7894c6f · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation Are We on the Right Way for Evaluating Large Vision-Language Models?
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff8bb995-86a5-47ab-a9af-6e3f4a808755 · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation Microsoft COCO Captions: Data Collection and Evaluation Server
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7bd01bc-355d-4378-8430-4d52309701ae · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e3b7325-fb3c-403b-9538-584a770ebc6b · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation Le, Sergey Levine, and Yi Ma
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5dce37d-d9c2-4d0b-888a-cbfc4ac61195 · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation Documenting Large Webtext Corpora: A Case Study on the Colossal Clean Crawled Corpus
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6d79c24-7fe5-41ca-af19-30b895ee88de · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation Complex Video Reasoning and Robustness Evaluation Suite (CVRR-ES)
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 145fe6b8-60a8-4bd3-a933-6aa0a5eaad46 · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation NPHardEval: Dynamic Benchmark on Reasoning Ability of Large Language Models via Complexity Classes
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18484495-cf17-4f71-b15a-4214b86012b5 · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation Video-R1: Reinforcing Video Reasoning in MLLMs
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f864f67-399d-4d94-8d52-7712ddc73784 · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc258f49-a426-4b23-bc29-558dde71b26f · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation SPHINX-X: Scaling Data and Parameters for a Family of Multi-modal Large Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 373f9932-70b9-42e9-8684-9ed355346882 · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation Time travel in LLMs: Tracing data contamination in large language models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6dad6892-dcbe-4f3c-8327-94cc60b97c7b · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation Goodfellow, Jonathon Shlens, and Christian Szegedy
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 03daa9c5-f96a-43e8-944a-775293501894 · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation Flat minima
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 296a0b41-3fd2-40f9-ba56-9bbb5d0eb0b1 · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16b0f9da-ef7e-455e-a53e-c4a85c7e0b23 · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation On Large-Batch Training for Deep Learning: Generalization Gap and Sharp Minima
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d940628d-e01b-4cad-acd2-5bf23370c18f · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation LLaV A-OneVision: Easy visual task transfer
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c7cc41dc-b9de-4638-aa37-b8f895cd19a0 · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56be8d9d-0980-44d9-8a08-da5276953aa3 · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation VideoChat-R1: Enhancing Spatio-Temporal Perception via Reinforcement Fine-Tuning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 954ca561-503e-4eec-9ed3-dd3d1fe530e5 · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation Visual Instruction Tuning
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1214211-d31e-42b3-811c-b63afed3d312 · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation Visual instruction tuning.Advances in neural information processing systems, 36, 2024
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 655f1235-d44d-4cb3-89c3-0231e9e6d2cd · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation On the robustness of multimodal language model towards distractions, 2025
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9bc5acdd-12dc-415d-97c4-5eeadbcac686 · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation Is your video language model a reliable judge? In The Thirteenth International Conference on Learning Representations, 2025
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 786f99bd-9ae8-4afc-a830-939e58122de1 · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation MMBench: Is Your Multi-modal Model an All-around Player?
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3002a829-5539-413f-8145-f53522016021 · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation The Llama 3 herd of models
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 719807d8-f564-4bf8-84ec-85ffd002dfec · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation Ok-vqa: A visual question answering benchmark requiring external knowledge
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c71a775-0782-4af8-98a8-9bb109f52d3c · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5588a6fb-1bff-444c-affa-7cb2e6634c9a · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation Introducing GPT-4.1
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e6db3cb7-00f5-467c-a8e7-ea121436aac4 · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation Introducing o3 and o4-mini: Our smartest models yet
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b7293655-e8c3-47b1-beba-d9e66f85cc81 · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation Chatterji, Faisal Ladhak, and Tatsunori Hashimoto
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e33612d5-bdef-4b89-8fd5-1966ba17a24f · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation Qwen2.5-VL Technical Report
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7fdd844e-eb3c-416b-9c9b-7bf7065b460d · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation LLaMA: Open and Efficient Foundation Language Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d741959-b5bc-45b1-a7a8-54dbea542666 · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation VL-Rethinker: Incentivizing Self-Reflection of Vision-Language Models with Reinforcement Learning
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83b791cf-4563-4bac-b69b-e823938961cd · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc628e12-ea92-4ca2-8b7b-a2916806fa5f · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49db2cd3-1be6-43e4-9ca3-4059d3f4df2b · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation Realworldqa
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f5299504-2ffd-4225-86e8-3561e2128509 · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation Dynamic multimodal evaluation with flexible complexity by vision-language bootstrapping
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 170e434f-ad70-4ad2-8b5b-37a9c533797f · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation A Survey on Multimodal Large Language Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e027e542-f702-40f6-9e38-0c9c8015b887 · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5879e056-8321-4c4d-9c80-1ed495a4e8ba · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bcc4f788-8ab4-443c-b4bf-217c5acd20de · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation DyVal: Dynamic Evaluation of Large Language Models for Reasoning Tasks
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 678d5dee-677b-490a-9f62-21926e94ddac · outbound
Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation reasoning MLLMs,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 87a9aca1-b6d3-44d1-a386-294b3273e171 · inbound
DMC-CF: Dynamic Multimodal CounterFactual QA benchmark for Causal Reasoning Reasoning Multimodal Large Language Model: Data Contamination and Dynamic Evaluation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.