Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 26 inbound Pith citation observations for arXiv:2407.11691.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T21:49:27.762903Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
3
pith, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 9b2b6afd-e249-426a-9884-5d899c6e2245 · inbound
Number it: Temporal Grounding Videos like Flipping Manga VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bcafcba6-8f48-41d5-9d96-1e189e1fea68 · inbound
MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models
Reference 212
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5714edc0-0fb7-401a-aa5b-ba5e6dfac8b5 · inbound
VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 988de421-57cb-44d4-a950-f2f4a2a42171 · inbound
Advancing Myopia To Holism: Fully Contrastive Language-Image Pre-training VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4fada578-696f-4cd4-88b8-77005655e502 · inbound
CompCap: Improving Multimodal Large Language Models with Composite Captions VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a8af7fbc-f283-488f-937e-0cb0f76f67e7 · inbound
POINTS1.5: Building a Vision-Language Model towards Real World Applications VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 302ada15-5c9f-4496-a6b7-5971401e8e94 · inbound
Toward Robust Hyper-Detailed Image Captioning: A Multiagent Approach and Dual Evaluation Metrics for Factuality and Coverage VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61be594d-ee87-474d-8a2f-6031032e4763 · inbound
MVTamperBench: Evaluating Robustness of Vision-Language Models VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2638719a-6663-4c36-8752-37b15a7b659b · inbound
Prioritizing Image-Related Tokens Enhances Vision-Language Pre-Training VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d21366e8-78ce-4120-92db-4a02f47b2928 · inbound
Circle-RoPE: Cone-like Decoupled Rotary Positional Embedding for Large Vision-Language Models VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 37d8ce2c-fe48-4922-b4cb-40e1f2b309c9 · inbound
Argus Inspection: Do Multimodal Large Language Models Possess the Eye of Panoptes? VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac4ca1f2-65ff-41c0-a482-281972f3c65e · inbound
Mono-InternVL-1.5: Towards Cheaper and Faster Monolithic Multimodal Large Language Models VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models
Reference 132
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2bc60fe3-88be-415e-803c-f886f360bf8e · inbound
In-context Learning of Vision Language Models for Detection of Physical and Digital Attacks against Face Recognition Systems VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42bc6ae6-ce22-4d6e-9970-e4b71f5ef03c · inbound
Improving Large Vision and Language Models by Learning from a Panel of Peers VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2483dcad-774a-4fce-8d5c-ef8546d96383 · inbound
LaV-CoT: Language-Aware Visual CoT with Multi-Aspect Reward Optimization for Real-World Multilingual VQA VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 51089e0d-fd39-4343-8c03-d24a391be2b8 · inbound
ChartVerse: Scaling Chart Reasoning via Reliable Programmatic Synthesis from Scratch VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation fce6002d-6acf-4e5a-8652-57a1bf04b315 · inbound
S2H-DPO: Hardness-Aware Preference Optimization for Vision-Language Models VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 7b4a8009-1c28-46da-85b4-e173e73760d8 · inbound
GeoSym127K: Scalable Symbolically-verifiable Synthesis for Multimodal Geometric Reasoning VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 9fb821af-4b1f-4227-9b44-71bb0591aa1a · inbound
Beyond Mode Collapse: Distribution Matching for Diverse Reasoning VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 14906927-c213-4dfe-a9f6-9bdad2c30e1f · inbound
DeepInsight: A Unified Evaluation Infrastructure Across the Physical AI Stack VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 0e29ed91-b731-4575-acf7-6014584ac405 · inbound
LLM-Based Examination of Eligibility Criteria from Securities Prospectuses at the German Central Bank VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 2caa173d-f753-475c-a264-790ef9896e3c · inbound
Causal Connections: Leveraging Multilingual Fine-Tuning for Financial QA@FinCausal 2026 VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 76c71cff-bfa0-4cdb-bdd5-16f4e584dbb8 · inbound
MotionAtlas: Detailed Region Captioning for Motion-Centric Videos VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation ce412b0f-15e9-4e5c-b6d0-536b2bf5d278 · inbound
Nemotron-Labs-Diffusion: A Tri-Mode Language Model Unifying Autoregressive, Diffusion, and Self-Speculation Decoding VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 15cb102d-67ec-4f8c-9f8c-69894e76adbb · inbound
SLAPBench: Benchmarking Multimodal Large Language Models for Four-Finger SLAP Fingerprint Verification VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75b98791-32eb-4264-a0f5-e0a1d862d3ca · inbound
Deferred Exposure of Future Trajectories for Verifiable Reasoning in Autonomous Driving VLMs VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.