Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2410.08202.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:21:05.372302Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-16T09:16:17.330970Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 9f39929b-a390-4545-91ca-7d44a6c219a5 · inbound
Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization Mono-InternVL: Pushing the Boundaries of Monolithic Multimodal Large Language Models with Endogenous Visual Pre-training
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ea889e35-28b7-4e61-9495-e975b419072e · inbound
Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling Mono-InternVL: Pushing the Boundaries of Monolithic Multimodal Large Language Models with Endogenous Visual Pre-training
Reference 172
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 13d25a94-f9e4-46f9-ac01-cb6142da07aa · inbound
TimeCausality: Evaluating the Causal Ability in Time Dimension for Vision Language Models Mono-InternVL: Pushing the Boundaries of Monolithic Multimodal Large Language Models with Endogenous Visual Pre-training
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ef13d95-c7ca-4dad-b000-c89cc7399d84 · inbound
Visual Embodied Brain: Let Multimodal Large Language Models See, Think, and Control in Spaces Mono-InternVL: Pushing the Boundaries of Monolithic Multimodal Large Language Models with Endogenous Visual Pre-training
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c10ab9b-7207-44a0-b710-b1ec91a7e84f · inbound
GOBench: Benchmarking Geometric Optics Generation and Understanding of MLLMs Mono-InternVL: Pushing the Boundaries of Monolithic Multimodal Large Language Models with Endogenous Visual Pre-training
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e05ec7b-f3b2-4b73-9f76-0d1c6efa92ff · inbound
SMAR: Soft Modality-Aware Routing Strategy for MoE-based Multimodal Large Language Models Preserving Language Capabilities Mono-InternVL: Pushing the Boundaries of Monolithic Multimodal Large Language Models with Endogenous Visual Pre-training
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6c1c945-7804-4a3a-abcf-16ce18d90eec · inbound
Breaking Bad Molecules: Are MLLMs Ready for Structure-Level Molecular Detoxification? Mono-InternVL: Pushing the Boundaries of Monolithic Multimodal Large Language Models with Endogenous Visual Pre-training
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a59232dd-a6e0-4986-affb-a787d45ad1d7 · inbound
LaVi: Efficient Large Vision-Language Models via Internal Feature Modulation Mono-InternVL: Pushing the Boundaries of Monolithic Multimodal Large Language Models with Endogenous Visual Pre-training
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0432910-0c1c-432f-b169-d2edd481baa3 · inbound
Language-Unlocked ViT (LUViT): Empowering Self-Supervised Vision Transformers with LLMs Mono-InternVL: Pushing the Boundaries of Monolithic Multimodal Large Language Models with Endogenous Visual Pre-training
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf96d5d1-f44a-4e86-b065-e6a114834a41 · inbound
MagicVL-2B: Empowering Vision-Language Models on Mobile Devices with Lightweight Visual Encoders via Curriculum Learning Mono-InternVL: Pushing the Boundaries of Monolithic Multimodal Large Language Models with Endogenous Visual Pre-training
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd768d52-3dee-44e1-a7bf-da37ea51c617 · inbound
InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency Mono-InternVL: Pushing the Boundaries of Monolithic Multimodal Large Language Models with Endogenous Visual Pre-training
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.