Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2311.06607.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T19:58:19.334924Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T06:39:37.479106Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 0cb96e4f-8e4b-4928-83dd-191236d49a52 · inbound
A Survey on Multimodal Large Language Models Monkey: Image Resolution and Text Label Are Important Things for Large Multi-modal Models
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 77798ffe-5e13-4a9e-9b18-b33c9ab5be2a · inbound
MMBench: Is Your Multi-modal Model an All-around Player? Monkey: Image Resolution and Text Label Are Important Things for Large Multi-modal Models
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2dc1f223-840d-47b3-83b6-30d50bd8bf64 · inbound
InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks Monkey: Image Resolution and Text Label Are Important Things for Large Multi-modal Models
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7194c1c2-9fc8-4a1f-91ef-9ce83b2e3d93 · inbound
InternLM-XComposer2: Mastering Free-form Text-Image Composition and Comprehension in Vision-Language Large Model Monkey: Image Resolution and Text Label Are Important Things for Large Multi-modal Models
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d49e1eb5-1728-405e-ba7e-138d7f79f98a · inbound
A Survey on Hallucination in Large Vision-Language Models Monkey: Image Resolution and Text Label Are Important Things for Large Multi-modal Models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6b69ddb0-f51d-49e8-ad2a-daef9811f7b1 · inbound
MM1: Methods, Analysis & Insights from Multimodal LLM Pre-training Monkey: Image Resolution and Text Label Are Important Things for Large Multi-modal Models
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 245f14d7-b4f0-49f3-ad45-72fc13368c17 · inbound
Are We on the Right Way for Evaluating Large Vision-Language Models? Monkey: Image Resolution and Text Label Are Important Things for Large Multi-modal Models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1ed97519-ff46-4555-8a73-4419df871bfe · inbound
How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites Monkey: Image Resolution and Text Label Are Important Things for Large Multi-modal Models
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 013f2f03-743b-49cf-bbf8-339868f8f464 · inbound
InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output Monkey: Image Resolution and Text Label Are Important Things for Large Multi-modal Models
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e2c0c614-df4b-4e47-8fa7-185b0efbc18e · inbound
MME-RealWorld: Could Your Multimodal LLM Challenge High-Resolution Real-World Scenarios that are Difficult for Humans? Monkey: Image Resolution and Text Label Are Important Things for Large Multi-modal Models
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9d3a7fc1-83a7-4d6c-9880-70a9f42332fa · inbound
PDF-WuKong: A Large Multimodal Model for Efficient Long PDF Reading with End-to-End Sparse Sampling Monkey: Image Resolution and Text Label Are Important Things for Large Multi-modal Models
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 27a21b8e-3b2d-4ad6-8b6b-c8b3cfe7171e · inbound
Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling Monkey: Image Resolution and Text Label Are Important Things for Large Multi-modal Models
Reference 140
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ea2a5ef6-e5f0-4db7-bc73-a67dbf644f8d · inbound
FLARE: Fully Integration of Vision-Language Representations for Deep Cross-Modal Understanding Monkey: Image Resolution and Text Label Are Important Things for Large Multi-modal Models
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 3984b339-064a-43db-9912-503af55cc6e9 · inbound
InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models Monkey: Image Resolution and Text Label Are Important Things for Large Multi-modal Models
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 88447597-4568-469d-a634-9e20ed3c5b47 · inbound
Visual Funnel: Resolving Contextual Blindness in Multimodal Large Language Models Monkey: Image Resolution and Text Label Are Important Things for Large Multi-modal Models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 57eecde9-6e1e-4448-9fe3-a94055097536 · inbound
ReMoT: Reinforcement Learning with Motion Contrast Triplets Monkey: Image Resolution and Text Label Are Important Things for Large Multi-modal Models
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e15e1f92-b278-4d04-89e0-b91bb83ed2e0 · inbound
Qwen-RobotWorld Technical Report: Unifying Embodied World Modeling through Language-Conditioned Video Generation Monkey: Image Resolution and Text Label Are Important Things for Large Multi-modal Models
Reference 166
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0ed651f1-a7c3-44ae-8b8a-c60a789c91ac · inbound
HPP: Hierarchical Programmatic Probing for Long Video Understanding by Decoupling Perception and Reasoning Monkey: Image Resolution and Text Label Are Important Things for Large Multi-modal Models
Reference 279
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a6cd8e01-3f80-4ee3-86ce-87e3715168a5 · inbound
ESC: Emotional Self-Correction for Reliable Vision-Language Models Monkey: Image Resolution and Text Label Are Important Things for Large Multi-modal Models
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f03fc67c-d05c-4124-ba0c-b89f0ee94c29 · inbound
Qwen-Audio-VAE Technical Report Monkey: Image Resolution and Text Label Are Important Things for Large Multi-modal Models
Reference 179
Source-reported events for the cited work
Unavailable: canonical work link unavailable.