Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 48 inbound Pith citation observations for arXiv:2308.00692.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-12T21:29:35.371817Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-08T02:44:27.641390Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation e7bc3476-42d1-48eb-a25f-60a498b89f61 · inbound
A Survey on Multimodal Large Language Models LISA: Reasoning Segmentation via Large Language Model
Reference 144
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation e800d743-7f54-4c14-9bca-fbd2d8db1d48 · inbound
Improved Baselines with Visual Instruction Tuning LISA: Reasoning Segmentation via Large Language Model
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 440c846a-a9c6-47c4-a20a-2a8032f7999a · inbound
InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks LISA: Reasoning Segmentation via Large Language Model
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 74b58cc9-8d61-4a47-9fe9-799368bf3e75 · inbound
Q-Align: Teaching LMMs for Visual Scoring via Discrete Text-Defined Levels LISA: Reasoning Segmentation via Large Language Model
Reference 220
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 99565bf8-9a88-4cd7-a99c-c54b89af5064 · inbound
MoE-LLaVA: Mixture of Experts for Large Vision-Language Models LISA: Reasoning Segmentation via Large Language Model
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f2f92ab5-4202-4cf0-a0bd-328334a5928e · inbound
MobileVLM V2: Faster and Stronger Baseline for Vision Language Model LISA: Reasoning Segmentation via Large Language Model
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 3b836275-5181-4428-b431-2b6e90226b10 · inbound
MM1: Methods, Analysis & Insights from Multimodal LLM Pre-training LISA: Reasoning Segmentation via Large Language Model
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation e25276ac-475d-41b5-9f0b-b2043af2d0a1 · inbound
Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models LISA: Reasoning Segmentation via Large Language Model
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 90548e61-3ef4-40fe-99a3-20453b93b910 · inbound
How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites LISA: Reasoning Segmentation via Large Language Model
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d9973341-a54d-4963-89ee-fdaeb533ec6d · inbound
Hallucination of Multimodal Large Language Models: A Survey LISA: Reasoning Segmentation via Large Language Model
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 16432e8c-11ae-4110-8cbf-15c8e11bc4ab · inbound
Retrieval Augmented Recipe Generation LISA: Reasoning Segmentation via Large Language Model
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57d4be7d-a371-4c96-8e2e-ccfd64ff52f3 · inbound
Motion-Grounded Video Reasoning: Understanding and Perceiving Motion at Pixel Level LISA: Reasoning Segmentation via Large Language Model
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4001b619-8fb9-4b15-9588-efb73c8522e5 · inbound
Instruction-Guided Editing Controls for Images and Multimedia: A Survey in LLM era LISA: Reasoning Segmentation via Large Language Model
Reference 140
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f37db29-59a6-42d9-bc99-caed9a1506e3 · inbound
InsightEdit: Towards Better Instruction Following for Image Editing LISA: Reasoning Segmentation via Large Language Model
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e9ac932-b0fd-48a0-875c-929d4c08e1b0 · inbound
HyperSeg: Towards Universal Visual Segmentation with Large Language Model LISA: Reasoning Segmentation via Large Language Model
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation abd9a563-b0ef-4a01-992e-f118bbf77b27 · inbound
ChatRex: Taming Multimodal LLM for Joint Perception and Understanding LISA: Reasoning Segmentation via Large Language Model
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c081d1cf-2291-40a6-add3-b52aaad0501a · inbound
Agri-LLaVA: Knowledge-Infused Large Multimodal Assistant on Agricultural Pests and Diseases LISA: Reasoning Segmentation via Large Language Model
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 948e5ad0-0b29-4ebc-b27c-8d4548f9c9af · inbound
EditScout: Locating Forged Regions from Diffusion-based Edited Images with Multimodal LLM LISA: Reasoning Segmentation via Large Language Model
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c37f5ad3-671c-4868-afa1-58cd9c28b6cf · inbound
AIpparel: A Multimodal Foundation Model for Digital Garments LISA: Reasoning Segmentation via Large Language Model
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c35744ab-b3ba-4957-b82b-3cb9040a231d · inbound
InstructSeg: Unifying Instructed Visual Segmentation with Multi-modal Large Language Models LISA: Reasoning Segmentation via Large Language Model
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab6b5c80-ea47-4e38-ab44-fadae128fc91 · inbound
Are They the Same? Exploring Visual Correspondence Shortcomings of Multimodal LLMs LISA: Reasoning Segmentation via Large Language Model
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1f5dab7-0915-4a14-8676-0d8951c142fe · inbound
GeoPix: Multi-Modal Large Language Model for Pixel-level Image Understanding in Remote Sensing LISA: Reasoning Segmentation via Large Language Model
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57551ff5-db4a-46ff-b974-7b06e4df53af · inbound
Densely Connected Parameter-Efficient Tuning for Referring Image Segmentation LISA: Reasoning Segmentation via Large Language Model
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ba89b2b-c607-46b3-9c58-1a1a3c8095b5 · inbound
Pixel-Level Reasoning Segmentation via Multi-turn Conversations LISA: Reasoning Segmentation via Large Language Model
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69848a13-51bb-4a10-a771-a3e89c5162d9 · inbound
On the robustness of multimodal language model towards distractions LISA: Reasoning Segmentation via Large Language Model
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22f767ef-56a2-4155-a22e-5bbff6497ade · inbound
Visual Embodied Brain: Let Multimodal Large Language Models See, Think, and Control in Spaces LISA: Reasoning Segmentation via Large Language Model
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11113140-16c8-4c23-abf7-a76ebf07b908 · inbound
STORM: Benchmarking Visual Rating of MLLMs with a Comprehensive Ordinal Regression Dataset LISA: Reasoning Segmentation via Large Language Model
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6abe27b-963b-47ef-8e03-5daf892aa2b1 · inbound
VideoMolmo: Spatio-Temporal Grounding Meets Pointing LISA: Reasoning Segmentation via Large Language Model
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e5e4640-7317-478a-b48b-99a5e8a3afb8 · inbound
MedSeg-R: Reasoning Segmentation in Medical Images with Multimodal Large Language Models LISA: Reasoning Segmentation via Large Language Model
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 273f83b8-2f55-4422-b866-b56ff5e11ddb · inbound
ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation LISA: Reasoning Segmentation via Large Language Model
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1982561-f511-4d16-a33c-2e135b68235f · inbound
Inter2Former: Dynamic Hybrid Attention for Efficient High-Precision Interactive LISA: Reasoning Segmentation via Large Language Model
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 096df928-9178-48e5-8a28-324a82b4db17 · inbound
KptLLM++: Towards Generic Keypoint Comprehension with Large Language Model LISA: Reasoning Segmentation via Large Language Model
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 781612d5-9bc4-4af2-812d-dee76c2e90ef · inbound
VisionThink: Smart and Efficient Vision Language Model via Reinforcement Learning LISA: Reasoning Segmentation via Large Language Model
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78686686-890a-4340-87c9-c72b035b5a40 · inbound
DynImg: Key Frames with Visual Prompts are Good Representation for Multi-Modal Video Understanding LISA: Reasoning Segmentation via Large Language Model
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8230abef-3c1f-45fa-a6dd-28c842cf370b · inbound
Advancing Visual Large Language Model for Multi-granular Versatile Perception LISA: Reasoning Segmentation via Large Language Model
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9fca09ee-667d-40b0-9ce7-6108f3bd8d1f · inbound
ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver LISA: Reasoning Segmentation via Large Language Model
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 942099c8-ae51-4134-9c6f-d153f2cce095 · inbound
MINGLE: VLMs for Semantically Complex Region Detection in Urban Scenes LISA: Reasoning Segmentation via Large Language Model
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation e27b91bd-f254-44cf-a1a9-9939fb1c3a6e · inbound
VisCoP: Visual Probing for Video Domain Adaptation of Vision Language Models LISA: Reasoning Segmentation via Large Language Model
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67864beb-28a3-4eae-bb71-be687ea84ac1 · inbound
Molmo2: Open Weights and Data for Vision-Language Models with Video Understanding and Grounding LISA: Reasoning Segmentation via Large Language Model
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 708a6f30-9b32-4808-ba10-c61796725e5d · inbound
Chat-Scene++: Exploiting Context-Rich Object Identification for 3D LLM LISA: Reasoning Segmentation via Large Language Model
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 996455f2-3463-4e25-982e-58266447e76f · inbound
Moondream Segmentation: From Words to Masks LISA: Reasoning Segmentation via Large Language Model
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f26d572c-4106-4450-a5e5-e066623f7380 · inbound
WildDet3D: Scaling Promptable 3D Detection in the Wild LISA: Reasoning Segmentation via Large Language Model
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation c053b920-dbe8-4a68-a714-ab6cf6b669dc · inbound
Vision Harnessing Agent for Open Ad-hoc Segmentation LISA: Reasoning Segmentation via Large Language Model
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 563f29a5-3d2d-4865-9859-216e933c1bf7 · inbound
Through the PRISM: Principle-Aware, Interpretable, and Multi-Scale Evaluation of Visual Designs LISA: Reasoning Segmentation via Large Language Model
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f38ad578-1b2e-4d30-9651-b8bca572e351 · inbound
CAIRN: Cross-Room 3D Scene Understanding with Topology-Aware Large Multimodal Models LISA: Reasoning Segmentation via Large Language Model
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 92a261bf-a0aa-4bd0-af57-212de3b6fb8d · inbound
CAIRN: Cross-Room 3D Scene Understanding with Topology-Aware Large Multimodal Models LISA: Reasoning Segmentation via Large Language Model
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4cdf2020-110e-4e22-9c09-901df2ceb4c8 · inbound
Symbol and Footprint Database for Electronic Components by Agentic Recognition and Generation LISA: Reasoning Segmentation via Large Language Model
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dcdd51bf-6c3a-4e16-a828-df52d40381e8 · inbound
Vision-Language Grounding as Bidirectional Concept Correspondence LISA: Reasoning Segmentation via Large Language Model
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.