Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 22 inbound Pith citation observations for arXiv:2401.02330.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-12T19:14:24.357070Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T06:39:37.481871Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 238f4ae0-43af-44c0-ac1d-6f6f576eac31 · inbound
MobileVLM V2: Faster and Stronger Baseline for Vision Language Model LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation efa3fa37-7825-49ae-ae74-70459f1c9d1b · inbound
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model
Reference 137
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 1c655613-c311-4e0f-9bf6-b1980c9068c9 · inbound
MM1: Methods, Analysis & Insights from Multimodal LLM Pre-training LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model
Reference 135
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e83d3dbb-328f-4248-b67d-2a7d1d376388 · inbound
SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 01874c40-64b1-4a60-ad83-df77265fd892 · inbound
TinyVLA: Towards Fast, Data-Efficient Vision-Language-Action Models for Robotic Manipulation LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 064ebc94-c7b9-47c6-956c-da541c1e6655 · inbound
Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation bfa17e6d-eb0d-460e-9494-c54ee7549810 · inbound
Learn from Downstream and Be Yourself in Multimodal Large Language Model Fine-Tuning LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model
Reference 111
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 350fbf87-738c-4120-bf83-c1e3e53f517d · inbound
Generalist Virtual Agents: A Survey on Autonomous Agents Across Digital Platforms LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8433939e-def4-4455-b368-f57453b87aea · inbound
Understanding Museum Exhibits using Vision-Language Reasoning LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model
Reference 95
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01a51fb4-2b19-467e-95e9-18ad1872f3f2 · inbound
Olympus: A Universal Task Router for Computer Vision Tasks LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model
Reference 100
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf603baa-01b3-46dd-9ca7-8a40f53975ee · inbound
AlzheimerRAG: Multimodal Retrieval Augmented Generation for Clinical Use Cases using PubMed articles LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 066c069c-f4d9-4b05-9906-c9891d41afa8 · inbound
WalkVLM:Aid Visually Impaired People Walking by Vision Language Model LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b47da43e-68bb-4219-90ff-0a9e8e9856fe · inbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 069d95ed-154c-49f9-8fb8-b65a17131abb · inbound
LLMQuoter: Enhancing RAG Capabilities Through Efficient Quote Extraction From Large Contexts LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee3c0e73-0e11-4f6a-9d7f-ae60a25d8028 · inbound
Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 22b36fa5-7f81-4cef-a74a-5cb88c287ac6 · inbound
UniCMs: A Unified Consistency Model For Efficient Multimodal Generation and Understanding LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c94ca341-21c4-4a6b-970f-82d1e6a2e98f · inbound
Mogao: An Omni Foundation Model for Interleaved Multi-Modal Generation LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model
Reference 99
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 62bdc963-0009-484b-8998-f6b6db471a92 · inbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3d9b7ca-2369-4474-9e1e-20e63ef9d394 · inbound
Boosting Embodied AI Agents through Perception-Generation Disaggregation and Asynchronous Pipeline Execution LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f37af42-41e5-4448-9e3e-cb67becaef07 · inbound
From Plausibility to Verifiability: Risk-Controlled Generative OCR with Vision-Language Models LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e441337d-b762-42cb-9e2c-65302fb3803d · inbound
HPP: Hierarchical Programmatic Probing for Long Video Understanding by Decoupling Perception and Reasoning LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model
Reference 285
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 1f0fdd25-f117-41c8-8184-f0f80beee9a8 · inbound
SCOPE and SCION: A Benchmark and an Auditable Reference Pipeline for Schema Induction and Fusion from Text LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model
Reference 165
Source-reported events for the cited work
Unavailable: canonical work link unavailable.