Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T17:38:56.881537Z
Paper Citation Record · LEDGER
As of 15 August 2026, this Paper Citation Record lists 27 of 27 outbound references and 0 inbound Pith citation observations for arXiv:2412.08771.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T17:38:56.881537Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
27 of 27 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 3da70c1f-c91b-4a78-ac6f-1045cbf86f16 · outbound
LLaVA-Zip: Adaptive Visual Token Compression with Intrinsic Image Information LLaVA-OneVision: Easy Visual Task Transfer
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a876d2af-f907-4e5f-9bd7-fb52c28479a1 · outbound
LLaVA-Zip: Adaptive Visual Token Compression with Intrinsic Image Information Improved baselines with visual instruction tuning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2f2f848-a4e1-4498-ae03-80ec1bd017a6 · outbound
LLaVA-Zip: Adaptive Visual Token Compression with Intrinsic Image Information Visual instruction tuning
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af58c3fd-2e36-494f-9fd3-029e451b498c · outbound
LLaVA-Zip: Adaptive Visual Token Compression with Intrinsic Image Information Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 20390628-8090-428a-907a-0c1a15a05ccd · outbound
LLaVA-Zip: Adaptive Visual Token Compression with Intrinsic Image Information Less is More: A Simple yet Effective Token Reduction Method for Efficient Multi-modal LLMs
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b17128a6-cf98-48ea-bcac-7682aca48906 · outbound
LLaVA-Zip: Adaptive Visual Token Compression with Intrinsic Image Information Beyond LLaVA-HD: Diving into High-Resolution Large Multimodal Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 082004d0-bfd8-4a5a-a00f-5ae86d088b6f · outbound
LLaVA-Zip: Adaptive Visual Token Compression with Intrinsic Image Information HiRED: Attention-Guided Token Dropping for Efficient Inference of High-Resolution Vision-Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eefffb81-b279-4120-ada8-0c530d5cf11d · outbound
LLaVA-Zip: Adaptive Visual Token Compression with Intrinsic Image Information Matryoshka Multimodal Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b22601a-7576-4731-bf58-c7d046f3fc20 · outbound
LLaVA-Zip: Adaptive Visual Token Compression with Intrinsic Image Information Flamingo: a visual language model for few-shot learning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81522011-0c70-4c50-b4e6-f46a7adcdf3e · outbound
LLaVA-Zip: Adaptive Visual Token Compression with Intrinsic Image Information FocusLLaVA: A Coarse-to-Fine Approach for Efficient and Effective Visual Token Compression
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66b5eed6-4f90-499c-9ac5-849f2171c11c · outbound
LLaVA-Zip: Adaptive Visual Token Compression with Intrinsic Image Information Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51d5bba8-1fb1-493c-b520-d423a3a07391 · outbound
LLaVA-Zip: Adaptive Visual Token Compression with Intrinsic Image Information Learning transferable visual models from natural language supervision
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 886b59ed-d335-4c77-9271-7b21ec3dfb4a · outbound
LLaVA-Zip: Adaptive Visual Token Compression with Intrinsic Image Information Judging llm-as-a-judge with mt-bench and chatbot arena
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0462a536-8991-45ab-a93b-062bd723a3a9 · outbound
LLaVA-Zip: Adaptive Visual Token Compression with Intrinsic Image Information Lmms-eval: Accelerating the development of large multimoal models, March 2024
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ff118de8-15c3-40fb-b0f8-34609c1c92f9 · outbound
LLaVA-Zip: Adaptive Visual Token Compression with Intrinsic Image Information Gqa: A new dataset for real-world visual reasoning and com- positional question answering
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 1574031e-206e-434c-86c5-f0d001fa7d8b · outbound
LLaVA-Zip: Adaptive Visual Token Compression with Intrinsic Image Information Towards vqa models that can read
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 522ece48-3d03-42b5-834c-122ee1b67971 · outbound
LLaVA-Zip: Adaptive Visual Token Compression with Intrinsic Image Information Making the v in vqa matter: Elevating the role of image understanding in visual question answering
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21ed474f-098d-4ec1-aa90-6b8977507d4d · outbound
LLaVA-Zip: Adaptive Visual Token Compression with Intrinsic Image Information Evaluating Object Hallucination in Large Vision-Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9f6b0d1-199e-4e39-9103-f02365561dc9 · outbound
LLaVA-Zip: Adaptive Visual Token Compression with Intrinsic Image Information MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7136a7d0-242b-432d-bf2d-ed56bf59d1e0 · outbound
LLaVA-Zip: Adaptive Visual Token Compression with Intrinsic Image Information MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ba0145d-d6af-4293-9300-b83a1639c68b · outbound
LLaVA-Zip: Adaptive Visual Token Compression with Intrinsic Image Information SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation deac9b71-a821-4eb7-8923-97a186334a82 · outbound
LLaVA-Zip: Adaptive Visual Token Compression with Intrinsic Image Information Microsoft coco: Common objects in context
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63e0699d-71d1-4c86-a7f4-0503c5dbf140 · outbound
LLaVA-Zip: Adaptive Visual Token Compression with Intrinsic Image Information Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32b3dd32-1c17-48d5-b54e-60ce82f730ef · outbound
LLaVA-Zip: Adaptive Visual Token Compression with Intrinsic Image Information Ray: A distributed framework for emerging {AI} applications
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation d516df20-98d7-4ee1-b139-f49f5193a639 · outbound
LLaVA-Zip: Adaptive Visual Token Compression with Intrinsic Image Information Ffcv: Accelerating training by removing data bottlenecks
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 6e4c7e81-8d2f-4afb-a905-1232853e6952 · outbound
LLaVA-Zip: Adaptive Visual Token Compression with Intrinsic Image Information The Llama 3 Herd of Models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 885f5b81-3a33-4fd7-9e6d-36c80160fe4d · outbound
LLaVA-Zip: Adaptive Visual Token Compression with Intrinsic Image Information Video-LLaVA: Learning United Visual Representation by Alignment Before Projection
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.