Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T17:37:54.891321Z
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 0 inbound Pith citation observations for arXiv:2411.12787.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T17:37:54.891321Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
38 of 38 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d7f1bee5-c5be-4905-b8ca-a43edb0dd259 · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning Flamingo: a visual language model for few-shot learning
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a26616b5-c94c-419a-a46a-4fef259519f5 · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning Layer Normalization
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b986d2a-0096-41de-86b8-ffc475aaa73e · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5afec7ec-c18a-4f90-ae8a-94c76f4ef2bc · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66132ade-bf2d-4a37-b9ba-22066cb3ef22 · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning LLaVA-MoLE: Sparse Mixture of LoRA Experts for Mitigating Data Conflicts in Instruction Finetuning MLLMs
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f5db9c4-0d15-41d3-9a10-32912c279eef · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08921ef2-c25c-45a5-97d2-dde53a6983fd · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning Instructblip: Towards general- purpose vision-language models with instruction tuning,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0da724d6-0a78-47c3-bf3b-d571344a3133 · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning Parameter-efficient fine-tuning of large-scale pre-trained language models.Nature Machine In- telligence, 5(3):220–235, 2023
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6e169af-b8f8-4e5e-bdd5-e05a896e9cc3 · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning MouSi: Poly-Visual-Expert Vision-Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 791b1a21-74a6-4ede-ba5f-e9c4af2a9656 · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning Deep sparse rectifier neural networks
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 95eb8dfe-3c88-4319-944f-85912b3452dd · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning LoRA: Low-Rank Adaptation of Large Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation acacf141-d1a7-492d-b5b9-3c6f8cf9c8e3 · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning Harder tasks need more experts: Dynamic routing in moe models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f0c4c9c-b5a1-4e67-961f-33c073ec933b · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning RoDE: Linear Rectified Mixture of Diverse Experts for Food Large Multi-Modal Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2cae4764-886c-4fa6-92aa-ae68254158f4 · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning Unlocking textual and visual wisdom: Open-vocabulary 3d object detection enhanced by comprehensive guidance from text and image
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 7ad139bc-c6db-4689-ac7b-37cdb056f429 · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning Don't Deceive Me: Mitigating Gaslighting through Attention Reallocation in LMMs
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e8fdd22-c8c3-4c82-b9b2-96d3439bc3a4 · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning Lumen: Unleashing versa- tile vision-centric capabilities of large multimodal models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 7da838a1-9d43-4d58-99a5-22b63f4f3261 · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning Imagenet classification with deep convolutional neural net- works
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7b873ba-d7bf-4ae4-b205-0edc508b4c29 · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eccca891-7eb6-4474-bd9e-d6cc548c56e1 · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning Rouge: A package for automatic evaluation of summaries
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95104aa3-f445-495f-93da-ecc27bb27c04 · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning Visual instruction tuning, 2023
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 480ba212-1816-4dbd-9758-2dd3d696f467 · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning Improved baselines with visual instruction tuning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 9823dec1-cb94-42d7-a37d-f41e97f3ed15 · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f5cf207-ebb8-4ff4-a6c4-54c25a6b83e6 · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning AdaMoLE: Fine-Tuning Large Language Models with Adaptive Mixture of Low-Rank Adaptation Experts
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2bf07059-d448-4a95-9510-58703b1ff33f · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning DeepSeek-VL: Towards Real-World Vision-Language Understanding
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d93f0c8e-19ec-4a4a-b20d-2101feb765f4 · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning DINOv2: Learning Robust Visual Features without Supervision
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16286e8f-2876-4b89-9ad0-6c6ad8ab5879 · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning Flickr30k entities: Collecting region-to-phrase corre- spondences for richer image-to-sentence models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 4167a38d-3880-40f4-a5ef-164b61f5a131 · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning A Call for Clarity in Reporting BLEU Scores
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61538d55-5b18-46f5-b59e-c707a540807b · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning Learning transferable visual models from natural language supervi- sion
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 780a05aa-8369-4659-8f5a-6edb99520912 · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning Scienceqa: A novel resource for question answering on scholarly articles
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ecd51877-14f4-4d05-b97d-ea703d8943f3 · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 031c52a2-023b-4533-841e-c6bbe7f3c87c · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning Eyes wide shut? exploring the visual shortcomings of multimodal llms
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 264b8c67-1619-472e-9c4c-6c98e158f314 · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning LLaMA: Open and Efficient Foundation Language Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a8cebf6-8ead-4290-9bfd-00fc64a31cf5 · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning Mixture of LoRA Experts
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9663fe63-948c-4bfd-9f35-dc7e1365bfa1 · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning Vision transformer with deformable attention
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 6aee29cb-5595-4836-be2c-d468dbc7c14c · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning LLaVA-UHD: an LMM Perceiving Any Aspect Ratio and High-Resolution Images
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c651cc23-de9b-469f-b8a4-48eb6bd89114 · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning FoodLMM: A Versatile Food Assistant using Large Multi-modal Model
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bce9786f-9c3d-4cdd-8eeb-1345393caf8c · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b787204d-0eae-4bab-a0a2-258eadc739e8 · outbound
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning Deformable DETR: Deformable Transformers for End-to-End Object Detection
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.