Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T22:12:46.203510Z
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 6 inbound Pith citation observations for arXiv:2501.02699.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T22:12:46.203510Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:39:46.817981Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-11T12:33:33.048939Z
47 of 47 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c7b9e58d-426c-4226-af10-28feab2a975b · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a343b1a-2c71-4d75-9b27-31a539697c31 · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models From colouring-in to pointillism: revisiting semantic segmentation supervision,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a7042ddb-c0cd-432d-84ac-f696ee083c86 · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models Lan- guage models are few-shot learners
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation d3008bc1-1a52-4ac6-bb86-461bc590c0b5 · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models A simple framework for contrastive learning of visual representations
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 05bad6a0-b0f8-44e1-92d0-4b8b5337f409 · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models Pali: Scaling language-image learning in 100+ languages
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 295dfee2-7742-4d67-bf2c-1a7c9b24c125 · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models Reproducible scal- ing laws for contrastive language-image learning
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation bd63c6de-4d5c-4948-b8bd-8fc567f0c520 · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models Gonzalez, Ion Stoica, and Eric P
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 1b7ca14f-55bc-4981-9cdc-d76196d4bf43 · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models Scaling instruction- finetuned language models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 3ea23f45-5b55-4806-a84b-674e585aac84 · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models Instructblip: Towards general- purpose vision-language models with instruction tuning
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 0b528abc-5b6d-470e-bfd7-fdf62a34c9b4 · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3a372a6-26b2-42c9-be24-1ffdb16a0a21 · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models InternLM-XComposer2: Mastering Free-form Text-Image Composition and Comprehension in Vision-Language Large Model
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42ec67c8-c742-411e-9e4b-e5bce5ad1869 · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 480426e8-a53b-436f-8222-3f5ac6131c2e · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models A unified continual learn- ing framework with general parameter-efficient tuning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 17fd81d1-ee78-43f7-a47d-00aef5545268 · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models Bootstrap your own latent-a new approach to self-supervised learning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 6fe60b39-e4c5-445c-9b7d-12051a3fbf12 · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models Hal- lusionbench: An advanced diagnostic suite for entangled language hallucination and visual illusion in large vision- language models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation bb3d1da8-92c7-442b-aa18-140b45367798 · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models Dimension- ality reduction by learning an invariant mapping
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 300379b3-abc7-4e0a-bc4e-018c85f0604a · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models Momentum contrast for unsupervised visual rep- resentation learning
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation fe3f92e1-3b6e-4386-b31a-43b7c5d11564 · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models Lora: Low-rank adaptation of large language models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 06141fea-6bb7-4ecc-a2b3-a91be3ae543f · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models BRA VE: Broadening the visual encoding of vision-language models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 5a82ff62-5666-49b7-a6b4-b69f49c4fde7 · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 5475cd02-fe9e-41d1-a807-c216b85c3f2d · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models Evaluating object hallucination in large vision-language models
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 9242b92d-a487-4efe-9c13-01176c3c2c61 · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models Microsoft coco: Common objects in context
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ecc45de-0f82-4f83-b926-50dd3a274eae · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models Improved baselines with visual instruction tuning, 2023
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation fdfe1b4f-0fff-4b23-93c4-10d09b7fdad3 · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models Visual instruction tuning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 2df76d0f-031b-485e-93e0-ae15b865f0ee · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models RoBERTa: A Robustly Optimized BERT Pretraining Approach
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63674953-6d78-4576-887f-d3fe1d4463bf · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models Gpt-4 technical report, 2023
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a788af3f-93ba-4716-844f-0c764e310eea · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models DINOv2: Learning Robust Visual Features without Supervision
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09d6a444-330f-4814-a7b1-7066a29cd35c · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models Grounding multimodal large language models to the world
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 32d80f44-e30b-4224-9dce-ec26c648ec6e · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models Learning transferable visual models from natural language supervision
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 84eeb2b9-2e1b-4566-a8b2-b566d629f2c2 · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models Exploring the limits of transfer learning with a unified text-to-text transformer
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation f04221c6-25e9-404c-a1fa-84982be34d33 · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models xgen-mm-phi3-mini-instruct model card, 2024
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 5af1e5ca-f966-4c73-a321-995674b0c369 · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models A-okvqa: A benchmark for visual question answering using world knowl- edge, 2022
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 635290f4-9207-4838-9dbf-00fc11950316 · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models Eva-clip: Improved training techniques for clip at scale,
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ef7d0e3-9af0-47fa-aeee-ede1cf3e9863 · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models Eyes wide shut? exploring the vi- sual shortcomings of multimodal llms
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b95bbb3a-0605-47f3-9421-11dd07e27506 · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96c09293-1ff0-4048-9783-5fa2c93009f0 · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models Pivot: Prompting for video con- tinual learning
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b1828c33-1c9c-4cc2-b7f8-29247a18e67e · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models Behind the magic, merlim: Multi- modal evaluation benchmark for large image-language mod- els, 2024
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b63a8b17-9a5d-4957-b9e6-21f37885b11d · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models CorDA: Context-Oriented Decomposition Adaptation of Large Language Models for Task-Aware Parameter-Efficient Fine-tuning
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 701f9a50-6733-426d-8c1a-a1422a1dcc10 · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models Sigmoid loss for language image pre-training
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation f5d1b0f0-9725-4153-b827-08613a188c29 · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models Galore: Memory- efficient llm training by gradient low-rank projection, 2024
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 8a0cf7ee-cb42-4269-8d08-5d482c3ed60e · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models Analyzing and mitigating object hallucination in large vision-language models
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation f597d41b-bfd7-46b8-a88d-c284693e4134 · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad788c7b-c3b7-46c7-9ac5-2784493283e4 · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models LLA” (LLaV A-1.5), “LLA*
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 9afe98ca-a354-49c4-a965-ac45f47ab924 · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models We compare the zero-shot and linear probing performance of EAGLE-tuned VLMs against the original models
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation d41bdf56-ba47-46dc-9a91-7014699d54b2 · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models Figure 3 presents 3 additional scenarios when EAGLE effectively reduces the hallucinations of the IT-VLMs
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 2d98a4e1-12c9-4405-a4ea-9d474b4d1b83 · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models The same hyperparameters are applied to both VLMs, EV A01-CLIP- g-14 and OpenAI CLIP-L-14-336
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b7eceb0f-73eb-4269-b15b-d19ff7b11c95 · outbound
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models Unresolved cited work
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 3f6fa8ed-8e71-4294-bb90-12668f424ada · inbound
Hallucination of Multimodal Large Language Models: A Survey EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models
Reference 159
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 83d1695a-9fc1-4e70-86bc-45c46c44e8d2 · inbound
MoDA: Modulation Adapter for Fine-Grained Visual Grounding in Instructional MLLMs EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e0cb9c0-6a7d-45ad-bf9a-db4528cb194b · inbound
Empowering Multimodal LLMs with External Tools: A Comprehensive Survey EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models
Reference 209
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e52a1569-1195-4c37-a629-46c10e323306 · inbound
Towards Mitigating Hallucinations in Large Vision-Language Models by Refining Textual Embeddings EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65cc5cbc-6c77-4dca-8e7e-da991fdcbdbd · inbound
CURE: Curriculum-guided Multi-task Training for Reliable Anatomy Grounded Report Generation EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e809645-c5ef-4a63-bc2b-21e91d2d3166 · inbound
DICA: Dual-Indicator Guided Contrastive Alignment in Multimodal Large Language Models EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.