Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T08:14:37.817377Z
Paper Citation Record · LEDGER
As of 15 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 0 inbound Pith citation observations for arXiv:2510.22067.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T08:14:37.817377Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
31 of 31 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9ae79580-080b-4527-abfb-919b202f836d · outbound
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation Qwen2.5-VL Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8977dd1-3c51-4d54-a3fc-f62988cf15b8 · outbound
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44510d44-ff8f-4659-942a-4ce9cc822da4 · outbound
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation Cracking the Code of Hallucination in LVLMs with Vision-aware Head Divergence
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9d43695-07d0-4e16-ba21-541b64c32ff8 · outbound
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation GPT-4o System Card
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 863d9f65-ccbb-4e1e-9ae2-fd0b085c9787 · outbound
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6cf07e3d-9b3d-4692-b185-ff93c934dce8 · outbound
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation The open images dataset v4: Unified image classification, object detection, and visual relationship detection at scale.In- ternational journal of computer vision, 128(7):1956–1981,
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2238118a-95e6-4ae2-a752-ed872caccf5f · outbound
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7c9ead3-b030-400f-8aee-7eb155274e66 · outbound
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation LLaRA: Supercharging Robot Learning Data for Vision-Language Policy
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2be8d0d-1efc-48f7-90dd-cdf54dd17cbd · outbound
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation Evaluating Object Hallucination in Large Vision-Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8117fa5c-6801-4783-9d2a-43967664a79a · outbound
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation Mitigating Hallucinations in Large Vision-Language Models via Summary-Guided Decoding
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ef232f3-ed38-462a-8175-99bae75ff279 · outbound
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation Object Hallucination in Image Captioning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6455772b-8bc2-4e39-a866-c54cd2bd2bd2 · outbound
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation Aligning Large Multimodal Models with Factually Augmented RLHF
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd86c048-156a-4a5a-8069-44b7abf78f3a · outbound
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation MINT: Mitigating Hallucinations in Large Vision-Language Models via Token Reduction
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f147be7-e61c-4e8b-8103-9c192ac89fc8 · outbound
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation MLLM can see? Dynamic Correction Decoding for Hallucination Mitigation
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1657e0f0-7d95-470a-bb6a-b6c83753092a · outbound
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation Efficient Streaming Language Models with Attention Sinks
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0925bab1-209a-4a2e-baae-b823c54a58ab · outbound
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation TARAC: Mitigating Hallucination in LVLMs via Temporal Attention Real-time Accumulative Connection
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71750f2c-0f0a-4adb-a554-14ff6c96b1d3 · outbound
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation MLLMs Know Where to Look: Training-free Perception of Small Visual Details with Multimodal LLMs
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c38204a-6773-4a42-8f8f-36161a6ecb54 · outbound
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation Tell Your Model Where to Attend: Post-hoc Attention Steering for LLMs
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62835c2e-a1ee-4832-8abc-d6737303a147 · outbound
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation Debiasing Multimodal Large Language Models via Penalization of Language Priors
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25dd4e4e-ab37-4922-b807-c21f269f41e1 · outbound
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation Looking Beyond Text: Reducing Language bias in Large Vision-Language Models via Multimodal Dual-Attention and Soft-Image Guidance
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc44c118-708b-4dd8-bf3a-ebcd0d0a57b6 · outbound
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d8c921f-02f7-49c7-b9f9-fbe8cb314f8b · outbound
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation Is there a frisbee in the image?
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c95a9c96-069b-4f01-9bf5-c8eef3349d5f · outbound
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation Unresolved cited work
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1ff3dfc-3f15-446e-98d6-3eb5779162d9 · outbound
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation Please just answer yes or no
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1010fef4-3055-45a3-9bd6-836145d28101 · outbound
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation Layers for Cross-Modal Fusion Enhancement.Since the benchmarks we consider lack dedi- cated validation sets for hyperparameter tuning, we follow Kang et al
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c248c1f3-1199-4590-9a48-c1ca94ec102e · outbound
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning
Reference 2014
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fec52c2c-783d-44ee-98a5-99ef68d68da4 · outbound
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation Self-Introspective Decoding: Alleviating Hallucinations for Large Vision-Language Models
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb5943d7-59fd-4b9d-b3c0-40c2b3307e46 · outbound
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation Drew A Hudson and Christopher D Manning
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19dfcfd5-bdc8-44fd-a210-9d3f803fef97 · outbound
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation Information Flow Routes: Automatically Interpreting Language Models at Scale
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc905e3d-6cc1-4f57-9794-b725bce5e1e8 · outbound
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3183994d-7ed0-4758-b10b-f3d17259ae19 · outbound
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation PerturboLLaVA: Reducing Multimodal Hallucinations with Perturbative Visual Training
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.