Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 39 inbound Pith citation observations for arXiv:2503.03321.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:45:28.622401Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T03:19:31.045973Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 26f19368-69be-4297-ad00-dc9f67a84f3a · inbound
Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6765169-2bf6-441f-932a-d0bc6ad9374f · inbound
Not All Tokens and Heads Are Equally Important: Dual-Level Attention Intervention for Hallucination Mitigation See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65ddff21-ce30-45c1-9483-4fd7cbd0887b · inbound
Examining Vision Language Models through Multi-dimensional Experiments with Vision and Text Features See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation faa9854e-6bc5-4828-aed1-bb21b25fe44d · inbound
HiDe: Rethinking The Zoom-IN method in High Resolution MLLMs via Hierarchical Decoupling See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6e7dfff8-e5c2-4649-a669-34d764219291 · inbound
HiDe: Rethinking The Zoom-IN method in High Resolution MLLMs via Hierarchical Decoupling See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3102a7f9-dd18-4a0b-b788-6034d468655c · inbound
Attention Misses Visual Risk: Risk-Adaptive Steering for Multimodal Safety Alignment See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 863d9f65-ccbb-4e1e-9ae2-fd0b085c9787 · inbound
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 465b55ba-ba7c-4563-a593-04e432d62705 · inbound
MVI-Bench: A Comprehensive Benchmark for Evaluating Robustness to Misleading Visual Inputs in LVLMs See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d1d25e1a-0373-4e3b-a027-a3115dc80d91 · inbound
MVI-Bench: A Comprehensive Benchmark for Evaluating Robustness to Misleading Visual Inputs in LVLMs See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36680381-e702-4798-bef3-21fc87835478 · inbound
Can Vision-Language Models Count? A Synthetic Benchmark and Analysis of Attention-Based Interventions See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 837504ba-23fd-4c0a-827d-e9080d9327e1 · inbound
EAGLE: Expert-Augmented Attention Guidance for Tuning-Free Industrial Anomaly Detection in Multimodal Large Language Models See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 019e877f-d788-46fd-a323-8e61986d8f3c · inbound
Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 34e3fb74-d1dc-4b4e-9f29-d0e0bfefbb2d · inbound
Counting to Four is still a Chore for VLMs See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f38698a3-d098-4a92-b8b8-cbf920c25201 · inbound
The Cost of Language: Centroid Erasure Exposes and Exploits Modal Competition in Multimodal Language Models See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2c89199c-f754-48a0-8a7c-b8b831309cbb · inbound
Sink-Token-Aware Pruning for Fine-Grained Video Understanding in Efficient Video LLMs See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 96e39971-fcc6-44f3-9e74-9d7d8f9df146 · inbound
Latent Denoising Improves Visual Alignment in Large Multimodal Models See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 15b448f1-d203-4a4e-871d-8f668ec1eb7d · inbound
Combating Visual Neglect and Semantic Drift in Large Multimodal Models for Enhanced Cross-Modal Retrieval See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4566a0c8-e827-4395-bf5d-cb8ec66cd7a8 · inbound
Large Vision-Language Models Get Lost in Attention See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a3d5c165-df40-40d4-9ac9-b47f11a0ca60 · inbound
RAVE: Re-Allocating Visual Attention in Large Multimodal Models See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8f45ea5d-edad-4a0f-9064-b4ea98aaa6fe · inbound
RAVE: Re-Allocating Visual Attention in Large Multimodal Models See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 145b7c44-0cda-4d95-90c0-ff50baaa55d8 · inbound
MLLMs Know When Before Speaking: Revealing and Recovering Temporal Grounding via Attention Cues See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 941f5569-b466-421a-8c5b-ffb153b9be89 · inbound
Inference Time Optimization with Confidence Dynamics See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f7cf9279-2f32-40c9-be26-fe63dbc86178 · inbound
Addressing Exacerbated Attention Sink for Source-Free Cross-Domain Few-Shot Learning See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a41cd955-68d3-41d9-a696-11c4074591a9 · inbound
When Graph Tokens Sink: A Mechanistic Analysis of Graph Language Models See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b37d17ce-7aef-4aaa-8d64-58ba4f066f4e · inbound
Mechanistic Insights into Functional Sparsity in Multimodal LLMs via CoRe Heads See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b039fe84-d7ec-463f-a462-d74ffb2c4031 · inbound
Reason Twice: Segmentation via Candidate Discovery and Comparative Reasoning See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation afb174be-9df7-440c-af56-d881b879226d · inbound
From Senses to Decisions: The Information Flow of Auditory and Visual Perception in Multimodal LLMs See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f663055d-df4d-4e1e-b486-d9cf7ee7eb4a · inbound
Reroute, Don't Remove: Recoverable Visual Token Routing for Vision-Language Models See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d90ea871-600b-40cf-900d-1c1f81e05318 · inbound
Last But Not Least: Boundary Attention CalibratiON for Multimodal KV Cache Compression See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation dbfcec07-4c4c-43c2-92b9-4538d13d9812 · inbound
The Hidden Evolution of Disguised Visual Context inside the VLM See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5082a377-8813-4fe3-ac05-7b19dd3bb3f9 · inbound
VisReflect: Latent Visual Reflection for Fine-Grained Perception in Long Visual Context See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9d60cd50-a443-4612-955b-80f85d85c4cc · inbound
ADAPT: Attention Dynamics Alignment with Preference Tuning for Faithful MLLMs See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 669815dd-0801-42ec-8b36-de554e29f9db · inbound
Information-Regularized Attention for Visual-Centric Reasoning See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6ff1c143-7362-4272-bccc-416129289b67 · inbound
The Ebb and Flow of Multimodal Focus: Scheduling Visual Relay Windows for Grounded VLM Reasoning See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2cba484-7271-45db-ada6-1c7465442908 · inbound
The Ebb and Flow of Multimodal Focus: Scheduling Visual Relay Windows for Grounded VLM Reasoning See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dfaf8944-d82a-4aba-bcd1-997dde3b7c7d · inbound
DrawingVQA: A Real-World Benchmark for Multi-Depth Visual-Textual Reasoning on Construction Drawings See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9d423ba-0d22-4dc5-b4a2-33fed6b76e9e · inbound
ST-Veto: Spatio-Temporal Token Veto for Diffusion MLLMs via Taylor Prediction and Visual Grounding See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5b66a19-8b7a-43b8-b13f-0909d0cc1a4a · inbound
Text Template Tokens Are Implicit Semantic Registers in Diffusion Transformers See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd3f7e0f-08b8-4be6-bc56-4994d17e9b8b · inbound
Not All Redundant Tokens Are Alike: Analyzing Visual Token Pruning through Token Roles See What You Are Told: Visual Attention Sink in Large Multimodal Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.