Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:47:07.903561Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 7 inbound Pith citation observations for arXiv:2507.01790.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:47:07.903561Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T01:00:26.919995Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-01T17:15:52.016694Z
41 of 41 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 81f33fad-f608-4169-ab62-9fcec184c70b · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? Multimodal biomedical ai
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6d595cf1-56ec-47b9-ba72-ddeadc186849 · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? Understanding intermediate layers using linear classifier probes
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c09864a-642d-4096-b634-7f795a35f720 · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? Flamingo: a visual language model for few-shot learning
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 26f38637-60c2-403e-8948-5ecb129deaba · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? Qwen2.5-VL Technical Report
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04f238e6-90fb-493e-80c5-fe5d5fd89292 · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? Probing classifiers: Promises, shortcomings, and advances
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a8ff95b2-42ca-4a9b-a95a-f968dcc63c28 · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? On the robustness of large multimodal models against image adversarial attacks
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5dd82d92-11e5-4409-87d0-506b45b03509 · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 921bb0d0-f044-467b-b493-b92e6524b6c4 · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? Words or Vision: Do Vision-Language Models Have Blind Faith in Text?
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7efe53c-d129-4982-91bd-484f48087a92 · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? Imagenet: A large-scale hierarchical image database
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b330963-4990-4baa-a8ad-3333f6fcf127 · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? The pascal visual object classes (voc) challenge
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8e1b6d8-03f9-410f-b75b-59f6b38387d5 · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? Pixels versus priors: Controlling knowledge priors in vision-language models through visual counterfacts
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c58d2bf2-5ac0-4aac-9135-1e7583da58b1 · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? What do VLM s NOTICE ? a mechanistic interpretability pipeline for G aussian-noise-free text-image corruption and evaluation
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 219fdcb1-a9dc-46e2-894a-0bfa23f61067 · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? How does GPT -2 compute greater-than?: Interpreting mathematical abilities in a pre-trained language model
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation eacdb6e6-0dc5-4f5a-a378-24aacb8a8d4d · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? Unresolved cited work
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5f5bd9f7-d638-4adb-a967-910e90828056 · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? Adam: A Method for Stochastic Optimization
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f6745f4-7f2e-40e3-9df7-942bb196b3cc · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? Learning multiple layers of features from tiny images.(2009), 2009
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 948663c9-e9f1-4b37-8be8-10e5ca903714 · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? Beyond the doors of perception: Vision transformers represent relations between objects
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c6a79849-c175-4e02-a595-9753ee31a037 · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? LL a VA -onevision: Easy visual task transfer
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a7aab324-ef49-4635-a7d8-1afb4b707de1 · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? Inference-time intervention: Eliciting truthful answers from a language model
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88497041-3c36-4e81-98c0-491e01213a2f · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? Benchmarking Multi-modal Semantic Segmentation under Sensor Failures: Missing and Noisy Modality Robustness
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1a037fbd-c560-40f1-aafa-e548c41e9a26 · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? Improved baselines with visual instruction tuning, 2023
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49b6407e-6b08-4795-83df-bbc8955a76a8 · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? The quest for the right mediator: A history, survey, and theoretical grounding of causal interpretability
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 835729aa-e31d-4519-90a5-7c51c4c1ff2d · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? Towards interpreting visual information processing in vision-language models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 78824f13-9509-4b94-8d80-070aee0ff6c5 · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? Same task, different circuits: Disentangling modality-specific mechanisms in vlms
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4045827d-8ad6-4910-aef7-c948e23bd086 · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? Introducing operator
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a68d5511-6e85-4683-8fb3-0e6847d566ba · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? GPT-4 Technical Report
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f206f484-5183-40d2-a623-9fcab1b725fd · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? Interpreting the linear structure of vision-language model embedding spaces
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd9f27a9-779d-4f5d-9fa9-6b5c4f9df9a6 · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? V -measure: A conditional entropy-based external cluster evaluation measure
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 490fa9ea-1e47-4873-b46e-c7600c1a4bab · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? On the Adversarial Robustness of Multi-Modal Foundation Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27494d1a-1258-4ee3-8986-4321e9c8df62 · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? What do you learn from context? probing for sentence structure in contextualized word representations
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 13cb4d14-85a6-4c25-978d-617ffba7998b · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? Bert rediscovers the classical nlp pipeline
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 355d1c63-4f5c-4747-8764-fd6dccae3cb5 · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? Investigating gender bias in language models using causal mediation analysis
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2dec1286-d704-46a3-a389-0369d48bd845 · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? Analyzing multi-head self-attention: Specialized heads do the heavy lifting, the rest can be pruned
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c22cf30-9bc8-4c25-b8de-56ac146da63a · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? The caltech-ucsd birds-200-2011 dataset
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85216b34-3275-4f8e-8904-190d27a6f832 · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be17bcbc-6a94-4caf-9ef8-f2d0e315a1ea · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? Multimodal Inconsistency Reasoning (MMIR): A New Benchmark for Multimodal Reasoning Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cce0599b-a114-4793-bb89-cf434f2a65d5 · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? Characterizing mechanisms for factual recall in language models
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22a2a9d2-601a-437d-925d-a220b9229005 · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? Does vision-and-language pretraining improve lexical grounding? In Findings of the Association for Computational Linguistics: EMNLP 2021, pages 4357--4366, 2021
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1027564e-6fa4-4942-bd0d-549b07549037 · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? Emergence of abstract state representations in embodied sequence modeling
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 658b2ef7-f1d6-4d56-b06a-25407e037b18 · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? Calling a spade a heart: Gaslighting multimodal large language models via negation, 2025
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c47deb8-34ed-426f-b729-1e8ce3c23b4d · outbound
How Do Vision-Language Models Process Conflicting Information Across Modalities? Unraveling Cross-Modality Knowledge Conflicts in Large Vision-Language Models
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9767cf6-0900-4952-9bfb-3c0ca9bc9db0 · inbound
Can Large Multimodal Models Actively Recognize Faulty Inputs? A Systematic Evaluation Framework of Their Input Scrutiny Ability How Do Vision-Language Models Process Conflicting Information Across Modalities?
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c9086df-9a57-439d-a560-1ad63b475312 · inbound
TRANSPORTER: Transferring Visual Semantics from VLM Manifolds How Do Vision-Language Models Process Conflicting Information Across Modalities?
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 524773b5-315c-4eac-8bf8-a010888435df · inbound
Beyond Text-Dominance: Understanding Modality Preference of Omni-modal Large Language Models How Do Vision-Language Models Process Conflicting Information Across Modalities?
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 854e9a80-1a6a-48de-8edb-a3c3607b15ab · inbound
Vision-Default, Prior-Override: Causal Mechanisms of Perception-Knowledge Conflict in Vision-Language Models How Do Vision-Language Models Process Conflicting Information Across Modalities?
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation cf1015e8-e448-412d-a1fd-3f9e5624e96b · inbound
ScAle: Attention Head Scaling as a Minimal Adapter for Spatial Reasoning in Vision Language Models How Do Vision-Language Models Process Conflicting Information Across Modalities?
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 218f5833-c688-4f6d-b7e9-af96fcff062a · inbound
Attending to Multimodal Generation One Token at a Time How Do Vision-Language Models Process Conflicting Information Across Modalities?
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0acd5ca-061f-4b96-8016-9cd40cbb5726 · inbound
Linguistic Context Recodes Visual Representations in Vision-Language Models How Do Vision-Language Models Process Conflicting Information Across Modalities?
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.