Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:2109.04448.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-12T11:06:10.026636Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-20T15:03:24.780269Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation d646a383-9468-4e36-af97-baa081b32688 · inbound
Cross-modal Information Flow in Multimodal Large Language Models Vision-and-Language or Vision-for-Language? On Cross-Modal Influence in Multimodal Transformers
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 168d1ec8-1e15-43a4-9108-42b99c2a9024 · inbound
Explainable and Interpretable Multimodal Large Language Models: A Comprehensive Survey Vision-and-Language or Vision-for-Language? On Cross-Modal Influence in Multimodal Transformers
Reference 151
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa39ec64-9548-49a8-8b08-436f511eb967 · inbound
A Review of Multimodal Explainable Artificial Intelligence: Past, Present and Future Vision-and-Language or Vision-for-Language? On Cross-Modal Influence in Multimodal Transformers
Reference 203
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 251beba0-629c-4e9b-9f7d-5a2103929d7a · inbound
Medical Context Distorts Decisions in Clinical Vision Language Models Vision-and-Language or Vision-for-Language? On Cross-Modal Influence in Multimodal Transformers
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.