Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2404.18624.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T10:22:23.223689Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-01T22:06:16.588829Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation c65432de-aa7e-42d4-b967-18ee6a206c22 · inbound
On the Risk of Misleading Reports: Diagnosing Textual Biases in Multimodal Clinical AI Do Vision & Language Decoders use Images and Text equally? How Self-consistent are their Explanations?
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7eeb6ec-974e-42a0-905d-091b4df582d1 · inbound
When to Call an Apple Red: Humans Follow Introspective Rules, VLMs Don't Do Vision & Language Decoders use Images and Text equally? How Self-consistent are their Explanations?
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 92e8d1eb-a8ea-4443-b7dd-505820fb9694 · inbound
Mitigating Action-Relation Hallucinations in LVLMs via Relation-aware Visual Enhancement Do Vision & Language Decoders use Images and Text equally? How Self-consistent are their Explanations?
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 88efbdee-8564-4b3e-ba26-903f1452f388 · inbound
Medical Context Distorts Decisions in Clinical Vision Language Models Do Vision & Language Decoders use Images and Text equally? How Self-consistent are their Explanations?
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 362d4bae-2551-4d28-8683-906cb4e3e56e · inbound
Attention-guided Fine-tuning of Multimodal Large Language Models Improves Chain-of-Thought Reasoning Do Vision & Language Decoders use Images and Text equally? How Self-consistent are their Explanations?
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.