Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2301.05226.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:35:55.800954Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-22T22:42:13.506064Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation f804bb78-8fa3-4b9b-918e-eff6e51e1459 · inbound
Multimodal Chain-of-Thought Reasoning in Language Models See, Think, Confirm: Interactive Prompting Between Vision and Language Models for Knowledge-based Visual Reasoning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fdbd6c0a-404a-4908-9f9f-0152b88d2277 · inbound
Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey See, Think, Confirm: Interactive Prompting Between Vision and Language Models for Knowledge-based Visual Reasoning
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0cc0e3a7-0202-4161-86e2-959667e1bc7e · inbound
When 'YES' Meets 'BUT': Can Large Models Comprehend Contradictory Humor Through Comparative Reasoning? See, Think, Confirm: Interactive Prompting Between Vision and Language Models for Knowledge-based Visual Reasoning
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 431d7a65-b4a8-4db9-8a57-5d0e8b9fee88 · inbound
MME-Reasoning: A Comprehensive Benchmark for Logical Reasoning in MLLMs See, Think, Confirm: Interactive Prompting Between Vision and Language Models for Knowledge-based Visual Reasoning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6eee8200-8caf-4adc-98e2-211589b90ad5 · inbound
VFaith: Do Large Multimodal Models Really Reason on Seen Images Rather than Previous Memories? See, Think, Confirm: Interactive Prompting Between Vision and Language Models for Knowledge-based Visual Reasoning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f50ac4c8-1124-4a79-b6c8-e32f1c9947e9 · inbound
MagiC: Evaluating Multimodal Cognition Toward Grounded Visual Reasoning See, Think, Confirm: Interactive Prompting Between Vision and Language Models for Knowledge-based Visual Reasoning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e53b242c-3402-4e09-9564-67467c7bba45 · inbound
GoViG: Goal-Conditioned Visual Navigation Instruction Generation via Multimodal Reasoning See, Think, Confirm: Interactive Prompting Between Vision and Language Models for Knowledge-based Visual Reasoning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation da99e7ad-f238-43cd-86f2-4d729a8d488d · inbound
LaRe: Latent Refocusing for Multimodal Reasoning See, Think, Confirm: Interactive Prompting Between Vision and Language Models for Knowledge-based Visual Reasoning
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 535c2719-7ef5-4792-bad4-758da2518608 · inbound
MIND: Multi-rationale INtegrated Discriminative Reasoning Framework for Multi-modal Large Models See, Think, Confirm: Interactive Prompting Between Vision and Language Models for Knowledge-based Visual Reasoning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.