Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-10T09:01:35.425918Z
Paper Citation Record · LEDGER
As of 23 July 2026, this Paper Citation Record lists 11 of 11 outbound references and 2 inbound Pith citation observations for arXiv:2604.16256.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-10T09:01:35.425918Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-07-23T06:31:01.910684+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-13T07:43:50.633779Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-05-13T07:47:33.319161Z
11 of 11 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation e9d5d633-9a7c-456d-9a8e-76e60b701bc8 · outbound
Do Vision-Language Models Truly Perform Vision Reasoning? A Rigorous Study of the Modality Gap OpenAI GPT-5 System Card
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 4f39df01-7084-46e2-8517-4267c377a059 · outbound
Do Vision-Language Models Truly Perform Vision Reasoning? A Rigorous Study of the Modality Gap EasyARC: Evaluating Vision Language Models on True Visual Reasoning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 86b2b7ce-77b7-49d2-80a1-a10ee83e03a1 · outbound
Do Vision-Language Models Truly Perform Vision Reasoning? A Rigorous Study of the Modality Gap Emogen: Emotional image content generation with text-to-image diffusion models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation c8299b56-70f6-4268-af4e-2416ece439b4 · outbound
Do Vision-Language Models Truly Perform Vision Reasoning? A Rigorous Study of the Modality Gap URLhttps://doi.org/10.1007/978-3-031-73242-3 10
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 03aa5c2d-3571-4e31-b762-40cedb41e5df · outbound
Do Vision-Language Models Truly Perform Vision Reasoning? A Rigorous Study of the Modality Gap You are given a **cross math puzzle** in a **textual markdown grid format**
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation c43f78a9-55c5-4b25-84e4-7c2adf8cc0d8 · outbound
Do Vision-Language Models Truly Perform Vision Reasoning? A Rigorous Study of the Modality Gap You are given a **cross math puzzle** in a **image format**
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation b6a6c397-ee48-4675-9520-29d59af3fc58 · outbound
Do Vision-Language Models Truly Perform Vision Reasoning? A Rigorous Study of the Modality Gap You are given a **cross math puzzle** in **both image format and textual markdown format**
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 56fad0d0-b3f6-4e86-bb69-4f921626dc61 · outbound
Do Vision-Language Models Truly Perform Vision Reasoning? A Rigorous Study of the Modality Gap Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 86efefc3-01d1-40c7-ac84-e50071f55c7d · outbound
Do Vision-Language Models Truly Perform Vision Reasoning? A Rigorous Study of the Modality Gap Unresolved cited work
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation 1b155ff9-751c-4abf-b921-151e36394255 · outbound
Do Vision-Language Models Truly Perform Vision Reasoning? A Rigorous Study of the Modality Gap Unresolved cited work
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation ac2646c3-2185-4d7c-8750-61950c38d326 · outbound
Do Vision-Language Models Truly Perform Vision Reasoning? A Rigorous Study of the Modality Gap You are given a **cross math puzzle** in a **textual markdown grid format**
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation bb1354e6-797d-4137-88e5-ff5f2f9e9713 · inbound
SeePhys Pro: Diagnosing Modality Transfer and Blind-Training Effects in Multimodal RLVR for Physics Reasoning Do Vision-Language Models Truly Perform Vision Reasoning? A Rigorous Study of the Modality Gap
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.
Observation a8d8672a-8a01-4291-b1b9-c2915943612b · inbound
SeePhys Pro: Diagnosing Modality Transfer and Blind-Training Effects in Multimodal RLVR for Physics Reasoning Do Vision-Language Models Truly Perform Vision Reasoning? A Rigorous Study of the Modality Gap
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.