Pith. sign in

Paper Citation Record · LEDGER

Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2305.02317.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2305.02317 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:10:11.266165Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T23:49:02.294828Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 8d2a8bfd-bce7-4e65-abcb-f1cb399801d9 · inbound

A Survey on Multimodal Large Language Models cites this paper.

A Survey on Multimodal Large Language Models Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 187

Resolution
verified exact
arxiv_id, observed 2026-05-16T02:56:42.155815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T02:56:41.658658Z digest=sha256:f4f7edb6e3d36b61ba88b1fd2304e966244974638f1f39e5b27d6d89279858c5

Observation 6a9d2b29-e7e5-4a5f-aade-acd9df9586fa · inbound

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey cites this paper.

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:18:53.689340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T17:18:52.996467Z digest=sha256:47603799dca6e1ab66044b3204acd9d086ca5bce439435693874ada8611e9ed2

Observation 01fe981e-0f3a-4d7e-97af-0a6f31164355 · inbound

CoT-VLA: Visual Chain-of-Thought Reasoning for Vision-Language-Action Models cites this paper.

CoT-VLA: Visual Chain-of-Thought Reasoning for Vision-Language-Action Models Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-16T05:21:44.988534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T05:21:44.903048Z digest=sha256:523e7dcece0136c7d11998c082ce82629f55920fa6790bc3da83c9073fc4c1fd

Observation 824d62e9-3f98-438e-9cc0-bbd5e9de4f34 · inbound

VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought cites this paper.

VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T15:10:11.266165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:10:11.266165Z digest=sha256:b6071e0d752796f500a5a036b7f1c85da5b79f655543d87482485a648651929f

Observation 2747b5cd-31f1-4a9a-a7d0-9b27b197983b · inbound

ReFineVLA: Reasoning-Aware Teacher-Guided Transfer Fine-Tuning cites this paper.

ReFineVLA: Reasoning-Aware Teacher-Guided Transfer Fine-Tuning Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:23:17.734824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:23:17.734824Z digest=sha256:64f182389b64dd317b35bca34380de7fb6e46f5c2e0721f4bc55fd55ac3d223f

Observation 251461e2-d61c-43f5-9dff-f1df33622f8f · inbound

Point-RFT: Improving Multimodal Reasoning with Visually Grounded Reinforcement Finetuning cites this paper.

Point-RFT: Improving Multimodal Reasoning with Visually Grounded Reinforcement Finetuning Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:05.316112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:05.316112Z digest=sha256:524913a2ac681c4b72d87833c900b6c6a30944f71892c3f012a7a03042222570

Observation 991f54e3-8231-4670-bdee-c8beff1d6ab8 · inbound

Argus: Vision-Centric Reasoning with Grounded Chain-of-Thought cites this paper.

Argus: Vision-Centric Reasoning with Grounded Chain-of-Thought Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T12:42:25.021502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:42:25.021502Z digest=sha256:7ac2e8c30c041da3d388ea5f5c2d3fdcd34e89b45daa5d63fa1bde86b5338b09

Observation d6973e5d-3d08-4214-bbe4-a5764c3ead37 · inbound

ViCrit: A Verifiable Reinforcement Learning Proxy Task for Visual Perception in VLMs cites this paper.

ViCrit: A Verifiable Reinforcement Learning Proxy Task for Visual Perception in VLMs Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T04:40:10.790536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:40:10.790536Z digest=sha256:38f3e61c7788ff06c919f2cc46dbb7c591a28b6aeb40e4b40e56ef431b2ce571

Observation 9515bdd8-7253-490f-8bd9-b5d53f00a0c8 · inbound

Argus Inspection: Do Multimodal Large Language Models Possess the Eye of Panoptes? cites this paper.

Argus Inspection: Do Multimodal Large Language Models Possess the Eye of Panoptes? Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T11:19:02.751761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:19:02.751761Z digest=sha256:f312448825e4f4a78116f172ca66011b6e98579f240d0bdb2717feedfac01108

Observation e1e41564-69ac-402c-b641-dd779435f174 · inbound

MagiC: Evaluating Multimodal Cognition Toward Grounded Visual Reasoning cites this paper.

MagiC: Evaluating Multimodal Cognition Toward Grounded Visual Reasoning Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T18:51:05.057773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:51:05.057773Z digest=sha256:cbaad11cf903c2773992adbf0c8b599db82b9f6d96127aa25f25911f010e95b5

Observation 1d9134d9-2d67-4ef6-8865-6454a3103638 · inbound

WebWatcher: Breaking New Frontier of Vision-Language Deep Research Agent cites this paper.

WebWatcher: Breaking New Frontier of Vision-Language Deep Research Agent Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-15T18:56:23.987203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T18:56:23.817544Z digest=sha256:67e0f753314ecd14b588251c160e8022f21fbd2758c7b64fb3e58d4ccb7b7511

Observation ad075480-7c56-48d2-b986-c70603077360 · inbound

Less is More Tokens: Efficient Math Reasoning via Difficulty-Aware Chain-of-Thought Distillation cites this paper.

Less is More Tokens: Efficient Math Reasoning via Difficulty-Aware Chain-of-Thought Distillation Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-05T05:34:40.348235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:34:40.348235Z digest=sha256:159d2f99c9e19a8944ba47a9cdff3096f277330100000c1d4a1f310258be445b

Observation bcaae467-e363-4982-b444-879b2eaf533f · inbound

Mitigating Visual Context Degradation in Large Multimodal Models: A Training-Free Decoupled Agentic Framework cites this paper.

Mitigating Visual Context Degradation in Large Multimodal Models: A Training-Free Decoupled Agentic Framework Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-18T12:32:36.347236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-18T12:31:25.257879Z digest=sha256:f7a5bcd7197e04a38e6a26697e76b4d57283f82336f098fc5b75245191f81148

Observation c9501840-9dc0-4a57-8174-2ed738536823 · inbound

ReFineVLA: Multimodal Reasoning-Aware Generalist Robotic Policies via Teacher-Guided Fine-Tuning cites this paper.

ReFineVLA: Multimodal Reasoning-Aware Generalist Robotic Policies via Teacher-Guided Fine-Tuning Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T09:48:48.348511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T05:06:38.517652Z digest=sha256:1290bfb6512ff2fccb79200449893d750e1bd0a589538ef51c781b9fb0ea2327

Observation ab011063-c40e-4571-a5aa-9352f7f6effe · inbound

R-CoV: Region-Aware Chain-of-Verification for Alleviating Object Hallucinations in LVLMs cites this paper.

R-CoV: Region-Aware Chain-of-Verification for Alleviating Object Hallucinations in LVLMs Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:36:04.240674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T01:28:14.039910Z digest=sha256:08a6a32150f7ab182597bfa51b1432bf41750f827449ea07af936a8a844fd1c5

Observation 595f3cef-e4df-4b0c-95c9-73f5a1ade0d6 · inbound

Visual-OPSD: Cross-Modal On-Policy Self-Distillation for Efficient Unified Multimodal Reasoning cites this paper.

Visual-OPSD: Cross-Modal On-Policy Self-Distillation for Efficient Unified Multimodal Reasoning Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T23:49:02.296271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-26T21:47:51.437284Z digest=sha256:9c34a51b2e84e99b6bcc7bbf180f2075ca6e4bd61bd24322fbca5904a99e2754

Observation 1e0eeb18-b51d-4fbc-9a48-fd58cc5f6a4c · inbound

ReShift: Aha-Moment-Driven Reasoning-Level Backdoor Attacks on Vision-Language Models cites this paper.

ReShift: Aha-Moment-Driven Reasoning-Level Backdoor Attacks on Vision-Language Models Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T11:56:54.961436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-02T11:48:35.972372Z digest=sha256:34204305be1a43aea0c5d4e5fe1ae4b4989671ba547c1186e040c57edbb8afd4