Pith. sign in

Paper Citation Record · LEDGER

The Role of Chain-of-Thought in Complex Vision-Language Reasoning Task

As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2311.09193.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2311.09193 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T13:38:14.123422Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T08:09:40.792964Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 2a6769f1-860c-4fbf-a0f5-be13c7bc6c1d · inbound

CoMT: A Novel Benchmark for Chain of Multi-modal Thought on Large Vision-Language Models cites this paper.

CoMT: A Novel Benchmark for Chain of Multi-modal Thought on Large Vision-Language Models The Role of Chain-of-Thought in Complex Vision-Language Reasoning Task

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-11T13:38:14.123422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:38:14.123422Z digest=sha256:6d98fdee1de4226b655e8c3078e8208a9ad12f3059f56f33458a95baf7ce116e

Observation 11a83eb5-0294-4e10-8ef6-7f0b1f0beb4a · inbound

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey cites this paper.

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey The Role of Chain-of-Thought in Complex Vision-Language Reasoning Task

Reference 224

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:18:53.515245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-15T17:18:52.996467Z digest=sha256:cf15cdcc90512cccc0b5291126b953dfd7e44bd5df5471be42bab2948bcc5ef0

Observation 9aa33499-ec22-44ca-8371-05a812bd4fe2 · inbound

RSVP: Reasoning Segmentation via Visual Prompting and Multi-modal Chain-of-Thought cites this paper.

RSVP: Reasoning Segmentation via Visual Prompting and Multi-modal Chain-of-Thought The Role of Chain-of-Thought in Complex Vision-Language Reasoning Task

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T11:08:45.273322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:08:45.273322Z digest=sha256:de65ca2412b654f0028ba7248b4f2aa70a81da70ca92f4bd73f3511df6ae02f8

Observation 518bea90-bf82-478e-9b14-1bafb206c194 · inbound

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought cites this paper.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought The Role of Chain-of-Thought in Complex Vision-Language Reasoning Task

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.092686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.092686Z digest=sha256:2b719d300280188862c7773ff2ca8bd29547fea36facdba9b7149898c20f6079

Observation 75444443-fc2f-4161-a82e-a565c3d2ca3c · inbound

ViTCoT: Video-Text Interleaved Chain-of-Thought for Boosting Video Understanding in Large Language Models cites this paper.

ViTCoT: Video-Text Interleaved Chain-of-Thought for Boosting Video Understanding in Large Language Models The Role of Chain-of-Thought in Complex Vision-Language Reasoning Task

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T17:49:34.898886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:49:34.898886Z digest=sha256:09eaa4fc52850d5d3264f4981e19e538d36cf48d19cbe043b84e75302d5d3214

Observation a19da922-fc32-42d2-91b9-f0850cb05a4b · inbound

Visual Language Models as Zero-Shot Deepfake Detectors cites this paper.

Visual Language Models as Zero-Shot Deepfake Detectors The Role of Chain-of-Thought in Complex Vision-Language Reasoning Task

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-06T11:45:28.127701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:45:28.127701Z digest=sha256:7bcab24c7a544d7d7c2539f9290ad6923ded7bc2d21bf9feec34d30ee428bb04

Observation 3c2137cc-250d-4e5f-84a9-8b6c6eec4581 · inbound

Test-Time Matching: Unlocking Compositional Reasoning in Multimodal Models cites this paper.

Test-Time Matching: Unlocking Compositional Reasoning in Multimodal Models The Role of Chain-of-Thought in Complex Vision-Language Reasoning Task

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-18T09:51:13.574111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T09:48:39.943352Z digest=sha256:2156d7c95825ef752da1a311dcbc40cbc001f356dfe99621d1bbf158ae9477d5

Observation f03be398-58c4-4f23-b262-afa04cd7a487 · inbound

MIND: Multi-rationale INtegrated Discriminative Reasoning Framework for Multi-modal Large Models cites this paper.

MIND: Multi-rationale INtegrated Discriminative Reasoning Framework for Multi-modal Large Models The Role of Chain-of-Thought in Complex Vision-Language Reasoning Task

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-03T18:25:07.155502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:25:07.155502Z digest=sha256:90ccc30067f10f2615010e35ac3283debc70d87ce7a3e1490e1c25ee16cdfa34

Observation 4b1dbbce-d29e-4f78-9868-133ff82b7b81 · inbound

Improving Reasoning in Vision-Language Models via Perception Verified Self-Training cites this paper.

Improving Reasoning in Vision-Language Models via Perception Verified Self-Training The Role of Chain-of-Thought in Complex Vision-Language Reasoning Task

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-04T08:09:40.794691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-26T12:14:35.109298Z digest=sha256:0184cf7650f47e0b9808559e86136539dedcbcef1dced2ccdd28e29d29a540c3

Observation bde207e3-4750-47f7-8f65-dd79e4b47f46 · inbound

Improving Reasoning in Vision-Language Models via Perception Verified Self-Training cites this paper.

Improving Reasoning in Vision-Language Models via Perception Verified Self-Training The Role of Chain-of-Thought in Complex Vision-Language Reasoning Task

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-01T09:35:39.641794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-07-01T06:29:51.635039Z digest=sha256:5a93976a3125297fae603a303e65eaf9e1d9941da4d0e3611c5726f6f1d18f7f