Pith. sign in

Paper Citation Record · LEDGER

CoCoT: Contrastive Chain-of-Thought Prompting for Large Multimodal Models with Multiple Image Inputs

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2401.02582.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2401.02582 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:27:05.064735Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T09:51:13.622166Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation af53e087-66fa-4a22-80f5-43a94b385d8b · inbound

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey cites this paper.

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey CoCoT: Contrastive Chain-of-Thought Prompting for Large Multimodal Models with Multiple Image Inputs

Reference 67

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:18:53.703843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T17:18:52.996467Z digest=sha256:765974c20b57ecd643749fa6ff28378262c01dbbaf97001b914229e6527dc236

Observation 5793574b-1f18-435d-8eb9-a6f95939c91f · inbound

WeThink: Toward General-purpose Vision-Language Reasoning via Reinforcement Learning cites this paper.

WeThink: Toward General-purpose Vision-Language Reasoning via Reinforcement Learning CoCoT: Contrastive Chain-of-Thought Prompting for Large Multimodal Models with Multiple Image Inputs

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:05.064735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:05.064735Z digest=sha256:807cb98b1aede5f566fedea7f7a8c2d6d406750aca07299b5a41396f180ba4b2

Observation db9e8f2c-82da-40c2-9525-65d048c539d0 · inbound

Extracting Multimodal Learngene in CLIP: Unveiling the Multimodal Generalizable Knowledge cites this paper.

Extracting Multimodal Learngene in CLIP: Unveiling the Multimodal Generalizable Knowledge CoCoT: Contrastive Chain-of-Thought Prompting for Large Multimodal Models with Multiple Image Inputs

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T23:41:44.018505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:41:44.018505Z digest=sha256:badaba8cac4765443ba9f36e28b31aaaa5e190289129a6b3cb76bb9c5f874ab4

Observation 90347a6c-eb5a-4ac5-9b91-3134c0d0cba6 · inbound

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought cites this paper.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought CoCoT: Contrastive Chain-of-Thought Prompting for Large Multimodal Models with Multiple Image Inputs

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.765868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.765868Z digest=sha256:bcbc3d5fdad42f1e3c072958f3fce60423223e701c01f41b43496dbb9025f08b

Observation 3d26c41d-b290-45bc-adf5-eab104e8dd3b · inbound

PDB-Eval: An Evaluation of Large Multimodal Models for Description and Explanation of Personalized Driving Behavior cites this paper.

PDB-Eval: An Evaluation of Large Multimodal Models for Description and Explanation of Personalized Driving Behavior CoCoT: Contrastive Chain-of-Thought Prompting for Large Multimodal Models with Multiple Image Inputs

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T14:36:55.887586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:36:55.887586Z digest=sha256:4412c3ad9be3eec68afec05f28a0067b2d6fa0b594e88fe1ad2e8ee839ada2cb

Observation 6e96639b-de49-4f13-b5ff-0d16025061b7 · inbound

Test-Time Matching: Unlocking Compositional Reasoning in Multimodal Models cites this paper.

Test-Time Matching: Unlocking Compositional Reasoning in Multimodal Models CoCoT: Contrastive Chain-of-Thought Prompting for Large Multimodal Models with Multiple Image Inputs

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-18T09:51:13.625048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-18T09:48:39.943352Z digest=sha256:2b18c474ac1dafb952400ad51ee6e65c53658ff4d1220e01a94adf52b4c20aeb

Observation cd670699-7307-452d-85fd-260683163256 · inbound

VisRAG2.0: Mitigating Visual Hallucinations via Evidence-Guided Multi-Image Reasoning in Visual Retrieval-Augmented Generation cites this paper.

VisRAG2.0: Mitigating Visual Hallucinations via Evidence-Guided Multi-Image Reasoning in Visual Retrieval-Augmented Generation CoCoT: Contrastive Chain-of-Thought Prompting for Large Multimodal Models with Multiple Image Inputs

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T10:40:43.075026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:40:43.075026Z digest=sha256:025c1289510bf1c842d2bb73c1798a45a2cc36d16b4740b397c45b6c3f5dfd94

Observation f0f39b06-f2e2-487d-9d91-f9b1e1246ea0 · inbound

MIND: Multi-rationale INtegrated Discriminative Reasoning Framework for Multi-modal Large Models cites this paper.

MIND: Multi-rationale INtegrated Discriminative Reasoning Framework for Multi-modal Large Models CoCoT: Contrastive Chain-of-Thought Prompting for Large Multimodal Models with Multiple Image Inputs

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-03T18:25:08.021152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:25:08.021152Z digest=sha256:edd0a440e2d282e1a93eeb7e85156f13c74c0d4dbe9cdd0d9c4bb8c875afbfc0