Pith. sign in

Paper Citation Record · LEDGER

Language Models Can See: Plugging Visual Controls in Text Generation

As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2205.02655.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2205.02655 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T15:58:41.918019Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T08:12:30.132637Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation fa8d95b6-64c7-48ed-bb78-d00c76b7cf0e · inbound

Socratic Models: Composing Zero-Shot Multimodal Reasoning with Language cites this paper.

Socratic Models: Composing Zero-Shot Multimodal Reasoning with Language Language Models Can See: Plugging Visual Controls in Text Generation

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-05-16T09:50:00.642396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T09:50:00.546571Z digest=sha256:95b986e0fa2a9b72e69559fb59b476c34b2a6fd6dace662f5bde242598673a4d

Observation ca1f1af3-11e0-4760-af35-63c1e14e746b · inbound

PandaGPT: One Model To Instruction-Follow Them All cites this paper.

PandaGPT: One Model To Instruction-Follow Them All Language Models Can See: Plugging Visual Controls in Text Generation

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-16T09:01:37.131850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T09:01:37.096206Z digest=sha256:40c9c4836a1dba1ccf8368c079ef885684ef1ddf487089779c1313deebac4a0b

Observation ebaf16d6-d807-4cd0-a7a6-54c22a0c4998 · inbound

Intelligent Agents with Emotional Intelligence: Current Trends, Challenges, and Future Prospects cites this paper.

Intelligent Agents with Emotional Intelligence: Current Trends, Challenges, and Future Prospects Language Models Can See: Plugging Visual Controls in Text Generation

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-18T08:12:30.136017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T08:11:31.181704Z digest=sha256:bd7af8767a730bb507b60ec8e833c5052560bbc7bae43e7ad0f25294b87a663b

Observation 205e5e13-5c70-47b0-94ea-196f0d545c4d · inbound

Modality Gap-Driven Subspace Alignment Training Paradigm For Multimodal Large Language Models cites this paper.

Modality Gap-Driven Subspace Alignment Training Paradigm For Multimodal Large Language Models Language Models Can See: Plugging Visual Controls in Text Generation

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:17:36.102757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T08:17:29.924860Z digest=sha256:a766505f07d192cde966f14ad5557a6e65a60f741be53bbb4372db9b40eb9f46

Observation 690e0f5a-ab45-4e6e-87ce-04e46591f37a · inbound

Modality Gap-Driven Subspace Alignment Training Paradigm For Multimodal Large Language Models cites this paper.

Modality Gap-Driven Subspace Alignment Training Paradigm For Multimodal Large Language Models Language Models Can See: Plugging Visual Controls in Text Generation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-03T05:34:29.001268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:34:29.001268Z digest=sha256:4fe2d6345c2a2f7a12de8896a060bb039c135eedd1b87e71bef03601fb12fb95

Observation 29dba067-8766-412f-99e3-33e03a8f1616 · inbound

Adjudicated Captioning: Multi-Agent Alignment Scoring and Consensus-Distilled Beam Arbitration for Strict Zero-Shot Image Captioning cites this paper.

Adjudicated Captioning: Multi-Agent Alignment Scoring and Consensus-Distilled Beam Arbitration for Strict Zero-Shot Image Captioning Language Models Can See: Plugging Visual Controls in Text Generation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T15:58:41.918019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:58:41.918019Z digest=sha256:1cc9281fff041e8f62bb5b9ede77a8a4032b2824e2a45e3261cf9fbb30b3ad9b