Pith. sign in

Paper Citation Record · LEDGER

What does CLIP know about a red circle? Visual prompt engineering for VLMs

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2304.06712.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2304.06712 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-14T04:22:35.763669Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-24T03:33:50.786236Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ac6869bc-71bd-42a0-afd8-8f19df449366 · inbound

The Dawn of LMMs: Preliminary Explorations with GPT-4V(ision) cites this paper.

The Dawn of LMMs: Preliminary Explorations with GPT-4V(ision) What does CLIP know about a red circle? Visual prompt engineering for VLMs

Reference 117

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T23:26:06.303819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T23:26:06.183574Z digest=sha256:ef39a5c7c3b15e2fa1de2e8a5729aece1154551c9b8a8b23382496acca7b0bca

Observation d66c9429-a5b5-4afc-ba0d-b70f8ad9d70b · inbound

RAR: Retrieving And Ranking Augmented MLLMs for Visual Recognition cites this paper.

RAR: Retrieving And Ranking Augmented MLLMs for Visual Recognition What does CLIP know about a red circle? Visual prompt engineering for VLMs

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-24T03:33:50.788761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-24T03:31:58.848180Z digest=sha256:6cc6798877fb2d58f1c5e778e8c96a5c6281fdbc108244e3554b64663a0b3f3a

Observation dcb66970-d321-4245-98b2-a09e0c80b1e4 · inbound

BLINK: Multimodal Large Language Models Can See but Not Perceive cites this paper.

BLINK: Multimodal Large Language Models Can See but Not Perceive What does CLIP know about a red circle? Visual prompt engineering for VLMs

Reference 67

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:18:15.551709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T20:18:15.439163Z digest=sha256:bc9777d8adba88b8379e3e5752f20aa0a8267dcaab2e23a18c586dafdcd01277

Observation 4af9e36e-5d89-4cab-8991-79139306fd1e · inbound

Panther: Illuminate the Sight of Multimodal LLMs with Instruction-Guided Visual Prompts cites this paper.

Panther: Illuminate the Sight of Multimodal LLMs with Instruction-Guided Visual Prompts What does CLIP know about a red circle? Visual prompt engineering for VLMs

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-12T15:51:51.622598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:51:51.622598Z digest=sha256:10b480ffa19fac6d67123394d1a004799d8e9145827b3e393688246bf54d5719

Observation 13c64e38-2859-4455-b351-b67f34e1f291 · inbound

A Holistically Point-guided Text Framework for Weakly-Supervised Camouflaged Object Detection cites this paper.

A Holistically Point-guided Text Framework for Weakly-Supervised Camouflaged Object Detection What does CLIP know about a red circle? Visual prompt engineering for VLMs

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-10T21:12:13.974754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:12:13.974754Z digest=sha256:bdb42eda92fb3c718f6af6d48724cea8872264dc4b336efd232a2439ee14f67a

Observation cef1a705-6e1d-4fe4-b967-8f7369bcafec · inbound

LPOI: Listwise Preference Optimization for Vision Language Models cites this paper.

LPOI: Listwise Preference Optimization for Vision Language Models What does CLIP know about a red circle? Visual prompt engineering for VLMs

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T13:45:56.874054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:45:56.874054Z digest=sha256:0182742a19eb227dcc7489d11858ca0300756d4660168e792ac9a44b8145f014

Observation 28371d84-9f29-4461-852d-8375f006c0d9 · inbound

ConText: Driving In-context Learning for Text Removal and Segmentation cites this paper.

ConText: Driving In-context Learning for Text Removal and Segmentation What does CLIP know about a red circle? Visual prompt engineering for VLMs

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T11:01:11.648041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:01:11.648041Z digest=sha256:94ee0bf4dce44bf3a64ab98f59b8406f2603becc60476fb4cc890e8d8479aa59

Observation b7290177-dbcc-48bf-afe5-da6924556643 · inbound

Decouple before Align: Visual Disentanglement Enhances Prompt Tuning cites this paper.

Decouple before Align: Visual Disentanglement Enhances Prompt Tuning What does CLIP know about a red circle? Visual prompt engineering for VLMs

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T10:17:04.350951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:17:04.350951Z digest=sha256:7b2828091dadab705abf0f0148d8e05a267efa7a60f1463bd0c8251036da3ae7

Observation 15dba7cd-b3ed-4867-80aa-c3808a26b4e3 · inbound

Finding Needles in Images: Can Multimodal LLMs Locate Fine Details? cites this paper.

Finding Needles in Images: Can Multimodal LLMs Locate Fine Details? What does CLIP know about a red circle? Visual prompt engineering for VLMs

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-05T23:35:22.821052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:35:22.821052Z digest=sha256:a39de1784c7a13f7e1043aeda1ec44161e71420858c99b3acddca57689d47654

Observation 8a56b3ae-f222-4922-919f-cef4794ab2c3 · inbound

Diagnosing as Cardiologists Do: ECG Agents with Doctor-Grounded Priors for Clinical Reasoning Across Diseases and Populations cites this paper.

Diagnosing as Cardiologists Do: ECG Agents with Doctor-Grounded Priors for Clinical Reasoning Across Diseases and Populations What does CLIP know about a red circle? Visual prompt engineering for VLMs

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-14T04:22:35.763669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:22:35.763669Z digest=sha256:bdd363114ef1191f5b148a566b67749ab72156d1f9b922c8a9e25941f3bd416b