Pith. sign in

Paper Citation Record · LEDGER

VL-ICL Bench: The Devil in the Details of Multimodal In-Context Learning

As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2403.13164.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2403.13164 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T00:46:54.936952Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T05:07:39.144487Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 37dfd2de-3af0-4c01-b9b1-233381126529 · inbound

Image-Text Relation Prediction for Multilingual Tweets cites this paper.

Image-Text Relation Prediction for Multilingual Tweets VL-ICL Bench: The Devil in the Details of Multimodal In-Context Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T23:18:27.679231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:18:27.679231Z digest=sha256:ffca9ad43ecc7419faad990c354ae60965c028260ba0757b982f08e7a8234876

Observation 964eca0e-c530-40d4-8056-76e070db338b · inbound

Mimicking or Reasoning: Rethinking Multi-Modal In-Context Learning in Vision-Language Models cites this paper.

Mimicking or Reasoning: Rethinking Multi-Modal In-Context Learning in Vision-Language Models VL-ICL Bench: The Devil in the Details of Multimodal In-Context Learning

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T05:25:30.457496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:25:30.457496Z digest=sha256:5025b6b32336df101cbb417d8eafd6e33bbd75405d937c9c0e657b0488252626

Observation 993d4abe-3f16-40b6-9fbd-24509c847c6e · inbound

True Multimodal In-Context Learning Needs Attention to the Visual Context cites this paper.

True Multimodal In-Context Learning Needs Attention to the Visual Context VL-ICL Bench: The Devil in the Details of Multimodal In-Context Learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T15:29:35.474334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:29:35.474334Z digest=sha256:4b86e020a0ef92db9fb798211d23980e116fe57e81d31ff7e084761467917372

Observation ced217b6-825d-472f-9eb2-4d0ad76de108 · inbound

UniICL: Systematizing Unified Multimodal In-context Learning through a Capability-Oriented Taxonomy cites this paper.

UniICL: Systematizing Unified Multimodal In-context Learning through a Capability-Oriented Taxonomy VL-ICL Bench: The Devil in the Details of Multimodal In-Context Learning

Reference 53

Resolution
unresolved
no resolver link, observed 2026-07-13T18:42:33.178622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T18:42:33.178622Z digest=sha256:c572cbfeaa80b404eb6ea1deef315b66e2d90d397ab5d1e2602dacb622459d0a

Observation 0a893d77-d2e7-4217-8561-96b08fc881bf · inbound

Why Multimodal In-Context Learning Lags Behind? Unveiling the Inner Mechanisms and Bottlenecks cites this paper.

Why Multimodal In-Context Learning Lags Behind? Unveiling the Inner Mechanisms and Bottlenecks VL-ICL Bench: The Devil in the Details of Multimodal In-Context Learning

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:45:28.192426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-10T13:41:37.942145Z digest=sha256:00b50fddeef4d7185863612a0f8c56c7d63e4f186c0b780692f856415c07e845

Observation 561106d6-f70a-4b57-a4b3-a5b9d8c14d46 · inbound

QCalEval: Benchmarking Vision-Language Models for Quantum Calibration Plot Understanding cites this paper.

QCalEval: Benchmarking Vision-Language Models for Quantum Calibration Plot Understanding VL-ICL Bench: The Devil in the Details of Multimodal In-Context Learning

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:46:24.436457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-07T16:22:12.446406Z digest=sha256:78562572b71ddcaee64846672218bfc5b234cd933f3a6e7757d751eb0d4c50e5

Observation b02ca60a-0dd4-4684-b16e-eef35874ecc2 · inbound

Personal Visual Context Learning in Large Multimodal Models cites this paper.

Personal Visual Context Learning in Large Multimodal Models VL-ICL Bench: The Devil in the Details of Multimodal In-Context Learning

Reference 85

Resolution
malformed identifier
arxiv_id, observed 2026-05-12T07:06:37.331198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-12T03:42:15.402131Z digest=sha256:dca855265ff051e065b8179d2ff7b9778f0dfa4f47c581afc6884ce2b2c784e4

Observation 352b51bc-8eb7-447d-a147-9b4e415cfae9 · inbound

MMCL-Bench: Multimodal Context Learning from Visual Rules, Procedures, and Evidence cites this paper.

MMCL-Bench: Multimodal Context Learning from Visual Rules, Procedures, and Evidence VL-ICL Bench: The Devil in the Details of Multimodal In-Context Learning

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:59:27.481171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-14T20:55:50.873271Z digest=sha256:ecfd25e3009ac1a141907c2e09fd2c91aa38c516b3f63b07f23442632674bec6

Observation d3659923-4b80-4e71-8a48-fc26a48f0453 · inbound

Dive into the Scene: Breaking the Perceptual Bottleneck in Vision-Language Decision Making via Focus Plan Generation cites this paper.

Dive into the Scene: Breaking the Perceptual Bottleneck in Vision-Language Decision Making via Focus Plan Generation VL-ICL Bench: The Devil in the Details of Multimodal In-Context Learning

Reference 59

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T02:56:29.254693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-28T10:28:41.952330Z digest=sha256:7a943b60d9903220c1fdca0ca6f1a076508b0390346e290047a705866adbdab7

Observation 719236bc-1804-4ea1-9b81-84eb3d7b2a54 · inbound

Quo Vadis, Visual In-Context Learning? A Unified Benchmark Across Domains and Tasks cites this paper.

Quo Vadis, Visual In-Context Learning? A Unified Benchmark Across Domains and Tasks VL-ICL Bench: The Devil in the Details of Multimodal In-Context Learning

Reference 111

Resolution
malformed identifier
arxiv_id, observed 2026-07-03T05:07:39.146050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-27T13:24:48.138643Z digest=sha256:1b81f481e650a49ebba99871e597a556a152fbe808df671e471cc2288f2f87b6

Observation fcd82c62-f782-4b8b-8aa5-f253d14000c7 · inbound

MAG: MAnifold Guided Semi-Supervised Multi-modal In-Context Learning cites this paper.

MAG: MAnifold Guided Semi-Supervised Multi-modal In-Context Learning VL-ICL Bench: The Devil in the Details of Multimodal In-Context Learning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-16T00:46:54.936952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:46:54.936952Z digest=sha256:fa75bc3d0b4e71ea647801eea617177e6bdb9d2a992bb474e47c54d4f93c48d0