Pith. sign in

Paper Citation Record · LEDGER

VL-ICL Bench: The Devil in the Details of Multimodal In-Context Learning

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2403.13164.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2403.13164 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:25:30.457496Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T05:07:39.144487Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 964eca0e-c530-40d4-8056-76e070db338b · inbound

Mimicking or Reasoning: Rethinking Multi-Modal In-Context Learning in Vision-Language Models cites this paper.

Mimicking or Reasoning: Rethinking Multi-Modal In-Context Learning in Vision-Language Models VL-ICL Bench: The Devil in the Details of Multimodal In-Context Learning

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T05:25:30.457496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:25:30.457496Z digest=sha256:18b71524016f783e8f7c704b3c398201e1f3ddff0fa02bc5bafb6f4c3f6b19b6

Observation 993d4abe-3f16-40b6-9fbd-24509c847c6e · inbound

True Multimodal In-Context Learning Needs Attention to the Visual Context cites this paper.

True Multimodal In-Context Learning Needs Attention to the Visual Context VL-ICL Bench: The Devil in the Details of Multimodal In-Context Learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T15:29:35.474334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:29:35.474334Z digest=sha256:1f968baae074e50c310ba8d3f7d13431052cb12a860777d9fcb46c26d60928e4

Observation ced217b6-825d-472f-9eb2-4d0ad76de108 · inbound

UniICL: Systematizing Unified Multimodal In-context Learning through a Capability-Oriented Taxonomy cites this paper.

UniICL: Systematizing Unified Multimodal In-context Learning through a Capability-Oriented Taxonomy VL-ICL Bench: The Devil in the Details of Multimodal In-Context Learning

Reference 53

Resolution
unresolved
no resolver link, observed 2026-07-13T18:42:33.178622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T18:42:33.178622Z digest=sha256:3f0eb714255f2e60af85fdc6a588ad7b90dd8f21e543c76b5b44482dcf514c8d

Observation 0a893d77-d2e7-4217-8561-96b08fc881bf · inbound

Why Multimodal In-Context Learning Lags Behind? Unveiling the Inner Mechanisms and Bottlenecks cites this paper.

Why Multimodal In-Context Learning Lags Behind? Unveiling the Inner Mechanisms and Bottlenecks VL-ICL Bench: The Devil in the Details of Multimodal In-Context Learning

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:45:28.192426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T13:41:37.942145Z digest=sha256:83df23282527b98a281772dfb95b6376905ebff516752debebdfe7f9e4ecac7a

Observation 561106d6-f70a-4b57-a4b3-a5b9d8c14d46 · inbound

QCalEval: Benchmarking Vision-Language Models for Quantum Calibration Plot Understanding cites this paper.

QCalEval: Benchmarking Vision-Language Models for Quantum Calibration Plot Understanding VL-ICL Bench: The Devil in the Details of Multimodal In-Context Learning

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:46:24.436457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-07T16:22:12.446406Z digest=sha256:6aefdafa9ef78e135de7a1d3b0eaa24e59b7a5659196f116c3ea022a74428118

Observation b02ca60a-0dd4-4684-b16e-eef35874ecc2 · inbound

Personal Visual Context Learning in Large Multimodal Models cites this paper.

Personal Visual Context Learning in Large Multimodal Models VL-ICL Bench: The Devil in the Details of Multimodal In-Context Learning

Reference 85

Resolution
malformed identifier
arxiv_id, observed 2026-05-12T07:06:37.331198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T03:42:15.402131Z digest=sha256:938367423b9c8ce94b45ee48f827201cd72af01655acd7e05e2d522697aa3867

Observation 352b51bc-8eb7-447d-a147-9b4e415cfae9 · inbound

MMCL-Bench: Multimodal Context Learning from Visual Rules, Procedures, and Evidence cites this paper.

MMCL-Bench: Multimodal Context Learning from Visual Rules, Procedures, and Evidence VL-ICL Bench: The Devil in the Details of Multimodal In-Context Learning

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:59:27.481171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-14T20:55:50.873271Z digest=sha256:d01488e9aae5a8135659462acdd28099518138a40336835abff5b8ea72cc15e6

Observation d3659923-4b80-4e71-8a48-fc26a48f0453 · inbound

Dive into the Scene: Breaking the Perceptual Bottleneck in Vision-Language Decision Making via Focus Plan Generation cites this paper.

Dive into the Scene: Breaking the Perceptual Bottleneck in Vision-Language Decision Making via Focus Plan Generation VL-ICL Bench: The Devil in the Details of Multimodal In-Context Learning

Reference 59

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T02:56:29.254693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-28T10:28:41.952330Z digest=sha256:9451e7734eb857d87caecac79b13b938f674f3dff3c1b37867dcb04d86d6de0c

Observation 719236bc-1804-4ea1-9b81-84eb3d7b2a54 · inbound

Quo Vadis, Visual In-Context Learning? A Unified Benchmark Across Domains and Tasks cites this paper.

Quo Vadis, Visual In-Context Learning? A Unified Benchmark Across Domains and Tasks VL-ICL Bench: The Devil in the Details of Multimodal In-Context Learning

Reference 111

Resolution
malformed identifier
arxiv_id, observed 2026-07-03T05:07:39.146050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T13:24:48.138643Z digest=sha256:0cf30fa0cf543fae97f4e82128e0af8e223a8b7781533ca44321354c2dbf1ad1