Pith. sign in

Paper Citation Record · LEDGER

Multimodal Task Vectors Enable Many-Shot Multimodal In-Context Learning

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2406.15334.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.15334 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T15:05:19.256230Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T05:34:40.674297Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation da2197ba-17e3-46ba-8e97-32736727f1de · inbound

Teaching VLMs to Localize Specific Objects from In-context Examples cites this paper.

Teaching VLMs to Localize Specific Objects from In-context Examples Multimodal Task Vectors Enable Many-Shot Multimodal In-Context Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T16:39:40.536776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:39:40.536776Z digest=sha256:6837927cd5fc838fa9e9cb3628ff3464ca3f41a1b43ad3c7a191753a84fcda83

Observation 5be0f649-ec2a-418a-abd1-682a70217bc8 · inbound

Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features cites this paper.

Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features Multimodal Task Vectors Enable Many-Shot Multimodal In-Context Learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T10:24:06.611337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:24:06.611337Z digest=sha256:d371bd56e4206d61318b0d96b70b152fdb78f765142066c09704da1570979cd6

Observation e06625a7-57d5-468d-8939-a0636341a180 · inbound

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation cites this paper.

Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation Multimodal Task Vectors Enable Many-Shot Multimodal In-Context Learning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T04:50:49.099961Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:50:49.099961Z digest=sha256:612544cc658bb598d36f3a2724813b64cbcff527e6489d41821452f1be6e1348

Observation 277abd81-4632-4e7f-956a-54db7f93624a · inbound

MAPLE: Many-Shot Adaptive Pseudo-Labeling for In-Context Learning cites this paper.

MAPLE: Many-Shot Adaptive Pseudo-Labeling for In-Context Learning Multimodal Task Vectors Enable Many-Shot Multimodal In-Context Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:43.013943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:43.013943Z digest=sha256:80cedeb82d8ec7e5fb3192c84ab370f0d64b95ce0e87e54f3251b540f157a874

Observation ae254501-6b43-4905-9190-77d8e08dc682 · inbound

Less is More Tokens: Efficient Math Reasoning via Difficulty-Aware Chain-of-Thought Distillation cites this paper.

Less is More Tokens: Efficient Math Reasoning via Difficulty-Aware Chain-of-Thought Distillation Multimodal Task Vectors Enable Many-Shot Multimodal In-Context Learning

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-08-05T05:34:40.679690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T05:34:40.305028Z digest=sha256:fe835aa192b39393e3e51c3594d085d5d6f2cb6a650e2faea52ea7b8a44e257f

Observation 76b08a8e-690f-4024-9cb9-390ca9ab9bd3 · inbound

In-Context Collapse in Vision-Language Models and How to Mitigate it? cites this paper.

In-Context Collapse in Vision-Language Models and How to Mitigate it? Multimodal Task Vectors Enable Many-Shot Multimodal In-Context Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T15:05:19.256230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:05:19.256230Z digest=sha256:21aac2f1dcfdfc17705d747eccdf774195c21b48aceb3c6c5ddea4435162aa9f

Observation 6f03b30b-636b-4216-abed-a42feae33873 · inbound

When Is a Task Vector Enough? An Empirical Theory of Implicit Multimodal ICL cites this paper.

When Is a Task Vector Enough? An Empirical Theory of Implicit Multimodal ICL Multimodal Task Vectors Enable Many-Shot Multimodal In-Context Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-14T12:25:20.439241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T12:25:20.439241Z digest=sha256:3bc43d58ec08a6d1ee985736a9bbe53d38d52d721484a90ecfaa483101bfbfba