Pith. sign in

Paper Citation Record · LEDGER

Many-Shot In-Context Learning in Multimodal Foundation Models

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2405.09798.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.09798 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:45:25.419580Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-25T04:25:19.781518Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 8689c888-24a6-4676-8f49-0fd6649775e0 · inbound

MuirBench: A Comprehensive Benchmark for Robust Multi-image Understanding cites this paper.

MuirBench: A Comprehensive Benchmark for Robust Multi-image Understanding Many-Shot In-Context Learning in Multimodal Foundation Models

Reference 26

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T01:09:30.475892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T01:09:30.360275Z digest=sha256:e35b0a1689bd68833417f45f5e307f670c4a7264604fc74c5da375ce847dbebe

Observation c324eeda-50c0-4d0a-a6d4-63947787e77b · inbound

LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models cites this paper.

LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models Many-Shot In-Context Learning in Multimodal Foundation Models

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T06:01:53.841270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T06:01:53.730356Z digest=sha256:25b575bb9e84e6de5ae57701ef50487d39637ce382cf2f2498e28e7eafcafd2e

Observation 41d89f27-34a1-4bbc-9c4f-63c89b39b4f0 · inbound

Selecting Demonstrations for Many-Shot In-Context Learning via Gradient Matching cites this paper.

Selecting Demonstrations for Many-Shot In-Context Learning via Gradient Matching Many-Shot In-Context Learning in Multimodal Foundation Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T10:45:25.419580Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:45:25.419580Z digest=sha256:414390266e4fb1763cae29d8f0dd3f55d40cdaccdf3cd38d19aa3ab7483189b9

Observation 973035ca-77dd-43cd-8391-1e740949f05d · inbound

How Well Does GPT-4o Understand Vision? Evaluating Multimodal Foundation Models on Standard Computer Vision Tasks cites this paper.

How Well Does GPT-4o Understand Vision? Evaluating Multimodal Foundation Models on Standard Computer Vision Tasks Many-Shot In-Context Learning in Multimodal Foundation Models

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T05:57:08.083275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T05:55:09.188048Z digest=sha256:5f16f6410985e703bbda56e4097b4fd7f262b96c0cb2c5a4b370cb15eb77bc06

Observation 81541ab2-1b20-4b79-8230-ef2d6bf3d74c · inbound

True Multimodal In-Context Learning Needs Attention to the Visual Context cites this paper.

True Multimodal In-Context Learning Needs Attention to the Visual Context Many-Shot In-Context Learning in Multimodal Foundation Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T15:29:34.320527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:29:34.320527Z digest=sha256:361091f2101b4502bf4b5e6c587f8c2eca95d46259706e85cd2841fdf22177f4

Observation f949f0e5-46a8-445c-be4f-4a7c575cc8a6 · inbound

True Multimodal In-Context Learning Needs Attention to the Visual Context cites this paper.

True Multimodal In-Context Learning Needs Attention to the Visual Context Many-Shot In-Context Learning in Multimodal Foundation Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T15:29:34.377976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:29:34.377976Z digest=sha256:1e7b8a2e85c187516dda76e490a83ca117e6f0138da9567c59bfce85e86a8886

Observation 95eb9f2f-dd60-4303-995a-fb97ec87b6d8 · inbound

Towards Compute-Optimal Many-Shot In-Context Learning cites this paper.

Towards Compute-Optimal Many-Shot In-Context Learning Many-Shot In-Context Learning in Multimodal Foundation Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T15:20:38.661092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:20:38.661092Z digest=sha256:90516aa2b933f859aebd19ff3c532563790e61dc970a95e850e9dfb203f7c0ee

Observation a01c224e-b043-443d-9d2f-17963259643c · inbound

Visual Species Recognition with Large Multimodal Models as Post-Hoc Correctors cites this paper.

Visual Species Recognition with Large Multimodal Models as Post-Hoc Correctors Many-Shot In-Context Learning in Multimodal Foundation Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-03T17:20:52.476377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T17:20:52.476377Z digest=sha256:6cb0593e60e3504c35b6bac3f18069400c7f5db7fefa5c462ab77232bf612f81

Observation c6c0965f-1f6e-458e-9f0e-fd5d4f94c0b7 · inbound

Personal Visual Context Learning in Large Multimodal Models cites this paper.

Personal Visual Context Learning in Large Multimodal Models Many-Shot In-Context Learning in Multimodal Foundation Models

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T07:06:37.577186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T03:42:15.402131Z digest=sha256:50be2efa5d0aabe9b9ea915447844febd9ab5b8508f49e91c6e206c3e09091c2

Observation c5992951-048a-4136-8f99-84469c93f655 · inbound

When Youth Enter the Algorithmic Wild: Discovering and Understanding Potentially Harmful Teen Videos on Douyin and Kwai cites this paper.

When Youth Enter the Algorithmic Wild: Discovering and Understanding Potentially Harmful Teen Videos on Douyin and Kwai Many-Shot In-Context Learning in Multimodal Foundation Models

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:25:19.783887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-25T04:20:27.456558Z digest=sha256:cb4ea4b1cb5537fdc7814e6326567b33c1cc62f24eedf79d917c5b2da309283e

Observation 610dae7e-06c2-4739-9085-c936289a3846 · inbound

In-Context Learning for Wound Classification with Small Multimodal Language Models cites this paper.

In-Context Learning for Wound Classification with Small Multimodal Language Models Many-Shot In-Context Learning in Multimodal Foundation Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-01T14:15:59.840877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:15:59.840877Z digest=sha256:a59868ba84b52d74618f3f050122a9fec0ef19131f1ee358bd83794e9b0bd31c

Observation 081ba260-c89f-4407-88bd-80a52c4d203e · inbound

Inference-Time Agentic Decision Rules Beat Longer Evolving Search for Multi-Image Medical Reasoning cites this paper.

Inference-Time Agentic Decision Rules Beat Longer Evolving Search for Multi-Image Medical Reasoning Many-Shot In-Context Learning in Multimodal Foundation Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T05:38:56.492179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T05:38:56.492179Z digest=sha256:4f1e05f9d727455955f701b42ca9c629d4bc37c1c963db8f2a3a34157810800b