Pith. sign in

Paper Citation Record · LEDGER

GRIP: Feedback-Guided Prompt Retrieval for Large Multimodal Models

As of 18 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 0 inbound Pith citation observations for arXiv:2606.12744.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.12744 v1

Coverage vector

measured 14 of 14 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-27T09:34:20.560870Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

14 of 14 outbound references displayed

  • verified exact9
  • verified fuzzy0
  • unresolved3
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0a5bdd23-4ccd-4312-aa10-fa9761ecb0dc · outbound

This paper cites Qwen2.5-VL Technical Report.

GRIP: Feedback-Guided Prompt Retrieval for Large Multimodal Models Qwen2.5-VL Technical Report

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T11:28:04.242745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-27T09:34:20.560870Z digest=sha256:690caadf288ec318d339685345a3549676646a02484d8f50ac16aab38078a23e

Observation 5ee41487-1255-490b-96aa-edb7c7393298 · outbound

This paper cites Can multimodal large language models truly perform multimodal in-context learning? In 2025 IEEE/CVF Winter Conference on Applications of Computer Vision (WACV), pp.

GRIP: Feedback-Guided Prompt Retrieval for Large Multimodal Models Can multimodal large language models truly perform multimodal in-context learning? In 2025 IEEE/CVF Winter Conference on Applications of Computer Vision (WACV), pp

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-27T09:34:20.560870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T09:34:20.560870Z digest=sha256:2ff46f787184a442889e0425efaea9366b593852e5cc2d2b0b0631da5c2da950

Observation 7a679b94-225d-4ed4-bbb4-77322b9887a4 · outbound

This paper cites Microsoft COCO Captions: Data Collection and Evaluation Server.

GRIP: Feedback-Guided Prompt Retrieval for Large Multimodal Models Microsoft COCO Captions: Data Collection and Evaluation Server

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-07-03T11:28:04.248622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-27T09:34:20.560870Z digest=sha256:625689c9963f71ba847c82ec50f51fa0e668b1500ec5f613a09aadac212e4cdf

Observation 83f0f176-fd2f-43e0-a44e-864edf16aafd · outbound

This paper cites The Llama 3 Herd of Models.

GRIP: Feedback-Guided Prompt Retrieval for Large Multimodal Models The Llama 3 Herd of Models

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-07-03T11:28:04.237290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-27T09:34:20.560870Z digest=sha256:1f4f41978a755966c9ec7e0c84c624f3a618359a7ccddf49bb9029bf100b4ab1

Observation 802d269e-ac42-4592-a967-8c7e58bcd151 · outbound

This paper cites Ilharco, M.

GRIP: Feedback-Guided Prompt Retrieval for Large Multimodal Models Ilharco, M

Reference 5

Resolution
metadata mismatch
doi, observed 2026-06-27T09:40:47.154456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-27T09:34:20.560870Z digest=sha256:a6fd2693e2f9a0f9460e7c634d51c05c5278241d638de8ec4d1ad35a1a517a61

Observation 868bcc42-3554-4402-8961-38e86c4f2eab · outbound

This paper cites Syntriever: How to Train Your Retriever with Synthetic Data from LLMs.

GRIP: Feedback-Guided Prompt Retrieval for Large Multimodal Models Syntriever: How to Train Your Retriever with Synthetic Data from LLMs

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-03T11:28:04.239902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-27T09:34:20.560870Z digest=sha256:02f7fa3bc63cffae89cf34715550d4690568698d3c59cb9d33e2aab34b59698d

Observation dc7c6749-ea1f-494a-b27d-2d9ee8fa0c43 · outbound

This paper cites SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension.

GRIP: Feedback-Guided Prompt Retrieval for Large Multimodal Models SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-07-03T11:28:04.234876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-27T09:34:20.560870Z digest=sha256:dcfb0cf3b37e6775fc23ee4ff3331d3fe2727ec4bfd033011dfc5dce6ec08296

Observation e637aabf-a534-4df9-b8f6-6443c32d8321 · outbound

This paper cites Dr.ICL: Demonstration-Retrieved In-context Learning.

GRIP: Feedback-Guided Prompt Retrieval for Large Multimodal Models Dr.ICL: Demonstration-Retrieved In-context Learning

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-03T11:28:04.254399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-27T09:34:20.560870Z digest=sha256:585e4ed48c61ff20fdaef91d9aa36c8b3759e28d4e4f8a4fc1cb40b117263c8b

Observation a5e59d3d-858c-49ba-9e33-dc4f44a7cdc1 · outbound

This paper cites an unresolved cited work.

GRIP: Feedback-Guided Prompt Retrieval for Large Multimodal Models Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-27T09:34:20.560870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T09:34:20.560870Z digest=sha256:2611879fff52267017aa2694e919dbfc09442aba59a692ddde350ac6ff3a5937

Observation 29cabaae-74e9-4750-8c8b-173e3350f50c · outbound

This paper cites Learning to retrieve prompts for in-context learning.

GRIP: Feedback-Guided Prompt Retrieval for Large Multimodal Models Learning to retrieve prompts for in-context learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-27T09:34:20.560870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T09:34:20.560870Z digest=sha256:e26b0293ddb4f17f5c185eb026a4f39586e2490c2fc0937b886d1cb0abe2320f

Observation 9ffb31f5-d8df-4999-9fb8-df3e97e286f1 · outbound

This paper cites URL https://aclanthology.org/2022.naacl-main.191.

GRIP: Feedback-Guided Prompt Retrieval for Large Multimodal Models URL https://aclanthology.org/2022.naacl-main.191

Reference 11

Resolution
verified exact
doi, observed 2026-06-27T09:40:47.156390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-27T09:34:20.560870Z digest=sha256:5bd94bffaabf0bbb4325ef23dd7e4655f5ba1fb8b7a6320a62fe4dc2255280cd

Observation 2f51f9f6-309d-46f3-81ee-2473e4e755fd · outbound

This paper cites Learning to Retrieve In-Context Examples for Large Language Models.

GRIP: Feedback-Guided Prompt Retrieval for Large Multimodal Models Learning to Retrieve In-Context Examples for Large Language Models

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-03T11:28:04.240181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-27T09:34:20.560870Z digest=sha256:df15247f41912ee57cb45c48011ccb453029bfca484e94cf61153a50ccdd3ab7

Observation 1190a19c-0add-4172-b484-ad9a0e91232a · outbound

This paper cites Demonstration Selection for In-Context Learning via Reinforcement Learning.

GRIP: Feedback-Guided Prompt Retrieval for Large Multimodal Models Demonstration Selection for In-Context Learning via Reinforcement Learning

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-03T11:28:04.248446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-27T09:34:20.560870Z digest=sha256:68780138024aebb014de601d6f56437d0a98a9bcceffc87fa8966f1418e93eed

Observation 8d9b819c-2ffe-4f6f-a328-8344868c714b · outbound

This paper cites MMICL: Empowering Vision-language Model with Multi-Modal In-Context Learning.

GRIP: Feedback-Guided Prompt Retrieval for Large Multimodal Models MMICL: Empowering Vision-language Model with Multi-Modal In-Context Learning

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-03T11:28:04.251441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-27T09:34:20.560870Z digest=sha256:cf592d03541c3e7c4cf3c9748f82f7029d430d048e6c80d2d9a5dd40bdfad603

Pith citing papers

No inbound Pith citation observations are available.