Pith. sign in

Paper Citation Record · LEDGER

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning

As of 12 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 0 inbound Pith citation observations for arXiv:2412.19289.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.19289 v3

Coverage vector

measured 50 of 50 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T00:47:52.532503Z

measured 50 of 50 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

50 of 50 outbound references displayed

  • verified exact6
  • verified fuzzy3
  • unresolved41
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 226f241b-a6e8-4021-897a-0b4095f4c58e · outbound

This paper cites , " * write output.state after.block = add.period write newline.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning , " * write output.state after.block = add.period write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.337518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.337518Z digest=sha256:c9e8777e8a52c68523e5394e889bacd9591833c64e01ea9d0dd4b200467f1139

Observation 0dae187d-1325-49b4-b159-77c04d02c43b · outbound

This paper cites write newline.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.342475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.342475Z digest=sha256:23eadcdd4fa886708ec9c1e671b7ac1b93430973c22b4da6116dfce88df49bbf

Observation 7a818e42-d0de-4553-b90a-7c042df8278c · outbound

This paper cites an unresolved cited work.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.347000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.347000Z digest=sha256:eff4cbbf31e98d57a4bb09577ee5e8d12d3c83a5147662dcf64d9c23f277a7a2

Observation cff45548-ce89-437c-9505-aa5874459687 · outbound

This paper cites SPICE: Semantic Propositional Image Caption Evaluation.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning SPICE: Semantic Propositional Image Caption Evaluation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.350966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.350966Z digest=sha256:94860a11b799b3f4495e8e34a1657da2bf9e6407efa97d15c71b93e4b0868d96

Observation d878a7b6-5845-47db-b99b-6ce4702cdc54 · outbound

This paper cites Exploring Visual Prompts for Adapting Large-Scale Models.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Exploring Visual Prompts for Adapting Large-Scale Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.355230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.355230Z digest=sha256:bfacaed86e185c3aa77a3a25aa1e14dbc3d0bea86bb5ec58e13156a78ae9c11f

Observation 28adb247-1126-4345-9a50-40070c3b7b43 · outbound

This paper cites CaMEL: Mean Teacher Learning for Image Captioning.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning CaMEL: Mean Teacher Learning for Image Captioning

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-11T00:47:52.932804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.359461Z digest=sha256:6a367ed68b87a5f7c49feed60670f79764f8622f57dd5b07202ffd7e82af7c50

Observation a8898f81-acc2-45bb-a3c5-fa492e59170d · outbound

This paper cites PaLI: A Jointly-Scaled Multilingual Language-Image Model.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning PaLI: A Jointly-Scaled Multilingual Language-Image Model

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.363799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.363799Z digest=sha256:9797ca5227851c1815531039a2032cb004bca766a1f056c02d1ce276e81c2ddd

Observation e671ddb3-7ad5-420c-9562-2c709005c4c4 · outbound

This paper cites E.; Stoica, I.; and Xing, E.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning E.; Stoica, I.; and Xing, E

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.367994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.367994Z digest=sha256:7430a647fd2c94c20c572f62354a8bb3627d1bb6213d4d1ebafd5a6adc482ec0

Observation 8aa8e136-3c00-4fcf-9234-2aebaf9832a4 · outbound

This paper cites J.; and Lavie, A.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning J.; and Lavie, A

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:47:53.053822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.372139Z digest=sha256:ec02b843c88562847980552bdc35401464ed939c959da12c98d5080e2c55993b

Observation 023456da-f5ef-4eab-86e7-e8132b51b89e · outbound

This paper cites Transferable Decoding with Visual Entities for Zero-Shot Image Captioning.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Transferable Decoding with Visual Entities for Zero-Shot Image Captioning

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-11T00:47:52.907331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.377156Z digest=sha256:e600979ca1d2ba8b3401a70db92b2685c445f63d4c8e8074f1f3a58463e84fd3

Observation b9500f88-5c56-4d37-8c59-36dc115bd496 · outbound

This paper cites Making Pre-trained Language Models Better Few-shot Learners.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Making Pre-trained Language Models Better Few-shot Learners

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.381700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.381700Z digest=sha256:f31c87cada4ea6df41e79856eaa6fe9aa5a4a7f8aa112e8b6472580a31f80a30

Observation 49d42f1d-8228-409b-9826-da1b7e270051 · outbound

This paper cites Language-only Efficient Training of Zero-shot Composed Image Retrieval.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Language-only Efficient Training of Zero-shot Composed Image Retrieval

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.386106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.386106Z digest=sha256:caee3f0e390127a9e3f9bd65bdb70040b9f78f8b2be8dac918377c820dbd7f0a

Observation f8c9a401-df39-491f-bc5d-a55c981f36aa · outbound

This paper cites Scaling Up Vision-Language Pre-training for Image Captioning.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Scaling Up Vision-Language Pre-training for Image Captioning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.390168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.390168Z digest=sha256:121d6893ca24e5b282d6fa7e43bf1cad81dc078530049341d0eca53fb889174c

Observation 0d0790bb-0b6f-4354-a632-5587ddb2eefb · outbound

This paper cites REVEAL: Retrieval-Augmented Visual-Language Pre-Training with Multi-Source Multimodal Knowledge Memory.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning REVEAL: Retrieval-Augmented Visual-Language Pre-Training with Multi-Source Multimodal Knowledge Memory

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.394111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.394111Z digest=sha256:0f889d0b49fb04e68029f6a2459ea919020f51f0157bc806b3538ee9857713cc

Observation cdcd0b62-1ac9-42fd-beca-718e10dd8200 · outbound

This paper cites Scaling Up Visual and Vision-Language Representation Learning With Noisy Text Supervision.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Scaling Up Visual and Vision-Language Representation Learning With Noisy Text Supervision

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.398127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.398127Z digest=sha256:21af03d430afc9a26514a24c6d1bf2d1e00e727716d38fae72421062bc7d7a76

Observation 68cfa15d-5001-412c-8014-6bba038e47b0 · outbound

This paper cites Visual Prompt Tuning.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Visual Prompt Tuning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.402265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.402265Z digest=sha256:0a3ddcd153b7c6cf3f2ab8554b491a7a46e98b38b227486bb468688e9ac5bfca

Observation f6844831-ca97-454d-923c-9795e2a6a256 · outbound

This paper cites Billion-scale similarity search with GPUs.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Billion-scale similarity search with GPUs

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.406187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.406187Z digest=sha256:4253b56e88e341b3843fefcb9b79c851d2345211e23b434008b3de6d8ddb1aa8

Observation fbf6f1ea-74aa-4474-9dad-76ac4b4f5126 · outbound

This paper cites Deep Visual-Semantic Alignments for Generating Image Descriptions.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Deep Visual-Semantic Alignments for Generating Image Descriptions

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.410230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.410230Z digest=sha256:339401575e5355915cd323115cc9d8e167c9bcf62bd80574399ba38a528cf2da

Observation 2e8b41ff-a450-4e9c-af0d-c5631d105fa0 · outbound

This paper cites Auto-Encoding Variational Bayes.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Auto-Encoding Variational Bayes

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.414573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.414573Z digest=sha256:21f16889250beff8ce7322d8b36704c1d3fb9d77dfa63e8f18ad5d28719f68eb

Observation 778da2bf-ee88-4ffc-a9d5-3a556e13dc02 · outbound

This paper cites The Power of Scale for Parameter-Efficient Prompt Tuning.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning The Power of Scale for Parameter-Efficient Prompt Tuning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.418671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.418671Z digest=sha256:1c8b567bc14523349a1335a6358d9ea6454330c857d9e2532d8eee3d316030bb

Observation 59ca8558-0ed1-43ae-bf88-aa1a3ad6f4a5 · outbound

This paper cites BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.422712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.422712Z digest=sha256:8f90e501c3fc864c404570a793d29e65c9a950aaa089f3dc86e0904b736daedc

Observation 230e9ed6-070d-49f1-8d6d-6f975a8b0aa3 · outbound

This paper cites EVCap: Retrieval-Augmented Image Captioning with External Visual-Name Memory for Open-World Comprehension.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning EVCap: Retrieval-Augmented Image Captioning with External Visual-Name Memory for Open-World Comprehension

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.426766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.426766Z digest=sha256:e8937f980fd4b0a6e505199c9561312b3b1fc8edeb8044f28a7789ccdfdbac76

Observation a9f4eac5-2562-4f10-b109-2dda3ee40556 · outbound

This paper cites an unresolved cited work.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-11T00:47:53.041629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.430620Z digest=sha256:e47d65dc3d8883e0eeda1fd697cc186bbc034b2784f4629600efafc78c006622

Observation fd13970e-17f6-4ac2-becd-daad3b68fff5 · outbound

This paper cites Oscar: Object-Semantics Aligned Pre-training for Vision-Language Tasks.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Oscar: Object-Semantics Aligned Pre-training for Vision-Language Tasks

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.434115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.434115Z digest=sha256:7188c80e5ec03f509288d64a58c591c2993139d09676f588639adfe2e59e62d8

Observation 0395bd76-e967-4b2f-b5fd-f7a6b66007e1 · outbound

This paper cites Prefix-Tuning: Optimizing Continuous Prompts for Generation.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Prefix-Tuning: Optimizing Continuous Prompts for Generation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.437835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.437835Z digest=sha256:f1c1d4a1a2f8013ddb24dcd53fa85296fa2ff918080a4f36da96b9328ecfb2b1

Observation 434cd9d7-14b2-4a18-83bf-708115380476 · outbound

This paper cites Microsoft COCO: Common Objects in Context.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Microsoft COCO: Common Objects in Context

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.441953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.441953Z digest=sha256:45a1f12d9f04c8d972a93fb863027fe93a45d5b15e99ec61595934e2eca2cca0

Observation 2f665086-7745-455f-ac61-ae254c66e6c2 · outbound

This paper cites Few-shot Learning with Multilingual Language Models.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Few-shot Learning with Multilingual Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.446185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.446185Z digest=sha256:f1ab62ca37c31bdbb69727ca07667cf07cff92928de5c335e84747bf24678b0f

Observation 93255380-34c1-443d-88a2-11ec4e3a7bee · outbound

This paper cites Visual Instruction Tuning.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Visual Instruction Tuning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.450290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.450290Z digest=sha256:53b04dd02e68d9076c3ac8b710ee458b6dabb6b2db4a2c6627861ca2fa239282

Observation 226c086b-add0-4e3f-839c-d969a4be0fc6 · outbound

This paper cites I-Tuning: Tuning Frozen Language Models with Image for Lightweight Image Captioning.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning I-Tuning: Tuning Frozen Language Models with Image for Lightweight Image Captioning

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-11T00:47:52.727828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.454181Z digest=sha256:ddc47c3b60e6f3a48458225e5a5379125af2d39e2ac3a8d13e75ea86fe09ee7c

Observation 230a2993-3479-4c04-8728-416ccb573566 · outbound

This paper cites MAPL: Parameter-Efficient Adaptation of Unimodal Pre-Trained Models for Vision-Language Few-Shot Prompting.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning MAPL: Parameter-Efficient Adaptation of Unimodal Pre-Trained Models for Vision-Language Few-Shot Prompting

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.458002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.458002Z digest=sha256:03d8c096bd8d3c48e052d3647af8690a0245608adc479c74a98a84bb3b9c7dfb

Observation 6f40e8af-c06c-4d64-87db-5b4eb3febfb7 · outbound

This paper cites ClipCap: CLIP Prefix for Image Captioning.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning ClipCap: CLIP Prefix for Image Captioning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.461697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.461697Z digest=sha256:948633a8057aa88f5cdfc7f4ab202bd4187ca22a78baa24fb0f6cdd88748c377

Observation e739e8bb-f0fe-4bc1-b724-fd797411b8b6 · outbound

This paper cites an unresolved cited work.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-11T00:47:53.030299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.465468Z digest=sha256:efc03bfa6fefd32246e035bcc32c0a7ab6b546cf25596e50bd0c9bdd46184d84

Observation e56161bb-1edb-48aa-b691-edaf369245f4 · outbound

This paper cites an unresolved cited work.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-11T00:47:53.018872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.468715Z digest=sha256:77a64b9351100a2a177dfb8ab32d980d1069cbdb0ca97d71eb1b30f193cb32cf

Observation 5841e810-0639-45f3-bec8-1f34f573c01e · outbound

This paper cites Flickr30k Entities: Collecting Region-to-Phrase Correspondences for Richer Image-to-Sentence Models.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Flickr30k Entities: Collecting Region-to-Phrase Correspondences for Richer Image-to-Sentence Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.472320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.472320Z digest=sha256:e648886642ba8ee1908fe1222c79be6fc41d3bbb8d7e96ae75be523ed8995a41

Observation afa7388d-2396-49cc-b606-d7566f664077 · outbound

This paper cites Learning Transferable Visual Models From Natural Language Supervision.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Learning Transferable Visual Models From Natural Language Supervision

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.476285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.476285Z digest=sha256:9ca553bbde3c8c82bd756d70ca3a298d480a03cd76e1a236fa4f15cfcd90df25

Observation 982a31ea-aaee-44bb-8600-1f34d2273480 · outbound

This paper cites an unresolved cited work.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-11T00:47:53.007004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.480296Z digest=sha256:5fdc55d24c80a09dcacfa5cb20017e67390e9c0488c95dc7422e90bcd7b4dc23

Observation 3ea24812-6756-427b-9c9a-6a89fb70bec9 · outbound

This paper cites LMCap: Few-shot Multilingual Image Captioning by Retrieval Augmented Language Model Prompting.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning LMCap: Few-shot Multilingual Image Captioning by Retrieval Augmented Language Model Prompting

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.483794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.483794Z digest=sha256:fa6e65d902bb631a454c6b42f92db63346c216baca272c34591512510c27566b

Observation a5fbff30-4ddb-4018-9f30-1e344aa2a986 · outbound

This paper cites SmallCap: Lightweight Image Captioning Prompted with Retrieval Augmentation.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning SmallCap: Lightweight Image Captioning Prompted with Retrieval Augmentation

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-08-11T00:47:52.660913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.487763Z digest=sha256:3c3f04fba213b7686cd6bbd8644615400b6a7c419a6d88d954ddbe89cdc57b32

Observation 658f9a71-d444-4e33-b749-0caa9763f830 · outbound

This paper cites P.; Elliott, D.; and Martins, B.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning P.; Elliott, D.; and Martins, B

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:47:52.996295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.491626Z digest=sha256:ff2cb97539b2f223e129f44554ab4df6d2ea48283795277a71172caeeb185d2a

Observation 54b2ca25-4afb-4bff-91d4-045daf8eb7ad · outbound

This paper cites EVA-CLIP: Improved Training Techniques for CLIP at Scale.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning EVA-CLIP: Improved Training Techniques for CLIP at Scale

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.495100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.495100Z digest=sha256:5d45d326b967c4471b1db3c90b031c30d6be3112cb02be77969bd74f2d0c81ec

Observation 99d81053-2031-459e-a814-9b90d8503bfb · outbound

This paper cites L.; and Parikh, D.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning L.; and Parikh, D

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:47:52.985218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.498807Z digest=sha256:80a365d5e421a1d0509b56e0ce2191c8cc713bf926138a0d5c4e0a6b0a232ce1

Observation 049e45b1-06fd-41c7-9935-b80f1758c45c · outbound

This paper cites an unresolved cited work.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-11T00:47:52.974325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.502607Z digest=sha256:83909fd65213b29bb53a9385ad6ffb40e45f5b61d046a42e1a8c5a2d40cfb87a

Observation cda9c693-6e3a-4b49-84b7-35ea6863a103 · outbound

This paper cites an unresolved cited work.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-11T00:47:52.963583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.506044Z digest=sha256:ca85911ffe9aefb716508046602f65c77de4a1123ea80f1a9739c51e7611dc7f

Observation e4514247-2ebd-4ed2-bc74-e7053b154cfd · outbound

This paper cites CogVLM: Visual Expert for Pretrained Language Models.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning CogVLM: Visual Expert for Pretrained Language Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.509718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.509718Z digest=sha256:a760f19b98ec5396244222ceb391397eb6420827631efec7aec25ec210ca92e0

Observation d03c0505-7d79-4c8e-a6d6-0fb2112b0a24 · outbound

This paper cites SimVLM: Simple Visual Language Model Pretraining with Weak Supervision.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning SimVLM: Simple Visual Language Model Pretraining with Weak Supervision

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.513342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.513342Z digest=sha256:9ae12f160e1fbc1cd1997c2bc9bd230da5b0dd662693fc5472c3a06a6e1bf202

Observation 9268e79b-dd81-410f-803d-edb3aa4c880f · outbound

This paper cites DualPrompt: Complementary Prompting for Rehearsal-free Continual Learning.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning DualPrompt: Complementary Prompting for Rehearsal-free Continual Learning

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-08-11T00:47:52.615222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.517018Z digest=sha256:1d1396ee72e7b541228312a807350d60237c3666a623cfec16a6fff45e8ca178

Observation 91256ac0-9dd6-41e1-9938-660fafd9122c · outbound

This paper cites Re-ViLM: Retrieval-Augmented Visual Language Model for Zero and Few-Shot Image Captioning.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Re-ViLM: Retrieval-Augmented Visual Language Model for Zero and Few-Shot Image Captioning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.520772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.520772Z digest=sha256:b65004b3523deed719dd28d786b2fd65993e0f89651a1fed54729389168c17f1

Observation c8eb9d5e-d19c-4e4b-bd0c-15a77cdf443e · outbound

This paper cites MeaCap: Memory-Augmented Zero-shot Image Captioning.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning MeaCap: Memory-Augmented Zero-shot Image Captioning

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-11T00:47:52.590167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.524701Z digest=sha256:7510a60a0e5dc323c20488537c389a97149b1aa11f0932b30ecb9c5a692a8c75

Observation a85aaa13-e464-4ee1-90cf-551ff68e6e21 · outbound

This paper cites OPT: Open Pre-trained Transformer Language Models.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning OPT: Open Pre-trained Transformer Language Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.528543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.528543Z digest=sha256:17d2b07960eb04ef2553e826a7e6f97a19225db391425b536fa29f3eb4121af7

Observation a0fcaac9-22ea-4392-b6f1-fb1b67373a17 · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.532503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.532503Z digest=sha256:bfa7cdaa3134f7de777ee0074030d9e33f924c75e1d9911a485908100a197fbc

Pith citing papers

No inbound Pith citation observations are available.