Pith. sign in

Paper Citation Record · LEDGER

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning

As of 21 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 0 inbound Pith citation observations for arXiv:2412.19289.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.19289 v3

Coverage vector

measured 50 of 50 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T00:47:52.532503Z

measured 50 of 50 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

50 of 50 outbound references displayed

  • verified exact6
  • verified fuzzy3
  • unresolved41
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 226f241b-a6e8-4021-897a-0b4095f4c58e · outbound

This paper cites , " * write output.state after.block = add.period write newline.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning , " * write output.state after.block = add.period write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.337518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.337518Z digest=sha256:ce7857da76dda6c11934088d69790a9bb19b0038c0f3aceba6cf4d706af47898

Observation 0dae187d-1325-49b4-b159-77c04d02c43b · outbound

This paper cites write newline.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.342475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.342475Z digest=sha256:d1f093dfe1e41e202deab64ccb44495c34c7b23c0bbc724e2f90c1ba84eca1e4

Observation 7a818e42-d0de-4553-b90a-7c042df8278c · outbound

This paper cites an unresolved cited work.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.347000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.347000Z digest=sha256:af65714a9dc4b3a517cd2cde2de46a7581abb409dd86fe130ad9283b16510615

Observation cff45548-ce89-437c-9505-aa5874459687 · outbound

This paper cites SPICE: Semantic Propositional Image Caption Evaluation.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning SPICE: Semantic Propositional Image Caption Evaluation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.350966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.350966Z digest=sha256:07b930ef4c6a0cb389a4a1104d9e02e31e8c66fa8437dc1aedfa4c6029101d24

Observation d878a7b6-5845-47db-b99b-6ce4702cdc54 · outbound

This paper cites Exploring Visual Prompts for Adapting Large-Scale Models.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Exploring Visual Prompts for Adapting Large-Scale Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.355230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.355230Z digest=sha256:4d015fd768efee00ba0cacb1d4b67afb8369997824fdb01568a5085655c99b03

Observation 28adb247-1126-4345-9a50-40070c3b7b43 · outbound

This paper cites CaMEL: Mean Teacher Learning for Image Captioning.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning CaMEL: Mean Teacher Learning for Image Captioning

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-11T00:47:52.932804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.359461Z digest=sha256:3e597f510653cf261322d12396d49df5925b86fcc6857723601c56d7ae3db22b

Observation a8898f81-acc2-45bb-a3c5-fa492e59170d · outbound

This paper cites PaLI: A Jointly-Scaled Multilingual Language-Image Model.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning PaLI: A Jointly-Scaled Multilingual Language-Image Model

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.363799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.363799Z digest=sha256:9112c40fbde0a12e6b31dac25b68c893c505ab5a090c2c043cca1f6d53581be5

Observation e671ddb3-7ad5-420c-9562-2c709005c4c4 · outbound

This paper cites E.; Stoica, I.; and Xing, E.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning E.; Stoica, I.; and Xing, E

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.367994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.367994Z digest=sha256:9c7337f643abc8326df21a56a19eb790f27e0605fe2c12e27dc5ea23bbd27dd0

Observation 8aa8e136-3c00-4fcf-9234-2aebaf9832a4 · outbound

This paper cites J.; and Lavie, A.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning J.; and Lavie, A

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:47:53.053822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.372139Z digest=sha256:9f2eb8aa5fc061bf33a87f57b44f540ba114a7818092b99e846a6fdc75eebbc6

Observation 023456da-f5ef-4eab-86e7-e8132b51b89e · outbound

This paper cites Transferable Decoding with Visual Entities for Zero-Shot Image Captioning.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Transferable Decoding with Visual Entities for Zero-Shot Image Captioning

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-11T00:47:52.907331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.377156Z digest=sha256:2197000a851cee6555d088ec7abcf9114e35376613ea21b7de2811b601219ba6

Observation b9500f88-5c56-4d37-8c59-36dc115bd496 · outbound

This paper cites Making Pre-trained Language Models Better Few-shot Learners.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Making Pre-trained Language Models Better Few-shot Learners

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.381700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.381700Z digest=sha256:4f602e7c599816c79c0b6c2df3e753b3d54981d2bf1736a1032c69cdbc7c7df3

Observation 49d42f1d-8228-409b-9826-da1b7e270051 · outbound

This paper cites Language-only Efficient Training of Zero-shot Composed Image Retrieval.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Language-only Efficient Training of Zero-shot Composed Image Retrieval

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.386106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.386106Z digest=sha256:b1a68dc3827bea06be2b8506295c24e7f7c0bee8c25d6007121122e915bbb319

Observation f8c9a401-df39-491f-bc5d-a55c981f36aa · outbound

This paper cites Scaling Up Vision-Language Pre-training for Image Captioning.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Scaling Up Vision-Language Pre-training for Image Captioning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.390168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.390168Z digest=sha256:2797221d6970b158de9a63c969d26256ea01acbd86a8e74d3aac45da30eac622

Observation 0d0790bb-0b6f-4354-a632-5587ddb2eefb · outbound

This paper cites REVEAL: Retrieval-Augmented Visual-Language Pre-Training with Multi-Source Multimodal Knowledge Memory.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning REVEAL: Retrieval-Augmented Visual-Language Pre-Training with Multi-Source Multimodal Knowledge Memory

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.394111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.394111Z digest=sha256:f25b19fa9f8d6896c0a003635b8bd423156ecb1f6b1557dbacea88aa7ff4c281

Observation cdcd0b62-1ac9-42fd-beca-718e10dd8200 · outbound

This paper cites Scaling Up Visual and Vision-Language Representation Learning With Noisy Text Supervision.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Scaling Up Visual and Vision-Language Representation Learning With Noisy Text Supervision

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.398127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.398127Z digest=sha256:c5e375d2a2ab8d98a11aecceb787c9f1ecb0a53676e8830347cd8147040110f5

Observation 68cfa15d-5001-412c-8014-6bba038e47b0 · outbound

This paper cites Visual Prompt Tuning.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Visual Prompt Tuning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.402265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.402265Z digest=sha256:402b69477a8bc6a024c69557ab8894b38f2555670dcb89f04a1f5650aa57913b

Observation f6844831-ca97-454d-923c-9795e2a6a256 · outbound

This paper cites Billion-scale similarity search with GPUs.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Billion-scale similarity search with GPUs

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.406187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.406187Z digest=sha256:f0360897c1d3584b8cb1aea3d2b56a1798cd5fc5910cb00782acc4364b3e2a62

Observation fbf6f1ea-74aa-4474-9dad-76ac4b4f5126 · outbound

This paper cites Deep Visual-Semantic Alignments for Generating Image Descriptions.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Deep Visual-Semantic Alignments for Generating Image Descriptions

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.410230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.410230Z digest=sha256:02342cce7673442c62574fccb95cefd1e6b2d1c497e7641fe277dca419513365

Observation 2e8b41ff-a450-4e9c-af0d-c5631d105fa0 · outbound

This paper cites Auto-Encoding Variational Bayes.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Auto-Encoding Variational Bayes

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.414573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.414573Z digest=sha256:c44b776ecb54d10c50cf3bb43b54db047d20919a5253e4ba992fca36f71bb81c

Observation 778da2bf-ee88-4ffc-a9d5-3a556e13dc02 · outbound

This paper cites The Power of Scale for Parameter-Efficient Prompt Tuning.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning The Power of Scale for Parameter-Efficient Prompt Tuning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.418671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.418671Z digest=sha256:dfe1c24f9c1d21d0936f2d1167d4979c36e9d96b3a23787debccb1cd314d9e71

Observation 59ca8558-0ed1-43ae-bf88-aa1a3ad6f4a5 · outbound

This paper cites BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.422712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.422712Z digest=sha256:a6ad739b392c21f86aea64f2841f771cef259dad2ce8c948b865c01939f50108

Observation 230e9ed6-070d-49f1-8d6d-6f975a8b0aa3 · outbound

This paper cites EVCap: Retrieval-Augmented Image Captioning with External Visual-Name Memory for Open-World Comprehension.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning EVCap: Retrieval-Augmented Image Captioning with External Visual-Name Memory for Open-World Comprehension

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.426766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.426766Z digest=sha256:a645e417547babb0ec74b2ad0c42a9aaf078b55406f519dc38051d63ce4432d5

Observation a9f4eac5-2562-4f10-b109-2dda3ee40556 · outbound

This paper cites an unresolved cited work.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-11T00:47:53.041629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.430620Z digest=sha256:b09d33fdf982732ae1a69a3dffbd4772456a6b303231825f14ec2bade43a87c1

Observation fd13970e-17f6-4ac2-becd-daad3b68fff5 · outbound

This paper cites Oscar: Object-Semantics Aligned Pre-training for Vision-Language Tasks.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Oscar: Object-Semantics Aligned Pre-training for Vision-Language Tasks

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.434115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.434115Z digest=sha256:46f3a27b8f1252af9c8849a9456008799094d7440cbd749cebf4585c378ae6a2

Observation 0395bd76-e967-4b2f-b5fd-f7a6b66007e1 · outbound

This paper cites Prefix-Tuning: Optimizing Continuous Prompts for Generation.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Prefix-Tuning: Optimizing Continuous Prompts for Generation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.437835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.437835Z digest=sha256:65b8950bbbb2c0f736a5342c6cf0858c56e828e1d2810a8470904b4977574cbc

Observation 434cd9d7-14b2-4a18-83bf-708115380476 · outbound

This paper cites Microsoft COCO: Common Objects in Context.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Microsoft COCO: Common Objects in Context

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.441953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.441953Z digest=sha256:2b3ac290e56d74abb87484a4b384389332dde39bd0a0c67e89df12e8b0e244af

Observation 2f665086-7745-455f-ac61-ae254c66e6c2 · outbound

This paper cites Few-shot Learning with Multilingual Language Models.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Few-shot Learning with Multilingual Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.446185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.446185Z digest=sha256:3095f2bac9ba5c893070cca04039d6c9067d95c7e6efe205d17d448db26173cd

Observation 93255380-34c1-443d-88a2-11ec4e3a7bee · outbound

This paper cites Visual Instruction Tuning.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Visual Instruction Tuning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.450290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.450290Z digest=sha256:d7a380b4e57985458a9dbcab17252997da483f51aba0d25f721d52c7093f9253

Observation 226c086b-add0-4e3f-839c-d969a4be0fc6 · outbound

This paper cites I-Tuning: Tuning Frozen Language Models with Image for Lightweight Image Captioning.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning I-Tuning: Tuning Frozen Language Models with Image for Lightweight Image Captioning

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-11T00:47:52.727828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.454181Z digest=sha256:1e46c196d03f9d9fadd731ff05bc7c500a43a5edeb0bd400f00de9f195ce0c66

Observation 230a2993-3479-4c04-8728-416ccb573566 · outbound

This paper cites MAPL: Parameter-Efficient Adaptation of Unimodal Pre-Trained Models for Vision-Language Few-Shot Prompting.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning MAPL: Parameter-Efficient Adaptation of Unimodal Pre-Trained Models for Vision-Language Few-Shot Prompting

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.458002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.458002Z digest=sha256:aba63335982038eb5fccff37c019ceb641ced1d6a58ad12f1ce83c90d1c40993

Observation 6f40e8af-c06c-4d64-87db-5b4eb3febfb7 · outbound

This paper cites ClipCap: CLIP Prefix for Image Captioning.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning ClipCap: CLIP Prefix for Image Captioning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.461697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.461697Z digest=sha256:6e8f5fcf74d7eb4b2e89311c433c346b3b91f42654f40dce73355437dc2b15f1

Observation e739e8bb-f0fe-4bc1-b724-fd797411b8b6 · outbound

This paper cites an unresolved cited work.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-11T00:47:53.030299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.465468Z digest=sha256:65f996aa8aabd0f90e2d5a11730caa5aecb2954d11936fb3de82ad42b5134b19

Observation e56161bb-1edb-48aa-b691-edaf369245f4 · outbound

This paper cites an unresolved cited work.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-11T00:47:53.018872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.468715Z digest=sha256:a066a5c3922e6ab94190ec7d399266e30df4fa31c2ec48f8fa0b2ed6570eec40

Observation 5841e810-0639-45f3-bec8-1f34f573c01e · outbound

This paper cites Flickr30k Entities: Collecting Region-to-Phrase Correspondences for Richer Image-to-Sentence Models.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Flickr30k Entities: Collecting Region-to-Phrase Correspondences for Richer Image-to-Sentence Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.472320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.472320Z digest=sha256:9b264d3043945772f2b8166224a077a7f98096b1bde065f6c238373a4b10d1a9

Observation afa7388d-2396-49cc-b606-d7566f664077 · outbound

This paper cites Learning Transferable Visual Models From Natural Language Supervision.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Learning Transferable Visual Models From Natural Language Supervision

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.476285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.476285Z digest=sha256:6908629e0e18438f714786b974b439a3f2417ad04f38f93658295025e090e833

Observation 982a31ea-aaee-44bb-8600-1f34d2273480 · outbound

This paper cites an unresolved cited work.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-11T00:47:53.007004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.480296Z digest=sha256:e14d2ee35d75295a7748f314ce19b9b3c5e1bef910e6377ac990fd75c7dd66f9

Observation 3ea24812-6756-427b-9c9a-6a89fb70bec9 · outbound

This paper cites LMCap: Few-shot Multilingual Image Captioning by Retrieval Augmented Language Model Prompting.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning LMCap: Few-shot Multilingual Image Captioning by Retrieval Augmented Language Model Prompting

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.483794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.483794Z digest=sha256:d7e4191178624ae69eac42e96e7c341263d766bb1243c7b18850c2b399bc7ff3

Observation a5fbff30-4ddb-4018-9f30-1e344aa2a986 · outbound

This paper cites SmallCap: Lightweight Image Captioning Prompted with Retrieval Augmentation.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning SmallCap: Lightweight Image Captioning Prompted with Retrieval Augmentation

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-08-11T00:47:52.660913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.487763Z digest=sha256:1dcf61f6f947244940cabe32c965f87660b9b072e534db80a3c754c2717a31b8

Observation 658f9a71-d444-4e33-b749-0caa9763f830 · outbound

This paper cites P.; Elliott, D.; and Martins, B.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning P.; Elliott, D.; and Martins, B

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:47:52.996295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.491626Z digest=sha256:d066f17e9271927909d7b2843333b4bf9e48f727be0e392ba6c90f35aaa9b5b8

Observation 54b2ca25-4afb-4bff-91d4-045daf8eb7ad · outbound

This paper cites EVA-CLIP: Improved Training Techniques for CLIP at Scale.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning EVA-CLIP: Improved Training Techniques for CLIP at Scale

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.495100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.495100Z digest=sha256:b45fba5b25c5bdd496f2ed2d82babd00abd8957e2762b5406bde49a6e05dcf86

Observation 99d81053-2031-459e-a814-9b90d8503bfb · outbound

This paper cites L.; and Parikh, D.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning L.; and Parikh, D

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:47:52.985218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.498807Z digest=sha256:6e8705eae16141ed48e6a894a4aae63c57cfa9fcb05fc6844961b2decd47ae79

Observation 049e45b1-06fd-41c7-9935-b80f1758c45c · outbound

This paper cites an unresolved cited work.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-11T00:47:52.974325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.502607Z digest=sha256:d14d5d62a73921c4b6d7815922b086a2129c82e88479f9c140df33f2a614b812

Observation cda9c693-6e3a-4b49-84b7-35ea6863a103 · outbound

This paper cites an unresolved cited work.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-11T00:47:52.963583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.506044Z digest=sha256:5c712e92b57e5a5b3ffa2f58d168767e0d80c4de87f6e3125d7a449b4f72f1d7

Observation e4514247-2ebd-4ed2-bc74-e7053b154cfd · outbound

This paper cites CogVLM: Visual Expert for Pretrained Language Models.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning CogVLM: Visual Expert for Pretrained Language Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.509718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.509718Z digest=sha256:a985d6ce1fe54ce8a8752c8198acc9edb4220aa41239999541612c7c9881f90a

Observation d03c0505-7d79-4c8e-a6d6-0fb2112b0a24 · outbound

This paper cites SimVLM: Simple Visual Language Model Pretraining with Weak Supervision.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning SimVLM: Simple Visual Language Model Pretraining with Weak Supervision

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.513342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.513342Z digest=sha256:f77854550991686af3dbea527b0a9171ddb83f092893fb33c863bc21910aa643

Observation 9268e79b-dd81-410f-803d-edb3aa4c880f · outbound

This paper cites DualPrompt: Complementary Prompting for Rehearsal-free Continual Learning.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning DualPrompt: Complementary Prompting for Rehearsal-free Continual Learning

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-08-11T00:47:52.615222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.517018Z digest=sha256:f32d38e513a2e3caf6780b8b9567b9f080f46efb961dcdede52136d3db98a2dc

Observation 91256ac0-9dd6-41e1-9938-660fafd9122c · outbound

This paper cites Re-ViLM: Retrieval-Augmented Visual Language Model for Zero and Few-Shot Image Captioning.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning Re-ViLM: Retrieval-Augmented Visual Language Model for Zero and Few-Shot Image Captioning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.520772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.520772Z digest=sha256:56913ab2be55e5452670a72afbcd6173d0dcb6326f724e1c3d7a0c760e262772

Observation c8eb9d5e-d19c-4e4b-bd0c-15a77cdf443e · outbound

This paper cites MeaCap: Memory-Augmented Zero-shot Image Captioning.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning MeaCap: Memory-Augmented Zero-shot Image Captioning

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-11T00:47:52.590167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-11T00:47:52.524701Z digest=sha256:c5928aeb37d198136b669d7e9dc3d0b073c2ea13968ae138499206d2e8fbbdc2

Observation a85aaa13-e464-4ee1-90cf-551ff68e6e21 · outbound

This paper cites OPT: Open Pre-trained Transformer Language Models.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning OPT: Open Pre-trained Transformer Language Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.528543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.528543Z digest=sha256:2f56f2970423e7d6dd7d37a4c31684351e8c8be5ee7a2900c2769d0d45927f7c

Observation a0fcaac9-22ea-4392-b6f1-fb1b67373a17 · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-11T00:47:52.532503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:47:52.532503Z digest=sha256:0ba80af5eebfa4902cf86b83b8585841f8e80e3a2eb18b471c5da8b628bf3dd3

Pith citing papers

No inbound Pith citation observations are available.