Pith. sign in

Paper Citation Record · LEDGER

MedBLIP: Fine-tuning BLIP for Medical Image Captioning

As of 16 August 2026, this Paper Citation Record lists 16 of 16 outbound references and 0 inbound Pith citation observations for arXiv:2505.14726.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.14726 v1

Coverage vector

measured 16 of 16 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:13:32.927976Z

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

16 of 16 outbound references displayed

  • verified exact0
  • verified fuzzy12
  • unresolved4
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 52a08e91-b55a-4fec-8031-7271062a4efb · outbound

This paper cites Imageclef medical caption.

MedBLIP: Fine-tuning BLIP for Medical Image Captioning Imageclef medical caption

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:33.554643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T20:13:32.724178Z digest=sha256:04b59a3edba81a6ff69d4406777bfa77117e6712afd7c499199bde543d5057ef

Observation a2ef3908-02cf-4097-ace6-f9bcd19b80d8 · outbound

This paper cites MedBLIP: Bootstrapping Language-Image Pre-training from 3D Medical Images and Texts.

MedBLIP: Fine-tuning BLIP for Medical Image Captioning MedBLIP: Bootstrapping Language-Image Pre-training from 3D Medical Images and Texts

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T20:13:32.740940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:13:32.740940Z digest=sha256:718869e77c8efda005b93321971a96d96913c5fbdb20b217c1e0ba7320a5ee83

Observation 8dd8ddfe-f920-420f-9c8d-76df1daa4639 · outbound

This paper cites ROCOv2: Radiology Objects in COntext Version 2, an Updated Multimodal Image Dataset.

MedBLIP: Fine-tuning BLIP for Medical Image Captioning ROCOv2: Radiology Objects in COntext Version 2, an Updated Multimodal Image Dataset

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T20:13:32.745462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:13:32.745462Z digest=sha256:fecd9482bc8ba20653f1e3e6622344f0328ce235a0a81ce20960858e049f8f7d

Observation b516f6d3-a0eb-4b36-9d3a-dcb2bc335f2b · outbound

This paper cites an unresolved cited work.

MedBLIP: Fine-tuning BLIP for Medical Image Captioning Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:13:33.544409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T20:13:32.750068Z digest=sha256:6e64ef2db55f77f6a05db3435518103cb34dfaaf45cbaf6607fbc6ac601e0109

Observation da202823-4ddc-4fc5-b6ee-34f032a43c7b · outbound

This paper cites Visual cluster grounding for im- age captioning.

MedBLIP: Fine-tuning BLIP for Medical Image Captioning Visual cluster grounding for im- age captioning

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:33.453304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T20:13:32.753592Z digest=sha256:27e9b6571a83f419d4a98454d55d75728ec4a5ac18634daa072f4bc26740d759

Observation a8fe9d58-5f33-4175-9f45-921113faad7d · outbound

This paper cites Blip: Bootstrapping language-image pre-training for unified vision-language understand- ing and generation.

MedBLIP: Fine-tuning BLIP for Medical Image Captioning Blip: Bootstrapping language-image pre-training for unified vision-language understand- ing and generation

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:33.441687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T20:13:32.758221Z digest=sha256:8cb61abc2391966c6bb5c28eae1687ebfe26d75ad8522607f55e342072823cdf

Observation 12366f7e-e31a-46cf-a69c-2f16ea042089 · outbound

This paper cites Blip-2: Bootstrapping language- image pre-training with frozen image encoders and large language models, 2023.

MedBLIP: Fine-tuning BLIP for Medical Image Captioning Blip-2: Bootstrapping language- image pre-training with frozen image encoders and large language models, 2023

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:33.358867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T20:13:32.762242Z digest=sha256:5621432beba8527972fe402cc4fb35a2a24db41d4135ba9f6a4e7a9507111381

Observation 45325507-5ec9-44d0-874c-f06930639176 · outbound

This paper cites A Systematic Review of Deep Learning-based Research on Radiology Report Generation.

MedBLIP: Fine-tuning BLIP for Medical Image Captioning A Systematic Review of Deep Learning-based Research on Radiology Report Generation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T20:13:32.766363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:13:32.766363Z digest=sha256:524f1cf91ed2c4b03b58a15a17fd0a83894a01267d22ea8f0bb2aaf9ef1d64a0

Observation fc29797f-f7ac-4114-b23b-2bfa22c9b06c · outbound

This paper cites Image caption generation using vision transformer and gpt architecture.

MedBLIP: Fine-tuning BLIP for Medical Image Captioning Image caption generation using vision transformer and gpt architecture

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:33.348350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T20:13:32.769797Z digest=sha256:b3ee578551e7f05eead698dbae34d00cbcb57688bd1bfbf1ba2acb4704b5c94f

Observation 5605740a-3fb5-4a6d-a518-b6e8ff03679e · outbound

This paper cites Uit-darkcow team at image- clefmedical caption 2024: Diagnostic captioning for radiology images efficiency with transformer models, 2024.

MedBLIP: Fine-tuning BLIP for Medical Image Captioning Uit-darkcow team at image- clefmedical caption 2024: Diagnostic captioning for radiology images efficiency with transformer models, 2024

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:33.337341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T20:13:32.783213Z digest=sha256:d64c3858bd63fed358efbe7a603056d80fe2224ddf8ea3620c5dfc0d0fce8543

Observation 4c5fdba8-87d0-4c4c-b3c9-843d71bd3ef6 · outbound

This paper cites Learning transferable visual models from natural language supervision.

MedBLIP: Fine-tuning BLIP for Medical Image Captioning Learning transferable visual models from natural language supervision

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:33.321788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T20:13:32.881076Z digest=sha256:cdf08902e8650e33a296f1f75a08cd2d8e1f973be9ae95edcedeff98aea49df3

Observation c4a0c560-1a90-495b-80da-fded6481e663 · outbound

This paper cites Overview of Image- CLEFmedical 2024 – Caption Prediction and Con- cept Detection.

MedBLIP: Fine-tuning BLIP for Medical Image Captioning Overview of Image- CLEFmedical 2024 – Caption Prediction and Con- cept Detection

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:33.229655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T20:13:32.912693Z digest=sha256:b16583b940217b56c67ede780344842c44523cbbbbf96766d227227048019dfd

Observation c9ab49c0-bcab-48df-8758-cb33a02aa8b6 · outbound

This paper cites Medical image captioning with blip2 and opt-6.7b.

MedBLIP: Fine-tuning BLIP for Medical Image Captioning Medical image captioning with blip2 and opt-6.7b

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:33.218375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T20:13:32.916647Z digest=sha256:5ffed4c149f906403ddb3e8c3f4085057d24e386af312f480d6f7451b5b5cee7

Observation 25fc2a32-17f0-4ec5-869c-8366d9482e98 · outbound

This paper cites Iu chest x-ray collection.

MedBLIP: Fine-tuning BLIP for Medical Image Captioning Iu chest x-ray collection

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:33.206537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T20:13:32.920567Z digest=sha256:d8bbe88cc641d17a995dd31307f0ac4d51d6c80434bb69e14fc4d8fad4268412

Observation 777eca17-66c6-402c-9673-a1dc9fa53c01 · outbound

This paper cites Show and tell: A neural image caption generator.

MedBLIP: Fine-tuning BLIP for Medical Image Captioning Show and tell: A neural image caption generator

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:33.095971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T20:13:32.924365Z digest=sha256:d6acc06176b40e9fd696d56c96d18bad288122a93639b33b9511112386e9d8df

Observation b4a2ee70-d757-4f6d-840a-327d8bf3a609 · outbound

This paper cites Show, attend and tell: Neural image caption generation with visual attention.

MedBLIP: Fine-tuning BLIP for Medical Image Captioning Show, attend and tell: Neural image caption generation with visual attention

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:13:33.039414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T20:13:32.927976Z digest=sha256:463b0c699b4e3642e37c8adb9f2a9ad4f6ee6c3277cb7d30b7227a54b92393c6

Pith citing papers

No inbound Pith citation observations are available.