Pith. sign in

Paper Citation Record · LEDGER

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy

As of 7 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 3 inbound Pith citation observations for arXiv:2506.09958.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.09958 v1

Coverage vector

measured 54 of 54 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:42:02.726920Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:57:51.206063Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-12T03:56:21.604762Z

Reference resolution

54 of 54 outbound references displayed

  • verified exact4
  • verified fuzzy5
  • unresolved42
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1c52b270-b9ff-4969-b23b-a8c9878cbd98 · outbound

This paper cites Foundation metrics for evaluating effectiveness of healthcare conversations powered by generative AI.npj Digital Med., 7(82):1–14, March 2024.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Foundation metrics for evaluating effectiveness of healthcare conversations powered by generative AI.npj Digital Med., 7(82):1–14, March 2024

Reference 1

Resolution
malformed identifier
no resolver link, observed 2026-08-07T04:42:02.560852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.560852Z digest=sha256:3e56f56fd684cc391f110d4441df3e33effd1911a89dcb581b822bece4dd352b

Observation a2c6ce8e-7a85-4ee8-bff5-7d04aa538d59 · outbound

This paper cites Flamingo: a Visual Language Model for Few-Shot Learning.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Flamingo: a Visual Language Model for Few-Shot Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.564894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.564894Z digest=sha256:a578d6dc3793fb330c574599b242e8de8aac4cf0776e2febab3cf643c18a1c0d

Observation a011d625-0122-4965-8059-336b9f3babff · outbound

This paper cites A deep learning framework for quality assessment and restoration in video endoscopy.arXiv, April 2019.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy A deep learning framework for quality assessment and restoration in video endoscopy.arXiv, April 2019

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.568482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.568482Z digest=sha256:ac6b60e4217f4ea6da4175f94d3b741b333c96c77769fed7cd4ff4aeba6637ee

Observation 742143e1-a8a9-4ca3-8b5c-06eadc2925fb · outbound

This paper cites Qwen2.5-VL Technical Report.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Qwen2.5-VL Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.572023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.572023Z digest=sha256:63e170d58dfab49358e182c459dca94f31dc641202a445464e616210217492b5

Observation 9b01ca2c-d2e7-4f42-b1c4-a6b01b57a0d9 · outbound

This paper cites Vision–Language Model for Visual Question Answering in Medical Imagery.Bioengineering, 10(3):380, March 2023.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Vision–Language Model for Visual Question Answering in Medical Imagery.Bioengineering, 10(3):380, March 2023

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.575479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.575479Z digest=sha256:692615f058699a588465a20a9e757e35742bfaf9b90f725a780fbbfeaf27f230

Observation c8071b00-5997-42eb-b362-7c04b953944f · outbound

This paper cites Smedsrud, Steven Hicks, Debesh Jha, Sigrun L.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Smedsrud, Steven Hicks, Debesh Jha, Sigrun L

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.578778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.578778Z digest=sha256:ce027fa664f9e688605d56ae0c89603846d7d456c0ec64bdc558328d814347e8

Observation c193399a-2514-4e52-8e1b-6bf57876587d · outbound

This paper cites Iglovikov, and Alexandr A.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Iglovikov, and Alexandr A

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.581769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.581769Z digest=sha256:636a30928b3833cf1f5ae0e8019f13cb0c59e15b79e3cf442737a105d21eac97

Observation 4b9b7d8e-612e-4d26-8bcb-8f5485610ce8 · outbound

This paper cites A Simple Framework for Contrastive Learning of Visual Representations.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy A Simple Framework for Contrastive Learning of Visual Representations

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.584675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.584675Z digest=sha256:2ae3855bd1cbd7dceaa5327b736d48c876c751567485448e49b75223148937f8

Observation b1d98937-3b9b-4a47-92f4-fc1afd4bd9c3 · outbound

This paper cites R-LLaVA: Improving Med-VQA Understanding through Visual Region of Interest.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy R-LLaVA: Improving Med-VQA Understanding through Visual Region of Interest

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.588945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.588945Z digest=sha256:4df8e60d203352a1a36b35f34cb4f2ad94572f0a5122a2c55ed546a58b7e997d

Observation e0f0c41c-5985-4413-907c-2d2629c589ea · outbound

This paper cites Generative Models in Medical Visual Question Answering: A Survey.Appl.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Generative Models in Medical Visual Question Answering: A Survey.Appl

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.591678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.591678Z digest=sha256:7476d540abcf93b503452956e2af97a33f2c2c1132adb6ba131d2449cc9bdf01

Observation a3b1350f-f88c-4339-96fa-4f4a2a8c620f · outbound

This paper cites LLM-based NLG Evaluation: Current Status and Challenges.Computational Linguistics, pages 1–27, 2025.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy LLM-based NLG Evaluation: Current Status and Challenges.Computational Linguistics, pages 1–27, 2025

Reference 11

Resolution
malformed identifier
no resolver link, observed 2026-08-07T04:42:02.594102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.594102Z digest=sha256:b83c2c3a7dd0d5e9dd1af83f4d9ead3510d973aaede95261ed118e2b852e7e61

Observation 1c7bb54d-d855-4835-93c3-d9ae3e933a37 · outbound

This paper cites Hicks, Vajira Thambawita, P ˚ al Halvorsen, and Michael A.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Hicks, Vajira Thambawita, P ˚ al Halvorsen, and Michael A

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.597646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.597646Z digest=sha256:a2474f748384fef5207b7f2122818b352e93df05012280ba9a43371d4f01ba2d

Observation b8cbc6fb-8c62-40bc-bf01-60df620470cb · outbound

This paper cites Medgemma hugging face, May 2025.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Medgemma hugging face, May 2025

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:42:04.839627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:42:02.600361Z digest=sha256:6a8277d2bcd26281610cda20394cc5d1dffed53bddaef137756e434b1dbd7971

Observation 2d9bf78c-e0a5-4e92-8415-b4f5e3352157 · outbound

This paper cites LaPA: Latent Prompt Assist Model For Medical Visual Question Answering.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy LaPA: Latent Prompt Assist Model For Medical Visual Question Answering

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-08-07T04:42:03.038603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:42:02.603313Z digest=sha256:069fc90bf10f5e278cad7908dfbf7b3cb351a5f5d88e3205e2f57e4f985a891b

Observation 059da0cf-f138-4789-8bb7-8129831f33a1 · outbound

This paper cites DiN: Diffusion Model for Robust Medical VQA with Semantic Noisy Labels.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy DiN: Diffusion Model for Robust Medical VQA with Semantic Noisy Labels

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-07T04:42:03.024840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:42:02.606380Z digest=sha256:893c19090355e4852cb25f0968d3462c93d4e216d371a57d70ed0bfb6eb4510c

Observation 73acb3fc-ee4d-4efb-9116-8e12b4afe5dd · outbound

This paper cites Vision-language models for medical report generation and visual question answering: a review.Front.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Vision-language models for medical report generation and visual question answering: a review.Front

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.609354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.609354Z digest=sha256:5dfb12294798ad405d3a0d7ff9f3db35203c0192fd81ae8230dd82482d2d7d55

Observation ce258bb0-d390-4241-ad2d-c3a69acb344b · outbound

This paper cites DeBERTaV3: Improving DeBERTa using ELECTRA-Style Pre-Training with Gradient-Disentangled Embedding Sharing.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy DeBERTaV3: Improving DeBERTa using ELECTRA-Style Pre-Training with Gradient-Disentangled Embedding Sharing

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.612276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.612276Z digest=sha256:9704bf7d06ae05657565127581a4b205b2dae08a70d5984e989d65134d5bb6f2

Observation ba445f8d-489f-472a-be4c-234269813756 · outbound

This paper cites Overview of imageclefmedical 2023-medical visual question answering for gastrointestinal tract.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Overview of imageclefmedical 2023-medical visual question answering for gastrointestinal tract

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:42:04.613594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:42:02.617533Z digest=sha256:830b4821490439962c7149800c7cb664be11305194d54dc4c2d7d47dec859511

Observation 0dc39b54-26f4-4063-9f9c-6e19054be7dc · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy LoRA: Low-Rank Adaptation of Large Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.620794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.620794Z digest=sha256:6358d66726a07336ebb6edc887201a777530409d1a95bb89bf1e2732acfbb2bb

Observation 8aee339d-ca74-4715-94d7-19d6be5718b2 · outbound

This paper cites Summers, and Yingying Zhu.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Summers, and Yingying Zhu

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.624394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.624394Z digest=sha256:559c9e328b2fb1b68bafffa3bbecae0f3a6a1671921038ff46579de1419120d8

Observation 84594a3c-9911-445d-91f4-5d41e69b0f21 · outbound

This paper cites Sadman Hafiz, Jamin Rahman Jim, Md.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Sadman Hafiz, Jamin Rahman Jim, Md

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.627513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.627513Z digest=sha256:482f8ef37d395c4a719e21268fa4b5a568f216f2d53dac32ced0d2b33d80a761

Observation eff2e767-9c63-482a-920e-2a60e8a62874 · outbound

This paper cites Hicks, Vajira Thambawita, Enrique Garcia-Ceja, Michael A.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Hicks, Vajira Thambawita, Enrique Garcia-Ceja, Michael A

Reference 23

Resolution
verified exact
doi, observed 2026-08-07T04:42:02.991850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:42:02.631069Z digest=sha256:10012f0bd2c4bcd1d010e98aa29120a9894ac602bf049bf70789f11515f4bf3b

Observation bb2fe99e-3373-4845-ab3f-2a81e219314a · outbound

This paper cites Lau, Soumya Gayen, Asma Ben Abacha, and Dina Demner-Fushman.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Lau, Soumya Gayen, Asma Ben Abacha, and Dina Demner-Fushman

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.633750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.633750Z digest=sha256:d89a83278a4958b346865c8a114327251c006cf25d8eaaabb4aaa4b65ddab2fc

Observation d8645d43-43fa-4eed-ad94-0683aa80f03b · outbound

This paper cites Meteor: an automatic metric for MT evaluation with high levels of correlation with human judgments.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Meteor: an automatic metric for MT evaluation with high levels of correlation with human judgments

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.636692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.636692Z digest=sha256:573047c7b39cf8c717465976562df371cb12c7e46c7f78c90d308c23c83905f7

Observation 9b4184d9-7b52-4b7d-a755-5e76e336007f · outbound

This paper cites LLaVA-Med: Training a Large Language-and-Vision Assistant for Biomedicine in One Day.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy LLaVA-Med: Training a Large Language-and-Vision Assistant for Biomedicine in One Day

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.639446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.639446Z digest=sha256:92d3753f942aae4c7cf186adae299e65d2b858ec2ed444c534e0244b61fa0155

Observation 6374a26f-cdb1-4954-b58a-21362fd715d5 · outbound

This paper cites Less Could Be Better: Parameter-efficient Fine-tuning Advances Medical Vision Foundation Models.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Less Could Be Better: Parameter-efficient Fine-tuning Advances Medical Vision Foundation Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.642311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.642311Z digest=sha256:e51a80d57b75dfec1aa97445108540c6a8fc41617f5a3d3c2225d671b0436e35

Observation 5f2a347e-1273-423f-8ea4-455a9a976e34 · outbound

This paper cites Candidate-Heuristic In-Context Learning: A new framework for enhancing medical visual question answering with LLMs.Information Processing & Management, 61(5):103805, September 2024.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Candidate-Heuristic In-Context Learning: A new framework for enhancing medical visual question answering with LLMs.Information Processing & Management, 61(5):103805, September 2024

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.645547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.645547Z digest=sha256:591a590bde090452eb23f28cdae32ff5cf3c99959badcd40d80900ee2d77be82

Observation b3ee46df-b7b5-4e5a-86c2-0948147c2b90 · outbound

This paper cites Rouge: A package for automatic evaluation of summaries.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Rouge: A package for automatic evaluation of summaries

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.648531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.648531Z digest=sha256:2e48c24a75b9ad468baed80ea0a41af9a35b73913bd4828ecc64272eae562566

Observation 2c6ec00c-b26a-4396-806b-9c645233b806 · outbound

This paper cites Medical visual question answering: A survey.Artif.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Medical visual question answering: A survey.Artif

Reference 30

Resolution
metadata mismatch
raw_fallback, observed 2026-08-07T04:42:03.491149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:42:02.651405Z digest=sha256:fab73adf08600bf113bc22b4539359c1061bc1e1a5d0b25191bd0a89974ad438

Observation 4927d41e-e237-47b8-aa8b-a2ea1a5640b1 · outbound

This paper cites Slake: A Semantically-Labeled Knowledge-Enhanced Dataset For Medical Visual Question Answering.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Slake: A Semantically-Labeled Knowledge-Enhanced Dataset For Medical Visual Question Answering

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.654115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.654115Z digest=sha256:9ffb74d5bd313c0908ea13fc6f8c095aa7f324074fc9a936ec8b2bdc0e45f896

Observation 454e4dff-4980-4af8-8a5a-ac47abd2aee0 · outbound

This paper cites Visual Instruction Tuning.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Visual Instruction Tuning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.657045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.657045Z digest=sha256:c726570e5beb73bef91e0f8b15693d6f92537d3e4bd97ce9f77e4f0a4d1f57b4

Observation 35c11c11-694e-4dd3-92a9-a62c5d80325b · outbound

This paper cites Peft: State-of- the-art parameter-efficient fine-tuning methods.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Peft: State-of- the-art parameter-efficient fine-tuning methods

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:42:04.463784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:42:02.659998Z digest=sha256:c74aabcc5cee77c0a586c6673571778143603b65f5449b1b304978e97a550209

Observation 2f0cbe2a-c590-420f-9b06-cbd8b6b11149 · outbound

This paper cites Med-Flamingo: a Multimodal Medical Few-shot Learner.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Med-Flamingo: a Multimodal Medical Few-shot Learner

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.665674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.665674Z digest=sha256:fdb10d51cdf16a569ec5434e4f387a6693aed69fff4f51f82221f86cec50282f

Observation 27c44666-2662-4c50-8ac3-b0c032f09de0 · outbound

This paper cites GPT-4 Technical Report.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy GPT-4 Technical Report

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.668362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.668362Z digest=sha256:68034e7f41cfc19bd0ee0be72aabc0b5e0b8806f1200170839a1479afac1c15f

Observation 63bbee4d-3107-4a92-94b2-4d782acf47e3 · outbound

This paper cites Chaudhari, et al.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Chaudhari, et al

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.671397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.671397Z digest=sha256:12b9f9edf68edeb312a481465c3563e7cea57be39139e43fbd1720b05766dc53

Observation de414271-f059-4fcd-86d9-dd21f162ea55 · outbound

This paper cites BLEU: a method for automatic evaluation of ma- chine translation.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy BLEU: a method for automatic evaluation of ma- chine translation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.673870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.673870Z digest=sha256:3298d9fc643562816dfa84be5db5031e1fc3f5ce9b329a6c614a3beba44ae9a5

Observation c9f0e415-519b-4534-b083-ff7561b001d0 · outbound

This paper cites chrF: character n-gram F-score for automatic MT evaluation.ACL Anthology, pages 392–395, September.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy chrF: character n-gram F-score for automatic MT evaluation.ACL Anthology, pages 392–395, September

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:42:04.314693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:42:02.676750Z digest=sha256:84e5cc510662de243d0f8052fac57c401b5c0046b624106a1330507add268a36

Observation 71e43144-29fe-42b6-bee7-f243b85df2ec · outbound

This paper cites ZeRO: Memory Optimizations Toward Training Trillion Parameter Models.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy ZeRO: Memory Optimizations Toward Training Trillion Parameter Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.682342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.682342Z digest=sha256:2dbe3f53e79f073a443ddf3016f194930175cd21cd9a24263f1a6db147bcf510

Observation 12a40f5f-3c5e-486e-86aa-fae06d9bfec2 · outbound

This paper cites Vision-Language and Large Language Model Performance in Gastroenterology: GPT, Claude, Llama, Phi, Mistral, Gemma, and Quantized Models.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Vision-Language and Large Language Model Performance in Gastroenterology: GPT, Claude, Llama, Phi, Mistral, Gemma, and Quantized Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.685440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.685440Z digest=sha256:dfd16c58c37c547218ee5168d19378881acf4e89d7f472eab88a0100bccffb1a

Observation 3d5da4d0-b840-4da3-a59f-969e4b5f2476 · outbound

This paper cites BLEURT: Learning Robust Metrics for Text Generation.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy BLEURT: Learning Robust Metrics for Text Generation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.689211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.689211Z digest=sha256:130cc56859411936c13ffe06e02219edb8a9a7cf4e37307c5f6a5786bd4e1963

Observation 24d49581-6a9f-44fb-848d-238a30735c5b · outbound

This paper cites Pfohl, Heather Cole-Lewis, et al.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Pfohl, Heather Cole-Lewis, et al

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.692791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.692791Z digest=sha256:5979a3de8e356d3a96a0a0ca47705fe7f0fae25f28cc8462373b123f079597d6

Observation c7b19ffe-ac7f-4563-9322-33b9c37baa25 · outbound

This paper cites Qwen3 Technical Report.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Qwen3 Technical Report

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.695513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.695513Z digest=sha256:6ef6e35d8440612271af804e3921b1648d30cf346d339ab16348cd410a8b8701

Observation ea035ea9-0fd3-475e-bb1f-20b208f33a15 · outbound

This paper cites Sanders, Yuchen Liu, Kennarey Seang, Bach Xuan Tran, Atanas G.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Sanders, Yuchen Liu, Kennarey Seang, Bach Xuan Tran, Atanas G

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.698427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.698427Z digest=sha256:d8d901ce33a4545b045d99d3e762aef3be1edcc3356dacba0e9d5419e31f87e8

Observation 61c26755-4ced-440d-9ee4-61af9d046faf · outbound

This paper cites Wilkinson, Michel Dumontier, IJsbrand Jan Aalbersberg, Gabrielle Appleton, Myles Axton, Arie Baak, Niklas Blomberg, Jan-Willem Boiten, Luiz Bonino da Silva Santos, Philip E.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Wilkinson, Michel Dumontier, IJsbrand Jan Aalbersberg, Gabrielle Appleton, Myles Axton, Arie Baak, Niklas Blomberg, Jan-Willem Boiten, Luiz Bonino da Silva Santos, Philip E

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.701389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.701389Z digest=sha256:e11165b2346045640d882e71d66c383deafd01496a3b431da449dfaf343b23df

Observation fab20e9a-a383-49f6-9eb6-f9752df995cf · outbound

This paper cites an unresolved cited work.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:42:04.164751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:42:02.704076Z digest=sha256:e76787f9b1de5158dcbdcb021015f98f3bb8e18c5c368dda1716a191bd233f46

Observation e22cee96-9e5c-420f-825a-ebdb89da6b06 · outbound

This paper cites Parameter-Efficient Fine-Tuning Methods for Pretrained Language Models: A Critical Review and Assessment.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Parameter-Efficient Fine-Tuning Methods for Pretrained Language Models: A Critical Review and Assessment

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.706457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.706457Z digest=sha256:def654dc78b5b3921f298e7bfaa8069f44937cbe28e8f7beea24ed9ccab2c8cd

Observation ede2d0b1-2c9f-4fa3-b54b-9e6713528d75 · outbound

This paper cites MedFrameQA: A Multi-Image Medical VQA Benchmark for Clinical Reasoning.arXiv, May 2025.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy MedFrameQA: A Multi-Image Medical VQA Benchmark for Clinical Reasoning.arXiv, May 2025

Reference 49

Resolution
verified exact
doi, observed 2026-08-07T04:42:02.854127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:42:02.709606Z digest=sha256:02c8dfd231adcb8c0211470eecac86dc06ac287a13f0a02e5cafc9cb36f80c43

Observation f5190ec1-fb2e-44d6-8a72-8eb1b5c4be42 · outbound

This paper cites Fine-grained Adaptive Visual Prompt for Generative Medical Visual Question Answering.AAAI, 39(9):9662–9670, April 2025.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Fine-grained Adaptive Visual Prompt for Generative Medical Visual Question Answering.AAAI, 39(9):9662–9670, April 2025

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.712182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.712182Z digest=sha256:d3618989f334589cc99b19aa268330a2e4e4bf1b31bc5f33f03157f9cc61dc8a

Observation 5f569dee-710c-4b5a-8fbc-8f61c8bd0d7e · outbound

This paper cites Medical Visual Question Answering via Conditional Reasoning.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Medical Visual Question Answering via Conditional Reasoning

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:42:04.013804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:42:02.715622Z digest=sha256:022c663b2bf1f662ec7d87355772f61d45665dc255c4e471330ee8b27b139b63

Observation 9685044a-1a2a-4a01-b705-e61c3f63213f · outbound

This paper cites BERTScore: Evaluating Text Generation with BERT.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy BERTScore: Evaluating Text Generation with BERT

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.720800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.720800Z digest=sha256:76a61dec3e18c1f8bc7b853be0f490ade0f7418e2ab56913ac8fa1a8e3f983e0

Observation fb0332a3-b5c3-4810-82d8-acbe4ada430d · outbound

This paper cites PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.723789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.723789Z digest=sha256:778fcdcb923e8fb9cfa53b3506c4f5babfde67cac87bfe95e6f8bb8e4c356143

Observation 24983bd3-10c5-4138-8d99-7ff0ad5ad655 · outbound

This paper cites SWIFT:A Scalable lightWeight Infrastructure for Fine-Tuning.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy SWIFT:A Scalable lightWeight Infrastructure for Fine-Tuning

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.726920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.726920Z digest=sha256:4e55ed643c326b9150061f36197d695b4f280918c8560dbf3d9c3d7c37a66f0d

Observation 9973781f-b476-41fa-82b6-60c86ed56032 · outbound

This paper cites an unresolved cited work.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Unresolved cited work

Reference 2015

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.679485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.679485Z digest=sha256:52a8da401ca630bfff2dd21992befb504c489b1c5d6c8a0147c3a1ea9b8b968b

Observation 7fd676bb-74cc-4ad6-8035-32ac396b560a · outbound

This paper cites an unresolved cited work.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy Unresolved cited work

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.718367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.718367Z digest=sha256:189fb661da9a9110eb087eeb563555e9290e05062c79ab9e57692ef22c23ff6d

Pith citing papers

Observation f0beebc1-9999-4b7e-97b3-738c0a256f0c · inbound

Multimodal AI for Gastrointestinal Diagnostics: Tackling VQA in MEDVQA-GI 2025 cites this paper.

Multimodal AI for Gastrointestinal Diagnostics: Tackling VQA in MEDVQA-GI 2025 Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T15:57:51.206063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:57:51.206063Z digest=sha256:b7487d50487dd0c6e5137ad3959f461ffb2a7b6d2ce8115da68b707b5138490f

Observation 16a6d077-dde9-4722-9d55-5f74cf186cc2 · inbound

RadThinking: A Dataset for Longitudinal Clinical Reasoning in Radiology cites this paper.

RadThinking: A Dataset for Longitudinal Clinical Reasoning in Radiology Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:56:21.606430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T03:55:55.359488Z digest=sha256:dbfeab3cb2daf146abeecf00d4369dfc857b340656688f51c7c5dc3259875d0e

Observation e4faedd8-0162-413c-94cc-6b6cde36c72b · inbound

Measuring and Improving Complex-Atomic Answer Consistency in Endoscopic VQA cites this paper.

Measuring and Improving Complex-Atomic Answer Consistency in Endoscopic VQA Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T16:57:30.636879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:57:30.636879Z digest=sha256:0ff10194f11296c707cbe16ac3b580ac258220b5e1942bf92250c43a0283ec14