Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-01T23:47:50.268649Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 0 inbound Pith citation observations for arXiv:2607.15241.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-01T23:47:50.268649Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
43 of 43 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation e4ebc507-d62e-44e1-bf94-5f791056258f · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA Medico 2025: Visual Question Answering for Gas- trointestinal Imaging,
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d16a7216-3a40-44fc-9818-4085e9660cdb · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA Kvasir- VQA-x1:A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 5dec4e0d-4588-455c-bb74-84a48943852c · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA LoRA: Low-rank adaptation of large language models,
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9dcffa5-4f19-4203-a93e-ec59e81f761d · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA Qlora: efficient finetuning of quantized llms,
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26945d1b-331b-498f-8161-b968d2a1e030 · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA Parameter-Efficient Fine-Tuning Methods for Pretrained Language Models: A Critical Review and Assessment,
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f05b3ad-951d-4b70-a0a6-f6accc12d987 · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA A Survey on Medical Large Language Models: Tech- nology, Application, Trustworthiness, and Future Directions,
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da9bb747-cbe2-4829-9677-7b9f50b3bf53 · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA VQA-Med: Overview of the medical visual ques- tion answering task at imageclef 2019,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f643b1a8-7b63-49ad-9eb8-d0d57fb0216f · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA Medical visual question answering at imageclef-vqa med,
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5a89455-2369-4dff-9f7d-0e0ca65a1abb · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA Medical visual question answering: A survey,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9151dd43-1ebb-4203-a109-1c10c3687e90 · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA Overview of ImageCLEFmedical 2025– Visual Question Answering and Synthetic Image Generation for Gastrointestinal Tract,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5339cb1b-4d73-48f7-81ab-d46c12c1717d · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA Kvasir-VQA: A Text-Image Pair GI Tract Dataset,
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5cd66580-afca-4f33-b966-b0ed6097ac59 · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA Exploring Vision-Language Models for Medical VQA on Gastrointestinal Images: A LoRA Fine-Tuning Study,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d0c46b8-516a-4dcb-b199-cdff2161440a · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA LoRA-Enhanced PaliGemma for Efficient Visual Question Answering in Gastrointestinal Imaging,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d614b8a-4db9-48eb-91a6-cee613eb9f89 · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA Multimodal Explanations: Justifying Decisions and Pointing to the Evidence,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1360515-aeb9-4a11-9ef2-c5df9b42e6f5 · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA Towards faithfully interpretable NLP systems: How should we define and evaluate faithfulness?
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea716100-f6f1-4e61-a763-5206f7367ae0 · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA Towards Faithful Model Explanation in NLP: A Survey,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd9a8709-8a16-4e3c-88e5-016dfb7b7596 · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68ccbab9-1bf5-4d6f-b779-5f07e51735f3 · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA From Answers to Explanations: Self-Probing Efficiently Fine-Tuned Vision-Language Models for Medical VQA at Medico 2025,
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 681e5945-f427-4ea2-ac2d-4f2a37530fb9 · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA Curriculum-Guided Fine-Tuning for Multimodal VQA in GI Endoscopy (Team Lama4Vision),
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5ca3c1e-3da0-492c-a78d-8d709f5d43aa · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA Qwen3 Technical Report
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7236eafa-da93-4966-8f40-deb8c5dc4ebe · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA Medico 2025: Visual Question Answering for Gastrointestinal Imaging,
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 493f3572-b12f-474c-9412-4f751380382f · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA ROUGE: A Package for Automatic Evaluation of Sum- maries,
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a82743ab-0539-45c4-9570-acededdcdc94 · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA METEOR: An Automatic Metric for MT Evaluation with Improved Correlation with Human Judgments,
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 110fcc60-9559-4100-bdb4-b083f2e2a41b · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA chrF++: words helping character n-grams,
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6abbe02b-a3a5-43f5-82a6-6458c5dd0dff · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA BLEU: a Method for Automatic Evaluation of Machine Translation,
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24df5ea4-9ff7-43fc-aa81-9d2341311d4b · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA BERTScore: Evaluating Text Generation with BERT,
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f988d8ec-5902-40a0-8837-923347d84bd7 · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA Evaluate: A library for easily evaluating machine learn- ing models and datasets,
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0d5858b-8470-4f7e-bb7b-c91f24e72121 · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA Multi-Task Learning for Visually Grounded Reasoning in Gastrointestinal VQA,
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1bec4554-0cb9-4d1a-a21d-ea1453b008dd · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA Medico 2025: Visual Question Answering (with Multimodal Explanations) for Gastrointestinal Imaging,
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a0891e2-cfa0-4662-abf6-c849e2cb7fd8 · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA Enhancing Encoder-Decoder Architecture to Visual Question Answering Task for Gastrointestinal Images,
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 959a3a71-2560-4da1-9f26-fae5a11d93e7 · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA BLIP-2-based Visual Question Answering with Multimodal Explanations for Gastrointestinal Imaging,
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc4b55f5-baa4-464c-9c54-7fe22d247155 · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA X-VQA for GI Diagnostics: Multimodal Visual Question Answering with Confidence-Aware Explanations,
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01ce6139-da6c-4c35-be3f-bb35e21e2f5c · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA Florence-2: Advancing a Unified Representation for a Variety of Vision Tasks
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c948514-a116-437c-8b8f-0b9341fe9c28 · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA PaliGemma: A versatile 3B VLM for transfer
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21735fab-982d-4862-af1c-9ea6dc7c9c46 · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA PaliGemma 2: A Family of Versatile VLMs for Transfer
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c965876-d298-4e34-88c4-468c00889d1b · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 964f9428-2927-4230-98fd-36e4e4208916 · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e52d2d06-c041-4263-afe5-8600e111f4ea · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edd60e04-6a9a-4b1d-a3df-b01248df2e21 · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA Image Segmentation Using Text and Image Prompts,
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8429f8ad-901e-4982-a43f-f08385b01322 · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA Grad-CAM: Visual Explanations from Deep Networks via Gradient-Based Localization,
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 120f8509-c27b-444f-b5f5-9c9a2c5543c4 · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA Quantifying attention flow in transformers,
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20d0c7a3-558b-4114-a3c7-48f4f355eed8 · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA Curriculum Learning: A Survey,
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aaa4a646-4320-4ef4-a2f7-34572a60d652 · outbound
Beyond the Leaderboard: Design Lessons for Trustworthy Multimodal VQA LLM-FP4: 4-bit floating-point quantized transformers,
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.