Pith. sign in

Paper Citation Record · LEDGER

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models

As of 7 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 0 inbound Pith citation observations for arXiv:2606.06534.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.06534 v1

Coverage vector

measured 51 of 51 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-28T03:25:43.795795Z

measured 51 of 51 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

51 of 51 outbound references displayed

  • verified exact2
  • verified fuzzy0
  • unresolved49
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 74d47ca5-2c91-4fd7-9135-13ebe0635bce · outbound

This paper cites METEOR: An auto- matic metric for MT evaluation with improved correlation with human judgments.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models METEOR: An auto- matic metric for MT evaluation with improved correlation with human judgments

Reference 1

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:1dd4cd0932b3eab46f68ddb3e826dbd2bf5ee5591fc4e2db4ee1180aebdaeda6

Observation 2b2a52be-90a6-47cc-981b-714cc74ff394 · outbound

This paper cites Saliency-driven ex- plainable deep learning in medical imaging: Bridging visual explainability and statistical quantitative analysis.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Saliency-driven ex- plainable deep learning in medical imaging: Bridging visual explainability and statistical quantitative analysis

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:ef882a30fd91bee75cece3df18002b1cc24dd52d5fd5fbafa0bc2de534a70d83

Observation 51275b1c-441c-42fe-858a-14b0949f2acf · outbound

This paper cites Swin-unet: Unet-like pure transformer for medical image segmentation,.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Swin-unet: Unet-like pure transformer for medical image segmentation,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:12f6c906649496f6f003c0503d19fde8c07f2bfa9629847587f90184e32a4f38

Observation 91f781fa-0d2e-4c96-a0f2-1faa47920e78 · outbound

This paper cites Yuille, and Yuyin Zhou.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Yuille, and Yuyin Zhou

Reference 4

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:a8edcc3607ff2aae6d83bed3cc0a1024f452ce91f70ae594bfbff7fdd07a60f3

Observation 7bd4b319-66be-4d07-bda1-d3f45ae4f9ca · outbound

This paper cites Pretraining vision-language model for difference visual question answering in longitudinal chest x-rays, 2024.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Pretraining vision-language model for difference visual question answering in longitudinal chest x-rays, 2024

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:7774fc4d10597a5752c32c980d1b118102a699992a30c5d8e458f192a70e2850

Observation bc9f6c91-c73d-42d2-ba16-e37b245fb42e · outbound

This paper cites Co-saliency detection with co-attention fully convo- lutional network, 2020.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Co-saliency detection with co-attention fully convo- lutional network, 2020

Reference 6

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:b8c360291d936ee142234fecb95b43ab6f39e07d86653845e75ab5100597f91a

Observation d181d43a-b611-4f7b-9ebc-9baad5b07ae4 · outbound

This paper cites Goldberger, Luis A.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Goldberger, Luis A

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:f6fd1f7ebfacdf05d1c438814c997bca4ea02bf25f88edcd8c7ac9aaa7d1f838

Observation c86a62c0-8ce1-4c15-8aae-0d70fe088e7c · outbound

This paper cites Unetr: Transformers for 3d medical image segmentation, 2021.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Unetr: Transformers for 3d medical image segmentation, 2021

Reference 8

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:474b94db3ba190c3b907f1fbc7a8afa51779a85f84638feb48b927cbd50d676b

Observation e3fe7c9d-c58e-420d-8700-a1da7bf06f7b · outbound

This paper cites Medical-Diff-VQA: A Large- Scale Medical Dataset for Difference Visual Question An- swering on Chest X-Ray Images.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Medical-Diff-VQA: A Large- Scale Medical Dataset for Difference Visual Question An- swering on Chest X-Ray Images

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:66437312095a611ca7c2d987bc59e5e312038e8512fce2acbd7c64744cc8d959

Observation 6c9d75ea-3233-47f6-98c4-78a0b99edb59 · outbound

This paper cites Summers, and Yingying Zhu.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Summers, and Yingying Zhu

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:2ade0f8153463e3316df7b294e2a935271acb424807a1f2df2672a8515846efb

Observation 11dc461c-1060-4c72-9d04-ef181bcdabea · outbound

This paper cites Lungren, and Serena Yeung.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Lungren, and Serena Yeung

Reference 11

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:1393dba7579105fbebf1ffba7c8a9f7089eb4193bafb1bdbbafea1b09d7f85ca

Observation 934357a1-b883-4a80-93ac-f370c87c1121 · outbound

This paper cites Elsabawy, Alaa Ebraheem Elnakeeb, and Nora Elrashidy.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Elsabawy, Alaa Ebraheem Elnakeeb, and Nora Elrashidy

Reference 12

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:ea303f61584928fcc0f8f88f856626747f0b785965c89a238df4f63e8dbf1886

Observation 27c63b98-c214-46b1-b89a-2cdd05acd10f · outbound

This paper cites Jaeger, Simon A.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Jaeger, Simon A

Reference 13

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:f40bae005adf42883d080acdfc062a6745d2a97e85a58567e07192d1d74722b5

Observation fde7716b-f967-4b1b-a077-940cebdcc29b · outbound

This paper cites Spatial transformer networks, 2016.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Spatial transformer networks, 2016

Reference 14

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:97a2f71a0fd7d6d7003c87f98c0eb0d62d14c0574e32df43a052b21157dc91ca

Observation 0523bd36-fb98-4235-93fd-531a21b51864 · outbound

This paper cites One map does not fit all: Evaluating saliency map explanation on multi-modal medical images, 2021.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models One map does not fit all: Evaluating saliency map explanation on multi-modal medical images, 2021

Reference 15

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:50e89adb07408fc5885a0868ffd4e95ae8e20cb2630390e0d55a62d4cca22aff

Observation d5a853e3-44ae-4802-9425-e56c169b3966 · outbound

This paper cites an unresolved cited work.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:ebce72f8177db5cc88a629110d6d34ae07a51d46fa1db1a1bbb8d45e21fc48e9

Observation b4688c77-7f5f-4380-874a-6b81748c4380 · outbound

This paper cites an unresolved cited work.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:ef96d4481aa188eba7e19e6453cfdb946e484dc145d4cc66f3f062435106fa26

Observation 8220432c-8e88-4bc3-a533-d3fdacf5f3a5 · outbound

This paper cites Schroeder, and Tolga Tasdizen.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Schroeder, and Tolga Tasdizen

Reference 18

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:9d26033f7c85d5181f0d487d8211a796645a04cf90feab9124cd026a7506f157

Observation 1f53a779-0798-4d8b-85d4-034fdb709b7a · outbound

This paper cites Llava-med: Training a large language- and-vision assistant for biomedicine in one day, 2023.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Llava-med: Training a large language- and-vision assistant for biomedicine in one day, 2023

Reference 19

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:df7d124a88191c52fdf7617bbe66c34cef4c2d10f350f57c5a0f1b1dd453d2b9

Observation c2ee01bd-0e98-4943-b73a-93bc9cdf388b · outbound

This paper cites Meddinov3: How to adapt vision foundation models for medical image segmentation?, 2025.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Meddinov3: How to adapt vision foundation models for medical image segmentation?, 2025

Reference 20

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:fcd18a04ea145acb34dd0fd1718642a57febe00be600d80853b077c8365350c6

Observation b6c84719-01e2-456f-9302-b544575a68be · outbound

This paper cites ROUGE: A package for automatic evaluation of summaries.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models ROUGE: A package for automatic evaluation of summaries

Reference 21

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:f5cb9b98b612d240050ac4d4ba3b03b352fa6bceb80489e141c933ba142f6e09

Observation 97f9b9d2-3043-4489-a972-1ad9f0f42121 · outbound

This paper cites Medical visual question answering: A survey.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Medical visual question answering: A survey

Reference 22

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:216c21aa5849f47fa1ab39e0a66ad859e1f08af51572db786efb8751d9d27e2a

Observation 1869c21f-ae86-4b43-b53c-d7fa1b948aee · outbound

This paper cites Swin Transformer: Hierarchical Vision Transformer using Shifted Windows.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Swin Transformer: Hierarchical Vision Transformer using Shifted Windows

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-07-02T11:36:55.007448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:37cfe11c4703fb91de103105701fd5d049d87c754788c5c0e20cef4ba8f11a2d

Observation e04beac6-0512-4291-82ff-5d90dc46169a · outbound

This paper cites Thakoor, Padraig Corcoran, Ying Chen, and Hantao Liu.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Thakoor, Padraig Corcoran, Ying Chen, and Hantao Liu

Reference 24

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:e9724cae10f8ec92ad9cb857fa69cca5bfb08f5dc2dda4047a65127b4cd3cadf

Observation 39f84702-46da-4619-a97c-15148d71ed61 · outbound

This paper cites Spot the Difference: Difference Visual Ques- tion Answering with Residual Alignment.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Spot the Difference: Difference Visual Ques- tion Answering with Residual Alignment

Reference 25

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:1402e6e3747f7db898617846895c7746304570f45479751adce47e3377c9cf3e

Observation 054d7f9f-76c4-4fc1-9a2b-c4b963f59e03 · outbound

This paper cites Segment anything in medical images.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Segment anything in medical images

Reference 26

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:bb806370c26f907da8d6436224589213322f2919369eb3d34bffe819c2b0655b

Observation 95b9bcda-96df-4f2f-b010-ca3fd161dde5 · outbound

This paper cites Unveiling differences: A vision encoder-decoder model for difference medical visual question answering.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Unveiling differences: A vision encoder-decoder model for difference medical visual question answering

Reference 27

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:cd19e26501b2773c9ba9610579c85a010ed2d4902b889807885630c74c308f12

Observation 264c93f3-360f-451d-8c74-1d24535bbdec · outbound

This paper cites Lon- gitudinal change detection on chest x-rays using geometric correlation maps.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Lon- gitudinal change detection on chest x-rays using geometric correlation maps

Reference 28

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:41fa8e5fbd115c2dfe30d18bf4c713bb7233ccad91ab184e96fb989de7aff215

Observation 88a22f4e-59b4-4205-81ff-69b79861cfe1 · outbound

This paper cites Dinov2: Learning robust visual features with- out supervision, 2024.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Dinov2: Learning robust visual features with- out supervision, 2024

Reference 29

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:69583c0224e679e1f22bcaefc0bc442b91ebb7a9cc0374a66fb1dd93fb185230

Observation 84981248-b1dc-4cbd-86d2-b784145ffa63 · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Bleu: a method for automatic evaluation of machine translation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:2b7e5f34953c9e2d797c339f08327954fc88dbff2b3d94d3436638feb69a8e53

Observation 18544029-ce04-4f5c-baff-4e53f8dff19c · outbound

This paper cites Castro, Anton Schwaighofer, Matthew P.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Castro, Anton Schwaighofer, Matthew P

Reference 31

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:01ced8cdc2385919474e11e3e392c90cf703aa3c7ef3587cb257184b88dacb64

Observation 591b912d-04b2-4651-a75c-3a7b14d4efc5 · outbound

This paper cites Language models are unsuper- vised multitask learners.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Language models are unsuper- vised multitask learners

Reference 32

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:9ca20bd08076669c6fc51663e47f3adc8efee86803316d4d656582ff554b09a6

Observation 02876894-caf1-4aea-8e3b-f09242f8ff4e · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models SAM 2: Segment Anything in Images and Videos

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-07-02T11:36:55.005058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:160fc5ffeae9fa430790fcef6dab1707541b784808855e33dfd593c4b1f76af9

Observation e278d090-5dee-4d61-a74c-3a428e676e0a · outbound

This paper cites Imagenet-21k pretraining for the masses.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Imagenet-21k pretraining for the masses

Reference 34

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:3a676b046be1023b372cb27f8597862c5dab02e64232cca05205fd9fa68702e9

Observation bdcb7e87-9a8c-4d19-b475-71c4813e17fe · outbound

This paper cites an unresolved cited work.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:85c8da93b1048d8931f766b5d5d875b1066e7aa8d62a881333d44e3b61bece44

Observation 8728e0ed-05eb-4e73-a973-b16ac678f260 · outbound

This paper cites Peeken, Daniel Rueckert, and Benedikt Wiestler.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Peeken, Daniel Rueckert, and Benedikt Wiestler

Reference 36

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:3088eda0440c57582b9ab75826c9a172070cd4cd46f6eb9c70af3b75f3162be0

Observation 622e632b-0da0-4e91-b66a-39d2d35b899f · outbound

This paper cites Selvaraju, Michael Cogswell, Abhishek Das, Ramakrishna Vedantam, Devi Parikh, and Dhruv Ba- tra.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Selvaraju, Michael Cogswell, Abhishek Das, Ramakrishna Vedantam, Devi Parikh, and Dhruv Ba- tra

Reference 37

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:7c6136f14a2894a81bcf40669dc1b39a66670faee7713f934e0446774763030c

Observation 077699b7-b671-4dd8-b622-4844c7919fd8 · outbound

This paper cites an unresolved cited work.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:046bd989f718aa5ce02c01d4f779b6fa3c68ac1383d3ba9fbb417bbc9d8666ab

Observation bf188445-87e5-498d-8228-447902e4cfc3 · outbound

This paper cites General pur- pose image encoder dinov2 for medical image registration,.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models General pur- pose image encoder dinov2 for medical image registration,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:a9c96d190c85862e0daf93843891135c432bf7df6d7af868460fba5512b79287

Observation 3c421942-e1cc-413e-a010-d6004bea366f · outbound

This paper cites Gomez, Lukasz Kaiser, and Illia Polosukhin.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Gomez, Lukasz Kaiser, and Illia Polosukhin

Reference 40

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:4d660d1d5eec7a1b528f136f68db24c49e2d809869b0b9fcfb9766f434e3e102

Observation 5860de1d-01fc-4911-9a7a-1468c976362f · outbound

This paper cites Lawrence Zitnick, and Devi Parikh.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Lawrence Zitnick, and Devi Parikh

Reference 41

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:c933b5763c3a703a8efb48cc17f21fca54d7b17574101a214dc5dcb2b79e6493

Observation ba046c6d-d957-4f3e-b218-e80c04d7e89e · outbound

This paper cites Co-attention aligned mutual cross-attention for cloth- changing person re-identification.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Co-attention aligned mutual cross-attention for cloth- changing person re-identification

Reference 42

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:d93be4e929a565d820f1ae8db5d1d9b1fe9d82c48b882ebd70923fe2fd565e8e

Observation f1749510-a4bd-4b68-921c-979ad0171957 · outbound

This paper cites Transformers: State-of-the-art natural language processing.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Transformers: State-of-the-art natural language processing

Reference 43

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:bf95dd660d2016e749e6e10408bb8b4fd9254c5c441442dc5fe21b17d7841210

Observation 84550c3e-a025-426c-b595-d866e21a55a7 · outbound

This paper cites Sabel, and Tobias Lasser.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Sabel, and Tobias Lasser

Reference 44

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:6ad81ce9eb98e29f9b7b1241bda39c4fe7ef6712d0105f588ad92bb3caeeb764

Observation e985809a-19be-4a5d-83e4-2e69caacb30d · outbound

This paper cites Segdino: An efficient design for medical and natural image segmentation with dino-v3, 2025.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Segdino: An efficient design for medical and natural image segmentation with dino-v3, 2025

Reference 45

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:e8b2bb6d847cceb55dc762a15d4f19371ccb46cd739fecba7cd9505bf285cc81

Observation ed7e4634-fac8-4f3c-a633-5b7e58a22d88 · outbound

This paper cites Exploring visual relationship for image captioning, 2018.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Exploring visual relationship for image captioning, 2018

Reference 46

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:8f1d17db04bd65d27af0e231167814018e8c42b5294927f2fd8626d11d4032ae

Observation 2bb3aff1-5a06-4a3d-a5e6-03b2ec3248a5 · outbound

This paper cites Describing and localizing multiple changes with transformers, 2021.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Describing and localizing multiple changes with transformers, 2021

Reference 47

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:255688c5b21a2c46a9e748963f5699980c8d63d5e641c92a4a793a5f1c6fbfee

Observation 6562b13e-0276-43d4-9b18-bc9c44f5c405 · outbound

This paper cites Mazomenos.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Mazomenos

Reference 48

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:77d63288431b6f23c57d90b554d285ac8279a4b8ae3fb6055c4edc4c0504bdc0

Observation 0e20a436-75d9-4705-a7d7-47152ef41d91 · outbound

This paper cites Lungren, Akshay Chaudhari, Ser- ena Yeung-Levy, Curtis P.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Lungren, Akshay Chaudhari, Ser- ena Yeung-Levy, Curtis P

Reference 49

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:43457990f00b229a4b667ea03007a54edc31d8b79aab4fe9b07dcdcae82ab101

Observation 415f5052-df69-4804-a1fe-c2197f10a732 · outbound

This paper cites PMC-VQA: Vi- sual instruction tuning for medical visual question answer- ing, 2024.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models PMC-VQA: Vi- sual instruction tuning for medical visual question answer- ing, 2024

Reference 50

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:b3af494e7c5b44a69cf58e346a8271248f20a946e6931dea628b7d55271397f9

Observation 90c129eb-c1cd-4b8c-bdb8-b777f38ed186 · outbound

This paper cites Medical sam 2: Segment medical images as video via segment anything model 2, 2024.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Medical sam 2: Segment medical images as video via segment anything model 2, 2024

Reference 51

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:9cdd93cd0af97c2ecededdb26fe56e8581de479c83f76f3f5769f62f6ec87040

Pith citing papers

No inbound Pith citation observations are available.