Pith. sign in

Paper Citation Record · LEDGER

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models

As of 20 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 0 inbound Pith citation observations for arXiv:2606.06534.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.06534 v1

Coverage vector

measured 51 of 51 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-28T03:25:43.795795Z

measured 51 of 51 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

51 of 51 outbound references displayed

  • verified exact2
  • verified fuzzy0
  • unresolved49
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 74d47ca5-2c91-4fd7-9135-13ebe0635bce · outbound

This paper cites METEOR: An auto- matic metric for MT evaluation with improved correlation with human judgments.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models METEOR: An auto- matic metric for MT evaluation with improved correlation with human judgments

Reference 1

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:92c93da34fcf8b047cd362634467ed45938c16f0ff57282e15b645497440939f

Observation 2b2a52be-90a6-47cc-981b-714cc74ff394 · outbound

This paper cites Saliency-driven ex- plainable deep learning in medical imaging: Bridging visual explainability and statistical quantitative analysis.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Saliency-driven ex- plainable deep learning in medical imaging: Bridging visual explainability and statistical quantitative analysis

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:532eec38ffdabcc944d124119c155ce6aba3584e1baef13ed5e99c22b6379f1c

Observation 51275b1c-441c-42fe-858a-14b0949f2acf · outbound

This paper cites Swin-unet: Unet-like pure transformer for medical image segmentation,.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Swin-unet: Unet-like pure transformer for medical image segmentation,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:9357264d4f8a8f3980b58ce4f6ec5b6b12ffb74f5d5b21d6d389a70713af861f

Observation 91f781fa-0d2e-4c96-a0f2-1faa47920e78 · outbound

This paper cites Yuille, and Yuyin Zhou.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Yuille, and Yuyin Zhou

Reference 4

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:0cd859da8d52c63ed103081d1106f28bce05bfbe27c171a0db5e0abccda04353

Observation 7bd4b319-66be-4d07-bda1-d3f45ae4f9ca · outbound

This paper cites Pretraining vision-language model for difference visual question answering in longitudinal chest x-rays, 2024.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Pretraining vision-language model for difference visual question answering in longitudinal chest x-rays, 2024

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:0199a1f2a26044fce6450db6706064847046591355538653263fb95c230f86ba

Observation bc9f6c91-c73d-42d2-ba16-e37b245fb42e · outbound

This paper cites Co-saliency detection with co-attention fully convo- lutional network, 2020.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Co-saliency detection with co-attention fully convo- lutional network, 2020

Reference 6

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:1460b2c80f8fabe0424024dd3302922cf2e8d82ab890194886374997f3f358e1

Observation d181d43a-b611-4f7b-9ebc-9baad5b07ae4 · outbound

This paper cites Goldberger, Luis A.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Goldberger, Luis A

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:952fa2ba1ea03f738daddff67d76cc827869e3fe182dc75bf46ef394575078e7

Observation c86a62c0-8ce1-4c15-8aae-0d70fe088e7c · outbound

This paper cites Unetr: Transformers for 3d medical image segmentation, 2021.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Unetr: Transformers for 3d medical image segmentation, 2021

Reference 8

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:35ffe33434e0a72a215fbb74f4da65333666ef59fc58783782223d41fbbdd12c

Observation e3fe7c9d-c58e-420d-8700-a1da7bf06f7b · outbound

This paper cites Medical-Diff-VQA: A Large- Scale Medical Dataset for Difference Visual Question An- swering on Chest X-Ray Images.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Medical-Diff-VQA: A Large- Scale Medical Dataset for Difference Visual Question An- swering on Chest X-Ray Images

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:1fef93af9c250f7b7a402c590de51b96958ac9d327138b56dd646ff63bf19767

Observation 6c9d75ea-3233-47f6-98c4-78a0b99edb59 · outbound

This paper cites Summers, and Yingying Zhu.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Summers, and Yingying Zhu

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:2cfe4412f8da75d4431b20b299a1c1dca9def51e905d6c6803acdc1a3658930f

Observation 11dc461c-1060-4c72-9d04-ef181bcdabea · outbound

This paper cites Lungren, and Serena Yeung.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Lungren, and Serena Yeung

Reference 11

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:14f4f14ae5e43a1260cbcd089cae56833a68671c0251078bf504fecbd4e59b1a

Observation 934357a1-b883-4a80-93ac-f370c87c1121 · outbound

This paper cites Elsabawy, Alaa Ebraheem Elnakeeb, and Nora Elrashidy.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Elsabawy, Alaa Ebraheem Elnakeeb, and Nora Elrashidy

Reference 12

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:92ef33e43b97ec2bae0ddf5e427153456d68e33f7019f13068672f843edc27a0

Observation 27c63b98-c214-46b1-b89a-2cdd05acd10f · outbound

This paper cites Jaeger, Simon A.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Jaeger, Simon A

Reference 13

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:cf8353e5d419fbd5c14a5025be805b9592d77c364a65517acf66e876496404fb

Observation fde7716b-f967-4b1b-a077-940cebdcc29b · outbound

This paper cites Spatial transformer networks, 2016.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Spatial transformer networks, 2016

Reference 14

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:0ba07063a6ba97a30a022343199303373abe25d3aad8f2eee1d3fdd0ad8b6551

Observation 0523bd36-fb98-4235-93fd-531a21b51864 · outbound

This paper cites One map does not fit all: Evaluating saliency map explanation on multi-modal medical images, 2021.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models One map does not fit all: Evaluating saliency map explanation on multi-modal medical images, 2021

Reference 15

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:03eccb063c134279a2a9649078221d5db837ff899f01f3ed58386709c3613311

Observation d5a853e3-44ae-4802-9425-e56c169b3966 · outbound

This paper cites an unresolved cited work.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:a450d5715938b3e8207d5530f2692bce1266aea8038222d52f0c026fca2a0683

Observation b4688c77-7f5f-4380-874a-6b81748c4380 · outbound

This paper cites an unresolved cited work.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:dfcfc47885db2f1511fd8f3367bce9fd1b67a256173e4604f7c576ab1538afb1

Observation 8220432c-8e88-4bc3-a533-d3fdacf5f3a5 · outbound

This paper cites Schroeder, and Tolga Tasdizen.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Schroeder, and Tolga Tasdizen

Reference 18

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:16b3453a63cb994f2265e3280b9eedc13dafbf3df10f27fb16efa44ea0f5cc2e

Observation 1f53a779-0798-4d8b-85d4-034fdb709b7a · outbound

This paper cites Llava-med: Training a large language- and-vision assistant for biomedicine in one day, 2023.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Llava-med: Training a large language- and-vision assistant for biomedicine in one day, 2023

Reference 19

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:7d44b1342f2f3330dd0d856020f9e94184b37dff956d4065d76b2747f43f1d81

Observation c2ee01bd-0e98-4943-b73a-93bc9cdf388b · outbound

This paper cites Meddinov3: How to adapt vision foundation models for medical image segmentation?, 2025.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Meddinov3: How to adapt vision foundation models for medical image segmentation?, 2025

Reference 20

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:1f459e2243361338ed6db78e81e9c098bb2e1bba53b229f9aea5fcb98b3edc6f

Observation b6c84719-01e2-456f-9302-b544575a68be · outbound

This paper cites ROUGE: A package for automatic evaluation of summaries.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models ROUGE: A package for automatic evaluation of summaries

Reference 21

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:c69f4c9fcf1733ae3adc99a6b57f19209bff3589cbfd88d50295727fe1305fc9

Observation 97f9b9d2-3043-4489-a972-1ad9f0f42121 · outbound

This paper cites Medical visual question answering: A survey.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Medical visual question answering: A survey

Reference 22

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:35c5e426e2c578627c4b25a762a19bb3c80ace10c2f9a503084436f72f17a882

Observation 1869c21f-ae86-4b43-b53c-d7fa1b948aee · outbound

This paper cites Swin Transformer: Hierarchical Vision Transformer using Shifted Windows.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Swin Transformer: Hierarchical Vision Transformer using Shifted Windows

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-07-02T11:36:55.007448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:1569319a8c39061ec86745058a6a1256888c7de6dbc39ecb7d26ea9611f6e555

Observation e04beac6-0512-4291-82ff-5d90dc46169a · outbound

This paper cites Thakoor, Padraig Corcoran, Ying Chen, and Hantao Liu.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Thakoor, Padraig Corcoran, Ying Chen, and Hantao Liu

Reference 24

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:a233094e3c7c1f96b0ea39939b8cc27633f43c6b7427bc8b37a990bff53849d3

Observation 39f84702-46da-4619-a97c-15148d71ed61 · outbound

This paper cites Spot the Difference: Difference Visual Ques- tion Answering with Residual Alignment.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Spot the Difference: Difference Visual Ques- tion Answering with Residual Alignment

Reference 25

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:edf31e2dec3841c6747534903d5b0c59422424363968ce6cd95918d7df95aa28

Observation 054d7f9f-76c4-4fc1-9a2b-c4b963f59e03 · outbound

This paper cites Segment anything in medical images.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Segment anything in medical images

Reference 26

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:8792306b0f1321b1e14710f63c2c6fd390ea06984104719e1f70171bcfa8ef0a

Observation 95b9bcda-96df-4f2f-b010-ca3fd161dde5 · outbound

This paper cites Unveiling differences: A vision encoder-decoder model for difference medical visual question answering.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Unveiling differences: A vision encoder-decoder model for difference medical visual question answering

Reference 27

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:1fe6f8e7f2fbf509f8f20fad0cd259d5ced8b84812531690656658211b99cb2a

Observation 264c93f3-360f-451d-8c74-1d24535bbdec · outbound

This paper cites Lon- gitudinal change detection on chest x-rays using geometric correlation maps.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Lon- gitudinal change detection on chest x-rays using geometric correlation maps

Reference 28

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:dd443dbe5b859aea2ac31bf5327bc98a344b5e7610f3171bd86b3531f814609f

Observation 88a22f4e-59b4-4205-81ff-69b79861cfe1 · outbound

This paper cites Dinov2: Learning robust visual features with- out supervision, 2024.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Dinov2: Learning robust visual features with- out supervision, 2024

Reference 29

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:ce1f6994c99d3681da38c74a0a354647ffeac151191fbf8e296773a3b7efd9a9

Observation 84981248-b1dc-4cbd-86d2-b784145ffa63 · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Bleu: a method for automatic evaluation of machine translation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:7a90ce17f27a69ce976ae1042250ab75c82da52824ab0ecb3fab5e18b954e14e

Observation 18544029-ce04-4f5c-baff-4e53f8dff19c · outbound

This paper cites Castro, Anton Schwaighofer, Matthew P.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Castro, Anton Schwaighofer, Matthew P

Reference 31

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:b9213b19bd87215b5c864a6fdc9742dec56c987b31e1eec7474020752e502f31

Observation 591b912d-04b2-4651-a75c-3a7b14d4efc5 · outbound

This paper cites Language models are unsuper- vised multitask learners.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Language models are unsuper- vised multitask learners

Reference 32

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:a18be4a93b3aef24c2cd93a0d4ad3edd0eff3ef6fc4162890520ecaa853130af

Observation 02876894-caf1-4aea-8e3b-f09242f8ff4e · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models SAM 2: Segment Anything in Images and Videos

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-07-02T11:36:55.005058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:71bc86a899a5c6f5416ac1fe7f36e48152438a8358e0531b21fddcb6630d3048

Observation e278d090-5dee-4d61-a74c-3a428e676e0a · outbound

This paper cites Imagenet-21k pretraining for the masses.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Imagenet-21k pretraining for the masses

Reference 34

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:523dd6d69cfd21ae9d3c19cd07c6c1c4dc7aa0a8090f23086ffad801840cfc2e

Observation bdcb7e87-9a8c-4d19-b475-71c4813e17fe · outbound

This paper cites an unresolved cited work.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:84898245a3b2ff94f0be69ed381e668548027fecc2f10e4dab6cb6945e29a1d8

Observation 8728e0ed-05eb-4e73-a973-b16ac678f260 · outbound

This paper cites Peeken, Daniel Rueckert, and Benedikt Wiestler.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Peeken, Daniel Rueckert, and Benedikt Wiestler

Reference 36

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:5d662b58ae36a1dedef02c7af3d62038d6b8adcb999df7f5c691276aa971e047

Observation 622e632b-0da0-4e91-b66a-39d2d35b899f · outbound

This paper cites Selvaraju, Michael Cogswell, Abhishek Das, Ramakrishna Vedantam, Devi Parikh, and Dhruv Ba- tra.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Selvaraju, Michael Cogswell, Abhishek Das, Ramakrishna Vedantam, Devi Parikh, and Dhruv Ba- tra

Reference 37

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:a07ab3a7ce8f9e1d481cd2dacc85415825088f62416d331fec4b63fab97def5a

Observation 077699b7-b671-4dd8-b622-4844c7919fd8 · outbound

This paper cites an unresolved cited work.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:a0038dbf21bfa816de2fb2c7e27c2f9807a00a173a5c8f41d4884a5ffa4952d8

Observation bf188445-87e5-498d-8228-447902e4cfc3 · outbound

This paper cites General pur- pose image encoder dinov2 for medical image registration,.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models General pur- pose image encoder dinov2 for medical image registration,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:c855eedb6c86816542ba2440d02c564d9c6282726c3c8ec86cdb078935a2c2f0

Observation 3c421942-e1cc-413e-a010-d6004bea366f · outbound

This paper cites Gomez, Lukasz Kaiser, and Illia Polosukhin.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Gomez, Lukasz Kaiser, and Illia Polosukhin

Reference 40

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:10a3e8ec51fdf746b728b53e8a57186f3fc74b5bb11fd0ca0fd1b0adfb30241f

Observation 5860de1d-01fc-4911-9a7a-1468c976362f · outbound

This paper cites Lawrence Zitnick, and Devi Parikh.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Lawrence Zitnick, and Devi Parikh

Reference 41

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:1e849f6894b1e0b752409d9a81b8c4b3608950fcd5f9cc0bfab7109f1220d469

Observation ba046c6d-d957-4f3e-b218-e80c04d7e89e · outbound

This paper cites Co-attention aligned mutual cross-attention for cloth- changing person re-identification.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Co-attention aligned mutual cross-attention for cloth- changing person re-identification

Reference 42

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:598bb539b63a43fa788ab5283899e241a53b194b241497051d4e67e01cf200ae

Observation f1749510-a4bd-4b68-921c-979ad0171957 · outbound

This paper cites Transformers: State-of-the-art natural language processing.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Transformers: State-of-the-art natural language processing

Reference 43

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:e06095d1a89360a9c3beab780458bfdf0696a6190d69023b8c91d1104d73eb2d

Observation 84550c3e-a025-426c-b595-d866e21a55a7 · outbound

This paper cites Sabel, and Tobias Lasser.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Sabel, and Tobias Lasser

Reference 44

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:24a8f3d0501612da8e72991c056916452cc16719099a435f1e8f6f8803e71fa2

Observation e985809a-19be-4a5d-83e4-2e69caacb30d · outbound

This paper cites Segdino: An efficient design for medical and natural image segmentation with dino-v3, 2025.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Segdino: An efficient design for medical and natural image segmentation with dino-v3, 2025

Reference 45

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:01aa03c846fefc06bebb53e46de7ffe7c19452269094fc85c02f30c7d0900bdc

Observation ed7e4634-fac8-4f3c-a633-5b7e58a22d88 · outbound

This paper cites Exploring visual relationship for image captioning, 2018.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Exploring visual relationship for image captioning, 2018

Reference 46

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:4a539790f47d796f6968062c7f0b0291821bdfe11dbe58b8da69a02f5b7a786f

Observation 2bb3aff1-5a06-4a3d-a5e6-03b2ec3248a5 · outbound

This paper cites Describing and localizing multiple changes with transformers, 2021.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Describing and localizing multiple changes with transformers, 2021

Reference 47

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:381c4900d67526da227120eaaf3c19b98c2a99d50d71a84b8523d52ee40226c8

Observation 6562b13e-0276-43d4-9b18-bc9c44f5c405 · outbound

This paper cites Mazomenos.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Mazomenos

Reference 48

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:096646ce967acd1587eb7f820a550523af8be4d5a754b576079e618ab59da8c4

Observation 0e20a436-75d9-4705-a7d7-47152ef41d91 · outbound

This paper cites Lungren, Akshay Chaudhari, Ser- ena Yeung-Levy, Curtis P.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Lungren, Akshay Chaudhari, Ser- ena Yeung-Levy, Curtis P

Reference 49

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:0d271653f7bc8935244edbe9cb4c39cf7ebe06cc204954934324e4a7baf15179

Observation 415f5052-df69-4804-a1fe-c2197f10a732 · outbound

This paper cites PMC-VQA: Vi- sual instruction tuning for medical visual question answer- ing, 2024.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models PMC-VQA: Vi- sual instruction tuning for medical visual question answer- ing, 2024

Reference 50

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:f51a526fce893a7fbd2ecc3f5bc64cdf0b64364aceab62d3b0e487624d41e91d

Observation 90c129eb-c1cd-4b8c-bdb8-b777f38ed186 · outbound

This paper cites Medical sam 2: Segment medical images as video via segment anything model 2, 2024.

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models Medical sam 2: Segment medical images as video via segment anything model 2, 2024

Reference 51

Resolution
unresolved
no resolver link, observed 2026-06-28T03:25:43.795795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T03:25:43.795795Z digest=sha256:b37eee2b660aa27fde03be0e62bafba5476cac4024da2d9376ac127459ff5427

Pith citing papers

No inbound Pith citation observations are available.