Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T19:54:42.017223Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 1 inbound Pith citation observation for arXiv:2507.04333.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T19:54:42.017223Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T15:40:02.719163Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-06T15:40:04.174660Z
53 of 53 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation cdfddc5f-b514-4ce2-9404-2f56f22e5eaa · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Visual question answering in the medical domain based on deep learning approaches: A comprehensive study,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ddc0eda4-10d4-4da1-baf7-db45e5a6faca · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Vqa: Visual question answering,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e3ea3aeb-2a9b-4ddc-9717-3d168e17b137 · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b92d35e-3bf6-46ba-9d77-a7d3eaaba66f · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Vqa-med: Overview of the medical visual question answering task at imageclef 2019,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 786cbd93-d63f-4ab0-8bfd-573add25b31d · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Simvqa: Exploring simulated environments for visual question answering,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9ffd5b60-c7e6-45fd-99ee-88b59877c3f3 · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Miss: A generative pre-training and fine-tuning approach for med-vqa,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 29f587bf-5b17-4c0c-9b87-6a3ca3ce5fc5 · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing ViT-V-Net: Vision Transformer for Unsupervised Volumetric Medical Image Registration
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7400de3f-c074-4436-9a0c-ca365b656cb0 · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Counterfactual samples synthesizing for robust visual question answering,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9fd29176-115b-4431-987d-16ffce950dfc · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Generative bias for robust visual question answering,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5b554d4f-c8ed-40fe-aa2b-b17b739ff30d · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing BERT: Pre- training of Deep Bidirectional Transformers for Language Un- derstanding,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cbc6d438-742b-4940-9d4e-762ac3f04bb1 · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Mukea: Mul- timodal knowledge extraction and accumulation for knowledge- based visual question answering,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f0e9e18b-128d-4acf-93cb-99e080d08ab9 · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Cross-modal self- attention with multi-task pre-training for medical visual question answering,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b876a9f9-38b9-4a8e-9484-8c8da478034b · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Swapmix: Diagnosing and regularizing the over-reliance on vi- sual context in visual question answering,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8a9d93d2-7440-4750-903d-2566f9cad3b4 · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing LoRA: Low-Rank Adaptation of Large Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64ad1017-e69b-42ae-876b-bae7a793728c · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Expert knowledge-aware image dif- ference graph representation learning for difference-aware med- ical visual question answering,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 906e60f6-5ff6-4529-b100-95a3014009bc · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Interpretable medical image visual question answering via multi-modal relationship graph learning,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 77f33150-ef03-4d0b-87ad-4ccc5fcf1df1 · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Medical knowledge-based network for patient-oriented visual question answering,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 89b3a497-9e5f-45f0-a3de-9d2c292a3ee9 · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Mistral 7B
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6561164-6493-4360-a892-f3787ed728d3 · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Semi-Supervised Classification with Graph Convolutional Networks
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 113ddaa5-fc12-4d78-8e4d-26fc1b1faf5f · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Read Like a Radiologist: Efficient Vision-Language Model for 3D Medical Imaging Interpretation
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d7de044-3070-45b6-bfe0-20565c669ab1 · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Llava-Med: Training a large language-and-vision assistant for biomedicine in one day,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f31ac35f-ec96-4c78-b9a4-ae30d892f3db · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Dynamic graph enhanced contrastive learning for chest x-ray report genera- tion,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b40eeff9-4701-49db-ab1d-f42fbb92e21c · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Masked vision and language pre-training with unimodal and multimodal contrastive losses for medical visual question answering,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 21541f9a-325d-48c9-8b01-4df6ffe1a615 · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Oscar: Object-semantics aligned pre-training for vision-language tasks,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 710d6250-ef31-4234-bfe7-764f87f9badd · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Revive: Regional visual representation matters in knowledge-based visual question answering,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 72ba6c67-98ed-42ca-b63e-4123b107953e · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Medical visual question answering: A survey,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fb089f40-51d4-4c0a-bb68-89be47f4dbdc · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Slake: A semantically-labeled knowledge-enhanced dataset for medical visual question answering,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 821c37c8-e4e6-4e84-9883-afb1bf2bccfe · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Slake: A semantically-labeled knowledge-enhanced dataset for medical visual question answering,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 89d7ffc8-714c-48a7-b0a1-27b831079c4d · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Efficient Estimation of Word Representations in Vector Space
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8127206-7314-425d-af16-acddf30e4f62 · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Beyond the Hype: A dispassionate look at vision-language models in medical scenario
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4774f49-b889-499b-b9d0-9b830e42abf2 · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing K-pathvqa: Knowledge-aware multimodal representation for pathology visual question answering,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fd17081f-47fa-461e-ad15-2e96d563143b · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Relation Extraction with Word Graphs from N-grams,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 770c0f31-edf0-40ab-8ef9-19e9eb9feb72 · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing An improved attention for visual question answering,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 148ae62c-5e42-49a0-bb31-065cdada7d89 · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Learning Word Representations with Regularization from Prior Knowledge,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4befd6a5-425a-48f1-a82f-efe8a053ba66 · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Complementary Learning of Word Em- beddings,
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 488efb1a-27f0-4e35-ae39-7e62ebcce14d · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing ZEN 2.0: Continue Training and Adaption for N-gram Enhanced Text Encoders
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3dc585b8-0a6b-48f3-bfa9-262fd14aee85 · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing VL-BERT: Pre-training of Generic Visual-Linguistic Representations
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e78a2eac-f36e-41f8-b89d-5a2ee88659d0 · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing ChiMed-GPT: A Chinese medical large language model with full training regime and better alignment to human preferences,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0202e505-99c1-498b-bbe1-cecb04b40914 · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Chimed: A chinese medical corpus for question answering,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 867720fa-cb2a-42bc-b8c2-e00437be774d · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Recurrent Visual Feature Extraction and Stereo Attentions for CT Report Generation
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4ab7174-7f68-44d6-9f7a-88d57276949f · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Supertagging Combinatory Categorial Grammar with Attentive Graph Convolutional Networks,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation de52edc8-5f81-4374-9404-9bc3dee63d54 · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Dialogue Summarization with Mix- ture of Experts based on Large Language Models,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 63a2924d-a309-496b-98b9-688ae8d3b17f · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Diffusion networks with task-specific noise control for radiology report generation,
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 280cb60e-44f1-4dc4-bb1b-d84dfbd0a295 · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Learning multimodal contrast with cross-modal memory and reinforced contrast recognition,
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6905bfd4-4f7f-46c7-9282-4b9daf92d9e5 · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7dee86f9-1c66-46fc-bd8b-cd9f2f4090c1 · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Open-ended medical visual question answering through prefix tuning of language models,
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 81c6e27b-9066-43dd-b03f-e3a08abc30a6 · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Graph Attention Networks
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3e37c72-b1bf-4a66-addf-2a7da3fb9152 · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7666148e-657a-4fd3-88f8-2d53dcec53cb · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing MedCLIP: Contrastive Learning from Unpaired Medical Images and Text,
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 11cc7b50-8dfb-4230-b2e1-2ecc30e4b106 · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Towards Generalist Foundation Model for Radiology by Leveraging Web-scale 2D&3D Medical Data
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e315b091-23a7-4ebe-8bad-a4986323c8d1 · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing BERTScore: Evaluating Text Generation with BERT
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07772605-64a3-477a-9f22-9b39bb5078b0 · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b41a2c6c-2110-48c7-86bc-9b115a446bc7 · outbound
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing When ra- diology report generation meets knowledge graph,
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 031fe5a1-7450-4ea2-a2a6-e9de644ef2db · inbound
ChiMed 2.0: Advancing Chinese Medical Dataset in Facilitating Large Language Modeling Computed Tomography Visual Question Answering with Cross-modal Feature Graphing
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.