Pith. sign in

Paper Citation Record · LEDGER

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing

As of 7 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 1 inbound Pith citation observation for arXiv:2507.04333.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.04333 v1

Coverage vector

measured 53 of 53 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:54:42.017223Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:40:02.719163Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T15:40:04.174660Z

Reference resolution

53 of 53 outbound references displayed

  • verified exact0
  • verified fuzzy35
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation cdfddc5f-b514-4ce2-9404-2f56f22e5eaa · outbound

This paper cites Visual question answering in the medical domain based on deep learning approaches: A comprehensive study,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Visual question answering in the medical domain based on deep learning approaches: A comprehensive study,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:43.061470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:39.663180Z digest=sha256:e64cff563d14f6677f623e3d79199b6854e4dd97904b630583b9c3c9f46fdc50

Observation ddc0eda4-10d4-4da1-baf7-db45e5a6faca · outbound

This paper cites Vqa: Visual question answering,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Vqa: Visual question answering,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:43.043767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:39.709244Z digest=sha256:26b02201cc6748acd135718c9feb91228d9292e2f9aac95d968e077089dcd224

Observation e3ea3aeb-2a9b-4ddc-9717-3d168e17b137 · outbound

This paper cites M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T19:54:39.777120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:54:39.777120Z digest=sha256:909125c70f09b61990b26c7a9cd0b160eac6f1d65484d98bde833b9abcd427b7

Observation 4b92d35e-3bf6-46ba-9d77-a7d3eaaba66f · outbound

This paper cites Vqa-med: Overview of the medical visual question answering task at imageclef 2019,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Vqa-med: Overview of the medical visual question answering task at imageclef 2019,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:43.027402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:39.811230Z digest=sha256:7e31f96048ec6e3dba751fc99b709f1c4274708c419f4b7d64dc26d5289a7262

Observation 786cbd93-d63f-4ab0-8bfd-573add25b31d · outbound

This paper cites Simvqa: Exploring simulated environments for visual question answering,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Simvqa: Exploring simulated environments for visual question answering,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:43.008144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:39.907255Z digest=sha256:de686337d354ca3c5a17e98f8ac62e3b2eaaf7bc3d185a6d3bd4af66d9ada13e

Observation 9ffd5b60-c7e6-45fd-99ee-88b59877c3f3 · outbound

This paper cites Miss: A generative pre-training and fine-tuning approach for med-vqa,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Miss: A generative pre-training and fine-tuning approach for med-vqa,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:42.990324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:39.979389Z digest=sha256:ff929c327fca835b2b17a879ac1172619731cebd1986a54d927de42da0597dbc

Observation 29f587bf-5b17-4c0c-9b87-6a3ca3ce5fc5 · outbound

This paper cites ViT-V-Net: Vision Transformer for Unsupervised Volumetric Medical Image Registration.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing ViT-V-Net: Vision Transformer for Unsupervised Volumetric Medical Image Registration

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T19:54:40.019082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:54:40.019082Z digest=sha256:a07de32c1bca3c90fb76d12a7f85fce40988ace7a10e5c38c846959c22906250

Observation 7400de3f-c074-4436-9a0c-ca365b656cb0 · outbound

This paper cites Counterfactual samples synthesizing for robust visual question answering,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Counterfactual samples synthesizing for robust visual question answering,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:42.972457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:40.102925Z digest=sha256:d3445ad7877ed69e0e14bce61d8547263cd49038841113778551350a1e2349e0

Observation 9fd29176-115b-4431-987d-16ffce950dfc · outbound

This paper cites Generative bias for robust visual question answering,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Generative bias for robust visual question answering,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:42.955214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:40.264141Z digest=sha256:d8d3681ff18d48b4d97508322bf4db194bd7a1addecf1ce061465e8905110a0e

Observation 5b554d4f-c8ed-40fe-aa2b-b17b739ff30d · outbound

This paper cites BERT: Pre- training of Deep Bidirectional Transformers for Language Un- derstanding,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing BERT: Pre- training of Deep Bidirectional Transformers for Language Un- derstanding,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:42.938116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:40.394270Z digest=sha256:b4d46b15180d0e3925625d21de1fbaf16cd9e5a69fc7d9a0b07599b697d3a5e5

Observation cbc6d438-742b-4940-9d4e-762ac3f04bb1 · outbound

This paper cites Mukea: Mul- timodal knowledge extraction and accumulation for knowledge- based visual question answering,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Mukea: Mul- timodal knowledge extraction and accumulation for knowledge- based visual question answering,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:42.920561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:40.503431Z digest=sha256:52c96b5f95b09b7d9caeabeb6c9f45e08620d3ad56c1d20ef83d808caaed9e7e

Observation f0e9e18b-128d-4acf-93cb-99e080d08ab9 · outbound

This paper cites Cross-modal self- attention with multi-task pre-training for medical visual question answering,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Cross-modal self- attention with multi-task pre-training for medical visual question answering,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:42.903089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:40.551231Z digest=sha256:81b8cc5bc891a31472f5f4330454d80bdce6c67995149931cf121fb5c9c030ad

Observation b876a9f9-38b9-4a8e-9484-8c8da478034b · outbound

This paper cites Swapmix: Diagnosing and regularizing the over-reliance on vi- sual context in visual question answering,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Swapmix: Diagnosing and regularizing the over-reliance on vi- sual context in visual question answering,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:42.885488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:40.697192Z digest=sha256:1d0df5126f13418945b22b91cf9e52a78d1326d43bcb656d72d79069cfc10156

Observation 8a9d93d2-7440-4750-903d-2566f9cad3b4 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing LoRA: Low-Rank Adaptation of Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T19:54:40.822270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:54:40.822270Z digest=sha256:fbb9a6e11529bf5ea3d753044519cee2183b7c6ea659a512b674cfc1bd34c19d

Observation 64ad1017-e69b-42ae-876b-bae7a793728c · outbound

This paper cites Expert knowledge-aware image dif- ference graph representation learning for difference-aware med- ical visual question answering,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Expert knowledge-aware image dif- ference graph representation learning for difference-aware med- ical visual question answering,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:42.867569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:40.942191Z digest=sha256:eca7dd20c0bc8a3e9338406ad40aacac9fc6ee1b0633ecce38e97b5ab30c7175

Observation 906e60f6-5ff6-4529-b100-95a3014009bc · outbound

This paper cites Interpretable medical image visual question answering via multi-modal relationship graph learning,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Interpretable medical image visual question answering via multi-modal relationship graph learning,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:42.848294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:41.060460Z digest=sha256:2313ed2640d825c632f0ba468e40a8b875cd3f35ccafd08fba63f576b5f5da4c

Observation 77f33150-ef03-4d0b-87ad-4ccc5fcf1df1 · outbound

This paper cites Medical knowledge-based network for patient-oriented visual question answering,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Medical knowledge-based network for patient-oriented visual question answering,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:42.829605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:41.185112Z digest=sha256:e3f8d654d665075cab79565ea280940f880419ea5dd416fe9d8b8266a355638f

Observation 89b3a497-9e5f-45f0-a3de-9d2c292a3ee9 · outbound

This paper cites Mistral 7B.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Mistral 7B

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T19:54:41.237191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:54:41.237191Z digest=sha256:78e33685b85392a3cc94c400a562aafce1ba3dd2b11b69374f80edc1c19bc864

Observation e6561164-6493-4360-a892-f3787ed728d3 · outbound

This paper cites Semi-Supervised Classification with Graph Convolutional Networks.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Semi-Supervised Classification with Graph Convolutional Networks

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T19:54:41.321445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:54:41.321445Z digest=sha256:0172978969b1556a270e8cba3d233c23cddfff11efa1fa5187faeaa8f78024d6

Observation 113ddaa5-fc12-4d78-8e4d-26fc1b1faf5f · outbound

This paper cites Read Like a Radiologist: Efficient Vision-Language Model for 3D Medical Imaging Interpretation.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Read Like a Radiologist: Efficient Vision-Language Model for 3D Medical Imaging Interpretation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T19:54:41.430371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:54:41.430371Z digest=sha256:547b7877d49d7a0cb34ffaab5e6dfc46ecf57ad5edaaf342c1a272c3d4204224

Observation 8d7de044-3070-45b6-bfe0-20565c669ab1 · outbound

This paper cites Llava-Med: Training a large language-and-vision assistant for biomedicine in one day,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Llava-Med: Training a large language-and-vision assistant for biomedicine in one day,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:42.811977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:41.598103Z digest=sha256:9c101f17ffcacc9af97dfabb73e46b708bee0321ade08e0f6f3c5c2bd38d94a6

Observation f31ac35f-ec96-4c78-b9a4-ae30d892f3db · outbound

This paper cites Dynamic graph enhanced contrastive learning for chest x-ray report genera- tion,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Dynamic graph enhanced contrastive learning for chest x-ray report genera- tion,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:42.794830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:41.702930Z digest=sha256:3ff7fce30d0d4cd410779cb3f13913ba518184cf71323e5ee4623596eec78551

Observation b40eeff9-4701-49db-ab1d-f42fbb92e21c · outbound

This paper cites Masked vision and language pre-training with unimodal and multimodal contrastive losses for medical visual question answering,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Masked vision and language pre-training with unimodal and multimodal contrastive losses for medical visual question answering,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:42.778194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:41.758741Z digest=sha256:b124b3faa07edb18ef2600efec9d0b422802aec9e78d7237db26e11510ca39e0

Observation 21541f9a-325d-48c9-8b01-4df6ffe1a615 · outbound

This paper cites Oscar: Object-semantics aligned pre-training for vision-language tasks,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Oscar: Object-semantics aligned pre-training for vision-language tasks,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:42.759684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:41.839156Z digest=sha256:5da7c4d45adbca0466d32f1648ec1813d903c4c0eeffb4409050ff9aaf41c043

Observation 710d6250-ef31-4234-bfe7-764f87f9badd · outbound

This paper cites Revive: Regional visual representation matters in knowledge-based visual question answering,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Revive: Regional visual representation matters in knowledge-based visual question answering,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:42.740514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:41.845567Z digest=sha256:b5496b21527003c614afb714b9c5d2c7a33d0968f615b86b8c5b884f083630c2

Observation 72ba6c67-98ed-42ca-b63e-4123b107953e · outbound

This paper cites Medical visual question answering: A survey,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Medical visual question answering: A survey,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:42.722596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:41.850503Z digest=sha256:a4a579f8d848e355d021e782b531a4ad470ab930bf2c441adf60f6b44e69f997

Observation fb089f40-51d4-4c0a-bb68-89be47f4dbdc · outbound

This paper cites Slake: A semantically-labeled knowledge-enhanced dataset for medical visual question answering,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Slake: A semantically-labeled knowledge-enhanced dataset for medical visual question answering,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:42.705226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:41.855310Z digest=sha256:d9423bb5d47d93f8dd1c95c41a8dd4ffe41cb70f3535162507c4a8c2d86322db

Observation 821c37c8-e4e6-4e84-9883-afb1bf2bccfe · outbound

This paper cites Slake: A semantically-labeled knowledge-enhanced dataset for medical visual question answering,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Slake: A semantically-labeled knowledge-enhanced dataset for medical visual question answering,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:42.685304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:41.860397Z digest=sha256:e8a23a0a63422fde91dffa4e623d520e47fb01b5a7fffe5f288d15d57959e2fe

Observation 89d7ffc8-714c-48a7-b0a1-27b831079c4d · outbound

This paper cites Efficient Estimation of Word Representations in Vector Space.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Efficient Estimation of Word Representations in Vector Space

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T19:54:41.865463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:54:41.865463Z digest=sha256:6d188c85138c8e1c9a76c88237c825811676978baeb1f801d719ee390e877479

Observation b8127206-7314-425d-af16-acddf30e4f62 · outbound

This paper cites Beyond the Hype: A dispassionate look at vision-language models in medical scenario.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Beyond the Hype: A dispassionate look at vision-language models in medical scenario

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T19:54:41.871951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:54:41.871951Z digest=sha256:3439f1fdcc53b300785b9003bea79c3f8f761ea3ea7f6ad64a4da3388eb6c76d

Observation b4774f49-b889-499b-b9d0-9b830e42abf2 · outbound

This paper cites K-pathvqa: Knowledge-aware multimodal representation for pathology visual question answering,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing K-pathvqa: Knowledge-aware multimodal representation for pathology visual question answering,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:42.666927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:41.877348Z digest=sha256:5729bec728e40d06f431ff0c3aea8cbd84cba5441f4440a87c8af66da587e04a

Observation fd17081f-47fa-461e-ad15-2e96d563143b · outbound

This paper cites Relation Extraction with Word Graphs from N-grams,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Relation Extraction with Word Graphs from N-grams,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:42.648583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:41.883334Z digest=sha256:cc8733d457e293f705d864259c0c87fda6a86b517051d5f7923dbab789737afc

Observation 770c0f31-edf0-40ab-8ef9-19e9eb9feb72 · outbound

This paper cites An improved attention for visual question answering,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing An improved attention for visual question answering,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:42.629276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:41.890969Z digest=sha256:e636629e9c55fd0699c52cea67c5c1e0832985ef4ec4c510990f3b9ef4436eb6

Observation 148ae62c-5e42-49a0-bb31-065cdada7d89 · outbound

This paper cites Learning Word Representations with Regularization from Prior Knowledge,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Learning Word Representations with Regularization from Prior Knowledge,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:42.606857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:41.896806Z digest=sha256:0bc6119c6f297fafdb514a626416722450a4e2cbe1f74bc714fee663e5257ac6

Observation 4befd6a5-425a-48f1-a82f-efe8a053ba66 · outbound

This paper cites Complementary Learning of Word Em- beddings,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Complementary Learning of Word Em- beddings,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T19:54:41.903392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:54:41.903392Z digest=sha256:83fcb74121782e5011604e2a5bbfe262d805d187ab75c2c4ba474438b2195da4

Observation 488efb1a-27f0-4e35-ae39-7e62ebcce14d · outbound

This paper cites ZEN 2.0: Continue Training and Adaption for N-gram Enhanced Text Encoders.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing ZEN 2.0: Continue Training and Adaption for N-gram Enhanced Text Encoders

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T19:54:41.909167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:54:41.909167Z digest=sha256:2d18b0f4aafed94013d59b14b7d0d32d86e09965ab73fc9ab1e5641b3c9b1a2b

Observation 3dc585b8-0a6b-48f3-bfa9-262fd14aee85 · outbound

This paper cites VL-BERT: Pre-training of Generic Visual-Linguistic Representations.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing VL-BERT: Pre-training of Generic Visual-Linguistic Representations

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T19:54:41.914852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:54:41.914852Z digest=sha256:5418455e7382ad372d077dee513d274bcdf018a0ee694b512b896a6fbdd358a8

Observation e78a2eac-f36e-41f8-b89d-5a2ee88659d0 · outbound

This paper cites ChiMed-GPT: A Chinese medical large language model with full training regime and better alignment to human preferences,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing ChiMed-GPT: A Chinese medical large language model with full training regime and better alignment to human preferences,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:42.574834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:41.920856Z digest=sha256:6f4e933ee802aead205d6eb78f269258edd766fab336e0c8afd78a80a3e19974

Observation 0202e505-99c1-498b-bbe1-cecb04b40914 · outbound

This paper cites Chimed: A chinese medical corpus for question answering,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Chimed: A chinese medical corpus for question answering,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:42.555078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:41.926915Z digest=sha256:11070242a25b5b509a5b9ed61c5b29a55b523bfbcc2667dce60e972da9dcde0a

Observation 867720fa-cb2a-42bc-b8c2-e00437be774d · outbound

This paper cites Recurrent Visual Feature Extraction and Stereo Attentions for CT Report Generation.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Recurrent Visual Feature Extraction and Stereo Attentions for CT Report Generation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T19:54:41.933749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:54:41.933749Z digest=sha256:66ff214e152893672c9c75c141ccb5e8b10f654a6cfb8e01970aef7995648e7b

Observation b4ab7174-7f68-44d6-9f7a-88d57276949f · outbound

This paper cites Supertagging Combinatory Categorial Grammar with Attentive Graph Convolutional Networks,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Supertagging Combinatory Categorial Grammar with Attentive Graph Convolutional Networks,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:42.536940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:41.941352Z digest=sha256:573fab110a49eac476aeb2fd22723d605c16cb35c072c8542c3e83b9ffe53354

Observation de52edc8-5f81-4374-9404-9bc3dee63d54 · outbound

This paper cites Dialogue Summarization with Mix- ture of Experts based on Large Language Models,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Dialogue Summarization with Mix- ture of Experts based on Large Language Models,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:42.518424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:41.946962Z digest=sha256:bcaca532516f651435da5b5af007e1f35974848f5688eac107a3abb1352aa58c

Observation 63a2924d-a309-496b-98b9-688ae8d3b17f · outbound

This paper cites Diffusion networks with task-specific noise control for radiology report generation,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Diffusion networks with task-specific noise control for radiology report generation,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:42.495824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:41.953143Z digest=sha256:138a60717f5157ad8d3890dbda72f3c214abeebc28c7a6e1e68910e88f7cefa9

Observation 280cb60e-44f1-4dc4-bb1b-d84dfbd0a295 · outbound

This paper cites Learning multimodal contrast with cross-modal memory and reinforced contrast recognition,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Learning multimodal contrast with cross-modal memory and reinforced contrast recognition,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:42.476118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:41.959504Z digest=sha256:1eb26ebcec3d7f0600704e0d3dfff521b80637ed46d61d42a3548d2ab92484b3

Observation 6905bfd4-4f7f-46c7-9282-4b9daf92d9e5 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T19:54:41.966441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:54:41.966441Z digest=sha256:ed0bddb0d7dc7bf071f0b11546f0c2d27d3571018422eef728ce20404e29bad0

Observation 7dee86f9-1c66-46fc-bd8b-cd9f2f4090c1 · outbound

This paper cites Open-ended medical visual question answering through prefix tuning of language models,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Open-ended medical visual question answering through prefix tuning of language models,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:42.455142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:41.972313Z digest=sha256:961101e6eb78fe62b148b52eb7493f11408cc84db17a079fd50baa03da8596bc

Observation 81c6e27b-9066-43dd-b03f-e3a08abc30a6 · outbound

This paper cites Graph Attention Networks.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Graph Attention Networks

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T19:54:41.978028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:54:41.978028Z digest=sha256:94cffa924777f72c2255b3bd2bc6aff731e633a68c94e3df1a7ce092238bd353

Observation f3e37c72-b1bf-4a66-addf-2a7da3fb9152 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T19:54:41.985296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:54:41.985296Z digest=sha256:e21db2ecc9c03733f446d4c400daa7de2aeddefc7fe37ded94f17290f100d896

Observation 7666148e-657a-4fd3-88f8-2d53dcec53cb · outbound

This paper cites MedCLIP: Contrastive Learning from Unpaired Medical Images and Text,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing MedCLIP: Contrastive Learning from Unpaired Medical Images and Text,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:42.437037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:41.991989Z digest=sha256:55f955e4a8e6e8c71633d7ff0c0db47a3e863975ebb707d25b34107fa79f486f

Observation 11cc7b50-8dfb-4230-b2e1-2ecc30e4b106 · outbound

This paper cites Towards Generalist Foundation Model for Radiology by Leveraging Web-scale 2D&3D Medical Data.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing Towards Generalist Foundation Model for Radiology by Leveraging Web-scale 2D&3D Medical Data

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T19:54:41.997876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:54:41.997876Z digest=sha256:200f06e5d1ad3567223e268409704c5dc5c2e4da1ea39b041bd8d8f29b7ccb1b

Observation e315b091-23a7-4ebe-8bad-a4986323c8d1 · outbound

This paper cites BERTScore: Evaluating Text Generation with BERT.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing BERTScore: Evaluating Text Generation with BERT

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T19:54:42.003671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:54:42.003671Z digest=sha256:ec770a7573bf8a015a19c8214a370a2ef06451483fc4b8de032f9ad9d41ca664

Observation 07772605-64a3-477a-9f22-9b39bb5078b0 · outbound

This paper cites PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T19:54:42.010528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:54:42.010528Z digest=sha256:58d13ecf75fbc2909beae3875b326e2ba659f44466ae0fcf38f9b017de217d31

Observation b41a2c6c-2110-48c7-86bc-9b115a446bc7 · outbound

This paper cites When ra- diology report generation meets knowledge graph,.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing When ra- diology report generation meets knowledge graph,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:54:42.417984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:54:42.017223Z digest=sha256:8050389c253792065709a230f702ead3a01f9739694880da41cb3860a6121d58

Pith citing papers

Observation 031fe5a1-7450-4ea2-a2a6-e9de644ef2db · inbound

ChiMed 2.0: Advancing Chinese Medical Dataset in Facilitating Large Language Modeling cites this paper.

ChiMed 2.0: Advancing Chinese Medical Dataset in Facilitating Large Language Modeling Computed Tomography Visual Question Answering with Cross-modal Feature Graphing

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-06T15:40:04.259081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:40:02.719163Z digest=sha256:0b9ebf7d157c572f87cd95af99ee2c976a54d15e843b423408b5a7a19b65b211