Pith. sign in

Paper Citation Record · LEDGER

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting

As of 7 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 2 inbound Pith citation observations for arXiv:2509.03800.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.03800 v1

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T10:46:09.996574Z

measured 62 of 62 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T19:34:04.544567Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-28T19:22:34.657932Z

Reference resolution

60 of 60 outbound references displayed

  • verified exact1
  • verified fuzzy39
  • unresolved20
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a6fa80d1-f5d6-4f4a-8c11-f7590af040bf · outbound

This paper cites Merlin: A vision language foundation model for 3d computed tomography.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Merlin: A vision language foundation model for 3d computed tomography

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:04.243549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:04.243549Z digest=sha256:72fa11f8830f5b82beb81d8549065cd7d0901e5dc6d0c6287dd32cd376333264

Observation b33a7ab9-9f9f-45d3-8445-97b33ff2b16b · outbound

This paper cites A vision–language foundation model for the generation of realistic chest x-ray images.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting A vision–language foundation model for the generation of realistic chest x-ray images

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.958572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:04.330865Z digest=sha256:569ca31b82ce1a8bdc7e5f52ae7f99e8e77cb5ad1ec3a77ddce4dcf190a432a7

Observation cb37fff5-0668-4885-9707-e0ec24f3015e · outbound

This paper cites Making the most of text semantics to improve biomedical vision–language processing.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Making the most of text semantics to improve biomedical vision–language processing

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:04.426003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:04.426003Z digest=sha256:16316adc192f52129915266b6d1cabb503505116588039e4128a9b302865fe43

Observation d404663f-9670-4008-8297-2b334af143b6 · outbound

This paper cites Understanding and confronting our mis- takes: the epidemiology of error in radiology and strategies for error reduction.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Understanding and confronting our mis- takes: the epidemiology of error in radiology and strategies for error reduction

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.932829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:04.516487Z digest=sha256:ec32d44d4c05a504591cb5facc99da9ff1c590dca8cd9d83321d9ef23a9c2595

Observation c143d303-e11e-4e10-bd16-89cf43f4e210 · outbound

This paper cites Joint modeling of chest radiographs and radiology reports for pulmonary edema assessment.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Joint modeling of chest radiographs and radiology reports for pulmonary edema assessment

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.917747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:04.585735Z digest=sha256:fe3dee927d0b94f8388fabb7a5abbfd4d224391ef6deac390b381ec46aae2e6e

Observation 41ed5110-7327-4e3d-bafa-1a1622ef2ee5 · outbound

This paper cites Contrastive Localized Language-Image Pre-Training.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Contrastive Localized Language-Image Pre-Training

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:04.661105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:04.661105Z digest=sha256:436fd92b0532fd11ea56fd94ffdddab7422f23acbbc42a8dec8ee285db120d48

Observation 3f111173-6cf8-4d07-8f39-b1d2c4ddf619 · outbound

This paper cites A review of medical image data augmentation techniques for deep learning applications.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting A review of medical image data augmentation techniques for deep learning applications

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.900852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:04.776379Z digest=sha256:fcf7db3518904b044201edd3495361105c43017f1d488798213010a917913504

Observation f555adc7-97f4-414c-8a7b-5d9ba7d8b2d8 · outbound

This paper cites Machine-learning-based multiple abnormality prediction with large-scale chest computed tomography volumes.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Machine-learning-based multiple abnormality prediction with large-scale chest computed tomography volumes

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.884634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:04.842478Z digest=sha256:475afe3458acbf4e7c8fc166185cc59cf83a595be09d1cf68ffb2fe81713eaaf

Observation 85df312c-28df-4027-8a32-0cf6486e38a7 · outbound

This paper cites Multi-View and Multi-Scale Alignment for Contrastive Language-Image Pre-training in Mammography.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Multi-View and Multi-Scale Alignment for Contrastive Language-Image Pre-training in Mammography

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-05T10:46:10.536952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:04.962958Z digest=sha256:136227593cd3536be2ebb02df1e5cd4ad54dbb3bbdc818422a03eec93a0e2d6d

Observation 4feab371-5592-4901-b37e-221704c40765 · outbound

This paper cites The Llama 3 Herd of Models.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting The Llama 3 Herd of Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:05.092590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:05.092590Z digest=sha256:a913edd078b6cb4a5e86d2344bcd8876b4a32e3177c2c77d5c9fd245b58e43ef

Observation 13ac3f6d-a019-4f0b-825c-d36d7414fa43 · outbound

This paper cites Devel- oping generalist foundation models from a multimodal dataset for 3d computed tomography.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Devel- oping generalist foundation models from a multimodal dataset for 3d computed tomography

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:05.191575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:05.191575Z digest=sha256:2cb76af8c525d5b12b3378de30af39e5d237c3d0c86f7c996f2e51df93f1a747

Observation ac57e9c6-b566-4905-a756-dd5818e81d92 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting LoRA: Low-Rank Adaptation of Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:05.296884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:05.296884Z digest=sha256:2334b1e0110c695c2dc7b2551e0dae4e0a30a9383632cd1d96b6a5573528779c

Observation aa273c70-1ad5-4176-a12e-25ee2a518a90 · outbound

This paper cites Gloria: A multimodal global-local representation learning framework for label-efficient medical image recognition.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Gloria: A multimodal global-local representation learning framework for label-efficient medical image recognition

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:05.360284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:05.360284Z digest=sha256:0af129921b0578143c0f0ae027d1efe21c605145f54d7a5e33f1181d876682c3

Observation c7bf766a-ec65-41da-a482-156f75f1b8da · outbound

This paper cites Enhancing representation in medical vision-language foun- dation models via multi-scale information extraction techniques.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Enhancing representation in medical vision-language foun- dation models via multi-scale information extraction techniques

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.859199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:05.458401Z digest=sha256:97fc38fff9526e5daf74d9ab396c08e0c409ea06e0e7b1d44ad2839d99b5308e

Observation da68f205-0616-4a33-8ce2-23eb68858145 · outbound

This paper cites STU-Net: Scalable and Transferable Medical Image Segmentation Models Empowered by Large-Scale Supervised Pre-training.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting STU-Net: Scalable and Transferable Medical Image Segmentation Models Empowered by Large-Scale Supervised Pre-training

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:05.562057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:05.562057Z digest=sha256:877b3c80a15e2d11135158a395f68bfbd1f36c79f0e5c745abc4c840c5aab4da

Observation 4dbc9100-63b5-4b79-b258-fb0395cb7c62 · outbound

This paper cites nnu-net: a self-configuring method for deep learning-based biomedical image segmentation.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting nnu-net: a self-configuring method for deep learning-based biomedical image segmentation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:05.633725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:05.633725Z digest=sha256:3cc8af1f9af61afb361cb8d534c687df0398f88062acb0dadd82e15e19e5ac7c

Observation 0d20a01c-58e5-4de6-8e64-28d7416906b8 · outbound

This paper cites Fool me twice: delayed diagnoses in radiology with emphasis on perpetuated errors.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Fool me twice: delayed diagnoses in radiology with emphasis on perpetuated errors

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.834498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:05.740237Z digest=sha256:6aa827fcef7848e6c2eb05b32ab95d26f030a1288981e61ab831445be7c8c116

Observation 0ef9eb6e-3c40-4f58-8234-d3ac1161d8f4 · outbound

This paper cites Generating synthetic data for medical imaging.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Generating synthetic data for medical imaging

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.819212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:05.837515Z digest=sha256:d67bf2a06f83c2fc31f668715a508057020b98e9aee817e72efbfe932fa10d5e

Observation 35d0a13e-01ef-4672-bf63-cc177925cee6 · outbound

This paper cites Cxr-llava: a multimodal large language model for interpreting chest x-ray images.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Cxr-llava: a multimodal large language model for interpreting chest x-ray images

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.803136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:05.952126Z digest=sha256:94d91d3049d80b7cb4e88e482c26870fd55186c6440527f9b16b6352ceb34f2d

Observation 71cb91cd-edae-46b0-9c62-195ae4964bf7 · outbound

This paper cites Llava-med: Training a large language-and-vision assistant for biomedicine in one day.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Llava-med: Training a large language-and-vision assistant for biomedicine in one day

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:06.026153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:06.026153Z digest=sha256:7ccc36992eee25d94a0371009013a8a1505ca3a72d418efe8ee38a5400b3ce7f

Observation 6b547427-1974-4c16-a177-2ff59895bdc3 · outbound

This paper cites Artificial general intelligence for medical imaging analysis.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Artificial general intelligence for medical imaging analysis

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.775179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:06.144188Z digest=sha256:bd3eb0d6e0a4b498fd47bc54f53a548b92d3b1da1312e11c62ecf0d0680ec4dd

Observation 9920dd3d-25af-4526-9efc-147946bfd878 · outbound

This paper cites Ct-glip: 3d grounded language-image pretraining with ct scans and radiology reports for full-body scenarios.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Ct-glip: 3d grounded language-image pretraining with ct scans and radiology reports for full-body scenarios

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:06.214147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:06.214147Z digest=sha256:79a16a32ebd837b088a7050ec625c6c0226ddd5a1356c9e839e2cca1fbacc9c2

Observation 045e8575-14d9-458f-90dc-c8791db2b0af · outbound

This paper cites Pmc-clip: Contrastive language-image pre-training using biomedical documents.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Pmc-clip: Contrastive language-image pre-training using biomedical documents

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:06.347751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:06.347751Z digest=sha256:ae1bbad12fa065f23e158a16413df80eab286e9f66d012c8360a8fb0d0dc3f09

Observation 56045d27-ec37-4ba5-a6bd-9f2a0988974a · outbound

This paper cites Can Medical Vision-Language Pre-training Succeed with Purely Synthetic Data?.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Can Medical Vision-Language Pre-training Succeed with Purely Synthetic Data?

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:06.490159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:06.490159Z digest=sha256:64c93a6059eff9963f8f3f0ae540aff850601fcc5ce39bf4d43cdca258ca2232

Observation 7cfb17f5-68a2-4784-ae44-4289145e1d0d · outbound

This paper cites Improved baselines with visual instruction tuning.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Improved baselines with visual instruction tuning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:06.654963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:06.654963Z digest=sha256:6f91cb40f6cb7937ea3f6816d846d5519593c892372d1f6354e98ea952f51d27

Observation 08c8a7a7-1b04-4673-b65c-8d991d660f66 · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Representation Learning with Contrastive Predictive Coding

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:06.818244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:06.818244Z digest=sha256:778a40b282617a7c96b823c361a0b39cf31de5f486d76eb8541dd0cf58701748

Observation 051d5200-2466-433f-948e-fa46ba9c4c30 · outbound

This paper cites Unsupervised medical image translation with adversarial diffusion models.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Unsupervised medical image translation with adversarial diffusion models

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.736423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:06.953711Z digest=sha256:55abce6afacd2d5d3436891f59bed7db4c40c440171f115babd0ea531ffa4c39

Observation 5a6d5801-2d1a-43c9-bdba-d88c5914c4df · outbound

This paper cites On variational bounds of mutual information.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting On variational bounds of mutual information

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.720929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:07.058611Z digest=sha256:7217c1b42b8397af27272143408d680d42d8d473ed7f01d4337b8bda05863593

Observation 04d8ab2e-3c09-4969-8f54-f0c86ead5096 · outbound

This paper cites Learning transferable visual models from natural language supervision.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Learning transferable visual models from natural language supervision

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:07.172235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:07.172235Z digest=sha256:29c2d94c97d173b4a8301f6cdfb866d7bfeeecd8aafeeef76093b20acf070ea3

Observation 904762dd-f001-4ef0-96f7-33945a9c2c06 · outbound

This paper cites Study of thoracic ct in covid-19: the stoic project.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Study of thoracic ct in covid-19: the stoic project

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.695888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:07.275215Z digest=sha256:2ff2a470d75c8c4c044341c2b822e08c2fe40152b5c1ff6ed4c23ddede644749

Observation 7526b28c-0615-4fd0-8f25-a789a2a3a0d4 · outbound

This paper cites Deep learning in medical image analysis.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Deep learning in medical image analysis

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.680527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:07.401048Z digest=sha256:7136eed76bffcf2598af609a0851d7e8cb268535084d01df87f7761c4b4e86ef

Observation 8cea4486-36cd-46ca-a3e0-800c3835fb65 · outbound

This paper cites Large-scale and fine-grained vision- language pre-training for enhanced ct image understanding.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Large-scale and fine-grained vision- language pre-training for enhanced ct image understanding

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.666394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:07.524178Z digest=sha256:ab6047561ee9b7496b7c12780c82f948e672a30c3ae0423f3a3dc7f22b78a93b

Observation d6fa8c35-2d72-432f-8f29-319783f54832 · outbound

This paper cites Bioclip: A vision foundation model for the tree of life.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Bioclip: A vision foundation model for the tree of life

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.651773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:07.697454Z digest=sha256:4fe744e759db91feacaafaf1a3d17fcfe83650fc5317109e0bd4bd3293e63758

Observation 196f8bc6-9cd3-4d50-ba35-9d041cf1c193 · outbound

This paper cites XrayGPT: Chest Radiographs Summarization using Medical Vision-Language Models.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting XrayGPT: Chest Radiographs Summarization using Medical Vision-Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:07.826433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:07.826433Z digest=sha256:363867edf2f7b8cb70291a5fc4eb7b75c82aa46003f868df7d583407e76d0096

Observation fe0aa725-db38-4b9a-9523-c2a648ecd696 · outbound

This paper cites Communication errors in radiology–pitfalls and how to avoid them.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Communication errors in radiology–pitfalls and how to avoid them

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.637015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:07.970509Z digest=sha256:e96298abd97b85680017850453c94be49d2b7fd87c7943ecd0c0ea1e368bdcdd

Observation 75a176bc-5a08-4d29-9b3e-e14cd42ca5df · outbound

This paper cites Multi- granularity cross-modal alignment for generalized medical visual representation learning.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Multi- granularity cross-modal alignment for generalized medical visual representation learning

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.621985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:08.098509Z digest=sha256:bc7146ac891b64669d0dedc2d526bab52fb99e20c3a970e97b3bdfc88df228b3

Observation 01ab2a9b-31b2-4afe-9136-fa8d0a2ff7f7 · outbound

This paper cites Totalseg- mentator: robust segmentation of 104 anatomic structures in ct images.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Totalseg- mentator: robust segmentation of 104 anatomic structures in ct images

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.608638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:08.279430Z digest=sha256:c6207350491d3515f6459e743c775a6cb3e23ccdc4db7f6bf8e6802fb7451253

Observation 06f050c3-69ee-4cae-a1f7-a1026382da7c · outbound

This paper cites Medklip: Medical knowledge enhanced language-image pre-training for x-ray diagnosis.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Medklip: Medical knowledge enhanced language-image pre-training for x-ray diagnosis

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.591474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:08.403517Z digest=sha256:1eaa8b8e3dc9e6828e132e3b4ba1c56e3438fb2fe6ee9354f7f2531f25cd9615

Observation 99fd17d9-2eb1-4594-bf2f-41f070a5863a · outbound

This paper cites Unimiss: Universal medical self-supervised learning via breaking dimensionality barrier.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Unimiss: Universal medical self-supervised learning via breaking dimensionality barrier

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.577184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:08.565766Z digest=sha256:198f654c7f70fa65c983ca5ee41c839b6bc48e8e88dcb367c6676bfcccb4046a

Observation eb497786-556b-4405-a3d4-fa736c636040 · outbound

This paper cites Demystifying CLIP Data.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Demystifying CLIP Data

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:08.732381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:08.732381Z digest=sha256:e9ec90e59c89748c3e612c44d1c868058c87fd29c9459bf917a2f25ab1e4e37e

Observation 57eb387b-0cc9-4b4b-947e-0e076d79af67 · outbound

This paper cites Glipv2: unifying local- ization and vl understanding.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Glipv2: unifying local- ization and vl understanding

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.561088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:08.830792Z digest=sha256:3f62630b3521a7ea9edf41fc22afa7ca325d4d806064137e3df6cfc39c01962f

Observation 0871f1a1-4930-4509-b976-2ae4035b12ae · outbound

This paper cites Biomedgpt: A unified and generalist biomedical generative pre-trained transformer for vision, language, and multimodal tasks.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Biomedgpt: A unified and generalist biomedical generative pre-trained transformer for vision, language, and multimodal tasks

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.546935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:08.923378Z digest=sha256:125c9ed718b6a093e516ec9cf69670443080074c5859afacdc3da288b44cbf78

Observation fd585d6f-5915-4473-b7d0-29ad8ac20d8c · outbound

This paper cites RadGenome-Chest CT: A Grounded Vision-Language Dataset for Chest CT Analysis.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting RadGenome-Chest CT: A Grounded Vision-Language Dataset for Chest CT Analysis

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:09.008709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:09.008709Z digest=sha256:0bd282a2bb020d8267163d3ddafa8138dd24f084d49d820738a762f396b17712

Observation 5278f4c0-7029-4709-814e-c6f077d70839 · outbound

This paper cites Development of a large-scale medical visual question-answering dataset.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Development of a large-scale medical visual question-answering dataset

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.532120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:09.069233Z digest=sha256:8894b66a95133cad654b64de822a5241a4850e1a329cc110b6d7e8746c124789

Observation 52da4f73-68ac-4a46-ae3b-f2ce4f16ed22 · outbound

This paper cites Each of these claims is supported by theoretical analysis, ablation studies, and experimental results.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Each of these claims is supported by theoretical analysis, ablation studies, and experimental results

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.447833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:09.130493Z digest=sha256:621c776905fff7e7a693b3f3331b7f3921f2ca73ffc83c33c98ce21ad47b44db

Observation 2e310e4c-9ace-41a2-81bd-58b7deb97277 · outbound

This paper cites Limitations.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Limitations

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.216943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:09.166745Z digest=sha256:b7f2ae61de566423758c0dd9f89f41b498fe927400899aea875aa43496f03f4b

Observation 655d84c8-671a-4e46-8840-a2ce1a9282b3 · outbound

This paper cites Guidelines: • The answer NA means that the paper does not include theoretical results.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Guidelines: • The answer NA means that the paper does not include theoretical results

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:12.864349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:09.222371Z digest=sha256:7c6f55fb7f9d3cf10aaec22f16d5b213476723744be9c29ae5d191a261dfcedf

Observation 4babdb4f-17c1-4e46-af10-1ab8db572560 · outbound

This paper cites Guidelines: • The answer NA means that the paper does not include experiments.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Guidelines: • The answer NA means that the paper does not include experiments

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:12.564751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:09.304182Z digest=sha256:bf86177a5d60e1c5499b5386197f3e12a772e17ec16586c6b130bb7ea1c32ab8

Observation 42aa9b52-0cf8-4efa-b0ce-3849436a662a · outbound

This paper cites Guidelines: • The answer NA means that paper does not include experiments requiring code.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Guidelines: • The answer NA means that paper does not include experiments requiring code

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:12.259108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:09.363957Z digest=sha256:d558a2bb8ab2d81aef79f6af0ab37adefc84d4f9c0218525b7e7ccb038b5f639

Observation 81e41ff1-e8e3-43ed-9355-b8b9868615ca · outbound

This paper cites Guidelines: • The answer NA means that the paper does not include experiments.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Guidelines: • The answer NA means that the paper does not include experiments

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:12.017340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:09.437212Z digest=sha256:f097e3fd9a38a4b7be8cd24a09b7ed28ba851f74645fda1e65753b568a33b2eb

Observation e1fad113-1c30-4f52-88ea-f48e99eed218 · outbound

This paper cites Guidelines: • The answer NA means that the paper does not include experiments.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Guidelines: • The answer NA means that the paper does not include experiments

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:11.928213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:09.500264Z digest=sha256:076a92935778d9a894de97a4be6c4d0f49505bf845f07096a62bd71ef1b8d1e5

Observation 60d98338-22c4-40ca-957a-87b29c194562 · outbound

This paper cites Guidelines: • The answer NA means that the paper does not include experiments.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Guidelines: • The answer NA means that the paper does not include experiments

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:11.810015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:09.569808Z digest=sha256:673ba4d6f81919a0ea5f4dc82a310d2b0bdf70d6927d421c01cadfeeb5fc753e

Observation fe82fdc4-fa2e-4d71-bcbe-0c308e46cdd5 · outbound

This paper cites All data used in this study are from publicly available, de- identified medical datasets, and no personally identifiable information (PII) was accessed or used.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting All data used in this study are from publicly available, de- identified medical datasets, and no personally identifiable information (PII) was accessed or used

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:11.712125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:09.648785Z digest=sha256:52b7b5361ebea1f87db411d4b5eb13d3650368bd5bb84684b0ec28384dbeab5b

Observation 4cbe450f-4293-472e-80de-8a150b92989b · outbound

This paper cites Guidelines: 17 • The answer NA means that there is no societal impact of the work performed.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Guidelines: 17 • The answer NA means that there is no societal impact of the work performed

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:11.553592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:09.694965Z digest=sha256:93194daebc44e8561902990a34c386bcdeca3398a1425564f0e6c1f59ab91c61

Observation 91091c49-eb63-4fbe-b291-6edfceb106fd · outbound

This paper cites an unresolved cited work.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-05T10:46:11.419872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:09.756872Z digest=sha256:028983d38b15eabfe7677747c8a7043acdcb9b6ad03e0ebbf68424fc59945f1e

Observation 68c2e891-feb6-46ed-b4b6-2720a5c2d289 · outbound

This paper cites Guidelines: • The answer NA means that the paper does not use existing assets.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Guidelines: • The answer NA means that the paper does not use existing assets

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:11.269223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:09.813237Z digest=sha256:0bc45d87d53b0d98879d438031562f4db6c4ddb07efab8a1675d1cedb762e429

Observation 065dba93-e8cc-466b-ba0b-2abbd75001ec · outbound

This paper cites These assets will be released with accompanying documentation upon paper acceptance.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting These assets will be released with accompanying documentation upon paper acceptance

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:11.105435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:09.853132Z digest=sha256:7844e9785e6ef11b04e9c85148cc7c12876e9e4e8627838770833a0c92f48cfb

Observation 5cf22d1e-5be8-4e8f-9c0d-983ca3baaf39 · outbound

This paper cites All data used are from publicly available, de-identified medical datasets with appropriate licenses and do not involve any direct interaction with individuals.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting All data used are from publicly available, de-identified medical datasets with appropriate licenses and do not involve any direct interaction with individuals

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:10.989629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:09.939998Z digest=sha256:bd121310701c8a449d75b2d765fd0f58eb6855f02940bbb7a49cfd24c4366c62

Observation 304aea77-f964-4354-8c2a-855d2a6a01d3 · outbound

This paper cites Therefore, IRB approval was not required.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Therefore, IRB approval was not required

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:10.820337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:09.978316Z digest=sha256:1884952896a93832336b7094b79b7dc4c7ebf27ff0e1eeb5bb04cf5f20ceceeb

Observation 2905636d-5717-4d58-9d1d-4f6aeb97aac9 · outbound

This paper cites Answer: [Yes] Justification: Large language models such as GPT-4o and Qwen2.5, were used to rewrite radiology reports for improving semantic clarity during pretraining.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Answer: [Yes] Justification: Large language models such as GPT-4o and Qwen2.5, were used to rewrite radiology reports for improving semantic clarity during pretraining

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:10.677975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:46:09.996574Z digest=sha256:5f46fb93367b18d0aaf4121376cf078c9ee9a4f6ab18fbb627843a5b21a3147d

Pith citing papers

Observation 0cd03453-c97c-4b33-9b7c-1d3daeea6f74 · inbound

Self-Supervised Dynamical System Representations for Physiological Time-Series cites this paper.

Self-Supervised Dynamical System Representations for Physiological Time-Series MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T19:34:04.544567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T19:34:04.544567Z digest=sha256:b9e678fc66757ae5d71926e25a394ba0236af21004801cda7a6c33f0ce079ae9

Observation e8ecbf8a-69eb-48e9-9d43-16ea5263acd2 · inbound

ASAP: Advancing Medical Volumetric Representation Learning with Anatomy-aware Semantically-adaptive Pre-training cites this paper.

ASAP: Advancing Medical Volumetric Representation Learning with Anatomy-aware Semantically-adaptive Pre-training MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting

Reference 81

Resolution
verified exact
arxiv_id, observed 2026-06-28T19:22:34.659378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T19:16:42.139096Z digest=sha256:7de417b13444d2a1810da87e18e48b1a92b9b4b68dd0d4444f32c9b3fd51d905