Pith. sign in

Paper Citation Record · LEDGER

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting

As of 8 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 2 inbound Pith citation observations for arXiv:2509.03800.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.03800 v1

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T10:46:09.996574Z

measured 62 of 62 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T19:34:04.544567Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-28T19:22:34.657932Z

Reference resolution

60 of 60 outbound references displayed

  • verified exact1
  • verified fuzzy39
  • unresolved20
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a6fa80d1-f5d6-4f4a-8c11-f7590af040bf · outbound

This paper cites Merlin: A vision language foundation model for 3d computed tomography.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Merlin: A vision language foundation model for 3d computed tomography

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:04.243549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:04.243549Z digest=sha256:21ae1f15e2f0c38540ed8744ce1289c43a22e12403523ce2c15ca61ca3b6864a

Observation b33a7ab9-9f9f-45d3-8445-97b33ff2b16b · outbound

This paper cites A vision–language foundation model for the generation of realistic chest x-ray images.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting A vision–language foundation model for the generation of realistic chest x-ray images

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.958572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:04.330865Z digest=sha256:a774f5d771097fa47fbe6a877859617fcf196f33d095f6e915e68a1ac6559de5

Observation cb37fff5-0668-4885-9707-e0ec24f3015e · outbound

This paper cites Making the most of text semantics to improve biomedical vision–language processing.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Making the most of text semantics to improve biomedical vision–language processing

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:04.426003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:04.426003Z digest=sha256:4cb43b00259766741a4c3468656ba4b89b163a1d29c7c7c730352b54199ea467

Observation d404663f-9670-4008-8297-2b334af143b6 · outbound

This paper cites Understanding and confronting our mis- takes: the epidemiology of error in radiology and strategies for error reduction.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Understanding and confronting our mis- takes: the epidemiology of error in radiology and strategies for error reduction

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.932829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:04.516487Z digest=sha256:545e61e43fb9c2be09573728ecae5a73a8c01dca10999e143cd94f0f2c0ebe05

Observation c143d303-e11e-4e10-bd16-89cf43f4e210 · outbound

This paper cites Joint modeling of chest radiographs and radiology reports for pulmonary edema assessment.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Joint modeling of chest radiographs and radiology reports for pulmonary edema assessment

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.917747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:04.585735Z digest=sha256:7f38f180448502c08bbed94e2cd7ab45f5334ee48dd2e3bba039fd1cbed326bc

Observation 41ed5110-7327-4e3d-bafa-1a1622ef2ee5 · outbound

This paper cites Contrastive Localized Language-Image Pre-Training.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Contrastive Localized Language-Image Pre-Training

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:04.661105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:04.661105Z digest=sha256:08c110710ebc97c456a310ba35374b07e754c16e016b9d942e2a21d05cfd8a93

Observation 3f111173-6cf8-4d07-8f39-b1d2c4ddf619 · outbound

This paper cites A review of medical image data augmentation techniques for deep learning applications.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting A review of medical image data augmentation techniques for deep learning applications

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.900852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:04.776379Z digest=sha256:d37f68426a3d8c970a9fbbe62d417b1bb168c4613a9c06d861577e6ebcf12c74

Observation f555adc7-97f4-414c-8a7b-5d9ba7d8b2d8 · outbound

This paper cites Machine-learning-based multiple abnormality prediction with large-scale chest computed tomography volumes.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Machine-learning-based multiple abnormality prediction with large-scale chest computed tomography volumes

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.884634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:04.842478Z digest=sha256:37fbfc240ffe76a4d05db5cac190343ad2cba74741ade2ed71310073ae27a498

Observation 85df312c-28df-4027-8a32-0cf6486e38a7 · outbound

This paper cites Multi-View and Multi-Scale Alignment for Contrastive Language-Image Pre-training in Mammography.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Multi-View and Multi-Scale Alignment for Contrastive Language-Image Pre-training in Mammography

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-05T10:46:10.536952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:04.962958Z digest=sha256:3c5f47383a6dd7ea703ac90bfb765ac2b6ea3d8f263a4d6aea57d5ee80e1cf9d

Observation 4feab371-5592-4901-b37e-221704c40765 · outbound

This paper cites The Llama 3 Herd of Models.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting The Llama 3 Herd of Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:05.092590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:05.092590Z digest=sha256:0414699c797a2c6180929346bc163706793dc4f8bf2f73b667b836add7962230

Observation 13ac3f6d-a019-4f0b-825c-d36d7414fa43 · outbound

This paper cites Devel- oping generalist foundation models from a multimodal dataset for 3d computed tomography.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Devel- oping generalist foundation models from a multimodal dataset for 3d computed tomography

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:05.191575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:05.191575Z digest=sha256:b9767db56f9f3a892c480430898806c2a7574d7043f658ed0bbaadf14c2a0ca6

Observation ac57e9c6-b566-4905-a756-dd5818e81d92 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting LoRA: Low-Rank Adaptation of Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:05.296884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:05.296884Z digest=sha256:8822086c9253f308733d01802123ac40411fe31e294d7b7b1a55968afc0375c1

Observation aa273c70-1ad5-4176-a12e-25ee2a518a90 · outbound

This paper cites Gloria: A multimodal global-local representation learning framework for label-efficient medical image recognition.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Gloria: A multimodal global-local representation learning framework for label-efficient medical image recognition

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:05.360284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:05.360284Z digest=sha256:ad7409bf347296222b36de4c91818401ead93e71f11018a6ca5de2e46b31c4f1

Observation c7bf766a-ec65-41da-a482-156f75f1b8da · outbound

This paper cites Enhancing representation in medical vision-language foun- dation models via multi-scale information extraction techniques.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Enhancing representation in medical vision-language foun- dation models via multi-scale information extraction techniques

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.859199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:05.458401Z digest=sha256:71d73f6ebfe0bc2118f3cb9e7c8987d9011e9e79dfa39ba8af4fcf40de0bd9d6

Observation da68f205-0616-4a33-8ce2-23eb68858145 · outbound

This paper cites STU-Net: Scalable and Transferable Medical Image Segmentation Models Empowered by Large-Scale Supervised Pre-training.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting STU-Net: Scalable and Transferable Medical Image Segmentation Models Empowered by Large-Scale Supervised Pre-training

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:05.562057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:05.562057Z digest=sha256:5f76afaf2b8ad6a0c65ffc05c95de6a48e2c504da25619eb5ec20ed44554ebd3

Observation 4dbc9100-63b5-4b79-b258-fb0395cb7c62 · outbound

This paper cites nnu-net: a self-configuring method for deep learning-based biomedical image segmentation.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting nnu-net: a self-configuring method for deep learning-based biomedical image segmentation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:05.633725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:05.633725Z digest=sha256:83a816066c2d72b22ae4a12ed7d3f07c8a0ac5442d68d4141ed21771e40549d8

Observation 0d20a01c-58e5-4de6-8e64-28d7416906b8 · outbound

This paper cites Fool me twice: delayed diagnoses in radiology with emphasis on perpetuated errors.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Fool me twice: delayed diagnoses in radiology with emphasis on perpetuated errors

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.834498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:05.740237Z digest=sha256:23e8bb839b43a1e9408c47f81e3866244458815d8770ef4f4fea10284786d3ad

Observation 0ef9eb6e-3c40-4f58-8234-d3ac1161d8f4 · outbound

This paper cites Generating synthetic data for medical imaging.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Generating synthetic data for medical imaging

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.819212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:05.837515Z digest=sha256:5a83380b30b3503ed78d8ba4db7dd0a6381340917380f66344d05c80db50f6be

Observation 35d0a13e-01ef-4672-bf63-cc177925cee6 · outbound

This paper cites Cxr-llava: a multimodal large language model for interpreting chest x-ray images.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Cxr-llava: a multimodal large language model for interpreting chest x-ray images

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.803136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:05.952126Z digest=sha256:751785c41e8be6dadde9a4b9f9f8dff6487566c9d8da4da75e7a5ede74660159

Observation 71cb91cd-edae-46b0-9c62-195ae4964bf7 · outbound

This paper cites Llava-med: Training a large language-and-vision assistant for biomedicine in one day.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Llava-med: Training a large language-and-vision assistant for biomedicine in one day

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:06.026153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:06.026153Z digest=sha256:1dac488006132977707cfbc3454315a06dffc4fea2b9370eb222c317558f6488

Observation 6b547427-1974-4c16-a177-2ff59895bdc3 · outbound

This paper cites Artificial general intelligence for medical imaging analysis.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Artificial general intelligence for medical imaging analysis

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.775179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:06.144188Z digest=sha256:eea40403d24b5ca547ed10998812a8120084a4e139c2853a8a7c1cb21e709541

Observation 9920dd3d-25af-4526-9efc-147946bfd878 · outbound

This paper cites Ct-glip: 3d grounded language-image pretraining with ct scans and radiology reports for full-body scenarios.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Ct-glip: 3d grounded language-image pretraining with ct scans and radiology reports for full-body scenarios

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:06.214147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:06.214147Z digest=sha256:17fece55f3b6e2c29ce2d0fb26d4dc82e1e883200b4892fb133d91b99bbee0ff

Observation 045e8575-14d9-458f-90dc-c8791db2b0af · outbound

This paper cites Pmc-clip: Contrastive language-image pre-training using biomedical documents.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Pmc-clip: Contrastive language-image pre-training using biomedical documents

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:06.347751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:06.347751Z digest=sha256:1e90e757b565836f4d23b31e37cd9d293883201b06188d3c2fdc2108b98b77f4

Observation 56045d27-ec37-4ba5-a6bd-9f2a0988974a · outbound

This paper cites Can Medical Vision-Language Pre-training Succeed with Purely Synthetic Data?.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Can Medical Vision-Language Pre-training Succeed with Purely Synthetic Data?

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:06.490159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:06.490159Z digest=sha256:cf846ea47fb77b5811bd009d65d21e87c462bd29c96b5e1ca26d86ad1dbf2264

Observation 7cfb17f5-68a2-4784-ae44-4289145e1d0d · outbound

This paper cites Improved baselines with visual instruction tuning.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Improved baselines with visual instruction tuning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:06.654963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:06.654963Z digest=sha256:8b821fae9e14f2910e8289f4dd8bc5e1dd88a4c9f53eacf1cdcc0007d4aaede8

Observation 08c8a7a7-1b04-4673-b65c-8d991d660f66 · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Representation Learning with Contrastive Predictive Coding

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:06.818244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:06.818244Z digest=sha256:f8d3b106d928d5ace684aafc9671dfaf6610f65387addb4552009ac427ad4b2a

Observation 051d5200-2466-433f-948e-fa46ba9c4c30 · outbound

This paper cites Unsupervised medical image translation with adversarial diffusion models.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Unsupervised medical image translation with adversarial diffusion models

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.736423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:06.953711Z digest=sha256:e67494441d34bf949b762e9e51de7d97108900a1c45ffd8db1fed089bf502f7c

Observation 5a6d5801-2d1a-43c9-bdba-d88c5914c4df · outbound

This paper cites On variational bounds of mutual information.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting On variational bounds of mutual information

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.720929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:07.058611Z digest=sha256:3a3ca80c226525586a314f1a4c71d295027ed48c12dadd0b82af92474fc8f95c

Observation 04d8ab2e-3c09-4969-8f54-f0c86ead5096 · outbound

This paper cites Learning transferable visual models from natural language supervision.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Learning transferable visual models from natural language supervision

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:07.172235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:07.172235Z digest=sha256:2ffe0054aac911b223cee7fa70bd030bca16c8369c3e0d116944ce7094c2867e

Observation 904762dd-f001-4ef0-96f7-33945a9c2c06 · outbound

This paper cites Study of thoracic ct in covid-19: the stoic project.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Study of thoracic ct in covid-19: the stoic project

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.695888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:07.275215Z digest=sha256:33c95751b907676c2eab62a7160a420bc95d2d21d088ee9f61aa1ae551e22162

Observation 7526b28c-0615-4fd0-8f25-a789a2a3a0d4 · outbound

This paper cites Deep learning in medical image analysis.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Deep learning in medical image analysis

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.680527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:07.401048Z digest=sha256:1f7176cc770ed6f45e278c50d4149ed35151ea431e8b521e7e1f84d305ba3091

Observation 8cea4486-36cd-46ca-a3e0-800c3835fb65 · outbound

This paper cites Large-scale and fine-grained vision- language pre-training for enhanced ct image understanding.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Large-scale and fine-grained vision- language pre-training for enhanced ct image understanding

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.666394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:07.524178Z digest=sha256:02a1c36c0b8182f0639ef7ffa70365e3d897d9237847f53d0aba791502d4479a

Observation d6fa8c35-2d72-432f-8f29-319783f54832 · outbound

This paper cites Bioclip: A vision foundation model for the tree of life.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Bioclip: A vision foundation model for the tree of life

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.651773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:07.697454Z digest=sha256:5b8012a778ebbe36d8bb87b206dbe1eb3260641c49a3e3e6a5d4ee6a2bd3a2e8

Observation 196f8bc6-9cd3-4d50-ba35-9d041cf1c193 · outbound

This paper cites XrayGPT: Chest Radiographs Summarization using Medical Vision-Language Models.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting XrayGPT: Chest Radiographs Summarization using Medical Vision-Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:07.826433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:07.826433Z digest=sha256:4cfb89bbc085faaeaf69e9f07489f3d90f37138f84b34e6562d72a7a1c28d6d5

Observation fe0aa725-db38-4b9a-9523-c2a648ecd696 · outbound

This paper cites Communication errors in radiology–pitfalls and how to avoid them.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Communication errors in radiology–pitfalls and how to avoid them

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.637015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:07.970509Z digest=sha256:7831d7519fe91ddb6348d5078d69a783b9a2208eeb5bc52349f3d2dc6f90c507

Observation 75a176bc-5a08-4d29-9b3e-e14cd42ca5df · outbound

This paper cites Multi- granularity cross-modal alignment for generalized medical visual representation learning.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Multi- granularity cross-modal alignment for generalized medical visual representation learning

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.621985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:08.098509Z digest=sha256:d22e51e0ad6fa3ff0a95e7d8722f0976c12177d66fd5a9d00549e3d910047f6d

Observation 01ab2a9b-31b2-4afe-9136-fa8d0a2ff7f7 · outbound

This paper cites Totalseg- mentator: robust segmentation of 104 anatomic structures in ct images.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Totalseg- mentator: robust segmentation of 104 anatomic structures in ct images

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.608638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:08.279430Z digest=sha256:d87b6e52e10bd20786f92f68ed2a226d01da41b40f6592cd9bd813c0792bd299

Observation 06f050c3-69ee-4cae-a1f7-a1026382da7c · outbound

This paper cites Medklip: Medical knowledge enhanced language-image pre-training for x-ray diagnosis.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Medklip: Medical knowledge enhanced language-image pre-training for x-ray diagnosis

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.591474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:08.403517Z digest=sha256:b3593605ec268879ab550e4bde796892f46af28ad9405ec6b8a93d17a9777822

Observation 99fd17d9-2eb1-4594-bf2f-41f070a5863a · outbound

This paper cites Unimiss: Universal medical self-supervised learning via breaking dimensionality barrier.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Unimiss: Universal medical self-supervised learning via breaking dimensionality barrier

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.577184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:08.565766Z digest=sha256:b39dcdc8e008cf6f82688996ce156f7f9b649683041914735a89d94fd955d518

Observation eb497786-556b-4405-a3d4-fa736c636040 · outbound

This paper cites Demystifying CLIP Data.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Demystifying CLIP Data

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:08.732381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:08.732381Z digest=sha256:5a3c7db8c96331e2929a42d6839a80173200ac299d734118f4f08cdaf8e0dd05

Observation 57eb387b-0cc9-4b4b-947e-0e076d79af67 · outbound

This paper cites Glipv2: unifying local- ization and vl understanding.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Glipv2: unifying local- ization and vl understanding

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.561088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:08.830792Z digest=sha256:4cd976ae6d53ae13867757a50377d6215c65e80959a9276f950408f593abf2b4

Observation 0871f1a1-4930-4509-b976-2ae4035b12ae · outbound

This paper cites Biomedgpt: A unified and generalist biomedical generative pre-trained transformer for vision, language, and multimodal tasks.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Biomedgpt: A unified and generalist biomedical generative pre-trained transformer for vision, language, and multimodal tasks

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.546935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:08.923378Z digest=sha256:045fb2364fb487d950462a1eacd290f2695ce9fce7c67b3aac9b41ac53962b2e

Observation fd585d6f-5915-4473-b7d0-29ad8ac20d8c · outbound

This paper cites RadGenome-Chest CT: A Grounded Vision-Language Dataset for Chest CT Analysis.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting RadGenome-Chest CT: A Grounded Vision-Language Dataset for Chest CT Analysis

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:09.008709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:09.008709Z digest=sha256:e1b1af3d90eb5b3a2719deba298d3fae9819e1a34ff55e8a50f09ba2be78f9f6

Observation 5278f4c0-7029-4709-814e-c6f077d70839 · outbound

This paper cites Development of a large-scale medical visual question-answering dataset.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Development of a large-scale medical visual question-answering dataset

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.532120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:09.069233Z digest=sha256:19ef4d1914dabc55d714888d612e2c3620589da047c5f2348fe22c14fa229677

Observation 52da4f73-68ac-4a46-ae3b-f2ce4f16ed22 · outbound

This paper cites Each of these claims is supported by theoretical analysis, ablation studies, and experimental results.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Each of these claims is supported by theoretical analysis, ablation studies, and experimental results

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.447833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:09.130493Z digest=sha256:71712c4828080b7d702357396bcf40b6609317bb7ba74e35f8406bd7bbf111f8

Observation 2e310e4c-9ace-41a2-81bd-58b7deb97277 · outbound

This paper cites Limitations.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Limitations

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.216943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:09.166745Z digest=sha256:8228d88c2397121cfc3c4107b59cdc7a2b78463d9e7146e42dfbfb8f0f9bdeb8

Observation 655d84c8-671a-4e46-8840-a2ce1a9282b3 · outbound

This paper cites Guidelines: • The answer NA means that the paper does not include theoretical results.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Guidelines: • The answer NA means that the paper does not include theoretical results

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:12.864349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:09.222371Z digest=sha256:afcd688af7a2fbbbe1a1a3ea41bf2342aa9e0e2ee87b13d69e62b14e760ff5ab

Observation 4babdb4f-17c1-4e46-af10-1ab8db572560 · outbound

This paper cites Guidelines: • The answer NA means that the paper does not include experiments.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Guidelines: • The answer NA means that the paper does not include experiments

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:12.564751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:09.304182Z digest=sha256:81190a72b19ed311901f7c3b82d3f8605a18df5e9f18570c3c339d3013d63552

Observation 42aa9b52-0cf8-4efa-b0ce-3849436a662a · outbound

This paper cites Guidelines: • The answer NA means that paper does not include experiments requiring code.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Guidelines: • The answer NA means that paper does not include experiments requiring code

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:12.259108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:09.363957Z digest=sha256:b36a3062e41f211fb1a3c9b7dca771896b63ede7243c323139aad6156b99a99e

Observation 81e41ff1-e8e3-43ed-9355-b8b9868615ca · outbound

This paper cites Guidelines: • The answer NA means that the paper does not include experiments.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Guidelines: • The answer NA means that the paper does not include experiments

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:12.017340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:09.437212Z digest=sha256:74224a819f7140e818c4be3ca3ef6a63103db08d4d1778751d7c735117fd46fd

Observation e1fad113-1c30-4f52-88ea-f48e99eed218 · outbound

This paper cites Guidelines: • The answer NA means that the paper does not include experiments.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Guidelines: • The answer NA means that the paper does not include experiments

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:11.928213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:09.500264Z digest=sha256:feff470f5a662f60bb17d64a046dff46aca9c59b887afae8ada651fdc3000413

Observation 60d98338-22c4-40ca-957a-87b29c194562 · outbound

This paper cites Guidelines: • The answer NA means that the paper does not include experiments.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Guidelines: • The answer NA means that the paper does not include experiments

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:11.810015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:09.569808Z digest=sha256:1f250f6db5ea0ce2fbdb319abd9973b4413db61ab164053d4134de5db0377af9

Observation fe82fdc4-fa2e-4d71-bcbe-0c308e46cdd5 · outbound

This paper cites All data used in this study are from publicly available, de- identified medical datasets, and no personally identifiable information (PII) was accessed or used.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting All data used in this study are from publicly available, de- identified medical datasets, and no personally identifiable information (PII) was accessed or used

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:11.712125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:09.648785Z digest=sha256:b0859f30dacf2d49cd8086a69c5d64478a44921b642f6f4782710337a2065db6

Observation 4cbe450f-4293-472e-80de-8a150b92989b · outbound

This paper cites Guidelines: 17 • The answer NA means that there is no societal impact of the work performed.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Guidelines: 17 • The answer NA means that there is no societal impact of the work performed

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:11.553592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:09.694965Z digest=sha256:d6a5f34dafac39638d9249c03b5791854f973d07d292ce5779b0ea1dc9f92754

Observation 91091c49-eb63-4fbe-b291-6edfceb106fd · outbound

This paper cites an unresolved cited work.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-05T10:46:11.419872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:09.756872Z digest=sha256:f8022597ffe9b7adf032cfdb96e19b2c85cfd4164c6574cee1ac2923286dc8ed

Observation 68c2e891-feb6-46ed-b4b6-2720a5c2d289 · outbound

This paper cites Guidelines: • The answer NA means that the paper does not use existing assets.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Guidelines: • The answer NA means that the paper does not use existing assets

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:11.269223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:09.813237Z digest=sha256:6afef9df3a9a67cf60ddac5fd5e2e14fb2d198749dc25d518b6af0ed81084aba

Observation 065dba93-e8cc-466b-ba0b-2abbd75001ec · outbound

This paper cites These assets will be released with accompanying documentation upon paper acceptance.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting These assets will be released with accompanying documentation upon paper acceptance

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:11.105435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:09.853132Z digest=sha256:eb1d9fea801da5a508ad421215af5bb72edcc0cd59fa273b15603ac580bcf182

Observation 5cf22d1e-5be8-4e8f-9c0d-983ca3baaf39 · outbound

This paper cites All data used are from publicly available, de-identified medical datasets with appropriate licenses and do not involve any direct interaction with individuals.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting All data used are from publicly available, de-identified medical datasets with appropriate licenses and do not involve any direct interaction with individuals

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:10.989629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:09.939998Z digest=sha256:d1bccc02386aa24fd89f377d232275fe7df8857e25626b05dc3acd86f7365977

Observation 304aea77-f964-4354-8c2a-855d2a6a01d3 · outbound

This paper cites Therefore, IRB approval was not required.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Therefore, IRB approval was not required

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:10.820337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:09.978316Z digest=sha256:aa60df6ef927481e760f7701ed4dae17f71c96ac13154ecb03a2cd78bcf8e7ed

Observation 2905636d-5717-4d58-9d1d-4f6aeb97aac9 · outbound

This paper cites Answer: [Yes] Justification: Large language models such as GPT-4o and Qwen2.5, were used to rewrite radiology reports for improving semantic clarity during pretraining.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Answer: [Yes] Justification: Large language models such as GPT-4o and Qwen2.5, were used to rewrite radiology reports for improving semantic clarity during pretraining

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:10.677975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T10:46:09.996574Z digest=sha256:06fa50161b0ce3cac649e1ef8185b813f53a0c8b947a9d4212902b07b5519b9e

Pith citing papers

Observation 0cd03453-c97c-4b33-9b7c-1d3daeea6f74 · inbound

Self-Supervised Dynamical System Representations for Physiological Time-Series cites this paper.

Self-Supervised Dynamical System Representations for Physiological Time-Series MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T19:34:04.544567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T19:34:04.544567Z digest=sha256:ddc85d084eb2e932feee7dd63d93bb500c12c92bd60b405ae8ed5a0040b6294e

Observation e8ecbf8a-69eb-48e9-9d43-16ea5263acd2 · inbound

ASAP: Advancing Medical Volumetric Representation Learning with Anatomy-aware Semantically-adaptive Pre-training cites this paper.

ASAP: Advancing Medical Volumetric Representation Learning with Anatomy-aware Semantically-adaptive Pre-training MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting

Reference 81

Resolution
verified exact
arxiv_id, observed 2026-06-28T19:22:34.659378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T19:16:42.139096Z digest=sha256:ba8f97284fc278e1517285e6c49cf869c09165b9bdbb1bc2b2b63b428c4e2705