Pith. sign in

Paper Citation Record · LEDGER

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding

As of 15 August 2026, this Paper Citation Record lists 100 of 127 outbound references and 0 inbound Pith citation observations for arXiv:2607.24743.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.24743 v2

Coverage vector

measured 100 of 127 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-31T06:20:30.115352Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 127 outbound references displayed

  • verified exact3
  • verified fuzzy0
  • unresolved97
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 688ad64a-4194-4329-a302-5516cadadec9 · outbound

This paper cites Qwen2.5-VL Technical Report.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Qwen2.5-VL Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.650039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.650039Z digest=sha256:2c0f59599e92e4b9ce217115bb57c482f8bd53e7ae2f891782c78f10ff611ece

Observation 349f8bd2-2ded-401a-8486-9512b3deb042 · outbound

This paper cites Qwen3-VL Technical Report.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Qwen3-VL Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.655604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.655604Z digest=sha256:d0cd2228984c8767331098abd1c7de7be02252be464f46a1cc24dd8101695982

Observation 3dd06eb9-99b0-40f8-828d-c3f00778136c · outbound

This paper cites Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.660200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.660200Z digest=sha256:ce6cf22544c69c7c20f6e78e54a8ce6deb62d015c57e5ff70f05dd12512e8e0b

Observation a96911c1-0072-49d4-864d-26b1f8ab2d15 · outbound

This paper cites GPT-4o System Card.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding GPT-4o System Card

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.664434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.664434Z digest=sha256:a8078ae48ffd5b004c26e343061da8d4aa59e63a3cf4e4c620656ec88b00b617

Observation 63f3b688-a626-4a43-85fb-2654e1785562 · outbound

This paper cites Improved baselines with visual instruction tuning.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Improved baselines with visual instruction tuning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.668884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.668884Z digest=sha256:b1cc7ac714b650d95829bbe7b145322842e32196769e2de6529004e76d195b14

Observation aba874c2-a026-4fd2-b4a4-731a5210f7b0 · outbound

This paper cites Lingshu: A Generalist Foundation Model for Unified Multimodal Medical Understanding and Reasoning.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Lingshu: A Generalist Foundation Model for Unified Multimodal Medical Understanding and Reasoning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.673214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.673214Z digest=sha256:e41b16cee69b82a6c96ef13d8be7509dd163744c6f39a8d5ebb46b75b89f3551

Observation dd47a75d-3903-441a-8942-a4f2eb05c14d · outbound

This paper cites Hulu-med: A transparent generalist model towards holistic medical vision-language understanding.arXiv preprint arXiv:2510.08668, 2025.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Hulu-med: A transparent generalist model towards holistic medical vision-language understanding.arXiv preprint arXiv:2510.08668, 2025

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.678846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.678846Z digest=sha256:739f16007f457d7101105b06f703ba41033ea7ccc2e5883cfadb6b2518e2d5a0

Observation 5b734fb2-dc20-4584-8703-961c506ed543 · outbound

This paper cites MedGemma Technical Report.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding MedGemma Technical Report

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.683476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.683476Z digest=sha256:91112d14c98c09038fa131db87cf24f27402240f5a3946ca0dad570ead0d3865

Observation 8755a2ba-2e51-42bd-80de-a43a8b95054f · outbound

This paper cites Medvlm-r1: Incentivizing medical reasoning capability of vision-language models (vlms) via reinforcement learning.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Medvlm-r1: Incentivizing medical reasoning capability of vision-language models (vlms) via reinforcement learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.693151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.693151Z digest=sha256:fad1cccbfc971dc3969a2dc01684b6b9a18d9056c6cbd16d59447cf0a07cf2ea

Observation 0b3a0f5f-d5e8-46a1-a864-7d452677729f · outbound

This paper cites Med- r1: Reinforcement learning for generalizable medical reasoning in vision-language models.arXiv preprint arXiv:2503.13939, 2025.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Med- r1: Reinforcement learning for generalizable medical reasoning in vision-language models.arXiv preprint arXiv:2503.13939, 2025

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.697647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.697647Z digest=sha256:c6ede62b7b3fac0989fe21bb26ea485ed46435060d86f30d8e47592b6846809c

Observation 2c12473b-74ac-4ee9-8de7-42ac4d543ef4 · outbound

This paper cites HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.702013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.702013Z digest=sha256:6eb317179f48969b1e9377c027ff96263a549bc57c377a53bf8199093224a76e

Observation be92fe12-4dd8-4e57-8f9c-bfcd91512494 · outbound

This paper cites A generalist vision–language foundation model for diverse biomedical tasks.Nature Medicine, 30(11):3129–3141, 2024.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding A generalist vision–language foundation model for diverse biomedical tasks.Nature Medicine, 30(11):3129–3141, 2024

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.706803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.706803Z digest=sha256:73695df996b9dd87ecac53c9fcfa031cc53543924d597e591c016c2cf85767b5

Observation c19e5cb2-31e4-4fe3-a8a7-20a7cea10345 · outbound

This paper cites Omn- imedvqa: A new large-scale comprehensive evaluation benchmark for medical lvlm.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Omn- imedvqa: A new large-scale comprehensive evaluation benchmark for medical lvlm

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.711046Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.711046Z digest=sha256:96665f520be914506c0e65cabd2e5eee480af91cb323ae5ac819f63aac4023fc

Observation 18b50e56-e3fe-4c59-a3b3-b74111262837 · outbound

This paper cites Slake: A semantically-labeled knowledge- enhanced dataset for medical visual question answering.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Slake: A semantically-labeled knowledge- enhanced dataset for medical visual question answering

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.715200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.715200Z digest=sha256:aa8bdf88dbcd27183df925567436de1ce0c1bb869dffda1c523bd53af7f84ccf

Observation 4ad5dd33-9273-4382-9b97-c1255719f2d4 · outbound

This paper cites Generalist foundation models from a multimodal dataset for 3d computed tomography.Nature Biomedical Engineering, pages 1–19, 2026.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Generalist foundation models from a multimodal dataset for 3d computed tomography.Nature Biomedical Engineering, pages 1–19, 2026

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.719488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.719488Z digest=sha256:8660278f9bc77ff02c27c56d438bee1b20e420eb2472664c912516cf8e96b3e1

Observation 210aef49-4e8d-4ae8-a311-b98061527cc1 · outbound

This paper cites Measuring massive multitask language understanding.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Measuring massive multitask language understanding

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.724241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.724241Z digest=sha256:a51be84d1c08e7a8fab02bb19694e91f25a40ef44f5e95b7695c57684e7cede8

Observation d74a9669-f3fd-419a-b0e2-a8e961260548 · outbound

This paper cites Medxpertqa: Benchmarking expert-level medical reasoning and understanding.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Medxpertqa: Benchmarking expert-level medical reasoning and understanding

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.733500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.733500Z digest=sha256:5575dc312f01646555ee18a13a107f488ab9ff36230fedec7128a4b5bba4db9f

Observation 1623dc72-a475-45b3-9c22-5b5d251b6b30 · outbound

This paper cites Pubmedqa: A dataset for biomedical research question answering.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Pubmedqa: A dataset for biomedical research question answering

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.737585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.737585Z digest=sha256:595646da94f590de41d17460c067da37393073993deabc5d7fc88d73e9fbd5f3

Observation 4b315706-af7f-44c2-a99c-abd0e690293e · outbound

This paper cites Towards a holistic framework for multimodal LLM in 3D brain CT radiology report generation.Nature Communications, 16(1):2258, 2025.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Towards a holistic framework for multimodal LLM in 3D brain CT radiology report generation.Nature Communications, 16(1):2258, 2025

Reference 20

Resolution
verified exact
doi, observed 2026-07-31T06:20:56.545754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-31T06:20:29.741466Z digest=sha256:492cb0dc896c0c5014a829541934d5ff725f134a66b82bbe862d623eaf27dfd2

Observation a1ce074b-85e6-4ec6-8a9f-5a17c378d055 · outbound

This paper cites Med3dvlm: An efficient vision-language model for 3d medical image analysis.IEEE Journal of Biomedical and Health Informatics, 2025.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Med3dvlm: An efficient vision-language model for 3d medical image analysis.IEEE Journal of Biomedical and Health Informatics, 2025

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.745685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.745685Z digest=sha256:4908ee1ceeb3ce45856f3abc3d0f8e08a88a09cef0a7c23aeed6b95dfb9086c0

Observation 3a62ce93-2cff-423e-9fd2-51b951816f66 · outbound

This paper cites Towards generalist foundation model for radiology by leveraging web-scale 2d&3d medical data.Nature Communications, 16(1):7866, 2025.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Towards generalist foundation model for radiology by leveraging web-scale 2d&3d medical data.Nature Communications, 16(1):7866, 2025

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.749856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.749856Z digest=sha256:d24a9df37dcc2097316002ee94f806842806f02d457e5f099abdd8ba5b7d4ba5

Observation dfc17799-a2e0-4108-9c70-1c3f349ac05c · outbound

This paper cites Chaunzwa, Simon Bernatz, Ahmed Hosny, Raymond H.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Chaunzwa, Simon Bernatz, Ahmed Hosny, Raymond H

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.754042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.754042Z digest=sha256:26164073acc2db8ee6aa3b3fb9959fbfc96d8f6cde2afae704fe0086037141f9

Observation 48887214-bd09-4b60-990a-71e6decaaa8a · outbound

This paper cites Eagle: Exploring the design space for multimodal llms with mixture of encoders.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Eagle: Exploring the design space for multimodal llms with mixture of encoders

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.758357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.758357Z digest=sha256:bcde133a43077d1bac6be9835e7ca9141910845db074e69e591027c74af3211f

Observation d2ce34c7-52cc-43da-b23b-f4584169ce5e · outbound

This paper cites Eagle-2: Faster inference of language models with dynamic draft trees.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Eagle-2: Faster inference of language models with dynamic draft trees

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.763082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.763082Z digest=sha256:5bd998e1dce47546e2bbc6a18f52b4a60149eca8792c6e732a70c0bdf787bf38

Observation 7d5e7133-0421-4f48-8fdf-57c255b3cadc · outbound

This paper cites Cambrian-1: A fully open, vision-centric exploration of multimodal llms.Advances in Neural Information Processing Systems, 37:87310–87356, 2024.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Cambrian-1: A fully open, vision-centric exploration of multimodal llms.Advances in Neural Information Processing Systems, 37:87310–87356, 2024

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.767417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.767417Z digest=sha256:31792101b20bdd8d2a92fb4e85b13d7a613a589ad7700af7aff2b91cda557a6a

Observation f29ee6d0-04d5-475c-a188-3895e076f8bf · outbound

This paper cites Mini-gemini: Mining the potential of multi-modality vision language models.IEEE Transactions on Pattern Analysis and Machine Intelligence, 2025.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Mini-gemini: Mining the potential of multi-modality vision language models.IEEE Transactions on Pattern Analysis and Machine Intelligence, 2025

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.771725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.771725Z digest=sha256:9d2cb83e9f0ca439f9a73c0fb64024e0b15de1ac09651bb3c83b12929dddafeb

Observation 2d1c42fd-5af8-4da6-a155-5163e083b38c · outbound

This paper cites Prismatic vlms: Investigating the design space of visually-conditioned language models.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Prismatic vlms: Investigating the design space of visually-conditioned language models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.775668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.775668Z digest=sha256:dc3639a25b7fbddb6bcee9e691f91ee61abaccd075be4ea1a8a09979ecad4e68

Observation 10e4f40b-8f19-48a1-9d59-9655bd1b4cfd · outbound

This paper cites Feast your eyes: Mixture- of-resolution adaptation for multimodal large language models.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Feast your eyes: Mixture- of-resolution adaptation for multimodal large language models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.780141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.780141Z digest=sha256:84cb5f99c22762c37a52f8b0679fdf4c1b5f5496d1c7755d6e21a638ad12b668

Observation 5080d7c1-0871-46ea-b646-14bbdba9553b · outbound

This paper cites Rouge: A package for automatic evaluation of summaries.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Rouge: A package for automatic evaluation of summaries

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.784508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.784508Z digest=sha256:25dc9ba152d8dded10ba94b581d1807c166de818800b496158233de7d63ad873

Observation 6505c687-0dac-4779-b7d5-58a405e2bebd · outbound

This paper cites Meteor: An automatic metric for mt evaluation with improved correlation with human judgments.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Meteor: An automatic metric for mt evaluation with improved correlation with human judgments

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.788820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.788820Z digest=sha256:ee84574cb20a651aba30f77e86fce80c3465f2d095cfcd099fbab81889dca877

Observation da605372-68d6-42fe-97ac-461ba0f4ab64 · outbound

This paper cites Cider: Consensus-based image description evaluation.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Cider: Consensus-based image description evaluation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.792972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.792972Z digest=sha256:8c172fa58e475dcb4249d01ddc824e7773629c992bf28d73e38cdf6de4db7ad6

Observation 30f4c534-4b37-48e9-9daf-96ed1d01a57a · outbound

This paper cites Weinberger, and Yoav Artzi.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Weinberger, and Yoav Artzi

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.797458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.797458Z digest=sha256:bd94dc8508421118245411b36da1557ebaea76b4d481573c7ba8c65ce4e33b86

Observation 954d929a-361c-468c-b218-3f77940b49d5 · outbound

This paper cites Improving the factual correctness of radiology report generation with semantic rewards.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Improving the factual correctness of radiology report generation with semantic rewards

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.801602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.801602Z digest=sha256:8e624f44c8516212341a55f1192f0f97d3eb607c6d185628df36f120040da27a

Observation 74eeecf1-9592-4266-8097-c60e77b72016 · outbound

This paper cites Ratescore: A metric for radiology report generation.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Ratescore: A metric for radiology report generation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.805839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.805839Z digest=sha256:ca632deec09c6d1b8602a3e1b86a5792b5f333f73ab3463c60a8149e36d9c7f5

Observation 8a447070-b784-423a-a197-19e059605aa6 · outbound

This paper cites Green: Generative radiology report evaluation and error notation.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Green: Generative radiology report evaluation and error notation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.809882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.809882Z digest=sha256:d80477e76addf89721329a8a315461d76cadfa5cbbe52c315ada92f15bc4fed1

Observation 00563525-fba1-4245-bf9e-62c9a1c389c8 · outbound

This paper cites an unresolved cited work.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.814169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.814169Z digest=sha256:e4391114a6f7788cdf64a3051a45d9e07e3cbce1be9a776826eb0a6c2d6f08e0

Observation 1cf78e50-f8f3-4ae8-93c7-8aa8e981fd7f · outbound

This paper cites A survey of llm-based agents in medicine: How far are we from baymax?Findings of the Association for Computational Linguistics: ACL 2025, pages 10345–10359, 2025.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding A survey of llm-based agents in medicine: How far are we from baymax?Findings of the Association for Computational Linguistics: ACL 2025, pages 10345–10359, 2025

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.818221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.818221Z digest=sha256:2b2605dd91faae3266ba6158455fb3a2eb8e189a6c9b7ba69abbeacd048fb879

Observation bc66c442-7f39-4550-9600-e60ee3641b9b · outbound

This paper cites MedAgent-Pro: Towards Evidence-based Multi-modal Medical Diagnosis via Reasoning Agentic Workflow.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding MedAgent-Pro: Towards Evidence-based Multi-modal Medical Diagnosis via Reasoning Agentic Workflow

Reference 39

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.822404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.822404Z digest=sha256:9dd674f453a1064e68127122dce960619abf434b4fd4675530f953e9b7be5459

Observation 3c851925-2376-4bd7-9795-d9bd6eed143d · outbound

This paper cites Mmedagent: Learning to use medical tools with multi-modal agent.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Mmedagent: Learning to use medical tools with multi-modal agent

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.826778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.826778Z digest=sha256:14f814025b24d30290f3c2213ac195fa93fd6f4fd7590433b7e73d733e9fe2c2

Observation 34d450f4-b668-41e3-8946-f4ea977c0be5 · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.831234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.831234Z digest=sha256:9a7cd8fc33fd2d6704444b3f2a0857a0523dc059fc682efe62ca63ec651292a4

Observation 71b7a39e-7caf-4228-9a4b-f294a9286220 · outbound

This paper cites Bimedix2: Bio-medical expert lmm for diverse medical modalities.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Bimedix2: Bio-medical expert lmm for diverse medical modalities

Reference 42

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.835505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.835505Z digest=sha256:89186ebc15b9bc40c9eff55f56248451f15d913ef556422a7eb170d7294984d5

Observation 1999e115-1bfd-43c6-9356-644332f9ddce · outbound

This paper cites Healthgpt: A medical large vision-language model for unifying comprehension and generation via heterogeneous knowledge adaptation.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Healthgpt: A medical large vision-language model for unifying comprehension and generation via heterogeneous knowledge adaptation

Reference 43

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.839618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.839618Z digest=sha256:0268c4c290c7f0fc2797666a915c12258868277a07f761edfd15d3edb157702f

Observation 7f69231d-8d33-486d-929c-3cd549cabdd6 · outbound

This paper cites GSCo: Towards Generalizable AI in Medicine via Generalist-Specialist Collaboration.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding GSCo: Towards Generalizable AI in Medicine via Generalist-Specialist Collaboration

Reference 44

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.844422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.844422Z digest=sha256:d9646d4ea73035d8fed303b5ff8dbc385ff337e4dfe021e91983bba6e3ec5027

Observation 682d512d-c8a4-41c4-bb8d-8f2a23dd6e79 · outbound

This paper cites an unresolved cited work.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.849030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.849030Z digest=sha256:e0c70c2c5550656b07c392a2b1358319ad22bb1736680c208fb44581b4ed666c

Observation 0ced1486-3e33-439b-864f-1bb054377bfa · outbound

This paper cites 3d-rad: A comprehensive 3d radiology med-vqa dataset with multi-temporal analysis and diverse diagnostic tasks.arXiv preprint arXiv:2506.11147, 2025.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding 3d-rad: A comprehensive 3d radiology med-vqa dataset with multi-temporal analysis and diverse diagnostic tasks.arXiv preprint arXiv:2506.11147, 2025

Reference 47

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.864324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.864324Z digest=sha256:f6ec12c1d8532c03db0986d705d4521058e2f119180dd8ebcc627ae9a6ec1e61

Observation f19110b7-9d69-4be2-ba0c-6277f055a8b5 · outbound

This paper cites Amos: A large-scale abdominal multi-organ benchmark for versatile medical image segmentation.Advances in neural information processing systems, 35:36722–36732, 2022.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Amos: A large-scale abdominal multi-organ benchmark for versatile medical image segmentation.Advances in neural information processing systems, 35:36722–36732, 2022

Reference 48

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.869180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.869180Z digest=sha256:5eab7c77487d729bcee5c8788d34aba8065c4f9a86eb7bfa4da5c1d0038f0cce

Observation 5e5af2aa-0fb2-443a-928f-7d2d087eacd6 · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Bleu: a method for automatic evaluation of machine translation

Reference 49

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.874150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.874150Z digest=sha256:573138b20eab3f8d095d4db36eeb692cba431725ddc6204244a1513a3c6950ae

Observation 4304564b-fd87-4f6d-9dba-5e209839f70c · outbound

This paper cites RadGraph-XL: A large-scale expert-annotated dataset for entity and relation extraction from radiology reports.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding RadGraph-XL: A large-scale expert-annotated dataset for entity and relation extraction from radiology reports

Reference 50

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.878777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.878777Z digest=sha256:f74fa1f88c37a08b5e746733e0e1095b34752b486d599a74a6dc4ffac1ebd645

Observation 1543372b-c5d7-4443-a577-3f5076039b8c · outbound

This paper cites Switch transformers: Scaling to trillion parameter models with simple and efficient sparsity.Journal of Machine Learning Research, 23(120):1–39, 2022.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Switch transformers: Scaling to trillion parameter models with simple and efficient sparsity.Journal of Machine Learning Research, 23(120):1–39, 2022

Reference 51

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.883448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.883448Z digest=sha256:5ec5a17dbffaf9b51b312a0891072f23b47769170693b912b3b184e67a58f6b3

Observation bd93326b-7a8b-4b29-a647-693440de4ffc · outbound

This paper cites Position: Compositional generative modeling: A single model is not all you need.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Position: Compositional generative modeling: A single model is not all you need

Reference 52

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.887609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.887609Z digest=sha256:cf4cd813d0318c7fa26e0789796d6e02157991621265e8efd3ae737c89a629dd

Observation 2c418c22-aee1-4e89-aa7e-a474ae753f87 · outbound

This paper cites A convnet for the 2020s.Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2022.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding A convnet for the 2020s.Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2022

Reference 53

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.892067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.892067Z digest=sha256:f8951acd9a27f592baa8aa8b9d4335065567bdc79da38e35fb3cfdfd40417236

Observation f61d1b31-4371-4300-85f9-8f0dc26897f6 · outbound

This paper cites an unresolved cited work.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.896289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.896289Z digest=sha256:4abdad26f0a2fa785baeffb37cff120f994fb2a5a78087e43837a1121a7d8cf9

Observation 7be93ead-d69b-4dfd-ac23-e88a0f103224 · outbound

This paper cites Exploring scalable medical image encoders beyond text supervision.Nature Machine Intelligence, 7(1):119–130, 2025.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Exploring scalable medical image encoders beyond text supervision.Nature Machine Intelligence, 7(1):119–130, 2025

Reference 55

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.900617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.900617Z digest=sha256:107bef2dbc602c3a7fa4c3591634a4e45c01699acc207c98c1ad61d9f96a69f0

Observation a4f6110b-4507-4485-9621-b0c65035518c · outbound

This paper cites A multimodal biomedical foundation model trained from fifteen million image–text pairs.Nejm Ai, 2(1):AIoa2400640, 2025.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding A multimodal biomedical foundation model trained from fifteen million image–text pairs.Nejm Ai, 2(1):AIoa2400640, 2025

Reference 56

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.905366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.905366Z digest=sha256:c12fbb67329a1eb241ac2c04aea5ce4b8b0534a224aa59cb9fe563a55e797e47

Observation 4df93691-7ce3-48aa-9912-133539a07a24 · outbound

This paper cites Visionzip: Longer is better but not necessary in vision language models.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Visionzip: Longer is better but not necessary in vision language models

Reference 57

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.909560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.909560Z digest=sha256:0076cc5dd9d04374618964acbe3379ab4fa508849b22f8174f22f13562cc0c88

Observation 38502715-f4d4-4147-a183-59a62a097b6b · outbound

This paper cites Sparsevlm: Visual token sparsification for efficient vision-language model inference.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Sparsevlm: Visual token sparsification for efficient vision-language model inference

Reference 58

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.913652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.913652Z digest=sha256:694521d69ee20f08dd9c1ba85470f150db07bb2ae6e7a8e5270517d9e1b86336

Observation cfb0ae7e-81b0-4a1f-baf9-03bb01564451 · outbound

This paper cites Attention is all you need.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Attention is all you need

Reference 59

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.918023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.918023Z digest=sha256:e220e3ec7940e1b84baf103dd506c53a431e5ec9063ae1f303b9783fbc4f5700

Observation 31c4033e-3bc0-4396-b0fe-ac91b5baedfc · outbound

This paper cites Deepstack: Deeply stacking visual tokens is surprisingly simple and effective for lmms.Advances in Neural Information Processing Systems, 37:23464–23487, 2024.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Deepstack: Deeply stacking visual tokens is surprisingly simple and effective for lmms.Advances in Neural Information Processing Systems, 37:23464–23487, 2024

Reference 60

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.922247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.922247Z digest=sha256:20b66ffab38c9a7c00dbdfc9b68468f67cada2c92b2d627eaf0116d1698f6bb0

Observation b4cbec94-23d4-49eb-a553-a513b6a4887f · outbound

This paper cites Perception Encoder: The best visual embeddings are not at the output of the network.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Perception Encoder: The best visual embeddings are not at the output of the network

Reference 61

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.926532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.926532Z digest=sha256:3ac6d5c9f999045c5a2ce0cf0a04934b3a72ed0d9b2bc07e1f397ebde95ab926

Observation 3e70cf8e-e305-4cc9-89c9-212a3477561f · outbound

This paper cites Generalized radiograph representation learning via cross-supervision between images and free-text radiology reports.Nature Machine Intelligence, 4(1):32–40, 2022.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Generalized radiograph representation learning via cross-supervision between images and free-text radiology reports.Nature Machine Intelligence, 4(1):32–40, 2022

Reference 62

Resolution
verified exact
doi, observed 2026-07-31T06:20:56.429503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-31T06:20:29.931059Z digest=sha256:df441aed5f3f496279496d365393ed8f5d29b380f96c33cd4261b2616044d25e

Observation 64387d1d-4d25-4d61-a423-b2c7d9693099 · outbound

This paper cites Huatuogpt, towards taming language model to be a doctor.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Huatuogpt, towards taming language model to be a doctor

Reference 63

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.935881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.935881Z digest=sha256:da14a93a02608cf2351642fac4f53ba97f5476e781e93579ab2ef03bd0cdbf8c

Observation a4a7c094-d03c-48f6-8ade-236f3518099e · outbound

This paper cites State of what art? a call for multi-prompt llm evaluation.Transactions of the Association for Computational Linguistics, 12: 933–949, 2024.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding State of what art? a call for multi-prompt llm evaluation.Transactions of the Association for Computational Linguistics, 12: 933–949, 2024

Reference 64

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.940553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.940553Z digest=sha256:5e203692ec8b8fc3a52b2ddafcc0b842c6bd12232c203c0997a303f55840b17e

Observation 06246414-fd9f-43a5-8f22-9ea711e55c07 · outbound

This paper cites Surveillance, epidemiology, and end results (seer) program (www.seer.cancer.gov), 2025.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Surveillance, epidemiology, and end results (seer) program (www.seer.cancer.gov), 2025

Reference 65

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.944785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.944785Z digest=sha256:e0ad3f5c8553d684db60871b7ca882586a61d73dd3ca06efb5eeb0213b96fa03

Observation e0874377-4866-4051-8d9f-6a3603acbcb3 · outbound

This paper cites Look again, think slowly: Enhancing visual reflection in vision-language models.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Look again, think slowly: Enhancing visual reflection in vision-language models

Reference 66

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.949392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.949392Z digest=sha256:772f6d1210b59eebe237caab3dfe42e9b21ec6737e452afd7893b819e6932486

Observation 9b759028-84a6-47a2-ba64-7c60aba90e4a · outbound

This paper cites M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 67

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.953458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.953458Z digest=sha256:161d2b7a16c7603b8c18a5e62633c8c7bad0aa19dcaed41af452f3fdb765b1e9

Observation 5e511336-50fd-469f-8431-52dae50e0fdf · outbound

This paper cites Merlin: a computed tomography vision–language foundation model and dataset.Nature, pages 1–11, 2026.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Merlin: a computed tomography vision–language foundation model and dataset.Nature, pages 1–11, 2026

Reference 68

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.957764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.957764Z digest=sha256:03045d02756ee83bbc21bb35a95ba924ea8994afb2081779ee7250c6edc8a2c4

Observation 84157aed-d860-4493-a2e7-510412f936f5 · outbound

This paper cites Inspect: A multimodal dataset for patient outcome prediction of pulmonary embolisms.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Inspect: A multimodal dataset for patient outcome prediction of pulmonary embolisms

Reference 69

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.962001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.962001Z digest=sha256:f36e61dcf292ce228dbe70ac870820c5e02722ef9c03cdb1cf671fc0cf8a89b1

Observation 04772d10-3435-4fc5-8d68-98c7db1754e6 · outbound

This paper cites Vista3d: A unified segmentation foundation model for 3d medical imaging.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Vista3d: A unified segmentation foundation model for 3d medical imaging

Reference 70

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.966447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.966447Z digest=sha256:047f1ff15c47f621d627b3555e79f9339f9b032668ed6a0a8357b63f1139e6f5

Observation c365480e-31b9-4f43-9d62-45c4c010eace · outbound

This paper cites Totalsegmentator: robust segmentation of 104 anatomic structures in ct images.Radiology: Artificial Intelligence, 5(5):e230024, 2023.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Totalsegmentator: robust segmentation of 104 anatomic structures in ct images.Radiology: Artificial Intelligence, 5(5):e230024, 2023

Reference 71

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.970641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.970641Z digest=sha256:20eb8efbe8491a9779393be501edd8253cdbe1a7dbf83de9acec39f9bf434533

Observation 2c1b3231-9293-4041-ac81-a230c24044ba · outbound

This paper cites DMQR-RAG: Diverse Multi-Query Rewriting for RAG.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding DMQR-RAG: Diverse Multi-Query Rewriting for RAG

Reference 72

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.975213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.975213Z digest=sha256:c5c1972f6cf54c31a1685ba24b61bd06aa40dd0f0f8b8f5559d70b2559439987

Observation e46f354f-7c63-43b2-b8b3-82c6b1050814 · outbound

This paper cites UMLS knowledge sources, release 2024aa, 2024.http://www.nlm.nih.gov/ research/umls/licensedcontent/umlsknowledgesources.html(accessed 15 July 2024).

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding UMLS knowledge sources, release 2024aa, 2024.http://www.nlm.nih.gov/ research/umls/licensedcontent/umlsknowledgesources.html(accessed 15 July 2024)

Reference 73

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.979544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.979544Z digest=sha256:7a810b43cf58041cfae35a8bda1cf8c84bb31b6a084a5a3329ff3b607b3b4783

Observation 03895d10-25c3-4a81-8d4e-771d8ab7785a · outbound

This paper cites Building a knowledge graph to enable precision medicine.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Building a knowledge graph to enable precision medicine

Reference 74

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.983879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.983879Z digest=sha256:fcee02cd93aa0395633595bb7bb62c455996af85aebb34ee55be00d944ce2e40

Observation 5fb49e35-285a-4a91-a43c-3b8d4ba5cb26 · outbound

This paper cites Sayers, Evan E.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Sayers, Evan E

Reference 75

Resolution
verified exact
doi, observed 2026-07-31T06:20:56.390675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-31T06:20:29.988047Z digest=sha256:c75aa3b2750b1f7d22d3f0aef075348bf794664af14dc037cc1c14cec8e1df92

Observation 384e0de7-46d0-4a79-8e44-520119b13afa · outbound

This paper cites StatPearls Publishing, Treasure Island, FL, 2026.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding StatPearls Publishing, Treasure Island, FL, 2026

Reference 76

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.992608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.992608Z digest=sha256:cfb2f41326f294129116e7935f6595d7ef9db2a46b2eee0d4ab212cb23a4693c

Observation abe30840-efc1-49ed-bbe7-c2d0b02a7aa8 · outbound

This paper cites What disease does this patient have? a large-scale open domain question answering dataset from medical exams.Applied Sciences, 11 (14):6421, 2021.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding What disease does this patient have? a large-scale open domain question answering dataset from medical exams.Applied Sciences, 11 (14):6421, 2021

Reference 77

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.997011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.997011Z digest=sha256:d7b218ad42ba8945fc2b750b4a06e86749f7b60f7af10bdcbf6ec79fe61f2347

Observation a656dbeb-ed18-4bab-9941-bd063f1b57b4 · outbound

This paper cites Xiong, Q.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Xiong, Q

Reference 78

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.001898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.001898Z digest=sha256:18e7fe408ffa5782e5c5c01ca335426faf48b3c60922e0fb9ce02f046f25922c

Observation 3708dfa9-d6a4-4317-831a-718fb0560e1f · outbound

This paper cites Multi-modal ai for opportunistic screening, staging and progression risk stratification of steatotic liver disease.Nature Communications, 17(1):1562, 2026.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Multi-modal ai for opportunistic screening, staging and progression risk stratification of steatotic liver disease.Nature Communications, 17(1):1562, 2026

Reference 79

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.006389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.006389Z digest=sha256:3818fe6be0cd3041241d6000e13ffc26fda3522aa2715ca293a4c2a880c1dca0

Observation 6042fd0f-a957-43bd-8664-2c3c213c0fd4 · outbound

This paper cites Effective lymph nodes detection in ct scans using location debiased query selection and contrastive query representation in transformer.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Effective lymph nodes detection in ct scans using location debiased query selection and contrastive query representation in transformer

Reference 80

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.011190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.011190Z digest=sha256:29e42cb6395099f2a1267066f3984f14bf1256de90b1f2d61c866afd7103b13a

Observation 159ffc65-f324-454c-a538-0d1d3e18b61e · outbound

This paper cites On the limits of cross-domain generalization in automated x-ray prediction.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding On the limits of cross-domain generalization in automated x-ray prediction

Reference 81

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.016358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.016358Z digest=sha256:bf583612cf4ce55361493fa946092d3e1caaf86e4cae93383f0ec5c5ce55220e

Observation e919322b-8d10-48ed-996e-b91ea5d17d1c · outbound

This paper cites Kalra, and Pingkun Yan.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Kalra, and Pingkun Yan

Reference 82

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.020989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.020989Z digest=sha256:2c2ee3b83d7da9f943d556a1e3794ea19ef51bd8731c9f08cbd393eb511f9e2c

Observation a093b395-a272-4cb4-8da1-a57e4c359fca · outbound

This paper cites A Vision-Language Foundation Model to Enhance Efficiency of Chest X-ray Interpretation.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding A Vision-Language Foundation Model to Enhance Efficiency of Chest X-ray Interpretation

Reference 83

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.025041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.025041Z digest=sha256:0e936f40bd65d0749ad2dac103ae662674e6a26a3440a584885e73c9bc69a084

Observation b5b38be1-7521-4b5c-adfc-85ff63c74427 · outbound

This paper cites Pmc-clip: Contrastive language-image pre-training using biomedical documents.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Pmc-clip: Contrastive language-image pre-training using biomedical documents

Reference 84

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.029504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.029504Z digest=sha256:730e89ae254ddd3dca8ff1113fabf98cd0e079bf22a3c8c9a63d768c7fae95b2

Observation 58a121f4-155e-4ada-b1ab-f3173442f15f · outbound

This paper cites Friedrich.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Friedrich

Reference 85

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.033853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.033853Z digest=sha256:2b8b9a4d2133ff6643709b2f4e6bbba2e19c6622ba2e907c5d226548e39c5ca1

Observation c1add6ae-b71d-4d34-9289-1c6a762659e1 · outbound

This paper cites Seco de Herrera, et al.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Seco de Herrera, et al

Reference 86

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.038050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.038050Z digest=sha256:9671ce82eff280da1a748ce3899e6a78092cf05709875ac371e48ad612b76512

Observation 105fe18f-caf5-4881-ae5d-36a9392a8f7f · outbound

This paper cites Towards injecting medical visual knowledge into multimodal llms at scale.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Towards injecting medical visual knowledge into multimodal llms at scale

Reference 87

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.042211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.042211Z digest=sha256:b3b2e30d4e1fa71afa18e4bbac9b8aad10d8b19675f49c4178ce0fa46efd2a20

Observation 59a2440a-d3e0-4a68-9668-59930d2c371b · outbound

This paper cites Mimic-cxr, a de-identified publicly available database of chest radiographs with free-text reports.Scientific data, 6(1):317, 2019.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Mimic-cxr, a de-identified publicly available database of chest radiographs with free-text reports.Scientific data, 6(1):317, 2019

Reference 88

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.046871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.046871Z digest=sha256:0fe88bd43f5b3f03e769460880042e09f3676ba1a691e05d0634f4fdf357901a

Observation ccb2ac9a-77c4-4063-8d35-e192150b92b9 · outbound

This paper cites Medicat: A dataset of medical images, captions, and textual references.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Medicat: A dataset of medical images, captions, and textual references

Reference 90

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.056055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.056055Z digest=sha256:1b2f7e19aa1df2e6783d9c012442818e9d72f2a1d2240797ff22bc2156068b36

Observation d015b708-69f2-4042-8c6a-3fc7a8411b2b · outbound

This paper cites an unresolved cited work.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Unresolved cited work

Reference 91

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.060699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.060699Z digest=sha256:dcadd124ebe5b767bde159952b1619b6791e5a87cabbc16ed65bdc752a9b7c33

Observation 07e4553a-9360-42d5-95fa-27317f889765 · outbound

This paper cites an unresolved cited work.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Unresolved cited work

Reference 92

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.065156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.065156Z digest=sha256:cd1eaa9d922415517e9c68f678e6eee6202fc28dcb76fb3a0a6de582ca8b8291

Observation 63c57ade-5f9c-4fd5-affb-f688d969e969 · outbound

This paper cites Improved baselines with visual instruction tuning.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Improved baselines with visual instruction tuning

Reference 93

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.069484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.069484Z digest=sha256:c4c86f56df2ade24f2c4cb572349184142ccc9b703b35e9b152d6efc0d4bc07b

Observation 1e5965f2-1145-42a5-9ee1-e80072f3dab6 · outbound

This paper cites Molmo and pixmo: Open weights and open data for state-of-the-art vision-language models.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Molmo and pixmo: Open weights and open data for state-of-the-art vision-language models

Reference 94

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.073880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.073880Z digest=sha256:4b5931e421e42dc20bd67754a1638a02fbed88c7ebbcfa54a75e37abcb2115e3

Observation 01c68a65-9112-43cf-a98d-25e9db79fa09 · outbound

This paper cites Llava-med: Training a large language-and-vision assistant for biomedicine in one day.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Llava-med: Training a large language-and-vision assistant for biomedicine in one day

Reference 95

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.078185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.078185Z digest=sha256:138ae5dddff60c685f9df2764b3233c3215a5f50d292b01d885823cc78d3c664

Observation b64b981d-a68c-4836-8ffe-03af00a43128 · outbound

This paper cites Quilt- llava: Visual instruction tuning by extracting localized narratives from open-source histopathology videos.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Quilt- llava: Visual instruction tuning by extracting localized narratives from open-source histopathology videos

Reference 96

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.082969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.082969Z digest=sha256:efc7e5ca7c889a72b2ea98df0d9a96e01a7e607e5f2d4ed924168204422eae12

Observation f1f8acc9-da48-4bef-bd68-a6477220783d · outbound

This paper cites Hicks, Vajira Thambawita, Pål Halvorsen, and Michael A.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Hicks, Vajira Thambawita, Pål Halvorsen, and Michael A

Reference 97

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.087647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.087647Z digest=sha256:f086ad13e5581175d91236e0514965a56049153c8f61ac15c596ab2073ebdc85

Observation 4f143619-ee0e-4ef5-96f7-5b29a2406603 · outbound

This paper cites MIMIC-Ext-MIMIC-CXR-VQA: A complex, diverse, and large-scale visual question answering dataset for chest x-ray images, 2024.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding MIMIC-Ext-MIMIC-CXR-VQA: A complex, diverse, and large-scale visual question answering dataset for chest x-ray images, 2024

Reference 98

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.091803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.091803Z digest=sha256:cf1560664cf96646adcaf493bd4a895fda4c92339f320b3a2fac44e9e0444e30

Observation 9c99380f-b213-4281-ba1c-f98dce2d3856 · outbound

This paper cites Towards visual question answering on pathology images.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Towards visual question answering on pathology images

Reference 99

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.096393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.096393Z digest=sha256:f57b96719b70bd77fcdb1c43ca0342f0b27792169cb58a979fc29ea73a873d35

Observation ae49aa77-1870-44e2-8298-90c3dbdc2ca8 · outbound

This paper cites Development of a large-scale medical visual question-answering dataset.Communications Medicine, 4(1):23, 2024.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Development of a large-scale medical visual question-answering dataset.Communications Medicine, 4(1):23, 2024

Reference 100

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.101128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.101128Z digest=sha256:091119bcca5fb94b43d895d91176addb4fffb4f2021dc3300eb802756a3d7c80

Observation 715c13e9-60ac-49c9-b201-4f5b5e915322 · outbound

This paper cites Hasan, Vivek V.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Hasan, Vivek V

Reference 101

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.105764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.105764Z digest=sha256:b813a1cdaa4dc56c6489486ae198b270f6f4ebd21dc069f62d41dad9e07e4bf4

Observation 2be0753d-faab-421f-a9a6-656381fbae80 · outbound

This paper cites A dataset of clinically generated visual questions and answers about radiology images.Scientific data, 5(1):1–10, 2018.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding A dataset of clinically generated visual questions and answers about radiology images.Scientific data, 5(1):1–10, 2018

Reference 102

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.110246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.110246Z digest=sha256:180537242984766ce397c0e964fc5de6158ff7c0891da807e32f8102a1b527af

Observation f5b30254-a857-401d-8ab4-7216c357b7dc · outbound

This paper cites GMAI-VL-R1: Harnessing Reinforcement Learning for Multimodal Medical Reasoning.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding GMAI-VL-R1: Harnessing Reinforcement Learning for Multimodal Medical Reasoning

Reference 103

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.115352Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.115352Z digest=sha256:f98aa4a34cc3344fb870cbec5b7137eec3c6292a86c32d29966c4ca63208a210

Pith citing papers

No inbound Pith citation observations are available.