Pith. sign in

Paper Citation Record · LEDGER

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding

As of 13 August 2026, this Paper Citation Record lists 100 of 127 outbound references and 0 inbound Pith citation observations for arXiv:2607.24743.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.24743 v2

Coverage vector

measured 100 of 127 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-31T06:20:30.115352Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 127 outbound references displayed

  • verified exact3
  • verified fuzzy0
  • unresolved97
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 688ad64a-4194-4329-a302-5516cadadec9 · outbound

This paper cites Qwen2.5-VL Technical Report.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Qwen2.5-VL Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.650039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.650039Z digest=sha256:07a86c647c874be7ae1abfab94a818d801ca1866061da63eff7bc837b68e0a87

Observation 349f8bd2-2ded-401a-8486-9512b3deb042 · outbound

This paper cites Qwen3-VL Technical Report.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Qwen3-VL Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.655604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.655604Z digest=sha256:5b0e801d763e8c78f0f17c02b47f818443c1337a3a9dacbab5a6ff11cfb54668

Observation 3dd06eb9-99b0-40f8-828d-c3f00778136c · outbound

This paper cites Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.660200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.660200Z digest=sha256:9d5e541ad99b5177185977639b559724a558f65ab206eabbfad0f0ae062d7e19

Observation a96911c1-0072-49d4-864d-26b1f8ab2d15 · outbound

This paper cites GPT-4o System Card.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding GPT-4o System Card

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.664434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.664434Z digest=sha256:1d218b0f58464cf3fd20fe05910d2743faff5b078cee2faf73c1e677075b9d2e

Observation 63f3b688-a626-4a43-85fb-2654e1785562 · outbound

This paper cites Improved baselines with visual instruction tuning.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Improved baselines with visual instruction tuning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.668884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.668884Z digest=sha256:2858e02b099a913ea2a67fe7da5fde5e2ba9f4d6e12b8d356c75c98ef2ae582e

Observation aba874c2-a026-4fd2-b4a4-731a5210f7b0 · outbound

This paper cites Lingshu: A Generalist Foundation Model for Unified Multimodal Medical Understanding and Reasoning.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Lingshu: A Generalist Foundation Model for Unified Multimodal Medical Understanding and Reasoning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.673214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.673214Z digest=sha256:7c68a8a155c2d65b3bc2d8b2c9c679d444ac1bebf94522d9780df688818fa1a9

Observation dd47a75d-3903-441a-8942-a4f2eb05c14d · outbound

This paper cites Hulu-med: A transparent generalist model towards holistic medical vision-language understanding.arXiv preprint arXiv:2510.08668, 2025.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Hulu-med: A transparent generalist model towards holistic medical vision-language understanding.arXiv preprint arXiv:2510.08668, 2025

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.678846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.678846Z digest=sha256:f77934192ef2188dc577eec3a60ecab9f7b8f6c0606e293b0c01b3644d23d2f9

Observation 5b734fb2-dc20-4584-8703-961c506ed543 · outbound

This paper cites MedGemma Technical Report.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding MedGemma Technical Report

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.683476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.683476Z digest=sha256:57fcb55bce74ce53a54b55bea68102f6f12b2fab4165c159dd3f5e98a4d2326e

Observation 8755a2ba-2e51-42bd-80de-a43a8b95054f · outbound

This paper cites Medvlm-r1: Incentivizing medical reasoning capability of vision-language models (vlms) via reinforcement learning.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Medvlm-r1: Incentivizing medical reasoning capability of vision-language models (vlms) via reinforcement learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.693151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.693151Z digest=sha256:30eb064a80ddeccd46be93fe1e99aa7e9ed0e35881ed0c6d3b3db2f0fa7d0786

Observation 0b3a0f5f-d5e8-46a1-a864-7d452677729f · outbound

This paper cites Med- r1: Reinforcement learning for generalizable medical reasoning in vision-language models.arXiv preprint arXiv:2503.13939, 2025.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Med- r1: Reinforcement learning for generalizable medical reasoning in vision-language models.arXiv preprint arXiv:2503.13939, 2025

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.697647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.697647Z digest=sha256:370c9f5710c05240764c5176a49f7901abebca59a7b1ee18c57f0e603b9436b9

Observation 2c12473b-74ac-4ee9-8de7-42ac4d543ef4 · outbound

This paper cites HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.702013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.702013Z digest=sha256:6350ff86342e288986ee200db88305eaa54864ad06783eb43927cc4034248fd1

Observation be92fe12-4dd8-4e57-8f9c-bfcd91512494 · outbound

This paper cites A generalist vision–language foundation model for diverse biomedical tasks.Nature Medicine, 30(11):3129–3141, 2024.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding A generalist vision–language foundation model for diverse biomedical tasks.Nature Medicine, 30(11):3129–3141, 2024

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.706803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.706803Z digest=sha256:bf50a228b6bc33fea8a8c0dba144db6527dae06ef0efb3e06b7adbef269a34f6

Observation c19e5cb2-31e4-4fe3-a8a7-20a7cea10345 · outbound

This paper cites Omn- imedvqa: A new large-scale comprehensive evaluation benchmark for medical lvlm.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Omn- imedvqa: A new large-scale comprehensive evaluation benchmark for medical lvlm

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.711046Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.711046Z digest=sha256:d8d429c05f07c99cd3d95c42f665ee158e11fa0332a42fcaf57b42ad4f942058

Observation 18b50e56-e3fe-4c59-a3b3-b74111262837 · outbound

This paper cites Slake: A semantically-labeled knowledge- enhanced dataset for medical visual question answering.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Slake: A semantically-labeled knowledge- enhanced dataset for medical visual question answering

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.715200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.715200Z digest=sha256:70962506926c22776cbd54db971dd251ce83ed37a8c7889aaa88aabc6870adb1

Observation 4ad5dd33-9273-4382-9b97-c1255719f2d4 · outbound

This paper cites Generalist foundation models from a multimodal dataset for 3d computed tomography.Nature Biomedical Engineering, pages 1–19, 2026.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Generalist foundation models from a multimodal dataset for 3d computed tomography.Nature Biomedical Engineering, pages 1–19, 2026

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.719488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.719488Z digest=sha256:4450ba04cf3b3869d299d94da87004121a707c4a5edf4f233a7caa0c35f408dc

Observation 210aef49-4e8d-4ae8-a311-b98061527cc1 · outbound

This paper cites Measuring massive multitask language understanding.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Measuring massive multitask language understanding

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.724241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.724241Z digest=sha256:e44635f6d62e7eef39bf0f6f5c3d24d8222028cfcdf858a086ceb21be52a95b2

Observation d74a9669-f3fd-419a-b0e2-a8e961260548 · outbound

This paper cites Medxpertqa: Benchmarking expert-level medical reasoning and understanding.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Medxpertqa: Benchmarking expert-level medical reasoning and understanding

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.733500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.733500Z digest=sha256:db354314b3bcbaf71faac30d17fff6a2b2bfe3134c71353445842aedc38012c9

Observation 1623dc72-a475-45b3-9c22-5b5d251b6b30 · outbound

This paper cites Pubmedqa: A dataset for biomedical research question answering.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Pubmedqa: A dataset for biomedical research question answering

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.737585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.737585Z digest=sha256:b1e2c6adbd37824dfbfa3a56fcd4772bf788ebc9ea31b56aea61deb591d67c17

Observation 4b315706-af7f-44c2-a99c-abd0e690293e · outbound

This paper cites Towards a holistic framework for multimodal LLM in 3D brain CT radiology report generation.Nature Communications, 16(1):2258, 2025.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Towards a holistic framework for multimodal LLM in 3D brain CT radiology report generation.Nature Communications, 16(1):2258, 2025

Reference 20

Resolution
verified exact
doi, observed 2026-07-31T06:20:56.545754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-07-31T06:20:29.741466Z digest=sha256:51af9f3c8aadecbd2d64570f47ae938daa1cea25c6ff37fdbea00e5ed6ed8f9b

Observation a1ce074b-85e6-4ec6-8a9f-5a17c378d055 · outbound

This paper cites Med3dvlm: An efficient vision-language model for 3d medical image analysis.IEEE Journal of Biomedical and Health Informatics, 2025.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Med3dvlm: An efficient vision-language model for 3d medical image analysis.IEEE Journal of Biomedical and Health Informatics, 2025

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.745685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.745685Z digest=sha256:e3bf1357a07a93492a0d51485ec1abf88ec9955f7a21ffcf142e15c4e4c3c117

Observation 3a62ce93-2cff-423e-9fd2-51b951816f66 · outbound

This paper cites Towards generalist foundation model for radiology by leveraging web-scale 2d&3d medical data.Nature Communications, 16(1):7866, 2025.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Towards generalist foundation model for radiology by leveraging web-scale 2d&3d medical data.Nature Communications, 16(1):7866, 2025

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.749856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.749856Z digest=sha256:a43ef52b7b03b8fb589aa4dc1a4a0f16feb4e2fcc1cff7dc39926074e30fdda2

Observation dfc17799-a2e0-4108-9c70-1c3f349ac05c · outbound

This paper cites Chaunzwa, Simon Bernatz, Ahmed Hosny, Raymond H.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Chaunzwa, Simon Bernatz, Ahmed Hosny, Raymond H

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.754042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.754042Z digest=sha256:00e02d5cc8270246c00874cb9206db7d6d6f2eda5ec7c0bccda1116403f23b2f

Observation 48887214-bd09-4b60-990a-71e6decaaa8a · outbound

This paper cites Eagle: Exploring the design space for multimodal llms with mixture of encoders.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Eagle: Exploring the design space for multimodal llms with mixture of encoders

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.758357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.758357Z digest=sha256:939c8104af9eab389a4af5e7bc0a162652fbe3fedeacbaafbea44d6811c26790

Observation d2ce34c7-52cc-43da-b23b-f4584169ce5e · outbound

This paper cites Eagle-2: Faster inference of language models with dynamic draft trees.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Eagle-2: Faster inference of language models with dynamic draft trees

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.763082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.763082Z digest=sha256:27a75fe82ecb52f6be2c738093b637372a60ba8da761b9edd2b469b6ed65ee3f

Observation 7d5e7133-0421-4f48-8fdf-57c255b3cadc · outbound

This paper cites Cambrian-1: A fully open, vision-centric exploration of multimodal llms.Advances in Neural Information Processing Systems, 37:87310–87356, 2024.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Cambrian-1: A fully open, vision-centric exploration of multimodal llms.Advances in Neural Information Processing Systems, 37:87310–87356, 2024

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.767417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.767417Z digest=sha256:6ab36e6d2bbb741dadccf2b27491bc6d966a299302439478fcc9960640559cb9

Observation f29ee6d0-04d5-475c-a188-3895e076f8bf · outbound

This paper cites Mini-gemini: Mining the potential of multi-modality vision language models.IEEE Transactions on Pattern Analysis and Machine Intelligence, 2025.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Mini-gemini: Mining the potential of multi-modality vision language models.IEEE Transactions on Pattern Analysis and Machine Intelligence, 2025

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.771725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.771725Z digest=sha256:aeb46187afb603d28ccc0fa841f69e03c84f04848ccece148b9c3107eeeffcff

Observation 2d1c42fd-5af8-4da6-a155-5163e083b38c · outbound

This paper cites Prismatic vlms: Investigating the design space of visually-conditioned language models.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Prismatic vlms: Investigating the design space of visually-conditioned language models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.775668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.775668Z digest=sha256:68d1e0c284e9b33057d9430ba681ecc38c9ee5fada9b5c00d6a00dac82e4a5fe

Observation 10e4f40b-8f19-48a1-9d59-9655bd1b4cfd · outbound

This paper cites Feast your eyes: Mixture- of-resolution adaptation for multimodal large language models.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Feast your eyes: Mixture- of-resolution adaptation for multimodal large language models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.780141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.780141Z digest=sha256:ca59ad89fd089ab5b62838cd78b0016c82b2953dbd7238da516cdea950f80a21

Observation 5080d7c1-0871-46ea-b646-14bbdba9553b · outbound

This paper cites Rouge: A package for automatic evaluation of summaries.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Rouge: A package for automatic evaluation of summaries

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.784508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.784508Z digest=sha256:1f1c7cd16d51d537fbb428dc0a91ec9e438b333c1b790bcb65be97124c458e3e

Observation 6505c687-0dac-4779-b7d5-58a405e2bebd · outbound

This paper cites Meteor: An automatic metric for mt evaluation with improved correlation with human judgments.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Meteor: An automatic metric for mt evaluation with improved correlation with human judgments

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.788820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.788820Z digest=sha256:3d51ac568e7069dacb32eb1738a7ac8b91bb250b2166658e8312a4e1127f1f8e

Observation da605372-68d6-42fe-97ac-461ba0f4ab64 · outbound

This paper cites Cider: Consensus-based image description evaluation.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Cider: Consensus-based image description evaluation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.792972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.792972Z digest=sha256:7a0a547979b7ca63a6d4ace5795b565c7dcb18eaf1612921395ffe442913d1f6

Observation 30f4c534-4b37-48e9-9daf-96ed1d01a57a · outbound

This paper cites Weinberger, and Yoav Artzi.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Weinberger, and Yoav Artzi

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.797458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.797458Z digest=sha256:74373dc4435d0688f2b58afc3b61a58021e284b8c8126fb272269e6b7b02e9ee

Observation 954d929a-361c-468c-b218-3f77940b49d5 · outbound

This paper cites Improving the factual correctness of radiology report generation with semantic rewards.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Improving the factual correctness of radiology report generation with semantic rewards

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.801602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.801602Z digest=sha256:ec735376dbc0f150c423e669b637f2ca42660c6d1a3bbf83157500862f1e6841

Observation 74eeecf1-9592-4266-8097-c60e77b72016 · outbound

This paper cites Ratescore: A metric for radiology report generation.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Ratescore: A metric for radiology report generation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.805839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.805839Z digest=sha256:a785850dba4792dfaacb01260fed1e7f810ceec14bc203e447fa3db5e972e4a1

Observation 8a447070-b784-423a-a197-19e059605aa6 · outbound

This paper cites Green: Generative radiology report evaluation and error notation.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Green: Generative radiology report evaluation and error notation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.809882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.809882Z digest=sha256:bb60818a049b82be6537054286416a82281dddbfd30439122e4a33f559a11ee7

Observation 00563525-fba1-4245-bf9e-62c9a1c389c8 · outbound

This paper cites an unresolved cited work.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.814169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.814169Z digest=sha256:670257afd38ca87abd49668adbc71457abc9ad03d8233990911fdebdd1e5ce64

Observation 1cf78e50-f8f3-4ae8-93c7-8aa8e981fd7f · outbound

This paper cites A survey of llm-based agents in medicine: How far are we from baymax?Findings of the Association for Computational Linguistics: ACL 2025, pages 10345–10359, 2025.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding A survey of llm-based agents in medicine: How far are we from baymax?Findings of the Association for Computational Linguistics: ACL 2025, pages 10345–10359, 2025

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.818221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.818221Z digest=sha256:63587b40cf2531662dfac4bffcc188062acc3b961297ef0c4abb120f3cd924f7

Observation bc66c442-7f39-4550-9600-e60ee3641b9b · outbound

This paper cites MedAgent-Pro: Towards Evidence-based Multi-modal Medical Diagnosis via Reasoning Agentic Workflow.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding MedAgent-Pro: Towards Evidence-based Multi-modal Medical Diagnosis via Reasoning Agentic Workflow

Reference 39

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.822404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.822404Z digest=sha256:9b2285630c045ed73bdb4bb4d77a45ac3bfa31e1adb26804b6c19f9d5b474448

Observation 3c851925-2376-4bd7-9795-d9bd6eed143d · outbound

This paper cites Mmedagent: Learning to use medical tools with multi-modal agent.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Mmedagent: Learning to use medical tools with multi-modal agent

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.826778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.826778Z digest=sha256:571198590b9404e9fc6e79e229d588933bd86a3d2b883658ae25845a80e70a80

Observation 34d450f4-b668-41e3-8946-f4ea977c0be5 · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.831234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.831234Z digest=sha256:159bd18fe9ab05b0e972f6c59f4d7cc9888daaac0ea104473e23840a351e85c9

Observation 71b7a39e-7caf-4228-9a4b-f294a9286220 · outbound

This paper cites Bimedix2: Bio-medical expert lmm for diverse medical modalities.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Bimedix2: Bio-medical expert lmm for diverse medical modalities

Reference 42

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.835505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.835505Z digest=sha256:63d54aa2dc887dfb67c89aa4413a9a79957bc23dcab08eda086eb410caed9cc9

Observation 1999e115-1bfd-43c6-9356-644332f9ddce · outbound

This paper cites Healthgpt: A medical large vision-language model for unifying comprehension and generation via heterogeneous knowledge adaptation.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Healthgpt: A medical large vision-language model for unifying comprehension and generation via heterogeneous knowledge adaptation

Reference 43

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.839618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.839618Z digest=sha256:58e9d96e289efa1ec2c3b701ed7f97eacef31f8c040ced96050f4874d81c8713

Observation 7f69231d-8d33-486d-929c-3cd549cabdd6 · outbound

This paper cites GSCo: Towards Generalizable AI in Medicine via Generalist-Specialist Collaboration.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding GSCo: Towards Generalizable AI in Medicine via Generalist-Specialist Collaboration

Reference 44

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.844422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.844422Z digest=sha256:bcdba786f04f6cb1392b92683efb67eafcbe686fef8aab52445df9c5abebea5a

Observation 682d512d-c8a4-41c4-bb8d-8f2a23dd6e79 · outbound

This paper cites an unresolved cited work.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.849030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.849030Z digest=sha256:1d0c616004a41e5d58fd21f3749ec7246b395fb197e1f5deb7d0d028bd508052

Observation 0ced1486-3e33-439b-864f-1bb054377bfa · outbound

This paper cites 3d-rad: A comprehensive 3d radiology med-vqa dataset with multi-temporal analysis and diverse diagnostic tasks.arXiv preprint arXiv:2506.11147, 2025.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding 3d-rad: A comprehensive 3d radiology med-vqa dataset with multi-temporal analysis and diverse diagnostic tasks.arXiv preprint arXiv:2506.11147, 2025

Reference 47

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.864324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.864324Z digest=sha256:36e8bfd0395c7aace4469654526239239cfc9c8591dde83de9979f091843bfb8

Observation f19110b7-9d69-4be2-ba0c-6277f055a8b5 · outbound

This paper cites Amos: A large-scale abdominal multi-organ benchmark for versatile medical image segmentation.Advances in neural information processing systems, 35:36722–36732, 2022.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Amos: A large-scale abdominal multi-organ benchmark for versatile medical image segmentation.Advances in neural information processing systems, 35:36722–36732, 2022

Reference 48

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.869180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.869180Z digest=sha256:974f4ca3d9425f2b9a1183920e3a947bfcaddb2a38f3074435a4c656a3d52a2f

Observation 5e5af2aa-0fb2-443a-928f-7d2d087eacd6 · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Bleu: a method for automatic evaluation of machine translation

Reference 49

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.874150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.874150Z digest=sha256:d8954798dececa256db8dd9e6028e2cdc9573f1895549ded1f1d80d80bdcfc56

Observation 4304564b-fd87-4f6d-9dba-5e209839f70c · outbound

This paper cites RadGraph-XL: A large-scale expert-annotated dataset for entity and relation extraction from radiology reports.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding RadGraph-XL: A large-scale expert-annotated dataset for entity and relation extraction from radiology reports

Reference 50

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.878777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.878777Z digest=sha256:df505cdcdf446057dd5cefe1c9125a0f7b69ea19b86511074c3536140fe246e0

Observation 1543372b-c5d7-4443-a577-3f5076039b8c · outbound

This paper cites Switch transformers: Scaling to trillion parameter models with simple and efficient sparsity.Journal of Machine Learning Research, 23(120):1–39, 2022.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Switch transformers: Scaling to trillion parameter models with simple and efficient sparsity.Journal of Machine Learning Research, 23(120):1–39, 2022

Reference 51

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.883448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.883448Z digest=sha256:45f9bb428f87d9397430fedf8ec28d2777ce59c8f166ce607240150e3a3027f6

Observation bd93326b-7a8b-4b29-a647-693440de4ffc · outbound

This paper cites Position: Compositional generative modeling: A single model is not all you need.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Position: Compositional generative modeling: A single model is not all you need

Reference 52

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.887609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.887609Z digest=sha256:c80b2afe81b232a67f4f4658da906d740fa3cca104b76cb95a3fd462f8af32fd

Observation 2c418c22-aee1-4e89-aa7e-a474ae753f87 · outbound

This paper cites A convnet for the 2020s.Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2022.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding A convnet for the 2020s.Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2022

Reference 53

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.892067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.892067Z digest=sha256:e81a304a47013c85fe2ee9b343f361a90797ec349982b2921b708804a39053ac

Observation f61d1b31-4371-4300-85f9-8f0dc26897f6 · outbound

This paper cites an unresolved cited work.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.896289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.896289Z digest=sha256:b3ad51975ac524fd9e55b29b203ed6e73708bb5d170e65bed7d42c1c8e32fd7d

Observation 7be93ead-d69b-4dfd-ac23-e88a0f103224 · outbound

This paper cites Exploring scalable medical image encoders beyond text supervision.Nature Machine Intelligence, 7(1):119–130, 2025.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Exploring scalable medical image encoders beyond text supervision.Nature Machine Intelligence, 7(1):119–130, 2025

Reference 55

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.900617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.900617Z digest=sha256:37c0345f55f40744d4ca20a3eadd25b789d418592082dad3b71fcfa3cd30a464

Observation a4f6110b-4507-4485-9621-b0c65035518c · outbound

This paper cites A multimodal biomedical foundation model trained from fifteen million image–text pairs.Nejm Ai, 2(1):AIoa2400640, 2025.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding A multimodal biomedical foundation model trained from fifteen million image–text pairs.Nejm Ai, 2(1):AIoa2400640, 2025

Reference 56

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.905366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.905366Z digest=sha256:f56ff64dc4853b15161035250cfcbdd536680fbf61f5c9c400613e2450fcdbe2

Observation 4df93691-7ce3-48aa-9912-133539a07a24 · outbound

This paper cites Visionzip: Longer is better but not necessary in vision language models.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Visionzip: Longer is better but not necessary in vision language models

Reference 57

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.909560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.909560Z digest=sha256:5eb063922ced3f49aef904cd6aadb75113cc6afa11678e33c30960d78410fbed

Observation 38502715-f4d4-4147-a183-59a62a097b6b · outbound

This paper cites Sparsevlm: Visual token sparsification for efficient vision-language model inference.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Sparsevlm: Visual token sparsification for efficient vision-language model inference

Reference 58

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.913652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.913652Z digest=sha256:9b290265a97f8a69d5f77a43636c504e61eeca36740cbba8b9cfbae8bbb9d663

Observation cfb0ae7e-81b0-4a1f-baf9-03bb01564451 · outbound

This paper cites Attention is all you need.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Attention is all you need

Reference 59

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.918023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.918023Z digest=sha256:ea7de611c829d12ad0cd7b4af2866313e7060d5967952f2d156b950fb07fc65e

Observation 31c4033e-3bc0-4396-b0fe-ac91b5baedfc · outbound

This paper cites Deepstack: Deeply stacking visual tokens is surprisingly simple and effective for lmms.Advances in Neural Information Processing Systems, 37:23464–23487, 2024.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Deepstack: Deeply stacking visual tokens is surprisingly simple and effective for lmms.Advances in Neural Information Processing Systems, 37:23464–23487, 2024

Reference 60

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.922247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.922247Z digest=sha256:981062249b1b57af4558bea6715f44c23c1dfac9af5175ac67abad2e887ecf50

Observation b4cbec94-23d4-49eb-a553-a513b6a4887f · outbound

This paper cites Perception Encoder: The best visual embeddings are not at the output of the network.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Perception Encoder: The best visual embeddings are not at the output of the network

Reference 61

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.926532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.926532Z digest=sha256:f4c080ae591697455a5c7263ae81b79fba57c84a3ecbe6547a079e4f2e5d7b1f

Observation 3e70cf8e-e305-4cc9-89c9-212a3477561f · outbound

This paper cites Generalized radiograph representation learning via cross-supervision between images and free-text radiology reports.Nature Machine Intelligence, 4(1):32–40, 2022.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Generalized radiograph representation learning via cross-supervision between images and free-text radiology reports.Nature Machine Intelligence, 4(1):32–40, 2022

Reference 62

Resolution
verified exact
doi, observed 2026-07-31T06:20:56.429503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-07-31T06:20:29.931059Z digest=sha256:9045971571935d8cce2c3ae6004c1845dfd72756c69dae4fd8877ee886f395b2

Observation 64387d1d-4d25-4d61-a423-b2c7d9693099 · outbound

This paper cites Huatuogpt, towards taming language model to be a doctor.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Huatuogpt, towards taming language model to be a doctor

Reference 63

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.935881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.935881Z digest=sha256:501b7161fc3bf4550069227cab6ee977b2ba67514e32e4b8d3b88508bf936122

Observation a4a7c094-d03c-48f6-8ade-236f3518099e · outbound

This paper cites State of what art? a call for multi-prompt llm evaluation.Transactions of the Association for Computational Linguistics, 12: 933–949, 2024.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding State of what art? a call for multi-prompt llm evaluation.Transactions of the Association for Computational Linguistics, 12: 933–949, 2024

Reference 64

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.940553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.940553Z digest=sha256:ea83990eb17599f712422344d01c01f95694cb15daacbf83964909b2a0df6551

Observation 06246414-fd9f-43a5-8f22-9ea711e55c07 · outbound

This paper cites Surveillance, epidemiology, and end results (seer) program (www.seer.cancer.gov), 2025.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Surveillance, epidemiology, and end results (seer) program (www.seer.cancer.gov), 2025

Reference 65

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.944785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.944785Z digest=sha256:ee14682b779859913468dc102e94d47410865a43be1d0d3c786d1871067eea5a

Observation e0874377-4866-4051-8d9f-6a3603acbcb3 · outbound

This paper cites Look again, think slowly: Enhancing visual reflection in vision-language models.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Look again, think slowly: Enhancing visual reflection in vision-language models

Reference 66

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.949392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.949392Z digest=sha256:38ad514bcc64515516b792c1034aee28b3ec78939a02a0adb564727ef59cf179

Observation 9b759028-84a6-47a2-ba64-7c60aba90e4a · outbound

This paper cites M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 67

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.953458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.953458Z digest=sha256:12da7c6594a3361eb2c75b98b960b3604e0a381d034f6b600fd048f4422d7bf5

Observation 5e511336-50fd-469f-8431-52dae50e0fdf · outbound

This paper cites Merlin: a computed tomography vision–language foundation model and dataset.Nature, pages 1–11, 2026.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Merlin: a computed tomography vision–language foundation model and dataset.Nature, pages 1–11, 2026

Reference 68

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.957764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.957764Z digest=sha256:b258360d2b091b0de88b50a4565a541d2fab6bb8013c223382a59af4c665acf8

Observation 84157aed-d860-4493-a2e7-510412f936f5 · outbound

This paper cites Inspect: A multimodal dataset for patient outcome prediction of pulmonary embolisms.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Inspect: A multimodal dataset for patient outcome prediction of pulmonary embolisms

Reference 69

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.962001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.962001Z digest=sha256:2bdc78d8135e41e252219ef2fe850dfe6be242bf830e9dedc623c300fce9f69a

Observation 04772d10-3435-4fc5-8d68-98c7db1754e6 · outbound

This paper cites Vista3d: A unified segmentation foundation model for 3d medical imaging.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Vista3d: A unified segmentation foundation model for 3d medical imaging

Reference 70

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.966447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.966447Z digest=sha256:c7fba4004a97d2c22b17d33cf69613fe385d1c8582aa11ecb9034caffcf24255

Observation c365480e-31b9-4f43-9d62-45c4c010eace · outbound

This paper cites Totalsegmentator: robust segmentation of 104 anatomic structures in ct images.Radiology: Artificial Intelligence, 5(5):e230024, 2023.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Totalsegmentator: robust segmentation of 104 anatomic structures in ct images.Radiology: Artificial Intelligence, 5(5):e230024, 2023

Reference 71

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.970641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.970641Z digest=sha256:98b2b2b1e693d378a9d12fd7b778c60ca6f139472452f245fb2af6beeda3a07e

Observation 2c1b3231-9293-4041-ac81-a230c24044ba · outbound

This paper cites DMQR-RAG: Diverse Multi-Query Rewriting for RAG.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding DMQR-RAG: Diverse Multi-Query Rewriting for RAG

Reference 72

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.975213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.975213Z digest=sha256:ca32a52e1c6d69ab1a311ede15b6c007d9c930307fad78c659795acafd456e4f

Observation e46f354f-7c63-43b2-b8b3-82c6b1050814 · outbound

This paper cites UMLS knowledge sources, release 2024aa, 2024.http://www.nlm.nih.gov/ research/umls/licensedcontent/umlsknowledgesources.html(accessed 15 July 2024).

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding UMLS knowledge sources, release 2024aa, 2024.http://www.nlm.nih.gov/ research/umls/licensedcontent/umlsknowledgesources.html(accessed 15 July 2024)

Reference 73

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.979544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.979544Z digest=sha256:ecaa8accc20a0b7e326d2edf28de0f0c77003ea296d6089653c38fb515989414

Observation 03895d10-25c3-4a81-8d4e-771d8ab7785a · outbound

This paper cites Building a knowledge graph to enable precision medicine.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Building a knowledge graph to enable precision medicine

Reference 74

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.983879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.983879Z digest=sha256:64060ee8ca62ad0eafd8744243cf834d1bde19709b531dabe8a99e97011cfad6

Observation 5fb49e35-285a-4a91-a43c-3b8d4ba5cb26 · outbound

This paper cites Sayers, Evan E.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Sayers, Evan E

Reference 75

Resolution
verified exact
doi, observed 2026-07-31T06:20:56.390675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-07-31T06:20:29.988047Z digest=sha256:4d82765ba884a25ef66a0ec77b6b56707c81de924bd65900e172410bd0850b6d

Observation 384e0de7-46d0-4a79-8e44-520119b13afa · outbound

This paper cites StatPearls Publishing, Treasure Island, FL, 2026.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding StatPearls Publishing, Treasure Island, FL, 2026

Reference 76

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.992608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.992608Z digest=sha256:c727e29d62c897156752a424baf69845310d8ff1744d422cb9e18dea3c22b29a

Observation abe30840-efc1-49ed-bbe7-c2d0b02a7aa8 · outbound

This paper cites What disease does this patient have? a large-scale open domain question answering dataset from medical exams.Applied Sciences, 11 (14):6421, 2021.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding What disease does this patient have? a large-scale open domain question answering dataset from medical exams.Applied Sciences, 11 (14):6421, 2021

Reference 77

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.997011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.997011Z digest=sha256:18f66ad228914627364bc34f50954fa1a1124deb64580878591276e21d1f422d

Observation a656dbeb-ed18-4bab-9941-bd063f1b57b4 · outbound

This paper cites Xiong, Q.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Xiong, Q

Reference 78

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.001898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.001898Z digest=sha256:3bddaf1371dca3022a33bc971f03d305857bd6ba7bb4dc8b74d485f892c38841

Observation 3708dfa9-d6a4-4317-831a-718fb0560e1f · outbound

This paper cites Multi-modal ai for opportunistic screening, staging and progression risk stratification of steatotic liver disease.Nature Communications, 17(1):1562, 2026.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Multi-modal ai for opportunistic screening, staging and progression risk stratification of steatotic liver disease.Nature Communications, 17(1):1562, 2026

Reference 79

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.006389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.006389Z digest=sha256:5043fa8f1c626643e8fab6e89c423944d12249c2268eff57b18d9316fc7f31d0

Observation 6042fd0f-a957-43bd-8664-2c3c213c0fd4 · outbound

This paper cites Effective lymph nodes detection in ct scans using location debiased query selection and contrastive query representation in transformer.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Effective lymph nodes detection in ct scans using location debiased query selection and contrastive query representation in transformer

Reference 80

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.011190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.011190Z digest=sha256:63b3c224bfe1f52fb7d19c296d38a6607060aff42ade666c0d936e4a197b1958

Observation 159ffc65-f324-454c-a538-0d1d3e18b61e · outbound

This paper cites On the limits of cross-domain generalization in automated x-ray prediction.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding On the limits of cross-domain generalization in automated x-ray prediction

Reference 81

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.016358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.016358Z digest=sha256:21e442d308b40bae57b19badd38078fcd896aee2f76e41c7892a5bf3c191929f

Observation e919322b-8d10-48ed-996e-b91ea5d17d1c · outbound

This paper cites Kalra, and Pingkun Yan.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Kalra, and Pingkun Yan

Reference 82

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.020989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.020989Z digest=sha256:8172267d584a2ad8861582d6898e5b6bc967711d340eaa0a01dad5db69686a32

Observation a093b395-a272-4cb4-8da1-a57e4c359fca · outbound

This paper cites A Vision-Language Foundation Model to Enhance Efficiency of Chest X-ray Interpretation.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding A Vision-Language Foundation Model to Enhance Efficiency of Chest X-ray Interpretation

Reference 83

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.025041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.025041Z digest=sha256:a61e06abeb7f353a762f26a3d05f3ea798e18d3f791ea13775891b9d35f36e8f

Observation b5b38be1-7521-4b5c-adfc-85ff63c74427 · outbound

This paper cites Pmc-clip: Contrastive language-image pre-training using biomedical documents.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Pmc-clip: Contrastive language-image pre-training using biomedical documents

Reference 84

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.029504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.029504Z digest=sha256:53ee5c821b8129581cea610a73e54380e79e71419f92ba8bf02ee0d2dfae7cf5

Observation 58a121f4-155e-4ada-b1ab-f3173442f15f · outbound

This paper cites Friedrich.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Friedrich

Reference 85

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.033853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.033853Z digest=sha256:b7d822961d65a2e6c2b497c92925a6a5b0c21744c78406f3b8a5f9e2b297614d

Observation c1add6ae-b71d-4d34-9289-1c6a762659e1 · outbound

This paper cites Seco de Herrera, et al.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Seco de Herrera, et al

Reference 86

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.038050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.038050Z digest=sha256:59acb304d8719fbbd0d1932f04850fca1059b8d976623bebaa2975635c7ac29b

Observation 105fe18f-caf5-4881-ae5d-36a9392a8f7f · outbound

This paper cites Towards injecting medical visual knowledge into multimodal llms at scale.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Towards injecting medical visual knowledge into multimodal llms at scale

Reference 87

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.042211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.042211Z digest=sha256:23ae9c0b99da0f1d52c55891007d3525614ce13ffffeabc423e6d82a8c1c1ae9

Observation 59a2440a-d3e0-4a68-9668-59930d2c371b · outbound

This paper cites Mimic-cxr, a de-identified publicly available database of chest radiographs with free-text reports.Scientific data, 6(1):317, 2019.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Mimic-cxr, a de-identified publicly available database of chest radiographs with free-text reports.Scientific data, 6(1):317, 2019

Reference 88

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.046871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.046871Z digest=sha256:91ba0a4315b42ed3abef2b4b9073e658b09d1733edf57cfa16fb06bd427d028e

Observation ccb2ac9a-77c4-4063-8d35-e192150b92b9 · outbound

This paper cites Medicat: A dataset of medical images, captions, and textual references.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Medicat: A dataset of medical images, captions, and textual references

Reference 90

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.056055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.056055Z digest=sha256:140a655d8beba4010ee2f4e9eab9a857ab33d88f220867859d1a58754709105f

Observation d015b708-69f2-4042-8c6a-3fc7a8411b2b · outbound

This paper cites an unresolved cited work.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Unresolved cited work

Reference 91

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.060699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.060699Z digest=sha256:e352bc7649ef51a8935508d42207091b942be3a04c3ee7f88c3a7c72b8d7f930

Observation 07e4553a-9360-42d5-95fa-27317f889765 · outbound

This paper cites an unresolved cited work.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Unresolved cited work

Reference 92

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.065156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.065156Z digest=sha256:53aaa427fc2dbb4194e79a0014d7461e72b9e662d0b6494e182d6407b7622c9c

Observation 63c57ade-5f9c-4fd5-affb-f688d969e969 · outbound

This paper cites Improved baselines with visual instruction tuning.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Improved baselines with visual instruction tuning

Reference 93

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.069484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.069484Z digest=sha256:7e470dd1aff2f5e05816be358b414a0ba1a403ef60c984a1e1ecacbafeaefe2c

Observation 1e5965f2-1145-42a5-9ee1-e80072f3dab6 · outbound

This paper cites Molmo and pixmo: Open weights and open data for state-of-the-art vision-language models.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Molmo and pixmo: Open weights and open data for state-of-the-art vision-language models

Reference 94

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.073880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.073880Z digest=sha256:7924ae96d577ac86479acf616cfb9a7e4673c20e748cc33b63d73200047cabba

Observation 01c68a65-9112-43cf-a98d-25e9db79fa09 · outbound

This paper cites Llava-med: Training a large language-and-vision assistant for biomedicine in one day.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Llava-med: Training a large language-and-vision assistant for biomedicine in one day

Reference 95

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.078185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.078185Z digest=sha256:962a1656ee658e5e13448e2554ee55549630ba07904ca347f0263fee251cc3b7

Observation b64b981d-a68c-4836-8ffe-03af00a43128 · outbound

This paper cites Quilt- llava: Visual instruction tuning by extracting localized narratives from open-source histopathology videos.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Quilt- llava: Visual instruction tuning by extracting localized narratives from open-source histopathology videos

Reference 96

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.082969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.082969Z digest=sha256:b9b722b1150bdfa98d3afcd1416f60365bd10959d08615d57bbf3a44da2afade

Observation f1f8acc9-da48-4bef-bd68-a6477220783d · outbound

This paper cites Hicks, Vajira Thambawita, Pål Halvorsen, and Michael A.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Hicks, Vajira Thambawita, Pål Halvorsen, and Michael A

Reference 97

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.087647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.087647Z digest=sha256:ced41895719888acb92fe444adb177b58c3a7fa5a2a2c50a076c0e471c4a0833

Observation 4f143619-ee0e-4ef5-96f7-5b29a2406603 · outbound

This paper cites MIMIC-Ext-MIMIC-CXR-VQA: A complex, diverse, and large-scale visual question answering dataset for chest x-ray images, 2024.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding MIMIC-Ext-MIMIC-CXR-VQA: A complex, diverse, and large-scale visual question answering dataset for chest x-ray images, 2024

Reference 98

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.091803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.091803Z digest=sha256:8f8c67826b6af40dc56663f2a36386296afbb4f1684069a2b15adbb3ee8b918f

Observation 9c99380f-b213-4281-ba1c-f98dce2d3856 · outbound

This paper cites Towards visual question answering on pathology images.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Towards visual question answering on pathology images

Reference 99

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.096393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.096393Z digest=sha256:121c65293610a4ba0faaea1f8a1b39b975595af73ed535c70989213d9d0d6736

Observation ae49aa77-1870-44e2-8298-90c3dbdc2ca8 · outbound

This paper cites Development of a large-scale medical visual question-answering dataset.Communications Medicine, 4(1):23, 2024.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Development of a large-scale medical visual question-answering dataset.Communications Medicine, 4(1):23, 2024

Reference 100

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.101128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.101128Z digest=sha256:6d47a5e015ad2a5ee4b1d251894d60ce4a5222269c824814fee9788f494db8a2

Observation 715c13e9-60ac-49c9-b201-4f5b5e915322 · outbound

This paper cites Hasan, Vivek V.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding Hasan, Vivek V

Reference 101

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.105764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.105764Z digest=sha256:c31601fd573563268e6ef3caa4b33f427fb1c1702811754d5b7df06d77702af3

Observation 2be0753d-faab-421f-a9a6-656381fbae80 · outbound

This paper cites A dataset of clinically generated visual questions and answers about radiology images.Scientific data, 5(1):1–10, 2018.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding A dataset of clinically generated visual questions and answers about radiology images.Scientific data, 5(1):1–10, 2018

Reference 102

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.110246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.110246Z digest=sha256:1ef946980d35c0d6348c05581f2ccd59bb7fc35f4b06c6003843ef7de66f8db8

Observation f5b30254-a857-401d-8ab4-7216c357b7dc · outbound

This paper cites GMAI-VL-R1: Harnessing Reinforcement Learning for Multimodal Medical Reasoning.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding GMAI-VL-R1: Harnessing Reinforcement Learning for Multimodal Medical Reasoning

Reference 103

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:30.115352Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:30.115352Z digest=sha256:af4d310c86bc25455f6ac516fa8ee7ffd5ee758752a0b650932e69372e3361f2

Pith citing papers

No inbound Pith citation observations are available.