Pith. sign in

Paper Citation Record · LEDGER

Medical Large Vision Language Models with Multi-Image Visual Ability

As of 8 August 2026, this Paper Citation Record lists 26 of 26 outbound references and 0 inbound Pith citation observations for arXiv:2505.19031.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.19031 v1

Coverage vector

measured 26 of 26 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:24:13.175284Z

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

26 of 26 outbound references displayed

  • verified exact0
  • verified fuzzy13
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c7a72a24-b16b-4d8f-ae7a-61302ed27a6f · outbound

This paper cites Cancers 15(2), 545 (2023).

Medical Large Vision Language Models with Multi-Image Visual Ability Cancers 15(2), 545 (2023)

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:24:15.883639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:24:10.277537Z digest=sha256:324f1398ec1534bb653855bed4f83444c62648ab6f4335cfa792ad29e17041cb

Observation 5d681346-a2ad-4700-910d-eed18fd9a76b · outbound

This paper cites M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models.

Medical Large Vision Language Models with Multi-Image Visual Ability M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:10.371955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:24:10.371955Z digest=sha256:8e825957ac32ae9395633ce60577d84999cb40da775d83b359888957610b0b25

Observation 025a1d10-0b39-41c9-9cc0-e42556521ae7 · outbound

This paper cites an unresolved cited work.

Medical Large Vision Language Models with Multi-Image Visual Ability Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:24:15.703492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:24:10.493787Z digest=sha256:55946b571cf1513446c078391ea854378f75cff4f0e9ee6391239acd1fd037db

Observation ffdc48f3-5f7b-4508-be56-2f31a66e7b3a · outbound

This paper cites Biological psychiatry: cognitive neuroscience and neuroimaging1(3), 230–244 (2016).

Medical Large Vision Language Models with Multi-Image Visual Ability Biological psychiatry: cognitive neuroscience and neuroimaging1(3), 230–244 (2016)

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:24:15.526172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:24:10.632153Z digest=sha256:e7efc36dc7fcc7ef913f57fd8bdfe4bcf40fcf8c420f7f2b193b06869a234c4e

Observation 1c3fd77f-f21f-4e5e-b8fd-e07fd1824b3d · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition.

Medical Large Vision Language Models with Multi-Image Visual Ability In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:10.773576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:24:10.773576Z digest=sha256:b94c37f0ce47f9cb87a64e19687e30887a1364af53fb40de3f2010ca0434291c

Observation aee9ebf1-8ec0-4a20-9825-d405e1a28756 · outbound

This paper cites GPT-4o System Card.

Medical Large Vision Language Models with Multi-Image Visual Ability GPT-4o System Card

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:10.907291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:24:10.907291Z digest=sha256:509f48c3485d14d339ab30a89daa01bf367a3d96a8a2578c0315610ccc17c826

Observation 2c342d75-a7d0-49f1-889d-248d13e2b1b7 · outbound

This paper cites Radiology: Artificial Intelligence5(1), e220047 (2023).

Medical Large Vision Language Models with Multi-Image Visual Ability Radiology: Artificial Intelligence5(1), e220047 (2023)

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:24:15.266171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:24:11.049661Z digest=sha256:a919fe3d9dbe25156f1093669b2cc24a301c69ac9a648e24ade04eb10a4088e8

Observation f97a701f-c6ff-4342-8d1f-dd7be612bf25 · outbound

This paper cites Transactions on Machine Learning Research 2024 (2024), https://openreview.net/forum?id=skLtdUVaJa.

Medical Large Vision Language Models with Multi-Image Visual Ability Transactions on Machine Learning Research 2024 (2024), https://openreview.net/forum?id=skLtdUVaJa

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:24:15.099908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:24:11.166673Z digest=sha256:937941bb5cfcdcaf06bec8bee63a547471cfc945dcc29c40539831b28f070025

Observation 25d5fc71-13dd-4f49-be02-eaed059b0ca6 · outbound

This paper cites Archives of neurology66(10), 1254–1259 (2009).

Medical Large Vision Language Models with Multi-Image Visual Ability Archives of neurology66(10), 1254–1259 (2009)

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:24:14.898592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:24:11.284080Z digest=sha256:c007682d319cbca1c60d09592aa96827ecd6ac8aaa879f721fd7c6d17c4a83b1

Observation 306052f8-c750-4d47-a850-b86dc2893e7f · outbound

This paper cites Nature Machine Intelligence3(4), 288–298 (2021).

Medical Large Vision Language Models with Multi-Image Visual Ability Nature Machine Intelligence3(4), 288–298 (2021)

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:24:14.709365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:24:11.370365Z digest=sha256:efb4d328c8a37749f68b52519d94af68a51cf3154b14983ee6cc1ab7af3cc907

Observation 17feca68-e54e-4282-a59d-d5b7bef384f7 · outbound

This paper cites Scientific data 5(1), 1–10 (2018).

Medical Large Vision Language Models with Multi-Image Visual Ability Scientific data 5(1), 1–10 (2018)

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:11.501957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:24:11.501957Z digest=sha256:4843d2a335ff530274ffd74c3b77b59f450b05ffbf3c979f0d74ec5ffe348f22

Observation d05d9b3c-decd-4bea-bc2b-3aa238472727 · outbound

This paper cites Advances in Neural Information Processing Systems36 (2024).

Medical Large Vision Language Models with Multi-Image Visual Ability Advances in Neural Information Processing Systems36 (2024)

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:11.611023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:24:11.611023Z digest=sha256:a9d4a3ab02467aa392c663ad6c2bf48878c7838910313fcbddb76001dd433f17

Observation 86b06bad-a99f-4dfb-8562-aa8fbd802b5e · outbound

This paper cites In: Benchmarking, Measur- ing, and Optimizing: Third BenchCouncil International Symposium, Bench 2020, 10 X.

Medical Large Vision Language Models with Multi-Image Visual Ability In: Benchmarking, Measur- ing, and Optimizing: Third BenchCouncil International Symposium, Bench 2020, 10 X

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:24:14.491820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:24:11.730335Z digest=sha256:1dbc3420a8a63f14e6e9a3838d05fde349e31fe04f7e13e5832f8fee7ebb4e93

Observation 820bac15-5f05-4bf5-91b3-f4a328a81309 · outbound

This paper cites Advances in neural information processing systems36 (2024).

Medical Large Vision Language Models with Multi-Image Visual Ability Advances in neural information processing systems36 (2024)

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:11.870901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:24:11.870901Z digest=sha256:673e15edc0a9029cb01e18f12df3700064c35f3ee4a54467722658fd5407322c

Observation 09a076c8-e187-4eee-a58b-a509c02c1004 · outbound

This paper cites MMDU: A Multi-Turn Multi-Image Dialog Understanding Benchmark and Instruction-Tuning Dataset for LVLMs.

Medical Large Vision Language Models with Multi-Image Visual Ability MMDU: A Multi-Turn Multi-Image Dialog Understanding Benchmark and Instruction-Tuning Dataset for LVLMs

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:11.997396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:24:11.997396Z digest=sha256:5b3119d89a92ceed74145af34fb45946f3fe9f9464b91c5f58e481650aa57bf6

Observation 96e66d03-94e0-4e53-aca7-5f26a7acac5c · outbound

This paper cites DeepSeek-VL: Towards Real-World Vision-Language Understanding.

Medical Large Vision Language Models with Multi-Image Visual Ability DeepSeek-VL: Towards Real-World Vision-Language Understanding

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:12.107542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:24:12.107542Z digest=sha256:264b4494fdedf94d66d9b77325e1c070a4f2a09e5b7ea5a0467666f871b13316

Observation d293c135-9ec3-4d29-b091-a03a1a1b1634 · outbound

This paper cites MMIU: Multimodal Multi-image Understanding for Evaluating Large Vision-Language Models.

Medical Large Vision Language Models with Multi-Image Visual Ability MMIU: Multimodal Multi-image Understanding for Evaluating Large Vision-Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:12.189818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:24:12.189818Z digest=sha256:4bef1a2dd067357cc8317b46441a5f8d4f6dac2032419ee5c9fb8f97e3b6ace8

Observation 6e302729-daf7-45f2-841b-99a6399dfc97 · outbound

This paper cites In: Machine Learning for Health (ML4H).

Medical Large Vision Language Models with Multi-Image Visual Ability In: Machine Learning for Health (ML4H)

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:12.319913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:24:12.319913Z digest=sha256:604a7b163d050305375a4db69c5e7e8709eecd58be59c600eeb391219cff2023

Observation 2d5b2357-880a-4743-b83e-447329e2b090 · outbound

This paper cites Nature communications 13(1), 4566 (2022).

Medical Large Vision Language Models with Multi-Image Visual Ability Nature communications 13(1), 4566 (2022)

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:24:14.346533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:24:12.441876Z digest=sha256:35a4963dfb828c84da5870c9570bc6b44f89e9cb2cad31f856f3013a184e329d

Observation 7bb5efd9-b31c-4b48-962a-b9154d48d5fc · outbound

This paper cites BMC Pulmonary Medicine 20, 1–9 (2020).

Medical Large Vision Language Models with Multi-Image Visual Ability BMC Pulmonary Medicine 20, 1–9 (2020)

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:24:14.169298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:24:12.574854Z digest=sha256:1410fbb99f5e1f6e5714b115d8aed7542a55fe34fae793ef174fef6be1dd14dd

Observation 8d983b9b-b40e-4819-b24b-941264ec9c48 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition.

Medical Large Vision Language Models with Multi-Image Visual Ability In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:24:13.986290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:24:12.704747Z digest=sha256:3beb47e564509e0887043d8aba1bd236f2a0947581879773a068eed45eefc0b2

Observation 30b2d87d-8990-4fea-9def-7b0ea24c755b · outbound

This paper cites In: The Thirty-eighth An- nual Conference on Neural Information Processing Systems.

Medical Large Vision Language Models with Multi-Image Visual Ability In: The Thirty-eighth An- nual Conference on Neural Information Processing Systems

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:24:13.809310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:24:12.831889Z digest=sha256:ac473697cf4ec4a199a7254dfffe04d7e6d85c1ba75cc221492b87f9bfd6c2a5

Observation 6d4c5e61-8b59-41c8-bc7e-73dcb6d278d0 · outbound

This paper cites Scientific data9(1), 768 (2022).

Medical Large Vision Language Models with Multi-Image Visual Ability Scientific data9(1), 768 (2022)

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:24:13.609934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:24:12.918590Z digest=sha256:af53301518bd38398fc0f096d449d3c937e1bbf9770ce6a2c6c261845fd20321

Observation ebd5d5bf-00b3-49cd-9d89-424a19521f26 · outbound

This paper cites Towards Generalist Foundation Model for Radiology by Leveraging Web-scale 2D&3D Medical Data.

Medical Large Vision Language Models with Multi-Image Visual Ability Towards Generalist Foundation Model for Radiology by Leveraging Web-scale 2D&3D Medical Data

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:13.004322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:24:13.004322Z digest=sha256:ccd3c5758f08dc5bf8a3bc7210db51fb262c94da6890c36d48ab6c652becde1f

Observation 074a5f24-eda1-4e78-b2eb-e653be5b2fe9 · outbound

This paper cites In: Proceedings of the 41st International Conference on Machine Learning.

Medical Large Vision Language Models with Multi-Image Visual Ability In: Proceedings of the 41st International Conference on Machine Learning

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:24:13.461879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:24:13.073032Z digest=sha256:eb1bca84ba780dcdf561e0fd2531333bc2c087b0fbc1c3af85eccc492d6cdded

Observation 337ab775-8800-4d24-b14b-6e4fc401f9d2 · outbound

This paper cites PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering.

Medical Large Vision Language Models with Multi-Image Visual Ability PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:13.175284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:24:13.175284Z digest=sha256:d1d06bdecdf260b1416cac4f3a6ce88d2227d81215469dea4c85aec2e486030d

Pith citing papers

No inbound Pith citation observations are available.