Pith. sign in

Paper Citation Record · LEDGER

M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 83 inbound Pith citation observations for arXiv:2404.00578.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.00578 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 83 of 83 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 83 of 83 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:43:30.669345Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

16
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f184404e-2b64-4587-83fb-ee63fc83351c · inbound

MAIRA-Seg: Enhancing Radiology Report Generation with Segmentation-Aware Multimodal Large Language Models cites this paper.

MAIRA-Seg: Enhancing Radiology Report Generation with Segmentation-Aware Multimodal Large Language Models M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T18:42:17.663360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:42:17.663360Z digest=sha256:a1d267cf549f73a7521053c7d941d4bb216687e67cc9b20771e48649caefba5a

Observation 0df6d565-e58a-4be5-a0b9-3d01336fa0a3 · inbound

Large Language Model with Region-guided Referring and Grounding for CT Report Generation cites this paper.

Large Language Model with Region-guided Referring and Grounding for CT Report Generation M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T14:15:55.972678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:15:55.972678Z digest=sha256:47d2e5c28b65f7617798477224608b6571da784f8a41f92ff110ad98622d4711

Observation 4266b35d-f31e-4f48-b8ca-ed32fb35b455 · inbound

MG-3D: Multi-Grained Knowledge-Enhanced 3D Medical Vision-Language Pre-training cites this paper.

MG-3D: Multi-Grained Knowledge-Enhanced 3D Medical Vision-Language Pre-training M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-11T20:21:00.016626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:21:00.016626Z digest=sha256:e9dad4f8050639f81f3028a5af68b3f1f7ca224edf0b202a43a4c8a933457651

Observation e14dedcb-cd1d-49fe-9acf-919d218b88b4 · inbound

From Specific-MLLMs to Omni-MLLMs: A Survey on MLLMs Aligned with Multi-modalities cites this paper.

From Specific-MLLMs to Omni-MLLMs: A Survey on MLLMs Aligned with Multi-modalities M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-11T14:43:41.768579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:43:41.768579Z digest=sha256:912aeffe50d892054949264f43fb157d94d9a43e864fa49327a27ec4e045201e

Observation bd1a39e1-ef5d-4104-a2bd-d47a6d4d4e11 · inbound

Read Like a Radiologist: Efficient Vision-Language Model for 3D Medical Imaging Interpretation cites this paper.

Read Like a Radiologist: Efficient Vision-Language Model for 3D Medical Imaging Interpretation M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-11T13:03:53.081260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:03:53.081260Z digest=sha256:a4823d899b39268a4a011da7234cc8e7ce23e4676ebb1146e7cd2a08c4a2337e

Observation 7cb83e7c-f1bd-4cc5-81ab-ee7c9947bc7a · inbound

ObjVariantEnsemble: Advancing Point Cloud LLM Evaluation in Challenging Scenes with Subtly Distinguished Objects cites this paper.

ObjVariantEnsemble: Advancing Point Cloud LLM Evaluation in Challenging Scenes with Subtly Distinguished Objects M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T11:54:12.097240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:54:12.097240Z digest=sha256:a01dddfc5257bdc650d984399606cc866bf4151a810c213ed9f676057eacce64

Observation 68ba6cf1-c3d4-4c02-9d24-eeddc0dce09d · inbound

RadGPT: Constructing 3D Image-Text Tumor Datasets cites this paper.

RadGPT: Constructing 3D Image-Text Tumor Datasets M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T21:31:28.402503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:31:28.402503Z digest=sha256:4681c26df736e40bd1b12f49e93f0bd7c06f43770a03df6c1b33783dd91c717f

Observation 9c44e8e2-4920-4a7b-867d-393199847498 · inbound

Generalization of Medical Large Language Models through Cross-Domain Weak Supervision cites this paper.

Generalization of Medical Large Language Models through Cross-Domain Weak Supervision M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-09T17:36:27.848303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:36:27.848303Z digest=sha256:e8393ae901215a5ddb23111125d052487023dfca785d83d8536f1ebd6a71d864

Observation b1d6dc9e-ab59-4121-b2a5-e9048a3241bf · inbound

From large language models to multimodal AI: A scoping review on the potential of generative AI in medicine cites this paper.

From large language models to multimodal AI: A scoping review on the potential of generative AI in medicine M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 117

Resolution
unresolved
no resolver link, observed 2026-08-07T22:13:31.441167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T22:13:31.441167Z digest=sha256:59223e2a60aedae70cee872b327aafe1f4d2c42034107e5e3f0bae1474db312a

Observation 302528a7-045a-4bef-b088-bff4db262e57 · inbound

UniCAD: Efficient and Extendable Architecture for Multi-Task Computer-Aided Diagnosis System cites this paper.

UniCAD: Efficient and Extendable Architecture for Multi-Task Computer-Aided Diagnosis System M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T21:43:30.669345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T21:43:30.669345Z digest=sha256:3eaf3b9db0de23a3d6f6e51ab265f9eb25e4ab1cabce9ccf4add5fc7f5d81054

Observation 7d745b69-d3cf-4a0f-9d48-b46624b1e27f · inbound

CT-Agent: A Multimodal-LLM Agent for 3D CT Radiology Question Answering cites this paper.

CT-Agent: A Multimodal-LLM Agent for 3D CT Radiology Question Answering M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:28.332256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:28.332256Z digest=sha256:e02137a261de81cc774063f9862f07ca54b904171546103fa6075877a1717666

Observation 80578d81-f810-4025-a65d-6a6b47436b5f · inbound

Are Vision Language Models Ready for Clinical Diagnosis? A 3D Medical Benchmark for Tumor-centric Visual Question Answering cites this paper.

Are Vision Language Models Ready for Clinical Diagnosis? A 3D Medical Benchmark for Tumor-centric Visual Question Answering M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:26:28.812740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:26:28.812740Z digest=sha256:09a8f0e26f7ffc7417d72eb44e268ff641c067dc021390d6ee9ee25f86d70be0

Observation 5d681346-a2ad-4700-910d-eed18fd9a76b · inbound

Medical Large Vision Language Models with Multi-Image Visual Ability cites this paper.

Medical Large Vision Language Models with Multi-Image Visual Ability M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:10.371955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:24:10.371955Z digest=sha256:9ab6d6a881c74c4f0b13a0b1ac162d4d74641d1f2fd7af21c05c9d805aea100d

Observation 7a954a7c-d5a8-4442-ba12-c67648decd15 · inbound

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding cites this paper.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.628602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.628602Z digest=sha256:a462bafc30ec843344cf4159b23f158baf71f8d4b84a613077c15c8ebb463926

Observation 8ce087ba-5932-4f0f-b9eb-e4043082d566 · inbound

CAPO: Reinforcing Consistent Reasoning in Medical Decision-Making cites this paper.

CAPO: Reinforcing Consistent Reasoning in Medical Decision-Making M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T20:10:52.036790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:10:52.036790Z digest=sha256:d71997dda4f518e1a8699a0727686ed67f2365a5f603aabe04cf2d9026859302

Observation 0b4e8ea3-63f2-4cec-96ff-fcc70fa5150d · inbound

Domain Specific Benchmarks for Evaluating Multimodal Large Language Models cites this paper.

Domain Specific Benchmarks for Evaluating Multimodal Large Language Models M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 185

Resolution
unresolved
no resolver link, observed 2026-08-07T00:39:42.343862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:39:42.343862Z digest=sha256:139c0d5dbeac1dfcf24ef2a698ffb4c0dfa04cdfda560160581c0aff00e28c24

Observation 470dc869-5bc2-4946-ac5e-dde649d80ed8 · inbound

MedErr-CT: A Visual Question Answering Benchmark for Identifying and Correcting Errors in CT Reports cites this paper.

MedErr-CT: A Visual Question Answering Benchmark for Identifying and Correcting Errors in CT Reports M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T23:10:45.988146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:10:45.988146Z digest=sha256:a518840fa2486567d90ffcd8fca61cf25ef69acdd347d4649de28d0606dbf78b

Observation 798aebe8-d62b-4fb3-a650-7e1e5543e530 · inbound

Recurrent Visual Feature Extraction and Stereo Attentions for CT Report Generation cites this paper.

Recurrent Visual Feature Extraction and Stereo Attentions for CT Report Generation M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:12.203692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:32:12.203692Z digest=sha256:b2f3334945597cd2622f6bd07317eace988b361f8535df6d717b14d17024cf7c

Observation 3ec38a6d-29f8-4b41-ae3e-40ac27ff54f7 · inbound

Region-Aware Multimodal Large Language Model via SlowFast Tokenization and Pseudo-Mask Guidance for 3D CT Report Generation cites this paper.

Region-Aware Multimodal Large Language Model via SlowFast Tokenization and Pseudo-Mask Guidance for 3D CT Report Generation M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:45.574493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:45.574493Z digest=sha256:cbd4c6be2dfdb2e2fae868adc8fa610bf3baa4340dfd46b8341324b67854338a

Observation 0ae81d1b-914f-41bd-894c-5e888dec489c · inbound

CRISP-SAM2: SAM2 with Cross-Modal Interaction and Semantic Prompting for Multi-Organ Segmentation cites this paper.

CRISP-SAM2: SAM2 with Cross-Modal Interaction and Semantic Prompting for Multi-Organ Segmentation M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T21:53:29.391871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:53:29.391871Z digest=sha256:82b348085fd7ef116328cbda98e8c441c1d2b816af90cd76a00cefc256003b17

Observation 85e4cc4c-d148-4b74-ae54-7312b2313783 · inbound

Prompt Mechanisms in Medical Imaging: A Comprehensive Survey cites this paper.

Prompt Mechanisms in Medical Imaging: A Comprehensive Survey M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 198

Resolution
unresolved
no resolver link, observed 2026-08-06T22:02:25.661416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:02:25.661416Z digest=sha256:19455c8325af2d875cd6a419873cecf462bfadea9a847bdbb3f5c49f0982e3c8

Observation 26534c7d-5826-4238-84f1-39e4d1a7b55a · inbound

MedGround-R1: Advancing Medical Image Grounding via Spatial-Semantic Rewarded Group Relative Policy Optimization cites this paper.

MedGround-R1: Advancing Medical Image Grounding via Spatial-Semantic Rewarded Group Relative Policy Optimization M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T21:04:54.657204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:04:54.657204Z digest=sha256:21b88ee00f860a638728cb83b056f45d3fcb10ecd69093beafcfff4e97080336

Observation e3ea3aeb-2a9b-4ddc-9717-3d168e17b137 · inbound

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing cites this paper.

Computed Tomography Visual Question Answering with Cross-modal Feature Graphing M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T19:54:39.777120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:54:39.777120Z digest=sha256:9de8348c373d588498dcf6d62e3bb6d787788798878c2c67d95df15cde7becbb

Observation 3827a0b4-4f05-43e1-a11f-de13308982b9 · inbound

Uncertainty-Driven Expert Control: Enhancing the Reliability of Medical Vision-Language Models cites this paper.

Uncertainty-Driven Expert Control: Enhancing the Reliability of Medical Vision-Language Models M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T18:06:23.568402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:06:23.568402Z digest=sha256:9cee5292295ea5bde3ea0785a99ba5d03fea9be8a33240495893ecca356ba55f

Observation 4d1a4189-a266-44f6-bd16-250cd314af5f · inbound

A Comprehensive Survey of Electronic Health Record Modeling: From Deep Learning Approaches to Large Language Models cites this paper.

A Comprehensive Survey of Electronic Health Record Modeling: From Deep Learning Approaches to Large Language Models M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T16:42:18.939180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:42:18.939180Z digest=sha256:9423a4273aa6dc1a56025d9e3d3a78358e32a14126ee57dc6c12a04ee98412e1

Observation 429a6ed5-dac3-4571-ae83-b392dadae28f · inbound

Analysis of Image-and-Text Uncertainty Propagation in Multimodal Large Language Models with Cardiac MR-Based Applications cites this paper.

Analysis of Image-and-Text Uncertainty Propagation in Multimodal Large Language Models with Cardiac MR-Based Applications M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T16:39:48.004934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:39:48.004934Z digest=sha256:d91963eeeea055708621a81d8c5ff154cac6bc4304833ebf6984ebcd53d9777f

Observation ac148cf1-fe75-4dd7-9fa9-ff26107f2d1a · inbound

Cardiac-CLIP: A Vision-Language Foundation Model for 3D Cardiac CT Images cites this paper.

Cardiac-CLIP: A Vision-Language Foundation Model for 3D Cardiac CT Images M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T12:12:57.424476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:12:57.424476Z digest=sha256:070a7716a350a3e79111d2883b0cde0cc8d0d855b84a2bd23398774a690a03f5

Observation 5569499f-13f0-42fd-ab8f-8d331e38777e · inbound

Disorder-induced stress-flow misalignment in soft glassy materials revealed using multi-directional shear cites this paper.

Disorder-induced stress-flow misalignment in soft glassy materials revealed using multi-directional shear M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T23:25:36.790483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:25:36.790483Z digest=sha256:dcfc5345d7af2818f6db7057220667db182764887abaefe22515a1a2eaa1986d

Observation 1c8df07b-4282-4af2-aa22-24f49dff84f6 · inbound

VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine cites this paper.

VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T17:30:00.948324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:30:00.948324Z digest=sha256:37341cdacf6eddb64e8ea8f6d2b1a2a379242188ad913701a9303a39272ecafe

Observation c9f01c67-46b2-4be2-8c81-5e915db4d451 · inbound

Unified Supervision For Vision-Language Modeling in 3D Computed Tomography cites this paper.

Unified Supervision For Vision-Language Modeling in 3D Computed Tomography M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T12:29:50.230776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:29:50.230776Z digest=sha256:53d347e8b9da126878fb804f7f6eea25cd1016035cd710d28cf8129a378a0fd9

Observation 148242bc-beaf-4689-a852-a917e615e22d · inbound

Discrete Prompt Tuning via Recursive Utilization of Black-box Multimodal Large Language Model for Personalized Visual Emotion Recognition cites this paper.

Discrete Prompt Tuning via Recursive Utilization of Black-box Multimodal Large Language Model for Personalized Visual Emotion Recognition M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T13:29:52.943085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:29:52.943085Z digest=sha256:ea64be36c51882dd87e8e8301f6c8e9acc25da3d6f1070074cbe566d52907865

Observation 6aba8ddf-b582-48be-b877-5da0a05f5a3d · inbound

SimCroP: Radiograph Representation Learning with Similarity-driven Cross-granularity Pre-training cites this paper.

SimCroP: Radiograph Representation Learning with Similarity-driven Cross-granularity Pre-training M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T20:51:06.573174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:51:06.573174Z digest=sha256:730f45b082c16b99342f8435be2c1d317ae42e113158aab8ed8b09747013cd24

Observation e3004081-6425-4a0d-ae43-e085ed86ab35 · inbound

Enhancing 3D Medical Image Understanding with Pretraining Aided by 2D Multimodal Large Language Models cites this paper.

Enhancing 3D Medical Image Understanding with Pretraining Aided by 2D Multimodal Large Language Models M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-04T19:48:40.847417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:48:40.847417Z digest=sha256:5b7144c312566e160da6f75c5021b11371da8b6eb6210548be9c5e6cc9ef9340

Observation d6dd44cf-da80-4248-986b-23d2752283de · inbound

MITS: A Large-Scale Multimodal Benchmark Dataset for Intelligent Traffic Surveillance cites this paper.

MITS: A Large-Scale Multimodal Benchmark Dataset for Intelligent Traffic Surveillance M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T20:33:33.091166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:33:33.091166Z digest=sha256:069501b69bf417a9d05f716475ff8eea5311d1f2ed0f2d3259104f108ec183c0

Observation 005877fd-9168-4760-b693-704595f14130 · inbound

Toward a Vision-Language Foundation Model for Medical Data: Multimodal Dataset and Benchmarks for Vietnamese PET/CT Report Generation cites this paper.

Toward a Vision-Language Foundation Model for Medical Data: Multimodal Dataset and Benchmarks for Vietnamese PET/CT Report Generation M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T13:51:50.397726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:51:50.397726Z digest=sha256:334416f9caba17e0da347762d896b7ee87bce9584088df06c7a2ff2c1947d10c

Observation fe6b51a2-07cb-4b1c-b50b-ffbe728f3d2e · inbound

IBISAgent: Reinforcing Pixel-Level Visual Reasoning in MLLMs for Universal Biomedical Object Referring and Segmentation cites this paper.

IBISAgent: Reinforcing Pixel-Level Visual Reasoning in MLLMs for Universal Biomedical Object Referring and Segmentation M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-16T17:33:10.136009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-16T17:31:33.903063Z digest=sha256:f1f1a33ab5f07c4bc73160815497e29f2f5a667c7189739d889e54fb997e0a3d

Observation bd8a8576-6537-45f3-ae60-098771088330 · inbound

Beyond Medical Diagnostics: How Medical Multimodal Large Language Models Think in Space cites this paper.

Beyond Medical Diagnostics: How Medical Multimodal Large Language Models Think in Space M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-02T18:16:31.288447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:16:31.288447Z digest=sha256:796f7031c5bfe524593387a2a360af9cac79abcc2c02fcba85f2cbd4fc380006

Observation 3f6ea956-afc8-4a51-acc3-c0eb28d0c475 · inbound

Medical Image Spatial Grounding with Semantic Sampling cites this paper.

Medical Image Spatial Grounding with Semantic Sampling M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-14T21:05:40.877444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T21:05:40.877444Z digest=sha256:1b2ec2da88b5a7083625233e1f2a5f66bd0c1acf136acfa1793bf23831745f33

Observation 48436c65-7b87-48aa-a06b-56a5ae0b7220 · inbound

Machine Learning for Network Attacks Classification and Statistical Evaluation of Adversarial Learning Methodologies for Synthetic Data Generation cites this paper.

Machine Learning for Network Attacks Classification and Statistical Evaluation of Adversarial Learning Methodologies for Synthetic Data Generation M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-13T23:01:41.380237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T23:01:41.380237Z digest=sha256:35f04db1551728440ac48171c2336497485951022951ce4ab9de989da27f798d

Observation 4a339af0-f232-449c-bfe7-6bf03b6ac85a · inbound

Visual Instruction-Finetuned Language Model for Versatile Brain MR Image Tasks cites this paper.

Visual Instruction-Finetuned Language Model for Versatile Brain MR Image Tasks M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T20:43:15.199309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-13T20:39:11.270334Z digest=sha256:d75e3bba5a997b2b29fbe365058564e728a7c34638519b329bea433bb44b64fd

Observation 138461d1-9c0c-4b30-9c66-22adfc8bbb42 · inbound

Learning Robust Visual Features in Computed Tomography Enables Efficient Transfer Learning for Clinical Tasks cites this paper.

Learning Robust Visual Features in Computed Tomography Enables Efficient Transfer Learning for Clinical Tasks M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:03:00.237995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-13T16:59:24.785428Z digest=sha256:16c9f68af5d4465034b72c2c5a33b61e8a40318a4aff0adfbca2ad26508cca60

Observation cb2eda88-b33c-444e-ba19-e771842c9bfd · inbound

Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis cites this paper.

Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:20:58.323661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T16:43:02.337806Z digest=sha256:19fdd304c9886f1e8d54fe2ce6204391e0ccbd7d6b85706eb8c17abb93fe6ec1

Observation 86fcf4aa-8ba2-4b64-8b48-0db5b84b8ff9 · inbound

Representation geometry shapes task performance in vision-language modeling for CT enterography cites this paper.

Representation geometry shapes task performance in vision-language modeling for CT enterography M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:41:02.482725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T15:53:30.757910Z digest=sha256:089d0abbdbefce3755d248114e29ee985655335166ed7c081905b759c6f5dd77

Observation a46b70fe-bfd3-4dbb-ac7c-b72bff565382 · inbound

RadAgent: A tool-using AI agent for stepwise interpretation of chest computed tomography cites this paper.

RadAgent: A tool-using AI agent for stepwise interpretation of chest computed tomography M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-10T10:55:03.876696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T10:54:35.000783Z digest=sha256:270072e5d07e33383fd2eb38cb5786adc3d996859f3d564fc5d57a5d0d2a33de

Observation 9d036104-0408-4a7d-9184-2671cf3fd640 · inbound

RadAgent: A tool-using AI agent for stepwise interpretation of chest computed tomography cites this paper.

RadAgent: A tool-using AI agent for stepwise interpretation of chest computed tomography M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-12T19:46:06.596310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T19:46:06.596310Z digest=sha256:86bd6ffd61b79faf2d98d151b238135123262e3671fd907dbb24e220a0dd77d6

Observation 7506e170-b439-4c9d-90ed-1800a6e1e08a · inbound

Agentic Large Language Models for Training-Free Neuro-Radiological Image Analysis cites this paper.

Agentic Large Language Models for Training-Free Neuro-Radiological Image Analysis M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T08:27:51.878808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T08:25:57.646822Z digest=sha256:1773b66a39a4a178c44810779f271b7eafd3cf95efc8acb8780561b8941e8bc6

Observation 770f8a69-294e-4fd9-82e1-1ea0c6d32b50 · inbound

MedScribe: Clinically Grounded CT Reporting through Agentic Workflows cites this paper.

MedScribe: Clinically Grounded CT Reporting through Agentic Workflows M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T05:55:31.001916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-05-08T19:22:09.493354Z digest=sha256:cc6ff37a22c056e8c4cf6956c290a4a85f8eab96a5a83ec537def37bd8915189

Observation 44219f25-ef43-45c8-b1d2-95882c382539 · inbound

CXR-ContraBench: Benchmarking Negated-Option Attraction in Medical VLMs cites this paper.

CXR-ContraBench: Benchmarking Negated-Option Attraction in Medical VLMs M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:41:09.638022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-08T14:49:53.357083Z digest=sha256:0cb0ac376569a4439f490318c47d15b50348575945852922f94c5e3cfff462aa

Observation d3d81bc0-3430-4360-8cc9-34c57dae9ada · inbound

Lost in Volume: The CT-SpatialVQA Benchmark for Evaluating Semantic-Spatial Understanding of 3D Medical Vision-Language Models cites this paper.

Lost in Volume: The CT-SpatialVQA Benchmark for Evaluating Semantic-Spatial Understanding of 3D Medical Vision-Language Models M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T07:56:27.075904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-12T01:31:41.805037Z digest=sha256:69c3249b4e93bbbe8802dcfac1efbe3896e3ea3820bc94080f33a31da1451129

Observation d85bdc3c-c26b-417b-b411-e0023f455b7e · inbound

Lost in Volume: The CT-SpatialVQA Benchmark for Evaluating Semantic-Spatial Understanding of 3D Medical Vision-Language Models cites this paper.

Lost in Volume: The CT-SpatialVQA Benchmark for Evaluating Semantic-Spatial Understanding of 3D Medical Vision-Language Models M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T13:25:45.163252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-30T23:19:20.522732Z digest=sha256:371356fdbfd1e816e16a568ae6e5618a5bfb9660e8b4026c37379e8fbb7b2f43

Observation 529e51a5-2359-4a87-b6cd-7304dd5839c8 · inbound

DeepTumorVQA: A Hierarchical 3D CT Benchmark for Stage-Wise Evaluation of Medical VLMs and Tool-Augmented Agents cites this paper.

DeepTumorVQA: A Hierarchical 3D CT Benchmark for Stage-Wise Evaluation of Medical VLMs and Tool-Augmented Agents M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:41:43.481584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-12T04:02:36.159920Z digest=sha256:0278f3580fbc63010a4b592289c07beb589b9989a068dc6df49dfa4579bed9c6

Observation 22380ac2-6bc9-48f5-af9f-b6797fc2e88d · inbound

RadThinking: A Dataset for Longitudinal Clinical Reasoning in Radiology cites this paper.

RadThinking: A Dataset for Longitudinal Clinical Reasoning in Radiology M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:56:21.475486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-12T03:55:55.359488Z digest=sha256:4d403f4521db9577f1d09ccaf8aef9b1f57b7392f0a9cd3b5670fcd1a1cd9607

Observation 05262b06-a4d4-4746-ad24-b0d3da38668b · inbound

M3Net: A Macro-to-Meso-to-Micro Clinical-inspired Hierarchical 3D Network for Pulmonary Nodule Classification cites this paper.

M3Net: A Macro-to-Meso-to-Micro Clinical-inspired Hierarchical 3D Network for Pulmonary Nodule Classification M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:59:27.387088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-14T20:55:57.769186Z digest=sha256:60031975f4a500f4e0340ebfa76d7118416621c0ffcd33674ce329c8751f8d02

Observation 7fb91275-cba4-4666-9659-b98a272329ea · inbound

CA-GCL: Cross-Anatomy Global-Local Contrastive Learning for Robust 3D Medical Image Understanding cites this paper.

CA-GCL: Cross-Anatomy Global-Local Contrastive Learning for Robust 3D Medical Image Understanding M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T19:57:52.874218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-14T19:57:47.615619Z digest=sha256:27c883c19f21fa091b9621546a9f49954d59854bffab566722f0147565980258

Observation ca6b1399-553e-431c-ba0c-540cf4e34249 · inbound

CA-GCL: Cross-Anatomy Global-Local Contrastive Learning for Robust 3D Medical Image Understanding cites this paper.

CA-GCL: Cross-Anatomy Global-Local Contrastive Learning for Robust 3D Medical Image Understanding M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T01:19:20.424112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-04T01:10:31.958276Z digest=sha256:75af109013a3bc4060f7d94eb999a100ec22462cfad37db608193ac6887f12bb

Observation 9d3c9cbb-ecde-4fb4-a6df-a3eaeb18f26f · inbound

Segmentation, Detection and Explanation: A Unified Framework for CT Appearance Reasoning cites this paper.

Segmentation, Detection and Explanation: A Unified Framework for CT Appearance Reasoning M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-20T18:23:37.668034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-20T18:20:38.720544Z digest=sha256:4f3373c8bff9ad17b3a6635592e3928d7dfb2a301d7c21eef42474772de0f64f

Observation 1ae3aedd-9ba1-4ffd-ab96-2a7b455fcc9c · inbound

Regulating Anatomy-Aware Rewards via Trajectory-Integral Feedback for Volumetric Computed Tomography Analysis cites this paper.

Regulating Anatomy-Aware Rewards via Trajectory-Integral Feedback for Volumetric Computed Tomography Analysis M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T08:09:51.827961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-05-21T08:06:02.176934Z digest=sha256:6ce5527cf88be574fafb79b39512212eeffc5e46ebc623ff334dd351c96b8ef2

Observation 2c03b03f-12fe-4206-b2d3-81676cdeb256 · inbound

NeuroQA: A Large-Scale Image-Grounded Benchmark for 3D Brain MRI Understanding cites this paper.

NeuroQA: A Large-Scale Image-Grounded Benchmark for 3D Brain MRI Understanding M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-05-21T06:59:45.761587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-21T06:54:55.254082Z digest=sha256:2981031ece7b6d932b62d40a9baa15f32d661de7c2d746846ee79c0b4fab411c

Observation 506a630d-b3a8-4a77-9c5b-3bbbfaea9a63 · inbound

An Open Multi-Center Whole-Body FDG PET/CT Foundation Model for Tumor Segmentation cites this paper.

An Open Multi-Center Whole-Body FDG PET/CT Foundation Model for Tumor Segmentation M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:21:12.683535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-22T07:20:10.061711Z digest=sha256:684581b23bcf5e5b70f7bb361461b4f6dd9c5239f2a742d00bc372c7744546ce

Observation ec6342dc-bd07-4b37-824a-f33f47af1b31 · inbound

SliceWorld: A Predictive and Controllable World-State Model for CT Report Generation cites this paper.

SliceWorld: A Predictive and Controllable World-State Model for CT Report Generation M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-06-30T14:04:44.781779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-30T13:57:09.053990Z digest=sha256:2f72fcd496fc3f094781852e1a68ab768c61fdb0c6bd3036b5165007e4261141

Observation b7c10c63-759f-4782-83d1-4cf296c7de96 · inbound

MedVol-R1: Reward-Driven Evidence Grounding for Volumetric Reasoning Segmentation cites this paper.

MedVol-R1: Reward-Driven Evidence Grounding for Volumetric Reasoning Segmentation M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T18:33:50.891797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-29T18:26:10.873395Z digest=sha256:3f688d8d7f02d8e1b4e848eb96499826804ca22612b042134f93cce16069e9e4

Observation 39b0853b-74e0-4be6-bbfc-bf2773dab466 · inbound

Astra: a generalizable report generation foundation model for 3D computed tomography cites this paper.

Astra: a generalizable report generation foundation model for 3D computed tomography M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-06-29T00:02:50.405512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-28T23:16:12.442334Z digest=sha256:e9ac11d2a554f45d17e0ad60bc485dc8ec49e6ccd36669382be05c334f87dce4

Observation 71075a4f-7495-4349-a7e9-a8c86250dc02 · inbound

CardioLens: Revealing the Clinical Reality Gap of MLLMs via Multi-Sequence Cardiac MRI Evaluations cites this paper.

CardioLens: Revealing the Clinical Reality Gap of MLLMs via Multi-Sequence Cardiac MRI Evaluations M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:33:15.355195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-29T08:28:48.174508Z digest=sha256:8d944a648066fbbf6dba8ced68f7993d7e4694f55bac5dd5d83001cd8ae85119

Observation a0f3b2d1-8329-4eb8-8fb6-8a321e9d304f · inbound

ASAP: Advancing Medical Volumetric Representation Learning with Anatomy-aware Semantically-adaptive Pre-training cites this paper.

ASAP: Advancing Medical Volumetric Representation Learning with Anatomy-aware Semantically-adaptive Pre-training M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-06-28T19:22:34.679589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-28T19:16:42.139096Z digest=sha256:03229855bb4a9b4d44ddd9a8fd9a162d910c3b2d62b25dff7333fcd5d5c93fc6

Observation af9e1472-583d-4ac1-a46a-fffed3d884a6 · inbound

Multi-Granularity 3D Kidney Lesion Characterization from CT Volumes cites this paper.

Multi-Granularity 3D Kidney Lesion Characterization from CT Volumes M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-07-02T07:06:44.625727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-28T07:06:49.770326Z digest=sha256:dc64252bc1aca6dd6c773161b0c956330d247bfb36836200b59298e6798e551f

Observation defe729f-ada0-415d-8659-a556e91a25c9 · inbound

UniReason-Med: A Shared Grounded Reasoning Interface for 2D-to-3D Transfer in Medical VQA cites this paper.

UniReason-Med: A Shared Grounded Reasoning Interface for 2D-to-3D Transfer in Medical VQA M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 165

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T09:47:59.536498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-06-27T10:21:12.782864Z digest=sha256:6b853786f42cdb6ff4cfaf51c5ffb107701053663d7aef9a3cfea808da812906

Observation 36072f2b-9064-4ff3-85b0-5f89be424a96 · inbound

Venice-H1: Failure-Aware Query Re-Ranking with Multi-Scale Grid Signatures for Referring Image Segmentation cites this paper.

Venice-H1: Failure-Aware Query Re-Ranking with Multi-Scale Grid Signatures for Referring Image Segmentation M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-06-26T11:29:24.583964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-26T11:22:05.498644Z digest=sha256:d640578b75e62ed32366856f3ef7753e5e23ea758affc0efdd698fd4585788a4

Observation f994ea63-e58b-437c-b779-31ae20145b28 · inbound

E-MRL: Cross-view Aligned Evidence-driven Multimodal Reinforcement Learning for Reliable 3D Tumor Analysis cites this paper.

E-MRL: Cross-view Aligned Evidence-driven Multimodal Reinforcement Learning for Reliable 3D Tumor Analysis M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T12:49:52.204124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-26T06:02:30.459845Z digest=sha256:6d21850b8f50c5dc5210a70e92958a2777989b9da1f59409d7569012b39a52b3

Observation 3789f371-cba5-43a9-85ca-9b5e5f876c78 · inbound

MRI2Rep: Autoregressive Structured Report Generation for 3D Liver MRI cites this paper.

MRI2Rep: Autoregressive Structured Report Generation for 3D Liver MRI M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T20:00:08.150092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-25T20:55:03.803170Z digest=sha256:e436b0cf622d5345089770099b523016bb32b2361c98565f0f96da2d35c2414a

Observation 06ff476a-e955-41d5-84e8-9510e2c81c41 · inbound

TRACE: A Concept Bottleneck Model for Longitudinal 3D Glioblastoma Response Assessment cites this paper.

TRACE: A Concept Bottleneck Model for Longitudinal 3D Glioblastoma Response Assessment M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-06-30T06:54:19.947742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-30T06:52:42.092509Z digest=sha256:e183f4a58bf6831be731765cfa81fa9771929b8b08213e668ea2785f124c4f74

Observation f22dbbf4-8137-445d-8294-8f6817a00ba9 · inbound

TRACE: A Concept Bottleneck Model for Longitudinal 3D Glioblastoma Response Assessment cites this paper.

TRACE: A Concept Bottleneck Model for Longitudinal 3D Glioblastoma Response Assessment M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-07-03T22:39:00.916214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-03T22:29:18.856647Z digest=sha256:63b9a71308efaa36d1a2022fcccf7cdd67015fff743079d35abfc902d98bdda4

Observation fa7b3241-1b38-45b5-9007-bddae5600c32 · inbound

Towards Enhancing 3D Spatial Reasoning in Medical Multimodal Large Language Models cites this paper.

Towards Enhancing 3D Spatial Reasoning in Medical Multimodal Large Language Models M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T03:31:45.407270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:31:45.407270Z digest=sha256:0f178285bd72108748c50a2e498b6f07ec1a073424321ede30d55c896ffe9fd0

Observation 1e39b3dd-2d4e-4ff0-bd57-106651514303 · inbound

Multi-LLM Collaborative MRI Report Generation for Visual Instruction Tuning in Brain Oncology cites this paper.

Multi-LLM Collaborative MRI Report Generation for Visual Instruction Tuning in Brain Oncology M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T01:44:20.068459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:44:20.068459Z digest=sha256:230b6ed12719f3c10191d36bdcd0143e0f1c29ade755ba7f490139e25ed8c780

Observation 9e6e25e0-7723-432d-8ada-360944d57f86 · inbound

Cheap Probes Predict Expensive Training in 3D-CT Vision--Language Models cites this paper.

Cheap Probes Predict Expensive Training in 3D-CT Vision--Language Models M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T06:15:35.089393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:15:35.089393Z digest=sha256:7394306ec95cfdd3ea49e39b3babe79d0cf621d40b324a2563fd1c28c5fa7109

Observation 9b759028-84a6-47a2-ba64-7c60aba90e4a · inbound

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding cites this paper.

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 67

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:29.953458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:29.953458Z digest=sha256:161d2b7a16c7603b8c18a5e62633c8c7bad0aa19dcaed41af452f3fdb765b1e9

Observation c0c13d68-673d-4e3d-8620-2bbc419e8d3d · inbound

Rad-JEPA 3D: Radiology Joint-Embedding Predictive Model for 3D Computed Tomography cites this paper.

Rad-JEPA 3D: Radiology Joint-Embedding Predictive Model for 3D Computed Tomography M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-01T00:38:25.843369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:38:25.843369Z digest=sha256:94ff760f64c74bf121b52a5b508f92a364f56161c8278892fda40fe2925d08c5

Observation dfb6c282-c217-4b04-887e-8be7488d40c8 · inbound

MedARC: Training-Free Adaptive Redundancy Compression of Visual Tokens for 3D Medical Vision-Language Models cites this paper.

MedARC: Training-Free Adaptive Redundancy Compression of Visual Tokens for 3D Medical Vision-Language Models M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-01T13:31:49.000325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:31:49.000325Z digest=sha256:81388a66749a8813392e8003cd6d82d70fda44d48a090ae5d9135a128fcc4bda

Observation da340ddf-4a0f-4494-8801-ff5e61ec899e · inbound

Objective-Aligned Direct Answer SFT for Robust Multi-Frame Medical VQA cites this paper.

Objective-Aligned Direct Answer SFT for Robust Multi-Frame Medical VQA M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T05:36:44.874743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T05:36:44.874743Z digest=sha256:cc1bdda661a90f0377c279e06e67cf958648a7118d6a773dbf034613e726c199

Observation ddb36263-bcb2-4c54-af9f-00d976494409 · inbound

Learning How Much, Not Just What: Cross-Patient Burden Order for CT Vision-Language Pretraining cites this paper.

Learning How Much, Not Just What: Cross-Patient Burden Order for CT Vision-Language Pretraining M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-04T01:00:54.281987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:00:54.281987Z digest=sha256:03371bd933c295377504e8fc6fd459c9e03fd1eefb2181a2dc4a1847357621f2

Observation bdc3c00f-85c1-44b6-866d-4c7cdd0290d8 · inbound

ORCA: ORgan-Centroid Aggregation for Training-Free 3D CT Visual Token Compression cites this paper.

ORCA: ORgan-Centroid Aggregation for Training-Free 3D CT Visual Token Compression M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-04T00:46:03.549411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T00:46:03.549411Z digest=sha256:6d3d414a60003137ead106de9aef1316cb17214e9812a33329e1b6c010892e72

Observation ce51c620-3de8-41b5-bf9e-30bebf1ba67a · inbound

CARVE: Cross-Slice Anisotropic Reallocation of Visual Evidence for Efficient 3D Medical Volume Understanding cites this paper.

CARVE: Cross-Slice Anisotropic Reallocation of Visual Evidence for Efficient 3D Medical Volume Understanding M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T14:41:24.448336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:41:24.448336Z digest=sha256:ee25251a3b9b26b7d9da38ef1c12b12352767c2dfe6511fdbae0388f1d51343c

Observation ecc433c6-e8e9-4b99-a967-8151a028bf13 · inbound

Resolution Meets Reduction: Efficient Visual Context for 3D Radiology Report Generation cites this paper.

Resolution Meets Reduction: Efficient Visual Context for 3D Radiology Report Generation M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-14T04:32:15.838226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:32:15.838226Z digest=sha256:a4626a888c214b074f197d326af2f3bde98b4efbc2e991f341395cfccbb0b8b2

Observation 74802f81-024e-4f16-a846-98c0c0750a94 · inbound

HounsWorld: A Multimodal World Model for Hidden Patient-State Readout, Reconstruction, and Simulation cites this paper.

HounsWorld: A Multimodal World Model for Hidden Patient-State Readout, Reconstruction, and Simulation M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T21:00:35.740016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:00:35.740016Z digest=sha256:68f6a4f8194c6784a63491a7b926881d22f38ef7f2b556b25c22e4a1d4065b58