Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-10T16:43:02.337806Z
Paper Citation Record · LEDGER
As of 3 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 0 inbound Pith citation observations for arXiv:2604.10233.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-10T16:43:02.337806Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-03T06:30:56.289259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
49 of 49 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 71e8a687-9d7f-4d8c-84ea-29cc065c1030 · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Qwen2.5 Technical Report
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 885679bd-03eb-437a-8864-b7997f67bcc5 · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 7c24f226-36f0-4a2b-94c2-22c9bdf01b17 · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 28cc916f-d577-456e-a3b8-9367ed76b9f6 · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Sigmoid loss for language image pre-training
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 9e35857a-b8d2-4924-b5df-1874f6dda033 · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis DINOv2: Learning Robust Visual Features without Supervision
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation df48b65a-d775-4173-84f3-50775044ec35 · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Spatio-temporal and retrieval-augmented modelling for chest x-ray report generation
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 2972c0e6-4ace-42ae-aaf9-23cb34cffda4 · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 2acf86df-eca7-49e2-81ed-9137992b1607 · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Collaboration between clinicians and vision–language models in radiology report generation
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e442baf4-1c98-47f2-9b54-cf42217ebf8b · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Interpretable Bilingual Multimodal Large Language Model for Diverse Biomedical Tasks
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 1ab44201-8a4c-4ac6-b9d9-5401d5625989 · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Llava-med: Training a large language-and-vision assistant for biomedicine in one day
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation cb2eda88-b33c-444e-ba19-e771842c9bfd · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 6689e5fc-3971-4ffb-9962-0dbed9760043 · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Towards generalist foundation model for radiology by leveraging web-scale 2d&3d medical data
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 965cf1b4-ca81-431c-9e77-87e31a1d58ca · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Learning transferable visual models from natural language supervision
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e0e0c55b-773d-436a-bd43-ff8af3433adf · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Adaptive mixtures of local experts
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 71de56fe-cc81-40ff-a289-89ee9c2648a0 · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Learning to prompt for vision- language models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 3225faba-2b91-446b-98f6-6e90560e9314 · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Con- trastive learning of medical visual representations from paired images and text
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 04e4dda3-4bcc-49ae-aed6-70bb521fefd2 · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Procedure-aware surgical video- language pretraining with hierarchical knowledge augmentation
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation af982edf-2232-4d82-8cb3-a40740a428d1 · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Merlin: a computed tomography vision–language foundation model and dataset
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 33b693be-f363-4f52-9241-b30cf95c649b · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Triad: Vision foundation model for 3d magnetic resonance imaging
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f1f07db4-5a90-413c-83b1-1c8fcfcc2ea1 · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Learning neuroimaging models from health system-scale data
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 4388f767-c5fa-470f-9cb6-3ddaf465984f · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis LLaMA: Open and Efficient Foundation Language Models
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 0736d530-b06e-4bc9-a8b4-6e097667fb4f · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation ef70287c-a201-4846-9dc3-f2773fe76b92 · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Towards accurate differential diagnosis with large language models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 2d742989-f691-40bd-91ac-9f4e4c02ba5a · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Toward expert-level medical question answering with large language models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation b5421448-b2c4-4a5a-9045-54fee4451c4d · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Medla: A logic-driven multi-agent framework for com- plex medical reasoning with large language models
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation b72f02a3-faef-4772-bf46-a5d89c9d0f96 · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis A generalist vision–language foundation model for diverse biomedical tasks
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 196b1ef7-d7a2-4d0b-801f-1cb8542f6375 · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis arXiv preprint arXiv:2503.20047 , year=
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 7127ecd2-47d3-4ba9-bfde-76cdd8bf5981 · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Dynamic graph enhanced contrastive learning for chest x-ray report generation
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 27716b1b-cddf-4207-abf9-9df74fa96c31 · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis A medical multimodal large language model for future pandemics
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f34b5b53-04c8-42d3-92d0-3e879f4bb4a7 · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Multimodal generative ai for medical image interpre- tation
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 73c1a773-8cfb-4ca4-9173-e7e25da55669 · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Towards a holistic framework for multimodal llm in 3d brain ct radiology report generation
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation ec1470e2-6ae5-42a4-bf14-9aeadb954c13 · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis MIMIC-CXR-JPG, a large publicly available database of labeled chest radiographs
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 7f52210f-b72d-46ea-9240-3943cd91dd5c · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Generating Radiology Reports via Memory-driven Transformer
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 156ea7ad-b8e0-4798-94ce-fd58a2803567 · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Promptmrg: Diagnosis-driven prompts for medical report generation
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation be8c466a-9af3-481e-85be-b5bede69c09e · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Gmai-mmbench: A comprehensive multimodal evaluation benchmark towards general medical ai
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 10e844f4-7d35-4be5-93c3-562763853f2d · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Omnimedvqa: A new large-scale comprehensive evaluation benchmark for medical lvlm
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 5e236b72-212a-46ea-a51f-a2f8ccda19af · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Lmt++: Adaptively collaborating llms with multi- specialized teachers for continual vqa in robotic surgical videos
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 275fd5fc-8d0a-4034-b588-7936e5fd2672 · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Interactive and ex- plainable region-guided radiology report generation
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 96144497-895a-42ca-8828-9ccef4c52098 · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation ac59c447-9289-4d80-beff-b11658f6d429 · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Outrageously Large Neural Networks: The Sparsely-Gated Mixture-of-Experts Layer
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 382e72aa-1d4f-44b2-89d4-4a856dc7b8f2 · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Roformer: En- hanced transformer with rotary position embedding
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 85709c67-c6e4-4469-9b7b-63c0de2ee9cf · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Do vision transformers see like convolutional neural networks?
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation d72ed521-488d-4648-892b-802c79122c74 · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis PaCE: Unified Multi-modal Dialogue Pre-training with Progressive and Compositional Experts
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 6df0f719-58d3-4f66-977d-8a81f876bafd · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Scaling vision with sparse mixture of experts
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 95961ac8-9d37-4d97-8ab0-2d5c3ffe849c · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Mixture of Cluster-conditional LoRA Experts for Vision-language Instruction Tuning
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 90e1b575-01a0-404d-8864-38a4b2c59217 · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Med-MoE: Mixture of Domain-Specific Experts for Lightweight Medical Vision-Language Models
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation aefe669f-46d9-440c-8157-bff5acc24a51 · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis MoE-LLaVA: Mixture of Experts for Large Vision-Language Models
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e759b238-5fae-4070-848e-dcc4c4d264a7 · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Lora: Low-rank adaptation of large language models
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 18967fb0-cb06-4163-8415-1433c13aa1b3 · outbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis Zero: Memory optimizations toward training trillion parameter models
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
No inbound Pith citation observations are available.