Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T15:16:27.982410Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 85 of 85 outbound references and 11 inbound Pith citation observations for arXiv:2411.14522.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T15:16:27.982410Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T22:52:07.139701Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T01:39:23.953906Z
85 of 85 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation f0809c6c-e530-4956-bcdd-165374031977 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb68ff5b-f2c4-4c68-8a85-b779a8caedd9 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI The claude 3 model family: Opus, sonnet, haiku
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ebd4e356-e6fd-4315-94c0-9df366c4fc05 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Home - Retina Image Bank, 2024
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 939980e2-769d-4fab-bca9-f9733750eeb7 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI OpenFlamingo: An Open-Source Framework for Training Large Autoregressive Vision-Language Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 737d2341-da84-4bb5-a81c-6ec16d908136 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Blossom orca v3
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c358ec7-ce60-4f89-a2ff-f805515b47bd · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ff89aa8-1982-4eac-a181-7e472d9800e5 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI COIG-CQIA: Quality is All You Need for Chinese Instruction Fine-tuning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f1b43fe-ab6d-45a7-a574-e3609967799f · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Vqa-med: Overview of the medical visual question answering task at imageclef 2019
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8dc39d1f-77dc-4245-8dae-722d79aca42e · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Cosmopedia, 2024
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7abd5be5-6e47-4733-8029-1b3ca22517b0 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Leetcode dataset, 2023
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 0b54dee0-d328-40fe-8383-6bd19b4df47f · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI An augmented benchmark dataset for geometric question answering through dual parallel text en- coding
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 85d4dbcf-82ad-45f9-88e4-b9f37d25c8c8 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI CheXpert Plus: Augmenting a Large Chest X-ray Dataset with Text Radiology Reports, Patient Demographics and Additional Image Formats
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d842d6a9-bf21-4984-a324-0eb0a2b022ac · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00198be9-9255-40dc-a781-d94e84bcdfb4 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6452a82c-112c-497c-8344-a58c7e70e083 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI ShareGPT4V: Improving Large Multi-Modal Models with Better Captions
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9992be07-b41d-4250-8c0d-7bb53dc2a890 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI GMAI-MMBench: A Comprehensive Multimodal Evaluation Benchmark Towards General Medical AI
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67103c64-7a6c-4c77-a765-e4fa706aa943 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d99355f5-d6e6-4409-a946-df250c1b60a2 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1dcfadc6-fc6b-452b-bc74-a8276bdc61db · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Xtuner: A toolkit for efficiently fine-tuning llm
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e77aa06e-8cda-4729-9e83-b4e5ea7a6843 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Instructblip: Towards general- purpose vision-language models with instruction tuning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ad70fd0b-a3f1-4d5b-9d30-30cc3d69bbca · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Preparing a collection of radiology examinations for distribution and re- trieval
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation bc1e1e32-9ddb-49b9-9fdc-a15944844fbd · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Cogview: Mastering text-to-image generation via transformers
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 6acb6d61-6d11-4d9a-a197-33b0bf516e35 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Enhancing chat language models by scaling high- quality instructional conversations, 2023
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 3d36b248-f3b7-431d-9b55-156ba62e6399 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI InternLM-XComposer2: Mastering Free-form Text-Image Composition and Comprehension in Vision-Language Large Model
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9535ac0b-a992-421a-9ea8-04cfdbcbfd07 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI PaLM-E: An Embodied Multimodal Language Model
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 948e59e8-2f08-439e-8009-5b0feccbab3e · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Vlmevalkit: An open- source toolkit for evaluating large multi-modality models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8a090375-375b-410d-93ad-3a419b62b743 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Lima: Less is more for alignment, 2023
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation f82a13f6-f47a-4bd2-b682-ac2ec7b82f6f · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI LLaMA-Adapter V2: Parameter-Efficient Visual Instruction Model
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2eb5b23a-ea07-4bb6-9655-5760089ae459 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI GSCo: Towards Generalizable AI in Medicine via Generalist-Specialist Collaboration
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8bc2560-4eca-4777-828a-8138c7af8401 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI PathVQA: 30000+ Questions for Medical Visual Question Answering
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d05aa105-d6ff-4b0d-b562-f2a77d7db6d7 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Medical-diff-vqa: A large- scale medical dataset for difference visual question answer- ing on chest x-ray images, 2023
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ba74ef87-2fdb-4792-879b-d01e3d2f952d · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Omnimedvqa: A new large-scale comprehensive evaluation benchmark for medical lvlm
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ff3c5f8b-8fd7-4e65-9ae8-75f7947ad433 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Quilt-1m: One million image-text pairs for histopathology
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f30d8a76-7125-4b8e-ada2-812d6923897e · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Quilt-1m: One million image-text pairs for histopathology
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ace6742a-394a-42b2-bfdb-b3ea4158bf6e · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Mimic-cxr, a de- identified publicly available database of chest radiographs with free-text reports
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 846fe4a7-c08e-457f-b314-e69ff2ca614c · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Dvqa: Understanding data visualizations via ques- tion answering
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9550fcf-9ead-481f-afd1-ec4c9bd02839 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI A diagram is worth a dozen images
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 422491cf-ac34-4dec-a00f-76c566fb2159 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Ocr-free document understanding transformer
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 084ee0bd-36b7-4457-ac56-28d29ee3c24d · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI A dataset of clinically generated visual questions and answers about radiology images
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 2ae0acb5-eee4-4e2c-a796-97a2c323fb55 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Llava-med: Training a large language- and-vision assistant for biomedicine in one day
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 66cb7b1c-534e-4578-ae99-22a6b0fec26b · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0ead3e3-c8a0-4351-92a9-c8ec737d7e05 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Seeing and understanding: Bridging vision with chemical knowledge via chemvlm
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7dd5f11b-00a0-4169-8a00-84f9a18958a0 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Pmc-clip: Con- trastive language-image pre-training using biomedical docu- ments
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 614c8785-f960-42cf-a6f7-d6223ac5951b · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Slake: A semantically-labeled knowledge- enhanced dataset for medical visual question answering
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c8bc8183-cf94-451e-9376-1e6c221afa9f · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Visual instruction tuning, 2023
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c280c82f-eb05-476b-951e-368da4b01900 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Llava-next: Im- proved reasoning, ocr, and world knowledge, 2024
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb30a723-cae1-4d6e-af03-8c5af8a9c59c · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI LogiQA: A Challenge Dataset for Machine Reading Comprehension with Logical Reasoning
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9067b3d-23bc-4783-85ae-ad813262062f · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Qilin-Med-VL: Towards Chinese Large Vision-Language Model for General Healthcare
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5db7375-b621-4c75-ad9e-f5eec85839a8 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI DeepSeek-VL: Towards Real-World Vision-Language Understanding
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01181550-8cd0-4fb9-9e27-495406497941 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0494fc53-150b-42ec-85e3-4b4ae42c5b68 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Docvqa: A dataset for vqa on document images
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation cc0f8001-9fab-47d0-af8a-0862f495cf19 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Orca-math: Unlocking the potential of slms in grade school math, 2024
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 00368572-52cc-40f1-8a26-fa8e34513d9f · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Med-flamingo: a multimodal medical few-shot learner
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation bc8763a9-3187-47bd-a685-f42c07709cc5 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Learning transferable visual models from natural language supervi- sion
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01d7b33c-d69c-41b7-aa1f-5eef4d7074ad · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b819e4e5-efab-4e82-b283-9c263b4af9c7 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Seco de Herrera, et al
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation bac7800f-9f96-41eb-ba65-82294ed302ee · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Capabilities of Gemini Models in Medicine
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb1443b6-b347-4b3d-be41-72aae5223b28 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Large language models encode clinical knowledge
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 1b0907b5-9f76-4b44-b231-bc57dcbc8375 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI MedICaT: A Dataset of Medical Images, Captions, and Textual References
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation adfe8103-d840-4aa2-b589-e46d6b6d813b · outbound
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8bc57a48-baaf-4bbc-a4f5-019338a9bb5d · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Gemini: A Family of Highly Capable Multimodal Models
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c1ded91-8ed5-40c9-886a-ab80b647eba1 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Internlm: A multilingual language model with progressively enhanced capabilities, 2023
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 012f75d3-df3f-428e-bc0f-65b3b58fbe29 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Openhermes 2.5: An open dataset of synthetic data for generalist llm assistants, 2023
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 2d717345-e6c4-4d7c-b475-1959ea84eb6b · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI LLaMA: Open and Efficient Foundation Language Models
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 975b2ba1-b3bd-432c-902f-2bd51c9a0077 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91903ecb-3c1c-4234-917f-356fd40d1a59 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Towards gen- eralist biomedical ai
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9223ef79-fa0b-47a8-abdf-f2919e1588e2 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI PMC-LLaMA: Towards Building Open-source Language Models for Medicine
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fcb4f0f9-ac6e-44f9-afbb-d8f71ccee68e · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Towards Generalist Foundation Model for Radiology by Leveraging Web-scale 2D&3D Medical Data
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95d0985f-facc-4500-99cf-a5463af0032f · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI MedTrinity-25M: A Large-scale Multimodal Dataset with Multigranular Annotations for Medicine
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2fd5e40-6ffa-47b7-a945-2e14b621c775 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI LLaVA-UHD: an LMM Perceiving Any Aspect Ratio and High-Resolution Images
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9bea8762-ce0a-467e-8158-f1c2cf4504c6 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Firefly( 流 萤): 中 文 对 话 式 大 语 言 模 型
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 53169a36-9c93-4476-a02a-e31960efaaa1 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI SA-Med2D-20M Dataset: Segment Anything in 2D Medical Imaging with 20 Million masks
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 589337fe-1275-4edb-a5c6-ff4f47647968 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI mPLUG-Owl2: Revolutionizing Multi-modal Large Language Model with Modality Collaboration
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b35edca-080d-4e12-890b-22ac3a091234 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Yi: Open Foundation Models by 01.AI
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61836b9d-123c-4ef2-931e-d527a78ede2d · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for ex- pert agi
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e910100c-211f-4468-b854-0a80a34ee52c · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI A generalist vision–language foundation model for diverse biomedical tasks
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f476cd61-7b25-46f1-b91d-13dc0b7abaee · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f34ccd04-4706-454a-9b63-2557f7b4aa6e · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10508207-032f-4067-b8ca-c9d628aa9eaa · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI The first step involves an initial quality screening using a multimodal large language model (such as GPT-4o) to automatically assess the quality of the pairs
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation b2f6f970-88dc-4bb1-bcf7-ace042973598 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI Unresolved cited work
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 2e368cbb-8672-4b61-b8c3-b3217dec7ca5 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI If the average score of the 30 image-text pairs from a dataset is less than or equal to 3 (on a scale of 1 to 5), the dataset is classified as having satisfactory quality
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 03b51d9b-25f2-4485-aa2b-6889a99144c2 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI - Assess whether there are any medical errors or misleading information that could impact diagnostic or treatment deci- sions
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 09d4e0a6-fd43-4e39-b16f-f4adbfef6216 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI - Assess whether the language is fluent and natural, and if there are any overly complex sentence structures that might hinder information transmission
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 0de715af-47ef-4c43-9858-5ec3f5452e93 · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI - Assess whether any essential information is missing, omitted, or incomplete
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 3f55f2e7-f557-4a4f-a3b8-076e01b899ed · outbound
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI - Ensure that the imaging data is appropriately interpreted, the descriptions are clear and accurate, and they align with the medical diagnosis
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ad7a026c-852b-4e47-8001-2594a43bac8c · inbound
Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI
Reference 134
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation f0942baf-09d8-4fac-9bdc-b7bcfcf6d01c · inbound
Efficient Few-Shot Medical Image Analysis via Hierarchical Contrastive Vision-Language Learning GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c2296e6-84bd-4304-8071-13b82696856b · inbound
EndoChat: Grounded Multimodal Large Language Model for Endoscopic Surgery GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3e95069-62e4-42b4-a2a1-85888ad466fa · inbound
MM-Skin: Enhancing Dermatology Vision-Language Model with an Image-Text Dataset Derived from Textbooks GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8bacb87c-e481-4cec-ab32-258917010cad · inbound
Constructing Ophthalmic MLLM for Positioning-diagnosis Collaboration Through Clinical Cognitive Chain Reasoning GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea982593-0a13-4fd5-bd2a-332eb7dacf9b · inbound
Towards Better Dental AI: A Multimodal Benchmark and Instruction Dataset for Panoramic X-ray Analysis GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation caf1a38d-2a9a-425f-b46e-274e53eb0425 · inbound
MedSynapse-V: Bridging Visual Perception and Clinical Intuition via Latent Memory Evolution GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e9db9ebb-da10-459d-b82c-70780d3ad8e9 · inbound
MedSynapse-V: Bridging Visual Perception and Clinical Intuition via Latent Memory Evolution GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 0c5f7290-e351-41db-ab89-abcf432d32c3 · inbound
MedSynapse-V: Bridging Visual Perception and Clinical Intuition via Latent Memory Evolution GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation b74a0c42-9500-4b84-9145-94c246601ae6 · inbound
MedSynapse-V: Bridging Visual Perception and Clinical Intuition via Latent Memory Evolution GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 09c3a83a-c961-4552-979d-203866364e18 · inbound
Aloe-Vision: Robust Vision-Language Models for Healthcare GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.