Pith. sign in

Paper Citation Record · LEDGER

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding

As of 21 August 2026, this Paper Citation Record lists 64 of 64 outbound references and 8 inbound Pith citation observations for arXiv:2504.14692.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.14692 v1

Coverage vector

measured 64 of 64 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:46:51.794871Z

measured 72 of 72 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T14:41:24.473812Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T16:38:39.529470Z

Reference resolution

64 of 64 outbound references displayed

  • verified exact0
  • verified fuzzy4
  • unresolved60
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 148fea56-bed5-4359-a8db-f36124ef2780 · outbound

This paper cites online" 'onlinestring :=.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.381029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.381029Z digest=sha256:bfb058268880726d16f7b4498577531c0fa937da9a0eaf6d96ed7ca7a0d06afc

Observation dd974ef4-1c4d-473a-9237-103b572034ad · outbound

This paper cites write newline.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.386453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.386453Z digest=sha256:8e614e0344ec9a9ebce81ce5fa5f9f5338d77a40cc0ee6883664a39b7debf43a

Observation 1f363759-f80f-4c14-aca7-038fcac28479 · outbound

This paper cites an unresolved cited work.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:46:52.688403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-16T11:46:51.392191Z digest=sha256:2b9cc851bc403e25c99a285a39715c954fadba1ad631ff8c8703d13e597f523a

Observation 07856dd3-ad0f-4e42-8495-7a7a36f2ca8d · outbound

This paper cites M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.397088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.397088Z digest=sha256:e722542ba3d0571e4b36f8763216f7dba2c1390fd2a855b70646c5a482f38da6

Observation ac40fb4d-7a2a-4d68-bb49-63a402154fef · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.402356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.402356Z digest=sha256:fcb3799bc5155156ef61624a080b5a29fbd5ba06db0981cef0a3e8867fbab17e

Observation b1853748-fd0f-421f-a9c8-11f5cfeb7ac6 · outbound

This paper cites Qwen2.5-VL Technical Report.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Qwen2.5-VL Technical Report

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.406812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.406812Z digest=sha256:781297048e0294c436d1de42f28390dff91b643709bd5984570748d50b34e21c

Observation 66756e46-032c-4609-b04a-b2dcd784ec42 · outbound

This paper cites an unresolved cited work.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.411668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.411668Z digest=sha256:e6c5102a807dd2f5551011412886c33450655b9e773035c14b205a03d8c1987a

Observation 14158b15-68d5-44bd-b4dd-92bbf9d3825c · outbound

This paper cites an unresolved cited work.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:46:52.664275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-16T11:46:51.416449Z digest=sha256:d9b229e20da8182950d223ee66838ddca3b7892fde1e15d0c339b4d8c782c46c

Observation 85473601-cf8d-4695-8dca-4cf78e7f79b5 · outbound

This paper cites u ller-Stich, J \.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding u ller-Stich, J \

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:46:52.649867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-16T11:46:51.421451Z digest=sha256:8752fb04866b1dba875706c769f867bccb6e4a3653177e211a94928928b16209

Observation 46aea96a-929d-49e2-ae25-7d588c81a4bc · outbound

This paper cites An Introduction to Vision-Language Modeling.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding An Introduction to Vision-Language Modeling

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.425958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.425958Z digest=sha256:8fc87ce9ef299341933fbef946752fc07c92df615f2ed2e227dded00cb7f31e7

Observation 3cc28d30-d7b8-4649-b8f4-24f7a5d66950 · outbound

This paper cites HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.430911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.430911Z digest=sha256:1d06e1903cd09b8124f75d795296b4ae20f7d34afdbf7b33c75f8d24d8ecc841

Observation e423e9c0-476d-403b-82d7-b5daeec0d261 · outbound

This paper cites an unresolved cited work.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:46:52.633884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-16T11:46:51.435727Z digest=sha256:399fc3ad0b49d443e102b81fa887646a6b980af709622ff9cb26bb896010add2

Observation 99622553-f194-4d49-967b-e7c914f775e1 · outbound

This paper cites an unresolved cited work.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:46:52.618623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-16T11:46:51.440156Z digest=sha256:072ce93dbc7d78887b96feecbad9f7c25957e727661a346a7a1ec229b5936f4c

Observation 129c777f-5038-4a13-9e7e-47c0359761da · outbound

This paper cites an unresolved cited work.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.444497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.444497Z digest=sha256:58aad13a35b905660054434a6652c6a7d3cfe24d8e7e28ed15a10bc181b5f8f1

Observation 47921d9e-6a4d-404a-a888-e7ec066748c4 · outbound

This paper cites an unresolved cited work.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:46:52.593839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-16T11:46:51.449259Z digest=sha256:2c5013c7c2f0761807f0683218b833b38a82c00c68a327381c7236e6e5f94ca9

Observation cbf078c3-78a1-4d47-93e9-0dc50fdaeece · outbound

This paper cites PathVQA: 30000+ Questions for Medical Visual Question Answering.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding PathVQA: 30000+ Questions for Medical Visual Question Answering

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.453558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.453558Z digest=sha256:fd77f3c343c86d0c6e7bca97b15f9f72e871ce1603e6e873f7f51e83492eb369

Observation 7ffc3b07-fc5a-45f4-af46-373ae5d45b29 · outbound

This paper cites an unresolved cited work.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:46:52.578451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-16T11:46:51.458377Z digest=sha256:706dfdf50581239644020e0234c9748285831e8b31d4593e383cea5f6519b5de

Observation b19c06ad-1b1b-47b4-b698-e1b79db09802 · outbound

This paper cites Modality-Fair Preference Optimization for Trustworthy MLLM Alignment.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Modality-Fair Preference Optimization for Trustworthy MLLM Alignment

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.462900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.462900Z digest=sha256:6512547cbb4b53052263cd9a5a9a1679a12f9ec224cd892d57c34882855c1a51

Observation 4a98cedc-6069-4e54-9f7b-75007ca62099 · outbound

This paper cites Joint Visual and Text Prompting for Improved Object-Centric Perception with Multimodal Large Language Models.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Joint Visual and Text Prompting for Improved Object-Centric Perception with Multimodal Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.467928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.467928Z digest=sha256:c9b81e94002fc0e9297928a27674969189294efcf9011567d958e9233dfaeb4d

Observation 41251ba2-9e7a-423c-a589-896771bed1e7 · outbound

This paper cites an unresolved cited work.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:46:52.563088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-16T11:46:51.472193Z digest=sha256:87df3f8e590ba12d67c772b7e93b219e285b4b849cc3ced65cc41dfa4672de1c

Observation 8061ecdd-0f3f-4259-afc3-7cdbe5bd2928 · outbound

This paper cites an unresolved cited work.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.476441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.476441Z digest=sha256:276c90bd276355f31421f8d1230d766ba10e0bc1bb4cb44f92263a024c016139

Observation f3c0102d-e380-459a-9f12-eb67e8f64ddc · outbound

This paper cites an unresolved cited work.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:46:52.537932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-16T11:46:51.480552Z digest=sha256:6808dd64efb18afd92a908e38b6070713d479403f6d7cddd58e74e77bc5790d6

Observation 445db077-f930-4e18-9523-cd54e4e81cc0 · outbound

This paper cites Surgical-LLaVA: Toward Surgical Scenario Understanding via Large Language and Vision Models.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Surgical-LLaVA: Toward Surgical Scenario Understanding via Large Language and Vision Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.489661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.489661Z digest=sha256:3e0bd11153bbce5e34da680faacbcd1073a1e939b9f4a7fc66a7ed4a0823a684

Observation d8079ac0-bb31-41c1-93d5-8cb38fb5be6c · outbound

This paper cites Scaling Laws for Neural Language Models.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Scaling Laws for Neural Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.493981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.493981Z digest=sha256:bd83413c1c77ab50f0d233e258f85ded86411cef4d1982cb5f977d4b7c095226

Observation c3fad696-d559-4ef4-90eb-315cfb8816de · outbound

This paper cites an unresolved cited work.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:46:52.522165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-16T11:46:51.498779Z digest=sha256:db4c0628bcc9931ab4d96225e8e9b7c3465bc41a475f2018d070dc3d08c9d7a0

Observation 880137e8-d257-40af-8e02-3f8b419a374e · outbound

This paper cites an unresolved cited work.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.503346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.503346Z digest=sha256:5521707a99635e56e5a499559eead10c88f9437eb8b33b0c449c9b0135eadd03

Observation a247ab58-3b92-45bc-9c1b-563f93474780 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding LLaVA-OneVision: Easy Visual Task Transfer

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.507731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.507731Z digest=sha256:b9efcf21a80b22e70cb6e295b0d1c1d982ccde4a2a6f3c385eb5f2da3a1d8275

Observation 4294b7a9-2e0b-4480-89da-ec3068770c4b · outbound

This paper cites an unresolved cited work.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.512265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.512265Z digest=sha256:00b38d75720cea7acb3ff0ec68115652dc9934ab96c7c5a75e89310ac52ee362

Observation a6efbd20-0b0f-4c28-ac49-da7213e4d2a3 · outbound

This paper cites LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.516703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.516703Z digest=sha256:6cf27ca9b5180b41e619d7ad5d875d2a448efc56f7129d65971d3bb6d1fd253b

Observation abedcc29-b2ac-4c41-bb81-2cd3bc196867 · outbound

This paper cites VideoChat: Chat-Centric Video Understanding.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding VideoChat: Chat-Centric Video Understanding

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.521189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.521189Z digest=sha256:4690f011b5cc03e19556caf0ae61a4dfdee31517b2a6641b6b14718864fb9f9c

Observation f27b6792-f78c-4c0f-956d-46f334a9fb36 · outbound

This paper cites VisualBERT: A Simple and Performant Baseline for Vision and Language.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding VisualBERT: A Simple and Performant Baseline for Vision and Language

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.525912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.525912Z digest=sha256:76281def4400771647d6886de4e28a5da20d6421cd360db4204fef27461cc371

Observation d123188f-9aed-4257-85f1-26171be7f3bd · outbound

This paper cites Self-supervised vision-language pretraining for Medical visual question answering.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Self-supervised vision-language pretraining for Medical visual question answering

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.530611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.530611Z digest=sha256:040d40e793fce1f4fb206d5603c0e679ef3f2b0b6f02ac73b542c46b07025eec

Observation 52d4a6be-8b38-47ae-a8e4-713e4cb6032f · outbound

This paper cites Video-LLaVA: Learning United Visual Representation by Alignment Before Projection.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Video-LLaVA: Learning United Visual Representation by Alignment Before Projection

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.535205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.535205Z digest=sha256:55b44d47fdd27c69ddc88391e802a9c502e1858107e00f97fc98b1c295f562af

Observation b51dd739-d043-44dd-ac9a-a997be74eb9e · outbound

This paper cites an unresolved cited work.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.539743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.539743Z digest=sha256:5d4d365ffcf990010df6a19e2cc32ccdcf5196b55956658d73b03c3032932ce0

Observation 891913a4-0a1b-4293-b455-e7fb80613926 · outbound

This paper cites an unresolved cited work.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.544027Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.544027Z digest=sha256:0ef9851b77ee22aadbac59fa011dbda1b85468dace1200cac1a0f33d14da0160

Observation 016dd5bb-7b64-41d1-817b-3968f8444382 · outbound

This paper cites MedCoT: Medical Chain of Thought via Hierarchical Expert.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding MedCoT: Medical Chain of Thought via Hierarchical Expert

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.548058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.548058Z digest=sha256:b2c4b54756a3744905648a78bb05511f789aa0a4d34dd2d920de4e17ef927e9b

Observation af705958-0c4a-4d16-9e1c-7759fcdff00c · outbound

This paper cites Q2ATransformer: Improving Medical VQA via an Answer Querying Decoder.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Q2ATransformer: Improving Medical VQA via an Answer Querying Decoder

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.552772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.552772Z digest=sha256:e0becf5d9c738c0fd3d993f48e0a0a6062fd3426c8faf391e65f0d6f83f1bfd3

Observation d792189c-4168-452b-8f20-831cb981e679 · outbound

This paper cites Lu, Bowen Chen, Drew F.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Lu, Bowen Chen, Drew F

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:46:52.467308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-16T11:46:51.557339Z digest=sha256:41046662e543f977ad63716d0dd731a9d83103fe7e80d072360933e59b8c32df

Observation a9fd584e-2e29-4709-8c72-afbc60c0953e · outbound

This paper cites an unresolved cited work.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.561591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.561591Z digest=sha256:5d082d38fb62edbe743f799fd8c56c04c06806fba0f4724071d5aead246836de

Observation 4aca6703-b020-48a2-82f3-1af67ba97a5d · outbound

This paper cites Med-Flamingo: a Multimodal Medical Few-shot Learner.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Med-Flamingo: a Multimodal Medical Few-shot Learner

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.566281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.566281Z digest=sha256:5df5c138ba96f771c2487ff48519a28875e99721aa9516c709fbd20e4a754e15

Observation e214ff20-32ca-4bd6-9e93-4fd44a8a0e5b · outbound

This paper cites an unresolved cited work.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:46:52.443683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-16T11:46:51.570745Z digest=sha256:e4838c6e8f8fd6c5277c6f27931e11796f3c93ba66ab9e0f71935041f86fabd7

Observation 1281c12e-45e1-45ba-96a1-a83219c76d43 · outbound

This paper cites an unresolved cited work.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.578509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.578509Z digest=sha256:4bfb07bfdb3f3f2da47b5f0f1bb3ce090684ea516ae55df7675dcb02346624ba

Observation b91606fa-6c32-4c2e-b85f-6df8d59d2845 · outbound

This paper cites Medical Vision Generalist: Unifying Medical Imaging Tasks in Context.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Medical Vision Generalist: Unifying Medical Imaging Tasks in Context

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.583250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.583250Z digest=sha256:15190ea8bda15c58cff963ffad5b0fbd6f6bf7bc292892603a0880b1fded84c7

Observation bc804ba7-f25d-437d-81b9-65f6ca16ddf4 · outbound

This paper cites an unresolved cited work.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:46:52.418248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-16T11:46:51.587682Z digest=sha256:8551dcd7b0049b2f09725cc95695fabe55cabdaaacc90292e16674b5cb7c6d35

Observation 894e5a1c-bafa-4369-95d9-557089bf93e8 · outbound

This paper cites an unresolved cited work.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:46:52.403215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-16T11:46:51.591981Z digest=sha256:2870ea79d25f2f9b24e32567e5b1b00bef34dd35b8dda7f934a5a1641ff2f09b

Observation 9f2f69b8-501f-4fc7-b67f-f6cae2e29ce8 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.596373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.596373Z digest=sha256:a52aa36212e5647fef83621069fd3cf7cf0c00ce73a607e3a427a9d432bc4a2f

Observation ef600fc8-8111-4b5c-836b-d18500d4759e · outbound

This paper cites an unresolved cited work.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:46:52.388412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-16T11:46:51.600748Z digest=sha256:96c347d2cb85e9c079e094a88b35fc53e47510dfb41a6988063ec228112d39f9

Observation 147b83d0-170a-43fb-9f6f-ef804797ba15 · outbound

This paper cites Open-Ended Medical Visual Question Answering Through Prefix Tuning of Language Models.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Open-Ended Medical Visual Question Answering Through Prefix Tuning of Language Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.605152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.605152Z digest=sha256:9a2bf68f9fc4c64259c7f65ebc245f12d95570788e46297f95f02ddf0e23e3d9

Observation aa667c15-1875-4530-812e-e6c220c7083c · outbound

This paper cites an unresolved cited work.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:46:52.373153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-16T11:46:51.609691Z digest=sha256:6e28337bf24151823715c4cb944c8de0115f65b1fd8783323ad1c3cc80a15386

Observation 66fc39e6-4c0f-4ba2-9ccb-3cf476bb99ed · outbound

This paper cites an unresolved cited work.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:46:52.357827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-16T11:46:51.728314Z digest=sha256:c43a52aaff6dad11ed2d09151e7bc7b3c1f239bce2289d370b15a4da832c3944

Observation 8dba16dd-bb6b-4fc5-bcf3-3b3eff5c7592 · outbound

This paper cites Towards Generalist Foundation Model for Radiology by Leveraging Web-scale 2D&3D Medical Data.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Towards Generalist Foundation Model for Radiology by Leveraging Web-scale 2D&3D Medical Data

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.737894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.737894Z digest=sha256:7604f99d2f0b73935cff264499b59e1f4bad17d2ff8cc7c1f81947b91924d29c

Observation 0538dda7-a145-446c-9625-220e24645c94 · outbound

This paper cites NExT-GPT: Any-to-Any Multimodal LLM.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding NExT-GPT: Any-to-Any Multimodal LLM

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.742015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.742015Z digest=sha256:474f6e092b053efc3a2f29fd5bd475cbdd1b6775290fc94d456454d932c31875

Observation 5a44c963-ee0e-43fe-808c-fe84d4131982 · outbound

This paper cites an unresolved cited work.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Unresolved cited work

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.746610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.746610Z digest=sha256:093bc49ee2caa7dd96d1485034e4eb1b26107ff368fb180995f1f3ce6f1f3a8e

Observation 41fd8756-c43d-4b1e-ab0d-0a0161ecd82c · outbound

This paper cites Weerasinghe, Bill J Wright, Ari Robicsek, Brian D.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Weerasinghe, Bill J Wright, Ari Robicsek, Brian D

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:46:52.333598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-16T11:46:51.750904Z digest=sha256:41b0a4bb8f4224ef2414956a1c07f33f63f74e256ffa26b6caf16c848e962c72

Observation f425a2cf-0ac0-49c5-8a44-34d37f650abc · outbound

This paper cites an unresolved cited work.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:46:52.317879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-16T11:46:51.755194Z digest=sha256:05ab5d1891c44cb2829eec8cb8141b7045466f8b35127d336c4ee9765fbc5e8e

Observation 415b84be-3dee-48bf-808f-6a8d1feada7a · outbound

This paper cites an unresolved cited work.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:46:52.301882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-16T11:46:51.759404Z digest=sha256:96c68eb9fddbe6cbe352fa4f71f2ff5694b3dad87cc3c3dd894f90babe7fe356

Observation 63a02901-1b02-40d1-9d74-594410b8adf4 · outbound

This paper cites an unresolved cited work.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Unresolved cited work

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.763804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.763804Z digest=sha256:d9d7f9b0f6e0dd8c0fe920dd8cf3578c24ddbea9fcb9b1fa0e217c4acdac9f5a

Observation f0979304-9292-416c-b8d4-96e7553b5ea6 · outbound

This paper cites an unresolved cited work.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:46:52.276482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-16T11:46:51.768092Z digest=sha256:edc0a475c33e55c051c42c201a8a422ff347cceb7b511f6d4252674ce4f270b4

Observation df89d656-3339-4546-b06e-e44465cc6c24 · outbound

This paper cites AnyGPT: Unified Multimodal LLM with Discrete Sequence Modeling.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding AnyGPT: Unified Multimodal LLM with Discrete Sequence Modeling

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.772308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.772308Z digest=sha256:729c4e8be92379456a218719fc7a275b604155688ecbc85751b62b08876a67b7

Observation 1fcbaa1a-1135-43a4-b20a-dd7855de4b20 · outbound

This paper cites VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.777135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.777135Z digest=sha256:75fcba0a8a2c287e7fc7c397af6b1a856e75b2e70333e2d50d1f09a17c4a55e2

Observation 42e9329e-b62f-4801-9b93-395e5ef14711 · outbound

This paper cites BiomedGPT: A Generalist Vision-Language Foundation Model for Diverse Biomedical Tasks.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding BiomedGPT: A Generalist Vision-Language Foundation Model for Diverse Biomedical Tasks

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.781263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.781263Z digest=sha256:3673a58c68d7a7b6696d5b9c67879b21bece415eb3689b4a01daffeff8c45091

Observation aeea8383-4523-4fd7-8bed-28cb1ed4b43b · outbound

This paper cites an unresolved cited work.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Unresolved cited work

Reference 64

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:46:52.261116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-16T11:46:51.785710Z digest=sha256:d3d2b8c5e6422624c496d554a09f28b11fb949b2899537544414b8bd708b28b5

Observation 7e091883-7430-48d6-9484-6dc50edfd904 · outbound

This paper cites BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:51.790201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:51.790201Z digest=sha256:4e88efc1aa82a21248dcd90a7dadb428f342e0e57bff6030bf0cd0b3c817822d

Observation fe172a27-0067-4fa3-987c-0103f155c531 · outbound

This paper cites Chia, Siegfried K.

OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding Chia, Siegfried K

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:46:52.246447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-16T11:46:51.794871Z digest=sha256:daf590bbc7f889eb3900a1cdcf554e81be32db55923a6dcd97d1faa4a59280fb

Pith citing papers

Observation 56a0951a-4f6d-4b33-9402-ac866a615bd6 · inbound

HSCR: Hierarchical Self-Contrastive Rewarding for Aligning Medical Vision Language Models cites this paper.

HSCR: Hierarchical Self-Contrastive Rewarding for Aligning Medical Vision Language Models OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T12:00:48.841437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:00:48.841437Z digest=sha256:eaec92c7b954739415de32ddb53d003cf013bbcfe09dd7de3add4398b27c35d9

Observation 27c740b8-85ae-4dc4-80d7-f9830c07c77f · inbound

Fast or Slow? Integrating Fast Intuition and Deliberate Thinking for Enhancing Visual Question Answering cites this paper.

Fast or Slow? Integrating Fast Intuition and Deliberate Thinking for Enhancing Visual Question Answering OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:02:10.533963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:02:10.533963Z digest=sha256:5aa9998a52005b6e78881a5b6bb5e4755c5c83758f8095ef8678aa4d8428a595

Observation 2d9ee975-fa18-4d4f-be38-500cc7621c67 · inbound

V2T-CoT: From Vision to Text Chain-of-Thought for Medical Reasoning and Diagnosis cites this paper.

V2T-CoT: From Vision to Text Chain-of-Thought for Medical Reasoning and Diagnosis OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T23:11:45.867138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:11:45.867138Z digest=sha256:c8cae039abeb5580d310b3b4ac75f88b8111622ff565172638cee71c88f0e7b0

Observation 99b54aae-8589-477b-8e8c-086b0b51fb87 · inbound

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning cites this paper.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.125834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.125834Z digest=sha256:20084bcf4b1ce5cf3bd08efbd7387427588ad32d6886f4c10433982d0321271a

Observation 50fc61aa-7acc-4978-972a-8de4ba7509f6 · inbound

UniReason-Med: A Shared Grounded Reasoning Interface for 2D-to-3D Transfer in Medical VQA cites this paper.

UniReason-Med: A Shared Grounded Reasoning Interface for 2D-to-3D Transfer in Medical VQA OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding

Reference 235

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T09:47:59.584324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-06-27T10:21:12.782864Z digest=sha256:f0138e4b047c4e5b826ed30baf0869955afe0e8950ad008c9c907d18a6afe523

Observation 9d2c6cdc-44ad-42c9-972e-3a9b068e2c8b · inbound

AtomiMed: Hierarchical Atomic Fact-Checking for Universal Clinical-Aware Medical Report Evaluation cites this paper.

AtomiMed: Hierarchical Atomic Fact-Checking for Universal Clinical-Aware Medical Report Evaluation OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T11:55:43.152392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-07-01T03:27:41.497269Z digest=sha256:51028531755ca3e4a0fb5b09ad97152ee7f10787d764a5df0015ebacfbd2d9e0

Observation 88f672f8-129d-45c8-81fd-461c72b4e864 · inbound

MedStreamBench: A Time-Aware Benchmark for Streaming and Proactive Medical Video Understanding cites this paper.

MedStreamBench: A Time-Aware Benchmark for Streaming and Proactive Medical Video Understanding OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-03T16:38:39.531041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-07-03T16:37:09.666491Z digest=sha256:832342fd373b0b8e0ade388fd36c736611d5bc9446594779834dba649c85e220

Observation 17209b59-2074-45cf-bd6b-58347f3f807c · inbound

CARVE: Cross-Slice Anisotropic Reallocation of Visual Evidence for Efficient 3D Medical Volume Understanding cites this paper.

CARVE: Cross-Slice Anisotropic Reallocation of Visual Evidence for Efficient 3D Medical Volume Understanding OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T14:41:24.473812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:41:24.473812Z digest=sha256:95a44ee972394d5c35ad8325b74cfe60f05670742d838915553c43a495c1ac04