Pith. sign in

Paper Citation Record · LEDGER

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images

As of 7 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 0 inbound Pith citation observations for arXiv:2608.03322.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.03322 v1

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T21:07:09.893269Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

49 of 49 outbound references displayed

  • verified exact2
  • verified fuzzy23
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8139d1e2-3c3a-4309-a652-b50b58240e49 · outbound

This paper cites International conference on medical image computing and computer-assisted intervention , pages=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images International conference on medical image computing and computer-assisted intervention , pages=

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:14.440464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:07:07.291303Z digest=sha256:a92fabfe096a8e4489c91c8198f8c58214687178d0687043e06c56a039c26cf0

Observation 798da038-d359-4dd3-bc2b-29e39aaafa7f · outbound

This paper cites International Conference on Medical Image Computing and Computer-Assisted Intervention , pages=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images International Conference on Medical Image Computing and Computer-Assisted Intervention , pages=

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:14.235802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:07:07.323838Z digest=sha256:9a2ad5973dbdf94ac8056219ff965cbda57c2b4ed11bf4bcfebb283fe0218b3f

Observation 34ec7a43-e77e-4a12-864b-b65bc4ffd6fd · outbound

This paper cites IEEE Transactions on Pattern Analysis and Machine Intelligence , year=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images IEEE Transactions on Pattern Analysis and Machine Intelligence , year=

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:14.039872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:07:07.354778Z digest=sha256:bf58546978f4de9de9aa426d260560f0daca8bc8b8671a9f46bd7c24290be761

Observation f3ceb8d4-476c-4986-b790-39f15fb39ec7 · outbound

This paper cites LocateAnything: Fast and High-Quality Vision-Language Grounding with Parallel Box Decoding.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images LocateAnything: Fast and High-Quality Vision-Language Grounding with Parallel Box Decoding

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:07.387536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:07.387536Z digest=sha256:35e85763d477c2dc8e0653fced56c4f279e80e9d289649e57e7f075146cf549a

Observation 6ec4edd4-a18a-4843-b5a9-59152b7a7b83 · outbound

This paper cites Nature communications , volume=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Nature communications , volume=

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:13.872592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:07:07.433529Z digest=sha256:4b3364e2439d402e027525f140f720ddf973feacc70935f3b77dcb45ce65aba0

Observation 4f2736d8-74e2-4d5b-8d6c-91256a884cde · outbound

This paper cites Nature methods , volume=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Nature methods , volume=

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:13.738062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:07:07.468330Z digest=sha256:2ce30eb9568ace774a1e4134a107c3bc9bc31684502e2baa58e60fad8ae2cb05

Observation 59b2ecb1-120e-491d-b221-857adb44898f · outbound

This paper cites international conference on medical image computing and computer-assisted intervention , pages=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images international conference on medical image computing and computer-assisted intervention , pages=

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:13.575409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:07:07.508242Z digest=sha256:bef374f69122e7b444eb1dc5c88e6f43eb050190bfe7e992e9c8485e9f243178

Observation 57e88c4d-9892-4f82-997b-395e7c3a7058 · outbound

This paper cites MedMoE: Modality-Specialized Mixture of Experts for Medical Vision-Language Understanding.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images MedMoE: Modality-Specialized Mixture of Experts for Medical Vision-Language Understanding

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:07.540354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:07.540354Z digest=sha256:70a4408bd7e48019307ac294ba0aa7419881feccc362692a418e1610f6c1a8fd

Observation f388ae50-5da8-4c7b-828c-16459a5034e2 · outbound

This paper cites Rex-Thinker: Grounded Object Referring via Chain-of-Thought Reasoning.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Rex-Thinker: Grounded Object Referring via Chain-of-Thought Reasoning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:07.567191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:07.567191Z digest=sha256:4b2ff022efdf572fd41877cd7182b3fd411182bfa9068c1455dc9ec44da47b28

Observation a2fcd220-4af4-4ced-9811-a1969e258422 · outbound

This paper cites arXiv preprint arXiv:2510.04477 , year=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images arXiv preprint arXiv:2510.04477 , year=

Reference 10

Resolution
verified exact
raw_fallback, observed 2026-08-05T21:07:10.819578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:07:07.598807Z digest=sha256:067a3f962abdd0eaebefa5390a467c6fadae0ec66eeac814c54fb4cd2c8d4614

Observation 2f6eb89d-c272-4b9b-87fc-d06fe82f5a8f · outbound

This paper cites Journal of medical imaging , volume=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Journal of medical imaging , volume=

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:13.464628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:07:07.657756Z digest=sha256:62bf73eb74e7be32558cbe48913ea8f7a7f45520fe4443def8f4b68062a2eea4

Observation 544d608b-642c-49db-8be5-36ac71182503 · outbound

This paper cites Proceedings of the IEEE/CVF international conference on computer vision , pages=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Proceedings of the IEEE/CVF international conference on computer vision , pages=

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:07.693581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:07.693581Z digest=sha256:9d01d39fcc149f05af5e9a9336d06aed3f8e2a2fe8bd0922803535f989476014

Observation 7f3499f1-877e-4ae3-bcf6-7a721c8371a0 · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:13.334089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:07:07.735063Z digest=sha256:ea05c8b25e02939092661221a6421b6be7772f29af1a25b30b63f9c7ba7fe247

Observation 2bd2698e-bf62-4d58-94e7-124c85a2665d · outbound

This paper cites European conference on computer vision , pages=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images European conference on computer vision , pages=

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:13.172153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:07:07.778837Z digest=sha256:a01367b4bfb3bc6419eedd5535efb06d3ea7df888d8a094ba3856a5f59899568

Observation 6a4eae3b-817b-4988-a3fd-315e68536dc9 · outbound

This paper cites European conference on computer vision , pages=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images European conference on computer vision , pages=

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:07.802293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:07.802293Z digest=sha256:f33d5e1d4d7a6163bfed169fcad171de1c1ea46763e1c0c63704e8503bf8a802

Observation c8c46346-8a68-4dcd-8c56-1a320ba32416 · outbound

This paper cites International Conference on Learning Representations , volume=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images International Conference on Learning Representations , volume=

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:07.838697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:07.838697Z digest=sha256:eca47afb5b4702c8786f13a0c2c0a5d91ceea40d354b1e3865f5ab61d971e68f

Observation a378364e-1e5e-42da-acba-05ca18c631c6 · outbound

This paper cites arXiv preprint arXiv:2510.12798 , year=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images arXiv preprint arXiv:2510.12798 , year=

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:07.882591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:07.882591Z digest=sha256:2500f7ff2784212c893373c6efa929a7f4843a89d36ce755088c4abf2fb42d3b

Observation 085a2332-2867-4624-a043-da1bf3de5719 · outbound

This paper cites 2026 IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) , pages=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images 2026 IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) , pages=

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:12.986260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:07:07.931795Z digest=sha256:9c219bd12c8bd2a9bd424fb5c93c6092936f8a889d37a546dc8a60a7cd0ba187

Observation 110f6385-b605-4c39-94a6-bc57dde51b2b · outbound

This paper cites VividMed: Vision Language Model with Versatile Visual Grounding for Medicine.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images VividMed: Vision Language Model with Versatile Visual Grounding for Medicine

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:08.007254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:08.007254Z digest=sha256:07a56cfb82c1b159831c2b9d7dcd17b379a1b31e7251c0fd8eb9cb3c7be070be

Observation 49ad6160-661a-4c88-b080-3bc11726618b · outbound

This paper cites Proceedings of the Computer Vision and Pattern Recognition Conference , pages=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Proceedings of the Computer Vision and Pattern Recognition Conference , pages=

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:12.882338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:07:08.069248Z digest=sha256:1df70c191aff2aff91e17e84277bd312535912747dbf93a91abc9ce85896795d

Observation 92b64d50-bfb0-4007-8452-c2d84cb062ac · outbound

This paper cites Proceedings of the AAAI Conference on Artificial Intelligence , volume=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Proceedings of the AAAI Conference on Artificial Intelligence , volume=

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:12.776827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:07:08.158591Z digest=sha256:ea69d9999bb6d3501e209700956621115ac6839d1405c09498441730cb0b1c91

Observation 539ca9fe-d91c-4f3f-8426-223cdba068c4 · outbound

This paper cites Nature Communications , year=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Nature Communications , year=

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:12.672341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:07:08.218339Z digest=sha256:655cf53788dd18068fae00e10edcc16c767e22e16797fefbc310af85a064c107

Observation 3652678f-b974-46c2-a861-824a124d0c48 · outbound

This paper cites European conference on computer vision , pages=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images European conference on computer vision , pages=

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:08.287875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:08.287875Z digest=sha256:b51dea0a031eead28539244e0ef9f516846b4fb272d9edf0c8695cc2db9c0a2d

Observation d0f5a677-cf76-4f15-b909-147757ebb773 · outbound

This paper cites arXiv preprint arXiv:2601.06847 , year=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images arXiv preprint arXiv:2601.06847 , year=

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:08.385821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:08.385821Z digest=sha256:cea8b71f8c4543048d968bebdf9662fb3e79e3cd9a1912aff198cba0956d5792

Observation bfc923ed-8082-4118-8eb6-30463cbad083 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Advances in Neural Information Processing Systems , volume=

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:12.461937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:07:08.421805Z digest=sha256:1376f7d9ec69d6cf4738c2a64511956556d7d3c1bd0c8cf19308438cd575f2b0

Observation c88710ae-06f9-4c58-bc6f-b9a7e8c8a542 · outbound

This paper cites UniVG-R1: Reasoning Guided Universal Visual Grounding with Reinforcement Learning.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images UniVG-R1: Reasoning Guided Universal Visual Grounding with Reinforcement Learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:08.465735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:08.465735Z digest=sha256:5d695893efa59331a0e165090889f27d3865a6f06f3dda274f3c2d7798726b5a

Observation 0021fea5-eb31-41b4-b5db-4ee7239a1c2d · outbound

This paper cites Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:12.310068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:07:08.521244Z digest=sha256:78ae4a46ea7c4b0ca31894be0444ee6a3bf5ae8db3118424d49a8380013e5bb1

Observation 7baefc8b-2614-44e4-84bb-7fa50f48fa78 · outbound

This paper cites Better Eyes, Better Thoughts: Why Vision Chain-of-Thought Fails in Medicine.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Better Eyes, Better Thoughts: Why Vision Chain-of-Thought Fails in Medicine

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:08.563485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:08.563485Z digest=sha256:d85e04bea13de4fb65306dc80f9bf9bdb03ac2ff2151cfb79d9bdd55c91bc86e

Observation 0cdb2e9e-7bca-4770-9a48-2ba6af965c1b · outbound

This paper cites Nature communications , volume=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Nature communications , volume=

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:08.634558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:08.634558Z digest=sha256:545a5220e114a7a5616b97800a2e69cb30a73c997784c62270942c838cb57584

Observation 98196b98-4ce3-4d47-94f3-30ac11d264a5 · outbound

This paper cites Radiology: Artificial Intelligence , volume=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Radiology: Artificial Intelligence , volume=

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:08.695541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:08.695541Z digest=sha256:b35172a2d2d13677f55186bfafece17ef9f0ada967bfb9909450c36bb219a2f9

Observation b9446222-422b-47a8-bf0e-47cd25906429 · outbound

This paper cites Medical image analysis , volume=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Medical image analysis , volume=

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:08.748748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:08.748748Z digest=sha256:fa58f4ec9c162ab87ff6703255c6c6ae0e406854fdbfce86d01c8fde0a961f2b

Observation e28eec6a-38e8-4e3f-b8bf-f709c7521296 · outbound

This paper cites International Conference on Medical Image Computing and Computer-Assisted Intervention , pages=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images International Conference on Medical Image Computing and Computer-Assisted Intervention , pages=

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:12.099143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:07:08.786746Z digest=sha256:f0bab988603fd01a99665300d9a35aae712cb3e2edf49aacca7509793b445090

Observation 3728f14f-26b9-4c6b-b8b5-5b7838536633 · outbound

This paper cites Medical physics , volume=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Medical physics , volume=

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:11.934478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:07:08.815007Z digest=sha256:cca76d4ae5e842aeed7f8d172fd5fb54f5ee6f955ab1ea155cc086a5762a2af2

Observation 886c2177-2804-4377-a1c5-8790a793c558 · outbound

This paper cites Scientific data , volume=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Scientific data , volume=

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:11.770374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:07:08.851524Z digest=sha256:9cd54c4808af4c7c197126594ebfca0ab019656028ca372f8267cef849eaa203

Observation 160b1582-f6be-48de-a60d-5ff7ccd42fd2 · outbound

This paper cites Scientific data , volume=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Scientific data , volume=

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:11.537733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:07:08.877954Z digest=sha256:d53b08e930bf9b6c725a227f92c282ee1a42b24d9e266e9fa47fea739ab0c768

Observation 53f19895-266c-44f3-b1cd-03a41cd2adbe · outbound

This paper cites arXiv preprint arXiv:2305.19112 , year=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images arXiv preprint arXiv:2305.19112 , year=

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:08.919218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:08.919218Z digest=sha256:7ef40b0a75df0a2171f3c8725e704f173857b2626575204ba90fed58da8202ca

Observation c8535cf8-1809-4a3e-8a35-262a22cdbf41 · outbound

This paper cites International conference on multimedia modeling , pages=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images International conference on multimedia modeling , pages=

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:08.948707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:08.948707Z digest=sha256:6aa6b118a76dd296eb07250ddfbfbb93fd5b95c27439f08086c47b729092dd9e

Observation ee625d8d-b7fe-43bd-bcf2-062ab53d7f6f · outbound

This paper cites saliency maps from physicians , author=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images saliency maps from physicians , author=

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:11.392202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:07:08.996716Z digest=sha256:e2f744dac82ce7d9840290b561feb923e9eddcaaaab81abdb5f74af2ab7a647f

Observation 7cade63a-ecd3-4fc1-9294-c121ba9e450b · outbound

This paper cites Information Sciences , volume=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Information Sciences , volume=

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:11.308346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:07:09.028296Z digest=sha256:c08acb601db626b09dcad21cad9aaf006bcfc334b47f34e6355ff94b24c84220

Observation c60c39ba-7f7b-4009-bbea-c50fe9b3cc2b · outbound

This paper cites IEEE transactions on medical imaging , volume=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images IEEE transactions on medical imaging , volume=

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:11.227731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:07:09.065811Z digest=sha256:e0fa0a0db2eeca5ec7e067e6d4ddd4d8bce8d303fe5aa3b58d2e367ed5b110c4

Observation 5dfb962e-91c8-4480-b4ce-152c824890a8 · outbound

This paper cites IEEE Transactions on Medical imaging , volume=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images IEEE Transactions on Medical imaging , volume=

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:11.062611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:07:09.098336Z digest=sha256:9d3d0d0d1f3893be927bf805cba12004c7552614f732d4fa7c513c84c9ecef1c

Observation efe2f606-036d-4080-978e-e841ff6e7743 · outbound

This paper cites European conference on computer vision , pages=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images European conference on computer vision , pages=

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:09.170037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:09.170037Z digest=sha256:b10f3be1b1f10d42fb4d614422b4b94e834c26d02c5ccbc53ca68e83c0cb6118

Observation 9d5ab858-d88d-4c63-bfa5-96bcbb3c514d · outbound

This paper cites Ultralytics YOLO26: Unified Real-Time End-to-End Vision Models.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Ultralytics YOLO26: Unified Real-Time End-to-End Vision Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:09.213500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:09.213500Z digest=sha256:0f710bd33ad43894864297df979229f327ea6395e15fe5f5eb8abe7fbb0857f9

Observation a9cd5b55-8b47-41bd-be3f-a25d42f77648 · outbound

This paper cites LW-DETR: A Transformer Replacement to YOLO for Real-Time Detection.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images LW-DETR: A Transformer Replacement to YOLO for Real-Time Detection

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:09.302031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:09.302031Z digest=sha256:297e47f981407192c6d168df844268c479873464f3325f978de4dbadc6baf658

Observation 6c6cbf1d-23fb-4ccb-85cf-f5d563015e11 · outbound

This paper cites arXiv preprint arXiv:2603.18739 , year=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images arXiv preprint arXiv:2603.18739 , year=

Reference 45

Resolution
verified exact
raw_fallback, observed 2026-08-05T21:07:10.299260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:07:09.384132Z digest=sha256:acbc53ab1fbb63f498e3512b5c97ccf54d1b27c26db9677350657c62afafa2cf

Observation cb508499-4c1a-41a4-a672-9f366aa90408 · outbound

This paper cites DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:09.449564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:09.449564Z digest=sha256:ff0f709a214e8df095c0976414106350304e79dc5d190a2200550f034a365c7c

Observation db45a968-61e3-45d5-b5c7-a97ac6f009bd · outbound

This paper cites arXiv preprint arXiv:2512.17436 , year=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images arXiv preprint arXiv:2512.17436 , year=

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:09.599566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:09.599566Z digest=sha256:dee86b85a324fea3c27a8d4c55e30363c696f29992c01c1dfce73b338fd4e303

Observation 07a1f85b-7386-431d-b18e-40ed5284516e · outbound

This paper cites Ovis2.5 Technical Report.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Ovis2.5 Technical Report

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:09.717152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:09.717152Z digest=sha256:518e215ffcc6b284aa90424239272407310f54e4b18c1d891901790f55b821c6

Observation 6ebcfbae-2f97-459c-9f74-b717d95ccad2 · outbound

This paper cites Qwen3-VL Technical Report.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Qwen3-VL Technical Report

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:09.893269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:09.893269Z digest=sha256:5474485bd6cd8f95b189c11e711aefa283d3ee58cc949b82286d28775ddcf943

Pith citing papers

No inbound Pith citation observations are available.