Pith. sign in

Paper Citation Record · LEDGER

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG

As of 9 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 1 inbound Pith citation observation for arXiv:2508.06496.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.06496 v1

Coverage vector

measured 48 of 48 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:50:54.486096Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-14T06:30:16.612345Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

48 of 48 outbound references displayed

  • verified exact1
  • verified fuzzy38
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 14e827d4-3a7f-4295-8a7b-82a5f0ff9681 · outbound

This paper cites The inception team at vqa-med 2020: Pre- trained vgg with data augmentation for medical vqa and vqg.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG The inception team at vqa-med 2020: Pre- trained vgg with data augmentation for medical vqa and vqg

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:55.068303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.306983Z digest=sha256:fa66ee8602491089eaeef27b38859f86c373bc2bfb9b3482cf6d784ed35c7a92

Observation eb557038-88f4-4a43-b0ee-c850e2ccf7de · outbound

This paper cites Flamingo: a visual language model for few-shot learning.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Flamingo: a visual language model for few-shot learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T15:50:54.312689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:50:54.312689Z digest=sha256:c7b029c34dd2fb8be63ecaec26d8e6ed48aa2903e1228577e45d8b0c92d4429c

Observation 8602554d-77e0-436c-94d1-8e251e02e49a · outbound

This paper cites Bottom-up and top-down attention for image captioning and visual question answering.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Bottom-up and top-down attention for image captioning and visual question answering

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:55.051579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.317920Z digest=sha256:51f60fe01b7e10aab6649b438d76531322d1aa886cc1c22e033b848b59678625

Observation 80d00b1a-e995-4d4e-b5be-d985abfe1037 · outbound

This paper cites Vqa: Visual question answering.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Vqa: Visual question answering

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:55.040451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.321588Z digest=sha256:d077a19800d1039b986b51eb02faceea2cced79ef28f33259757c8a338313aec

Observation 6a4c2e4a-1e8d-441f-a620-affcf1635601 · outbound

This paper cites A-okvqa: A dataset for adap- tive open knowledge visual question answering.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG A-okvqa: A dataset for adap- tive open knowledge visual question answering

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:55.029805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.325609Z digest=sha256:24325e360b0b412ac151e3655c4cf0b6d28327f44ee5accc4e69821ec8fc29f6

Observation e792d4f3-7a7e-4d15-96f2-df33914590e8 · outbound

This paper cites Ocr-vqa: A dataset for visual question answering with text in the wild.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Ocr-vqa: A dataset for visual question answering with text in the wild

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:55.019110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.329110Z digest=sha256:66d80170986b975b80c88b43cd833f9e90fa12fa47d4d512b9d03e6d1895bcd8

Observation cb724ada-2609-4830-b5cc-349b420eec7c · outbound

This paper cites Okvqa: A dataset for open knowledge visual question answering.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Okvqa: A dataset for open knowledge visual question answering

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:55.008024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.334117Z digest=sha256:8b12bb288ce35ea4bdfabe15c64779b0e180f79812faaf5f37cdb08cbb9236a6

Observation d71c85ab-0fa7-44fa-a174-c2d049d6783a · outbound

This paper cites Veagle: Advancements in multimodal rep- resentation learning.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Veagle: Advancements in multimodal rep- resentation learning

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.997357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.337541Z digest=sha256:ea46428e2179cbc2a37856cf78f4696838110692e8b5cf1481f5bfe6b5c08d02

Observation 4f6b2003-8a87-4d07-904a-7beaaa28f0c6 · outbound

This paper cites Visualgpt: Data-efficient adaptation of pretrained language models for image captioning.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Visualgpt: Data-efficient adaptation of pretrained language models for image captioning

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.986231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.341065Z digest=sha256:9233a83dc3f6d682cb6001164582a935ff332f90eefa5a9e7a3dc0daf0143737

Observation e7ccec12-a2b3-4400-957f-a558af30c738 · outbound

This paper cites Unsupervised Learning of KB Queries in Task-Oriented Dialogs.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Unsupervised Learning of KB Queries in Task-Oriented Dialogs

Reference 10

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T15:50:54.599287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.344783Z digest=sha256:d1b1ed882b1f6aff341badff0fab956786248741beb40e3d4f1d3d9e75a0688e

Observation b4f33e8a-ebaa-41b2-958b-5da2bf0b5877 · outbound

This paper cites Multi-level Chaotic Maps for 3D Textured Model Encryption.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Multi-level Chaotic Maps for 3D Textured Model Encryption

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-08-06T15:50:54.578734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.348419Z digest=sha256:63c779d552d52ab6e1bf2ac9cf21ab51fec1c86e805e52837aab0fd66363b459

Observation 8cae71b0-d663-4c17-b572-2f014ef36f99 · outbound

This paper cites From local to global: A graph rag approach to query-focused sum- marization, 2025.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG From local to global: A graph rag approach to query-focused sum- marization, 2025

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.975927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.352614Z digest=sha256:1a5cdea91f33f3776549e24a48eb1a6c6d655cb5c774f16e112f2acabe6a6fcd

Observation 39130238-8533-4adb-8d0d-1a2a8dd27a6d · outbound

This paper cites an unresolved cited work.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:50:54.966356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.356485Z digest=sha256:efe74950c69713bb92ff506061d2d0a075941e477c3043f4b76ae52d7c71bcb6

Observation dbbee4d6-7915-4b50-a854-cd9903bea7e6 · outbound

This paper cites Prompt learning with knowledge graphs for zero- shot relation extraction.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Prompt learning with knowledge graphs for zero- shot relation extraction

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.955411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.360598Z digest=sha256:0c8f17704d9d586a3018ac438473cf50aa6d7c5ab705d147d2d8973e8cfe601f

Observation da219817-ea59-44c6-9930-56e48ceb66f2 · outbound

This paper cites Imagebind: One embedding space to bind them all.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Imagebind: One embedding space to bind them all

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.945171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.363866Z digest=sha256:7cd0d49d2e3927aad142c0278c591ccfd2af733853983afa5f3024e448820573

Observation 3cb271ab-e6c5-424f-881e-0af75fd87ede · outbound

This paper cites Girshick.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Girshick

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.935163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.366895Z digest=sha256:2185719173bd25e158cdcbf0e9ccf0ca5eff148e69916add77f8b033a9fcac26

Observation 06590fca-c945-4c04-95dc-75762879fd07 · outbound

This paper cites MedAlpaca -- An Open-Source Collection of Medical Conversational AI Models and Training Data.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG MedAlpaca -- An Open-Source Collection of Medical Conversational AI Models and Training Data

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T15:50:54.370611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:50:54.370611Z digest=sha256:3a504e9bdc698db0d215d821b891aba4f18360dc99ab48b392c400e44a73351b

Observation 2612983a-d807-440f-95a2-f987135eebee · outbound

This paper cites Bliva: A simple multimodal llm for better handling of text-rich visual questions.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Bliva: A simple multimodal llm for better handling of text-rich visual questions

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.923922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.374634Z digest=sha256:585802cb77880af002426a353f688a1723f989a1398aa68ad439120314dab35c

Observation b568daf0-229e-4f1d-9114-669fb58e0758 · outbound

This paper cites Interpretable medical image visual question answering via multi-modal relationship graph learning.Med- ical Image Analysis, 97:103279, 2024.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Interpretable medical image visual question answering via multi-modal relationship graph learning.Med- ical Image Analysis, 97:103279, 2024

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.914096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.378530Z digest=sha256:18355f7e60871369468f4a026b2b952cf85c469c454cd11b338f07a379528f7a

Observation cc70cdb8-06c0-4033-b745-1084bcd88cc1 · outbound

This paper cites Retrieval-augmented gener- ation for knowledge-intensive nlp tasks.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Retrieval-augmented gener- ation for knowledge-intensive nlp tasks

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.903043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.382745Z digest=sha256:8738633420cc1baa229e858c8a096ba24c1993ab71c702967b15b5d889a7c23d

Observation 8f399d90-b07c-4872-86ac-fd52636dae0e · outbound

This paper cites Llava-med: Training a large language- and-vision assistant for biomedicine in one day.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Llava-med: Training a large language- and-vision assistant for biomedicine in one day

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.891926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.386519Z digest=sha256:3a6da4fed813adddceb9a90a59c9ef837ae644d560d802a7e02fbb7d70c3a8f3

Observation ff4757b2-8f03-42d8-80f1-8cd705676a68 · outbound

This paper cites Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.881829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.390443Z digest=sha256:4ae7bc3194faf6daed651c69a1d20d5bee4a232ab53827db9661d8b11c0b1235

Observation 11c4b357-ce72-453d-8fb9-f29579305448 · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.870066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.394082Z digest=sha256:6289802108abb89b19cc36b02bf34014c74a029369789c4718adb37c80cf96c5

Observation 3f73eeb8-3a04-4743-ac04-452e461aecbd · outbound

This paper cites Masked vision and language pre-training with uni- modal and multimodal contrastive losses for medical visual question answering.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Masked vision and language pre-training with uni- modal and multimodal contrastive losses for medical visual question answering

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.858316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.399164Z digest=sha256:382c61295afd5f96efab2cabeecb93a0d3f9587e30627bbe3f965f6a124cfdb2

Observation 3cc3c98d-0afc-4647-826f-f6afb737aacb · outbound

This paper cites Lawrence Zitnick.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Lawrence Zitnick

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.846804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.402739Z digest=sha256:a76ef2e02a8cc30efed2ef3fd38712946b7d4913eff8bf0551467e681154727a

Observation 004cbde5-821d-475d-9258-b497ccce5156 · outbound

This paper cites Visual instruction tuning.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Visual instruction tuning

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.834938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.406219Z digest=sha256:547eadbe04f6f325517babe0407f52095d59be71862b8ad0b72c7935e8b86380

Observation 02c580d1-817a-4479-8799-9b05251950e6 · outbound

This paper cites Can we Trust Chatbots for now? Accuracy, reproducibility, traceability; a Case Study on Leonardo da Vinci's Contribution to Astronomy.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Can we Trust Chatbots for now? Accuracy, reproducibility, traceability; a Case Study on Leonardo da Vinci's Contribution to Astronomy

Reference 27

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T15:50:54.548150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.409579Z digest=sha256:a1b3c1915761fac9f57957b466ef1ccde3c2fc8eec2ee49c083dfe61db8d8d48

Observation 79a35448-1e64-4128-be24-e8998fcdfbd9 · outbound

This paper cites Foundation models for generalist medi- cal artificial intelligence.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Foundation models for generalist medi- cal artificial intelligence

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.822316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.413049Z digest=sha256:6ac330f4b43957f46a9d9c5d196857957e16fed534a2ed9af840209bdb723296

Observation df2e7e2e-9c2a-46e5-bc64-df92b5195c67 · outbound

This paper cites Med-flamingo: a multimodal medical few-shot learner.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Med-flamingo: a multimodal medical few-shot learner

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.809268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.417280Z digest=sha256:0113b51d2cb9d026ae4205b04f045883b4b8183ccbef6c3c6d15a0ce1eb3c4da

Observation 5e96866b-a842-4045-a505-371176785f11 · outbound

This paper cites K-pathvqa: Knowledge-aware multimodal rep- resentation for pathology visual question answering.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG K-pathvqa: Knowledge-aware multimodal rep- resentation for pathology visual question answering

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.797037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.420725Z digest=sha256:f02d937cfb789b5bd2643f839760bf52b475af48e505fba4df1eb3bcd76398a0

Observation 44b3a3d5-27fd-4a0a-bc36-d7e7d952c21d · outbound

This paper cites Overcoming data limi- tation in medical visual question answering.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Overcoming data limi- tation in medical visual question answering

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.785426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.423859Z digest=sha256:e0c12e7a8243bb731c0d4b00f894614494f5ba6332d8d920ab96deb887965658

Observation c237a9e1-a7c8-4057-b873-569db0cf5b05 · outbound

This paper cites St-vqa: Visual question answering with a focus on scene texts.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG St-vqa: Visual question answering with a focus on scene texts

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.773459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.427548Z digest=sha256:ba7137df7f6e4f70bf955221b61aad42c5898bb9ff8de9495569d76897b56d6c

Observation 2c23d8ba-83b4-42c1-8a05-38b2adcde634 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Learning transferable visual models from natural language supervi- sion

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T15:50:54.430867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:50:54.430867Z digest=sha256:0bb240cfe6dbb7d6b56fd0563da7b678bf395812b693b206edd28fcf2bc1da3c

Observation a36ae1f0-dd3f-49b4-87ba-1cfe568394dd · outbound

This paper cites Docvqa: A dataset for document visual question answer- ing.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Docvqa: A dataset for document visual question answer- ing

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.754225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.434466Z digest=sha256:3a05fc6c63312a6c19eb08409a2e68f95041963e2c43f0009888d0097ad56d57

Observation bf596faf-5546-48a2-b1bf-d2465cda53ee · outbound

This paper cites Maivar-t: Multimodal audio-image and video action recognizer using transformers.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Maivar-t: Multimodal audio-image and video action recognizer using transformers

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.742797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.438125Z digest=sha256:9ecfe9cb025248cb53edc11c223f5f41c286afebf940618af50fe1f9a54d67d7

Observation 549747e7-e4be-4ec6-9b43-30f5f64eeb55 · outbound

This paper cites Sharma, A.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Sharma, A

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.733101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.441123Z digest=sha256:afe59cd7173e1cb6cf98265cec912c2ed8ddef1985d891aeb06a7fc92bca5747

Observation a02c14d9-bb8f-41b0-9b45-e0a1cba96ad5 · outbound

This paper cites Multimodal few-shot learning with frozen language models.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Multimodal few-shot learning with frozen language models

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.722054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.444663Z digest=sha256:f3ebb21b080f6661438273606e77bb1b14ed2e8a513d6aa196ff4492c78a0351

Observation 347f422c-576d-4397-a796-06d37ee99ffa · outbound

This paper cites Attention is all you need.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Attention is all you need

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.711077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.447854Z digest=sha256:9d010bf1b62a1ebedea2d22809ca6656166e7760ca89bf340db4392e6b29c3fa

Observation a1a75ff0-74c1-46f4-ba75-805f6df4ab5a · outbound

This paper cites Medical graph rag: Towards safe medical large language model via graph retrieval-augmented generation.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Medical graph rag: Towards safe medical large language model via graph retrieval-augmented generation

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.699090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.451614Z digest=sha256:d07b7d820722bde97327a87632475614c600a73cf10b239e8e35ec2f5cdd9d53

Observation 92cc7a48-8678-4f67-8c0b-8788de69aa39 · outbound

This paper cites Mmed-rag: Versatile multimodal rag system for medical vi- sion language models.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Mmed-rag: Versatile multimodal rag system for medical vi- sion language models

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.687549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.454992Z digest=sha256:a8403f76fd6eaaf380fbdcbc910f607662faade54e9121e85b38b71b98ead1e3

Observation 6b28f0b1-c14a-4bd1-af65-fe1ac6c9db2d · outbound

This paper cites Rule: Reliable multimodal rag for factuality in medical vision language models.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Rule: Reliable multimodal rag for factuality in medical vision language models

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.676197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.458130Z digest=sha256:013df79aae7ab231504487033c48a113fa2bfb58a05c0fb657f123cc4894d8e6

Observation 60dfb2fe-7e4f-4ab4-b54e-b52444cd22fe · outbound

This paper cites Ramm: Retrieval-augmented biomedical visual question answering with multi-modal pre-training.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Ramm: Retrieval-augmented biomedical visual question answering with multi-modal pre-training

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.664309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.461891Z digest=sha256:785ab916868b07dad295d5ffe92f5131662f44c4733c4de0ff09def4452b3830

Observation 58fe69e3-9993-41a6-b5dd-69ff91bacd1d · outbound

This paper cites A generalist vision–language foundation model for diverse biomedical tasks.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG A generalist vision–language foundation model for diverse biomedical tasks

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.650767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.465268Z digest=sha256:8078a9f86212dd74909df806bc8e1ad5433849449a518c8593a783c36b25b05f

Observation e7c6b56e-e8e5-425d-836e-adad4a8b1ab4 · outbound

This paper cites BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T15:50:54.469168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:50:54.469168Z digest=sha256:f23c54be5f1d4e5e50f9c987d8d7f119dae866e316009383c5ad7ced57e512bf

Observation e7eee508-72dc-4667-9893-68d86374efbf · outbound

This paper cites Multimodal representation learning by alternating uni- modal adaptation.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Multimodal representation learning by alternating uni- modal adaptation

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.638093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.474110Z digest=sha256:f0d6713187ccd7571990c8633deadb357eabd2b0f15dd8a505da61f2a3067f41

Observation 409a3c91-1642-4f8a-a0c8-d3c77c495ad1 · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T15:50:54.477871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:50:54.477871Z digest=sha256:5a4a59f35b193297c8cc3682db312524974d9dd11b0fc486c171101ffd80345e

Observation f91b943d-e1e0-46c1-89c2-96937c131f6d · outbound

This paper cites an unresolved cited work.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:50:54.625452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.482486Z digest=sha256:552882f31393ab724b8685283a553c407c24c8d3135c18fd4cc447952c2ec730

Observation 93c184ca-0c8a-4628-bf3f-73fa882a2742 · outbound

This paper cites Crossclr: Cross-modal contrastive learning for multi-modal video representations.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Crossclr: Cross-modal contrastive learning for multi-modal video representations

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.612976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:50:54.486096Z digest=sha256:3a3685018fb9630cc1576750e1b4b3933a1087462f29ad4ce19f3cb8337a32a0

Pith citing papers

Observation f7428869-4c5b-484c-9168-979a7864cfe6 · inbound

The Path to Self-Evolving Clinical Systems: Scaling Medical Agents from Assistance to Autonomy cites this paper.

The Path to Self-Evolving Clinical Systems: Scaling Medical Agents from Assistance to Autonomy Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG

Reference 164

Resolution
unresolved
no resolver link, observed 2026-07-14T06:30:16.612345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:30:16.612345Z digest=sha256:ac0f131339c65fa2086d334a988247de9acea0d3569c9ac74c0994f93bd13f6c