Pith. sign in

Paper Citation Record · LEDGER

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG

As of 8 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 1 inbound Pith citation observation for arXiv:2508.06496.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.06496 v1

Coverage vector

measured 48 of 48 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:50:54.486096Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-14T06:30:16.612345Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

48 of 48 outbound references displayed

  • verified exact1
  • verified fuzzy38
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 14e827d4-3a7f-4295-8a7b-82a5f0ff9681 · outbound

This paper cites The inception team at vqa-med 2020: Pre- trained vgg with data augmentation for medical vqa and vqg.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG The inception team at vqa-med 2020: Pre- trained vgg with data augmentation for medical vqa and vqg

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:55.068303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.306983Z digest=sha256:639f562bf05ef9bd4d014d813e0ff14a044c0d39b6bef6ff423b430df7230464

Observation eb557038-88f4-4a43-b0ee-c850e2ccf7de · outbound

This paper cites Flamingo: a visual language model for few-shot learning.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Flamingo: a visual language model for few-shot learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T15:50:54.312689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:50:54.312689Z digest=sha256:2bf239bbeb156b8ee40268aa7889d24d46f1cdc4670dc77af3af94f707d9fa70

Observation 8602554d-77e0-436c-94d1-8e251e02e49a · outbound

This paper cites Bottom-up and top-down attention for image captioning and visual question answering.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Bottom-up and top-down attention for image captioning and visual question answering

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:55.051579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.317920Z digest=sha256:0d0248705e617796441ada9cab6fdda3acfc97e6574c8ebb2c7320edd4d14d94

Observation 80d00b1a-e995-4d4e-b5be-d985abfe1037 · outbound

This paper cites Vqa: Visual question answering.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Vqa: Visual question answering

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:55.040451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.321588Z digest=sha256:dbc2d95d7dcc794a4c8333e1e578eb5af7a61e54ded4d42b4def2bcec52f055d

Observation 6a4c2e4a-1e8d-441f-a620-affcf1635601 · outbound

This paper cites A-okvqa: A dataset for adap- tive open knowledge visual question answering.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG A-okvqa: A dataset for adap- tive open knowledge visual question answering

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:55.029805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.325609Z digest=sha256:5005938d01f90c9d71c0f0e08e9e369a1ebdaee8aecaf33e006da062eb313da3

Observation e792d4f3-7a7e-4d15-96f2-df33914590e8 · outbound

This paper cites Ocr-vqa: A dataset for visual question answering with text in the wild.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Ocr-vqa: A dataset for visual question answering with text in the wild

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:55.019110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.329110Z digest=sha256:6ac75b0adfb7c592fd667db4ce893229bd3fddf4c7a120eabe4d64e9fa1ce0b5

Observation cb724ada-2609-4830-b5cc-349b420eec7c · outbound

This paper cites Okvqa: A dataset for open knowledge visual question answering.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Okvqa: A dataset for open knowledge visual question answering

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:55.008024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.334117Z digest=sha256:6d53fa3e245296bee6cadec929ca6930ddd773263969a14098c439292ec86206

Observation d71c85ab-0fa7-44fa-a174-c2d049d6783a · outbound

This paper cites Veagle: Advancements in multimodal rep- resentation learning.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Veagle: Advancements in multimodal rep- resentation learning

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.997357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.337541Z digest=sha256:9eb424b48c53df730144c4154bd0fb0c78df05aa09706ea0ad6d011721a77bf9

Observation 4f6b2003-8a87-4d07-904a-7beaaa28f0c6 · outbound

This paper cites Visualgpt: Data-efficient adaptation of pretrained language models for image captioning.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Visualgpt: Data-efficient adaptation of pretrained language models for image captioning

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.986231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.341065Z digest=sha256:7a8f1fbfb0f8c595f36923d93d5a498cdc6e6a533cf0df64849522e8b1b44dd4

Observation e7ccec12-a2b3-4400-957f-a558af30c738 · outbound

This paper cites Unsupervised Learning of KB Queries in Task-Oriented Dialogs.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Unsupervised Learning of KB Queries in Task-Oriented Dialogs

Reference 10

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T15:50:54.599287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.344783Z digest=sha256:ff8c97055155244eaf820adc56f5247d6c6d9515809f11ebb9aae3a788c3df97

Observation b4f33e8a-ebaa-41b2-958b-5da2bf0b5877 · outbound

This paper cites Multi-level Chaotic Maps for 3D Textured Model Encryption.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Multi-level Chaotic Maps for 3D Textured Model Encryption

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-08-06T15:50:54.578734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.348419Z digest=sha256:3c7982c25d4c44dd52f68de73f4892327d30ec221227475c2b2d5b7b93a83652

Observation 8cae71b0-d663-4c17-b572-2f014ef36f99 · outbound

This paper cites From local to global: A graph rag approach to query-focused sum- marization, 2025.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG From local to global: A graph rag approach to query-focused sum- marization, 2025

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.975927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.352614Z digest=sha256:e3a5d16c9fdaee0c81d216bd09c7985902992c8196240371d51200d4d993fe19

Observation 39130238-8533-4adb-8d0d-1a2a8dd27a6d · outbound

This paper cites an unresolved cited work.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:50:54.966356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.356485Z digest=sha256:14bdb3af2f4ff598e38c20eff39c4794f29bcc6a7d5a4755ebb450564acb6b83

Observation dbbee4d6-7915-4b50-a854-cd9903bea7e6 · outbound

This paper cites Prompt learning with knowledge graphs for zero- shot relation extraction.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Prompt learning with knowledge graphs for zero- shot relation extraction

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.955411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.360598Z digest=sha256:155525441abb195044d972514239fed3619d107b60f87f76192f05b12e3ba22b

Observation da219817-ea59-44c6-9930-56e48ceb66f2 · outbound

This paper cites Imagebind: One embedding space to bind them all.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Imagebind: One embedding space to bind them all

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.945171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.363866Z digest=sha256:847ff66e3acfc3c8d4ac25abcad349b132f694abbe7111ee37b3241c6bf407f7

Observation 3cb271ab-e6c5-424f-881e-0af75fd87ede · outbound

This paper cites Girshick.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Girshick

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.935163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.366895Z digest=sha256:50e02ae3a5d853e490b969aeb3cc590c16e97c32997b19d07fae180ba1bc2d5c

Observation 06590fca-c945-4c04-95dc-75762879fd07 · outbound

This paper cites MedAlpaca -- An Open-Source Collection of Medical Conversational AI Models and Training Data.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG MedAlpaca -- An Open-Source Collection of Medical Conversational AI Models and Training Data

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T15:50:54.370611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:50:54.370611Z digest=sha256:621a37021033ce35f57df3105bf65db2bf1005855dc9e2ba3ca1074859c30b64

Observation 2612983a-d807-440f-95a2-f987135eebee · outbound

This paper cites Bliva: A simple multimodal llm for better handling of text-rich visual questions.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Bliva: A simple multimodal llm for better handling of text-rich visual questions

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.923922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.374634Z digest=sha256:d071fcce8d8fea1a09e0559b0a535fd158ffb2716d33f1fdc1975ee5457bd34a

Observation b568daf0-229e-4f1d-9114-669fb58e0758 · outbound

This paper cites Interpretable medical image visual question answering via multi-modal relationship graph learning.Med- ical Image Analysis, 97:103279, 2024.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Interpretable medical image visual question answering via multi-modal relationship graph learning.Med- ical Image Analysis, 97:103279, 2024

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.914096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.378530Z digest=sha256:b9604100938c9a91438ff43daaf34dbefe8d7e0d9bbb1bcec5b5d9776b222c50

Observation cc70cdb8-06c0-4033-b745-1084bcd88cc1 · outbound

This paper cites Retrieval-augmented gener- ation for knowledge-intensive nlp tasks.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Retrieval-augmented gener- ation for knowledge-intensive nlp tasks

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.903043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.382745Z digest=sha256:9317f14ad608173b216016ec71609de1143e1dc74f27b0a187a6b223bc8e7a2a

Observation 8f399d90-b07c-4872-86ac-fd52636dae0e · outbound

This paper cites Llava-med: Training a large language- and-vision assistant for biomedicine in one day.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Llava-med: Training a large language- and-vision assistant for biomedicine in one day

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.891926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.386519Z digest=sha256:bff4e0951070e49edd45190aa14e313b7bfc0432824bb73aaf074145f1f293d8

Observation ff4757b2-8f03-42d8-80f1-8cd705676a68 · outbound

This paper cites Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.881829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.390443Z digest=sha256:5f04822190c97b3de7b3a00a80a95d3ccd2ca59fe16d0729c3f612db833aae7d

Observation 11c4b357-ce72-453d-8fb9-f29579305448 · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.870066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.394082Z digest=sha256:efd8f598372fdccae592b9e8a9516ebdfc63ea823573c0e3f61d66d87fd0b3f8

Observation 3f73eeb8-3a04-4743-ac04-452e461aecbd · outbound

This paper cites Masked vision and language pre-training with uni- modal and multimodal contrastive losses for medical visual question answering.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Masked vision and language pre-training with uni- modal and multimodal contrastive losses for medical visual question answering

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.858316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.399164Z digest=sha256:a5257535575ed98e0778200d2a8e659767095b14e97540f6d4bd9814a1705381

Observation 3cc3c98d-0afc-4647-826f-f6afb737aacb · outbound

This paper cites Lawrence Zitnick.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Lawrence Zitnick

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.846804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.402739Z digest=sha256:2517d66084a5ae7a0ce5977968b0e77d30c78e21427e4fddc608cfc7a3b0d4fd

Observation 004cbde5-821d-475d-9258-b497ccce5156 · outbound

This paper cites Visual instruction tuning.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Visual instruction tuning

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.834938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.406219Z digest=sha256:557cebc37d7aeb8c761c871c4ce30520b1469ec4b4192ec4876424759abaa425

Observation 02c580d1-817a-4479-8799-9b05251950e6 · outbound

This paper cites Can we Trust Chatbots for now? Accuracy, reproducibility, traceability; a Case Study on Leonardo da Vinci's Contribution to Astronomy.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Can we Trust Chatbots for now? Accuracy, reproducibility, traceability; a Case Study on Leonardo da Vinci's Contribution to Astronomy

Reference 27

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T15:50:54.548150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.409579Z digest=sha256:5af4a23867314eac85ee6a9c5b4833901d4d83c55b17af6a3362d83bdb42d69e

Observation 79a35448-1e64-4128-be24-e8998fcdfbd9 · outbound

This paper cites Foundation models for generalist medi- cal artificial intelligence.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Foundation models for generalist medi- cal artificial intelligence

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.822316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.413049Z digest=sha256:62dccb318e8c01b06a3c14ad399ea9b71c4c207c87976fab6f4784ceef616e03

Observation df2e7e2e-9c2a-46e5-bc64-df92b5195c67 · outbound

This paper cites Med-flamingo: a multimodal medical few-shot learner.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Med-flamingo: a multimodal medical few-shot learner

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.809268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.417280Z digest=sha256:ce8155d1290cd2511b56a0226d9f91e3982f6c56d347cc8dfeb3b33d1e03d0a0

Observation 5e96866b-a842-4045-a505-371176785f11 · outbound

This paper cites K-pathvqa: Knowledge-aware multimodal rep- resentation for pathology visual question answering.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG K-pathvqa: Knowledge-aware multimodal rep- resentation for pathology visual question answering

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.797037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.420725Z digest=sha256:f8943e33c85a81bfa7f047a5f9da639802f567d84973b06670433e3d1bcc663c

Observation 44b3a3d5-27fd-4a0a-bc36-d7e7d952c21d · outbound

This paper cites Overcoming data limi- tation in medical visual question answering.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Overcoming data limi- tation in medical visual question answering

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.785426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.423859Z digest=sha256:bec93ba0a84751ef695618239cc19dc0d0737eb8bff925ce1697d2739745bda3

Observation c237a9e1-a7c8-4057-b873-569db0cf5b05 · outbound

This paper cites St-vqa: Visual question answering with a focus on scene texts.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG St-vqa: Visual question answering with a focus on scene texts

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.773459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.427548Z digest=sha256:3b4e328827dc5c156127a106cd132964120635fa4b279d0b1a0cdcff1d292b97

Observation 2c23d8ba-83b4-42c1-8a05-38b2adcde634 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Learning transferable visual models from natural language supervi- sion

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T15:50:54.430867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:50:54.430867Z digest=sha256:43e37e26cd287bf5bdbc65b1ef221a0cd8a72d42af948612b291f4c3f6c3f7a8

Observation a36ae1f0-dd3f-49b4-87ba-1cfe568394dd · outbound

This paper cites Docvqa: A dataset for document visual question answer- ing.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Docvqa: A dataset for document visual question answer- ing

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.754225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.434466Z digest=sha256:e401986fd4a030dcfedb58b889d3899e00c4896085b5880126fe1dd49a6de0e9

Observation bf596faf-5546-48a2-b1bf-d2465cda53ee · outbound

This paper cites Maivar-t: Multimodal audio-image and video action recognizer using transformers.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Maivar-t: Multimodal audio-image and video action recognizer using transformers

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.742797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.438125Z digest=sha256:d54749642bffd18995ae6cc1fc2a31e01e0c7dbf21a0cd8a65a18a98096ab336

Observation 549747e7-e4be-4ec6-9b43-30f5f64eeb55 · outbound

This paper cites Sharma, A.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Sharma, A

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.733101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.441123Z digest=sha256:9b42b9cc217e66a14476c68fd332a33b516cba39278d9ce6f8fb72fe747a1d89

Observation a02c14d9-bb8f-41b0-9b45-e0a1cba96ad5 · outbound

This paper cites Multimodal few-shot learning with frozen language models.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Multimodal few-shot learning with frozen language models

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.722054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.444663Z digest=sha256:80bd019f5ce25729a1801f3ddfd1b02b1bfe0c76d10c28c6a6da79b881e8f524

Observation 347f422c-576d-4397-a796-06d37ee99ffa · outbound

This paper cites Attention is all you need.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Attention is all you need

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.711077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.447854Z digest=sha256:ad7e20091cfe1f160244b32adffc68fd16642e60f2fc10440e80fbbc4edb18a4

Observation a1a75ff0-74c1-46f4-ba75-805f6df4ab5a · outbound

This paper cites Medical graph rag: Towards safe medical large language model via graph retrieval-augmented generation.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Medical graph rag: Towards safe medical large language model via graph retrieval-augmented generation

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.699090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.451614Z digest=sha256:f70edd8c7e29f88f0abcecb70ab0a39a189a0f79c78e1254cc185cb82d5b1d7e

Observation 92cc7a48-8678-4f67-8c0b-8788de69aa39 · outbound

This paper cites Mmed-rag: Versatile multimodal rag system for medical vi- sion language models.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Mmed-rag: Versatile multimodal rag system for medical vi- sion language models

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.687549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.454992Z digest=sha256:34b6b21829cd992e2d5e8f27e11aebc17568bf06a72c01d45a668eb997e30150

Observation 6b28f0b1-c14a-4bd1-af65-fe1ac6c9db2d · outbound

This paper cites Rule: Reliable multimodal rag for factuality in medical vision language models.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Rule: Reliable multimodal rag for factuality in medical vision language models

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.676197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.458130Z digest=sha256:c3f2b617b20f9b6957d3e8ca42f8da37498b4bc15ffa54b1d9e11c1091e5af2a

Observation 60dfb2fe-7e4f-4ab4-b54e-b52444cd22fe · outbound

This paper cites Ramm: Retrieval-augmented biomedical visual question answering with multi-modal pre-training.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Ramm: Retrieval-augmented biomedical visual question answering with multi-modal pre-training

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.664309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.461891Z digest=sha256:88afdec4f1d62025084258cb4e6cadf745db7de4fdb4f19b2e11e34ede59f39c

Observation 58fe69e3-9993-41a6-b5dd-69ff91bacd1d · outbound

This paper cites A generalist vision–language foundation model for diverse biomedical tasks.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG A generalist vision–language foundation model for diverse biomedical tasks

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.650767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.465268Z digest=sha256:9a3600d1903316054f6f0a3b409872dd296b589beba91e9362471d4b894e6c4f

Observation e7c6b56e-e8e5-425d-836e-adad4a8b1ab4 · outbound

This paper cites BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T15:50:54.469168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:50:54.469168Z digest=sha256:c278ce980e918e180e87cc740aadfc62dfdc1359ef38f0473f7ff4ccb8d1dc84

Observation e7eee508-72dc-4667-9893-68d86374efbf · outbound

This paper cites Multimodal representation learning by alternating uni- modal adaptation.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Multimodal representation learning by alternating uni- modal adaptation

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.638093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.474110Z digest=sha256:a82c3a311846dafd58736d2edafa32250960f55e4bb60367ba7ec905f28eaa62

Observation 409a3c91-1642-4f8a-a0c8-d3c77c495ad1 · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T15:50:54.477871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:50:54.477871Z digest=sha256:c475345afe0d54d737a82e362a35a432cb667db4ac822b31088723908533e507

Observation f91b943d-e1e0-46c1-89c2-96937c131f6d · outbound

This paper cites an unresolved cited work.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:50:54.625452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.482486Z digest=sha256:afb871ff1a2cf525b45c7b6fad7424a9a032a11c028b280c4052828ce19e7ed5

Observation 93c184ca-0c8a-4628-bf3f-73fa882a2742 · outbound

This paper cites Crossclr: Cross-modal contrastive learning for multi-modal video representations.

Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG Crossclr: Cross-modal contrastive learning for multi-modal video representations

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:50:54.612976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:50:54.486096Z digest=sha256:5cfc680f791d25edb30c3a0ab7416a44e4542145c904915f20b6bfb6b6d40c61

Pith citing papers

Observation f7428869-4c5b-484c-9168-979a7864cfe6 · inbound

The Path to Self-Evolving Clinical Systems: Scaling Medical Agents from Assistance to Autonomy cites this paper.

The Path to Self-Evolving Clinical Systems: Scaling Medical Agents from Assistance to Autonomy Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG

Reference 164

Resolution
unresolved
no resolver link, observed 2026-07-14T06:30:16.612345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:30:16.612345Z digest=sha256:4a55746452e00fe4fa467f418c8d9a28bc44993c5b5ed655f20ed0da1b2f410f