Pith. sign in

Paper Citation Record · LEDGER

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning

As of 9 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 0 inbound Pith citation observations for arXiv:2508.18687.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.18687 v1

Coverage vector

measured 32 of 32 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T16:21:43.229838Z

measured 32 of 32 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

32 of 32 outbound references displayed

  • verified exact0
  • verified fuzzy7
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0d56dbef-d2c6-4400-93da-305577108b73 · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.069676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.069676Z digest=sha256:ebf84daff613849e41303fc3f49974ceabc49abbd69d915f8e7a5076297e880b

Observation 874730ed-2367-469f-bf19-a068bfbae59d · outbound

This paper cites Bioengineering (2023).

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning Bioengineering (2023)

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T16:21:43.750709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T16:21:43.076104Z digest=sha256:4d820e9f1eb42200c88e21eaa1c6ee7ccd41c72d285096156e30d148b4d56d99

Observation 81d137b3-ba22-44e8-8a5c-0dda87616818 · outbound

This paper cites Stable LM 2 1.6B Technical Report.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning Stable LM 2 1.6B Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.081686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.081686Z digest=sha256:d8bd41cc229243a9e6a9a3ab1625569c7f5155c6bb7d1faf4526b0eb233e3361

Observation 856c2817-485b-4fd5-93b5-fd73e42c9f16 · outbound

This paper cites HuatuoGPT-o1, Towards Medical Complex Reasoning with LLMs.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning HuatuoGPT-o1, Towards Medical Complex Reasoning with LLMs

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.088270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.088270Z digest=sha256:2c9ba84d18f21cdbc722952d99755f80a7c7499776e5cec6c339c7ff536ff27c

Observation 654ea1b8-186c-4fda-b34c-6ae2f4e6c556 · outbound

This paper cites HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.094255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.094255Z digest=sha256:4af165e4dfc5576bbd816866314ae00c235bd1dbd9c55ab994ae8240dbc659c3

Observation 35404bd6-df0c-4732-b267-bbc2486977b6 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (June 2019).

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (June 2019)

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T16:21:43.736485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T16:21:43.099994Z digest=sha256:30b3529358fac60f7c53d28b600d1d024f644f99dcee85806a67d3121c147d77

Observation 4fcd4a0c-82f5-4899-97cb-2edef63093a3 · outbound

This paper cites an unresolved cited work.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:21:43.721162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T16:21:43.105329Z digest=sha256:8cb87c220388bc990467d83c86fd8efe8ebf5c28096c0f17550694e23daf2517

Observation a2af4b97-a9d2-4d81-bc67-fef2b11bbc37 · outbound

This paper cites PathVQA: 30000+ Questions for Medical Visual Question Answering.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning PathVQA: 30000+ Questions for Medical Visual Question Answering

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.109874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.109874Z digest=sha256:ff10fc3185a43e70d0968e7d5302daa1ca4d503903794bc020ce5e99b992c229

Observation af194243-1dd1-4209-8d07-1ff654613c4f · outbound

This paper cites GPT-4o System Card.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning GPT-4o System Card

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.114782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.114782Z digest=sha256:26e744ecf46ccb08aa0620dc1c2b22353e193703163165fbc171b9150bc59e19

Observation ded7381b-d726-490c-96a9-da1e5a36551c · outbound

This paper cites CAPO: Reinforcing Consistent Reasoning in Medical Decision-Making.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning CAPO: Reinforcing Consistent Reasoning in Medical Decision-Making

Reference 10

Resolution
metadata mismatch
local_arxiv, observed 2026-08-05T16:21:43.493258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T16:21:43.119934Z digest=sha256:de6ad342458eafa3aac8f7e2d293f9b7c4f848f1891a4e4e286f10e1723727d0

Observation 99b54aae-8589-477b-8e8c-086b0b51fb87 · outbound

This paper cites OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.125834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.125834Z digest=sha256:ffa779c001a59019df296eb022370c271c0d759c4394ad231ad28a6d361a5d80

Observation 842011c5-775f-46b1-9dac-e03e853ca6aa · outbound

This paper cites Modality-Fair Preference Optimization for Trustworthy MLLM Alignment.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning Modality-Fair Preference Optimization for Trustworthy MLLM Alignment

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.131742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.131742Z digest=sha256:b9869b3dad27e1d6188a1c0a346f400925e4714d7cfdb1f1b48666225745443b

Observation 824ab91f-d14b-4f32-a7e4-eed74b56c12e · outbound

This paper cites HSCR: Hierarchical Self-Contrastive Rewarding for Aligning Medical Vision Language Models.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning HSCR: Hierarchical Self-Contrastive Rewarding for Aligning Medical Vision Language Models

Reference 13

Resolution
metadata mismatch
local_arxiv, observed 2026-08-05T16:21:43.441693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T16:21:43.136812Z digest=sha256:72fcf4e7a7e00f59c6d9c25894fa74c68a86653054ad5a2fc58447a3df338a0c

Observation 94835c0f-3e84-4c54-8a0f-e2081b6cdb4f · outbound

This paper cites In: Find- ingsoftheAssociationforComputationalLinguistics:EMNLP2024.pp.3843–3860 (2024).

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning In: Find- ingsoftheAssociationforComputationalLinguistics:EMNLP2024.pp.3843–3860 (2024)

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T16:21:43.704488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T16:21:43.141682Z digest=sha256:045667218e0500a0935ad31ce337c82d6d6187e657da6a3d67d3dc900700d9d7

Observation f77f716f-6263-4232-bdff-30f7e95c68e4 · outbound

This paper cites Advances in neural information processing systems33, 18661–18673 (2020).

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning Advances in neural information processing systems33, 18661–18673 (2020)

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.146007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.146007Z digest=sha256:0c28e63160e686a544607ce71ab5d180e9cbc7dd1e9a2c5563f944e41e9604d5

Observation d715c7cf-4939-4695-a727-68261835492d · outbound

This paper cites Scientific data 5(1), 1–10 (2018).

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning Scientific data 5(1), 1–10 (2018)

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.150464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.150464Z digest=sha256:d71f58e6e6ea333ca995d3dcabfce6cfec17d7388518adb7979a4ca2b4888bd9

Observation 0572ec1c-ded5-419b-8a4b-4299eeab4699 · outbound

This paper cites Advances in Neural Information Processing Systems36 (2024).

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning Advances in Neural Information Processing Systems36 (2024)

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T16:21:43.669282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T16:21:43.154662Z digest=sha256:8cecd56f3c5f51e0fee4b31872d4056ece68e88dfa0043045c442a1cc9219390

Observation 5fbe4744-d153-480e-a805-bed249f89ba5 · outbound

This paper cites Self-supervised vision-language pretraining for Medical visual question answering.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning Self-supervised vision-language pretraining for Medical visual question answering

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.159467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.159467Z digest=sha256:7f6d13b50ed411eaba453030bac577ee13c413e41c70bfe4cd3aed725b5336af

Observation 70745962-1e3d-4423-ac72-54c6a5d21985 · outbound

This paper cites IEEE (2021).

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning IEEE (2021)

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T16:21:43.654054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T16:21:43.164536Z digest=sha256:42be82d62894ccb0e8a3425ce4cfca8797d3f7a8ef3553c630c91484f5bf3bc4

Observation f214512f-ec9d-428b-ab54-bfc31d4fa80e · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR).

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.169297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.169297Z digest=sha256:b07582adf0e6c7b7a580c352bf8f7a1a7a0cad6707f31e112bb76cd924e3cdbc

Observation 9408c02f-6a75-459a-bb0a-cf1fc51d0c4a · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T16:21:43.629787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T16:21:43.175304Z digest=sha256:4dbdcfdec03f024ed44eda792235d7859767fb4b5561dd3816b5c74353e1b68e

Observation ca3a2c91-7e22-4437-a36e-e40dd505295a · outbound

This paper cites MedCoT: Medical Chain of Thought via Hierarchical Expert.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning MedCoT: Medical Chain of Thought via Hierarchical Expert

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.179799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.179799Z digest=sha256:98644b4ca254d85d5af549857bfd8988cba84199f36732af6647878f054e695e

Observation 74af216d-72f2-47fe-9127-c179c3ad6d18 · outbound

This paper cites Q2ATransformer: Improving Medical VQA via an Answer Querying Decoder.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning Q2ATransformer: Improving Medical VQA via an Answer Querying Decoder

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.184749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.184749Z digest=sha256:c05ca8e05cf5ffea99fc595c9d7bda2b7ee82a83cf58239e788a71f784f748be

Observation 34b7e499-6ab9-435f-949f-d673064c32be · outbound

This paper cites In: Machine Learning for Health (ML4H).

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning In: Machine Learning for Health (ML4H)

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.189927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.189927Z digest=sha256:3e19d958794e03829bcaf72b42c6b122e3746874c5f42bc3a54000e328495a60

Observation b21ae348-288f-4af6-a30b-26bb5361d0fb · outbound

This paper cites Sunny and Dark Outside?! Improving Answer Consistency in VQA through Entailed Question Generation.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning Sunny and Dark Outside?! Improving Answer Consistency in VQA through Entailed Question Generation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.194719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.194719Z digest=sha256:e308d624d8e89382ea8975aecbbf718c1b709bbaf8ea0ddf24edbcd7dd7c198e

Observation 694e967a-fb44-4d47-90f9-98f13fa09ce7 · outbound

This paper cites Towards Expert-Level Medical Question Answering with Large Language Models.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning Towards Expert-Level Medical Question Answering with Large Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.199298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.199298Z digest=sha256:343fbb4b98e233b95b1ad91fb3cdbad4e7027ae3e6170e43d2fcf7cc0ec80d30

Observation c46c0616-a5cb-43c2-888f-13cef66c9cd9 · outbound

This paper cites Open-Ended Medical Visual Question Answering Through Prefix Tuning of Language Models.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning Open-Ended Medical Visual Question Answering Through Prefix Tuning of Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.204129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.204129Z digest=sha256:e828a672d653a6ca22743eb9c46c60b7c40fcf239fa205703d20f46f45a47d0a

Observation 1378a18a-d0e6-461a-bcc4-512d825f694d · outbound

This paper cites STLLaVA-Med: Self-Training Large Language and Vision Assistant for Medical Question-Answering.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning STLLaVA-Med: Self-Training Large Language and Vision Assistant for Medical Question-Answering

Reference 28

Resolution
metadata mismatch
local_arxiv, observed 2026-08-05T16:21:43.325896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T16:21:43.209676Z digest=sha256:9decde804f0d008e3a66e48489f0e502e05cce5c04b056f11056ce5eafadde7b

Observation e873239e-c954-40da-ad86-1aa1f2e1c26a · outbound

This paper cites In: European Conference on Computer Vision.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning In: European Conference on Computer Vision

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T16:21:43.604315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T16:21:43.214578Z digest=sha256:d026303e172a836ebd6066add95a579be0871b3d52d1e4fc193ed78f30177fd4

Observation f02ac44f-b821-4240-910b-61c7fd4ba075 · outbound

This paper cites BiomedGPT: A Generalist Vision-Language Foundation Model for Diverse Biomedical Tasks.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning BiomedGPT: A Generalist Vision-Language Foundation Model for Diverse Biomedical Tasks

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.219267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.219267Z digest=sha256:7202a1722b2ecbf896891575479545787f236d05dc3900f4a14b49c3cf4ef04c

Observation 0681a122-b310-45ab-8758-383c3c0fb0bf · outbound

This paper cites BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.224387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.224387Z digest=sha256:618082675bb2d03ee484d954f37d13f6ed3a1622d1d8b149db472649f654ca5b

Observation ea78d15b-2a13-4fa5-b3dd-391b32e2c7ee · outbound

This paper cites PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.229838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.229838Z digest=sha256:16144fbbb9ae09b9bd71420f66e1236278f03fe3c657772f45eb06254c4f6b24

Pith citing papers

No inbound Pith citation observations are available.