Pith. sign in

Paper Citation Record · LEDGER

Frequency-Domain Dual-Branch Fusion for Medical Visual Question Answering

As of 17 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 0 inbound Pith citation observations for arXiv:2608.08307.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.08307 v1

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T00:13:39.491417Z

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

17 of 17 outbound references displayed

  • verified exact0
  • verified fuzzy13
  • unresolved4
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 016354c4-33f6-4e12-92ad-9660a9b5b785 · outbound

This paper cites Global Filter Networks for Image Classification.

Frequency-Domain Dual-Branch Fusion for Medical Visual Question Answering Global Filter Networks for Image Classification

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T00:13:40.109687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T00:13:39.054669Z digest=sha256:59091ffaa47497872defb4b2a36fde703d3186236a488d27be1ab0ece7834c9d

Observation 88d564f1-9523-44b3-a116-48c0d85aed22 · outbound

This paper cites Frequency Spectrum Is More Effective for Multimodal Representation and Fusion: A Multimodal Spectrum Rumor Detector.

Frequency-Domain Dual-Branch Fusion for Medical Visual Question Answering Frequency Spectrum Is More Effective for Multimodal Representation and Fusion: A Multimodal Spectrum Rumor Detector

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T00:13:40.101618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T00:13:39.158071Z digest=sha256:dc9934516fa99350aba6f43bde1faa1950c940852d7e665560e050f1b7c89806

Observation a24e571a-c1c3-4372-8040-e62d11523c91 · outbound

This paper cites FNet: Mixing Tokens with Fourier Transforms.

Frequency-Domain Dual-Branch Fusion for Medical Visual Question Answering FNet: Mixing Tokens with Fourier Transforms

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T00:13:40.093400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T00:13:39.220603Z digest=sha256:0486c2697a8ee0135236861790e874c8c7b2876b3289089d92f18cc800686345

Observation 4fc6e437-d83f-442c-b60d-e17f52e98944 · outbound

This paper cites Multi-Modal Masked Autoencoders for Medical Vision-and-Language Pre-Training.

Frequency-Domain Dual-Branch Fusion for Medical Visual Question Answering Multi-Modal Masked Autoencoders for Medical Vision-and-Language Pre-Training

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T00:13:40.086510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T00:13:39.240989Z digest=sha256:5993dbd880b71221c00da61ea0f35365f69b13a4b7eac9230732d820a350cc64

Observation 79effcc1-548b-44e1-aec0-ebadcae6f186 · outbound

This paper cites Self-Supervised Vision-Language Pretraining for Medical Visual Question Answering.

Frequency-Domain Dual-Branch Fusion for Medical Visual Question Answering Self-Supervised Vision-Language Pretraining for Medical Visual Question Answering

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T00:13:39.910268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T00:13:39.248870Z digest=sha256:3dab7338fb146cd6b6a48754e58ec55e38551ddb223918aac70916fdbd9abfa9

Observation 706e5344-5c1a-452f-a329-6a176a707231 · outbound

This paper cites Masked Vision and Language Pre-training with Unimodal and Multimodal Contrastive Losses for Medical Visual Question Answering.

Frequency-Domain Dual-Branch Fusion for Medical Visual Question Answering Masked Vision and Language Pre-training with Unimodal and Multimodal Contrastive Losses for Medical Visual Question Answering

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T00:13:39.801189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T00:13:39.252512Z digest=sha256:833eea5e21469802b29fa434de51198db536f649b470c556d09ede9e7ce7d74f

Observation d11be1ca-b46e-465c-a147-b274b80ad646 · outbound

This paper cites PeFoMed: Parameter Efficient Fine-tuning of Multimodal Large Language Models for Medical Imaging.

Frequency-Domain Dual-Branch Fusion for Medical Visual Question Answering PeFoMed: Parameter Efficient Fine-tuning of Multimodal Large Language Models for Medical Imaging

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T00:13:39.258091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T00:13:39.258091Z digest=sha256:e75626e606f7b73a156e2654e61b6187400302782e7d103b568a1ed0b7e47162

Observation 5be4630c-dc74-4c26-8579-da606f0146d0 · outbound

This paper cites BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs.

Frequency-Domain Dual-Branch Fusion for Medical Visual Question Answering BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T00:13:39.261649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T00:13:39.261649Z digest=sha256:7796e829432dbae90ae2c47315d6538709323fdbf9c81ca6c5d092e85ccc0a5a

Observation 9ef2c51f-7452-49d5-85ea-82b4045323ad · outbound

This paper cites BioBART: Pretraining and Evaluation of A Biomedical Generative Language Model.

Frequency-Domain Dual-Branch Fusion for Medical Visual Question Answering BioBART: Pretraining and Evaluation of A Biomedical Generative Language Model

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T00:13:39.793972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T00:13:39.264951Z digest=sha256:a955ceac2ef9295759e892f5f84e61ad0f1c664ccc8f80c41e6700ef36f1a334

Observation 478d2e88-58b5-4157-a16f-42386a00ed7f · outbound

This paper cites PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering.

Frequency-Domain Dual-Branch Fusion for Medical Visual Question Answering PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T00:13:39.269044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T00:13:39.269044Z digest=sha256:de88dc455cfe66b322cc1caf465c9dfc49a9498edf451c1e7efe0e76702014d0

Observation 933a2b5c-8fd1-4f09-a697-3ff03e3774d4 · outbound

This paper cites and Gayen, Soumya and Ben Abacha, Asma and Demner-Fushman, Dina , year =.

Frequency-Domain Dual-Branch Fusion for Medical Visual Question Answering and Gayen, Soumya and Ben Abacha, Asma and Demner-Fushman, Dina , year =

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T00:13:39.786363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T00:13:39.273287Z digest=sha256:ad9e1d1b454f09e0694cb41fbf4ee7135ce5e99644f5ac18a93d745db609702e

Observation d8a91ceb-8074-47fa-8f19-2420ef93c1ad · outbound

This paper cites SLAKE: A Semantically-Labeled Knowledge-Enhanced Dataset for Medical Visual Question Answering.

Frequency-Domain Dual-Branch Fusion for Medical Visual Question Answering SLAKE: A Semantically-Labeled Knowledge-Enhanced Dataset for Medical Visual Question Answering

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T00:13:39.775962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T00:13:39.277784Z digest=sha256:1220ef4272b7126a333dcd16b6ab5774c35cb09c9a42efa811310c0a017174ae

Observation 0c1d6849-1e59-4c9b-ba29-0481204fce4a · outbound

This paper cites Learning Transferable Visual Models From Natural Language Supervision.

Frequency-Domain Dual-Branch Fusion for Medical Visual Question Answering Learning Transferable Visual Models From Natural Language Supervision

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T00:13:39.765755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T00:13:39.315333Z digest=sha256:15b9d75df4a3f5d5c5d98b066e514ba1ff327931982ecffae347af739d373a3e

Observation 47c5002e-2bc9-48d8-92ba-37f7098f8a70 · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

Frequency-Domain Dual-Branch Fusion for Medical Visual Question Answering Representation Learning with Contrastive Predictive Coding

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T00:13:39.409329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T00:13:39.409329Z digest=sha256:62287f4f743cbd6deb75ed6175cf6a64a2d281caa2fef58d09a7b490e6f79b89

Observation 4d14bd01-c34d-40d2-9e37-39c85a7da69f · outbound

This paper cites Decoupled Weight Decay Regularization.

Frequency-Domain Dual-Branch Fusion for Medical Visual Question Answering Decoupled Weight Decay Regularization

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T00:13:39.757693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T00:13:39.474319Z digest=sha256:f550f987e30be50fc2685d9a922de70586a2d865e788d616d8489ddd39407b5f

Observation fc847a0c-2e92-442a-bd21-aa160dfe7b94 · outbound

This paper cites Proceedings of the IEEE International Conference on Computer Vision (ICCV) , pages =.

Frequency-Domain Dual-Branch Fusion for Medical Visual Question Answering Proceedings of the IEEE International Conference on Computer Vision (ICCV) , pages =

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T00:13:39.623717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T00:13:39.481236Z digest=sha256:98a91188b5ca0dc55f762ed432c8dce4a20421375b7c7e4766c6f14acb2f17bc

Observation 227f7878-44e5-4410-8e45-670bfaa33e2e · outbound

This paper cites Visualizing Data using t-SNE.

Frequency-Domain Dual-Branch Fusion for Medical Visual Question Answering Visualizing Data using t-SNE

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T00:13:39.540209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T00:13:39.491417Z digest=sha256:ab24761224c6190204ede4a6f63ffe67815ec44df27667a9f44a785136b28d9d

Pith citing papers

No inbound Pith citation observations are available.