Pith. sign in

Paper Citation Record · LEDGER

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation

As of 8 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 1 inbound Pith citation observation for arXiv:2507.07568.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.07568 v1

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:40:59.075135Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-10T14:09:06.139625Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-10T14:10:28.670574Z

Reference resolution

43 of 43 outbound references displayed

  • verified exact1
  • verified fuzzy31
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 78479a4d-5929-44c4-a896-406487b8c21b · outbound

This paper cites Bottom-up and top-down attention for image captioning and visual question answering.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Bottom-up and top-down attention for image captioning and visual question answering

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:54.634276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:54.634276Z digest=sha256:62b3656e33f0542f1ba6c99728af2898042f08f85bdccca2d4a3c387de030398

Observation 66c350e7-409e-48e6-b3ab-d8610a0fc357 · outbound

This paper cites Instance-level expert knowledge and aggregate discrimina- tive attention for radiology report generation.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Instance-level expert knowledge and aggregate discrimina- tive attention for radiology report generation

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:05.610400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:54.698556Z digest=sha256:f2e4f32d3c44024887ce60abd84cebe320c70da5a5fbbb06cf3d73dbc6aad382

Observation a3ec215e-a41d-40fe-81a3-32556a435ba6 · outbound

This paper cites Fine-Grained Image-Text Alignment in Medical Imaging Enables Explainable Cyclic Image-Report Generation.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Fine-Grained Image-Text Alignment in Medical Imaging Enables Explainable Cyclic Image-Report Generation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:54.786069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:54.786069Z digest=sha256:abcd4eb3f6dbf18b9f6e4ebf58dc362e16f460abbd5294208ccbf92d67e3de72

Observation d504dc60-88db-48ec-b52f-f26993d8d51c · outbound

This paper cites Generating Radiology Reports via Memory-driven Transformer.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Generating Radiology Reports via Memory-driven Transformer

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:54.869300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:54.869300Z digest=sha256:b0ece251a025f88babb49971e304ceb48db72ec79327da3c11bbb03b86df2a4c

Observation 9d6a7184-3743-452e-b37b-29994f9f3119 · outbound

This paper cites Cross-modal Memory Networks for Radiology Report Generation.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Cross-modal Memory Networks for Radiology Report Generation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:54.940751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:54.940751Z digest=sha256:b2a697e66880092e6f004fd90fb88fdcf44b12d27e5dcf8e0ad508290ff42970

Observation 52a97cc1-afeb-4dec-95d9-80a01c7c7f76 · outbound

This paper cites Meshed-memory transformer for image cap- tioning.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Meshed-memory transformer for image cap- tioning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:55.064495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:55.064495Z digest=sha256:d33f23208786ccca5abb7c4adca7aa4a269c05d433b628d4bd02b1b11b79e98c

Observation e2f3b51a-843d-4293-9f9d-eec123549feb · outbound

This paper cites To- wards diverse and natural image descriptions via a condi- tional gan.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation To- wards diverse and natural image descriptions via a condi- tional gan

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:05.462231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:55.148019Z digest=sha256:3e61df91e385da02dae0ec80a6c3fdbc4aee255cb5dc69d1ea595a4bf55024e6

Observation e41aad48-9aba-4b14-8f8b-079ca4ba5034 · outbound

This paper cites Meteor 1.3: Automatic metric for reliable optimization and evaluation of machine translation systems.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Meteor 1.3: Automatic metric for reliable optimization and evaluation of machine translation systems

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:05.278577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:55.209257Z digest=sha256:861a5de3cad6dff850efdbf63d14d794f0a32280e0d00244a5eca90bd9458ee4

Observation 444720c9-34f8-4b33-ba88-03685b39b303 · outbound

This paper cites Long short-term memory.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Long short-term memory

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:05.109519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:55.323303Z digest=sha256:8954d3ea2a52c045add8d0f3a00a925cc2f4c8271252f80bb2a9a4a99b46013a

Observation 69c3d97e-2961-4c1f-a040-89f0b82b9409 · outbound

This paper cites Scaling up vision-language pre-training for image captioning.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Scaling up vision-language pre-training for image captioning

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:04.969520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:55.433036Z digest=sha256:65f961fff8946039327f8a85ec5acb9cc18330a8da1647aeeeb638878a73ce20

Observation fd4f6c4e-9dfe-4597-9e56-c1c3019e36e6 · outbound

This paper cites Kiut: Knowledge-injected u-transformer for radiology report generation.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Kiut: Knowledge-injected u-transformer for radiology report generation

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:04.768171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:55.534153Z digest=sha256:36a9ae2bff16c6a225ba7115d3a32262ba6a00bd33619426e248e2e05b64c8fc

Observation b3f8e06e-977d-4568-84cd-db6808602b1d · outbound

This paper cites Chexpert: A large chest radiograph dataset with uncertainty labels and expert comparison.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Chexpert: A large chest radiograph dataset with uncertainty labels and expert comparison

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:04.578931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:55.655311Z digest=sha256:5ad4bb6ff7bda6d3d5c8c983fbc9c43b506d28a195fff9436a9f0accdd33f3fc

Observation 66150774-7d3f-42a1-96e6-d9944ecf8717 · outbound

This paper cites Promptmrg: Diagnosis-driven prompts for medical report generation.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Promptmrg: Diagnosis-driven prompts for medical report generation

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:04.442761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:55.760521Z digest=sha256:613cf12954cfdc3d0498a3b3797a02425ac13b9ae407e9cd91798322cf76c566

Observation 5b4ef1b4-7b98-4a38-993f-91d441f63684 · outbound

This paper cites Zero-shot camouflaged object detection.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Zero-shot camouflaged object detection

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:04.318825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:55.898674Z digest=sha256:3ae839834a70704c53331364c91611a4cbaafc687f066f2c5e777fb6c691c9f0

Observation baffaa0e-0a8a-4797-b2e5-e7060b176690 · outbound

This paper cites Dynamic graph enhanced contrastive learning for chest x-ray report generation.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Dynamic graph enhanced contrastive learning for chest x-ray report generation

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:03.750014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:56.011195Z digest=sha256:5e60666e5e9ce9527b8d890eb4d8625573b4d205f9efc2c0477165fd0c8204a8

Observation db6feba9-9b44-4b77-87e1-a16bb8ee3b6e · outbound

This paper cites Unify, align and refine: Multi- level semantic alignment for radiology report generation.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Unify, align and refine: Multi- level semantic alignment for radiology report generation

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:03.303434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:56.141380Z digest=sha256:385c474e16b6790c277fd91a4c9b188c752e3c9b2849912e05ffa4f105dae26a

Observation 445c98da-8d75-4fc3-9628-def98b1f73b0 · outbound

This paper cites Rouge: A package for automatic evaluation of summaries.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Rouge: A package for automatic evaluation of summaries

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:56.262544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:56.262544Z digest=sha256:99e2a0d50a3084ff141c40dd27073f74a44de0953f4d01bbc3bbae474c0e1a96

Observation 3e1d39db-f330-47b2-b6be-84e91f921b60 · outbound

This paper cites Exploring and distilling posterior and prior knowl- edge for radiology report generation.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Exploring and distilling posterior and prior knowl- edge for radiology report generation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:03.132064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:56.361215Z digest=sha256:f615549fe96f4d84c1cad8a15bf8b538100423aef54d7bf1a8776a2750c80311

Observation 26596a90-efe7-46db-97ae-9bbbeb7ec035 · outbound

This paper cites Grounding dino: Marry- ing dino with grounded pre-training for open-set object de- tection, 2024.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Grounding dino: Marry- ing dino with grounded pre-training for open-set object de- tection, 2024

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:02.954214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:56.507414Z digest=sha256:c83bc686a81abc7aa4ddfa68009f64f63e174c8619a80b0826aceb15c31fbe4e

Observation c30f39ce-c50a-4523-a79e-52880908d818 · outbound

This paper cites Knowing when to look: Adaptive attention via a visual sen- tinel for image captioning.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Knowing when to look: Adaptive attention via a visual sen- tinel for image captioning

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:02.739319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:56.639657Z digest=sha256:091c35dea68eb22ec4686c2592335e437801658d335c49e70c4b9ee56a9de04c

Observation f1115d3c-2adf-4404-beba-d14d20bf12da · outbound

This paper cites Im- proving chest x-ray report generation by leveraging warm starting.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Im- proving chest x-ray report generation by leveraging warm starting

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:02.508737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:56.748965Z digest=sha256:56d0053a43bdf2bd20945fbc16c04e168862b6bab52b9ba301b8b76bf8324362

Observation 7600f648-b485-4dec-a247-a68a9f9f67cb · outbound

This paper cites Progressive Transformer-Based Generation of Radiology Reports.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Progressive Transformer-Based Generation of Radiology Reports

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:40:59.265156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:56.851674Z digest=sha256:425f6e85ba7c3e9b739b7fb9430a79a22e6d991569a780ae7690e3957f87241c

Observation 481197d6-1da4-4f23-a372-1101e123d7bb · outbound

This paper cites An Introduction to Convolutional Neural Networks.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation An Introduction to Convolutional Neural Networks

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:56.975156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:56.975156Z digest=sha256:32ea3cd944babfab2b7620303fc190ce67caffc2d3090e1cc6e669e28bde0ab5

Observation 8a47ff2b-b8ab-4ff7-8152-aeba2f888650 · outbound

This paper cites X-linear attention networks for image captioning.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation X-linear attention networks for image captioning

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:02.280231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:57.092731Z digest=sha256:bc8678d075a46cbd65c6cf8f87c336d285323ed923cf89a1b383d71e46f561fa

Observation 399b2de2-7ed2-4d59-b049-64f5531a590b · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Bleu: a method for automatic evaluation of machine translation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:57.192845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:57.192845Z digest=sha256:fc3b7fb8051ada1a9a716f97a04b36220d9283bf5fe1366ac289ceddcc2a230a

Observation 2279dcf1-cd36-4f63-90c1-8bcec903b8a5 · outbound

This paper cites Self-critical sequence training for image captioning.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Self-critical sequence training for image captioning

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:02.077441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:57.300223Z digest=sha256:b4333b5a968b0f8e0287c2cfd48fc7fb9d2ff6f5abf1271ac9cbdec1f1f42719

Observation 6cbdd847-1624-48ba-bea7-e050a2f5ce2f · outbound

This paper cites CheXbert: Combining Automatic Labelers and Expert Annotations for Accurate Radiology Report Labeling Using BERT.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation CheXbert: Combining Automatic Labelers and Expert Annotations for Accurate Radiology Report Labeling Using BERT

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:57.387664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:57.387664Z digest=sha256:8ceac9511b2c755eaa0c7e931914b4e616a172153b7d6639e89803a2bce093a1

Observation bb509744-ab22-429f-9980-1e8dc68f5de1 · outbound

This paper cites Interactive and explainable region-guided radiol- ogy report generation.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Interactive and explainable region-guided radiol- ogy report generation

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:01.885758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:57.521412Z digest=sha256:3f38f7a2759d496eb35806c0c96980c78adcad23fb5cf4ee05e074d2e8198bbb

Observation 9f787231-8d18-4ab0-b334-5b4bb60b2e60 · outbound

This paper cites Attention is all you need.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Attention is all you need

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:57.653944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:57.653944Z digest=sha256:f00d022b135bdeb9a1f1fa90d047eba00eff98141e641fd5fa5567ea5a74a99f

Observation 726f4580-4d9c-4fb2-a79d-9b8813b37521 · outbound

This paper cites Hergen: El- evating radiology report generation with longitudinal data,.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Hergen: El- evating radiology report generation with longitudinal data,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:01.679143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:57.808493Z digest=sha256:1e875bd25e835dc8227fd4bcb0878e26d17ad4ccce334863b424b659c27f5827

Observation dc82586b-a8f1-4b1b-a1e6-cee1f9b0c2b3 · outbound

This paper cites Cross-modal pro- totype driven network for radiology report generation.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Cross-modal pro- totype driven network for radiology report generation

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:01.372332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:57.876827Z digest=sha256:4963bcdb8e78d00f4381f837d2ea52055f3f4264608a9212759cac4988950a82

Observation b743310e-a002-4e63-b0ea-cbe3e98c47af · outbound

This paper cites Multi-view feature fusion and visual prompt for remote sensing image captioning.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Multi-view feature fusion and visual prompt for remote sensing image captioning

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:01.242470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:57.968629Z digest=sha256:baf1055cc889a18992291df3c75dcfbf709511484db522d72f3ebe8a2a9e06aa

Observation 777f3ed9-0a61-4858-a7bf-c2a69703f466 · outbound

This paper cites Medclip: Contrastive learning from unpaired medical images and text, 2022.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Medclip: Contrastive learning from unpaired medical images and text, 2022

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:01.050486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:58.066074Z digest=sha256:b179cfc62a5d7568342cc8ab336d72c9b6c012ef199b99be0f9741a08119622b

Observation 81bd850c-94a1-4f1e-9f70-de01d224ba5b · outbound

This paper cites Metransformer: Radiology report generation by transformer with multiple learnable expert tokens.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Metransformer: Radiology report generation by transformer with multiple learnable expert tokens

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:00.765205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:58.159836Z digest=sha256:f45c28566abb72b9beafcc23d0eb455f670552b39678420b21450b5dd155db1e

Observation 00e48e2c-8dbf-4bc5-a4ec-39c6a2f85fcf · outbound

This paper cites Medklip: Medical knowledge enhanced language-image pre-training in radiology, 2023.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Medklip: Medical knowledge enhanced language-image pre-training in radiology, 2023

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:00.516888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:58.283944Z digest=sha256:b01bc670ec7479e59fc67f8a35f66fdcf9c35ff00e0bb7dec2b400ff68450b61

Observation ca36f8da-a3e4-4651-88e2-b907d02e5427 · outbound

This paper cites Clinical-bert: Vision-language pre-training for radiograph diagnosis and reports generation.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Clinical-bert: Vision-language pre-training for radiograph diagnosis and reports generation

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:00.236712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:58.395579Z digest=sha256:b12a9987bf2d41788cef24399cb474da6706b05f8ff40271aeb0bcb64f6eba48

Observation a1c43eb3-35d8-4341-8c96-d4285d96a71d · outbound

This paper cites Knowledge matters: Chest radiology report genera- tion with general and specific knowledge.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Knowledge matters: Chest radiology report genera- tion with general and specific knowledge

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:00.110219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:58.484139Z digest=sha256:1f341b9ec372ab294f54263316ae3e8be9d6bb11c505d3704eb25c248421cb66

Observation 3a4ab974-bc06-472f-a825-8bb259196c28 · outbound

This paper cites Radiology report generation with a learned knowledge base and multi-modal alignment.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Radiology report generation with a learned knowledge base and multi-modal alignment

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:40:59.972046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:58.584779Z digest=sha256:6f3d774a73372973ca9826290028c9933db33fbd0401547f1067d750b87440ce

Observation fc0a99fe-33e2-4751-b791-a7d51508a97f · outbound

This paper cites Improving hyperbolic representations via gromov- wasserstein regularization.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Improving hyperbolic representations via gromov- wasserstein regularization

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:40:59.865065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:58.684853Z digest=sha256:1c8fd63bcc4acf9dee571e3be7ce18484550cea6d89917cea2e721bd75683bef

Observation 61a4302a-1b9b-4391-b4d0-5c6efedcccd2 · outbound

This paper cites Otseg: Multi-prompt sinkhorn attention for zero-shot semantic segmentation.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Otseg: Multi-prompt sinkhorn attention for zero-shot semantic segmentation

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:40:59.719371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:58.767838Z digest=sha256:11276a3000e51a38d52173b35eea2931f4370c8b9686ff7e133e2a1f1adf0d86

Observation 6f210f77-71be-4632-afa1-6bd23d153e4a · outbound

This paper cites CoCa: Contrastive Captioners are Image-Text Foundation Models.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation CoCa: Contrastive Captioners are Image-Text Foundation Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:58.884213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:58.884213Z digest=sha256:ffcac157901879e0138177885cfc2e05ee708afb17716a97a71e781639679347

Observation 59a90ce7-959b-4231-9389-27b178fc4b20 · outbound

This paper cites Anatomy-guided weakly- supervised abnormality localization in chest x-rays.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Anatomy-guided weakly- supervised abnormality localization in chest x-rays

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:40:59.576735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:58.967913Z digest=sha256:71d20444375ea81d934eda5b25e08174577e35da97a8b9d27d5b3f01ddd3c6df

Observation 21525339-621f-46bf-876d-c7d8cecf6915 · outbound

This paper cites Sam-guided enhanced fine-grained encoding with mixed semantic learning for medical image captioning.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Sam-guided enhanced fine-grained encoding with mixed semantic learning for medical image captioning

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:40:59.428509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:59.075135Z digest=sha256:f6b7ee54fdd853848eda1da4f441ff0a62fed5b4f21ea3c63bf78d688e32b744

Pith citing papers

Observation 5c7c268a-d65a-4457-a4fa-2afbffd5e22a · inbound

Enhancing Reinforcement Learning for Radiology Report Generation with Evidence-aware Rewards and Self-correcting Preference Learning cites this paper.

Enhancing Reinforcement Learning for Radiology Report Generation with Evidence-aware Rewards and Self-correcting Preference Learning Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:10:28.672612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-10T14:09:06.139625Z digest=sha256:736eb1899c4608c626eae6b7f35eab1dd1576535b0ad25caa0a205465ea6bfb3