Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:21:57.827375Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 0 inbound Pith citation observations for arXiv:2505.15401.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:21:57.827375Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
46 of 46 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation ee5f9f08-db8f-49f3-aa17-aae05f21407d · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities Visual Question An- swering for Wishart H-Alpha Classification of Polarimetric SAR Images
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 72cd52bd-54da-406e-a04e-f8f67f362add · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities Bottom-up and top-down attention for image captioning and visual question answering
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation cbb3dcdb-0f39-4df0-af7e-50773d900597 · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities VQA: Visual question answering
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 14b93957-a845-4c1d-901a-f47bb9a90d79 · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities Language trans- formers for remote sensing visual question answering
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1569f4e1-b7de-4fde-a3df-03e7d78eda25 · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities Prompt-RSVQA: Prompting visual context to a language model for remote sensing visual question answering
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3a3acf35-426a-4159-b21a-8a5961d67822 · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities Multi-task prompt-RSVQA to explicitly count objects on aerial images
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 87075da6-15b3-43bc-a8d6-88ade9d3c2be · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities The curse of language biases in remote sensing VQA: the role of spatial attributes, language diversity, and the need for clear evaluation
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9c8c442e-917f-41f3-a6d5-0f50ec553a95 · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bad9fba5-c881-435c-bcba-c4c1ea3913c5 · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities PubMedCLIP: How Much Does CLIP Benefit Visual Ques- tion Answering in the Medical Domain? InEACL, pages 1151–1163, 2023
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 04287ee4-337b-413e-ab54-6e83bea5363a · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities S2 missionhttps : / / sentiwiki.copernicus.eu/web/s2- mission# S2Mission - RadiometricPerformanceS2 - Mission - Radiometric - Performancetrue
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ce95719a-c126-4d54-9a28-a069b0e37db5 · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities Cross- modal visual question answering for remote sensing data
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6167abf0-df47-4bd7-b9bb-470696cd188d · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities Making the V in VQA matter: El- evating the role of image understanding in visual question answering
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b14f041d-5c20-43d0-a7f6-b1d2c604712d · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities Overview of image- CLEF 2018 medical domain visual question answering task
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation acab82ca-f225-44be-82a5-af778643655f · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities PromptCap: Prompt-guided image captioning for SAR with GPT-3
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6b9b9f92-e9f7-404f-9574-b7ea41712f52 · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities CLEVR: A diagnostic dataset for compositional language and elementary visual reasoning
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9b8a1a1a-3e0b-4894-a8e1-7a382063846c · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities Q: How to specialize large vision-language models to data-scarce VQA tasks? a: Self-train on unlabeled images! InCVPR Proceedings, pages 15005–15015, 2023
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 158f617a-7fbd-4cdd-ab8d-5c937bae040b · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities Deep learning in multi- modal remote sensing data fusion: A comprehensive review
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 86e211b8-017c-46b3-899a-a9ae42449976 · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities VisualBERT: A Simple and Performant Baseline for Vision and Language
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1301e4cd-21d6-46a0-b1e2-3cfff6b3f5d0 · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities A comprehensive study of GPT-4V’s multimodal capabilities in medical imaging
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 12c92640-5b84-4a23-b599-6aafacffe9af · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities Medical visual question answering: A survey.Artificial In- telligence in Medicine, page 102611, 2023
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 01596c57-bf5b-485a-ba05-ac31a85d6fbd · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities RSVQA: Visual question answering for remote sensing data
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ea858d96-e9e7-4634-85fd-bdbb2f754d66 · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities RSVQA meets BigEarthNet: a new, large-scale, visual question an- swering dataset for remote sensing
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a031490f-64a6-4c73-95d2-fff894e5a46c · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities Deep learning and earth observation to support the sustainable development goals: Current approaches, open challenges, and future opportunities.GRS, 10(2):172–200,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e1a65670-4db2-4f13-850d-ef3ddd8017a3 · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities Learn- ing transferable visual models from natural language super- vision
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2d6b8fa-8ecb-491b-ab41-ac35a409a5b3 · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f51676e-46bc-46ab-a39e-c9b364ed7332 · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities How Much Can CLIP Benefit Vision-and-Language Tasks?
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33e86e99-5d9c-4be6-92c2-07820ee4a957 · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities BigEarthNet: A Large-Scale Benchmark Archive For Remote Sensing Image Understanding
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5414bcfe-b7c1-4eb3-9991-cbed8248c692 · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities BigEarthNet-MM: A large-scale, multimodal, multilabel benchmark archive for remote sensing image classification and retrieval [software and data sets].GRS, 9(3):174–180, 2021
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ac91fa7a-49c9-428c-bace-78be154dbe7d · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities LXMERT: Learning cross- modality encoder representations from transformers
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 479d1f48-f656-44c2-9d25-704d89862741 · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities Segmentation-guided attention for visual question answering from remote sensing images
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 33ff6f2d-ca6c-4bc6-8303-f71d08550ef9 · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities Can SAR improve RSVQA performance? In EUSAR, pages 1287–1292
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 28086c65-172e-49b7-a07b-3d98ad54f09c · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities A Visual Question Answering Method for SAR Ship: Breaking the Requirement for Multimodal Dataset Construction and Model Fine-Tuning
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 190efb43-6d6a-4629-a6e1-d16bb1bc55ff · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities Labsar, a one- gcp coregistration tool for sar–insar local analysis in high- mountain regions.Frontiers in Remote Sensing, 3:935137,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation eefacd51-dccc-4db1-a41f-169c64f1d508 · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities LabSAR, a one-GCP coregistration tool for SAR–InSAR local analysis in high- mountain regions.Frontiers in Remote Sensing, 3, 2022
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4370f8d1-0e16-41e1-bde3-84d33c886329 · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities Stacked attention networks for image question answering
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 481cc563-b99c-4592-8e44-1758d5739415 · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities Self- paced curriculum learning for visual question answering on remote sensing data
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4922c276-5ea5-4815-b7dc-0e8fc436f18a · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities Multi- lingual augmentation for robust visual question answering in remote sensing images
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3639b5e5-987d-42e6-a45c-1158e4d87256 · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities Frequency domain transfer learning for remote sensing visual question answering
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d2f071f0-c906-4382-8fdf-342f25dfc8a7 · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities Exploring data and models in SAR ship image captioning.IEEE Access, 10:pp
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 404125af-856a-44de-885a-48d7e81ab912 · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities Mutual attention inception network for remote sensing visual question answering.TGRS, 60:1–14, 2021
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2702e120-efee-4a32-aa7d-5874ef75d4fe · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities TRAR: Routing the attention spans in transformer for visual question answering
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 08bba89e-9e50-447d-9c45-c26b9122cbbc · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities To do so, we need the geographical position of the center of the VHR patch
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 06147c63-591a-457c-8970-996923b93856 · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities Unresolved cited work
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation aa5e0a4b-d6f6-4b0c-ab9b-3e3eb4239863 · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities To find the correct swath, the projection of the geographical point is applied, using the meta-data linked to each swath
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 784ca467-5426-4cdd-b736-b85e603c9cbd · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities The S1 images need to be debursted (removing of the black line and of the overlap) to get a continuous image before to extract the S1 patch that is inputted in the model
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6b0b2b99-4316-4d9d-9902-264e0b7a7d45 · outbound
Visual Question Answering on Multiple Remote Sensing Image Modalities A tail- value elimination procedure is performed on each chan- nel separately using statistics information extracted over the whole dataset
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
No inbound Pith citation observations are available.