Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T15:05:57.863886Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 1 inbound Pith citation observation for arXiv:2411.14704.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T15:05:57.863886Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-10T00:02:12.727566Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-10T00:29:47.839616Z
55 of 55 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 4c3c0f33-d76b-44b2-828f-248f4bb5975a · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Remote sensing big data computing: Challenges and opportunities,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b520db60-7253-48a9-ab56-357ceba5605a · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Big data for remote sensing: Challenges and opportunities,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 186ad7d1-a193-4143-a9b6-b76fada333c3 · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Understanding urban landuse from the above and ground perspectives: A deep learning, multi- modal solution,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 43844814-7191-468d-a558-7acd3ab66674 · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Hyperspectral data analysis for arid vegetation species: Smart & sustainable growth,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 6fa11a44-fe85-41c9-a49a-5b0c45c4ce5f · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Google earth engine cloud computing platform for remote sensing big data applications: A comprehensive review,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 960bfcb1-ca03-4497-8df8-27754115cfda · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Nwpu- captions dataset and mlca-net for remote sensing image captioning,
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 654ef5dd-7f5b-4121-bd63-54b3ae71712e · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Textrs: Deep bidirectional triplet network for matching text to remote sensing images,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 313642ac-f437-493f-823a-de51ef14e235 · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval A deep semantic alignment network for the cross-modal image-text retrieval in remote sensing,
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a042db3-5854-4d0e-9c98-8f731f04820f · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Fusion-based correlation learning model for cross-modal remote sensing image retrieval,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b1447892-a6cf-4ffb-84d5-2d9dfe23f503 · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Cross spectral image reconstruction using a deep guided neural network,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f87b27fd-5a3f-4f20-b745-98abb2d47e4f · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Image super-resolution using t-tetromino pixels,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a2bafe0a-f97a-414e-9580-6e01c0556ead · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Exploring a fine-grained multiscale method for cross-modal remote sensing image retrieval,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation deb1efde-12a0-4211-994d-3d1cc98e2126 · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Remote sensing cross-modal text-image retrieval based on global and local information,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14cdfa4f-a120-4b0e-ae72-c3c707dcabf0 · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval A lightweight multi-scale crossmodal text-image retrieval method in remote sensing,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c8db1e29-8418-4854-ae4c-169435be8548 · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Exploring uni-modal feature learning on entities and relations for remote sensing cross-modal text-image re- trieval,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 70b6b766-acec-470d-b78a-489c3a11e6af · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Hypersphere-based remote sensing cross-modal text-image retrieval via curriculum learning,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d19e02e5-e545-43e2-b5ad-b9799914777c · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f20a153-5c0f-4009-a05c-947e74b345ca · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Long short-term memory recurrent neural network architectures for large scale acoustic modeling,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation beb708e3-24d7-4cce-ae2e-c8810b2f1c7d · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Empirical Evaluation of Gated Recurrent Neural Networks on Sequence Modeling
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6dce598e-f33e-42cb-8ffc-3cf4bb0e6501 · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Attention is all you need,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dce4e68e-4451-4a12-980a-ede130182e20 · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Multiscale salient alignment learning for remote sensing image-text retrieval,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 71c63976-4336-441b-8ff7-2f40c3b50181 · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Interacting- enhancing feature transformer for cross-modal remote sensing image and text retrieval,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 9465be24-0dca-4812-a1c8-70fdddfadfc5 · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Align before fuse: Vision and language representation learning with momentum distillation,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation e6a07d92-cebb-495a-8826-6c0c53a5ed8c · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Deep saliency smoothing hashing for drone image retrieval,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 76ed2888-dad4-4d5a-9698-91351679a81c · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Multitask learning for sar ship detection with gaussian-mask joint segmentation,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 930a1883-858a-468c-ae0d-e8757f9c2a18 · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Swin transformer: Hierarchical vision transformer using shifted windows,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 2380a155-4554-40be-b0f7-8cda192742e4 · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Global context vision transformers,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 37aa4e64-5910-4588-9fbb-cf1ea37451dc · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Matching images and text with multi-modal tensor fusion and re- ranking,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 6d79f654-7652-4f77-a091-b4447c4ddbfb · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Vse++: Improving visual-semantic embeddings with hard negatives,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 84ac96ff-b178-43ae-82b3-893433494fb9 · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Exploring models and data for remote sensing image caption generation,
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fafb1176-24c3-427d-b2a1-e7c343df78e3 · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Deep semantic understanding of high resolution remote sensing image,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 44d0bb0b-68d2-4af5-bf61-64699935b562 · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval End-to-end convolutional semantic embeddings,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 519cb981-5d09-44c6-896e-a8e9aecc3cd6 · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Cross-modal semantic correlation learning by bi-cnn network,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f110aff3-2c77-467e-856d-caf394a2b75a · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Dual-path convolutional image-text embeddings with instance loss,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation e1216bf4-dd3e-4a46-a799-90b70d2aa7f3 · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Deep supervised cross-modal retrieval,
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03dc5032-001b-4cd2-9d2d-27ef02950545 · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Learning semantic concepts and order for image and sentence matching,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation ee5e49a4-9c8a-4ed9-be2d-5070c4577c11 · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Stacked cross attention for image-text matching,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 8fd228ad-e85b-43c2-a565-2c2e4aa3d91a · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Cross- modal attention with semantic consistence for image–text matching,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 38ae2995-86c3-4b14-93b2-da73ff3a6f7f · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Visual semantic reasoning for image-text matching,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 91c20eb1-1199-455f-a128-3099ece39e98 · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Image-text embedding learning via visual and textual semantic reasoning,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 1414cfaf-f4aa-449e-a1a5-f582b8bb01c7 · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks,
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ecf46abb-c98e-4236-af91-6bb7d5159778 · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Lxmert: Learning cross-modality encoder representations from transformers,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c1a80255-d52c-4e74-87fb-1c654dfcc07f · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Fashionbert: Text and image matching with adaptive loss for cross- modal retrieval,
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 84f014d7-3af0-4263-b883-94ef99a55daa · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Learning the best pooling strategy for visual semantic embedding,
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0e90948c-19b9-4e50-8e10-f5988d1ed44f · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Pixel-BERT: Aligning Image Pixels with Text by Deep Multi-Modal Transformers
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf64ed5d-685c-43e1-9734-415474ed2ece · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Learning transferable visual models from natural language supervision,
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2c38548-073c-429f-a7d2-c76397f0bc7a · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Vista: Vision and scene text aggregation for cross-modal retrieval,
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d6f8ddff-9a5c-43b6-82b6-5bc90cca8ffb · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Vilt: Vision-and-language transformer without convolution or region supervision,
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 597fa9db-e019-4720-8555-7da321fee29a · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Vlmo: Unified vision-language pre-training with mixture-of-modality-experts,
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 45903d78-eded-4900-a047-c4855e2cfac6 · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Knowledge-aided momentum contrastive learning for remote-sensing image text retrieval,
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 9b3afb81-747a-406f-858f-61a9b7dd3591 · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Multi- scale interactive transformer for remote sensing cross-modal image-text retrieval,
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 87a56034-0031-4957-9722-3c6a38a28750 · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Parameter-efficient transfer learning for remote sensing image-text retrieval,
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e1b2a61-fe9a-4da8-bcee-5a7041d2c5ce · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Integrating multisubspace joint learning with multilevel guidance for cross-modal retrieval of remote sensing images,
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ac19980-e07b-4711-8c32-64bc5507a56c · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Bert: Pre-training of deep bidirectional transformers for language understanding,
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 852a2f44-d06f-416e-843f-3fb4879354b6 · outbound
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Momentum contrast for unsupervised visual representation learning,
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation be331e05-aba8-4882-8bdd-c298a7fd228c · inbound
Fast-then-Fine: A Two-Stage Framework with Multi-Granular Representation for Cross-Modal Retrieval in Remote Sensing Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.