Pith. sign in

Paper Citation Record · LEDGER

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval

As of 13 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 1 inbound Pith citation observation for arXiv:2411.14704.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.14704 v1

Coverage vector

measured 55 of 55 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T15:05:57.863886Z

measured 56 of 56 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-10T00:02:12.727566Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-10T00:29:47.839616Z

Reference resolution

55 of 55 outbound references displayed

  • verified exact0
  • verified fuzzy41
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4c3c0f33-d76b-44b2-828f-248f4bb5975a · outbound

This paper cites Remote sensing big data computing: Challenges and opportunities,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Remote sensing big data computing: Challenges and opportunities,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.612459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.642747Z digest=sha256:35a82778002dfeefb2ae7165b606f40a654583a9f192927926adfe99c16d6e50

Observation b520db60-7253-48a9-ab56-357ceba5605a · outbound

This paper cites Big data for remote sensing: Challenges and opportunities,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Big data for remote sensing: Challenges and opportunities,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.599090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.648140Z digest=sha256:076910d244e52974b880904ada9055697e8310ca69ef7b548abd48a61cc6160c

Observation 186ad7d1-a193-4143-a9b6-b76fada333c3 · outbound

This paper cites Understanding urban landuse from the above and ground perspectives: A deep learning, multi- modal solution,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Understanding urban landuse from the above and ground perspectives: A deep learning, multi- modal solution,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.583419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.653137Z digest=sha256:3dbe55335d85dc5ae341cf3db9dfb49e710737f9ced00522108ca9a4ddc0cbd8

Observation 43844814-7191-468d-a558-7acd3ab66674 · outbound

This paper cites Hyperspectral data analysis for arid vegetation species: Smart & sustainable growth,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Hyperspectral data analysis for arid vegetation species: Smart & sustainable growth,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.569310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.660856Z digest=sha256:7097919565cd11a438e8df9b11a3cfe4db0d255bc7ed3963ba2120409842aa1a

Observation 6fa11a44-fe85-41c9-a49a-5b0c45c4ce5f · outbound

This paper cites Google earth engine cloud computing platform for remote sensing big data applications: A comprehensive review,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Google earth engine cloud computing platform for remote sensing big data applications: A comprehensive review,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.554913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.665468Z digest=sha256:e7ffd0b6bb67d2160da3a7914bd6ed03b39c6620b7e83de2fce2ed50b326c254

Observation 960bfcb1-ca03-4497-8df8-27754115cfda · outbound

This paper cites Nwpu- captions dataset and mlca-net for remote sensing image captioning,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Nwpu- captions dataset and mlca-net for remote sensing image captioning,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.669876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.669876Z digest=sha256:99161f8d025d28bb04b7dceafb634ac40abcf6ebfdbe4ef44736b3dd1a485da6

Observation 654ef5dd-7f5b-4121-bd63-54b3ae71712e · outbound

This paper cites Textrs: Deep bidirectional triplet network for matching text to remote sensing images,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Textrs: Deep bidirectional triplet network for matching text to remote sensing images,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.521644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.674957Z digest=sha256:507ff690b1057ad6548bb9e665b67aae1f16ee66f95440c7872a5a80320a307a

Observation 313642ac-f437-493f-823a-de51ef14e235 · outbound

This paper cites A deep semantic alignment network for the cross-modal image-text retrieval in remote sensing,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval A deep semantic alignment network for the cross-modal image-text retrieval in remote sensing,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.679755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.679755Z digest=sha256:bdd4592fffba16b763d572301033544aa5a270d24f823b11f13f59b1da8dd3df

Observation 0a042db3-5854-4d0e-9c98-8f731f04820f · outbound

This paper cites Fusion-based correlation learning model for cross-modal remote sensing image retrieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Fusion-based correlation learning model for cross-modal remote sensing image retrieval,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.499246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.683688Z digest=sha256:d7376e526b9f664a6da9f08d66d8abd2f5ae78f8e208169ae98d790564ca31ff

Observation b1447892-a6cf-4ffb-84d5-2d9dfe23f503 · outbound

This paper cites Cross spectral image reconstruction using a deep guided neural network,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Cross spectral image reconstruction using a deep guided neural network,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.486375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.687271Z digest=sha256:0ea35decafa8d15ab0bff49dad84ba45ed2b9e82a88ebd336dc45785c14478d4

Observation f87b27fd-5a3f-4f20-b745-98abb2d47e4f · outbound

This paper cites Image super-resolution using t-tetromino pixels,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Image super-resolution using t-tetromino pixels,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.472800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.691404Z digest=sha256:4be8fa00197ae2a040c9a7bce56209b101765be98c5e0370a8ceac6d97538d2a

Observation a2bafe0a-f97a-414e-9580-6e01c0556ead · outbound

This paper cites Exploring a fine-grained multiscale method for cross-modal remote sensing image retrieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Exploring a fine-grained multiscale method for cross-modal remote sensing image retrieval,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.695460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.695460Z digest=sha256:76eceaecc20ed32f61b3495c9e376574fa0cd0df2ee9659db49f07162a8ccd21

Observation deb1efde-12a0-4211-994d-3d1cc98e2126 · outbound

This paper cites Remote sensing cross-modal text-image retrieval based on global and local information,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Remote sensing cross-modal text-image retrieval based on global and local information,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.699844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.699844Z digest=sha256:3be4263ea9fa45296ae44e1a5d80311500c7532ef75ff20a07df2bfa2a7f246f

Observation 14cdfa4f-a120-4b0e-ae72-c3c707dcabf0 · outbound

This paper cites A lightweight multi-scale crossmodal text-image retrieval method in remote sensing,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval A lightweight multi-scale crossmodal text-image retrieval method in remote sensing,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.440911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.704005Z digest=sha256:2624980d2c038735ef5289701a23e9089136cfd2ca554d63ab1ff7ded0488b6e

Observation c8db1e29-8418-4854-ae4c-169435be8548 · outbound

This paper cites Exploring uni-modal feature learning on entities and relations for remote sensing cross-modal text-image re- trieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Exploring uni-modal feature learning on entities and relations for remote sensing cross-modal text-image re- trieval,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.427872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.707863Z digest=sha256:0d7b092f1d36f3110922c5fb76d9bb2768023b51d5869ceacd17ff89b15c5860

Observation 70b6b766-acec-470d-b78a-489c3a11e6af · outbound

This paper cites Hypersphere-based remote sensing cross-modal text-image retrieval via curriculum learning,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Hypersphere-based remote sensing cross-modal text-image retrieval via curriculum learning,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.413683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.711725Z digest=sha256:88773130cc8e690fae926b83f7e1c90e1135c26fd4f138f657fe041f72198e72

Observation d19e02e5-e545-43e2-b5ad-b9799914777c · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.715976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.715976Z digest=sha256:8bc1ab87fd4f182de93d24c3ec066039dddf2b62f710fb8f849a24f8f680b6ed

Observation 4f20a153-5c0f-4009-a05c-947e74b345ca · outbound

This paper cites Long short-term memory recurrent neural network architectures for large scale acoustic modeling,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Long short-term memory recurrent neural network architectures for large scale acoustic modeling,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.399313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.720276Z digest=sha256:78313dddecaf823019826510c5d9dc700c286d2b21caede0131c1deca7aa1979

Observation beb708e3-24d7-4cce-ae2e-c8810b2f1c7d · outbound

This paper cites Empirical Evaluation of Gated Recurrent Neural Networks on Sequence Modeling.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Empirical Evaluation of Gated Recurrent Neural Networks on Sequence Modeling

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.724413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.724413Z digest=sha256:fc5f1f09b99e5908661eeda54c103e0c78735a93c9beb4b0c736da406d55c7a4

Observation 6dce598e-f33e-42cb-8ffc-3cf4bb0e6501 · outbound

This paper cites Attention is all you need,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Attention is all you need,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.728685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.728685Z digest=sha256:227e4a93c5b651158af3481c29cbb457d1f7cc563652dd6a7c8734e0cbaa7497

Observation dce4e68e-4451-4a12-980a-ede130182e20 · outbound

This paper cites Multiscale salient alignment learning for remote sensing image-text retrieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Multiscale salient alignment learning for remote sensing image-text retrieval,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.378152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.732957Z digest=sha256:9023c87592b84dd04add98ac28260345c64471ceab96f3fa09561f2f4a872796

Observation 71c63976-4336-441b-8ff7-2f40c3b50181 · outbound

This paper cites Interacting- enhancing feature transformer for cross-modal remote sensing image and text retrieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Interacting- enhancing feature transformer for cross-modal remote sensing image and text retrieval,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.364329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.736707Z digest=sha256:52c61f10ebc66fdc0ce21d4e346620df4586dbed3328ff7558146caeab98e0f8

Observation 9465be24-0dca-4812-a1c8-70fdddfadfc5 · outbound

This paper cites Align before fuse: Vision and language representation learning with momentum distillation,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Align before fuse: Vision and language representation learning with momentum distillation,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.349930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.740423Z digest=sha256:95093f2d356ebda3811288f16a75cd983fe48bfa9ec20309ff533073255ec10c

Observation e6a07d92-cebb-495a-8826-6c0c53a5ed8c · outbound

This paper cites Deep saliency smoothing hashing for drone image retrieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Deep saliency smoothing hashing for drone image retrieval,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.332354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.744193Z digest=sha256:fd859e0820a35c04d27d36fed2138d4a0bd452c0b74175fab45b3c2ccd096cdb

Observation 76ed2888-dad4-4d5a-9698-91351679a81c · outbound

This paper cites Multitask learning for sar ship detection with gaussian-mask joint segmentation,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Multitask learning for sar ship detection with gaussian-mask joint segmentation,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.314377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.748050Z digest=sha256:4c07342869f7a143d35f6659688566906c298f67299e2e9846e1a115c9723f6d

Observation 930a1883-858a-468c-ae0d-e8757f9c2a18 · outbound

This paper cites Swin transformer: Hierarchical vision transformer using shifted windows,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Swin transformer: Hierarchical vision transformer using shifted windows,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.302154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.751703Z digest=sha256:4a07ec438d010b2dc74130e59c14d74519f1c3c820228f0d167c2f6444e51842

Observation 2380a155-4554-40be-b0f7-8cda192742e4 · outbound

This paper cites Global context vision transformers,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Global context vision transformers,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.289581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.755719Z digest=sha256:d4d449d9e975039479042d09f25ee0e08184481f5137437bcd5b4cebfed1a278

Observation 37aa4e64-5910-4588-9fbb-cf1ea37451dc · outbound

This paper cites Matching images and text with multi-modal tensor fusion and re- ranking,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Matching images and text with multi-modal tensor fusion and re- ranking,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.276858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.759505Z digest=sha256:c293482fde26d15e3436db39bcb5f4eec36ee846f7e9176419e8cc7d7e7b859f

Observation 6d79f654-7652-4f77-a091-b4447c4ddbfb · outbound

This paper cites Vse++: Improving visual-semantic embeddings with hard negatives,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Vse++: Improving visual-semantic embeddings with hard negatives,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.264258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.763722Z digest=sha256:978c054f801b4a1a6e5be2aa7e6a245edd840e7fbd428c3f80dc2a03a467be0c

Observation 84ac96ff-b178-43ae-82b3-893433494fb9 · outbound

This paper cites Exploring models and data for remote sensing image caption generation,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Exploring models and data for remote sensing image caption generation,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.767188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.767188Z digest=sha256:a656f95feef5907a5b1e90d9daacae30d725a0cbb53e0823a1fcab03b34be020

Observation fafb1176-24c3-427d-b2a1-e7c343df78e3 · outbound

This paper cites Deep semantic understanding of high resolution remote sensing image,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Deep semantic understanding of high resolution remote sensing image,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.242536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.770808Z digest=sha256:34616ca1920415ec8a2571d665f88b7bbc336741c55f54996af61ef97f597e1d

Observation 44d0bb0b-68d2-4af5-bf61-64699935b562 · outbound

This paper cites End-to-end convolutional semantic embeddings,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval End-to-end convolutional semantic embeddings,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.225883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.774534Z digest=sha256:8a32e275f04f17853ef8d6da3d2de911ec055bf9b66fb4ba22c059d95a781486

Observation 519cb981-5d09-44c6-896e-a8e9aecc3cd6 · outbound

This paper cites Cross-modal semantic correlation learning by bi-cnn network,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Cross-modal semantic correlation learning by bi-cnn network,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.210444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.778748Z digest=sha256:ec65a15b877dab0a997b2d2727ed61cf9cbb7eae6968fdb898ec26aad4997f47

Observation f110aff3-2c77-467e-856d-caf394a2b75a · outbound

This paper cites Dual-path convolutional image-text embeddings with instance loss,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Dual-path convolutional image-text embeddings with instance loss,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.195285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.782641Z digest=sha256:ce5753deeae69caee7a062fdf993a6ca319ebf9ea6355d1ff0545fb6ad557099

Observation e1216bf4-dd3e-4a46-a799-90b70d2aa7f3 · outbound

This paper cites Deep supervised cross-modal retrieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Deep supervised cross-modal retrieval,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.786515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.786515Z digest=sha256:15e4daa2a8519125179853059eb5fb6f6037a7c93a3eeae5db1bd895bfbc387f

Observation 03dc5032-001b-4cd2-9d2d-27ef02950545 · outbound

This paper cites Learning semantic concepts and order for image and sentence matching,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Learning semantic concepts and order for image and sentence matching,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.170650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.790384Z digest=sha256:2db832451accc4fa3254b4f9909279578b605bd93868ecbcc1eb338369bb5e4a

Observation ee5e49a4-9c8a-4ed9-be2d-5070c4577c11 · outbound

This paper cites Stacked cross attention for image-text matching,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Stacked cross attention for image-text matching,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.156138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.794558Z digest=sha256:650e09241c37910cf53b3fdb1f86130f091511c2edbb0a2dc701ef2c51a52287

Observation 8fd228ad-e85b-43c2-a565-2c2e4aa3d91a · outbound

This paper cites Cross- modal attention with semantic consistence for image–text matching,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Cross- modal attention with semantic consistence for image–text matching,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.142074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.798448Z digest=sha256:3f6add5a85d71bba4c5e56b89a9184d923f5ead28b07a717c4099c8975821967

Observation 38ae2995-86c3-4b14-93b2-da73ff3a6f7f · outbound

This paper cites Visual semantic reasoning for image-text matching,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Visual semantic reasoning for image-text matching,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.129587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.802192Z digest=sha256:0f9d5183e9c23ab76a3d66190bcf81c9b939a3ad817d72056d96f60a5776ff7b

Observation 91c20eb1-1199-455f-a128-3099ece39e98 · outbound

This paper cites Image-text embedding learning via visual and textual semantic reasoning,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Image-text embedding learning via visual and textual semantic reasoning,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.115153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.805867Z digest=sha256:978e4e7773fc56b3f0b27165dd3ca2fd5c0ef32fe6e7eb364d511c22c6434a9d

Observation 1414cfaf-f4aa-449e-a1a5-f582b8bb01c7 · outbound

This paper cites Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.809534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.809534Z digest=sha256:02e5a285203e8f6ce48f882d77a1b1b1d4ca199ded39a49da757b1a56acc21fa

Observation ecf46abb-c98e-4236-af91-6bb7d5159778 · outbound

This paper cites Lxmert: Learning cross-modality encoder representations from transformers,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Lxmert: Learning cross-modality encoder representations from transformers,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.093662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.813277Z digest=sha256:e862eb448108846b81c4ff6db930f3ff1e31ec1ba4439e4548684de8666df895

Observation c1a80255-d52c-4e74-87fb-1c654dfcc07f · outbound

This paper cites Fashionbert: Text and image matching with adaptive loss for cross- modal retrieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Fashionbert: Text and image matching with adaptive loss for cross- modal retrieval,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.080079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.816997Z digest=sha256:bc82eb25ec430a1e284d2d5669b05ac1dc68f5ce29fe6f3768c257cf1954312f

Observation 84f014d7-3af0-4263-b883-94ef99a55daa · outbound

This paper cites Learning the best pooling strategy for visual semantic embedding,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Learning the best pooling strategy for visual semantic embedding,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.066838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.820850Z digest=sha256:9ab33c77eb7f29767ff5495eded443e7ca16a1b4936eae7e746bd3ad91eeef4c

Observation 0e90948c-19b9-4e50-8e10-f5988d1ed44f · outbound

This paper cites Pixel-BERT: Aligning Image Pixels with Text by Deep Multi-Modal Transformers.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Pixel-BERT: Aligning Image Pixels with Text by Deep Multi-Modal Transformers

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.824935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.824935Z digest=sha256:1f17638afaed7953cfab4912910f1add303da1ece98d4a25b34216b4f1f22d5e

Observation bf64ed5d-685c-43e1-9734-415474ed2ece · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Learning transferable visual models from natural language supervision,

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.829050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.829050Z digest=sha256:5f4f3064e4cb8b7a0d3a8392b42d458d8b1ea7b0abeb7170e8b6a7e91632db2a

Observation e2c38548-073c-429f-a7d2-c76397f0bc7a · outbound

This paper cites Vista: Vision and scene text aggregation for cross-modal retrieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Vista: Vision and scene text aggregation for cross-modal retrieval,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.043211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.832696Z digest=sha256:0e06e24d4267fd6b269b12026813e41b2818b356b8bf6e012120838aacd949cc

Observation d6f8ddff-9a5c-43b6-82b6-5bc90cca8ffb · outbound

This paper cites Vilt: Vision-and-language transformer without convolution or region supervision,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Vilt: Vision-and-language transformer without convolution or region supervision,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.027761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.836567Z digest=sha256:d94da59e19091d8c63a4586da093f8945e894b72d7d6737ad3007adc09260653

Observation 597fa9db-e019-4720-8555-7da321fee29a · outbound

This paper cites Vlmo: Unified vision-language pre-training with mixture-of-modality-experts,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Vlmo: Unified vision-language pre-training with mixture-of-modality-experts,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:58.013236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.840435Z digest=sha256:d6a86e475d03efd6c22900e94b93ca1020573c193d894960c11a10a0fa0dd5e3

Observation 45903d78-eded-4900-a047-c4855e2cfac6 · outbound

This paper cites Knowledge-aided momentum contrastive learning for remote-sensing image text retrieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Knowledge-aided momentum contrastive learning for remote-sensing image text retrieval,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:57.999616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.844283Z digest=sha256:5fd813ed9c70abea9301c5491ab1dd24f7661381d71eb474b0e61aee3127736a

Observation 9b3afb81-747a-406f-858f-61a9b7dd3591 · outbound

This paper cites Multi- scale interactive transformer for remote sensing cross-modal image-text retrieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Multi- scale interactive transformer for remote sensing cross-modal image-text retrieval,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:57.983362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.848175Z digest=sha256:d88bd7f097495a46e556b2adf644120626b627bb62a5405a123864d671efc56d

Observation 87a56034-0031-4957-9722-3c6a38a28750 · outbound

This paper cites Parameter-efficient transfer learning for remote sensing image-text retrieval,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Parameter-efficient transfer learning for remote sensing image-text retrieval,

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.852135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.852135Z digest=sha256:48e24a4457678375cf6e7467c4fb9fa6a8dc0060c602816677beb5c72cd83104

Observation 8e1b2a61-fe9a-4da8-bcee-5a7041d2c5ce · outbound

This paper cites Integrating multisubspace joint learning with multilevel guidance for cross-modal retrieval of remote sensing images,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Integrating multisubspace joint learning with multilevel guidance for cross-modal retrieval of remote sensing images,

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:57.856091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:57.856091Z digest=sha256:0399bd7c79693d99d9f07242cf86c073695c63e1770c624f543092c3b4a55f76

Observation 1ac19980-e07b-4711-8c32-64bc5507a56c · outbound

This paper cites Bert: Pre-training of deep bidirectional transformers for language understanding,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Bert: Pre-training of deep bidirectional transformers for language understanding,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:57.951457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.859812Z digest=sha256:89edbfe742e968c01592faf4ee0c02629aab326813580ef9d5384f58f6dfb7ec

Observation 852a2f44-d06f-416e-843f-3fb4879354b6 · outbound

This paper cites Momentum contrast for unsupervised visual representation learning,.

Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval Momentum contrast for unsupervised visual representation learning,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:57.937860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T15:05:57.863886Z digest=sha256:280c5cbcb27e3097d8532451ad788b5fee51c5f317a5fcd64a891d02a7bddf51

Pith citing papers

Observation be331e05-aba8-4882-8bdd-c298a7fd228c · inbound

Fast-then-Fine: A Two-Stage Framework with Multi-Granular Representation for Cross-Modal Retrieval in Remote Sensing cites this paper.

Fast-then-Fine: A Two-Stage Framework with Multi-Granular Representation for Cross-Modal Retrieval in Remote Sensing Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-10T00:29:47.841047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-10T00:02:12.727566Z digest=sha256:145ed51b2821b7aac36a655c367cee675bf6df235aec34d6aa9df3961d86cf61